Descrição da vaga

Texto agregado para leitura rápida. Confira sempre a fonte original ao enviar a candidatura.

Our customer's product is an AI-powered platform that helps businesses make better decisions and work more efficiently. It uses advanced analytics and machine learning to analyze large amounts of data and provide useful insights and predictions. The platform is widely used in various industries, including healthcare, to optimize processes, improve customer experiences, and support innovation. It integrates easily with existing systems, making it easier for teams to make quick, data-driven decisions to deliver cutting-edge solutions.

Requirements

  • Bachelor's or Master's degree in Computer Science or a related field
  • Strong Python coding skills - 7+ years
  • 2+ years of hands-on experience with machine learning and production LLM systems
  • Experience building backend APIs with FastAPI, async patterns, rate limiting, and SQLAlchemy - 3+ years
  • Experience designing maintainable and extensible systems using dependency injection, interfaces, and abstract base classes
  • Experience with vector databases such as Pinecone, Weaviate, or Chroma, as well as hybrid search
  • Strong understanding of RAG architectures, including retrieval, reranking, context assembly, and response generation
  • Hands-on experience with LangChain and LangGraph for building and orchestrating LLM workflows
  • Advanced Python skills, including async/await, type hints, Pydantic, and SOLID principles
  • MLOps experience with MLflow, model versioning, and A/B testing; experience with Langfuse is a plus
  • Experience in NLP and computer vision, including document understanding, OCR, and GPT-4 Vision
  • Experience building feature pipelines, real-time and batch inference systems, and model serving
  • Hands-on experience with Hugging Face is required; experience with LlamaIndex is a plus
  • Familiarity with database technologies such as SQL
  • Good problem-solving skills and the ability to work in a fast-paced, team-oriented environment

Nice to have skills:

  • Understanding of DevOps, CI / CD including: Docker containerization, Azure DevOps pipelines or GitHub Actions, Kubernetes (nice to have);
  • Data security including: Multi-tenant data isolation, Secure key management (Azure Key Vault), Audit trail implementation;
  • Experience in designing on cloud platform including: Azure (strongly preferred): Azure OpenAI, Blob Storage, Key Vault, Container Registry, AWS or GCP;
  • Experience in data engineering in Big Data systems including: Large-scale data processing, ETL/ELT pipelines
  • Rate limiting and quota management for high-throughput API usage
  • Cost management and optimization for LLM usage at scale
  • Document processing expertise (PDF extraction, OCR tooling)
  • Production incident management and on-call experience
  • Testing strategies for non-deterministic LLM outputs (e.g., golden datasets, fuzzy matching)
  • Domain knowledge in regulated industries (e.g., healthcare/pharma workflows, regulatory compliance) is a plus

Responsibilities:

  • Build, refine, and use ML Engineering platforms and components; develop and implement scalable backend systems, APIs, and microservices using FastAPI
  • Implement MLOps including model KPI measurement, tracking, model drift detection, and model feedback loops
  • Deploy and operationalize ML and Deep Learning models, with a strong focus on LLMs and Generative AI
  • Integrate Azure OpenAI (GPT-4, GPT-4 Vision) and other LLM providers with proper retry logic and error handling
  • Maintain up-to-date knowledge of state-of-the-art technologies such as LLMs, GenAI, and transformer architectures
  • Scale machine learning algorithms to work on massive data sets under strict SLAs
  • Build and orchestrate model pipelines including feature engineering, inferencing, and continuous model training
  • Write backend application code in Python and SQL using strong object-oriented principles and asynchronous programming (asyncio, async/await)
  • Implement dependency injection patterns and layered architecture (Service, Foundation, Orchestration, DAL)
  • Build LLM observability (e.g., Langfuse) to track prompts, tokens, costs, and latency
  • Develop prompt management systems with versioning and fallback mechanisms
  • Implement Celery (or similar) workflows for asynchronous task processing and complex pipelines
  • Build multi-tenant architectures with client data isolation
  • Implement cost optimization strategies for LLM usage (prompt caching, batch processing, token optimization)
  • Integrate third-party APIs and services (e.g., document/OCR services, cloud storage, enterprise systems)
  • Collaborate with client-facing teams to understand business context and contribute to technical requirement gathering
  • Write production-ready code that is testable, maintainable, and accounts for edge cases and errors
  • Ensure high quality of deliverables by following architecture/design guidelines, coding best practices, and periodic design/code reviews
  • Write unit tests and higher-level tests to handle expected edge cases and errors gracefully
  • Troubleshoot backend application code using structured logging and distributed tracing
  • Use bug tracking, code review, version control, and other tools to organize and deliver work
  • Participate in scrum calls and agile ceremonies, communicating progress, issues, and dependencies
  • Document application changes and updates, including API documentation via OpenAPI/Swagger
  • Research and evaluate emerging architecture patterns and technologies through rapid learning, proofs-of-concept, and prototypes

Benefits

  • Awesome projects with an impact
  • Udemy courses of your choice
  • Team-buildings, events, marathons & charity activities to connect and recharge
  • Workshops, trainings, expert knowledge-sharing that keep you growing
  • Clear career path
  • Absence days for work-life balance
  • Flexible hours & work setup - work from anywhere and organize your day your way

Vagas relacionadas

Seleção por stack em comum com esta oportunidade

LinkedIn

Desenvolvedor Backend Pleno (Node.js)

Marinho Tech Talent Greater Salvador 25 candidaturas Hoje

Salário estimado

R$ 7k - 10k/mês

Pleno CLT

Voltar para todas as vagasBACKENDDesenvolvedor Backend Pleno (Node.js)Marinho Tech TalentDescrição da VagaProcuramos um Desenvolvedor Backend Pleno com experiência em Node.js para construir e otimizar a lógica de nossos servidores. Você trabalhará em um ambiente dinâmico, colaborando com equipes de ...

Ver Detalhes
Remoto LinkedIn

Software Engineer - Backend

BMP São Paulo 25 candidaturas Hoje

Salário estimado

R$ 9k - 13k/mês

Pleno CLT

Você é um apaixonado por negócios e Tecnologia e está em busca de novos desafios? Então, a BMP é o seu lugar! Estamos crescendo cada dia mais e buscando pessoas com espírito de liderança, criativas e atentas à inovação para fazer parte do nosso time. Se você está pronto para fazer parte de um ambien...

Ver Detalhes
LinkedIn

Mid Level Python Developer , Brazil

CI&T Brazil 25 candidaturas Hoje

Salário estimado

R$ 7k - 10k/mês

Pleno CLT

At CI&T, we help large enterprises transform the potential of AI into real business impact with AI Deployment, AI-native execution, and tech-integrated business solutions.With 30 years of experience in technological transformation, we accelerate innovation with expertise in Agentic SDLC, Application...

Ver Detalhes