Back to search
Nagarro Himalayas · Posted 2d ago

Senior Staff Engineer (LLM)

Full time Remote

LLM Engineering Machine Learning Engineering Python Development AI Engineering
Continue to application Add your email once, then Caio opens the original posting.

Indexed description

Role Overview: The LLM Engineer will join an existing development team to build and ship LLM-powered features in a complex, production application used at scale. This is a hands-on, full-stack role spanning backend services, APIs, and the LLM systems (retrieval, agents, evaluation) that power them. You are expected to work as an agentic engineer”using AI coding tools and autonomous agents to write code, automate workflows, and optimize how the team delivers. You will collaborate with global teams across multiple time zones and own features end to end.

In this role, you will:

  • Bring senior-level Python and LLM engineering expertise to the team.
  • Execute both planning and hands-on technical work independently.
  • Collaborate effectively with Product Owners and other stakeholders to solve complex problems. Work cross-functionally to deliver impactful solutions across teams.
  • Continuously develop your technical expertise and stay current with new technologies.
  • Bring curiosity and drive to expand your skills and knowledge.
  • Use a data-driven approach to solve technical challenges and make informed decisions.
  • Apply systems-level thinking that integrates data science and engineering principles.
  • Take full ownership of the features and projects you work on, delivering high-quality solutions independently.

Must-Have Skills:

  • Hands-on, daily use of AI-assisted and agentic coding tools (e.g., Claude Code, Cursor, GitHub Copilot, autonomous coding agents) to write and refactor code, automate workflows, and optimize engineering processes.
  • Strong experience with Python, particularly in building REST APIs using frameworks like FastAPI.
  • Grounding in NLP and machine learning as they relate to building LLM systems
  • Strong experience working with key LLM models APIs (e.g. OpenAI, Anthropic) Experience building, deploying, and securing MCP servers at scale.
  • Understanding of multi-agent systems and their applications in complex problem-solving scenarios.
  • Designing and implementing RAG systems end to end: vector databases, semantic search, retrieval quality, and chunking strategy.
  • Experience with prompt writing for various use cases
  • Experience with generative solutions released to prod, at scale, beyond POCs
  • Proficiency with server-side events, event-driven architectures, and messaging systems.
  • Strong critical thinking and systems thinking skills, with experience debugging, optimizing, and making sound engineering decisions across complex backend systems, not just solving isolated problems.
  • Solid understanding of security best practices for backend systems, including authentication and data protection.
  • Other Qualifications:
  • 2+ years of experience developing and experimenting with LLMs
  • 8+ years of experience developing APIs with Python

Nice-to-Have Skills:

  • Experience with LLM guardrails Experience with LLM Frameworks (e.g. LangChain, LlamaIndex)
  • Experience with LLM monitoring and observability
  • Experience developing AI/ML technologies within large and business critical applications
  • Building evaluation into LLM systems: eval harnesses, regression suites, LLM-as-judge, and offline/online quality metrics.

Must have Skills : Python, FastAPI, LLM Application Frameworks, Vector Databases and Embeddings,

Originally posted on Himalayas

Free. 20 seconds. No password. See every match in this search.

Create a free Caio profile to unlock more results and save your role and location preferences.

Unlock free search
Want help applying to roles like this? Search Caio for free. If repetitive applications get heavy, Managed Job Search adds supervised execution for $99/month.
View Managed Job Search