Forward Deployed Engineer III, Applied AI
Indexed description
Note: By applying to this position you will have an opportunity to share your preferred working location from the following: Zürich, Switzerland; London, UK.Minimum qualifications:
- Bachelor’s degree or equivalent practical experience.
- 2 years of experience with software development in one or more programming languages (e.g., Python or C++).
- 2 years of experience with GenAI techniques (e.g., LLMs, Multi-Modal, Large Vision Models) or with GenAI-related concepts (language modeling, computer vision).
- 1 year of experience with AI/ML infrastructure (e.g., model deployment, model evaluation, model optimization, data processing, or debugging).
- Experience developing and deploying multilingual natural language processing models.
- Master's degree or PhD in Computer Science, AI, Machine Learning, or a related technical field.
- Expertise in debugging Agent logic (ReAct loops, Chain of Thought) and optimizing tool selection.
- Experience implementing multi-agent systems using frameworks (e.g., LangGraph, CrewAI, or Google’s Agent Development Kit (ADK)) and complex patterns like ReAct, self-reflection, and hierarchical delegation.
- Experience connecting agents to enterprise knowledge bases and optimizing RAG chunking to prevent hallucinations.
As a Forward Deployed Engineer (FDE) in Applied AI, you are the "Agent Engineer" and the primary delivery arm for our customers' most critical AI initiatives. You will take initial conversational prototypes and transform them into production-ready solutions, owning the end-to-end engineering life-cycle.
The Cloud Applied AI (AAI) powers business growth with Gemini Enterprise. Our portfolio includes Gemini Enterprise for Customer Experience (Shopping Agent, CX Agent Studio, Agent Assist, Vertex AI Search - Commerce, Customer Experience Insights), along with other vertical and domain packaged solutions. We enable high adoption and speed to value by building solutions that are quickly deployed, delivering new 0-to-1 capabilities with startup agility. Team members operate at the forefront of AI, collaborating directly with model builders with unprecedented speed. Join us to work on cutting-edge projects and shape the future of AI in a fast-paced, collaborative, and impactful environment.
Responsibilities
- Serve as the delivery lead for conversational AI pilots, taking initial proof-of-value code and refactoring it for production by replacing stubs with secure, real-world integrations.
- Own the delivery of the initial Customer User Journeys (CUJs), ensuring that conversational flows are not just functional, but optimized for (top customer/industry priorities).
- Architect and optimize complex agentic workloads, focusing on reasoning loops, tool selection, and reducing latency while maintaining production-grade security and networking.
- Implement critical infrastructure, including rate limiting, error management, and regional failover strategies; develop monitoring systems and alerting dashboards during "Go-Live" windows to ensure stability as traffic scales.
- Architect security perimeters (VPC, IAM, CMEK) to satisfy CISO requirements and optimize RAG strategies to balance latency, accuracy, and cost.
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search