Gen AI Engineer (LLM &RAG)
Indexed description
Role Experience
4+ years in ML/NLP or software engineering, with 1.5+ years hands-on building LLM / RAG applications. Proven delivery of a RAG system: document ingestion, embeddings, vector search, prompt orchestration and evaluation. Experience with hallucination control, grounding, citation of sources, and structured evaluation of GenAI quality. Familiarity with fine-tuning / domain adaptation and with prompt-injection and jailbreak defense.
Core Skills & Capabilities :Microsoft stack (reference build):
Azure OpenAI Service, Azure AI Search, Azure AI Content Safety, prompt flow and evaluation in Azure AI Foundry. Copilot Studio for conversational experiences on the KIB intranet.
Open-source / custom stack:Python with LangChain / LlamaIndex / Semantic Kernel; Hugging Face Transformers and open embedding models.Open vector databases (pgvector, Qdrant, Chroma, Milvus); open evaluation tooling (RAGAS, DeepEval); local serving with Ollama / vLLM.Open and fine-tunable models (Llama, Mistral) with LoRA / PEFT techniques.
Security, access & data management:Implements guardrails and content filtering (Azure AI Content Safety or open equivalents such as Llama Guard / NeMo Guardrails) against prompt injection, data leakage and unsafe output.Enforces retrieval-time access control so a user only ever sees content their identity is entitled to (document-level and row-level security passed from Entra ID / the source system).Prevents sensitive data (PII, client, PCI) from being logged or sent to models outside the approved boundary; applies masking and redaction in the pipeline.Builds source-citation and audit trails so every answer can be traced to approved material — essential for regulatory defensibility.
Skills: llm,rag,genai
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search