AI Architect – Generative AI, LLMs & AWS
Indexed description
Job Description: AI Architect – Generative AI, LLMs & AWS
Location: West Palm Beach, FL
Role Summar
yWe are looking for a hands-on AI Architect to lead the design, development, enhancement, and deployment of an enterprise AI platform powered by Large Language Models (LLMs). The ideal candidate will combine strong software engineering skills with deep expertise in Generative AI, Agentic AI, and AWS cloud services to build scalable, secure, and production-ready AI applications
.This role requires active involvement in solution architecture, coding, cloud deployment, performance optimization, and mentoring engineering teams
.
Key Responsibiliti
- esDesign, develop, and enhance an enterprise AI platform integrating multiple Large Language Models (LLMs
- ).Architect and implement scalable AI solutions using modern Agentic AI frameworks and Retrieval-Augmented Generation (RAG) architecture
- s.Build intelligent AI agents capable of reasoning, planning, tool execution, and workflow orchestratio
- n.Develop secure, scalable, and highly available cloud-native AI applications on AW
- S.Design and implement REST APIs and microservices to expose AI capabilities to enterprise application
- s.Integrate AI services with enterprise systems, databases, APIs, and third-party platform
- s.Deploy, monitor, and optimize AI workloads on AWS ensuring performance, scalability, security, and cost efficienc
- y.Work closely with product owners and engineering teams to translate business requirements into technical solution
- s.Drive architecture reviews, code quality, CI/CD automation, and engineering best practice
- s.Evaluate emerging LLMs, AI frameworks, and cloud services to continuously improve the AI platfor
- m.Mentor developers and provide technical leadership across AI initiative
s.Required Technical Skil
lsGenerative
- AIStrong hands-on experience with OpenAI GPT, Anthropic Claude, Llama, Mistral, Amazon Nova, or similar foundation model
- s.Experience building enterprise-grade LLM application
- s.Expertise in Retrieval-Augmented Generation (RAG
- ).Prompt engineering, prompt optimization, embeddings, semantic search, and model evaluatio
- n.Experience integrating multiple LLM providers and managing model orchestratio
n.Agentic
AIHands-on experience with one or more o
- f:LangCha
- inLangGra
- phCrew
- AIMicrosoft Semantic Kern
- elAutoG
- enAmazon Bedrock Agen
tsExperience developin
- g:Multi-agent workflo
- wsTool calli
- ngFunction calli
- ngMemory manageme
- ntPlanning and reasoning agen
tsAWS Clo
udStrong hands-on experience wit
- h:Amazon Bedro
- ckAmazon SageMak
- erAWS Lamb
- daECS/E
- KSAPI Gatew
- ayStep Functio
- nsAmazon
- S3Dynamo
- DBAmazon OpenSear
- chAmazon Auro
- raCloudWat
- chI
- AMV
- PCEventBrid
- geSecrets Manag
erExperience with Infrastructure as Code (Terraform, AWS CDK, or CloudFormation) is highly desirabl
e.Programmi
- ngPython (mandator
- y)FastAPI / Fla
- skREST AP
- IsMicroservic
- esDock
- erKubernet
- esG
- itCI/CD pipelin
esDatabases & Sear
chExperience wit
- h:PostgreS
- QLDynamo
- DBOpenSear
- chPineco
- neWeavia
- teChro
- maFAI
- SSMilv
usPreferred Qualificatio
- nsExperience building AI products from concept to productio
- n.Strong understanding of LLMOps, MLOps, observability, and AI monitorin
- g.Experience implementing AI guardrails, responsible AI, and enterprise security control
- s.Familiarity with event-driven and serverless architecture
- s.Knowledge of authentication, authorization, and API securit
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search