Back to search
Berlitz Linkedin · Posted 20d ago

Senior Machine Learning Engineer

Germany

Linkedin
Continue to application Add your email once, then Caio opens the original posting.

Indexed description

We're looking for a Senior Machine Learning Engineer to build the AI services behind our global online learning platform, giving learners real-time speech practice, conversation, and feedback. You'll own the models and AI integrations powering our speaking-practice and evaluation features: classical and fine-tuned small models, edge models, cloud-based models, and integrations with external providers. This is a coding role: you'll write, maintain, and update the ML services that ship to production, using AI coding tools as part of your daily workflow.


You'll work cross-functionally with the mobile and platform teams, providing model-serving APIs, evaluation harnesses, and AI architecture guidance, with high autonomy and low bureaucratic friction. You'll work across the full perception stack (speech, text, CV ), with room to go deep on the modalities most relevant to what we're building next.


Core Responsibilities

  • Modeling & evaluation. Build, evaluate, and productionize models across the ML lifecycle, from classical/statistical approaches to fine-tuning small, edge-friendly transformers. Evaluation is metric-driven, not just accuracy, with human-agreement baselines where applicable.
  • GenAI & LLM integration. Design and operate integrations with hosted LLM providers: prompt and evaluation design, LLM -as-judge patterns, provider routing/failover, and cost/latency tradeoffs.
  • Edge & cloud models. Build and run models both on-device and in the cloud, and choose deliberately between the two based on latency, privacy, and cost.
  • Multimodal perception. Work across speech ( ASR / TTS ), NLP , and CV pipelines as the product requires. Deep expertise in every modality isn't required; the judgment to reason about tradeoffs across all of them is.
  • Production engineering. Set up, maintain, and update ML services in production: serving constraints, monitoring, drift, and connecting offline model scores to real learner outcomes. This is hands-on coding work, not just modeling.
  • Data pipelines. Build and maintain the ETL pipelines that feed these models: pulling from storage and databases, cleaning and transforming data, and reducing noise in real-world learner data.


Required Skills & Qualifications

  • Experience: 5+ years building and shipping machine learning systems to production, spanning classical ML and applied deep learning.
  • Languages: Rust or Go preferred; C++ also valued. Python (or Julia ) for data science, modeling, and exploratory analysis.
  • ML frameworks: framework-agnostic ( PyTorch , JAX , TensorFlow ); real production experience in at least one is what matters.
  • Tools: Git for version control, code review, and collaborating across parallel teams.
  • Daily, hands-on use of AI coding tools/agents as part of how you build and ship, not an occasional aid.
  • Hands-on experience integrating hosted LLM providers into production systems: prompting, evaluation, cost/latency tradeoffs, multi-provider routing.
  • Comfortable building models from scratch and fine-tuning pretrained ones, choosing rigorously between them based on the problem's actual constraints.
  • Strong communication: explaining modeling tradeoffs to engineers and product impact to non-technical stakeholders.


Nice to have

  • Model compression, quantization, or distillation for on-device deployment.
  • Hands-on experience in a couple of NLP , speech ( ASR / TTS ), or CV ; full expertise in every modality is not expected.
  • Data ETL and manipulation experience: S3 or similar object storage, Postgres or other databases, data transformation, noise reduction.


Why This Role

  • Full-stack perception. Speech, text, and CV all live under one small team: broad exposure rather than a narrow lane.
  • Greenfield. No legacy ML pipelines or technical debt; you help design the model-serving architecture from scratch.
  • Real constraints, real judgment. Edge deployment and multi-provider LLM routing mean the job is about tradeoffs, not just model accuracy.
Free. 20 seconds. No password. See every match in this search.

Create a free Caio profile to unlock more results and save your role and location preferences.

Unlock free search
Want help applying to roles like this? Search Caio for free. If repetitive applications get heavy, Managed Job Search adds supervised execution for $99/month.
View Managed Job Search