Senior Backend Engineer – LLM Inference
Indexed description
We strive to make a positive mark on the world through the infrastructure we build and give leading teams a service they can truly depend on. Headquartered in Helsinki, we operate globally with offices in London and San Francisco.
Join Verda while it’s still being built - not once it’s finished.
About The Role
At Verda, we are building the platform that makes LLM inference reliable and accessible on our AI cloud.
As a Senior Backend Engineer, you will develop backend services and Kubernetes-native applications that support our inference offering. You will work closely with infrastructure and AI engineering teams to improve the platform’s reliability, scalability and developer experience.
This is a hands-on role with ownership from design and implementation through deployment and production operation. You will help shape technical decisions as the platform grows.
Your Responsibilities
- Design, build and maintain production backend services in Go.
- Develop Kubernetes-native applications and automation for reliable platform operations.
- Build APIs and integrations that support our inference services and wider cloud platform.
- Improve scalability, performance and reliability through testing, observability and operational improvements.
- Collaborate across teams on technical design, delivery and production troubleshooting.
- Experience designing, shipping and operating production backend services, with the judgment to own work from an initial design through production rollout.
- Strong Go experience. You have shipped and maintained production Go services and are comfortable with concurrency, context propagation, API design and error handling.
- A solid understanding of asynchronous execution, concurrency, HTTP and long-lived connections. You can reason about timeouts, resource limits, cancellation and failure propagation.
- Distributed systems experience, including caching, consistency, retries, idempotency and reliable event processing.
- Strong SQL and relational data modelling skills, including transactions, schema migrations and PostgreSQL.
- Hands-on Kubernetes-native development experience, including controllers or operators, custom resources and reconciliation loops. You understand idempotent reconciliation, desired versus observed state, retries and resource lifecycle handling.
- Experience deploying and troubleshooting services on Kubernetes, including Helm, networking, readiness, configuration and safe rollouts.
- A disciplined approach to validation: meaningful tests, load measurements and failure scenarios, with clear communication about what has and has not been verified.
- Clear written English and the ability to explain design decisions, trade-offs and operational behavior to other teams.
- Rust experience, particularly in asynchronous or network services.
- Experience with API gateways, networking or traffic management.
- Familiarity with LLM API protocols and inference systems such as vLLM, SGLang, NVIDIA Dynamo or llm-d.
- Experience with Valkey/Redis, NATS JetStream or comparable distributed state and messaging systems.
- Experience with reliable event processing or usage metering systems.
- OpenTelemetry, Prometheus and capacity planning for workloads with variable request costs.
- Service discovery, GitOps or deployments across multiple Kubernetes clusters and regions.
- Python experience for integration with existing inference tooling.
- Cash and equity compensation along with various fringe benefits.
- Profitable operations with rapid, sustained growth.
- 40+ nationalities, with 6 different ones on the management team.
- A real chance to make an impact and work alongside world class engineers, researchers, and partners across the global AI ecosystem.
- Work mode: Based in Helsinki, Finland or London, UK or remote in Europe
- Level: Senior
- Employment type: Full time and permanent
Please submit your application through our Careers page. We don't accept applications sent by email.
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search