Junior Site Reliability Engineer
Indexed description
Exceptional Engineering: Our engineering team is far from average; it's a powerhouse of talent. Comprising former FAANG employees and programming champions, they bring deep expertise and a shared commitment to building world-class products. Our strength lies in engineering solutions that leverage the power of AI, Data, and distributed computing. We thrive in a fast-paced startup environment, where agility, creativity, and quality drive everything we do.
Global Presence, Local Innovation: Cosmose AI is a global company with a presence in 8 countries, transforming content engagement through cutting-edge technology. Our flagship product, KAIKAI Lock Screen, delivers AI-powered, personalized content to users worldwide, unlocking new and exciting opportunities for product placement.
We are looking for a Junior Site Reliability Engineer to join our engineering team. In this role, you will contribute to building reliable infrastructure, optimizing performance, and automating processes to support our innovative, AI-driven solutions. If you’re passionate about automation, optimizing performance, and ensuring system scalability, we’d love to have you on board!
TL;DR Why us:
- Big Scale & Startup Culture: Join a company that combines the scale of a major player with the agility and innovation of a startup
- Exceptional Team: Collaborate with a team of former employees from tech giants like Google, Meta, ByteDance, and Microsoft, as well as ACM ICPC programming champions
- Stock Options: As we grow and succeed together, you'll have the chance to benefit from it
- Flat Structure: Enjoy a flat organizational structure that promotes collaboration and provides opportunities to learn from seasoned technical leadership
- Automate Routine Tasks: Develop tools to automate administrative tasks, reducing manual intervention and improving efficiency.
- Optimize System Performance: Create automated solutions to monitor and maintain system performance, ensuring reliability and scalability.
- Enhance Security: Develop and deploy automated security measures, vulnerability scans, and compliance checks to protect our infrastructure and data.
- Operate a Cross-Cloud Network.: Monitor, troubleshoot, and expand an overlay network spanning multiple cloud providers and continents, ensuring performant and reliable connectivity.
- Cloud: AWS
- Operating systems: Linux
- Programming: Python
- Orchestration: Kubernetes, Helm
- Monitoring and alerting: Prometheus, Grafana, VictoriaMetrics
- HTTP server: HAProxy, Nginx
- CI&CD: Jenkins, Ansible
- Databases: Redis, Postgres, Clickhouse
- Message Queues: Kafka, RabbitMQ
- IaC - Terraform
- Solid understanding of Linux operating systems
- Programming skills, particularly in Python
- Experience in managing a Linux- or Unix-based server
- Conceptual understanding of orchestration tools such as Kubernetes
- A collaborative mindset with excellent problem-solving skills and attention to detail
- Good communication within the team (Polish and English are required)
- Knowledge of monitoring tools like Prometheus and Grafana.
- Practical experience with cloud platforms like AWS
- Familiarity with infrastructure as code (IaC) tools like Terraform.
- Experience with CI/CD pipelines
We’re building something big - and we’re just getting started.
If you’re excited about shaping the future of AI-powered mobile content and working with a team that’s pushing boundaries, we’d love to meet you.
👉 Be part of the Cosmose story. Apply now!
Read More On
TechCrunch, Business Insider
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search