SRE Mid
Indexed description
๐ผ What You'll Be Doing
๐ Ensure the availability, stability, and performance of business-critical production platforms.
โ๏ธ Automate operational tasks and infrastructure processes to improve efficiency and reliability.
๐ Monitor infrastructure and applications, proactively identifying issues before they impact users.
๐ Investigate production incidents, perform root cause analysis, and implement long-term solutions.
๐ ๏ธ Manage and optimize Linux and Windows environments, middleware, databases, and enterprise infrastructure.
โ๏ธ Support and improve container-based platforms and modern infrastructure environments.
๐ Develop monitoring, alerting, and observability solutions to improve platform visibility.
๐ค Work closely with Development, Infrastructure, Security, and Operations teams to improve system reliability and deployment processes.
๐ก Continuously improve platform resilience through automation, standardization, and DevOps best practices.
๐ ๏ธ What We're Looking For
โ 6+ years of experience in Site Reliability Engineering, DevOps, Infrastructure Engineering, or similar roles.
๐ง Strong experience administering Linux and Windows environments.
โ๏ธ Experience with enterprise schedulers such as BMC Control-M or similar.
๐ป Strong scripting skills using Shell, Python, and/or PowerShell.
๐ Experience with middleware technologies including:
- JBoss
- Tomcat
- Apache
- IIS
๐๏ธ Experience with relational databases such as:
- Oracle
- SQL Server
- Sybase
- MySQL
๐ Good knowledge of SQL.
๐ Experience with monitoring and observability platforms such as Zabbix, Grafana, or similar tools.
๐พ Experience with infrastructure backup and recovery processes.
๐ Good understanding of:
- LDAP
- Active Directory
- Kerberos
- Firewalls
- Reverse Proxies
- Networking fundamentals
๐ณ Experience working with container-based infrastructure.
โญ Nice to Have
โ๏ธ Experience with Kubernetes and container orchestration.
๐ Experience with CI/CD pipelines and Infrastructure as Code.
๐ค DevOps automation experience.
๐ Knowledge of SRE principles, SLIs, SLOs, and error budgets.
๐ Experience supporting highly available enterprise production environments.
๐ Who You Are
๐ก Passionate about reliability engineering and automation.
๐ Proactive, curious, and continuously looking for ways to improve systems.
๐ Strong analytical and troubleshooting skills.
๐ค A collaborative team player who enjoys working across engineering teams.
๐ Committed to continuous learning and technical excellence.
โก Comfortable operating in fast-paced, mission-critical environments.
๐ฃ๏ธ Strong communication skills with the ability to collaborate across multidisciplinary teams.
๐ฌ๐ง English proficiency at B2 level or above.
๐ Why Join Us?
๐ Build Reliable Platforms
Help ensure the availability and resilience of business-critical enterprise systems.
โ๏ธ Modern Engineering Practices
Work with automation, observability, container platforms, enterprise infrastructure, and DevOps methodologies.
๐ Collaborative Engineering Culture
Join an international team where innovation, knowledge sharing, and continuous improvement are encouraged.
๐ Grow Your Career
Take on challenging technical projects while expanding your expertise in SRE, cloud technologies, and platform engineering.
๐ป Real Technical Ownership
Influence operational excellence, platform reliability, and engineering best practices across the organization.
โจ Make an Impact
Your work will directly improve the performance, scalability, and reliability of systems used by thousands of users every day.
๐ Additional Information
๐ International and collaborative working environment.
โฐ Dynamic, fast-paced projects with exciting technical challenges.
๐ฌ English (B2 or above) is mandatory.
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search