DevOps Engineer
Indexed description
https://www.devsamurai.com/en/
About the RoleWe are looking for a Junior DevOps Engineer to take responsibility for the reliability, availability, and day-to-day operation of our cloud-based applications.
You will primarily work with Google Cloud Platform (GCP) and work closely with developers to keep our websites and services stable, secure, and available.
This role is ideal for someone with solid hands-on experience who wants to grow deeper into cloud infrastructure and DevOps.
Responsibilities- Manage and maintain production infrastructure primarily on GCP.
- Take ownership of website uptime, availability, and reliability.
- Monitor system health, performance, logs, and alerts.
- Investigate and resolve production incidents, including downtime, errors, and performance issues.
- Support application deployments and ensure releases are performed safely.
- Maintain and improve monitoring, alerts, and operational processes.
- Automate repetitive operational and deployment tasks.
- Work closely with developers to identify and resolve infrastructure-related issues.
- Identify potential reliability problems and proactively improve system stability.
- 1–3 years of hands-on experience in DevOps, Cloud Engineering, System Administration, or a related role.
- Solid hands-on experience with Google Cloud Platform (GCP).
- Good understanding of common GCP services and cloud infrastructure.
- Good understanding of Linux, networking, DNS, HTTP/HTTPS, and SSL/TLS.
- Experience monitoring and troubleshooting production web applications.
- Ability to investigate incidents using logs, metrics, and monitoring tools.
- Experience with CI/CD and application deployment processes.
- Basic scripting skills with Bash, Python, or similar.
- Good understanding of Git.
- Strong troubleshooting and problem-solving skills.
- Strong ownership and accountability for production systems and uptime.
- Willingness to participate in production support or on-call activities when required.
- Experience with Docker.
- Experience with GCP services such as App Engine, Compute Engine, Cloud SQL, Cloud Storage, VPC, and Load Balancing.
- Experience with monitoring tools such as Prometheus, Grafana, or similar.
- Experience with Terraform or other Infrastructure as Code tools.
- Experience improving application availability, performance, or infrastructure reliability.
- Our websites and services remain stable and highly available.
- Production issues are detected and resolved quickly.
- Deployments are reliable and have minimal impact on users.
- Monitoring provides clear visibility into system health.
- Recurring incidents are identified and prevented.
- Manual operational work is gradually automated.
GCP → Uptime & Reliability → Monitoring → Troubleshooting → Deployment → Automation
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search