Site Reliability Engineering (Sre)
Indexed description
At Fyld, we believe the future is built with people, technology, and strong values.
We are a Portuguese consulting company that operates with transparency, respect, and a focus on everyone's growth.
Here, every project is an opportunity to build reliable, scalable and resilient technology solutions.
Our philosophy?
Code for Big Solutions
— because we don't just build technology to make things work, we build systems that remain reliable as they scale.
We're looking for an experienced
Site Reliability Engineer (SRE)
with strong technical skills and a passion for reliability, automation and operational excellence.
By joining Fyld, you'll be part of a team where every member is challenged to innovate, collaborate and grow.
Your Profile
Bachelor's degree in Computer Science, Software Engineering, Information Technology, Engineering or a related field
Proven experience as an
SRE, DevOps Engineer, Platform Engineer or similar role
Strong understanding of
site reliability engineering principles
, including availability, reliability, scalability and resilience
Experience defining and working with
SLIs, SLOs, SLAs and error budgets
Hands-on experience with
observability
, including metrics, logs, traces and alerting
Experience with tools such as
Prometheus, Grafana, OpenTelemetry, ELK or equivalent observability platforms
Strong experience with
cloud platforms
, such as AWS, Azure or GCP
Experience with
Kubernetes and Docker
in production environments
Strong knowledge of
Infrastructure as Code
, using Terraform, OpenTofu, Ansible or similar technologies
Strong automation and scripting skills using
Python, Go, Bash, PowerShell or similar
Experience with incident management, troubleshooting and
root cause analysis
Experience conducting post-incident reviews and implementing preventive measures
Knowledge of performance monitoring,
capacity planning and system optimisation
Understanding of networking fundamentals, including TCP/IP, DNS, load balancing and
Familiarity with CI/CD and modern software delivery practices
Strong focus on automation and reducing manual operational work
Ability to work closely with software engineering, infrastructure, security and product teams
Strong analytical and problem-solving skills
Fluency in English
Nice To Have
Experience operating large-scale or business-critical production environments
Knowledge of
GitOps
and tools such as Argo CD or Flux
Experience with service meshes or distributed systems
Knowledge of chaos engineering and resilience testing
Experience with Kubernetes platform engineering
Familiarity with cloud cost optimisation and FinOps
Relevant certifications such as
CKA, AWS Certified DevOps Engineer or Google Cloud Professional Cloud DevOps Engineer
Our Values
Critical thinking and autonomy
Focus on meaningful solutions and real results
Respect, transparency, and continuous growth
What You'll Find At Fyld
A collaborative team that values individual progress
Access to continuous training and certifications
A culture of proximity, respect, and recognition
Challenging, purpose-driven projects where your talent as a
Site Reliability Engineer
has real impact
Want to be part of it? Send your CV to
Fyld. Code for what really matters.
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search