Back to search
PRACYVA Linkedin · Posted 3d ago

GCP/AWS Networking SRE

United Kingdom

Linkedin
Continue to application Add your email once, then Caio opens the original posting.

Indexed description

We are looking for a Senior Cloud Network SRE / Cloud Platform Engineer with strong expertise in GCP/AWS networking, incident management, Kubernetes, and platform reliability engineering. The ideal candidate should be capable of troubleshooting complex cloud networking issues, leading production incident resolution, and driving reliability improvements through automation and observability.

Key Responsibilities

  • Troubleshoot and resolve complex GCP/AWS network incidents in production environments.
  • Perform deep-dive analysis using TCP Dumps, packet captures, and network diagnostics tools.
  • Diagnose and resolve issues related to TCP/IP, DNS, Routing, Firewalls, VPNs, Load Balancers, Network Overlays, Network Segregation, and Peer-to-Peer communication.
  • Support and manage Kubernetes platforms (GKE/EKS).
  • Build and maintain monitoring, alerting, and observability solutions.
  • Perform incident response, problem management, and Root Cause Analysis (RCA).
  • Automate operational tasks using Python and Shell scripting.
  • Manage Infrastructure as Code (Terraform) and CI/CD pipelines.

Mandatory Skills

  • Strong hands-on experience in GCP and/or AWS Networking.
  • Expertise in TCP/IP, DNS, Routing, Firewalls, VPNs, Load Balancing, and Packet Analysis.
  • Experience with tcpdump, and network troubleshooting tools.
  • Strong knowledge of Network Overlays, Network Segmentation, and Kubernetes Networking.
  • Hands-on experience with Kubernetes (GKE/EKS).
  • Experience with Dynatrace, or Cloud Monitoring tools.
  • Python and Shell scripting expertise.
  • Terraform and CI/CD experience.
  • Strong Incident Management and RCA skills.

Good to Have

  • Banking/Financial Services experience.
  • Service Mesh (Istio) knowledge.
  • GCP/AWS/CKA/Terraform Certifications.
  • Understanding of SLI/SLO/SLA and SRE practices.

Screening Priority

  1. Cloud Networking & Troubleshooting
  2. Kubernetes Networking
  3. Incident Management & RCA
  4. Observability & Monitoring
  5. Python Automation
  6. Terraform & DevOps
  7. GCP/AWS Platform Engineering

Free. 20 seconds. No password. See every match in this search.

Create a free Caio profile to unlock more results and save your role and location preferences.

Unlock free search
Want help applying to roles like this? Search Caio for free. If repetitive applications get heavy, Managed Job Search adds supervised execution for $99/month.
View Managed Job Search