Back to search
Xebia Linkedin · Posted 6d ago

SRE Infrastructure Lead

Louisville

Linkedin
Continue to application Add your email once, then Caio opens the original posting.

Indexed description

Must be local to Louisville, KY


Xebia is a pioneering Software Engineering and IT consultancy company, transforming and executing at the intersection of Domain and Technology to create digital leaders for our people, clients, partners, and communities.

With over 20 years of experience, we help the world’s top 250 companies and category leaders overcome digital challenges, embrace innovation, adopt new technology, and implement new business models. In addition to high-quality consulting, we also provide offshoring and nearshoring services.

We are organized in complementary chapters – teams with a tremendous amount of knowledge and experience within a particular field, such as Agile, DevOps, Data and AI, Cloud, Software Technology, Low Code, and Microsoft Solutions.

Our mission can be captured in one word: Authority which means being a recognized leader in the service lines we serve. As professionals, as a global company, and as a digital pioneer.

The only way we can achieve this is by constantly reinventing ourselves, hiring exceptional talent, claiming the position of thought leader, and sharing our knowledge. That’s what got us to where we are today, and we’re incredibly proud of that.

Xebia is always on the lookout for extraordinary people. We believe that quality without compromise starts with our people. We care about mutual respect, appreciating cultural differences, and complementing, strengthening, and motivating each other to become an authority in your field.

Supporting the passion of exceptionally smart people is the foundation of our organization. Everything we do – our mission, values, strategic innovation, hiring process, and knowledge management – is anchored on this principle. At Xebia, we learn, grow, explore and create new frontiers in business together. For more details please visit www.xebia.com


About the job:-

Overview

We are seeking an experienced Infrastructure SRE Lead to lead the transition of infrastructure operations, establish dedicated operational support, and enhance the reliability, scalability, and security of a cloud-based platform. This role requires strong expertise in infrastructure operations, cloud environments, automation, observability, incident management, and cross-functional collaboration.

Key Responsibilities

  • Conduct a comprehensive assessment of the current infrastructure landscape, documenting operational processes, support responsibilities, and platform dependencies across cloud, edge, network, device management, observability, deployments, and incident response.
  • Analyze existing environments and design the separation of infrastructure, including dedicated tenants, environments, access controls, monitoring, routing, security boundaries, and operational processes where required.
  • Partner with existing infrastructure teams to gain a deep understanding of day-to-day operations through knowledge transfer sessions and operational shadowing, including incident management, release processes, escalation procedures, and support workflows.
  • Lead the transition of infrastructure responsibilities to a dedicated support team by establishing ownership models, operational runbooks, access management, service expectations, escalation paths, and handoff procedures.
  • Provide ongoing production infrastructure support, ensuring high availability, performance, and operational excellence.
  • Continuously improve platform reliability, observability, automation, security, scalability, and infrastructure maturity.
  • Collaborate with engineering, security, DevOps, and operations teams to drive platform improvements and operational efficiencies.
  • Participate in incident response, root cause analysis, and implementation of preventive measures.
  • Develop and maintain infrastructure documentation, operational standards, and best practices.

Required Qualifications

  • Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field.
  • 7+ years of experience in infrastructure engineering, platform operations, or site reliability engineering.
  • Strong experience supporting cloud-based infrastructure (AWS, Azure, or GCP).
  • Experience with networking, infrastructure operations, monitoring, observability, and production support.
  • Hands-on experience with Infrastructure as Code (Terraform, CloudFormation, or similar).
  • Experience with CI/CD pipelines and deployment automation.
  • Strong understanding of identity and access management, security controls, and operational governance.
  • Experience with incident management, troubleshooting, root cause analysis, and operational support.
  • Excellent documentation, communication, and stakeholder management skills.
  • Ability to work in cross-functional teams and manage infrastructure transition initiatives.

Preferred Qualifications

  • Experience with Kubernetes, Docker, or containerized platforms.
  • Familiarity with monitoring and observability tools such as Datadog, Grafana, Prometheus, Splunk, or New Relic.
  • Experience with scripting or automation using Python, Bash, or PowerShell.
  • Knowledge of ITIL practices and service management.
  • Experience leading infrastructure transformation or operational transition projects.


Free. 20 seconds. No password. See every match in this search.

Create a free Caio profile to unlock more results and save your role and location preferences.

Unlock free search
Want help applying to roles like this? Search Caio for free. If repetitive applications get heavy, Managed Job Search adds supervised execution for $99/month.
View Managed Job Search