Back to search
Selby Jennings Linkedin · Posted 19d ago

Site Reliability Engineer

New York, United States

Linkedin
Continue to application Add your email once, then Caio opens the original posting.

Indexed description

Join a highly technical environment where you'll play a key role in maintaining and enhancing the stability of critical trading infrastructure. This position is ideal for someone who enjoys diagnosing complex issues, working closely with end users, and continuously improving system performance in a fast-paced setting. You'll partner with a variety of teams across technology and business functions while gaining exposure to sophisticated infrastructure and low-latency systems.

Responsibilities

  • Support and enhance the reliability, availability, and performance of business-critical trading platforms.
  • Investigate production incidents and drive resolution across infrastructure, applications, and workflow processes.
  • Partner with trading, quantitative, development, and operational teams to improve system efficiency and user experience.
  • Serve as a primary point of escalation for complex production issues and drive root-cause analysis efforts.
  • Monitor distributed Linux-based environments, identifying risks and performance bottlenecks before they impact users.
  • Develop and maintain automation, operational tooling, deployment workflows, monitoring solutions, and platform management processes.
  • Participate in operational support responsibilities and contribute to ongoing service improvements.
  • Work closely with engineering teams to implement, test, and deploy technology enhancements.
  • Continuously evaluate existing processes and recommend improvements that increase stability, scalability, and operational effectiveness.

Qualifications

  • 3+ years of professional experience supporting production environments, platform engineering, site reliability, DevOps, or infrastructure operations.
  • Strong scripting and automation experience using Python and Linux shell scripting.
  • Deep understanding of Linux systems administration, performance analysis, and troubleshooting.
  • Experience supporting business-critical applications in a high-availability environment.
  • Strong analytical and problem-solving skills with the ability to investigate issues across multiple technology layers.
  • Exposure to software development concepts and compiled languages is beneficial.
  • Ability to work effectively with both technical and non-technical stakeholders.
  • Strong sense of accountability, urgency, and ownership.
  • Comfortable managing multiple priorities in a dynamic, time-sensitive environment.
  • Bachelor's degree in Computer Science, Engineering, a related discipline, or equivalent practical experience.
  • Curious mindset with a passion for learning new technologies and improving operational processes.

Free. 20 seconds. No password. See every match in this search.

Create a free Caio profile to unlock more results and save your role and location preferences.

Unlock free search
Want help applying to roles like this? Search Caio for free. If repetitive applications get heavy, Managed Job Search adds supervised execution for $99/month.
View Managed Job Search