Site Reliability Engineer
Indexed description
LearningSpring is The School Choice Management Platform that helps states, schools, families, SGOs, and providers run modern education freedom programs with clarity and accountability. It unifies school discovery, eligibility, applications, payments, and compliance, enabling partners to operate ESA, voucher, and tax credit scholarship programs at scale on a single platform.
LearningSpring's solutions, including the School Choice Marketplace, Education Freedom Wallet™, Federal Tax Credit Management, and complete School Choice Management for States, help families find their best-fit learning options, help schools and SGOs steward scholarships responsibly, and help states deliver transparent, auditable programs.
LearningSpring is looking for a Site Reliability Engineer who enjoys improving systems, automating repetitive work, and helping both engineers and employees be more effective.
This role is intentionally split between platform engineering (approximately 60%) and internal technology operations (approximately 40%). You'll help operate and improve our AWS-based platform while also owning the internal systems that keep our engineering organization productive. As LearningSpring grows, you'll have increasing opportunities to expand your platform engineering responsibilities.
You'll work closely with our Platform Engineering team to improve infrastructure, deployment reliability, monitoring, and operational tooling while owning identity management, employee onboarding, SaaS administration, and internal technical operations.
We don't believe operational work should be repetitive. If you find yourself performing the same task more than a few times, we expect you to ask how it can be automated. Success in this role comes from improving systems, reducing manual effort, and making both our engineering team and the rest of the company more effective.
Key Responsibilities
- Administer and maintain internal business systems including Google Workspace, Slack, GitHub, Linear, Jira, Atlassian, Zoom, Outline, Customer.io, HubSpot, and other SaaS platforms.
- Own identity and access management, including user provisioning, SSO, MFA, access reviews, and employee onboarding and offboarding.
- Automate repetitive operational and administrative tasks using Bash, Python, APIs, and workflow automation.
- Assist with security and compliance initiatives, including SOC 2 activities, Drata, access controls, endpoint security, and audit preparation.
- Support and improve our AWS platform, including ECS, RDS, Route 53, and related cloud infrastructure.
- Contribute infrastructure improvements using Terraform and GitHub Actions under the guidance of senior platform engineers.
- Build, maintain, and improve monitoring, alerting, dashboards, and operational tooling using AWS-native services and future third-party observability platforms as needed.
- Participate in incident response, troubleshooting, root cause analysis, and continuous operational improvement.
- Improve deployment reliability, CI/CD processes, and engineering workflows.
- Help optimize cloud infrastructure performance and cost.
- Create and maintain operational documentation, runbooks, and internal knowledge resources.
- Collaborate across Engineering and Operations to continuously improve reliability, security, and employee productivity.
Key Deliverables
- Accurate identity and access management across engineering and business platforms.
- Reliable, secure, and well-maintained cloud infrastructure and internal systems.
- Automated onboarding and offboarding processes that scale with company growth.
- Improved monitoring, alerting, and operational visibility.
- Measurable improvements in deployment reliability and operational efficiency.
- Reduced manual operational work through automation.
- Well-documented infrastructure, systems, and operational procedures.
- Continued improvement of LearningSpring's security and compliance posture.
Required Qualifications
- 2–5 years of experience in Site Reliability Engineering, Platform Engineering, Systems Administration, DevOps, or a similar operations-focused role.
- Proven working knowledge of Google Workspace
- Experience working with AWS cloud services in a production environment.
- Experience with Docker and containerized applications.
- Familiarity with Infrastructure as Code concepts, preferably Terraform.
- Experience using GitHub and GitHub Actions.
- Strong understanding of Linux system administration fundamentals, networking, DNS, TLS certificates, and cloud infrastructure concepts.
- Experience administering cloud-based productivity platforms such as Google Workspace and modern SaaS applications.
- Strong understanding of identity and access management, SSO, and MFA.
- Experience writing automation scripts using Bash, Python, or similar languages.
- Excellent troubleshooting, documentation, and communication skills.
- A collaborative, low-ego mindset with a willingness to learn, take ownership, and contribute wherever needed.
Preferred Qualifications
- Experience supporting SOC 2 or similar security and compliance frameworks.
- Experience supporting AWS ECS-based environments.
- Experience with AWS monitoring and observability tools such as CloudWatch.
- Familiarity with PagerDuty or similar incident management platforms.
- Experience integrating SaaS applications using APIs or workflow automation.
- Startup or high-growth technology company experience.
- Interest in growing into a mid-level platform or Site Reliability Engineering role.
Physical Requirements
- Sitting or standing: The ability to remain in a stationary position for extended periods.
- Using hands and fingers: The ability to operate standard office equipment and keyboards.
- Close visual acuity: The ability to see clearly at close distances, which is crucial for computer work.
Your Education
Where you went to school matters less than what you’ve learned since. We value curiosity, resilience, and self-awareness over pedigree. If you’re someone who pushes limits, seeks out complex problems, and knows when to ask for help, you’ll fit right in.
Please send your resume to [email protected] with the title of the role you are applying for in the subject line.
- Own employee onboarding, offboarding, and internal identity management with minimal oversight.
- Maintain secure, reliable internal technology systems with high availability.
- Improve monitoring, alerting, and operational visibility across our platform.
- Increase deployment reliability and reduce operational friction for engineering teams.
- Reduce manual operational work through automation and process improvements.
- Improve security posture through effective access management, documentation, and compliance support.
- Build confidence and technical ownership across both internal operations and cloud infrastructure.
- Progress toward independently owning increasingly complex platform engineering responsibilities.
- We value integrity: We lead with honesty, fairness, and transparency in every action and interaction. By treating each person with dignity, empathy, and respect, we strengthen the trust that underpins our relationships with families, educators, and community partners. Integrity guides both our decisions and how we serve our mission to support every learner’s growth.
- We have a growth mindset: We approach every challenge as an opportunity to innovate. Instead of asking why something can’t be done, we figure out how to make it happen: testing, learning, and iterating along the way.
- We are mission-driven: Everything we do helps more children access the right educational environment. Our work is essential, timely, and directly tied to helping students reach their potential.
- We act on evidence: We use data and measurable outcomes to guide our decisions and improve continuously. Every strategy and product we design reflects evidence-based practice and the ongoing pursuit of real-world impact.
- We are active problem solvers and innovators: We champion creative thinking and challenge convention to find new, effective solutions. Guided by curiosity and purpose, we use emerging tools, including AI and data analytics, to tailor experiences for parents and improve outcomes for students, adapting quickly to meet the evolving needs of families, schools, and states.
For individuals hired to work in Colorado, LearningSpring is required by law to include a reasonable estimate of the compensation range for this role. This compensation range is specific to the State of Colorado and includes the range of factors considered in making compensation decisions, including but not limited to skill sets, experience and training, certifications, etc.
Colorado Pay Range: The anticipated salary range for this role is $105,000–$130,000 annually.
Final compensation will be determined based on a variety of factors, including relevant experience, skills, education, and past performance. In addition to base salary, this position may also be eligible for a variable bonus and equity.
Our benefits include
- Competitive Medical, Dental, and Vision coverage
- A generous PTO policy
- A robust list of paid holidays
No visa sponsorship is available for this position.
Employment with LearningSpring is contingent upon the satisfactory completion of several pre-employment requirements, including a background check and a professional reference check.
LearningSpring, Inc. does not discriminate in employment opportunities or practices based on race, color, creed, sex, gender, gender identity or expression, pregnancy, childbirth or related medical conditions, religion, veteran and military status, marital status, registered domestic partner status, age, national origin or ancestry, physical or mental disability, medical condition (including genetic information or characteristics), sexual orientation, or any other characteristic protected by applicable federal, state, or local laws.
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search