Texas Sports Academy Logo

Texas Sports Academy

Senior Site Reliability Engineer

Posted An Hour Ago
Remote
Hiring Remotely in United States
Senior level
Remote
Hiring Remotely in United States
Senior level
Audit infrastructure, deployment pipelines, monitoring, alerting, incident response, on-call practices, and internal tools. Produce actionable audit reports, implement code and configuration fixes, improve SLOs and reliability practices, advise on scalable architecture, and partner with engineers on implementation and handoff. The role is a fully remote, part-time consulting engagement with potential for full-time conversion.
The summary above was generated by AI

Texas Sports Academy is a K-12 school designed for serious student-athletes who want both elite academics and high-level athletic development. Students learn twice as fast as traditional schools in just 2 hours a day, using the same 2-Hour Learning model as Alpha School. We are an AI-first school where AI is woven into how students learn, how the team works, and how we build and scale everything we do. That frees up their entire afternoon for serious training, where they work alongside former pro and D1 athletes coaching them at the highest standard.

We are hiring a Senior Site Reliability Engineer on a part-time consulting contract to audit our systems and make the changes required to keep them running as we scale. You will look at everything: infrastructure, deployment, monitoring, incident response, on-call, and how our internal tools actually behave under load. You will document what you find in clean, direct written audits, and you will ship the code and configuration changes needed to fix what is broken.

What You'll Do
  • Audit Our Systems End to End: Look at our infrastructure, deployment pipelines, monitoring, alerting, and incident response. Find where we are brittle, where we will break at the next scale step, and where we are wasting money.
  • Write Clear Audit Reports: Document what you found, why it matters, and how to fix it. Written work that our engineering team can act on directly.
  • Ship the Fixes: Do not stop at recommendations. Write the code, ship the config changes, and stand up the monitoring, alerting, and deployment improvements yourself where it makes sense.
  • Level Up Our Reliability Practices: Help us mature how we handle on-call, incidents, SLOs, and postmortems. Bring in the standards you know work at scale.
  • Advise on Architecture for Scale: Look ahead at where we are going and flag the infrastructure decisions we need to make now to be ready.
  • Partner With Our Engineering Team: Work directly with our engineers. Pair, review, and hand off cleanly so what you build actually sticks.

Requirements
  • Senior-Level SRE Experience: Eight or more years of hands-on site reliability, infrastructure, or production engineering work. You have owned real systems at real scale.
  • Deep Cloud Infrastructure Expertise: You are fluent in AWS at a serious level. You know networking, IAM, VPCs, and the failure modes cold.
  • Monitoring and Observability Mastery: You have built monitoring, alerting, and observability from the ground up with tools like Datadog, Grafana, Prometheus, or comparable. You know what a good alert looks like.
  • CI/CD and Deployment Pipelines: You have built and maintained real deployment pipelines and know how to make them fast, safe, and reversible.
  • Ships Code: You are a real engineer, not just an advisor. You are comfortable writing production code and configuration and taking responsibility for it.
  • AI-First Mindset: You use AI daily as part of your work flows. You are comfortable using AI coding tools to move faster and review your own work.
  • Strong Written Communication: You can write an audit report that an engineering team acts on directly.
  • Location and Setup: Fully remote, open globally. Reliable internet and a quiet space you can work 

from.

Bonus Points
  • On-Call and Incident Response Leadership: You have led on-call rotations and run real postmortems on real outages.
  • Cost Optimization Wins: You have real stories of cutting cloud spend meaningfully without breaking anything.
  • Security-Adjacent Work: You understand where reliability and security intersect and can flag issues on both sides.

BenefitsEngagement Details
  • Type: Part-time consulting contract
  • Duration: Ongoing, with the possibility of moving into a full-time role based on fit and results
  • Hours: Part-time to start, flexible based on scope
  • Location: Fully remote, anywhere in the world
Why Join Us?

You get to make the calls that keep the system running as we scale, and see your changes go live fast. Your written audits will be read, your code will ship, and the work you do will directly protect the day-to-day experience of real kids in real classrooms. If you are a senior SRE who wants a focused, high-signal engagement where your recommendations actually get implemented (by you), this is a good one to take.

Texas Sports Academy is an equal opportunity employer. We hire for character, capability, and mission alignment.

Similar Jobs

22 Days Ago
Remote or Hybrid
Senior level
Senior level
Digital Media • Gaming • Information Technology • Software • Sports • Esports • Big Data Analytics
Build, scale, and maintain cloud and on-premises infrastructure using IaC and configuration management. Implement observability, deployment platforms, and tooling; define infrastructure patterns and standards. Lead incident response, mentor engineers, and collaborate with Security and Networking to improve reliability and performance.
Top Skills: Amazon EksAnsibleApi GatewayCdnChefDnsGoKubernetesLoad BalancingPythonRancherReverse ProxyRubyTerraformVpc
19 Days Ago
In-Office or Remote
Senior level
Senior level
eCommerce • On-Demand • Software • Manufacturing
Lead architecture and operation of cloud infrastructure and Kubernetes (EKS). Drive large-scale automation with Terraform and GitOps, improve reliability and observability (Grafana/Prometheus/Loki/Tempo), participate in on-call incident response, enforce security and cost-optimization practices, and mentor mid-level SREs while partnering with product teams.
Top Skills: ArgocdAtlantisAuroraAWSCiliumEcrEksGitGithub ActionsGrafanaHelmHelm ChartIamImage ScanningJenkinsKubernetesLinuxLokiMimirMongoDBMySQLPostgresPrometheusPythonRdsRedisS3SqsTempoTerraformVpc
20 Days Ago
Remote
Senior level
Senior level
Information Technology • Software
Design, implement, and maintain stable, highly available infrastructure. Build and support CI/CD pipelines, deploy enterprise projects on AWS, automate build/release and monitoring, troubleshoot performance, participate in on-call rotation, and ensure security and compliance while collaborating across teams.
Top Skills: AWSBashDockerGitKubernetesLinuxTerraform

What you need to know about the Colorado Tech Scene

With a business-friendly climate and research universities like CU Boulder and Colorado State, Colorado has made a name for itself as a startup ecosystem. The state boasts a skilled workforce and high quality of life thanks to its affordable housing, vibrant cultural scene and unparalleled opportunities for outdoor recreation. Colorado is also home to the National Renewable Energy Laboratory, helping cement its status as a hub for renewable energy innovation.

Key Facts About Colorado Tech

  • Number of Tech Workers: 260,000; 8.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Lockheed Martin, Century Link, Comcast, BAE Systems, Level 3
  • Key Industries: Software, artificial intelligence, aerospace, e-commerce, fintech, healthtech
  • Funding Landscape: $4.9 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Access Venture Partners, Ridgeline Ventures, Techstars, Blackhorn Ventures
  • Research Centers and Universities: Colorado School of Mines, University of Colorado Boulder, University of Denver, Colorado State University, Mesa Laboratory, Space Science Institute, National Center for Atmospheric Research, National Renewable Energy Laboratory, Gottlieb Institute

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account