via Workable
$90K - 140K a year
Audit infrastructure and pipelines to resolve scalability bottlenecks and mature reliability practices including on-call rotations and incident response.
Eight or more years in site reliability or production engineering with deep AWS expertise and strong written communication skills.
Texas Sports Academy is a K-12 school designed for serious student-athletes who want both elite academics and high-level athletic development. Students learn twice as fast as traditional schools in just 2 hours a day, using the same 2-Hour Learning model as Alpha School. We are an AI-first school where AI is woven into how students learn, how the team works, and how we build and scale everything we do. That frees up their entire afternoon for serious training, where they work alongside former pro and D1 athletes coaching them at the highest standard. We are hiring a Senior Site Reliability Engineer on a part-time consulting contract to audit our systems and make the changes required to keep them running as we scale. You will look at everything: infrastructure, deployment, monitoring, incident response, on-call, and how our internal tools actually behave under load. You will document what you find in clean, direct written audits, and you will ship the code and configuration changes needed to fix what is broken. What You'll Do Audit Our Systems End to End: Look at our infrastructure, deployment pipelines, monitoring, alerting, and incident response. Find where we are brittle, where we will break at the next scale step, and where we are wasting money. Write Clear Audit Reports: Document what you found, why it matters, and how to fix it. Written work that our engineering team can act on directly. Ship the Fixes: Do not stop at recommendations. Write the code, ship the config changes, and stand up the monitoring, alerting, and deployment improvements yourself where it makes sense. Level Up Our Reliability Practices: Help us mature how we handle on-call, incidents, SLOs, and postmortems. Bring in the standards you know work at scale. Advise on Architecture for Scale: Look ahead at where we are going and flag the infrastructure decisions we need to make now to be ready. Partner With Our Engineering Team: Work directly with our engineers. Pair, review, and hand off cleanly so what you build actually sticks. Senior-Level SRE Experience: Eight or more years of hands-on site reliability, infrastructure, or production engineering work. You have owned real systems at real scale. Deep Cloud Infrastructure Expertise: You are fluent in AWS at a serious level. You know networking, IAM, VPCs, and the failure modes cold. Monitoring and Observability Mastery: You have built monitoring, alerting, and observability from the ground up with tools like Datadog, Grafana, Prometheus, or comparable. You know what a good alert looks like. CI/CD and Deployment Pipelines: You have built and maintained real deployment pipelines and know how to make them fast, safe, and reversible. Ships Code: You are a real engineer, not just an advisor. You are comfortable writing production code and configuration and taking responsibility for it. AI-First Mindset: You use AI daily as part of your work flows. You are comfortable using AI coding tools to move faster and review your own work. Strong Written Communication: You can write an audit report that an engineering team acts on directly. Location and Setup: Fully remote, open globally. Reliable internet and a quiet space you can work from. Bonus Points On-Call and Incident Response Leadership: You have led on-call rotations and run real postmortems on real outages. Cost Optimization Wins: You have real stories of cutting cloud spend meaningfully without breaking anything. Security-Adjacent Work: You understand where reliability and security intersect and can flag issues on both sides. Engagement Details Type: Part-time consulting contract Duration: Ongoing, with the possibility of moving into a full-time role based on fit and results Hours: Part-time to start, flexible based on scope Location: Fully remote, anywhere in the world Why Join Us? You get to make the calls that keep the system running as we scale, and see your changes go live fast. Your written audits will be read, your code will ship, and the work you do will directly protect the day-to-day experience of real kids in real classrooms. If you are a senior SRE who wants a focused, high-signal engagement where your recommendations actually get implemented (by you), this is a good one to take. Texas Sports Academy is an equal opportunity employer. We hire for character, capability, and mission alignment.
This job posting was last updated on 9/3/2026