Senior Site Reliability Engineer

Talkiatry

Completely RemoteFull TimeInformation Technology
Posted 3 days ago

Job description

About the Company

Talkiatry is transforming mental health care by making high-quality psychiatry more accessible, human, and sustainable for both patients and clinicians. As the largest dedicated psychiatry practice in the U.S., we use a technology-powered platform to empower clinicians and improve patient experiences.

Responsibilities

  • Define and roll out SRE practices including SLOs/SLIs, error budgets, and reliability standards
  • Build and improve observability through metrics, logging, distributed tracing, dashboards, and alerting
  • Drive down outage frequency by surfacing systemic reliability risks and partnering with teams for remediation
  • Reduce toil through automation, infrastructure-as-code, and self-service tooling
  • Own the health and usability of observability tooling, including documentation and training
  • Run production readiness reviews and partner with leadership on capacity planning

Requirements

  • 7+ years in software or infrastructure engineering
  • Substantial hands-on SRE or production reliability experience
  • Proven track record of reducing incidents and improving detection
  • Experience defining SLOs/SLIs and using error budgets
  • Deep observability expertise (e.g., Datadog, Prometheus, Grafana)
  • Strong experience operating production systems on AWS
  • Proficiency with infrastructure-as-code (e.g., Terraform)
  • Ability to build automation and tooling using Python, TypeScript, or similar
  • Excellent communication and influence skills

Preferred Qualifications

  • Experience standing up an SRE function for the first time at a startup or scale-up
  • Familiarity with TypeScript/Node.js, React, AWS EKS, and RDS
  • Experience with Kubernetes or container orchestration
  • Background in healthcare, tele-health, or regulated environments (HIPAA)

Skills & tools

SREAWSTerraform

What the team is looking for

Use this list as a quick fit check before you apply.

  1. 017+ years software/infrastructure engineering
  2. 02Hands-on SRE experience
  3. 03SLO/SLI expertise
  4. 04Observability expertise
  5. 05AWS experience
  6. 06Infrastructure-as-code proficiency
  7. 07Python or TypeScript
NeverApplyAd

Wake up to a shortlist, not a search results page.

NeverApply scores every new listing against your CV, salary floor and visa. A handful of real matches by morning.

Get your daily matches