NOCD Logo

NOCD

Senior Site Reliability Engineer

Posted One Month Ago
In-Office
Chicago, IL, USA
150K-200K Annually
Senior level
In-Office
Chicago, IL, USA
150K-200K Annually
Senior level
Lead platform and infrastructure engineering: build full-stack services, design AWS/Terraform infrastructure, manage Kubernetes/Docker, implement CI/CD, improve reliability and security, mentor engineers, and drive compliance for HIPAA and SOC 2.
The summary above was generated by AI

About NOCD

NOCD is the #1 telehealth provider for the treatment of obsessive-compulsive disorder (OCD). OCD is one of the most severe, prevalent, and misunderstood mental health conditions. NOCD creates access to online therapy for people with OCD through our telehealth platform. In the NOCD app, Members can quickly access and schedule live, face-to-face video therapy sessions with our national network of licensed Therapists that specialize in Exposure and Response Prevention Therapy (ERP) - considered the "gold standard" in OCD treatment. 
At NOCD, we help people reclaim their lives with clinically proven OCD treatment, by removing barriers to OCD care, and reducing the stigma associated with OCD. We’re changing the world and need other like-minded individuals to accelerate and expand our efforts.


Chicago, IL (Hybrid 3X a week)

Senior Site Reliability Engineer (SRE)

Chicago, IL (Hybrid)

Opportunity Overview

NOCD is looking for a Senior Site Reliability Engineer (SRE) to help build and scale the technology infrastructure powering our digital behavioral health platform.

This is a highly hands-on engineering role at the intersection of software engineering, cloud infrastructure, platform engineering, reliability, and security. You’ll help design and operate systems that are secure, scalable, observable, and resilient while enabling our engineering teams to ship software faster and more reliably.

You’ll have significant ownership in shaping our architecture, engineering standards, infrastructure, and developer experience as NOCD continues to scale.

What You'll Do

Software Engineering & Systems

  • Design, build, and maintain reliable production systems and services.
  • Write production-quality code in Python, TypeScript, or similar languages.
  • Develop and maintain APIs, microservices, automation, and internal engineering tools.
  • Contribute to architecture decisions, technical design reviews, code reviews, and engineering standards.
  • Apply software engineering principles to infrastructure, automation, and reliability challenges.

Cloud & Platform Engineering

  • Design, build, and operate AWS infrastructure supporting production applications.
  • Manage infrastructure as code using Terraform.
  • Build and maintain containerized environments using Docker and Kubernetes.
  • Design and improve CI/CD pipelines, deployment automation, and release processes.
  • Build internal tooling and platform capabilities that improve developer productivity.
  • Help establish scalable infrastructure patterns that can support continued company growth.

Reliability & Observability

  • Own and improve the reliability, availability, performance, and scalability of production systems.
  • Develop monitoring, alerting, logging, and observability strategies across our infrastructure and applications.
  • Lead incident response, troubleshooting, and root-cause analysis for production issues.
  • Establish and improve operational practices around incident management, postmortems, and preventative remediation.
  • Identify reliability risks and proactively improve system resilience, capacity, and performance.
  • Help define and monitor appropriate SLIs, SLOs, and operational metrics.

Security & Compliance

  • Implement cloud and application security best practices across infrastructure and production systems.
  • Partner with Security and Engineering to support HIPAA, SOC 2, and other compliance requirements.
  • Implement appropriate controls around access management, secrets, encryption, logging, and infrastructure security.
  • Help identify and remediate infrastructure and application security risks.

Technical Leadership

  • Own technical initiatives from design through production.
  • Partner closely with Software Engineering, Product, Security, and other teams to solve complex technical problems.
  • Mentor engineers and contribute to a strong engineering culture.
  • Help establish engineering and operational best practices as the company scales.
  • Balance reliability, security, engineering velocity, and business priorities when making technical decisions.
Required Qualifications
  • 7+ years of professional software engineering experience, with significant experience in SRE, platform engineering, DevOps, or cloud infrastructure.
  • Bachelor’s degree in Computer Science, Computer Engineering, Software Engineering, or a related technical field, or equivalent professional experience.
  • Strong software engineering fundamentals and experience writing production-quality code.
  • Strong hands-on experience with AWS and cloud architecture.
  • Strong experience with Terraform or other infrastructure-as-code tools.
  • Strong experience with Docker and Kubernetes.
  • Proficiency in Python, TypeScript, or a similar programming language.
  • Experience designing and operating CI/CD pipelines and deployment automation.
  • Strong understanding of distributed systems, APIs, networking, databases, and cloud architecture.
  • Experience with monitoring, logging, observability, and production troubleshooting.
  • Experience participating in or leading incident response and root-cause analysis.
  • Strong understanding of software reliability, scalability, availability, and performance.
  • Demonstrated ability to own technical initiatives and work effectively across engineering teams.
Preferred Qualifications
  • Experience working in healthcare, fintech, or another regulated environment.
  • Experience supporting HIPAA, SOC 2, HITRUST, or similar compliance frameworks.
  • Experience with Datadog, CloudWatch, Prometheus, Grafana, Splunk, or similar observability platforms.
  • Experience with GitHub Actions, Jenkins, ArgoCD, or GitOps.
  • Experience designing highly available or multi-region AWS architectures.
  • Experience with microservices and event-driven architectures.
  • Experience with infrastructure security, DevSecOps, IAM, secrets management, and encryption.
  • Experience establishing SLIs, SLOs, SLAs, and error budgets.
  • Experience building internal developer platforms or developer tooling.
  • Experience with disaster recovery, capacity planning, and performance engineering.
  • Experience working in a high-growth startup environment.
What We Offer
  • Comprehensive benefits package, including medical, dental, vision coverage, and 401(k) match
  • 11 observed company holidays a year
  • PTO based on an accrual system
  • Engaging startup environment with an outstanding mission-driven team atmosphere
  • Downtown Chicago office with an on-site gym
  • Noto provides 12 weeks of fully paid parental leave for the primary caregiver, and 6 weeks of fully paid leave for the secondary caregiver, for qualifying full-time employees
Pay Transparency

The expected pay range for this position is $150,000 to $200,000. Actual pay will be based on the individual’s qualifications and experience. This role is also eligible for annual performance-based incentives tied to individual achievement and company-wide goals.

HQ

NOCD Chicago, Illinois, USA Office

225 N. Michigan Ave, Ste 1430, Chicago, IL, United States, 60611

Similar Jobs

17 Days Ago
Remote or Hybrid
United States
Senior level
Senior level
Fintech • Software
The Senior Site Reliability Engineer ensures SaaS platforms remain reliable, performant, secure, and scalable. Responsibilities include building cloud infrastructure, implementing monitoring and alerting, automating operational runbooks and deployments, managing Infrastructure as Code, applying AI-powered observability and remediation, supporting Kubernetes and cloud networking, and leading incident triage and root-cause analysis during 24/7 on-call rotations.
Top Skills: AIAiopsAksAnsibleAppdynamicsAWSAzureAzure DevopsBashC# .NetCi/CdCloud NetworkingCloudopsCosmos DbDatadogDynatraceEksFirewallsHarnessIdera Sql Diagnostic ManagerInfrastructure As CodeJavaJenkinsKubernetesLinuxLoad BalancingNew RelicPowershellPythonRedgate Sql MonitorSolarwinds Database Performance AnalyzerSQLTerraformWindows
18 Days Ago
Easy Apply
Remote or Hybrid
USA
Easy Apply
180K-220K Annually
Senior level
180K-220K Annually
Senior level
Healthtech • Information Technology • Software • Telehealth
Develop, monitor, and maintain distributed production systems and AWS-based microservices infrastructure. Build automation, tooling, and repeatable processes that improve uptime, scalability, security, and operational efficiency. Support product engineering teams with performance, scaling, incident diagnosis, and production debugging. Analyze and tune systems, code, and networking while participating in on-call operations and blameless post-mortems.
Top Skills: AWSDnsDockerGCPGenaiHttp/HttpsKubernetesLoad BalancersNtpReverse ProxiesTcp/IpTlsWeb Application Firewalls
25 Days Ago
Easy Apply
Hybrid
Chicago, IL, USA
Easy Apply
180K-200K Annually
Senior level
180K-200K Annually
Senior level
Fintech • News + Entertainment • Software • Financial Services
Build fault-tolerant, self-healing infrastructure and internal tooling for a brokerage platform. Expand observability using Prometheus, Honeycomb, OpenTelemetry, and Grafana; scale HashiCorp Nomad services through capacity planning and load testing; establish SLOs, error budgets, and alerting; analyze Linux and networking performance; support production on-call operations; and mentor engineers in site reliability practices.
Top Skills: ConsulElixirGrafanaHashicorp NomadHoneycombJavaLinuxMulticastOpentelemetryPrometheusPythonRubyTcp/IpUdpVault

What you need to know about the Chicago Tech Scene

With vibrant neighborhoods, great food and more affordable housing than either coast, Chicago might be the most liveable major tech hub. It is the birthplace of modern commodities and futures trading, a national hub for logistics and commerce, and home to the American Medical Association and the American Bar Association. This diverse blend of industry influences has helped Chicago emerge as a major player in verticals like fintech, biotechnology, legal tech, e-commerce and logistics technology. It’s also a major hiring center for tech companies on both coasts.

Key Facts About Chicago Tech

  • Number of Tech Workers: 245,800; 5.2% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: McDonald’s, John Deere, Boeing, Morningstar
  • Key Industries: Artificial intelligence, biotechnology, fintech, software, logistics technology
  • Funding Landscape: $2.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Pritzker Group Venture Capital, Arch Venture Partners, MATH Venture Partners, Jump Capital, Hyde Park Venture Partners
  • Research Centers and Universities: Northwestern University, University of Chicago, University of Illinois Urbana-Champaign, Illinois Institute of Technology, Argonne National Laboratory, Fermi National Accelerator Laboratory

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account