Replicant Logo

Replicant

Senior Site Reliability Engineer

Posted 3 Days Ago
Remote
Hiring Remotely in USA
Senior level
Remote
Hiring Remotely in USA
Senior level
Build and improve Replicant’s AI-native platform, including site reliability, CI/CD, developer tooling, observability, incident management, cloud infrastructure, and autonomous-agent harnesses. Own production infrastructure reliability at scale, reduce operational toil, improve deployment workflows, participate in on-call rotation, and shape platform engineering patterns. The role uses TypeScript/Node.js, Python, Terraform, Kubernetes, Helm, GCP, and modern monitoring tools.
The summary above was generated by AI

At Replicant, we believe AI should work for people, starting with customer service. That’s why we built a platform that helps contact centers resolve more requests, proactively identify issues, and improve agent performance with AI-powered conversation intelligence and AI agents that act like your best reps.

Our AI agents handle millions of calls every month for Fortune 500 companies and high-growth innovators. From processing payments to booking appointments and authenticating users, they help customers get what they need instantly, 24/7. Meanwhile, our real-time conversation insights help contact center leaders coach better and improve every interaction.

We are leading the shift from legacy systems to AI-first service, powered by large language models (LLMs) and designed for enterprise scale, security, and empathy. If you’re excited by the potential of LLMs, voice AI, and building category-defining technology with a kind, ambitious team, you’ll love it here.

Our SRE team builds - not just supports - an AI-native platform, and they own exciting domains like platform and harness engineering, site reliability, cloud infrastructure, CI/CD, DevEx, observability, incident management, and COGS (e.g. cloud spend visibility). We're looking for a Site Reliability Engineer who has opinions about how these domains should work and wants agency in shaping where they go. If you are energized by enabling teams to succeed through systems- and patterns-level work, come build Replicant’s platform with us!

What You’ll Do

  • Contribute to patterns, design, and implementation of our domains; help shape the future of platform engineering at Replicant.

  • Build and improve systems that help reduce toil and enable Replicant's production infrastructure to remain available and operable under large-scale, real-time conversational AI traffic.

  • Extend and iterate our agent harness: Unsupervised AI agents are currently used by about 10% of the dev team - help us grow that number. The agent harness includes CI, sandboxes, guardrails, and validation (e.g. agent-first eval loops).

  • Own and improve our CI/CD pipelines and surrounding developer tooling: build and test performance, deployment ergonomics, and paved paths for new services.

  • Participate in on-call rotation and incident management to ensure platform uptime and quality. (SRE owns the base infrastructure, not the applications; non-business-hours pages are rare)

What You'll Bring

  • 6+ years’ experience in software development enablement roles.

  • Solid experience owning CI/CD platforms end to end - including domains like caching, architecture, and developer self-service.

  • Effective use of AI tools such as Claude and Cursor for coding, troubleshooting, and reasoning. You pair these skills with a defensible opinion on where to avoid using AI tools.

  • Familiarity with Node/TypeScript including making code changes (e.g. exposing new metrics), Python and Terraform for automation, and developing in a Kubernetes/Helm ecosystem.

  • Practical experience with observability: logs/metrics/tracing, monitoring/alerting, incident management process, and tooling.

  • Experience working in fully remote teams - tell us how you’ve made one work better.

  • Bonus:

    • Harness engineering experience - building platforms for autonomous agents.

    • Production-at-scale experience with GCP.

    • Telephony and SIP architectures, FreeSWITCH in particular.

Our stack is TypeScript/Node and Python running on Kubernetes - primarily on GCP (we are multi-cloud), with GitLab CI, Helm, Terraform, Datadog, Prometheus, and Grafana.

For all full-time employees, we offer:

🌴 In-person connection that counts: company-wide offsites and smaller team gatherings designed to make remote work feel personal

🖥️ Tech & learning stipend: Conferences, books, courses — interested? We’ll fund them

📍 Remote by design: We’re distributed — no guilt about life events, we trust you to manage your calendar

🏋️ Health & wellness: Flexible vacations, paid sabbatical after 5 years, comprehensive benefits, plus a stipend to support your physical and mental well-being

💸 Compensation that matches your impact: competitive salaries in the company you’re helping to build

📈 Equity with upside: We believe in shared ownership—You’ll own a real piece of a fast-growing AI company

Our Values

Replicant has three core values. It is critical that everyone who joins the team feels excited and moved by these values as every new team member makes an impact on our culture.

Blade Runners: We take ownership and pride to influence the outcomes of our goals. We are successful, and like a Blade Runner, use the tools at our disposal to reach our objectives. We value open and honest communication and proactively seek feedback along the way. We are a company driven to grow and achieve both individually and as a team.

Bread Makers: We are humble and strive toward an egalitarian culture. No task is too big or too small. We work together to achieve our goals and develop our company mission. We believe that the whole is greater than the sum of its parts in everything that we do.

Självdistans (Self-Distance): Självdistans is Swedish for self-distance. It's the ability to critically reflect on oneself and one's relations from an external perspective. With this in mind, we act with objectivity and always remember that we are not our work. There's no perfect science to growing a team or business, but we trust everyone at Replicant to point out our blind spots and humbly admit their own.

Replicant is proud to be an equal opportunity employer. We are committed to fostering an inclusive, diverse and equitable workplace that is built on trust, support and respect. We welcome all individuals and do not discriminate on the basis of gender identity and expression, race, ethnicity, disability, sexual orientation, colour, religion, creed, gender, national origin, age, marital status, pregnancy, sex, citizenship, education, languages spoken or veteran status. Accommodation is available upon request at any point during our recruitment process. If you require an accommodation, please speak to your talent acquisition partner or email us at [email protected] and we’ll work to meet your needs.

Similar Jobs

10 Days Ago
Easy Apply
Remote
USA
Easy Apply
191K-226K Annually
Senior level
191K-226K Annually
Senior level
Big Data • Healthtech • HR Tech • Machine Learning • Software • Telehealth • Big Data Analytics
Own the reliability, performance, resilience, observability, and security of AWS and Kubernetes infrastructure supporting products and AI/ML workloads. Define SLOs, lead incident response and root-cause analysis, build Terraform automation, optimize cloud costs, reduce operational toil, and establish deployment standards that help engineers ship reliably. Participate in on-call rotations and maintain HIPAA-compliant infrastructure.
Top Skills: AWSClaudeDatadogGitlabGoHipaaIstioKubernetesNatsPostgresPythonSoc 2TerraformTypescript
10 Days Ago
In-Office or Remote
92K-164K Annually
Senior level
92K-164K Annually
Senior level
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Architects, develops, operates, and maintains secure, resilient cloud infrastructure across Azure and AWS commercial and government environments. Builds Kubernetes platforms, infrastructure-as-code, observability tooling, monitoring, alerting, automation, and self-healing capabilities. Partners with engineering teams on SLOs, SLAs, deployments, incident response, performance testing, cost controls, security, and compliance. Participates in a 24/7 on-call rotation and improves operational processes and tooling.
Top Skills: ArgocdAWSAzureAzure MonitorDynatraceEncryptionFluxGitGitlabGraphanaHelmIaasIamKubernetesOwaspPaasPkiPrometheusPulumiRestful ServicesSplunkTerraformVisual Studio Code
10 Days Ago
In-Office or Remote
92K-164K Annually
Senior level
92K-164K Annually
Senior level
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Architects, builds, operates, and improves Azure commercial and government cloud infrastructure. Develops platform services, Kubernetes environments, infrastructure automation, observability, monitoring, alerting, and resiliency frameworks. Partners with engineering teams to define reliability objectives, supports production systems through an on-call rotation, and resolves root causes using automation and self-healing. The role requires strong cloud, Kubernetes, infrastructure-as-code, security, networking, and production operations expertise.
Top Skills: ApmArgocdAzureAzure Government CloudAzure MonitorDynatraceEncryptionFluxGitGrafanaHelmIaasIamInfrastructure As CodeInternet ProtocolsKubernetesNetworkingOwaspPaasPkiPrometheusPulumiRestful ServicesSlasSlisSlosSplunkTerraformVisual Studio Code

What you need to know about the Chicago Tech Scene

With vibrant neighborhoods, great food and more affordable housing than either coast, Chicago might be the most liveable major tech hub. It is the birthplace of modern commodities and futures trading, a national hub for logistics and commerce, and home to the American Medical Association and the American Bar Association. This diverse blend of industry influences has helped Chicago emerge as a major player in verticals like fintech, biotechnology, legal tech, e-commerce and logistics technology. It’s also a major hiring center for tech companies on both coasts.

Key Facts About Chicago Tech

  • Number of Tech Workers: 245,800; 5.2% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: McDonald’s, John Deere, Boeing, Morningstar
  • Key Industries: Artificial intelligence, biotechnology, fintech, software, logistics technology
  • Funding Landscape: $2.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Pritzker Group Venture Capital, Arch Venture Partners, MATH Venture Partners, Jump Capital, Hyde Park Venture Partners
  • Research Centers and Universities: Northwestern University, University of Chicago, University of Illinois Urbana-Champaign, Illinois Institute of Technology, Argonne National Laboratory, Fermi National Accelerator Laboratory

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account