Improve AWS production infrastructure reliability, observability, performance, and operational maturity. Build Terraform infrastructure, enhance CI/CD, automate operational work, manage incident response and on-call operations, lead postmortems, improve application resilience, support capacity planning and database reliability, and collaborate on security hardening and compliance. Mentor engineers and promote reliability practices across the organization.
Senior Site Reliability Engineer
Improve the reliability, performance, and operational maturity of a platform that supports the future of educational fundraising.
CONTRACT-TO-HIRE REMOTE - UNITED STATES SEATTLE / WEST COAST PREFERRED
About the role
Our client is looking for a hands-on Senior Site Reliability Engineer to improve the reliability, performance, and operational maturity of our platform. With our migration to AWS complete, this role will focus on strengthening our production environment: improving observability, automating infrastructure and operational work, enhancing incident response, and partnering with product engineers to build resilient systems. This is a high-impact role for someone who understands both infrastructure and application development. You will work across the stack, contribute code when appropriate, and help ensure our systems remain secure, scalable, and dependable as we grow.
What you'll do
• Operate, maintain, and improve our production infrastructure in AWS.
• Build and maintain infrastructure as code using Terraform.
• Improve monitoring, alerting, dashboards, and service-level indicators using New Relic or comparable observability platforms.
• Reduce alert noise and build systems that identify problems before customers are affected.
• Participate in the 24/7 on-call rotation and help coordinate the response to production incidents.
• Lead blameless postmortems and ensure corrective actions result in durable improvements.
• Partner with product engineers to diagnose performance and reliability issues throughout the application stack.
• Improve application resilience through appropriate use of timeouts, retries, queuing, backpressure, and idempotency.
• Improve CI/CD pipelines and deployment practices using platforms such as GitHub Actions, GitLab CI, or CircleCI.
• Automate repetitive operational work and reduce engineering toil.
• Create and maintain runbooks, system diagrams, troubleshooting guides, and production documentation.
• Support capacity planning, performance testing, database reliability, and production-readiness reviews.
• Collaborate with Security and Engineering teams on infrastructure hardening, access controls, logging, and compliance-related operational practices.
• Mentor engineers and promote effective reliability practices across the Engineering organization
What we're looking for
• 10+ years of overall software engineering, infrastructure, or systems experience, including at least 5 years in an SRE, Platform Engineering, DevOps, or production operations role.
• Previous professional software development experience and the ability to read, debug, and contribute to application code. • Strong, hands-on experience operating production workloads in AWS. • Experience building and maintaining infrastructure with Terraform or a similar infrastructure-as-code tool.
• Strong observability skills using New Relic, Datadog, or another modern monitoring platform. • Experience with incident response, on-call operations, postmortems, and production troubleshooting.
• Experience building or maintaining CI/CD pipelines.
• Working knowledge of networking, Linux, distributed systems, and relational databases.
• Strong judgment when balancing immediate operational needs with long-term maintainability.
• Clear communication skills and the ability to collaborate effectively across engineering disciplines.
• A track record of using automation to improve reliability and create leverage for other engineers.
Bonus points
• Experience with Ruby or Ruby on Rails.
• Strong PostgreSQL administration or performance-tuning experience.
• Experience operating enterprise SaaS products at scale.
• Familiarity with SLOs, SLIs, error budgets, capacity modeling, and load testing.
• Experience with payments, fintech, or other highly regulated systems.
• Experience supporting SOC 2 or similar security and compliance programs. Role details
• Contract-to-hire.
• Remote within the United States.
• Seattle-area or West Coast candidates are preferred to support occasional in-person collaboration, but exceptional candidates elsewhere should also be considered.
• Participation in a shared on-call rotation is required.
Similar Jobs
Machine Learning • Payments • Security • Software • Financial Services
Designs, implements, and supports hybrid on-premises and cloud infrastructure for federal IT operations. Leads cloud migrations, manages virtualization and identity platforms, automates infrastructure, and ensures compliance with FedRAMP, NIST, and FISMA. Monitors performance, availability, capacity, backup, recovery, and security while supporting Zero Trust initiatives. Collaborates with technical teams, documents procedures, and provides ongoing customer infrastructure support.
Top Skills:
Active DirectoryAnsibleAws GovcloudAzure Active Directory (Entra Id)FedrampFismaHyper-VLinuxMicrosoft Azure GovernmentNistPowershellTerraformVMwareWindows ServerZero Trust
Aerospace • Information Technology • Software • Cybersecurity • Design • Defense • Manufacturing
Provides sustaining process engineering support for semiconductor solar cell wafer and assembly production. Maintains process control using SPC, troubleshoots equipment and manufacturing anomalies, and drives improvements to reduce cycle time and cost while improving product performance. Supports solar technology R&D, process and product development, material characterization, safety initiatives, and cross-functional Operations objectives. This is a 100% onsite, second-shift role requiring U.S. Person status and the ability to obtain a Secret clearance.
Top Skills:
CC++Dry EtchingHigh-Temperature AnnealingLabviewPhotolithographySemiconductor ProcessingSQLStatistical Process Control (Spc)Thin-Film DepositionVBAWet Etching
Artificial Intelligence • Big Data • Cloud • Information Technology • Software • Big Data Analytics • Automation
Manage enterprise adoption, expansion, and revenue growth of the Dynatrace platform within assigned federal civilian accounts, particularly DHS. Build strategic account plans, map stakeholders, develop C-level relationships, manage complex sales cycles, and execute land-and-expand strategies. Coordinate global sales, support, services, partner, and executive resources while driving long-term cloud and corporate strategies. The role requires enterprise software sales experience, Federal account expertise, MEDDIC knowledge, and up to 30% regional travel.
Top Skills:
Cloud ComputingDynatraceObservability
What you need to know about the Chicago Tech Scene
With vibrant neighborhoods, great food and more affordable housing than either coast, Chicago might be the most liveable major tech hub. It is the birthplace of modern commodities and futures trading, a national hub for logistics and commerce, and home to the American Medical Association and the American Bar Association. This diverse blend of industry influences has helped Chicago emerge as a major player in verticals like fintech, biotechnology, legal tech, e-commerce and logistics technology. It’s also a major hiring center for tech companies on both coasts.
Key Facts About Chicago Tech
- Number of Tech Workers: 245,800; 5.2% of overall workforce (2024 CompTIA survey)
- Major Tech Employers: McDonald’s, John Deere, Boeing, Morningstar
- Key Industries: Artificial intelligence, biotechnology, fintech, software, logistics technology
- Funding Landscape: $2.5 billion in venture capital funding in 2024 (Pitchbook)
- Notable Investors: Pritzker Group Venture Capital, Arch Venture Partners, MATH Venture Partners, Jump Capital, Hyde Park Venture Partners
- Research Centers and Universities: Northwestern University, University of Chicago, University of Illinois Urbana-Champaign, Illinois Institute of Technology, Argonne National Laboratory, Fermi National Accelerator Laboratory



