Antares Capital LP Logo

Antares Capital LP

Vice President, Reliability Engineering & Technology Operations

Posted 25 Days Ago
Be an Early Applicant
In-Office
Chicago, IL, USA
175K-225K Annually
Senior level
In-Office
Chicago, IL, USA
175K-225K Annually
Senior level
Leads reliability engineering and technology operations, overseeing production operations, Azure and Kubernetes platforms, observability, incident response, automation, resiliency, disaster recovery, vendors, and operational transformation. Establishes reliability practices, SLOs, error budgets, service ownership, and performance metrics while leading engineering-minded teams and partnering with business and technology stakeholders. Drives AI-enabled operations, platform resilience, continuous improvement, and executive incident management.
The summary above was generated by AI
About Antares Capital

Antares Capital is a leading alternative credit manager and a trusted financing partner to private equity sponsors and middle-market companies. We are committed to building resilient, scalable, and modern technology platforms that support our business and clients.

As part of our continued technology transformation, we are seeking a Vice President, Reliability Engineering & Technology Operations to lead the evolution of our production operations, reliability engineering, observability, and operational automation capabilities.

This is a strategic leadership role for an engineering-minded leader who thrives at the intersection of software engineering, cloud infrastructure, platform operations, and operational excellence. The ideal candidate combines strong technical depth with exceptional execution skills and has experience building highly reliable systems while leading teams through modernization and transformation initiatives.

The Opportunity

Technology is central to Antares' growth strategy. We are investing heavily in cloud platforms, engineering excellence, AI-enabled workflows, automation, and modern operational practices.

As the leader of Reliability Engineering & Technology Operations, you will be responsible for the availability, performance, scalability, and resilience of critical business platforms. You will partner closely with Engineering, Infrastructure, Cybersecurity, Data, and Business stakeholders to ensure our systems remain secure, observable, scalable, and operationally mature.

You will help shape the future of technology operations by introducing reliability engineering practices, expanding observability, leveraging AI-driven operational capabilities, and reducing operational overhead through automation.

This role is ideal for someone who has grown through engineering, cloud, platform, infrastructure, or DevOps leadership roles and understands how to bridge engineering and operations to deliver exceptional business outcomes.

Key ResponsibilitiesReliability Engineering Leadership
  • Establish and lead Antares' Reliability Engineering function.
  • Define and implement strategies that improve system reliability, resiliency, scalability, and operational excellence.
  • Partner with engineering teams to embed reliability practices throughout the software development lifecycle.
  • Drive adoption of modern operational practices including service ownership, operational readiness reviews, SLOs, error budgets, and post-incident learning.
Technology Operations
  • Lead production operations across critical business applications and technology platforms.
  • Establish clear support models, escalation paths, ownership boundaries, and service management processes.
  • Oversee operational readiness, release support, change management, and platform health.
  • Continuously improve operational maturity through metrics, automation, process simplification, and engineering collaboration.
Azure Cloud & Kubernetes Operations
  • Provide technical leadership for cloud-based platforms running in Microsoft Azure.
  • Partner with Infrastructure and Engineering teams to optimize reliability, scalability, and operational efficiency.
  • Support containerized workloads and Kubernetes-based environments.
  • Drive best practices around cloud architecture, capacity planning, platform resilience, security, governance, and cost optimization.
  • Ensure cloud platforms are designed and operated to meet business continuity and availability objectives.
Incident Response & Problem Management
  • Serve as the executive incident leader during major production events.
  • Coordinate cross-functional teams during outages and high-severity incidents.
  • Manage communications with business stakeholders and technology leadership.
  • Establish disciplined root cause analysis processes and ensure corrective actions are executed.
  • Drive long-term reduction in recurring incidents and operational risk.
AI-Powered Operations & Automation
  • Champion an automation-first and AI-enabled approach to technology operations.
  • Identify opportunities to leverage AI for incident triage, alert correlation, knowledge management, operational analytics, and runbook execution.
  • Partner with engineering teams to develop intelligent automation and self-healing capabilities.
  • Evaluate emerging AIOps, agentic AI, and automation technologies and drive adoption where appropriate.
  • Reduce manual operational effort through scripting, workflow automation, orchestration platforms, and AI-assisted tooling.
Strategic Delivery & Organizational Leadership
  • Lead and mentor high-performing technology operations and reliability engineering teams.
  • Build strong partnerships across Engineering, Infrastructure, Cybersecurity, Architecture, Data, and Business teams.
  • Translate operational challenges into actionable roadmaps and measurable initiatives.
  • Drive accountability, execution excellence, and continuous improvement across the organization.
  • Present operational trends, risks, recommendations, and performance metrics to technology leadership.
Vendor & Partner Management
  • Manage strategic relationships with technology vendors and managed service providers.
  • Ensure vendor accountability against service commitments and contractual obligations.
  • Lead vendor escalations during service-impacting events.
  • Integrate third-party support processes into internal operational workflows.
Resiliency & Disaster Recovery
  • Partner with Engineering and Infrastructure teams to ensure systems meet recovery objectives.
  • Improve operational readiness for disaster recovery and business continuity events.
  • Support development of recovery automation, replay capabilities, and resilient platform architectures.
  • Promote documentation and operational knowledge sharing to reduce dependency on tribal knowledge.
QualificationsRequired
  • 7+ years of experience in software engineering, platform engineering, cloud engineering, infrastructure engineering, DevOps, technology operations, or related technical disciplines.
  • 3-5+ years of experience leading engineering, reliability, platform, cloud, DevOps, or technology operations teams.
  • Strong experience operating and supporting applications within Microsoft Azure environments.
  • Hands-on experience with Kubernetes and container-based platforms.
  • Experience supporting distributed systems and cloud-native architectures.
  • Strong understanding of application architecture, system dependencies, and production support models.
  • Experience leading major incident response activities and outage management processes.
  • Experience implementing monitoring, observability, and alerting solutions using tools such as Datadog, Grafana, Splunk, Dynatrace, New Relic, or similar platforms.
  • Experience building automation solutions using scripting languages, APIs, orchestration tools, and workflow platforms.
  • Strong understanding of DevOps, CI/CD, release management, and software delivery practices.
  • Demonstrated ability to define, measure, and improve KPIs related to reliability, availability, operational efficiency, and service quality.
  • Excellent communication and stakeholder management skills with the ability to influence technical and business leaders.
Preferred
  • Experience building or leading Reliability Engineering, Production Engineering, DevOps, or Platform Engineering organizations.
  • Experience implementing AI-assisted operational workflows, AIOps platforms, AI agents, or intelligent automation solutions.
  • Experience with ServiceNow, Control-M, or equivalent enterprise operational platforms.
  • Experience operating within highly regulated environments.
  • Financial services experience, including asset management, lending, banking, private credit, or investment management.

The Fine Print

  • Must have unrestricted authorization to work in the United States.
  • Must be willing to comply with pre-employment screening, including but not limited to drug testing, reference verification, and background check.
  • Must be willing to work from the Chicago or New York office.
#LI-hybrid

A reasonable estimate of the current base salary range at the time of posting is below. Base salary does not include other forms of compensation or benefits. Actual base salary within the specified range is comprised of several components, including but not limited to applicant's skill, prior relevant experience, specific degrees and certifications, job responsibilities, market considerations and the location of the position.

This role is eligible for a discretionary annual bonus (based on company, business unit and individual performance).

Our benefit offerings include medical, dental and vision coverage, employer paid short & long-term disability and life insurance, 401(k), profit sharing, paid time off, Maven family & fertility benefit, parental leave (including adoption, surrogacy, and foster placement), as well as other voluntary benefits.

Base Salary Range

$175,000 - $225,000

To learn more, visit www.antares.com. Antares is an Equal Opportunity Employer. Employment decisions are made without regard to race, color, religion, national or ethnic origin, sex, sexual orientation, gender identity or expression, age, disability, protected veteran status or other characteristics protected by law.

HQ

Antares Capital LP Chicago, Illinois, USA Office

500 West Monroe St,, Chicago, IL, United States, 60661

Similar Jobs

Yesterday
Remote or Hybrid
Hoffman Estates, IL, USA
60K-70K Annually
Entry level
60K-70K Annually
Entry level
Artificial Intelligence • Fintech • Insurance • Marketing Tech • Software • Analytics
Participate in a claims specialist development program by investigating insurance claims, reviewing medical records, evaluating damages, determining liability and case severity, and negotiating settlements with customers or attorneys. The role includes comprehensive training, mentoring, and independent claims management while requiring strong analytical, communication, negotiation, empathy, and organizational skills.
Yesterday
In-Office
Schaumburg, IL, USA
92K-164K Annually
Senior level
92K-164K Annually
Senior level
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Administer, architect, optimize, secure, and automate PostgreSQL, MySQL, and Snowflake databases in Azure. The role develops resilient database solutions for healthcare applications, manages performance, access controls, encryption, backups, and disaster recovery, and creates technical documentation. It collaborates across Scrum teams, presents solutions to stakeholders, evaluates emerging technologies, uses approved AI tools, solves complex database challenges, and mentors junior engineers.
Top Skills: AgileAi ToolsAzure Database For Postgresql Flexible ServerGitJavaLinuxAzureMySQLPostgresPythonScrumSnowflakeSQL
Yesterday
Hybrid
Hoffman Estates, IL, USA
67K-126K Annually
Senior level
67K-126K Annually
Senior level
Artificial Intelligence • Fintech • Insurance • Marketing Tech • Software • Analytics
Investigates and evaluates moderately complex to complex workers compensation and commercial claims. Determines coverage, liability, compensability, reserves, and damages; negotiates settlements and coordinates litigation. The role also engages experts, conducts mediations and reviews, mentors team members, supports escalated issues, participates in quality reviews, and may assist with team workload assignments and technical direction.

What you need to know about the Chicago Tech Scene

With vibrant neighborhoods, great food and more affordable housing than either coast, Chicago might be the most liveable major tech hub. It is the birthplace of modern commodities and futures trading, a national hub for logistics and commerce, and home to the American Medical Association and the American Bar Association. This diverse blend of industry influences has helped Chicago emerge as a major player in verticals like fintech, biotechnology, legal tech, e-commerce and logistics technology. It’s also a major hiring center for tech companies on both coasts.

Key Facts About Chicago Tech

  • Number of Tech Workers: 245,800; 5.2% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: McDonald’s, John Deere, Boeing, Morningstar
  • Key Industries: Artificial intelligence, biotechnology, fintech, software, logistics technology
  • Funding Landscape: $2.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Pritzker Group Venture Capital, Arch Venture Partners, MATH Venture Partners, Jump Capital, Hyde Park Venture Partners
  • Research Centers and Universities: Northwestern University, University of Chicago, University of Illinois Urbana-Champaign, Illinois Institute of Technology, Argonne National Laboratory, Fermi National Accelerator Laboratory

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account