Ford Motor Company Logo

Ford Motor Company

Site Reliability Engineer - Observability Platform

Posted An Hour Ago
Be an Early Applicant
In-Office or Remote
Hiring Remotely in United States
85K-193K Annually
Senior level
In-Office or Remote
Hiring Remotely in United States
85K-193K Annually
Senior level
Design, build, and operate a global observability platform across hybrid cloud and on-premises environments. Responsibilities include developing monitoring pipelines, infrastructure-as-code, reliability automation, SLI/SLO frameworks, performance optimization, incident response, root-cause analysis, and production troubleshooting. The role partners with engineering teams to improve system resilience, reduce toil, integrate AI/ML for anomaly detection, and establish observability best practices. It also provides technical mentorship and guidance.
The summary above was generated by AI

We made history and now we work to transform the future – for our customers, our communities and our families. You'll see your work on the road every day, helping people move freely and pursue their dreams. At Ford, you can build more than vehicles. Come build what matters.


Enterprise Technology plays a critical part in shaping the future of mobility. If you’re looking for the chance to leverage advanced technology to redefine the transportation landscape, enhance the customer experience and improve people’s lives, this is the opportunity for you. Join us and challenge your IT expertise and analytical skills to help create vehicles that are as smart as you are.

 

The Observability Platform Team designs, builds, and operates the monitoring and observability infrastructure that underpins visibility into application performance across hybrid environments—on-prem and cloud. Our platform integrates AI-driven analytics with intuitive dashboards to give engineering teams the telemetry, metrics, logs, and traces they need to detect issues faster, reduce MTTR, and drive continuous performance optimization.

 

We're hiring an experienced Site Reliability Engineer to architect, extend, and scale our global observability platform. This role sits at the intersection of software and systems engineering, requiring strong skills in distributed systems design, infrastructure automation, and production operations to ensure high availability, scalability, and maintainability of our monitoring stack.

Responsibilities
  • Design and implement scalable observability pipelines spanning metrics, logging, tracing, and alerting
  • Define and operationalize Service Level Indicators (SLIs) and Service Level Objectives (SLOs), and establish error budgets to effectively drive maximum availability and uptime
  • Build reusable infrastructure-as-code templates and frameworks to standardize observability instrumentation and onboarding
  • Architect, design, and develop automation to improve the resilience, recoverability, availability, and scalability of supported applications
  • Leverage experience to safely perform destructive testing to seek and discover vulnerabilities
  • Develop tooling to improve reliability, quality, and time-to-market for software solutions
  • Identify and reduce or eliminate toil via automation to maximize time spent on engineering and innovation
  • Collaborate with development teams to design, build, and operate scalable and resilient software systems using cloud-native principles
  • Proactively identify stability risks and work with engineering leadership to establish appropriate mitigation plans
  • Regularly review key technical metrics such as transaction errors, logging, response times, caching strategies, conversion/bounce rates, capacity, and resource utilization
  • Conduct performance analysis and optimization of new and in-production systems, measuring and optimizing performance to get ahead of customer needs and drive continuous innovation
  • Solve complex architecture, design, and business problems by simplifying processes, optimizing systems, and removing bottlenecks
  • Recognize, validate, and evangelize emerging technologies and architectures that align with business objectives
  • Troubleshoot complex, distributed production systems and drive root-cause analysis for platform-level incidents
  • Participate in incident response, support, recovery, and postmortem analysis
  • Provide technical guidance and mentorship to other team members
  • Continuously evaluate and integrate AI/ML capabilities to enhance anomaly detection, alerting precision, and performance insights
  • Collaborate cross-functionally with engineering teams to embed observability best practices into system design and deployment workflows
Qualifications
  • Bachelor’s Degree in Computer Science or equivalent experience
  • 3+ years of experience in an SRE role
  • 5+ years of programming experience with one or more of: Python, Go, Java/Scala, C, or C++
  • 3+ years of experience building reusable infrastructure-as-code templates & frameworks in Terraform or ToFu
  • 3+ years of experience with APM and monitoring tools such as Dynatrace, New Relic, ELK, Splunk, Prometheus, Sensu, Nagios, Kafka, or DataDog
  • 3+ years of experience with J2EE, NoSQL/SQL datastores, Spring Boot, GCP/AWS/Azure, and Docker/Kubernetes in developing multi-tier applications
  • Experience with RESTful APIs and microservices platforms
  • Working knowledge of the TCP/IP stack, internet routing, and load balancing
  • Strong proficiency with Google Cloud Platform and its library of services
  • Experience with automated, test-driven development in CI/CD pipelines
  • Thorough understanding of software development and agile methodologies
  • Understanding of, and ability to implement, effective observability strategies to improve MTTD/MTTR (Mean Time to Detect/Resolve)

 

You may not check every box, or your experience may look a little different from what we've outlined, but if you think you can bring value to Ford Motor Company, we encourage you to apply!
As an established global company, we offer the benefit of choice. You can choose what your Ford future will look like: will your story span the globe, or keep you close to home? Will your career be a deep dive into what you love, or a series of new teams and new skills? Will you be a leader, a changemaker, a technical expert, a culture builder…or all of the above? No matter what you choose, we offer a work life that works for you, including:

  • Immediate medical, dental, vision and prescription drug coverage
  • Flexible family care days, paid parental leave, new parent ramp-up programs, subsidized back-up childcare and more
  • Family building benefits including adoption and surrogacy expense reimbursement, fertility treatments, and more
  • Vehicle discount program for employees and family members and management leases
  • Tuition assistance
  • Established and active employee resource groups
  • Paid time off for individual and team community service
  • A generous schedule of paid holidays, including the week between Christmas and New Year’s Day
  • Paid time off and the option to purchase additional vacation time.

 

This position is a range of salary grades 6-8 and ranges from $85,400-$192,900.    
Final determination of salary grade will be based on candidate's skills and experience, and base salary will be set within the applicable range according to job scope, responsibility and competitive market value.
Internal applicants: moving into this role may result in an adjustment to your current compensation based on the posted pay range for this role, taking into consideration your qualifications and other relevant factors.
For more information on salary and benefits, click here: https://fordcareers.co/GSR

 

Visa sponsorship is not available for this position.

 

Candidates for positions with Ford Motor Company must be legally authorized to work in the United States. Verification of employment eligibility will be required at the time of hire.

We are an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, religion, color, age, sex, national origin, sexual orientation, gender identity, disability status or protected veteran status. In the United States, if you need a reasonable accommodation for the online application process due to a disability, please call 1-888-336-0660.

#LI-Remote

#LI-PW1

About UsAt Ford Motor Company, we believe freedom of movement drives human progress. With our incredible plans for the future of mobility, we have a wide variety of opportunities for you to accelerate your career and help us define tomorrow’s transportation.About the TeamWe believe that freedom of movement drives human progress. Ford Information Technology (IT) is shaping the future of mobility by redefining the transportation landscape, enhancing the customer experience and improving people’s lives. Join the Ford family as we change the way the world moves.

Ford Motor Company Oak Park, Illinois, USA Office

Oak Park, United States

Similar Jobs at Ford Motor Company

An Hour Ago
In-Office or Remote
United States
85K-193K Annually
Mid level
85K-193K Annually
Mid level
Automotive
Develop and maintain global monitoring and observability platforms using Go, JavaScript, GCP, Kubernetes, OpenTelemetry, PostgreSQL, and Terraform. Improve reliability, scalability, performance, security, and disaster recovery for cloud services. Responsibilities include troubleshooting production systems, capacity planning, automation, on-call support, incident postmortems, code reviews, documentation, and vulnerability assessments.
Top Skills: Document DatabasesDynatraceGoGoogle Cloud PlatformInfrastructure As CodeJavaScriptKubernetesOpentelemetryPostgresRelational DatabasesTerraform
7 Hours Ago
In-Office or Remote
United States
133K-251K Annually
Expert/Leader
133K-251K Annually
Expert/Leader
Automotive
Leads Ford Credit’s enterprise AI architecture strategy and manages architects while directly designing production-ready AI and agentic systems. Establishes AI-native delivery practices, evaluation standards, governance, security controls, cloud-native patterns, and modernization roadmaps. Guides Google Cloud architecture, hybrid-system integration, AI risk management, reusable reference architectures, and platform capabilities. Advises executives and cross-functional teams while developing the architecture organization through coaching, delegation, and performance management.
Top Skills: Agentic SystemsAPIsAuthorizationBigQueryCloud RunData ProtectionGeminiGenerative AiGCPGoogle Kubernetes Engine (Gke)Hybrid CloudIdentity And Access ManagementLarge Language ModelsMainframe SystemsModel Lifecycle ManagementObservabilityPub/SubRetrieval-Augmented GenerationSaaSSecrets ManagementSoftware Supply-Chain SecurityThreat ModelingVertex AiZero Trust
9 Hours Ago
In-Office or Remote
United States
100K-193K Annually
Senior level
100K-193K Annually
Senior level
Automotive
Designs, builds, maintains, and evolves Informatica IDMC and MDM SaaS solutions, integrations, dataflows, and business-critical software. Integrates systems using API- and event-based GCP patterns, Java, and Spring Boot. Leads requirements breakdown, guides projects through deployment, applies automated testing and CI/CD, documents architectures, mentors teammates, and advances AI-assisted development practices.
Top Skills: BpelC4 ModelingCi/CdCloud App Integration (Cai)Cloud Data Integration (Cdi)Cloud Data Quality (Cdq)Customer 360 MdmData Integration & ReplicationETLGCPGitInformatica IdmcInformatica Mdm SaasJavaJIRAJSONMermaidNoSQLSafeSpring BootSQLXpathXquery

What you need to know about the Chicago Tech Scene

With vibrant neighborhoods, great food and more affordable housing than either coast, Chicago might be the most liveable major tech hub. It is the birthplace of modern commodities and futures trading, a national hub for logistics and commerce, and home to the American Medical Association and the American Bar Association. This diverse blend of industry influences has helped Chicago emerge as a major player in verticals like fintech, biotechnology, legal tech, e-commerce and logistics technology. It’s also a major hiring center for tech companies on both coasts.

Key Facts About Chicago Tech

  • Number of Tech Workers: 245,800; 5.2% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: McDonald’s, John Deere, Boeing, Morningstar
  • Key Industries: Artificial intelligence, biotechnology, fintech, software, logistics technology
  • Funding Landscape: $2.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Pritzker Group Venture Capital, Arch Venture Partners, MATH Venture Partners, Jump Capital, Hyde Park Venture Partners
  • Research Centers and Universities: Northwestern University, University of Chicago, University of Illinois Urbana-Champaign, Illinois Institute of Technology, Argonne National Laboratory, Fermi National Accelerator Laboratory

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account