DEFCON AI Logo

DEFCON AI

Senior Data Platform Engineer

Posted 4 Days Ago
Remote
Hiring Remotely in USA
165K-200K Annually
Senior level
Remote
Hiring Remotely in USA
165K-200K Annually
Senior level
Design and build a durable data storage platform that preserves source disagreements, provenance, uncertainty, historical versions, and evidence. Develop result and evidence stores, lineage tracking, versioned rules and configurations, reproducible decision records, and point-in-time entity views. Collaborate with entity-resolution and machine-learning teams to support probabilistic outputs. The role requires production data engineering experience, strong Python and SQL, PostgreSQL expertise, U.S. citizenship, and an active Secret clearance.
The summary above was generated by AI

ABOUT DEFCON AI

RESILIENCE IN THE FACE OF DISRUPTION. DEFCON AI is an insights company that leverages artificial intelligence, mathematical optimization, data analytics, and software engineering for resilient optimization of complex systems.
In today’s dynamically changing world, DEFCON AI’s technology aligns outcomes with operational goals, better decision making, and empowers customers to anticipate assess, and mitigate the impacts of disruptions.

If your answer to “the two records disagree” is never overwrite, this role was built for you. 

About the Role 

As a Senior Data Platform Engineer, you’ll own the durable view of every resolved entity on the platform: the storage layer that holds what we know, where each fact came from, how confident we are in it, and how that picture has changed over time. When two sources disagree, both are kept with their evidence. When someone asks what we believed on a given date, and why, the platform can answer. 

You’ll join the analytics and AI engineering team behind a system that ingests records from dozens of disparate sources, resolves them to the right entity, highlights what analysts should review first, and provides transparent, explainable recommendations that users can trust. Operating within a secure government cloud environment, the platform depends on a storage layer that never quietly erases a disagreement between sources. 

Upstream matching is probabilistic. A match arrives with a confidence score rather than a yes, and the storage layer carries that uncertainty forward rather than collapsing it into one clean record. Sources change shape without warning and there is no shared key to join on, so this is a current data-engineering problem rather than a classic warehousing one. 

The seniority this role calls for is about judgment, not tooling breadth: knowing why you never overwrite, and being able to reconstruct a past decision with its evidence a year later. If your instinct when records disagree is to preserve both and let a human decide, you’ll be at home here. 

This is a fully remote role with occasional travel to DEFCON AI headquarters, customer sites, and partner facilities as needed.


Key Responsibilities 

  • Design and build the storage layer that preserves source disagreement and history rather than resolving them away at write time 
  • Build the first release’s result and evidence store: saved per-person results linked to the originating record, prior results and human feedback retained separately, and configuration versions on every result 
  • Implement provenance and lineage across every node and edge against the platform’s evidence-record contract 
  • Store the versioned rule library and configuration, and record on every result the exact rule versions and ordered context supplied to any model-assisted step 
  • Ensure every write is traceable to its origin and every decision is reproducible 
  • Carry match confidence and other uncertainty forward through the platform rather than collapsing it into a single value 
  • Support point-in-time questions: what did we believe about this entity on this date, and on what evidence 
  • Define, with entity resolution and ML teammates, what the storage layer needs to receive and what it must serve downstream 

Required Qualifications 

  • 6+ years in data engineering or data platform engineering, including production experience with versioned, temporal, or historized data models 
  • Experience designing storage where history, provenance, and source disagreement are first-class rather than resolved away 
  • Comfort building on probabilistic upstream output rather than a clean shared key 
  • Strong Python and SQL, with production experience on PostgreSQL or comparable 
  • US Citizenship Required 
  • Active US Secret clearance 

Preferred Qualifications 

  • Experience building the storage layer downstream of an entity-resolution or record-linkage system 
  • Bi-temporal, event-sourced, graph, or temporal data modeling in production 
  • Federal, government, or regulated-industry experience where a historical decision had to be reproducible on demand 
  • Active Top Secret clearance 

What Success Looks Like 

  • A durable view that never silently overwrites a disagreement between sources 
  • “What did we believe on this date, and on what evidence” is answerable a year later 
  • The storage layer holds up as probabilistic match output flows in, rather than assuming a clean join 

What We Offer 

  • A fully remote, results-based environment 
  • Competitive salary, bonus, and equity package 
  • 100% employer paid, comprehensive health insurance including medical, dental, and vision for you and your family 
  • Unlimited PTO, with your manager’s approval 
  • Flexible work environment where you manage your work day 
  • 14 weeks of fully-paid parental leave 

Salary Range: $165,000—$200,000. This represents the typical salary range for this position based on experience, skills, and other factors. 


We’re an Equal Opportunity Employer: You’ll receive consideration for employment without regard to race, sex, color, religion, sexual orientation, gender identity, national origin, protected veteran status, or on the basis of disability. 
Applicant Data Disclosure   
By submitting an application, you acknowledge that Defcon AI uses third-party service providers to facilitate its recruitment and hiring processes. These providers include applicant tracking systems, candidate verification platforms, and fraud detection tools (collectively, "Hiring Platforms"). Your application materials, including your résumé, cover letter, work samples, responses to application questions, and any other information you submit, may be transmitted to and processed by these Hiring Platforms for the following purposes:  
  • Managing and administering your application throughout the hiring process; 
  • Verifying the accuracy and authenticity of application materials, including by cross-referencing information you provide against publicly available sources and proprietary databases; 
  • Identifying indicators of potentially fraudulent, fabricated, or materially misleading application content, including but not limited to discrepancies between submitted materials and publicly available professional profiles, geographic anomalies, and fabricated work histories. 
Applications that are flagged through this process as containing indicators of fraud or material misrepresentation may be declined from further consideration. If you have questions about the status of your application or the evaluation process, please contact [email protected].  
 
Defcon AI requires its Hiring Platform providers to process your information solely for the purposes described above and in accordance with applicable law. Your information will be retained only for as long as necessary to fulfill these purposes and any applicable legal obligations, after which it will be deleted in accordance with Defcon AI's data retention policies.
For more information about how your data is used, please refer to our Privacy Policy and Applicant Privacy Notice.  

 

Similar Jobs

11 Days Ago
Remote or Hybrid
USA
140K-215K Annually
Senior level
140K-215K Annually
Senior level
Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Architects, deploys, and operates scalable data platform infrastructure supporting analytics. Responsibilities include administering Airflow and Superset, managing Terraform infrastructure, building CI/CD pipelines, enforcing security and compliance, developing DBT data models, cataloging assets in OpenMetadata, writing Python automation, and mentoring engineers. The role requires extensive infrastructure and data engineering experience across cloud platforms, containers, data security, and workflow orchestration.
Top Skills: Apache AirflowApache SupersetAWSAzureCi/CdData ModelingDbtDockerGCPIamKubernetesMachine Learning PipelinesOciOpenmetadataPythonSQLTerraform
27 Days Ago
Remote or Hybrid
191K-334K Annually
Senior level
191K-334K Annually
Senior level
Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Build and operate a security data platform connecting SIEM, identity, asset, threat intelligence, and knowledge graph systems. Design core APIs, distributed data pipelines, schemas, risk scoring, entity resolution, and coverage queries. Own architecture and platform workstreams from scoping through delivery, collaborate with detection and identity teams, review code, mentor engineers, and drive partner adoption. The role requires expert Python, deep Splunk expertise, cybersecurity experience, and senior-level project ownership.
Top Skills: Asset InventoryAsynchronous Data PipelinesIdentity GovernanceKnowledge GraphsLarge Language ModelsMachine LearningMessage BusMitre Att&CkPythonRetrieval-Augmented GenerationSchema RegistrySIEMSplSplunkSplunk Apis
37 Minutes Ago
In-Office or Remote
3 Locations
168K-270K Annually
Senior level
168K-270K Annually
Senior level
Artificial Intelligence • Computer Vision • Hardware • Robotics • Metaverse
Lead the architecture, implementation, and operation of DGX Cloud’s shared data platform. Build scalable batch and streaming pipelines, data models, APIs, libraries, and orchestration capabilities for fleet telemetry and operational data. Own cross-team technical initiatives, production investigations, data quality, observability, security, and reliability standards. Provide hands-on coding, architectural leadership, mentorship, and guidance for migrations, distributed systems, cloud infrastructure, and production data products.
Top Skills: Analytical DatabasesChange Data CaptureCi/CdCloud InfrastructureContainer OrchestrationDistributed Data ProcessingETLEvent ProcessingGpu ClustersLakehouse ArchitectureObject StorageRelational DatabasesSQLStreamingWorkflow Orchestration

What you need to know about the Chicago Tech Scene

With vibrant neighborhoods, great food and more affordable housing than either coast, Chicago might be the most liveable major tech hub. It is the birthplace of modern commodities and futures trading, a national hub for logistics and commerce, and home to the American Medical Association and the American Bar Association. This diverse blend of industry influences has helped Chicago emerge as a major player in verticals like fintech, biotechnology, legal tech, e-commerce and logistics technology. It’s also a major hiring center for tech companies on both coasts.

Key Facts About Chicago Tech

  • Number of Tech Workers: 245,800; 5.2% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: McDonald’s, John Deere, Boeing, Morningstar
  • Key Industries: Artificial intelligence, biotechnology, fintech, software, logistics technology
  • Funding Landscape: $2.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Pritzker Group Venture Capital, Arch Venture Partners, MATH Venture Partners, Jump Capital, Hyde Park Venture Partners
  • Research Centers and Universities: Northwestern University, University of Chicago, University of Illinois Urbana-Champaign, Illinois Institute of Technology, Argonne National Laboratory, Fermi National Accelerator Laboratory

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account