Maximum of 25 job preferences reached.
Top Remote Principal Software Engineer Jobs in Chicago, IL
Machine Learning • Security • Software • Analytics • Defense
Provides technical leadership for real-time RF sensor software development, integration, testing, and validation. Designs high-performance C/C++ systems, integrates signal-processing algorithms, debugs distributed and multithreaded applications, and transitions MATLAB/Python prototypes into production software. The role includes technical mentoring, customer and government stakeholder engagement, proposal development, documentation, system demonstrations, visualization tools, and approximately 10% travel for integration and field-testing activities.
Top Skills:
C/C++GitLinuxMatlabPython
Artificial Intelligence • Cloud • Consumer Web • Productivity • Software • App development • Data Privacy
Owns company-critical, long-term technical problems across multiple teams and organizations. Defines technical strategy, architectures, roadmaps, and engineering standards for large-scale customer-facing systems. Remains hands-on in software development while driving AI adoption in products and engineering workflows. Partners with senior product, engineering, and executive leaders to resolve complex problems, improve reliability and performance, and shape Dropbox’s technical direction.
Top Skills:
Agentic FrameworksAIConcurrencyDatabasesDistributed SystemsFrontend SystemsLlm ApisMlMobile SystemsSearch SystemsSoftware DevelopmentStorage Systems
Big Data • Fitness • Healthtech • Information Technology • Software • Analytics
Build and own Arcadia’s internal developer platform across a large-scale data lakehouse and Kubernetes applications. Develop self-service APIs, templates, golden paths, CI/CD capabilities, observability, security controls, and reliability practices. Lead platform migrations, improve developer experience, manage stakeholder needs and costs, set technical direction, and coach senior engineers. The role requires hands-on software development alongside expertise in cloud infrastructure, Kubernetes, networking, infrastructure as code, SLOs, incident response, and platform product management.
Top Skills:
Access ControlsAPIsCi/CdCloud InfrastructureDatabasesEncryptionInfrastructure As CodeKubernetesNetworkingObservabilityQueuesSlosStorage
Information Technology • Cybersecurity
Lead the architecture and development of a production Agentic SOC that uses LLM-based agents to investigate security signals, gather evidence, make defensible decisions, and support SOC analysts. Own agent behavior, tool use, evaluation, guardrails, tenant isolation, auditability, reliability, and autonomous decision-making. Collaborate with analysts, detection engineers, product, and engineering teams while providing technical leadership and mentoring. Build scalable backend systems using distributed workflows, queues, cloud infrastructure, and data stores.
Top Skills:
AWSAzureClaude CodeLlmsPostgresRedisRubyRuby On Rails
Fintech • Financial Services
Lead the architecture and technical direction of Forward Financing’s core fintech platform. Design scalable distributed backend systems, guide cross-team initiatives, establish engineering best practices, resolve incidents, and mentor engineers. The role owns the technical roadmap and architectural vision while partnering across teams to improve reliability, observability, code quality, and delivery. Ruby on Rails expertise is required, along with strong communication, influence, agile development experience, and a broad backend or full-stack perspective.
Top Skills:
Ruby On Rails
Artificial Intelligence • Cloud • Sales • Security • Software • Cybersecurity • Data Privacy
Leads architecture and development of SailPoint’s large-scale, multi-tenant identity security SaaS platform. Translates business strategy into technical roadmaps, designs reliable microservices and infrastructure, drives modernization from monoliths to distributed systems, mentors engineers, reviews designs, delivers code, and communicates technology strategy to executives and stakeholders. The role also requires adoption and oversight of AI-assisted engineering tools and technologies.
Top Skills:
Amazon BedrockAWSClaude CodeCursorDistributed SystemsDockerGoGraphQLKafkaLarge Language ModelsMicroservicesNumpyPandasPostgresPythonPyTorchRedisRest ApisSaaSScikit-LearnTerraformTimescaledbVertex Ai
Security
Designs and develops internet-scale data processing, infrastructure management, and observability platforms across cloud and on-premises environments. Establishes engineering standards, resolves complex distributed architecture challenges, and provides technical leadership across the organization. The role requires experience with large-scale orchestration systems, internet protocols such as DNS, WHOIS, and RDAP, architectural trade-offs, and internet security technologies.
Top Skills:
AlertingCloud ComputingDistributed SystemsDnsLoggingMetricsOn-Premises InfrastructureOrchestration SystemsRdapWhois
Reposted YesterdaySaved
Artificial Intelligence • Software • Automation
Founding engineer responsible for architecting and building core product components, writing scalable Python code, designing and maintaining AI/ML infrastructure, implementing testing and CI/CD best practices, shipping backend production systems, collaborating on product roadmap, and mentoring future engineers.
Top Skills:
Ci/CdLangchainLlamaindexLlmsPythonTesting
Cloud • Security • Software • Generative AI
Leads the roadmap, architecture, security, and organizational adoption of CI/CD and release infrastructure. Designs scalable release gates, hardens supply chains, directs engineers through complex projects, manages severe incidents, and drives cross-functional technical alignment. The role remains hands-on through code reviews, prototyping, debugging, documentation, mentorship, and design reviews while primarily achieving impact through team leadership.
Top Skills:
Argo CdBuildkiteCi/CdGithub ActionsGitopsKubernetes
Software • Quantum Computing • Metaverse • Infrastructure as a Service (IaaS)
Design, build, and operate distributed storage and file-system platforms for cloud-scale AI training and inference. Architect storage and caching for Kubernetes and GPU workloads, improve performance, reliability, observability, and resiliency, diagnose cross-layer production issues, and lead architecture reviews. Partner across infrastructure, networking, operating systems, and AI platform teams while mentoring engineers and shaping long-term storage strategy.
Top Skills:
AzureCC#C++Collective Communication LibrariesContainer RuntimesCudaDistributed StorageFile SystemsGpu DriversInfinibandJavaJavaScriptKey-Value CachingKubernetesKubernetes Device PluginsPythonRdmaRoce
Software • Quantum Computing • Metaverse • Infrastructure as a Service (IaaS)
Lead the architecture, technical strategy, and roadmap for Microsoft's Windows ML compiler stack. Design high-performance compiler infrastructure, execution engines, optimization frameworks, and developer tools supporting CPU, GPU, and NPU workloads. Drive NPU enablement, graph optimization, scheduling, memory management, code generation, and hardware abstraction. Partner with silicon vendors and internal teams, influence AI hardware and platform strategy, lead cross-team initiatives, and mentor engineers developing on-device and generative AI capabilities.
Top Skills:
CC#C++CpusDevice DriversDirectxExecution RuntimesGpusJavaJavaScriptMl CompilersNpusOperating SystemsPythonWindows Ml
Artificial Intelligence • Cybersecurity
Own the architecture and technical direction of NodeZero’s database platform. Design scalable PostgreSQL and graph-based data systems, multi-tenant schemas, partitioning, replication, indexing, and zero-downtime migrations. Establish database standards, governance, performance practices, and migration tooling while writing production code. Mentor engineers, lead architecture reviews, and collaborate across data, inventory, API, and security teams to ensure reliable, isolated, and scalable data infrastructure.
Top Skills:
Aws NeptuneFlywayGoGolang-MigrateInfluxdbLtreeNeo4JPostgresPythonTimescaledb
New
Track Smarter, Apply Better.
Ditch the spreadsheets. Organize your job search with our freeApplication Tracker.
Use For Free
Software • Quantum Computing • Metaverse • Infrastructure as a Service (IaaS)
Develop and optimize GPU kernels, compilers, runtimes, and distributed inference systems for large language model serving. Improve latency, throughput, reliability, and hardware efficiency through profiling and low-level performance optimization. Integrate scalable improvements into production systems, collaborate across model, infrastructure, compiler, and hardware teams, and mentor engineers. Principal-level contributors additionally lead cross-stack initiatives and create reusable performance capabilities across models and hardware generations.
Top Skills:
Amd GpusC++CompilersDistributed Inference SystemsGpu KernelsGpu ProgrammingLlm InferenceNvidia GpusPythonRuntimesSglangVllm
Software • Quantum Computing • Metaverse • Infrastructure as a Service (IaaS)
Leads architecture and hands-on development for AI frameworks, performance systems, benchmarking, and developer tools. Investigates issues across models, compilers, runtimes, services, and accelerator hardware; improves model onboarding, runtime performance, reliability, hardware utilization, and Azure efficiency. Establishes scalable measurement and observability capabilities, influences cross-team technical strategy, delivers production solutions, and mentors engineers while advancing engineering quality and maintainability.
Top Skills:
Ai FrameworksAmd GpusC++CompilersDistributed InferenceAzureNvidia GpusPythonRuntimesSglangVllm
Software • Quantum Computing • Metaverse • Infrastructure as a Service (IaaS)
Work with customers and cross-functional teams to design, build, and deliver cloud-based solutions. Lead design, identify dependencies, mentor engineers, act as DRI and on-call, and improve availability, reliability, observability, and performance. Contribute to open-source assets and collaborate with Microsoft product teams to produce extensible, maintainable code and scalable solution patterns.
Top Skills:
CC#C++JavaJavaScriptMicrosoft CloudPython
Software • Quantum Computing • Metaverse • Infrastructure as a Service (IaaS)
Provides technical leadership for AI framework and model-serving initiatives at cloud scale. Owns architecture, roadmaps, performance optimization, benchmarking, observability, and multi-release execution across models, compilers, runtimes, services, GPUs, and silicon. Leads cross-team alignment, makes architectural decisions, mentors engineers, performs hands-on implementation and debugging, and drives measurable improvements in model onboarding, runtime performance, reliability, hardware utilization, and Azure capacity efficiency.
Top Skills:
Ai FrameworksAmd GpusAzureC++CompilersLlm Inference FrameworksMicrosoft SiliconNvidia GpusPythonRuntimes
Software • Quantum Computing • Metaverse • Infrastructure as a Service (IaaS)
Own the architecture, implementation, and operational excellence of major Azure Boost DPU software components. Develop low-level systems software spanning operating systems, virtualization, networking, storage, device management, and hardware interfaces. Lead complex cross-functional initiatives, optimize hyperscale performance and reliability, resolve system-level issues, establish engineering practices, influence platform architecture, and mentor engineers across Azure infrastructure teams.
Top Skills:
AzureCC#C++Cloud InfrastructureContainersDmaDpuHypervisorsIommuJavaJavaScriptLinuxNetworkingNvmePciePythonRdmaRustSmartnicSr-IovStorageVirtualizationWindows
Software • Quantum Computing • Metaverse • Infrastructure as a Service (IaaS)
Lead architecture and hands-on development for Microsoft's virtualization platform, focusing on hypervisor and VMM capabilities. Design, implement, and optimize systems software across Windows, Linux, macOS, Azure, and containers. Improve performance, security, reliability, testing, and operational readiness; diagnose low-level hardware/software issues; enable virtualization on new silicon; mentor engineers; and influence platform strategy across teams and hardware partners.
Top Skills:
AzureCC++ContainersHyper-VHypervisorLinuxmacOSOpenvmmRustVirtual Machine Monitor (Vmm)Windows
Software • Quantum Computing • Metaverse • Infrastructure as a Service (IaaS)
Lead architecture and hands-on development of Microsoft's AI Safety Platform, building observability, detection, investigation, threat-hunting, and distributed data-processing capabilities. Design attack detection infrastructure, correlate high-volume telemetry, apply AI security research, conduct security and privacy reviews, and partner across engineering, science, and incident response teams. Mentor engineers and provide technical leadership for complex AI security systems.
Top Skills:
Agent-Based SystemsAzure Data FactoryBatch ProcessingCC#C++Ci/CdData LakesDatabricksEmbeddingsInfrastructure As CodeJavaJavaScriptKubernetesKustoLarge Language ModelsPythonRetrieval-Augmented GenerationSparkStreaming PipelinesVector Databases
Software • Quantum Computing • Metaverse • Infrastructure as a Service (IaaS)
Work on end-to-end AI performance for LLMs: benchmark and optimize inference on GPUs and Microsoft silicon, build tooling for performance insights and model porting, implement and test components in AI/DNN frameworks, and collaborate with internal and external partners to reduce hardware footprint and accelerate deployments.
Top Skills:
AmdAzure OpenaiCC#C++CudaGpu Profiling ToolsJavaJavaScriptMicrosoft SiliconNvidiaOnnx RuntimePythonPyTorchRocmTensorFlowTriton
Software • Quantum Computing • Metaverse • Infrastructure as a Service (IaaS)
Lead the architecture and hands-on development of agentic observability capabilities for Kubernetes and Linux. Build production-quality software, interpret systems telemetry using eBPF, mentor engineers, coordinate cross-team delivery, support open-source initiatives, and participate in on-call operations. The role focuses on safe, observable, and evaluable AI-agent workflows, distributed systems reliability, and technical leadership across engineering teams.
Top Skills:
AzureCC++Distributed SystemsEbpfGoKubernetesLinuxOpen-Source SoftwarePythonRustSystems Telemetry
Software
Lead development of a massively distributed edge compute platform. Design fault-tolerant systems, external APIs, high-throughput traffic routing, and Kubernetes-based resource management. Establish architecture standards, reliability metrics, security practices, and technical governance. Collaborate with AI/ML, hardware, networking, and security teams; communicate risks to executives; manage customer requirements; review code; and mentor engineers on distributed systems best practices.
Top Skills:
AWSAzureContainerizationGCPKubernetes
Software • Quantum Computing • Metaverse • Infrastructure as a Service (IaaS)
Lead design, development, and operation of highly available AKS Resource Provider and control-plane services. Drive cluster lifecycle, upgrades, networking, security, scalability, observability, resiliency, and incident prevention for large-scale distributed cloud systems. Collaborate across Azure engineering teams, lead architectural decisions, improve tooling and automation, troubleshoot production issues, and mentor engineers while supporting reliable Kubernetes workloads across cloud and edge environments.
Top Skills:
AzureAzure Kubernetes Service (Aks)C#C++Cloud-Native InfrastructureContainersDnsGithub CopilotGoJavaKubernetesLinuxLoad BalancingPythonRoutingTcp/IpTlsWindows
Software • Quantum Computing • Metaverse • Infrastructure as a Service (IaaS)
Lead the architecture, development, testing, and deployment of hyperscale distributed storage systems optimized for AI and machine learning workloads. Drive innovations in scalability, performance, security, durability, and efficiency while collaborating with AI researchers and cloud infrastructure teams. Mentor engineers, guide technical strategy, and help shape Azure Blob Storage into an AI-native platform supporting zettabyte-scale storage.
Top Skills:
Ai/MlAzure Blob StorageC#C++Cloud InfrastructureDatabase SystemsDistributed SystemsJavaPython
Aerospace • Logistics • Security • Software • Cybersecurity
Principal Software Engineer responsible for operating and improving business-critical data pipelines and infrastructure. Duties include developing monitoring, alerting, logging, CI/CD workflows, scalable data integrations, automated testing, deployment automation, and production support. The role troubleshoots workflow failures, data-quality issues, performance bottlenecks, and infrastructure incidents while collaborating with data engineers, platform engineers, and architects. Preferred experience includes Databricks, Spark, AWS, Terraform, Unity Catalog, Delta Lake, observability, identity management, and enterprise data platform governance.
Top Skills:
Amazon MskAmazon S3Apache KafkaSparkAWSAws BatchDatabricksDatabricks Asset BundlesDatabricks WorkflowsDelta LakeGitGithub ActionsGitlab Ci/CdIdentity FederationPythonRest ApisRole-Based Access ControlSecrets ManagementService PrincipalsSql WarehousesTerraformUnity Catalog
Let Your Resume Do The Work
Upload your resume to be matched with jobs you're a great fit for.
Success! We'll use this to further personalize your experience.
Top Chicago, IL Companies Hiring Remote Principal Software Engineers
See AllPopular Job Searches
All Filters
Total selected ()
No Results
No Results
























