Arena (arena.ai) Logo

Arena (arena.ai)

Software Engineer - Data Infrastructure

Reposted One Month Ago
In-Office or Remote
Hiring Remotely in CA
150K-350K Annually
Senior level
In-Office or Remote
Hiring Remotely in CA
150K-350K Annually
Senior level
The role involves designing and building data pipelines for analyzing user feedback on AI models, ensuring data quality and infrastructure scalability.
The summary above was generated by AI
About Arena Intelligence

Arena is the platform for evaluating how AI models perform in the real world. Founded by researchers from UC Berkeley's SkyLab, we're on a mission to measure and advance the frontier of AI for real-world use, and to build the foundation for everyone to understand, shape, and benefit from it.


Tens of millions of people use Arena each month to evaluate how frontier systems handle the work they actually do. The preferences they share power the most transparent, rigorous, and human-centered evaluations in AI. Leading AI labs, enterprises, and independent researchers rely on our work and open datasets to understand how models behave in real workflows: agentic coding, creative generation, professional productivity, and beyond. We go beyond leaderboards and decompose what human experience reveals about AI, so models advance toward the work people actually do.


We're a team of researchers, academics, builders, and creatives from UC Berkeley, Google, Stanford, and DeepMind. We seek truth, move fast, and value craftsmanship, curiosity, and impact over hierarchy. We're building a company where thoughtful, curious people from all backgrounds can do their best work together, in an office culture that radiates excellence, energy, and focus.

About the Role

Arena Intelligence is seeking a Data Infrastructure Software Engineer to join our team and build the data pipelines and infrastructure that powers real-world AI evaluation. You'll play a crucial role in designing and building the data pipelines that process and analyze tens of millions user vote data, directly impacting how we understand and evaluate AI model performance. This role is ideal for someone who thrives in fast-moving environments and interested in building products to ensure accurate and fair evaluation of human preferences across different models, which will shape the direction of future AI development.

As an early member of our data engineering team, you'll partner closely with researchers, engineers, and product leadership to retrieve valuable data and insights from human votes and feedback. You'll help us move fast while staying rigorous, improving data quality, scaling our infrastructure to new levels, and deepening our ability to compare frontier models and predict human preferences.

You’ll
  • Design and build robust data pipelines to ingest, process, and transform user vote data to features essential for model performance evaluation.

  • Collaborate with researchers and product leadership to understand product goals and necessary data.

  • Design and implement solutions to generate result dashboards and reports, providing useful information for the public, model providers, and researchers.

  • Ensure the integrity, data quality, and reliability of the pipelines.

  • Scale our data infrastructure to accommodate increasing data volumes and evolving analytical needs.

You’ll have
  • 5+ years of experience in software engineering, with a dedicated focus on data engineering and big data technologies

  • Proficiency in SQL and at least one programming language commonly used for data analysis (Python (preferred), Scala, R).

  • Hands-on experience with data processing and pipeline frameworks (Apache Spark, Ray Data, etc.) and at least one popular big data analytics platform (Databricks, Snowflake).

  • Demonstrated experience in designing, implementing, optimizing, and debugging production data pipelines.

Bonus Points
  • Prior work in data analytics or datalake platforms.

  • Experience in advanced data analysis tools, such as Delta lake, streaming tables.

  • Exposure to machine learning is a plus.

What we offer
  • We offer competitive compensation and equity aligned to the markets where our team members are based. The base salary range will depend on the candidate’s permanent work location.

  • Comprehensive health and wellness benefits, including medical, dental, vision, and additional support programs.

  • The opportunity to work on cutting-edge AI with a small, mission-driven team

  • A culture that values transparency, trust, and community impact

Come help build the space where anyone can explore and help shape the future of AI.

Arena Intelligence provides equal employment opportunities (EEO) to all employees and applicants for employment without regard to race, color, religion, sex, national origin, age, disability, genetics, sexual orientation, gender identity, or gender expression. We are committed to a diverse and inclusive workforce and welcome people from all backgrounds, experiences, perspectives, and abilities.

Similar Jobs

One Month Ago
Remote or Hybrid
110K-300K Annually
Mid level
110K-300K Annually
Mid level
Financial Services
Build and maintain scalable financial data infrastructure, including ingestion, validation, normalization, and integrations. Develop systems that calculate portfolio performance metrics efficiently across high-volume datasets, improve reliability and observability, and support distributed processing. Create internal tools, dashboards, reusable schemas, asset data feeds, and self-service analytics capabilities. Collaborate cross-functionally to expand market-data coverage, support B2B integrations, and strengthen operational resilience.
Top Skills: Business Intelligence DashboardsData InfrastructureData Ingestion PipelinesData ValidationDistributed Systems
4 Months Ago
Remote
United States
175K-230K Annually
Senior level
175K-230K Annually
Senior level
Healthtech
As a Staff Software Engineer, you'll architect and build foundational platform services, focusing on data infrastructure, healthcare interoperability, and secure APIs while mentoring other engineers.
Top Skills: AWSAzureDatabricksGCPHl7 Fhir ApisJavaKubernetesOauth2OidcPythonSparkTerraform
2 Months Ago
In-Office or Remote
United States
160K-325K Annually
Mid level
160K-325K Annually
Mid level
Artificial Intelligence • Machine Learning • Natural Language Processing • Software • Generative AI
As a Software Engineer in Data Infrastructure, you will build and maintain high-performance data storage systems for AI training, collaborate with experts, and optimize data processing.
Top Skills: AirflowApache BeamBigQueryDbtFlinkGcsKubernetesPythonS3Spark

What you need to know about the Chicago Tech Scene

With vibrant neighborhoods, great food and more affordable housing than either coast, Chicago might be the most liveable major tech hub. It is the birthplace of modern commodities and futures trading, a national hub for logistics and commerce, and home to the American Medical Association and the American Bar Association. This diverse blend of industry influences has helped Chicago emerge as a major player in verticals like fintech, biotechnology, legal tech, e-commerce and logistics technology. It’s also a major hiring center for tech companies on both coasts.

Key Facts About Chicago Tech

  • Number of Tech Workers: 245,800; 5.2% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: McDonald’s, John Deere, Boeing, Morningstar
  • Key Industries: Artificial intelligence, biotechnology, fintech, software, logistics technology
  • Funding Landscape: $2.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Pritzker Group Venture Capital, Arch Venture Partners, MATH Venture Partners, Jump Capital, Hyde Park Venture Partners
  • Research Centers and Universities: Northwestern University, University of Chicago, University of Illinois Urbana-Champaign, Illinois Institute of Technology, Argonne National Laboratory, Fermi National Accelerator Laboratory

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account