Companies Arena Intelligence, Inc. Machine Learning Engineer

About the role

Arena Intelligence, Inc. · Hybrid

About Arena Intelligence

Arena is the platform for evaluating how AI models perform in the real world. Founded by researchers from UC Berkeley's SkyLab, we're on a mission to measure and advance the frontier of AI for real-world use, and to build the foundation for everyone to understand, shape, and benefit from it.


Tens of millions of people use Arena each month to evaluate how frontier systems handle the work they actually do. The preferences they share power the most transparent, rigorous, and human-centered evaluations in AI. Leading AI labs, enterprises, and independent researchers rely on our work and open datasets to understand how models behave in real workflows: agentic coding, creative generation, professional productivity, and beyond. We go beyond leaderboards and decompose what human experience reveals about AI, so models advance toward the work people actually do.


We're a team of researchers, academics, builders, and creatives from UC Berkeley, Google, Stanford, and DeepMind. We seek truth, move fast, and value craftsmanship, curiosity, and impact over hierarchy. We're building a company where thoughtful, curious people from all backgrounds can do their best work together, in an office culture that radiates excellence, energy, and focus.

About the Role

Arena Intelligence is seeking a Senior Machine Learning Engineer to help scale and strengthen the core infrastructure that powers real-world AI evaluation. You’ll play a foundational role in shaping how we build, deploy, and improve our model benchmarking systems, working across data pipelines, inference APIs, and new evaluation methodologies. This is an opportunity to apply your technical expertise to a platform trusted by millions, and to help define how cutting-edge AI is assessed in the wild.

As one of the first ML engineers on the team, you’ll partner closely with researchers, engineers, and product leadership to turn new ideas into reliable systems. You’ll help us move fast while staying rigorous, improving reproducibility, scaling up to new modalities, and deepening our ability to understand and compare frontier models.

You’ll

  • Architect and build what will become our core modeling for data and evaluation products

  • Own the full stack data, model training, and eval pipelines

  • Help grow a culture of feedback and rapid product iteration as we build new features as a tight-nit team

  • Conduct research into state-of-the-art evaluation methods and contribute to the long-term vision for a centralized, scalable evaluation platform.

You’ll have

  • Strong programming skills with the ability to work across the stack in a typical recommendation system or LLM stack

  • Experience in deep learning, language models or reward model training

  • Experience in working with LLM for fine tuning, prompt engineering, function calling etc

  • Self-motivated with a willingness to take ownership of tasks

  • A passion for shipping quality products

  • 4+ years of industry experience or relevant projects

  • Solid understanding of statistics, and various tools and methodologies for evaluating uncertainty in a way that is specific to the given product being shipped

What we offer

  • We offer competitive compensation and equity aligned to the markets where our team members are based. The base salary range will depend on the candidate’s permanent work location.

  • Comprehensive health and wellness benefits, including medical, dental, vision, and additional support programs.

  • The opportunity to work on cutting-edge AI with a small, mission-driven team

  • A culture that values transparency, trust, and community impact

Come help build the space where anyone can explore and help shape the future of AI.

Arena Intelligence provides equal employment opportunities (EEO) to all employees and applicants for employment without regard to race, color, religion, sex, national origin, age, disability, genetics, sexual orientation, gender identity, or gender expression. We are committed to a diverse and inclusive workforce and welcome people from all backgrounds, experiences, perspectives, and abilities.

Ready to apply to Arena Intelligence, Inc.?
Apply to Arena Intelligence, Inc.

Similar jobs

Samsara
Lead Machine Learning Engineer - ML Infrastructure
Samsara
⚡ Apply early Remote - US · location restricted $200,200–$357,500
● New 👁 Seen ✓ Applied 1d ago
Samsara
Lead Machine Learning Engineer - ML Infrastructure
Samsara
⚡ Apply early Remote - Canada · location restricted CA$196,000–CA$269,500
● New 👁 Seen ✓ Applied 1d ago
DigitalOcean
Staff Forward Deployed Engineer, AI/ML
DigitalOcean
⚡ Apply early Seattle Hybrid $195,000–$239,000
● New 👁 Seen ✓ Applied 4d ago
DigitalOcean
Staff Forward Deployed Engineer, AI/ML
DigitalOcean
⚡ Apply early San Francisco Onsite $195,000–$239,000
● New 👁 Seen ✓ Applied 4d ago
Discord
Senior Software Engineer, Machine Learning (Ads)
Discord
⚡ Apply early San Francisco Bay Area Onsite
● New 👁 Seen ✓ Applied 5d ago
Kodiak
Staff Machine Learning Engineer - Data
Kodiak
⚡ Apply early San Francisco Bay Area Onsite $200,000–$265,000
● New 👁 Seen ✓ Applied 5d ago
Block
Staff Applied Machine Learning Engineer - Intelligent Data, Signals & Systems
Block
⚡ Apply early Bay Area, CA, United States of... Onsite $276,800–$415,200
● New 👁 Seen ✓ Applied 1w ago
Block
Staff Applied Machine Learning Engineer - Fraud & Abuse
Block
⚡ Apply early Bay Area, CA, United States of... Onsite $276,800–$415,200
● New 👁 Seen ✓ Applied 1w ago
Arena Intelligence, Inc.
Senior Software Engineer, ML Infrastructure
Arena Intelligence, Inc.
⚡ Apply early Bay Area Hybrid $150,000–$350,000
● New 👁 Seen ✓ Applied 1w ago

Sign up for suggestions tailored to the jobs you open and the searches you save.

Apply now
🤖

Whoa — hold up

JobsRadar was built for real people having a rough time in their job search — not for automated requests. You're clicking way too fast and you're now temporarily blocked.

Come back later. If you're genuinely job hunting, we've got your back — just act like a human.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Get the worldwide-remote edge.

Join our Telegram channel for the stuff that helps you land the role — salary benchmarks, the weekly market pulse, and new-feature drops. No spam, just signal.

Join the channel — it's free