📍
San Francisco, CA

Machine Learning Research Scientist/ Engineer, Agents

No experience
Technology & Digital
Software engineering
Posted:
December 29, 2025

Scale

Data labelling and model evaluation platform
72.7
Palpable Score
Apply >view company >

About Scale

At Scale AI, our mission is to accelerate the development of AI applications. For 8 years, Scale has been the leading AI data foundry, helping fuel the most exciting advancements in AI, including: generative AI, defense applications, and autonomous vehicles. With our recent Series F round, we’re accelerating the abundance of frontier data to pave the road to Artificial General Intelligence (AGI), and building upon our prior model evaluation work with enterprise customers and governments, to deepen our capabilities and offerings for both public and private evaluations.

About the ACE team

The Agent Capabilities & Environments (ACE) team, part of Scale’s Research organization, brings together customer-facing Researchers and Applied AI Engineers. Our core mission includes research on agent environments and RL reward signals, benchmarking autonomous agent performance across real-world scenarios and environments, creating robust data programs to improve Large Language Models (LLMs) agentic capabilities and building foundational tools and frameworks for evaluating models as agents. ACE focuses on autonomous agents that dynamically interact with diverse external environments, including code repositories, GUI interfaces, browsers, and more.

About This Role

This role is at the intersection of cutting-edge AI research and practical application, with a focus on studying the data types essential for building state-of-the-art agents, such as browser and SWE agents. The ideal candidate will explore the data landscape needed to advance intelligent, adaptable AI agents, guiding the data strategy at Scale to drive innovation. This position requires not only expertise in LLM agents and planning algorithms but also creativity in addressing novel challenges related to data, interaction, and evaluation. You will contribute to impactful research publications on agents, collaborate with customer researchers, and work alongside the engineering team to translate these advancements into real-world, scalable solutions.

Ideally you’d have:

Nice to have:

About the company

Scale

Company overview
Scale builds data infrastructure and tooling used to train, evaluate, and deploy AI systems, including work tied to RLHF, model evaluation, and enterprise AI workflows. Scale sells products like the Scale Data Engine and supports both private-sector and government customers building AI applications. Scale positions the company around “reliable AI systems” and operational excellence alongside software. Scale also runs a large set of roles across engineering, applied AI, operations, and go-to-market teams tied to AI delivery.

Locations and presence

Scale lists San Francisco as the headquarters and commonly hires into hubs like San Francisco and New York, with some roles also listing Seattle. Scale’s careers pages tend to specify location on each role rather than publishing a single, company-wide remote or hybrid policy in one place.

Palpable Score

72.7
/ 100
Scale has real early-career entry points through a dedicated university hub, recurring intern and new grad roles, and a named new grad program for Strategic Projects. Scale is better than many AI startups on transparency, with a published SWE hiring flow and a salary band on at least one flagship new grad posting. Early-career outcomes and stability are the main limiters because public signals point to high intensity and the company has had recent layoffs, while early-career conversion and promotion metrics are not published.
view full company profile >

Related jobs

📍
Needham, MA
SharkNinja
Supply Planner
January 26, 2026
view job >
📍
Needham, MA
SharkNinja
Fall 2026: APAC NPD Commercial Readiness Co-op (July to December)
January 26, 2026
view job >
📍
Needham, MA
SharkNinja
Fall 2026: Brand Marketing Co-op, Ninja (July to December)
January 26, 2026
view job >
📍
Needham, MA
SharkNinja
Fall 2026: Brand Marketing Co-op, Shark (July to December)
January 26, 2026
view job >