Anthropic

Research Engineer, Takeoff Intel

Anthropic · Remote-Friendly (Travel Required) | San Francisco, CA

Full-timeStaff+

About the Role

About Anthropic Anthropic's mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About the team At Anthropic, we are delegating a growing share of AI development to AI systems themselves. Takeoff Intel is the team that measures this recursion from the inside. We're part of the Anthropic Institute. We design evaluations of AI R&D capabilities, build the internal telemetry Anthropic uses to track how much of its own model development is becoming AI-assisted, and develop the quantitative methods that turn those signals into a calibrated picture of where capability growth is heading, so that Anthropic and the wider world have accurate situational awareness on this acceleration. Our work appears in Anthropic's model system cards (we own the AI R&D capability assessments and adapted Epoch's Capabilities Index to our evals); all the data in When AI Builds Itself comes from our team. Internally, our measurements shape research priorities and safety planning; externally, they contribute to Anthropic's public reporting on the pace of AI progress and to collaborations with third-party evaluators. We're a small team that works closely with pretraining, RL, economics, and policy researchers across the company. If you're passionate about measurement accuracy, and feel urgency about safety and situational awareness, you should consider joining us. About the role As a Research Engineer on Takeoff Intel you'll build and run the evaluation and measurement instruments that make this research possible. This is a generalist role on a small team: you'll work across evals infrastructure, large-scale data processing, and analysis tooling, and you'll prioritize shipping. We build instruments that answer real questions and help set priorities, not dashboards that surface noise. We value working prototypes, rapid iteration, accuracy and good prioritization. We often need to go from a vague research question to a running instrument quickly. We're hiring at both junior and senior levels. Responsibilities • Design, build, and run capability evaluations and measurement instruments at scale • Build the data and analysis pipelines that turn large volumes of model outputs and telemetry into reliable metrics • Prototype new instruments fast, validate them, and decide what to keep • Review and supervise AI-written code as a normal part of the workflow • Work closely with research scientists on the team and with partner teams to define what's worth measuring • Contribute to internal write-ups and public reporting You may be a good fit if you • Have shipped an evaluation, data product, or research library end to end • Prototype fast and are comfortable throwing code away • Handle messy, large-volume data without over-engineering • Have run experiments on large language models, not just moved their outputs around • Can work from a vague question rather than a spec • Communicate results clearly and collaborate closely with the researchers whose questions your instruments answer Strong candidates may also have • Built evaluation harnesses or benchmark infrastructure for LLMs • Experience with large-scale ML or data infrastructure (self-driving, observability, or similar) alongside ML exposure • Built tools or libraries that other researchers rely on • A track record of catching what AI-written code gets wrong Some examples of our work • Anthropic ECI: our adaptation of Epoch Capabilities Index published in all recent system cards to measure capability acceleration • AI R&D capability assessments in the Claude system cards • When AI Builds Itself: all data in the article comes from our team The annual compensation range for this role is listed below. For sales roles, the range provided is the role's On Target Earnings ("OTE") range, meaning that the range includes both the sales commissions/sales bonuses target and annual base salary for the role. Annual Salary:$350,000—$850,000 USDLogistics Minimum education: Bachelor's degree or an equivalent combination of education, training, and/or experience Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices. Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a vi
AI Safety / LLM500-1000 employeesSan Francisco, CAFounded 2021💰 Series E

Anthropic PBC is an American artificial intelligence (AI) company headquartered in San Francisco. It has developed a family of large language models (LLMs) named Claude. Anthropic operates as a public benefit corporation, which researches and develops AI to "study their safety properties at the technological frontier" and use this research to deploy safe models for the public.

PythonPyTorchJaxTypeScriptReact
Competitive salary · Equity · Health/dental/vision

💬 Developer Questions

Ask the team a question — answers show up here

🎯

What does the interview process look like?

🤖

What AI/vibe coding tools does the team use daily?

👥

How big is the engineering team?

Is the team fully async or are there required meetings?

🚀

What does onboarding look like for remote hires?

🔧

Can you share more about the tech stack and architecture?

📈

What does career growth look like in this role?

📅

What does a typical day look like?

💰

Is there a salary range you can share?

📊

Is equity or stock options part of the package?

🌍

Are there timezone requirements or preferences?

🛂

Do you sponsor work visas?

🏢 Is this your listing? Claim it to answer questions

Similar Jobs

Helpful resources

Hiring for a similar role? Post your job here — it's free →