Research Engineer, Computer Use
Anthropic · San Francisco, CA | New York City, NY | Seattle, WA · AI Research & Engineering · listed June 30, 2026
The shape of it
Seniority
Not stated
Where
Hybrid
Stated pay
$500,000 – $850,000 USD
Requirements listed
4
Length
888 words
In the posting’s own words
The Computer Use team focuses on teaching Claude to see, use, and understand computer interfaces. As a Research Engineer on the team, you'll work on advancing our models' ability to reliably and safely operate real software. We're looking for someone who's genuinely excited about both the research and the product sides of computer use.
What it asks for · 4
- Software engineering experience and proficiency in Python
- Experience training, fine-tuning, or evaluating machine learning models
- Strong communication skills and a collaborative working style
- Care about the societal impacts and safety of your work
Also a plus
- Experience training models for computer use or other agentic capabilities
- Experience with reinforcement learning, particularly in long-horizon or sparse-reward settings
- Familiarity with multimodal model training
- Experience building evaluations or benchmarks for agentic systems
- Experience building reinforcement learning environments, simulation systems, or large-scale ML infrastructure
- Experience working closely with product teams to drive model improvements
What the job covers
- Design and run experiments to improve Claude's perception and agentic capabilities
- Develop robust, reliable evaluation frameworks for measuring our models' ability to complete complex computer tasks
- Build and improve computer use and vision reinforcement learning training environments
- Create pipelines and tools to test and validate complex RL environments
- Collaborate with teams across the model training and infrastructure stack to improve our production training setup
- Partner with product teams to bring research advances into production
Tools and skills named
Models & research
- Reinforcement learning3×
- Machine learning2×
- Evaluations
- Fine-tuning
Languages
- Python
Words the posting leans on
- model8×
- computer7×
- experience7×
- training6×
- claude5×
- product4×
- improve3×
- reinforcement learning3×
- research3×
- agentic capabilities2×
- complex2×
- evaluation2×
- experience building2×
- experience training2×
- infrastructure2×
- model improvements2×
Counted from the posting after the mission statement and the legal notices are set aside. The ones near the top are the ones a screener is looking for.
The posting, your resume, and the gaps between them. One click loads all three.
More open at Anthropic
- Account Executive, AI NativeNew York City, NY; San Francisco, CA | New York City, NY
- Account Executive - DNBSingapore
- Account Executive, Public SectorSydney, Australia
- Account Executive - Public Sector (ASEAN)Singapore
- Account Executive, StartupsSan Francisco, CA | New York City, NY
- Accounting, Revenue Internal ControlsSan Francisco, CA | Seattle, WA