Anthropic
Research Engineer, Computer Use
About this role
Anthropic is seeking a Research Engineer for its Computer Use team to enhance AI models' capabilities in operating software reliably and safely. The role combines research and product development, ensuring improvements impact both internal and customer applications.
What you'll do
- Conduct experiments to enhance AI perception and capabilities
- Develop evaluation frameworks for complex software tasks
- Create reinforcement learning training environments
- Build tools for testing RL environments
- Collaborate with teams to enhance training setups
- Work with product teams to implement research findings
What they're looking for
- Proficiency in Python
- Experience in machine learning model training
- Strong communication abilities
- Interest in the societal impact of AI
- Experience in reinforcement learning
- Ability to develop multimodal training
- Familiarity with building evaluation benchmarks
- Collaboration with product teams
Benefits
- Annual salary range of $500,000 - $850,000
- Visa sponsorship available
- Hybrid work policy
- Inclusive work environment
- Opportunities for professional growth
- Focus on societal impacts of AI
Opens the application — the Jobs AI extension fills it for you. Set up autofill
Opens the official application on the employer’s site. No login required.
Anthropic
Anthropic builds Claude, an AI assistant, and is hiring for engineering roles across infrastructure, data systems, and security that support both AI research operations and the company's internal technology needs. The company seeks infrastructure engineers, systems integrators, data scientists, and security specialists to build production-scale systems for training data pipelines, financial operations, developer productivity measurement, research infrastructure, and server firmware security.
- Website
- anthropic.com
Likely interview questions
- Can you walk us through a project where you trained or fine-tuned a machine learning model? What challenges did you face and how did you overcome them?
- Describe your experience building evaluation frameworks or benchmarks. How do you measure success when evaluating complex model behaviors?