Skip to main content

Research Engineer, Visual Knowledge Work

Anthropic
Hybrid - NYC, SF, or Seattle, at least 25% in officeUpdated 1d ago
Compensation
$350k–$850k
Published range · Top quartile for Engineering (582 listings)
Location
Hybrid - NYC, SF, or Seattle, at least 25% in office
Remote eligibility
Employment
Full-time
Mid-level
Role family
Engineering
AI / ML
Apply on job-boards.greenhouse.io
Job actionsApply now
Job actionsApply now

About the job

Anthropic's mission is to create reliable, interpretable, and steerable AI systems. We are looking for a research engineer to join the Vision team, owning the end-to-end process of creating training data and RL environments targeting visual knowledge work.

What you'll do

  • Own the data strategy for vision capabilities end-to-end, from building evals and scaling RL environments.
  • Manage technical relationships with external data vendors, including writing task specifications, evaluating visual data and annotation quality, and iterating on reward design.
  • Develop and improve QA frameworks that catch reward hacking and ensure environment quality at scale.
  • Run generalization experiments to measure how data strategy changes improve multimodal capabilities on held-out evaluations.
  • Partner with pretraining, RL, and product teams.

You may be a good fit if you

  • Have 7+ years of ML, computer vision, and software engineering experience.
  • Have experience with reinforcement learning, reward design, or training data curation for large language or vision-language models.
  • Are familiar with the architecture, training, and operation of large vision language models.
  • Are comfortable managing technical vendor relationships and iterating quickly on feedback.
  • Are results-oriented, with a bias towards flexibility and impact.
  • Care about the societal impacts of your work.

Strong candidates may also have

  • Experience designing evals or benchmarks for LLMs or vision language models.
  • Large-scale pretraining, SL, and RL on language models.
  • Deep learning research on images, video, or other modalities.
  • Developing complex agentic systems using LLMs.
  • Large-scale ETL and data pipeline development.

Representative projects

  • Writing a vendor-facing specification for a new family of visual RL training tasks.
  • Running experiments to determine ideal training datamixes and parameters for a synthetically generated vision dataset.
  • Finetuning Claude to maximize its performance using a particular set of agent tools/skills.

Compensation

Annual Salary: $350,000—$850,000 USD.

Logistics

Location: New York City, NY; San Francisco, CA; Seattle, WA. Hybrid policy: all staff in one of our offices at least 25% of the time. Minimum education: Bachelor’s degree or equivalent. Required field: relevant to the role. Minimum years of experience: correlates with internal job level. Visa sponsorship: we do sponsor visas.

Skills & tags

What you can verify before applying

Compare the essentials before you leave: pay, remote scope, employment type, source, and the employer apply destination.