Research Engineer, Applied Finetuning
New York, New York County, New York, 10261, USA
Listed on 2026-07-01
-
Software Development
Machine Learning/ ML Engineer, AI Engineer (Applied/Software)
As a Research Engineer or Research Scientist in Applied Fine tuning, you will directly train the models we launch to the public via Claude.
AI and our API. In this role, you will design and iterate on state‑of‑the‑art fine tuning techniques, such as Constitutional AI and RLHF, to train our production Claude models. You will implement new algorithms, run experiments on data mixes, design evaluations, and improve our production model training pipeline. This role offers the opportunity to contribute to cutting‑edge research while also having a direct and measurable impact on the company’s success.
- Implement and optimize fine tuning pipelines to efficiently train production‑scale language models with techniques like Constitutional AI
- Develop novel prompts and prompting strategies to improve and test model behaviors
- Collaborate with other research teams to translate novel fine tuning techniques into our production model training process, ensuring models are helpful, honest, and harmless
- Design and run a new evaluation that tests Claude’s reasoning capabilities
- Collaborate with a research team to develop a robust evaluation for a new model capability they are developing
- Stay current with state‑of‑the‑art research in AI and machine learning, and propose ways to apply these advancements to production systems
- Have significant Python programming experience and machine learning experience
- Are results‑oriented, with a bias towards flexibility and impact
- Pick up slack, even if it goes outside your job description
- Enjoy pair programming
- Want to learn more about machine learning research
- Care about the societal impacts of your work
- Have clear written and verbal communication
- Fine‑tuning large language models with supervised learning or reinforcement learning
- Developing evaluations for language models
- Complex shared codebases and RL infrastructure
- Authoring research papers in machine learning, NLP, or AI alignment or similar industry experience
Annual Salary: $315,000—$510,000 USD
Location-based hybrid policy:
Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices.
US visa sponsorship:
We do sponsor visas! However, we aren’t able to successfully sponsor visas for every role and every candidate; operations roles are especially difficult to support. But if we make you an offer, we will make every effort to get you into the United States, and we retain an immigration lawyer to help with this.
Anthropic’s compensation package consists of three elements: salary, equity, and benefits. We are committed to pay fairness and aim for these three elements collectively to be highly competitive with market rates.
Equity - For eligible roles, equity will be a major component of the total compensation. We aim to offer higher‑than‑average equity compensation for a company of our size, and communicate equity amounts at the time of offer issuance.
US Benefits- Optional equity donation matching.
- Comprehensive health, dental, and vision insurance for you and all your dependents.
- 401(k) plan with 4% matching.
- 22 weeks of paid parental leave.
- Unlimited PTO – most staff take between 4‑6 weeks each year, sometimes more!
- Stipends for education, home office improvements, commuting, and wellness.
- Fertility benefits via Carrot.
- Daily lunches and snacks in our office.
- Relocation support for those moving to the Bay Area.
- Optional equity donation matching.
- Private health, dental, and vision insurance for you and your dependents.
- Pension contribution (matching 4% of your salary).
- 21 weeks of paid parental leave.
- Unlimited PTO – most staff take between 4‑6 weeks each year, sometimes more!
- Health cash plan.
- Life insurance and income protection.
- Daily lunches and snacks in our office.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).