×
Register Here to Apply for Jobs or Post Jobs. X

Applied Scientist - LLM, Alexa Conversational Modelling Intelligence

Job in Berlin, Coos County, New Hampshire, 03570, USA
Listing for: Amazon
Full Time position
Listed on 2026-07-20
Job specializations:
  • IT/Tech
    Machine Learning/ ML Engineer, AI Engineer (Applied/Software), Data Scientist
Salary/Wage Range or Industry Benchmark: 150000 - 230000 USD Yearly USD 150000.00 230000.00 YEAR
Job Description & How to Apply Below

Overview

As an Applied Scientist II in the Alexa Conversational Modelling Intelligence team within Alexa AI, you will drive model post-training for Large Language Models that power Alexa+. You will adopt and adapt state-of-the-art techniques—including supervised fine-tuning, reinforcement learning, preference optimization, and knowledge distillation—running rigorous experiments and translating findings into production-ready solutions that directly improve the customer experience for millions of users worldwide.

You will own the full model development cycle from data curation through training, evaluation, and deployment. Your day-to-day will involve developing evaluation methods and metrics, diagnosing model defects, optimizing model training pipelines, and iterating on recipes to move concrete quality and efficiency benchmarks. You ll write clean, reproducible code, contribute to shared tooling, and collaborate closely with scientists and engineers to bring models from experimentation to scale.

You are technically curious, experiment-driven, and motivated by real customer impact. You are an expert in LLM post-training. You will also advance the state of the art by publishing at top-tier NLP/ML conferences (ACL, EMNLP, NeurIPS, ICML, ICLR) — contributing to the broader research community while grounding your work in measurable outcomes.

Key responsibilities
  • Own the full model development cycle — from data curation through training, evaluation, and deployment.
  • Develop and apply post-training techniques: supervised fine-tuning, reinforcement learning, preference optimization, and knowledge distillation.
  • Build evaluation methods and metrics, and diagnose model defects to target the highest-impact improvements.
  • Optimize model training pipelines and iterate on recipes to move concrete quality and efficiency benchmarks.
  • Write high-quality documentation on methods and experiment outcomes, and communicate findings clearly to stakeholders.
A day in the life

Post-training is one of the most active frontiers in LLMs right now. The field has moved from scaling pretraining to getting more out of models afterward through RL, reasoning recipes, and preference optimization. You ll work on these techniques directly, on a product used by millions of customers every day. A typical day: review overnight training runs and dashboards, dig into model defects to form hypotheses, then curate data and iterate on a recipe, improving shared tooling along the way.

You ll sync with scientists and engineers to unblock the path to production, and write up your findings for stakeholders. It s fast-moving — a good idea can reach millions of customers within weeks.

About the team

The Alexa Conversational Modelling Intelligence team builds industry-leading LLM-based conversational technologies that customers love. Our mission is to push the envelope in LLMs for Alexa to deliver the best-possible customer experience. As an Applied Scientist, you ll contribute directly to that mission through model development and experimentation.

Qualifications Basic Qualifications
  • Have publications on top-tier conferences, such as CVPR, ICCV, ECCV or NeurIPS
  • Experience working with large, complex data sets
  • Experience working effectively with science, data processing, and software engineering teams
  • Experience in written and verbal communication skills to communicate with technical and non-technical audiences, including senior leadership
  • Experience building and deploying LLM solutions in production or at scale
  • Hands-on experience with Large Language Models training and fine-tuning via pre-training, SFT, and/or RLHF/preference optimization
  • Experience with LLM evaluation — building benchmarks, LLM-as-a-judge, or defect/quality analysis
  • Familiarity with modern training/inference infrastructure (e.g., distributed training, RL frameworks, model serving)
Preferred Qualifications
  • Experience in published research at top conferences
  • Experience with large, complex data sets
  • Experience working with science, data processing, and software engineering teams
  • Experience in written and verbal communication skills to communicate with technical and non-technical audiences, including…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary