More jobs:
AI Research Scientist – LLMs; Remote
Remote / Online - Candidates ideally in
San Jose, Santa Clara County, California, 95199, USA
Listed on 2026-07-29
San Jose, Santa Clara County, California, 95199, USA
Listing for:
Synthires
Remote/Work from Home
position Listed on 2026-07-29
Job specializations:
-
Research/Development
Data Scientist, AI Business & Operations, AI Evaluation
Job Description & How to Apply Below
LLM Research Scientist (Pre-training & Post-Training)
Location: Remote
Engagement Type: Hourly Contract
Compensation: $100–$120/hour
About the OpportunityThis opportunity is for experienced Machine Learning Researchers and LLM Research Scientists interested in contributing to advanced AI research and evaluation projects. The role focuses on training, fine‑tuning, and improving large language models (LLMs), conducting empirical research, and advancing state-of-the‑art foundation model capabilities.
You will work on challenging research problems spanning LLM pre-training, post-training, data curation, model evaluation, alignment, and optimization
.
- Train transformer-based language models from scratch and fine-tune open-weight foundation models.
- Design and optimize pre-training and post-training pipelines for large language models.
- Construct, curate, and optimize large-scale training datasets from raw web and other data sources.
- Develop data filtering, deduplication, quality classification, and curriculum learning strategies.
- Diagnose and resolve optimization failures, convergence issues, and training instabilities.
- Build and evaluate supervised fine-tuning (SFT), preference optimization (DPO, RLHF, RLAIF), and reward modeling pipelines.
- Design evaluation benchmarks and analyze model performance using rigorous experimental methodologies.
- Collaborate with AI researchers to improve model reasoning, alignment, efficiency, and overall performance.
- 3+ years of machine learning research experience (PhD research qualifies).
- Strong expertise in one or more of the following areas:
- LLM Post-Training
- Reinforcement Learning for LLMs
- Model Alignment and AI Safety
- LLM Evaluation and Benchmark Development
- Strong experience with PyTorch, JAX, Tensor Flow
, or similar machine learning frameworks. - Deep understanding of transformer architectures, large language models, optimization techniques, and modern deep learning methodologies.
- Excellent analytical, research, and scientific communication skills.
- Ability to work independently in a remote research environment.
- PhD in Computer Science, Machine Learning, Artificial Intelligence, Natural Language Processing, or a related field.
- Degree from a Top-100 university
, experience at a FAANG or leading AI company, or an equivalent research track record through publications or impactful open‑source contributions. - Experience with:
- Scaling Laws
- LLM Evaluation
- AI Alignment
- AI Safety Research
- Publications in leading AI conferences or significant open‑source contributions.
- Competitive compensation of $100–$120/hour.
- Weekly payments.
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
Search for further Jobs Here:
×