×
Register Here to Apply for Jobs or Post Jobs. X
More jobs:

AI Research Scientist

Job in Ova, Magoffin County, Kentucky, USA
Listing for: 1P284 THE CARLYLE GROUP EMPLOYEE CO., LLC
Full Time position
Listed on 2026-08-10
Job specializations:
  • Research/Development
    AI Evaluation
  • IT/Tech
    AI Evaluation
Salary/Wage Range or Industry Benchmark: 190000 - 220000 USD Yearly USD 190000.00 220000.00 YEAR
Job Description & How to Apply Below
Location: Ova

Position Summary

We are seeking an AI Research Scientist to lead the development of evaluation frameworks that measure and improve the reasoning capabilities of Large Language Models (LLMs) for investment decision-making. This individual will partner closely with investment professionals, AI researchers, and AI engineers to capture expert reasoning, design rigorous benchmarks, build high-quality evaluation datasets, and develop methodologies that guide model training, post-training, and deployment.

This role sits at the intersection of AI research, investment domain expertise, and applied machine learning, with a primary focus on ensuring our models reason accurately, consistently, and reliably across complex private market investment workflows. This individual should:
Think like an AI researcher but build practical systems. Rigorously evaluate model reasoning instead of relying solely on benchmark scores. Enjoy creating datasets, experiments, and evaluation methodologies. Work effectively with domain experts to extract tacit knowledge. Independently drive ambiguous research problems from concept through implementation.

ResponsibilitiesLLM Evaluation & Benchmark Development
  • Design and implement comprehensive evaluation frameworks for investment-focused LLMs.
  • Develop quantitative and qualitative metrics to measure reasoning quality, factuality, consistency, calibration, and investment decision quality.
  • Build representative benchmark datasets covering real-world investment workflows.
  • Design test suites that measure model performance across varying investment scenarios and edge cases.
  • Run and analyze evaluation results to identify opportunities to improve model performance.
Data Collection & Dataset Development
  • Develop strategies for collecting, curating, annotating, and maintaining high-quality training and evaluation datasets.
  • Translate investment reasoning into structured datasets suitable for training and evaluation.
  • Define annotation guidelines and quality assurance processes.
  • Partner with investment professionals to capture expert reasoning and decision-making processes.
Post-Training & Model Improvement
  • Support supervised fine-tuning (SFT), preference optimization (DPO/RLHF), reinforcement learning, and other post-training methodologies.
  • Design evaluation loops that measure gains from post-training efforts.
  • Identify model failure modes and recommend improvements.
  • Collaborate with AI engineers to deploy evaluation pipelines into production workflows.
Cross-Functional Collaboration
  • Work closely with investment teams to understand investment processes and reasoning.
  • Partner with AI engineers to integrate evaluation systems into model development pipelines.
  • Collaborate with AI researchers on new evaluation methodologies and experimental designs.
  • Communicate research findings and recommendations to technical and non-technical stakeholders.
Minimum Qualifications
  • M.S. or Ph.D. in Computer Science, Machine Learning, Artificial Intelligence, Statistics, or related field (or equivalent experience).
  • Experience developing LLM evaluation frameworks, benchmarks, or experimentation methodologies.
  • Experience collecting, curating, annotating, and managing datasets for AI systems.
  • Strong understanding of model evaluation metrics, experimental design, and statistical analysis.
  • Experience with SFT, DPO/RLHF, reinforcement learning, or related post-training techniques.
  • Strong Python programming skills and experience with modern ML frameworks.
  • Experience working with foundation models and modern LLM tooling.
  • Excellent communication and cross-functional collaboration skills.
Preferred Qualifications
  • Experience applying AI to finance or investment research.
  • Experience building human preference datasets.
  • Research in NLP, AI alignment, LLM evaluation, or reasoning.
  • Publications in leading AI conferences or journals.
  • Ability to rapidly learn complex technical and business domains.
Benefits/Compensation
  • The anticipated base salary range for this role is $190,000 to $220,000.
  • Retirement benefits
  • Health insurance
  • Life insurance and disability
  • Paid time off
  • Paid holidays
  • Family planning benefits
  • Various wellness programs
  • Annual discretionary incentive…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary