×
Register Here to Apply for Jobs or Post Jobs. X

Lead Data Scientist Applied AI

Job in Santa Clara, Santa Clara County, California, 95053, USA
Listing for: Socket.dev
Full Time position
Listed on 2026-09-21
Job specializations:
  • IT/Tech
    AI Engineer (Applied/Software), Machine Learning/ ML Engineer, AI Evaluation, Data Scientist
Salary/Wage Range or Industry Benchmark: 150000 - 170000 USD Yearly USD 150000.00 170000.00 YEAR
Job Description & How to Apply Below

The Role

We are seeking a Lead Data Scientist who can build advanced AI systems and demonstrate, with evidence, why they are accurate, reliable and appropriate for production. This role combines strong statistical and machine learning expertise with hands‑on experience designing and delivering production AI solutions.

You will work across Generative AI, Agentic AI, computer vision, forecasting and optimization. Your central responsibility will be to measure model performance and uncertainty, explain model behavior, identify risks and failure modes, and connect technical results to business outcomes that leaders can use to make decisions.

The ideal candidate is comfortable running structured experiments, presenting error analysis to technical and non‑technical stakeholders, and taking models from problem definition through deployment and monitoring. You should know when Generative AI is the right solution, when a simpler statistical approach is more effective, and how to support that decision with data.

What You Will Do Scientific Rigor and Model Evaluation
  • Justify modeling decisions with evidence
    :
    Frame business problems, form hypotheses, run structured experiments and select between classical machine learning, deep learning and Generative AI using measured performance, cost and risk.
  • Quantify risk and uncertainty
    :
    Measure confidence intervals, error rates, hallucination rates, bias, drift and failure modes. Define numerical production‑readiness criteria for each use case.
  • Improve explainability
    :
    Use feature attribution, SHAP, LIME, error analysis, slice analysis and evaluation dashboards to explain model behavior and limitations.
  • Translate results into business impact
    :
    Connect model metrics to revenue, cost, speed and risk. Quantify expected return, trade‑offs and downside scenarios for decision‑makers.
AI and Machine Learning Delivery
  • Design complete AI solutions
    :
    Own problem framing, data strategy, modeling, evaluation, deployment and monitoring for machine learning, LLM and agentic systems.
  • Build models for unstructured data
    :
    Develop production‑quality solutions using image, video, text, audio, sensor and time‑series data.
  • Deliver computer vision solutions
    :
    Build detection, classification, segmentation, OCR and tracking systems with measurable performance benchmarks.
  • Develop forecasting solutions
    :
    Create time‑series, demand and behavioral forecasting models with documented accuracy, error bands and integration into operational workflows.
  • Build with foundation models
    :
    Use Claude and other LLMs for prompting, fine‑tuning, systematic evaluation and integration into multi‑step, tool‑using and multi‑agent workflows with memory and guardrails.
What We Are Looking For
  • Bachelor's or Master's degree in Computer Science, Data Science, Machine Learning, Statistics or a related discipline, or equivalent practical experience.
  • A track record of approximately 10 or more AI and machine learning projects deployed to production, with clear ownership of approach selection, evaluation and risk assessment.
  • Strong statistical and algorithmic foundations, including hypothesis testing, experiment design, evaluation methodology and uncertainty quantification.
  • Hands‑on experience with RAG, tool‑use patterns and agentic frameworks such as Lang Graph, Llama Index, Auto Gen or CrewAI, along with Model Context Protocol where relevant.
  • Experience evaluating and protecting LLM and agentic systems through guardrails, systematic testing, hallucination measurement and failure analysis.
  • Practical experience with explainability, bias and fairness assessment, model monitoring and drift detection.
  • Familiarity with MLOps practices, including experiment tracking, CI/CD for machine learning, model registries and production monitoring.
  • Exper…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary