×
Register Here to Apply for Jobs or Post Jobs. X

Data Scientist

Job in San Francisco, San Francisco County, California, 94199, USA
Listing for: Happyrobot Inc.
Full Time position
Listed on 2026-07-19
Job specializations:
  • IT/Tech
    Data Analyst, AI Engineer (Applied/Software), Machine Learning/ ML Engineer, Data Scientist
Salary/Wage Range or Industry Benchmark: 150000 - 190000 USD Yearly USD 150000.00 190000.00 YEAR
Job Description & How to Apply Below

About Happy Robot

Happy Robot is the infrastructure for enterprises to build and orchestrate AI work forces. Our AI workers don't just communicate - they make decisions, take action, and run operations autonomously across voice, email, and enterprise systems. Born in Y Combinator (S23) and backed by a16z and Base
10 with over $60M raised, we power critical operations for global enterprises worldwide.

Our platform is battle-tested in the most demanding environments - where AI has real consequences. We started in logistics, built our own voice stack, models, and orchestration layer from the ground up, and are now bringing that infrastructure to every enterprise that runs the real economy. Learn more about our vision in our manifesto.

About the Role

You’ll help make data a core part of how we build and improve Happy Robot’s products.

You’ll work closely with Product, Engineering, and Machine Learning teams to measure how changes to our models, agents, and product features affect real-world performance. You’ll define meaningful metrics, design experiments, and conduct deeper analyses to understand how our agents create value for clients.

Your work will range from evaluating A/B tests and model changes to analyzing millions of conversations and workflows. You’ll turn complex data into clear insights that influence our product and ML roadmaps.

What You’ll Do
  • Define and track product, feature, and agent-level metrics.

  • Design, run, and interpret A/B tests for model changes, prompts, agent behavior, workflows, and product features.

  • Measure how the performance of our agents affects client outcomes, such as task completion, operational efficiency, response quality, and automation rates.

  • Connect offline model evaluations with production performance and real-world customer impact.

  • Conduct deep analyses across conversations, workflows, and product usage to identify opportunities and explain differences in performance.

  • Investigate anomalies and regressions, perform root-cause analyses, and recommend improvements.

  • Build statistical models, simulations, and analytical frameworks to support product and ML decisions.

  • Partner with Engineering to improve instrumentation, data quality, experimentation systems, and analytical data models.

  • Build dashboards and self-serve tools that help teams understand product and agent performance.

  • Communicate findings and recommendations clearly to technical and non-technical stakeholders.

Must Have
  • 4+ years of experience in Data Science, Product Analytics, or another highly quantitative product role.

  • Strong experience with experimental design, A/B testing, statistics, causal inference, and hypothesis-driven analysis.

  • Advanced proficiency in SQL and Python.

  • Experience defining and operationalizing product and feature metrics.

  • Ability to translate ambiguous product questions into rigorous analyses and actionable recommendations.

  • Strong product instincts and the ability to distinguish statistical significance from meaningful product or customer impact.

  • Experience partnering closely with Product, Engineering, or Machine Learning teams.

  • Strong written and verbal communication skills.

  • High attention to detail and commitment to analytical accuracy.

  • Founder mindset: ownership, independence, curiosity, and willingness to go deep.

Nice to Have
  • Experience working with large language models, AI agents, generative AI, or other probabilistic ML products.

  • Experience measuring the production impact of model, prompt, retrieval, or orchestration changes.

  • Familiarity with ML evaluation systems and the relationship between offline evaluations and online metrics.

  • Experience analyzing conversational, NLP, speech, or other unstructured data.

  • Experience with enterprise or B2B products.

  • Experience combining quantitative analysis with qualitative methods such as conversation reviews, customer feedback, surveys, or user research.

  • Familiarity with modern analytics infrastructure, data warehouses, experimentation platforms, and business intelligence tools.

  • Prior experience in a fast‑growing startup or other highly ambiguous environment.

Why join us?
  • Opportunity to work at a high-growth AI startup
    , backed by top investors.

  • Rapidly growing and…

To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary