Loading your workspace…
Rounds/Retell/Research Scientist - LLM·#23020
Live · accepting submissions
Log in to see accurate information
You're viewing this role as a guest. Sign in to see your application status, referral credits, and personalized match details.
Retell logo
AI·Series A·Redwood City, California

Research Scientist - LLM

at RetellAI
Location
Redwood City, California
Salary
$225,000 - $400,000
Type
Full-Time
About Retell
Retell AI is reimagining the call center with first-principles voice AI. The platform powers AI voice agents for thousands of companies, automating sales, support, and logistics calls that once required large human teams. Backed by Y Combinator, Alt Capital, and Carya Venture Partners, Retell claims $80M ARR with a team of 50 (the founder's framing), and is valued at over $1.5B. Public sources indicate $60M ARR and a team size between 100–500, but the founder's version is the one used in candidate-facing copy. Customers include CVS/Aetna, American Airlines, Lenovo, and Grab. The company has been recognized as #1 Best Places to Work in the Bay Area (San Francisco Business Times 2026), Top 50 AI Apps (a16z, 2025), and #3 Fastest-Growing Software Company (G2 Best Software Awards 2026). The founding team brings experience from Stripe payments and IMO medalist backgrounds. The vision: build a modern CX platform where entire contact centers are powered by AI "workers"—frontline agents, QA analysts, and managers—continuously executing, monitoring, and improving every customer interaction. Retell is positioned as one of the fastest-growing voice AI companies globally, with a sharp focus on real-world impact and technical depth.
Series A • 51 – 200
Stage & size
AI
Industry
2023
Founded
About This Role

The mandate: Advance the LLM side of Retell's voice agents — reasoning, latency, and conversational quality in real-time systems. Explore new techniques, train and iterate on models, design evaluation frameworks for subjective conversational quality, and get the results into production. The JD body is identical to the Audio req. The differentiator is the specialization axis: this seat is the LLM-side hire — post-training, reasoning, instruction following, agentic behavior, inference efficiency. The Audio seat is the speech-side hire — ASR, TTS, streaming audio. Recruiters should source these as two distinct pools with modest overlap at multimodal. Both are founding research seats. There is no research org yet; whoever takes this defines what LLM research means at Retell. What You'll Own

  • Research and experimentation — new techniques across LLMs and audio models targeting reasoning, latency, and conversational quality in real-time systems

  • Model training — build and iterate on models and pipelines; innovate on training paradigms, methods, and inference

  • Evaluation and benchmarking — novel eval frameworks, datasets, and metrics for complex real-world voice tasks

  • Research-to-production — work directly with engineering to deploy findings

  • Human feedback loops — methods for folding human evaluation into model improvement, especially on subjective conversational quality

  • Frontier tracking — bring new ideas into Retell's product and infrastructure Requirements Hard gates:

  • Master's in CS, ML, AI, or related required. PhD preferred. Equivalent research-level engineering experience considered — bypass the degree only for a genuinely research-grade portfolio.

  • Advanced ML research: LLM pre-training or post-training, transcription model training, TTS, or multimodal systems. Industry or academia.

  • PyTorch fluency, model architecture depth, and the underlying math. Round 4 tests all three live.

  • On-site in Redwood City — relocation fully covered Profile:

  • Goes from open-ended problem to working prototype without a spec

  • Translates research into systems that survive production

  • Communicates complex ideas cross-functionally Strong bonuses:

  • First- or co-author publications at NeurIPS, ICML, ICLR, ACL, EMNLP, COLM, or equivalent. For this req specifically, NeurIPS/ICML/ICLR/ACL/COLM carry more signal than Interspeech.

  • Post-training depthRLHF, DPO, RLAIF, preference modeling, instruction tuning, reward modeling. This is the highest-value specialization for the role.

  • Inference efficiency — speculative decoding, KV cache optimization, quantization, distillation, streaming generation. Latency is a first-class constraint here.

  • Agentic LLM work — tool use, function calling, multi-turn state, long-horizon task execution

  • LLM evaluation research, especially for subjective or open-ended quality

  • Competition awards Location and visa:

  • On-site, Redwood City. 100% relocation provided.

  • Sponsorship: Yes — H-1B, TN, L-1, E-3, F-1 (OPT/CPT), and O-1. O-1 is listed on this req. Lead with it for published international researchers — most startups won't touch it. Anti-patterns

  • ML engineers who fine-tune off-the-shelf models and call it research. The PyTorch and theory round exposes this in ten minutes.

  • Prompt engineers with a "LLM researcher" title

  • Pure academics with no interest in production or weak engineering — the Backend + AI Practical round is real

  • Researchers who need a large team, mature infrastructure, and a handed-down agenda

  • Anyone who only wants to publish. Retell isn't a lab.

  • Remote requirements — no exception surfaced Who Will Thrive Here

Someone with real LLM research credentials who's tired of waiting in the compute queue and the publication cycle at a big lab. They want their post-training run in front of 50M calls next month. They'll derive the objective and then debug the inference server. Ex-frontier-lab researchers wanting founding scope, and strong NLP/ML PhDs who want production stakes and a latency constraint that makes the problem harder, are the two clearest profiles.

Job Details
Experience
3+ Years
Salary
$225,000 - $400,000
Visa Sponsorship
Yes
Employment Type
Full-Time
Work Arrangement
In office
Work Intensity
9-9-5
Benefits & Perks
$70/day Doordash Credit For Unlimited Meals And Snacks
$200/month Wellness Reimbursement
$300/month Commuter Reimbursement
$75/month Phone Bill Reimbursement
$50/month Internet Reimbursement
100% Relocation Provided
Green Flags
First/co-author publications at top ML/AI conferences
Prior experience shipping ML models to production in a startup or high-growth environment
Deep PyTorch expertise and comfort with model internals
Red Flags
No hands-on experience with advanced ML research (LLMs, TTS, multimodal, etc.)
Lacks a graduate degree or equivalent research-level engineering experience
No evidence of bridging research to production (pure academic/theory background only)
Ideal Companies
O
OpenAI
G
Google DeepMind
A
Anthropic
M
Meta AI
M
Microsoft Research
A
Amazon Alexa AI

Required Candidate Q&A

Question 1
Are you able to relocate to Redwood City, CA and work on-site full-time? (100% relocation provided.)
Question 2
What is your current visa status, and do you require sponsorship? What is your earliest start date?
Question 3
Walk through a research project (LLM, TTS, or multimodal) you led or co-authored—what was the core technical challenge and how did you solve it?
Question 4
Describe a time you translated a research idea into a production system. What trade-offs did you make?
Question 5
Have you published at top-tier ML/AI conferences or won notable ML competitions? Please specify.
Question 6
How do you design evaluation frameworks for open-ended or subjective ML tasks?
Question 7
Why Retell—what about real-time voice AI and this team pulls you?

Candidate scorecard

· 10 criteria
Advanced ML research background (LLM pre-training/post-training, transcription, TTS, or multimodal systems)
Deep technical foundation in PyTorch, model architectures, and ML math
Master's degree in CS, ML, AI, or related field required; PhD preferred (or equivalent research-level experience)
Rounds · Confidential to recruiting partners · Last refreshed Aug 21, 2026