Baseten Applied AI
Post-Training Research Scientist
- Location San Francisco
- Compensation 210,000 – 285,000 USD / year
- Type FullTime
- Team EPD
- Posted 2026-03-17
Original posting ↗ You apply on the company site — we never collect applications.
Extracted automatically from public job postings. Always verify details on the original posting before applying.
Hard requirements to check first
- Clearance:not mentioned in the posting
- Work auth:not mentioned in the posting
Skills
SparkAgents
Excerpt from the original posting
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE This role sits at the frontier of our research agenda. You will pursue open problems at the intersection of post-training methodology and performant inference, and then collaborate with research engineering to translate findings into production systems. A meaningful portion of your time will be dedicated to research that deepens our understanding of how models learn, alignment, and architectural efficiency — questions that may not have immediate product application. The remainder will be directed toward research that solves concrete problems for Baseten's platform and customers, who are the fastest growing AI companies in the world like Cursor, Lovable, and Notion. We are looking for someone with sharp research taste and genuine creative instinct for problem selection. Someone who can identify questions that matter, design clean experiments to answer them, and push the state of the art. The environment here is not theoretical, but rather research that can be validated with eager customers who are serving billions of tokens a second. The Manifesto: https://labs.baseten.co/manifesto RECENT RESEARCH - Towards infinite context windows: neural KV cache compaction - Dense, on-policy or both? - Repeated kv cache for long-running agents - Distillation without the dark – replicating black-box on-policy distillation on Baseten RESPONSIBILITIES - Define and pursue a research agenda spanning both foundational and applied work, with the applied component connected to Baseten's platform and customer needs. - Design and execute rigorous experiments, frequently at meaningful scale (multi-node, trillion parameter models). - Work with customers to translate domain-specific requirements into research problems, where relevant to your agenda. - Publish at top venues (NeurIPS, ICML, ICLR) and establish Baseten's research presence. - Collaborate with model performance and training infrastructure teams to…