Cartesia · *HQ - San Francisco, CA

Applied Researcher, Audio Post-Training at Cartesia — *HQ - San Francisco, CA

Full-time*HQ - San Francisco, CA$200,000–$350,000/yearPosted 2026-07-20Apply on Ashby

Full job description

About Cartesia

About the Role

On the Audio Post-Training team, you’ll be building and improving the capabilities that define how the rest of the world interacts with our generative audio models. This team is where customer needs meet research, and covers the full spectrum of modeling from ideation through productionization. On any given day, you might design evaluations to reliably measure new capabilities, build processing pipelines to improve data quality, experiment with finetuning and reinforcement learning approaches to refine model behavior, and more.

This role is broad, and cross functional. Members of this team should combine broad research experience with a deep care about building to solve for customer needs and a strong sense of end-to-end ownership. You should be excited to synthesize customer complaints into a holistic understanding of capability gaps, to drive research efforts across data, model training, and evaluation to close those gaps, and to communicate those improvements to product and customer stakeholders. Ultimately, you will be responsible for creating the model experience that the rest of the world sees.

Your Impact

  • Collaborate with product teams to understand and prioritize customer asks
  • Cut through the ambiguity of vaguely described behavioral problems to make concrete research plans.
  • Ideate and experiment across the full modeling stack, including data processing, synthetic data, SFT, RL, and evals to solve for high priority model capabilities
  • Root cause failures in production models and understand how to fix them in future model iterations
  • Decide which features and capabilities are ready for public launch

What You Bring

  • Strong fundamentals in software engineering, machine learning, debugging complex systems, and the ability + desire to learn quickly.
  • Experience building and ensuring quality of large multilingual datasets.
  • Experience training and debugging generative models (speech, text, or multimodal), especially SFT, RL, synthetic data, and evaluation (both human and automated).
  • Excitement about solving problems grounded in real customer needs, not just benchmarks.
  • Bonus points if you have native proficiency in other languages!

More Details

Our Benefits (US Employees Only)

🏦 401(k)

🦖 Your own personal Yoshi

Our Commitment to Equal Opportunity

Required skills

  • databricks
  • reinforcement learning
  • artificial intelligence
  • rest
  • machine learning