Sanas

Member of Technical Staff, Research Evaluations

Sanas is pioneering the future of human communication. Founded by a team of Stanford researchers and entrepreneurs with deep industry experience, Sanas has developed the world's first real-time speech AI platform capable of accent translation, noise cancellation, speech enhancement, cross-language communication, and more.

Sanas makes conversations clearer, more inclusive, and more effective, removing barriers that prevent people from being understood, regardless of accent, background noise, or native language.

Sanas is currently one of the fastest growing startups in Silicon Valley, growing from $16M to $50M ARR in 2025. The company's core business is profitable and is on track to end 2026 with >$120M ARR. Our team combines deep expertise in model innovation and systems engineering with a design-minded product engineering culture to build and ship cutting-edge AI models and experiences — entirely in-house.

Sanas is a 130 person team, established in 2020. In this short span, we've successfully secured over $100 million in funding. Our innovation has been supported by the industry's leading investors, including Insight Partners, Google Ventures, Quadrille Capital, General Catalyst, Quiet Capital, and other influential investors. Our reputation is further solidified by collaborations with numerous Fortune 100 companies. With Sanas, you're not just adopting a product; you're investing in the future of communication.

If you’re looking to have a significant role in roadmapping and driving technical directions, if you’re looking to deploy challenging and big ideas without much overhead or slowness, if you're looking to leave your mark on an ambitious, generational mission to change how the worlds thinks about speech + AI, then Sanas is a well-suited place for you.

About the Role

Sanas is building a full Speech AI suite, all working together as one platform. As that surface area grows, so does the need for a single, rigorous owner of how we know everything is working.

As Evaluations Lead, you'll design the evaluation frameworks and benchmarking systems that answer that question — sitting at the intersection of research, product, and infrastructure to build the metrics, systems, and studies that hold our models accountable. This role suits someone who pairs scientific rigor with real technical execution. Your work will shape how Sanas builds and evaluates its models across every one of these products, making sure progress is measured not just by static benchmarks, but by the harder, more meaningful qualities — understanding, naturalness, and adaptability in real-world interaction.

Your Impact

  • Identify and define the model capabilities and behaviors that actually matter for evaluation — not just what's easy to measure
  • Build and ship evaluation pipelines with robust statistical analysis and clear, actionable reporting
  • Partner directly with model training and research teams to embed evaluation into the development loop itself
  • Prototype new user studies and behavioral experiments that ground evaluation in how these models actually get used

What You Bring

  • Experience designing or implementing evaluation frameworks for generative models — audio, text, or multimodal
  • Strong technical and analytical skills, with the ability to take an open-ended research idea and turn it into a production-ready system
  • Creativity in defining novel, quantitative metrics for qualities that are inherently subjective
  • Genuine excitement for building evaluation systems that bridge research and real-world use
  • Equal parts curious and rigorous — driven by actually figuring out how to measure meaningful progress, not just reporting a number
  • The ability to build it yourself. This is an engineering role — you'll be writing the pipelines and tooling, not just specifying them

Nice-to-Haves

  • AI modeling experience — someone who has trained, fine-tuned, or shipped models themselves brings a level of judgment to evaluation design that's hard to substitute, and is highly valued for this role
  • Background in audio modeling — Speech-to-Text, Text-to-Speech, or similar
  • Multilingual — especially relevant for evaluating multilingual systems, where understanding the semantic nuance across languages, not just the literal accuracy, is core to getting evaluation right


Engineering

Palo Alto, CA

Teilen auf:

NutzungsbedingungenDatenschutzCookiesPowered by Rippling