
Voice data for AI
Besimple AI is building the audio data and evaluation infrastructure for real-world voice AI deployment. We help AI teams collect, evaluate, and improve voice models across languages, accents, environments, and complex human interactions.
The founders are ex-Meta product and engineering leaders from MIT and Brown. We are a small, high-ownership team working directly with frontier AI labs and fast-moving AI companies to push the state of the art in audio models, voice agents, and human-in-the-loop evaluation.
We are looking for a full-time Mobile Engineer lead to own the mobile experience for our contributor data collection platforms.
This is a high-ownership role at the intersection of mobile engineering, audio recording, user experience, product, and AI data quality. You will own the mobile experiences that allow people around the world to record, review, submit, and improve audio data used to train and evaluate frontier voice AI systems.
This is not a generic app development role. You will work on real production workflows where mobile UX directly affects audio quality, contributor throughput, customer delivery, and model performance.
You will be successful if Besimple can collect high-quality audio data through mobile workflows that are:
Your core metrics may include:
This is a high-impact role at an early-stage AI company. You will not just build screens; you will help define how real-world audio data gets collected, evaluated, and improved for the next generation of voice AI.
You will work directly with the founders and users, own product-critical mobile workflows, and shape the engineering foundation for a platform used by AI teams building real production voice systems.
Source used for structure and company context: original YC job posting.
At Besimple AI, we’re making it radically easier for teams to build and ship reliable AI by fixing the hardest part of the stack: data. Good evaluation, training and safety data require domain experts, robust tooling and meticulous QA. AI teams and labs come to us to get high quality data so they can launch AI safely. We’re a YC X25 company based in Redwood City, CA, already powering evaluation and training pipelines for leading AI companies across customer support, search, and education. Join now to be close to real customer impact, not just demos.
High-quality, human-reviewed data is still the single biggest driver of model quality, but most teams are stuck with old tools and legacy processes that do not scale to modern, multimodal, agentic workflows. Besimple replaces that mess with instant custom UIs, tailored rubrics, and an end-to-end human-in-the-loop workflow that supports text, chat, audio, video, LLM traces, and more. We meet teams where they are—whether they need on-prem deployments and granular user management or a fast cloud setup—to turn evaluation into a continuous capability rather than a one-time project.
Founders previously built the annotation platform that supported Meta’s Llama models. We’ve seen how world-class annotation systems shape model quality and iteration speed; we’re bringing those lessons to every AI team that needs to ship with confidence. You’ll work directly with the founders and users, owning problems end-to-end—from an interface that unlocks a tough rubric, to a workflow that reduces disagreement, to a AI judge system that improves quality.
If you’re excited by systems that combine product design, human judgment, and applied AI—and you want to build the data and evaluation layer that keeps AI trustworthy—come build with us. See how fast teams can go from raw logs to a robust, human-in-the-loop eval pipeline—and how that changes the way they ship AI.