Data & ML infrastructure
Pipelines, annotation tooling, and platform work that turns raw Indic speech and text into model-ready datasets.
Careers
We're a small team building the training data behind the world's under-served languages - and we're always looking for people who care about this problem as much as we do. We don't do long job descriptions; we look at what you'd bring.
Where you fit
Pipelines, annotation tooling, and platform work that turns raw Indic speech and text into model-ready datasets.
Native experts across Indic languages, dialects, and domains - defining what "correct" means and enforcing it.
Building and running the curated community of vetted native speakers our data depends on.
Operations, partnerships, design - if you can move this mission forward, tell us what you'd own.
Get in touch
Tell us what you'd build, fix, or bring to Coremantle - engineering, linguistics, community, or something we haven't thought of - and attach your resume.