About

We started with the hardest audio, and the people who handle it.

Perit is a training-data company built around one idea: the person who does the job is the right person to record it, transcribe it and grade it. That started with support calls. It is widening to video, image, text — and to embodied data for robots.

The story

It started with calls that went badly.

Speech models were failing quietly. A transcript read well, a digit in an account number was wrong, and nobody noticed until three systems downstream. Clean read-speech corpora could not catch it, and crowd raters had no idea what a good call sounded like. So we built a bench of people who did — transcribers and aligners who work to a written guide and pass a test before they touch a paid file — and put it on a platform we run ourselves, Foundry.

That bench has now delivered more than four thousand hours of speech across nine locales, at fifty hours a working day, with every batch shipping the numbers it was produced at. The same machinery — recruiting to a locale, consent captured on the file, a calibration gate, QA sampling, a quality report — does not care whether the file is an audio clip, a photograph, a paragraph or a first-person video of someone folding a towel.

So the next step is the obvious one. Video, image and text collection to spec, annotation of any data type on the same bench, and a Physical Intelligence line — egocentric video, depth, teleoperation and handheld-gripper demonstrations — for teams training robots. We say plainly which of those has shipped and which is being scoped, because a number nobody measured is worth nothing to the person buying it.

Perit is operated by Zyno AI Inc and backed by Y Combinator. No frontier lab owns a piece of it, and none sits on the cap table — your held-out data does not end up next to a competitor's.

Delivered so far

Every number here is speech work already shipped.

4,000+
Hours of audio delivered
650
Transcribers and aligners on the bench
50 hrs
Delivered every working day
95%+
QA-sample accuracy, checked against WER
What we believe

Four things we would rather lose a deal over.

Operators, not a crowd

The people who record, transcribe and grade are recruited for the sector or the locale, and the rubric is written by someone who signs off on that work for a living.

Delivered, not promised

A number appears on this site after the hours ship, never before. Where a line is new, the page says so.

One consent, on the file

Captured before anyone speaks or films, naming what the data trains, withdrawable before payout, and referenced in the manifest of every item delivered.

Independent

Backed by Y Combinator, owned by nobody who trains a frontier model. Independence is a structural fact, not a claim.

Where the bench is

9 locales, recruited one at a time.

Backed by Y CombinatorF2026

Independent, and staying that way.

No frontier lab owns a piece of Perit, and none of them sits on our cap table. Your held-out audio does not end up next to a competitor's. Backed by Y Combinator, F26.

Facts
Company
Zyno AI Inc
Brand
Perit AI
Backed by
Y Combinator, F2026
Platform
Foundry — our own annotation and recording workspace
Bench
650 transcribers and aligners
Delivered
4,000+ hours of speech, 9 locales