Plus embodied data for robotics — egocentric, depth, teleoperation, gripper — in early access.
Data collectionEvery number here is speech work already shipped. When we take on a new modality it will be listed the same way — after it ships, not before.
50 on fixed-term contracts, the rest freelance — and hiring against every open sector.
Freelance bench, timestamped to the word so a model can be scored on when it heard something, not just what.
Speech delivered at scale. Video, image and text collected to spec.
Speech delivered · video, image, text openThe Foundry bench — transcription, alignment and labels, any data type.
Audio delivered · image, video, text on the same benchName the input your model fails on. We collect it, transcribe or label it, and license the set with the spec and terms published.
Priced by the unit it is delivered in

No dataset catalogue to guess from. You describe the failure; we scope, pilot and deliver against it, with the contributor rate on the quote.
Full delivery spec and licence termsNot the dataset — the input your model gets wrong. A clip, a photo, a paragraph is enough to start.
Quantity, locales, devices, conditions, annotation depth — and the contributor rate you will be paying, stated as its own line.
Five percent of the volume, or ten to thirty episodes for embodied data, delivered first with a quality report. If the spec was wrong, the re-run is on us.
Drops into your bucket in the region you chose, with the manifest, consent references and reviewer chain on every item.
Egocentric video, depth, teleoperation and handheld-gripper demonstrations on the same recruiting, consent and QA machinery as our speech work. No volumes listed until the first collections ship.
Lowercase, fillers kept, one written convention for every tag, and a timestamp on every word — so accuracy and latency can be scored apart.
Read the methodhaan so i i filed the claim on monday [pause] but the portal is still showing pending only
let me pull that up for you ma'am [keyboard] can you confirm the claim reference
it's uh [pause] c l four two nine [overlap] eight—
[overlap] four two nine eight, yes i have it here
The bench is built from people who worked the floor, the ward and the claims desk — ex-contact-centre QA leads, claims handlers, ward clerks. They are paid by the unit, at rates published on the listing.

Two crafts, one bench. Transcribers write what was actually said, fillers and all. Aligners put a start and end time on every word. A rubric author is someone who signs off on these calls at work, and every rubric has to be applicable by a stranger.
Contributors record on their own phones, in their own rooms, in the accent they speak. Blue is where the bench works today; anything else we recruit to order against locale quotas.
We would rather publish how a run works and no numbers than numbers nobody measured. The design is fixed and public now.
Two tracks — transcription and live conversation — run across ten sectors on held-out audio, scored against the person who does that job.
Protocol published · first run openLive speech-to-speech calls with trained speakers who interrupt, change their mind and hand over the phone mid-call.
Design partnersAnyone can join the recording pool. Pick a session, speak in your own accent, and get paid once it passes review — no studio, no interview, no shift to show up for.