Speech & voice AI
Specify regional language coverage, accents, code-switching and acoustic conditions. Assess transcript conventions against your recognition task.
Explore speech & voice ai ↗Fiund helps AI teams find and license real-world data for training and evaluation, with sources, usage rights and delivery scoped for each project.
English podcast recordings with natural dialogue and solo speech. Browse a collection to see its formats and available annotations.
Browse all datasetsBrowse the full collection of English podcast audio, video, and transcripts.
Back-and-forth dialogue, with individual speaker tracks where recorded.
Single-speaker recordings for long-form speech and transcription tasks.
DATA WE WORK WITH
Explore first-person video, multilingual speech, music, documents and more. Start with the source material and coverage you need.
Explore data types ↗First-person video that follows a person through real tasks, with hands, objects and the surrounding environment in view.
Spanish-language speech and audiovisual data scoped around your regions, conversation types and model task.
Road and driving footage for studying traffic, changing environments and vehicle-view perception.
Educational recordings that explain subjects through spoken instruction, supporting visuals and examples.
Production components that connect a musical recording with separate source parts and symbolic note information where available.
Business records and correspondence that preserve the relationships between products, customers, orders and operations.
Explore how recordings, transcripts, and documentation fit together. The exact files and annotations are agreed for your collection.

See the recording environment and camera views. Individual speaker video is available where separately recorded.
Illustrative example with a synthetic transcript. Request a collection sample to evaluate actual recordings.
Request a sampleBuilt around your model task
Start with the problem your model needs to solve. Explore relevant data types, sample requirements and delivery considerations.
Specify regional language coverage, accents, code-switching and acoustic conditions. Assess transcript conventions against your recognition task.
Explore speech & voice ai ↗Select recordings where the audio and picture describe the same event. Confirm source relationships and timing information in a representative sample.
Explore multimodal ai ↗Specify the tools, objects and environments your model needs to observe. First-person sequences can retain the context around a decision or movement.
Explore physical & embodied ai ↗Specify document types, layouts and the fields your model needs to recognize. Agree whether source files, text representations or both are required.
Explore document ai ↗Source fit, access and preparation are confirmed for your brief.
Review the source, usage rights, and delivery details before you license a collection.
How we source dataReview where recordings came from and how each file connects to its source.
Source documentationAgree the permitted AI uses and review ownership, consent, and any restrictions.
Rights reviewConfirm the media, annotations, file specifications, and documentation you will receive.
Delivery processStart with your model task. We’ll help you define the source, license, and delivery for your project.
Discuss your data needsShare your model task, languages, recording requirements, and timeline so we can assess the fit.
Evaluate available recordings and annotations. We identify anything that would need additional sourcing or preparation.
Agree usage rights, restrictions, pricing, and the exact files and annotations included in your project.
Receive the selected media and agreed documentation, with a file manifest that identifies what was delivered.
Need another language, recording setup, or annotation? Tell us the requirements. We’ll assess what can be sourced or commissioned.
Discuss your requirements
Earn from your audio and video through licensing. Keep ownership of your work while Fiund helps connect it with buyers and define the permitted uses.
License your contentWhat to expect before you request a sample or license a collection.
Explore English podcast recordings with audio, video, and transcripts, including natural dialogue and solo speech. Each collection page explains its content, formats, and use cases. Other categories may require sourcing or commissioning.
Yes. Send a sample request identifying the collection and your intended evaluation. We will confirm the available material and review scope. The interactive homepage demonstration uses editorial imagery and a synthetic transcript; actual collection samples are reviewed separately.
Available transcript files and other annotations are listed for the selected collection. Precise alignment, speaker mapping, additional labels, and human review are confirmed for the chosen files or scoped as additional work. Source media and commissioned refinements are identified separately.
A signed agreement that lets an AI team train models on media someone else owns — audio, video, images or text — in exchange for payment. The licence spells out what can be trained, for how long, and what rights travel with the data, so nobody is relying on fair-use guesses.
You send a spec: modality, language, hours, recording conditions. We match it against libraries whose owners have agreed to license, clear the rights, and deliver the data under a licence that grants AI-training use explicitly. One agreement covers the whole delivery.
Original audio, video, creator UGC, motion capture, sensor data, task demonstrations, and other media may be licensable when the supplier controls the relevant rights and the proposed AI use, identifiable people, embedded works, privacy, and contractual restrictions are addressed. A category page is not an availability claim; a particular collection still has to pass review.
Yes, through a written buyer licence. The agreement should identify the dataset, approved AI purposes, production or evaluation scope, recipients, term, territory, retention, and restrictions. A download, public post, or evaluation sample is not automatically a production-training grant.
Public accessibility alone does not establish permission for a proposed AI use. Buyers should examine the source, applicable permissions, privacy considerations, restrictions, and intended use. Our rights guides explain the questions to raise during diligence.
Per licence. You keep ownership while fiund handles AI licensing under one contributor agreement, and you are paid each time an AI team licenses your material. What a deal is worth depends on modality, exclusivity and volume — the payouts guide explains how deals are structured.
Every permission needed for AI training is secured in writing before data changes hands: ownership confirmed, training rights granted explicitly in the licence, and voice and likeness consent covered wherever people are identifiable.
We connect each approved asset to its source, owner or authorized licensor, applicable agreement, consent status, known restrictions, and file-level manifest. Unresolved material is held back instead of being described as cleared.
No. fiund’s standard contributor terms do not authorize a buyer to create a synthetic replica, clone, or voice or likeness model intended to imitate an identifiable person, or use that identity to imply endorsement.
There is no universal rate card. Pricing is scoped to the content type, task fit, volume, source quality, metadata, consent and diligence work, permitted uses, delivery requirements, and any requested exclusivity. A quote follows the actual brief and verified inventory.
fiund targets an initial scope and availability response within two business days. That response distinguishes current inventory from data that would need to be sourced or commissioned; it is not a promise that every collection can be delivered in that window.
An approved delivery can include the buyer licence, a versioned file manifest, data dictionary, technical specifications, provenance and consent references, known restrictions, and integrity evidence for the exact transferred set.

Tell us what you need to evaluate. We’ll confirm the available samples and next steps.
Discuss your data needsCollection availability, permissions, and final scope are confirmed against your brief.