Licensed audio
and video for AI.

FIUND helps AI teams license recorded conversations for training and evaluation, with samples, usage rights, and delivery agreed for each project.

Explore our
conversation datasets.

English podcast recordings with natural dialogue and solo speech. Browse a collection to see its formats and available annotations.

Browse all datasets
Editorial AI image
Human conversation.Audio, video, and transcripts.

What’s included
in a delivery.

Explore how recordings, transcripts, and documentation fit together. The exact files and annotations are agreed for your collection.

Illustrative data example
AI-generated illustration of a conversation in a recording studio.
HUMAN CONVERSATIONOne recording.
Multiple ways to study it.
Video

Review the source video.

See the recording environment and camera views. Individual speaker video is available where separately recorded.

  • Original recording files
  • Combined or individual camera views
  • Resolution and formats confirmed for your selection

AI imagery and a synthetic transcript illustrate the structure. Request a collection sample to evaluate actual recordings.

Request a sample

Know where your data
comes from.

Review the source, usage rights, and delivery details before you license a collection.

How we source data

How to license
data from FIUND.

Start with a sample. We’ll help you define the collection, license, and delivery for your project.

Request a sample
01

Tell us what you need.

Share your model task, languages, recording requirements, and timeline so we can assess the fit.

02

Review a sample.

Evaluate available recordings and annotations. We identify anything that would need additional sourcing or preparation.

03

Set the license and scope.

Agree usage rights, restrictions, pricing, and the exact files and annotations included in your project.

04

Receive the agreed files.

Receive the selected media and agreed documentation, with a file manifest that identifies what was delivered.

Need a custom dataset?

Need another language, recording setup, or annotation? Tell us the requirements. We’ll assess what can be sourced or commissioned.

Discuss your requirements
Editorial AI image of a speaker listening at a microphone in a home recording space.
FOR CONTENT OWNERS

License your content
to AI teams.

Earn from your audio and video through licensing. Keep ownership of your work while FIUND helps connect it with buyers and define the permitted uses.

License your content

Guides to
AI data licensing.

Explore all resources

Questions about
licensing data?

What to expect before you request a sample or license a collection.

What is currently in the FIUND catalog?

Explore English podcast recordings with audio, video, and transcripts, including natural dialogue and solo speech. Each collection page explains its content, formats, and use cases. Other categories may require sourcing or commissioning.

Can I review a sample before licensing?

Yes. Send a sample request identifying the collection and your intended evaluation. We will confirm the available material and review scope. The interactive homepage demonstration uses editorial imagery and a synthetic transcript; actual collection samples are reviewed separately.

Which annotations are included?

Available transcript files and other annotations are listed for the selected collection. Precise alignment, speaker mapping, additional labels, and human review are confirmed for the chosen files or scoped as additional work. Source media and commissioned refinements are identified separately.

What is AI training data licensing?

A signed agreement that lets an AI team train models on media someone else owns — audio, video, images or text — in exchange for payment. The licence spells out what can be trained, for how long, and what rights travel with the data, so nobody is relying on fair-use guesses.

How do AI teams license audio and video data?

You send a spec: modality, language, hours, recording conditions. We match it against libraries whose owners have agreed to license, clear the rights, and deliver the data under a licence that grants AI-training use explicitly. One agreement covers the whole delivery.

What kinds of content can be licensed for AI training?

Original audio, video, creator UGC, motion capture, sensor data, task demonstrations, and other media may be licensable when the supplier controls the relevant rights and the proposed AI use, identifiable people, embedded works, privacy, and contractual restrictions are addressed. A category page is not an availability claim; a particular collection still has to pass review.

More about licensing and permitted use

Can an AI team buy training data with commercial rights?

Yes, through a written buyer licence. The agreement should identify the dataset, approved AI purposes, production or evaluation scope, recipients, term, territory, retention, and restrictions. A download, public post, or evaluation sample is not automatically a production-training grant.

Is scraped data safe to train on?

Public accessibility alone does not establish permission for a proposed AI use. Buyers should examine the source, applicable permissions, privacy considerations, restrictions, and intended use. Our rights guides explain the questions to raise during diligence.

How do data owners get paid?

Per licence. You keep ownership while fiund handles AI licensing under one contributor agreement, and you are paid each time an AI team licenses your material. What a deal is worth depends on modality, exclusivity and volume — the payouts guide explains how deals are structured.

What does rights-cleared mean?

Every permission needed for AI training is secured in writing before data changes hands: ownership confirmed, training rights granted explicitly in the licence, and voice and likeness consent covered wherever people are identifiable.

How does fiund verify provenance and consent?

We connect each approved asset to its source, owner or authorized licensor, applicable agreement, consent status, known restrictions, and file-level manifest. Unresolved material is held back instead of being described as cleared.

Does fiund authorize voice or likeness cloning?

No. fiund’s standard contributor terms do not authorize a buyer to create a synthetic replica, clone, or voice or likeness model intended to imitate an identifiable person, or use that identity to imply endorsement.

How much does it cost to license training data?

There is no universal rate card. Pricing is scoped to the content type, task fit, volume, source quality, metadata, consent and diligence work, permitted uses, delivery requirements, and any requested exclusivity. A quote follows the actual brief and verified inventory.

How quickly can fiund respond to a dataset brief?

fiund targets an initial scope and availability response within two business days. That response distinguishes current inventory from data that would need to be sourced or commissioned; it is not a promise that every collection can be delivered in that window.

What documentation arrives with licensed data?

An approved delivery can include the buyer licence, a versioned file manifest, data dictionary, technical specifications, provenance and consent references, known restrictions, and integrity evidence for the exact transferred set.

Go deeper with FIUND

Find the right data
for your model.

Tell us what you need to evaluate. We’ll confirm the available samples and next steps.

Request a sample
HUMAN AT THE SOURCE.Have a library to license?

Editorial imagery is AI-generated. Collection availability, permissions, and final scope are confirmed against your brief.