We license audio & video data

Find the data. Fund the people who made it.

A licensing marketplace for real-world audio and video. AI teams get training data that isn't already in the crawl. The people who own it get paid, and keep their rights.

Signed rights Owner approved No scraping
What we license

Real-world media, sourced from the people who own it.

We focus on material that never made it into a public crawl — recorded by real people, in real conditions, and licensed directly from them.

Audio

Most speech corpora are scripted and read aloud. We source the other kind — people actually talking to each other, thinking out loud, talking over one another.

interviewsmulti-speakeraccentscrosstalk

Rights-cleared by design

Every asset carries a signed licence and, where people are identifiable, voice and likeness consent. The paperwork travels with the file.

signed licenceconsent on filediligence-ready

A global supply network

Creators, studios and archives on the supply side — approving every use, keeping ownership, and getting paid per licence.

creatorsstudiosarchives
Two sides, one marketplace

Whichever side you're on, the deal is the same shape.

Data you can defend in diligence

Volume gets you a demo. Chain of title gets you a model you can ship.

  • Not already in the crawl

    Non-public material sourced directly from owners, so you aren't paying for something your model has already seen.

  • Rights arrive with the file

    Every asset carries a signed licence from its owner granting AI training rights explicitly. We can show you the paperwork.

  • Send us a brief

    If the dataset doesn't exist yet, tell us the spec and we go source it. That's the part the name is about.

Request a dataset

Your archive is worth more than you think

If you've been recording for years, you're sitting on an asset. Licensing it shouldn't mean losing it.

  • You keep ownership

    You license, you don't sell. The material stays yours and you set the terms it goes out under.

  • You approve every use

    Nothing is licensed to a buyer or a use case you haven't signed off on.

  • You get paid per licence

    Transparent rates, transparent terms, and no exclusivity lock-in unless you want it and it's priced for it.

Licence your library
How it works

Three steps, no mystery.

STEP 01

Tell us what you need

A buyer sends a brief — modality, languages, conditions, volume, licence terms. Or an owner sends us what they have and we tell them honestly whether there's a market for it.

STEP 02

We source and clear it

We go to the people who hold the material and paper the rights before anything moves. Consent, likeness and training rights are handled up front, not retroactively.

STEP 03

Delivery and payout

The dataset arrives with its documentation attached. The owner gets paid. Everyone can see exactly what was licensed and to whom.

Rights & provenance

Licensed at the source. Every single time.

Training data deals almost never fall apart over quality. They fall apart because nobody can prove where the material came from or what rights came with it. We built around that problem first.

Signed at the source

A written agreement with the owner of every asset we license — no exceptions, no assumptions.

Training rights, explicitly

Granted in the licence itself rather than inferred from a general content or distribution agreement.

Voice and likeness consent

Covered separately wherever people are identifiable, because copyright alone doesn't reach it.

Nothing scraped

No crawling, and nothing sourced from behind a login. Provenance records are available during diligence.

AssetOwnerLicence
clip_08812.wavindependentsigned ✓
clip_08813.wavstudiosigned ✓
reel_00417.mp4creatorsigned ✓
reel_00418.mp4creatorsigned ✓
set_00092.ziparchivesigned ✓
Every row backed by a signed agreement, available in diligence

Let's talk about what you actually need.

Whether you're building a model or sitting on an archive, the first conversation is short and specific.