Speech datasets tailored to your model requirements

You specify the languages, speakers, and format. We scope recruitment, recording, transcription, QA, rights, and delivery in a written project plan.

5.0 on Datarade 4.4 on Trustpilot SKI Verified Supplier Backed by
Data provenance

The voice is only half the dataset.

The rest is proof: who spoke, how it was recorded, the rights attached, and a transcript your model can trust.

Spirelight delivers both.

Dataset record DK-JUT / 000412
Validated
clip_0412.wav 00:04.2
00:00 48 kHz · 24 bit 00:04.2
Human-verified transcript

“I’d like to change my booking to Friday morning.”

Language
da-DK
Dialect
Jutlandic
Region
Aarhus
Speaker
27 · M
Capture
Shure MV7 · 38 dB SNR
Rights
Signed consent
9/9 fields verified WAV + JSONL
Recruitment and recording workflow

Recruit the right speakers and collect recordings in one controlled flow.

We source contributors by language, dialect, age, gender, region, device, or other project criteria. Speakers record through browser-based tools with prompts, consent, audio checks, and project guidelines built into the workflow.

Live project tracking

Track collected hours, QA status, and delivery batches while the project runs.

Review progress during production and receive validated batches with audio, transcripts, metadata, and manifests. Your team can test early batches and adjust the collection before the full dataset is finished.

What we do

Speech datasets tailored to your requirements

Need Danish conversations, German call-center speech, French dialect coverage, wake words, commands, or audio plus video recordings? A written brief can define the workflow from speaker recruitment to final delivery: speech data collection services for custom recording requirements, and audio annotation services for structuring agreed labels and transcripts.

spirelight · session 01 / 03

Synthetic interface example only. Names, locations, ages, recording status, and audio specifications are fictional and are not customer data, live availability, or operating metrics.

Live session · Iberian Spanish

Scripted monologues, dialect-tagged

Recording
  • MRMaria · Madrid · 32Done
  • JPJavier · Sevilla · 41Recording
  • ALAna · Bilbao · 27Queued
WAV · 48 kHz · stereo Prompt set 02 / 12
spirelight · transcript 02 / 03
en-IE_002_dialogue_03.json QA · 2 reviewers
  1. 00:00.42 S1 Could you walk me through the booking flow you used last Tuesday?
  2. 00:03.10 S2 Sure, I opened the app, tapped the search bar, then… flagged
  3. 00:06.94 S1 Got it. Any pauses or hesitations there?
  4. 00:09.38 S2 Yeah, [pause 1.2s] I had to scroll to find the right date.
Word-level timestamps · Speaker-aware · Diarized
manifest.json 03 / 03
{
"project": { 3 fields }, {
"id": "sl-9241",
"language": "es-ES",
"hours": 3000
},
"audio": { 3 fields }, {
"format": "wav",
"sample_rate": 48000,
"channels": 2
},
"transcripts": { click to expand }, {
"format": "jsonl",
"timestamps": "word",
"diarized": true
},
"delivery": { click to expand } {
"channel": "s3-bucket",
"checksums": "sha256",
"batches": true
}
}
01

Speech collection

A collection can include monologues, dialogues, wake words, commands, scripted prompts, roleplays, and natural conversations, with speaker criteria and feasibility confirmed in the written scope.

  • Remote, moderated, on-site, or studio-style recording, subject to feasibility
  • Monologues, dialogues, commands, and roleplays
  • Audio-only or synchronized audio plus video
02

Transcription and annotation

An annotation scope can include machine-assisted or human-reviewed transcripts, timestamps, speaker labels, domain terminology, and rules matched to the agreed model requirements.

  • Word-level or segment-level timestamps
  • Speaker labels and dialogue structure
  • Human review based on your QA criteria
03

Dataset delivery

The delivery plan can package audio, transcripts, metadata, consent references, QA notes, and manifests in the formats agreed with your engineering team.

  • WAV, JSON, JSONL, CSV, or custom formats
  • Metadata schemas matched to your spec
  • Bucket transfer, API handoff, or batch delivery
Why Spirelight

Collect the speech data your model is missing

Four reasons teams use Spirelight for custom speech data collection.

Speaker reach
01 / 04

Recruit speakers by language, dialect, and profile.

Define the speakers you need by language, dialect, region, age, gender, device, environment, or other project criteria.

Contributor recruitment across 50+ languages and 30+ markets.

Controlled recording setups when quality matters.

We can run remote, on-site, or studio-style sessions using defined microphones, devices, rooms, scripts, and acoustic requirements.

From single-speaker sessions to multi-day collection projects.

Review early batches and adjust the project as it runs.

Your team can test early deliveries, identify gaps, and update speaker targets, prompts, or guidelines before the full dataset is complete.

Mid-project adjustments without restarting production.

Strong coverage in Nordic and harder-to-source European languages.

We regularly recruit and review speakers in markets where off-the-shelf datasets are limited, including Nordic languages, regional dialects, and smaller European language varieties.

Recruiters and reviewers across Denmark, Sweden, Norway, Finland, and Iceland.

Your speech data partner for

fine-tuning and language expansion.

A written project can coordinate contributor recruitment, recording workflows, transcription, QA, and delivery for a specified language-expansion brief.

Use cases

Speech data for common voice AI use cases

The workflow is similar across projects, but the speakers, prompts, recording conditions, annotations, and deliverables change with each use case.

Wake words and commands

Wake words and voice commands.

Collect commands, activation phrases, device instructions, and short utterances across languages, accents, microphones, and environments.

Multilingual

Multilingual ASR and TTS expansion.

Build language, accent, and dialect coverage for speech recognition, synthetic voice, and speech evaluation datasets.

Voice agents and emotion

Voice agents and emotion-aware AI.

Collect conversations, roleplays, customer service scenarios, emotional speech, and domain-specific interactions for more natural voice systems.

Team

The team running your data collection project

Spirelight combines commercial project design, recruitment operations, platform engineering, transcription workflows, QA, and delivery management in one team.

01 / 09

Andreas Kromann

CEO · Commercial lead

Andreas works with clients to turn model requirements into concrete data collection projects.

He defines the project scope, speaker targets, recruitment approach, and delivery expectations before production starts.

Emil Thorsson

CFO · Operations

Emil supports operations, documentation, compliance coordination, and project delivery.

He helps structure the process so recruitment, consent, production, and handoff stay aligned.

Gustav Aggeboe

CTO · Platform architecture

Gustav leads the platform architecture behind Spirelight.

He builds the systems used to manage recording, transcription, QA, metadata, contributor workflows, and dataset delivery.

Joyi Ulfat

Senior Project Manager

Joyi manages project execution across contributors, reviewers, and delivery teams.

She keeps production moving, follows up on daily progress, and helps ensure each project meets its agreed requirements.

Mateo Thelen

Project Manager

Mateo coordinates contributors, recording workflows, and production tasks.

He helps translate project requirements into daily execution and keeps the different parts of the workflow aligned.

Pekka Larjovuori

Crowd Source Expert

Pekka supports recruitment strategy and contributor operations.

He helps source speakers for projects with specific language, dialect, regional, or profile requirements.

Victor Melchior

Sales · Market expansion

Victor leads sales and market research, mapping where Spirelight's speech data work fits new clients and regions.

He runs country expansion research and opens conversations in the markets we move into next.

Yusif Aliyev

Legal · Policy & compliance

Yusif leads legal, internal policy, and compliance at Spirelight.

He maintains the internal policies and compliance processes that keep consent, data handling, and contracts in order.

Michael J. Jørgensen

B2B partners & talent recruitment

Michael leads B2B sales and manages Spirelight's partnerships with freelancers and companies.

He builds the relationships that bring in new clients and contributors, and keeps our freelancer and company partners aligned with each project's needs.

Guides

Understand speech data before you buy it

Practical buyer guides for teams building voice AI: what speech data is, how much a project may need, and which licensing terms to evaluate.

View all guides

Services and advanced topics: speech data collection services, audio annotation services, speaker diarization, speech data licensing and license agreements, telephony speech datasets, what speech data collection costs, or what data annotation is.

Popular collection configurations: Russian collection configuration, Amharic collection configuration, Egyptian Arabic collection configuration, Marathi collection configuration, Italian collection configuration, Greek collection configuration, or review all collection configurations.

Looking for paid contributor work? Create a free AI training profile, or read how people make money training AI. Projects open in waves and matching depends on each brief.

Transcription

Transcription staffed by language and dialect

Native contributor teams matched to the language and regional variant in the recording, with the written convention agreed before production starts.

View all transcription languages

See how the global contributor workforce behind every transcription project is recruited and managed.

Get started

Tell us the speech data you need

Send over your speaker profiles, language needs, and recording conditions. Our team will review the brief and respond with feasibility questions and next steps.

Scoped
speaker recruitment for each project brief
Confirmed
language and locale feasibility before launch
Defined
quality and acceptance criteria in the scope
Written
scope, usage rights, and delivery terms