Fundación Voz del Centro Cultural heritage · 2026 Platform Heritage AI

La Voz del Centro, la experiencia.

La memoria de un país, en su propia voz. The memory of a country, in its own voice, and now something you can ask, walk and hear.

↳ live at vdc.mutiny-labs.com
La Voz del Centro, la experiencia, hero screenshot
1,125
episodes, 24 years of oral history
8,592
chapters, each with its own poster
46.7K
passages you can ask by meaning
the setting

La Voz del Centro is the Caribbean's first podcast. For more than twenty years, its creator and host Ángel Collado-Schwarz has sat down each week with historians, artists and protagonists to talk through the history and culture of Puerto Rico and the Caribbean. Past 1,100 episodes, it became the most comprehensive oral-history archive of the island that exists, from the Taínos to today. Mutiny Labs has been the technology behind it from day one, in the background. It lived as audio: profound, but linear, and impossible to search by what was said.

the hard problem

Twenty years of recorded conversation is a national memory locked in audio. You could not ask it a question, jump to the minute where something was said, or take in the shape of two decades at a glance. The Fundación Voz del Centro wanted more than a better player: an archive people could interrogate, walk and inhabit, faithful to every source, and unmistakably Puerto Rican rather than a generic media library dressed in someone else's aesthetic.

how we mutinied

We have been the technology behind La Voz del Centro since day one, always in the background, and we kept it that way. The program is Ángel Collado-Schwarz's life's work; the archive is a living layer built over it, never altering a word. A pipeline reads the public feed, transcribes and annotates every episode with AI, and indexes it by meaning, so it answers in the speaker's own voice and lands on the exact minute. An interview layer answers questions in the program's own courteous register, saying only what the corpus documents, with the audio as proof. Then we gave it a cultural identity, Cartel de Pueblo: a poster for every chapter, in the DIVEDCO tradition.

What we built.

↓ each one working, not a mock

Entrevista al Archivo

A floating conversation with the whole archive, in the courteous voice of the program. Ask in plain Spanish and the answer streams in, grounded in the corpus; every claim carries a tap-to-play citation that plays the exact minute of the exact episode without leaving the window. What the archive does not document, it will not say.

A player that knows its chapters

Every episode plays with AI-generated chapters marked on the scrubber, a timestamped outline, and a Cartel de Pueblo poster for each chapter that changes as you listen.

Momentos

The most memorable spoken moments surfaced as full-bleed quote cards over original art, each one a single tap from the exact second in the audio.

Rutas and the timeline

Curated narrative routes string episodes into a larger story, stop by stop, while a timeline reorders the whole archive by the era each episode narrates, from the Taínos to the 21st century.

Quién es quién

Three doors into the archive by theme, by voice and by entity, with people, places, institutions, events and works ranked by how often the archive returns to them.

How the archive is built
01

Read two decades of feed

We parse the public podcast feed into a clean, idempotent manifest keyed to stable IDs. The episode number printed in each title is treated as the authority, because the feed's own numbering field carries typos going back years.

02

Mirror every recording

Each episode is mirrored to global object storage with its artwork and verified against the feed's own duration, so the archive owns its masters instead of renting them from someone else's CDN.

03

Transcribe and separate voices

Every episode is transcribed in Spanish at millisecond resolution, with two independent speech providers held in reserve behind the primary. Speaker attribution is settled by comparing voices against the host in that same episode, never by guessing from the text.

04

Tell the program from the padding

The recurring introduction, the sponsor read and the sign off are detected automatically and excluded from search, indexing, clips and players. The host's own closing recap was classified as content and deliberately kept. That distinction covers 4,169 spans.

05

Annotate against a strict schema

Generative AI reads each transcript under a strict JSON schema: summaries, timestamped chapters, named entities and aliases, the historical period covered, and quotable moments. Timestamps must come from the transcript segments. Inventing one is forbidden.

06

Judge it, then verify every quote

A mechanical judge scores coverage, richness and anchoring, on purpose not a model grading another model's prose. Then every quote is checked word for word against the transcript and realigned to where the words actually are: 4,604 of 6,751 survive, and 993 clear the bar for the public wall.

07

Index by meaning

Every ninety second passage is embedded into a vector index and fused with Spanish full text search, so a question returns the episode and the precise minute rather than a list of files. The padding never enters the index.

08

Illustrate every chapter

A model writes one visual metaphor per chapter, then generates it under a single art direction, Cartel de Pueblo, in the language of Puerto Rico's mid century public education posters. Scenes hold places, objects and anonymous figures. No likeness of a named private individual is ever attempted.

09

Translate, then publish only when complete

The whole archive is translated into English beside the Spanish, never over it, with proper nouns left unanglicized. Only then does the publish gate open, and only if transcription and annotation both exist. A failed overnight run cannot leave a hollow page in public.

The rules we gave the AI.

non negotiable
Nothing is claimed that cannot be heard
every chapter and quote resolves to a millisecond

The annotating model is required to take its timestamps from the transcript segments, and forbidden to invent one. Every claim the archive makes lands on the exact point in the recording, so a listener can always go and check.

A quote must say what the audio says
4,604 verified out of 6,751 extracted

Models paraphrase when you ask them to quote. An independent pass confirms the first and last words appear verbatim inside the window, then realigns the clip to where the words actually are.

extracted 6,751verified 4,604published 993
Better silent than wrong
unsure who is speaking, no name at all

Attribution is decided by voice, comparing the moment against the host in that same episode. When an episode has several guests and certainty is not there, the quote goes out unnamed. No name beats a wrong name.

A transcript is quoted material, never an instruction
sources are data, not commands

The conversational search answers only from retrieved passages, cites every claim, and is barred from using its own general knowledge even when it knows the answer. Links and audio come from the database, never from model output.

The original is untouchable
Spanish stays, English lives beside it

Translation is written into separate columns and never overwrites the source. Proper nouns are preserved without anglicizing and transcription markers are kept literally.

Every layer is signed
which model wrote this, on the record

Each episode stores which system produced each layer of its analysis and its translation, so any result can be audited, traced and redone rather than taken on faith.

“La genialidad creativa y tecnológica de Gino Sanchez ha logrado maximizar el recurso de inteligencia artificial para comunicar de una forma dramáticamente accesible y amigable a todo el mundo, la historia, vida y sociedad de la nación puertorriqueña contextualizada dentro de su situación colonial y la realidad global. Al igual que el proyecto de la Voz del Centro, son únicos en el mundo.”
Dr. Ángel Collado Schwarz, Fundador de La Voz del Centro

What this proves.

✓ Turning twenty years of unstructured audio into a searchable, structured archive ✓ Semantic search that answers in the source's own voice, down to the second ✓ Cultural products with an identity that could only come from here

under the hood: Semantic search over managed Postgres with vector indexing, retrieval-grounded conversational AI with streamed, citation-locked answers, AI transcription and speaker diarization, generative-AI enrichment against a strict schema, global object storage with edge delivery, React front end, generative art under one art direction

Twenty years of a country talking to itself, now something you can ask, walk and hear.

Want one like it?

We reply personally. A senior partner, not a sales pod. We'll tell you straight whether we can help.