Transcripts, metadata, captions, and translations from the Active Inference Institute's video library — a structured, source-namespaced dataset generated by Journal-Utilities.
-
Updated
Aug 21, 2026 - HTML
Transcripts, metadata, captions, and translations from the Active Inference Institute's video library — a structured, source-namespaced dataset generated by Journal-Utilities.
Lightweight public data registry for the Samuel & Audrey Media Network, listing datasets, archive records, websites, and research-facing resources with structured CSV/JSONL directory files, schema, citation, license, manifest, checksums, and llms files.
Official transcript corpus from the Nomadic Samuel YouTube channel, with 143 English travel video transcripts, metadata, SRT payloads, and CSV/JSONL exports.
English transcript corpus from the Samuel & Audrey YouTube travel and food channel, featuring 1,397 full video transcripts and 233,285 cue-level segment records from 2012–2026, with metadata, SRT payloads, CSV/JSONL exports, schema, citation, license, manifest, checksums, and llms files.
Spanish-first bilingual transcript corpus from the Samuel y Audrey YouTube channel, featuring 643 travel video records with Spanish and English transcripts, SRT payloads, metadata, CSV/JSONL exports, schema, citation, license, manifest, checksums, and llms files.
A lightweight CLI tool that extracts and displays transcripts from any YouTube video using Bun runtime.
Add a description, image, and links to the video-transcripts topic page so that developers can more easily learn about it.
To associate your repository with the video-transcripts topic, visit your repo's landing page and select "manage topics."