MyBook Podcasts (unlabeled) Unlabeled Russian language podcast episodes scraped from mybook.ru, packaged as Parquet shards with audio bytes embedded. Each row of the dataset contains: audio — the podcast episode (MP3), decoded on the fly via the Audio feature. ~30 flat scalar metadata columns: id , name , slug , lang , duration sec , audio bytes , rating , rating votes , rating scores , read count , reviews count , citations count , main author , main actor , publisher name , genres names , tags csv , series name , subscription id , written dt , updated at , share url , default cover path , preview audio url , annotation plain , annotation html . raw metadata json — the full original page JSON from mybook (89 fields with all nested structures: counters, niche category info, full author/ actor records, tag covers, etc) serialized as a string. Loading examples This dataset is built incrementally over time.
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy