• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
AI
Report

YODAS v3 speech dataset announced with 1.1 million hours of audio

A September 28 post says YODAS v3 is available on Hugging Face, with 48 kHz stereo audio, timestamped transcripts and translations, and 100-plus languages.

Hugging FaceHF
2 Sources, 12d ago, first seen 12d ago

TLDR

A September 28 post announces YODAS v3 on Hugging Face, claiming 1.1 million hours of audio and calling it the biggest audio dataset ever. It says the dataset is the first at this scale with 48 kHz stereo audio, and lists timestamped transcripts, translations, 100-plus languages and a CC-BY-3.0 license.

Combined views

1.3K

2 Sources, first seen 12d ago

34 reposts

Combined views

1.3K

2 Sources, first seen 12d ago

34 reposts

Sentiment

Positive——Negative

Summary

Not enough discussion yet.

No sentiment analysis available yet.

Featured Source

Sentiment

Positive——Negative

Summary

Not enough discussion yet.

No sentiment analysis available yet.

Related

Hugging Face download rates reportedly drop about 30% for many tracked top open LLMs

A user monitoring the models says download counting may have changed slightly, but the decline has continued.

Hugging Face adds shared discovery for reinforcement-learning environments

Hugging Face’s framework tags identify compatible tasksets, while execution stays with the framework and its runtime.

Liquid AI unveils d1, claimed to be the first model to outperform Jev on Hugging Face's Decision Index

Liquid AI says d1 wins on multilingual evaluations, is more robust against prompt injection and handles longer inputs more effectively. It lists access through the Liquid API.

2 Sources

Hugging Face@huggingfaceRT @chenwanch1: Who wants more speech training data? YODAS v3 is now available @huggingface It’s 1.1M hours - the biggest audio dataset e…12d
Kyle Kastner@kastnerkyleRT @chenwanch1: Who wants more speech training data? YODAS v3 is now available @huggingface It’s 1.1M hours - the biggest audio dataset e…11d
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    YODAS v3William ChenHugging Face

    2 Sources

    Hugging Face@huggingfaceRT @chenwanch1: Who wants more speech training data? YODAS v3 is now available @huggingface It’s 1.1M hours - the biggest audio dataset e…12d
    Kyle Kastner@kastnerkyleRT @chenwanch1: Who wants more speech training data? YODAS v3 is now available @huggingface It’s 1.1M hours - the biggest audio dataset e…11d
    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet