We're hiring!
*

WhisperSpeech makes its way to AI.dev

Mark Filion avatar

Mark Filion
December 07, 2023

Share this post:

Reading time:

Collabora is headed to San Jose, California, to take part in the inaugural edition of AI​.dev: Open Source GenAI & ML Summit, a new event which aims to bring together the brightest developers from around the world to shape the trajectory of open source AI.

Join us on Tuesday, December 12, as Jakub Piotr Cłapa dives into findings from WhisperSpeech, a new Open Source text-to-speech model developed by Collabora. Based entirely on properly licensed speech datasets and unrestricted Open Source code, the model's focus is to deliver the best natural-sounding Open Source speech synthesis solution for improved communication.

In this talk, Jakub will look at how Collabora scaled its models and training pipelines from hundreds to 80K+ hours of speech recordings, and will share lessons learned along the way. He'll also discuss some of the challenges encountered, including:

  • Gone in 16 minutes: the importance of small scale experiments.
  • Full throttle: is 100% GPU utilization enough?
  • Do you need a fancy framework? From single- to multi-GPU training.
  • Are SSDs fast enough? WebDataset brings a 10x improvement.
  • Does bigger always mean better? How to effortlessly scale AI models.
  • Clouds, enthusiasts or clusters? How to hunt down GPUs.
  • Defending moats. How is a gaming 4090 different from an H100?

If you plan on attending, please make sure to come say hello! Note that so you can also watch Jakub's talk remotely via the Room LL20D live stream.

Update: The video recording is now available, click on the link below to start watching!

Collabora @ AI​.dev: Open Source GenAI & ML Summit

Tricks Learned from Scaling WhisperSpeech Models to 80k+ Hours of Speech
Presented by Jakub Piotr Cłapa - Tuesday, December 12

 


Add a Comment

 

Search the newsroom

Latest News & Events

Collabora at Embedded World 2026: Open Source AI and Embedded Innovation

05/03/2026

As champions of open source development in the embedded community, Collabora will be at Booth 4-404 with an impressive lineup of live demonstrations…

RK3588 and RK3576 video decoders support merged in the upstream Linux Kernel

25/02/2026

Support for Rockchip’s VDPU381 and VDPU383 decoders is now upstream in Linux, bringing mainline H.264/HEVC decode support, robust IOMMU-reset…

Weston 15.0 is here: Lua shells, Vulkan rendering, and a smoother display stack

19/02/2026

Weston 15.0 has arrived, bringing a brand new Lua-based shell for fully customizable window management, an experimental Vulkan renderer,…

Open Since 2005 logo

Our website only uses a strictly necessary session cookie provided by our CMS system. To find out more please follow this link.

Collabora Limited © 2005-2026. All rights reserved. Privacy Notice. Sitemap.