Categories: Image

We open-sourced Sopro V2 Turbo – a 120M voice cloning TTS model that runs 5x faster than real time on CPU

Sopro V2 Turbo is an open-source TTS model that runs locally.

  • Clones a voice from 5-20s of audio
  • ~300ms to first audio on a laptop CPU
  • English, European Portuguese, French, German

Local web UI:

uvx --from sopro soprotts serve

There’s also a Python API and a browser package (@soprotts/onnx-web) for WebGPU/WASM.

Repo: https://github.com/samuel-vitorino/sopro Benchmarks + samples: https://research.haloneuro.ai/posts/sopro-v2

Edit: Hugging Face kindly created a Space, making it even easier for you to try the model. You can try it here: https://huggingface.co/spaces/hugging-apps/sopro-v2-turbo-tts

submitted by /u/SammyDaBeast
[link] [comments]

AI Generated Robotic Content

Share
Published by
AI Generated Robotic Content
Tags: ai images

Recent Posts

Qwen Image 2.1 – 1girl examples

Hey guys! So, I got early access to Qwen Image 2.1, and I did what…

19 hours ago

The Black Friday-ification of the 4th of July: what a decade of email data told us about America’s 250th

The Black Friday-ification of the 4th of July: what a decade of email data told…

19 hours ago

Forget the AI Slowdown—the Vulnerability Explosion Is Already Happening

AI labs are toying with an industry-wide pact to slow development. Meanwhile, widely available AI…

20 hours ago

Will AI models achieve the ability to improve autonomously? Leading labs say the scenario is near

Once a distant ambition for technology researchers, the prospect of artificial intelligence models teaching themselves…

20 hours ago

as far as is know it does t2i and i2i

submitted by /u/dev_ne [link] [comments]

2 days ago

Build And Understand a Vector Database From Scratch in 10 Easy Steps

In this article, you will learn how a vector database works under the hood by…

2 days ago