Menu

Post image 1
Post image 2
Post image 3
Post image 4
Post image 5
Post image 6
Post image 7
Post image 8
Post image 9
1 / 9
0

owensong/Inflect-Micro-v2 · Hugging Face

Hacker News·2 months ago
#fIHRK6QT
Reading 0:00
15s threshold

Inflect-Micro-v2 Complete local text-to-waveform speech synthesis under 10M parameters. Fixed-voice English TTS with deterministic seeds, long-text handling, and CPU or CUDA inference. A note from Owen I built and funded Inflect v2 independently. If this release finds a real audience, I would like to continue the project with a broader v3, which might include things like more langauges, voices, and stability improvements. If the model is useful to you, leaving a like on Hugging Face genuinely helps more people discover it. 9,356,513 deployable parameters · 37.53 MB FP32 · 24 kHz mono output Inflect v2 uses one public API across two sizes: Micro prioritizes quality below 10M parameters; Nano prioritizes footprint below 4M. Explore this model card Listen These are held-out text generations, not reconstructions of training audio. Each transcript is shown exactly as passed to the public frontend.…

Continue reading — create a free account

Join HashtagPLUS to read full articles, follow hashtags, vote, and join the conversation.

Read More