Hacker Newsnew | past | comments | ask | show | jobs | submit | toebee's commentslogin

if we can lower the pricing by not 50% but 10x, then I think it would be something people want. we are taking the bet that OSS models will take a huge chunk of market share not just in LLMs but in multimodal as well

If audio only then 10x is 100% possible right now based on math of the services.

+ Video is unlikely unless they are willing to give up margins.

Didn't take long at all for Google to smash out with 50% lol


thank you Dylan!

agreed. we've been doing some work around NVIDIA personaplex 7b, but its quality is quite far from GPT-Live-1, esp in terms of intelligence. Once a good OSS model is out, we'll be sure to be the first to serve it cheaply to the masses :)

awesome! will look into Darwin TTS. super interesting

the qwen3-asr inference repo is not OSSed as of now. we're planning to write a paper or tech report on it as it contains some general techniques for ASR inference.

how do I follow you? I have a small 5090 doing inference all the time and I barely use tts but a lot of asr, mostly whisper, I ported your tech report for tts and implemented some improvements on my whisper inference based on your tech report as well!

would love to talk sometime!


hey, thanks for letting us know! will look into the issue and see what went wrong.

very cool. will try for local use!

thanks for the interest! we have a blog post on exactly how we did it: https://narilabs.com/blog/qwen3-tts-speed-cost-frontier/

thanks for the feedback! will investigate and get it fixed

hey sorry about that, you can here a few of our voices here: https://narilabs.com/product/qwen3-tts/

The horizontal moving elements of examples become stuck and unable to be scolled once one of them is played. I'm using Vivaldi (chrome based) on Android

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: