How we made a text-to-speech model respond in sub-50 ms
Nari labs
,Low-latency, realtime multimodal model serving, starting with speech at 50 ms time to first audio.
关注
添加评论
点赞
收藏
点踩
分享
查看原文
评论
?
参与讨论