به Nostr بپیوندید
2026-05-26 07:21:31 CEST
in reply to

solo on Nostr: > (them) That's very low quality though, and our feedback from blind users indicated ...

> (them) That's very low quality though, and our feedback from blind users indicated that old TTS methods like that aren't usable (such as RHVoice).

> See > [> https://>; social.highenergymagic.net/@Gr> [email protected]/116638788131682457](/116638788131682457">https://social.highenergymagic.net//116638788131682457 )

> It will be improved, but there's a starting point for everything. We didn't want to delay it too much, but the current target was similar latency to Google's Speech Recognition & Synthesis, but perhaps it isn't as fast at higher speeds?

> Let's focus on actionable goals! Where could the latency be improved? Is it perhaps too slow at very high speeds?

> (me) maybe a setting could be offered so that users can choose which they want?

> (them) Maybe! Configuring the amount of "effort" to make it faster was something we wanted but it wasn't working out initially and we didn't want to delay the release for that.

> But what are the exact latency issues? We need clear targets.

> for example, currently the TTFA (time-to-first-audio) is ~150 ms on a Pixel 8a.