Jesse Zhang

speaker
306 appearances 1 recordings 1 series first heard Jan 2025 last heard Jan 2025

Jesse Zhang’s voice in public audio — every appearance, attributed to the second.

Trend

recordings per month · last 12 months
No recordings in the last 12 months.Older appearances are listed below; set an alert to hear about the next one.

Appearances

newest first · ▶ plays the moment
Big uh
big topic for us.
And as these companies get better, I mean we're working with them pretty closely right now on how
how you can actually like build these things well at scale.
Uh but as they get better, that that's also going to be huge for us to keep delivering these voice agents.
Lac latency is a big, big deal here, of course, uh with voice models.
So nowadays you have the the voice to voice models that we're playing around with.
Uh OpenAI is doing a lot of work here.
I think there's obviously a lot of trade offs there.
Voice to voice, latency is great.
Uh sometimes though with these production use cases, you do need the extra computation cycles.
So, you know, fetch data, do multiple model calls.
Um or y there's there might be other reasons that you c you can't do voice to voice and so okay, th that's that's one option that you you would consider.
Uh the other one is the the one you described where you're kind of going through your transcribing or yeah, doing
speech to text and then doing all the computation within text and then generating the voice at the end.
That always causes a little bit of extra latency, of course.
And so as you mentioned, a lot of folks have figured out fairly clever ways to to get around that.
You can, you know, start generating stuff first.
Um in in our use case you can always do something like, Hey, give me a sec, I'm looking up your data.
Um so these are all things we're playing around with.
Showing 141–160 of 306 · page 8 of 16 ← Previous Next →