Give it a shot, 3.1 live one in AI studio/API and max out reasoning - not the one in Gemini app...

dharma1 • yesterday at 10:32 PM • 1 reply • view on HN

Give it a shot, 3.1 live one in AI studio/API and max out reasoning - not the one in Gemini app it’s an older model.

Another option is to use pipecat with their VAD and separate STT and TTS and any (fast) LLM of your choice - but it’s more plumbing and not a true speech to speech model

Replies

stavros • yesterday at 11:59 PM

Haha, wow, I never thought I'd see a voice model that was too quick, but 3.1 live felt like it responded unnaturally quickly! I'm kind of blown away, I'd want to insert a 100ms delay to make it sound more natural, wow. I never thought I'd see that.

alt Hacker News

Replies