Posted by leumon 1 day ago
There's one or two I find more understated but I would love a 2026 SOTA V2V model that speaks clearly but without the artificial personality layered on.
Human interaction/theory of mind relies so much on non-verbal clues for interpreting emotion/intent and so for me having those neurons firing constantly while talking to an LLM just for an emotional no-op is exhausting to put up with for more than a couple minutes.
There's one male voice that would make me assume someone was sarcastically mocking me if I was talking to an actual person because it's just so over the top.
I don’t want to be aroused by my turn-by-turn street directions, thanks.
Let me tell you unlike every other mentioned model Gemini 3.8 Flash trial had to be reverted the same day. Instead of simply delegating tasks it would invent additional requirements and implementation details it knew nothing about and no amount of convincing not to do it would work. That's the first time a model failed on me so spectacularly despite having practically same Artificial Analysis Intelligence Index as another model that just worked (and higher than working DS Flash).
The reason I think it is relevant is: Live is likely even stupider model in every way possible (except hearing better than separate STT). So beware using it for agentic scenarios.