3 ms·
I think voice mode uses weaker models, just an FYI relative to the SOTA
by drakenot 6mo ago
I think voice mode uses weaker models, just an FYI relative to the SOTA
- scrollop 6mo agoDefinitely, seems like gpt 3
- SOLAR_FIELDS 6mo agoCan get around this with a local STT model and use text input but UX is probably clunkier
- pxc 6mo agoThe bigger problem for me is that the realtime voice modes lack tool use, so they can't look anything up or do anything. Model strength definitely also matters, but even dumb models can be helpful when they can look things up and try things out. And smart models that don't do those things kinda suck.