Llama 3.2 multimodal likely will be announced very soon h/t @danielhanchen
@testingcatalog
-
User update on Advanced Voice Mode availability
By
–
Nope – Advanced voice mode is not there yet (for me at least)
-
Implementing fallback logic for AI agent systems
By
–
Yes, I think this part needs a bit more polishing. It should fallback to standard if advanced fails after several attempts
-

New Gemini 1.5 Pro and Flash models released on Google AI Studio
By
–
ICYMI: New Gemini-1.5-Pro-002 and Gemini-1.5-Flash-8B-Exp-0924 models are now available on Google AI Studio
-

Android ChatGPT App Adds New Shortcuts for Voice and Camera
By
–

ICYMI: Latest Android ChatGPT app got new icon shortcuts pointing to Voice and Camera deep links Camera one is interesting here cuz it might become a shortcut into an upcoming voice UI with vision capability in the future! https://
t.co/EPXGZvEbQT -
Using AI to bridge the gap between thought and text
By
–
I was never a fun of typing, AI can understand my brain fart much faster than my own hands can transmit it into a keyboard in a meaningful way
-
Speculation on AI model audio generation capabilities
By
–
Do we expect it to be able to produce sounds any time soon? Like in early days of alpha
-
Predicting the future growth of the AI voice market
By
–
I am predicting a big "voice" market to be the case in the future
