TODAY'S AI NEWS: Google just released a new Gemini model with "thinking budget." Plus, more news from OpenAI, Alibaba, biotech company Profluent, and DeepMind. Here's everything you need to know:
@rowancheung
-
o3 Response Speed vs Google: When AI Feels Too Slow
By
–
It's funny cus deep down I feel the same way. But then o3 takes 2 mins to respond for a simple query I probably could've googled and I instantly switch
-
Gemini’s reasoning flaws revealed during public demo event
By
–
Demo would've been perfect if Gemini didn't recommend Lighthouse Park which is ~4 hours walk from the Convention Centre I'm assuming reasoning will get sorted before public launch though. Can't wait to try!
-

Grok Assistant Gets Memory Feature for Personalized Responses
By
–
Elon Musk's xAI started rolling out a memory feature into its Grok assistant (in beta)
— Rowan Cheung (@rowancheung) 17 avril 2025
Just like ChatGPT, Grok will reference past chats to provide personalized answers.
There's also a dedicated "forget" button to exclude specific chats from its memorypic.twitter.com/qCGeHkTfvNElon Musk's xAI started rolling out a memory feature into its Grok assistant (in beta) Just like ChatGPT, Grok will reference past chats to provide personalized answers. There's also a dedicated "forget" button to exclude specific chats from its memory
-

China’s Kling AI Releases 2.0 Models for Video and Images
By
–
China's Kling AI released two new models: KLING 2.0 Master for video generation and KOLORS 2.0 for images
— Rowan Cheung (@rowancheung) 17 avril 2025
Both come with improved prompt adherence, with KLING 2.0 standing out when dealing with prompts with sequential actions and complex motionspic.twitter.com/Bkyi29OUfZChina's Kling AI released two new models: KLING 2.0 Master for video generation and KOLORS 2.0 for images Both come with improved prompt adherence, with KLING 2.0 standing out when dealing with prompts with sequential actions and complex motions
-

Microsoft Rolls Out Copilot Vision in Edge Browser
By
–
Microsoft also started rolling out Copilot Vision in its Edge browser
— Rowan Cheung (@rowancheung) 17 avril 2025
It will read what's on screen to summarize aloud, working as a real-time collaborator/assistant when browsing the internet.
Best part: it's free—and opt-in (not active by default)!pic.twitter.com/5hVj5HtupYMicrosoft also started rolling out Copilot Vision in its Edge browser It will read what's on screen to summarize aloud, working as a real-time collaborator/assistant when browsing the internet. Best part: it's free—and opt-in (not active by default)!
-

Microsoft Adds Computer Use Capabilities to Copilot Studio
By
–
Microsft added computer use capabilities to its Copilot Studio The move will allow enterprises to build and deploy agents that can take UI actions on desktop and web apps! A big move!
-

Cohere Releases Embed 4 SOTA Multimodal Embedding Model
By
–
Cohere released Embed 4, a SOTA multimodal embedding model to add frontier search and retrieval capabilities to AI apps —128K-token context window
—Supports 100+ languages
—Optimized for data from regulated industries
—Up to 83% savings on storage costs -
Claude Research integrates Google Workspace for advanced data search
By
–
Anthropic added a Research feature in Claude with Google Workspace integration
— Rowan Cheung (@rowancheung) 17 avril 2025
Research will perform searches across the web and users’ connected work data
This data will also include users' emails, calendars, and docs, thanks to the Workspace linkpic.twitter.com/i46aYqnLwMAnthropic added a Research feature in Claude with Google Workspace integration Research will perform searches across the web and users’ connected work data This data will also include users' emails, calendars, and docs, thanks to the Workspace link
-

Google Expands Gemini Live Project Astra to All Android Users
By
–
Google is expanding Gemini Live's Project Astra capabilities—to all Android users!
— Rowan Cheung (@rowancheung) 17 avril 2025
This will enable users to engage with real-time visual AI and have multilingual conversations about anything seen and heard via their phone's camera or screen sharepic.twitter.com/TMj0A3pa7oGoogle is expanding Gemini Live's Project Astra capabilities—to all Android users! This will enable users to engage with real-time visual AI and have multilingual conversations about anything seen and heard via their phone's camera or screen share