Here is the Deep Think "Sparks unicorn" (This is created using TikZ, which is a language built for scientific diagrams & very much not for drawing. The original "Sparks of AGI" paper used the ability of the AI to draw a primitive unicorn as an example of unexpected AI abilities)
MULTIMODAL AI
-
OmniHuman from ByteDance Now Available on Replicate Platform
By
–
OmniHuman from ByteDance is now on Replicate.https://t.co/DgcTrHNMIY
— Replicate (@replicate) 1 août 2025
Give it an image with a character and some audio, and it'll create a video of them saying those words. It's good with emotion, gestures and can also bring backgrounds to life a little, like the fish in this… pic.twitter.com/5iHxZWdZmFOmniHuman from ByteDance is now on Replicate. https://
replicate.com/bytedance/omni
-human
… Give it an image with a character and some audio, and it'll create a video of them saying those words. It's good with emotion, gestures and can also bring backgrounds to life a little, like the fish in this -

Horizon-alpha: Solid Performance on Multiple AI Tasks
By
–
Having played with it a bunch, Horizon-alpha did a pretty solid version of Missile Command with Relativistic Effects with a few rounds of feedback, passed the Lem Test the first time (without reasoning), and drew a passable TikZ unicorn (if you know, you know). Very quick model.
-
Claude Mobile Adds Native Calendar and Messaging Tools
By
–
The first feature is the ability for Claude on mobile to trigger a human-in-the-loop native app dialog for adding a calendar event or sending a message or email – that's implemented using two new tools, event_create_v0 and message_compose_v0 https://t.co/qhksbvwxQj
— Simon Willison (@simonw) 31 juillet 2025The first feature is the ability for Claude on mobile to trigger a human-in-the-loop native app dialog for adding a calendar event or sending a message or email – that's implemented using two new tools, event_create_v0 and message_compose_v0
-
AI Artifacts Analysis on Claude Now Available
By
–
AI-powered Artifacts on Claude now allow users to upload images, text and PDF files for analysis.
— 🚨 AI News | TestingCatalog (@testingcatalog) 31 juillet 2025
Claude inside Claude 👀 https://t.co/b6FlNN6IXl pic.twitter.com/QBN1LsTcePAI-powered Artifacts on Claude now allow users to upload images, text and PDF files for analysis. Claude inside Claude
-
Flow AI Tool Discovers New Visual Prompting Capability
By
–
We are always playing and experimenting with our products and sometimes this leads to discovering new, unexpected capabilities.
— Google AI (@GoogleAI) 31 juillet 2025
Here @nmatares a designer on our AI filmmaking tool Flow, doodled on an image and uploaded it, unlocking a whole new way to prompt!
Try it out in… pic.twitter.com/46OKirnzw9We are always playing and experimenting with our products and sometimes this leads to discovering new, unexpected capabilities. Here @nmatares a designer on our AI filmmaking tool Flow, doodled on an image and uploaded it, unlocking a whole new way to prompt! Try it out in
-
AI Model Excels Across Languages and Receipt Recognition
By
–
Does super well on that too, any language too, including receipts.
-
Gemini video demo after running a prompt
By
–
3/ Gemini after I ran the prompt (video demo):https://t.co/CcsRScdm5t
— God of Prompt (@godofprompt) 31 juillet 20253/ Gemini after I ran the prompt (video demo):
-
DreamScene: 3D Gaussian Text-to-3D Scene Generation
By
–
DreamScene
— AK (@_akhaliq) 31 juillet 2025
3D Gaussian-based End-to-end Text-to-3D Scene Generation pic.twitter.com/9fEDUEj0uRDreamScene 3D Gaussian-based End-to-end Text-to-3D Scene Generation
-

Voice as the Original and Future Modality of Communication
By
–
"Voice is our past as much as it is our future" Although roon's "text is the universal interface" proved true for LLMs, I do think voice is the "OG modality" for communication between intelligent species: – Before writing was invented, we were speaking for over 100,000 years.
–
