Extremely frothy paper. The one representation to rule them all? https://t.co/L2OiXaJ7Kn
— Bilawal Sidhu (@bilawalsidhu) 29 avril 2026
Extremely frothy paper. The one representation to rule them all?

By
–
Extremely frothy paper. The one representation to rule them all? https://t.co/L2OiXaJ7Kn
— Bilawal Sidhu (@bilawalsidhu) 29 avril 2026
Extremely frothy paper. The one representation to rule them all?
By
–
Try text-to-motion v1 and v2: http://
replicate.com/uthana/text-to
-motion-vqvae-v1
… http://
replicate.com/uthana/text-to
-motion-diffusion-v2
…
By
–
3D animation and rigging models from @Uthana_Inc are on Replicate!
— Replicate (@replicate) 29 avril 2026
Turn text into production-ready 3D animation instantly. Just describe a move and get clean FBX/GLB files for Unity or Unreal.
Then, take any generated bipedal 3D character and automatically rig it in under 30… pic.twitter.com/v6k4vAZcZI
3D animation and rigging models from @Uthana_Inc are on Replicate! Turn text into production-ready 3D animation instantly. Just describe a move and get clean FBX/GLB files for Unity or Unreal. Then, take any generated bipedal 3D character and automatically rig it in under 30

By
–
DeepSeek’s multimodal model is now live, and some users are already able to try it out. So far, it’s performing pretty well. Now the question is: will this one also be open-sourced? Tomorrow might be a good time!

By
–
GPT 5.5 is significantly better than the earlier versions of GPT 5. Surprised that 5.6 is already on the way. GPT 5.5 has narrowed the gap with Claude Opus 4.7 albeit Claude in my experience as a still more careful in detailed attention to code and other matters. However, IMHO

By
–
So much important insight in this quote and chart. This lesson will be repeated with many other jobs over the coming decade.

By
–
It's especially strong at structured visuals:
> Infographics
> PPT slides
> Data-heavy posters and diagrams One model, one prompt, information-dense, formatted output. No separate design tools. No prompt chaining between systems.

By
–
SenseNova-U1 thinks in images > Not "text in – image out." It generates images natively, mid-reasoning, as part of the thinking chain. > Input question – Interleaved text + Image reasoning – Structured output.

By
–

SenseTime open-sourced SenseNova-U1, a multimodal image generation model built on NEO-Unify! This architecture drops the visual encoder and VAE entirely. It generates images natively as one system that can handle understanding, reasoning, and generation processes. @SenseTime_AI

By
–
We may experience an explosion in litigation around AI systems over the next 2 years. Claude is awesome in particular in speeding up coding and research as is GPT 5.5. But even the best LLMs make mistakes in production code that requires experienced devs to detect, and plug