One-Minute Generation with Test-Time Training This paper augments Transformers with Test-Time Training (TTT) layers—neural networks used as hidden states—to generate coherent one-minute videos from text storyboards. Problem: Long video generation is bottlenecked by
GENERATIVE AI
-

Hogwild! Inference: Parallel LLM Generation via Concurrent Attention
By
–
Hogwild! Inference: Parallel LLM Generation via Concurrent Attention This paper introduces Hogwild! Inference, a parallel LLM inference framework where multiple model instances collaborate by sharing a synchronized attention cache and dynamically adapting their strategies in
-

Top 10 Papers: Video Generation and Multimodal AI Models
By
–
Huge week for video generation and multimodal models, with detailed one-minute video generation and more efficient approaches to multimodality Check out the top 10 papers for the week – Hogwild! Inference: Parallel LLM Generation via Concurrent Attention
– One-Minute -
Gemini 2.0 Flash Enables Native Image Generation on Poe
By
–
Native image generation is available in Gemini 2.0 Flash Experimental at https://
poe.com/Gemini-2.0-Fla
sh-Exp
… and on all platforms where Poe is available. (3/3) -
Generate Text and Images Together in Single Request
By
–
You can also generate text and images together in one request. For example, it can create a short story about a topic while generating images for each scene based on your direction. (2/3) pic.twitter.com/KKbcfiz8Z1
— Poe (@poe_platform) 12 avril 2025You can also generate text and images together in one request. For example, it can create a short story about a topic while generating images for each scene based on your direction. (2/3)
-
Gemini 2.0 Flash Native Image Generation Capability
By
–
New: Gemini 2.0 Flash native image generation!
— Poe (@poe_platform) 12 avril 2025
This bot supports image output and conversational editing, allowing you to create and refine images by describing what you want. (1/3) pic.twitter.com/NfqlLAkTZpNew: Gemini 2.0 Flash native image generation! This bot supports image output and conversational editing, allowing you to create and refine images by describing what you want. (1/3)
-

DataAISummit 2026: 700+ Sessions on Data, AI, and Analytics
By
–
Join #DataAISummit for access to 700+ sessions on data intelligence, data warehousing, AI, governance and more! Register now to secure your spot at the world’s largest data, analytics and AI conference. Register now – early bird ends April 30! Here’s a preview of the sessions
-
SE2 and SE1 AI Model Updates: Smaller Releases Incoming
By
–
Those SE2 updates (VS 1.1 and upcoming VS 1.2) are much smaller than the usual SE1 updates. Also, the next SE1 update is very near.
-
Google offers free prompt engineering learning here
By
–
Learn prompt engineering here for free by Google. Impressive.
-
Chinese TTS Models Rankings on Huggingface Hub
By
–
本来想搜一下最近主打情感的中文 TTS 查资料的时候,顺便去 Huggingface 查了下 中文区 TTS
下载量,排名竟然非常出乎意料,下面是前五名 第一名竟然是 suno/bark
第二名还是 suno/bark-small
第三名是 spark-tts-0.5b
第四名是 fish-speech-1.5
第五名是
