AI Dynamics

Global AI News Aggregator

About

Lumina-mGPT: Multimodal Generative Pretraining for Text-to-Image

Lumina-mGPT Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining paper page: https://
huggingface.co/papers/2408.02
657
… We present Lumina-mGPT, a family of multimodal autoregressive models capable of various vision and language tasks, particularly

→ View original post on X — @_akhaliq