Scaling Autoregressive Multi-Modal Models: Pretraining and Instruction Tuning blog: https://
ai.meta.com/blog/generativ
e-ai-text-images-cm3leon/
… present CM3Leon (pronounced “Chameleon”), a retrieval-augmented, tokenbased, decoder-only multi-modal language model capable of generating and infilling both text and
@_akhaliq
-

CM3Leon: Scaling Autoregressive Multi-Modal Models for Text and Images
By
–
-

Hugging Face Papers Daily Email Released July 14
By
–
https://
huggingface.co/papers daily email just went out for 14 July -

HyperDreamBooth: Fast Text-to-Image Model Personalization with HyperNetworks
By
–
HyperDreamBooth: HyperNetworks for Fast Personalization of Text-to-Image Models paper page: https://
huggingface.co/papers/2307.06
949
… Personalization has emerged as a prominent aspect within the field of generative AI, enabling the synthesis of individuals in diverse contexts and styles, while -

Animate-A-Story: Retrieval-Augmented Video Generation for Storytelling
By
–
Animate-A-Story: Storytelling with Retrieval-Augmented Video Generation
— AK (@_akhaliq) 14 juillet 2023
paper page: https://t.co/Yvzd9mmMxN
Generating videos for visual storytelling can be a tedious and complex process that typically requires either live-action filming or graphics animation rendering. To… pic.twitter.com/4SDu3KAAWUAnimate-A-Story: Storytelling with Retrieval-Augmented Generation paper page: https://
huggingface.co/papers/2307.06
940
… Generating videos for visual storytelling can be a tedious and complex process that typically requires either live-action filming or graphics animation rendering. To -

Domain-Agnostic Tuning-Encoder for Fast Text-to-Image Personalization
By
–
Domain-Agnostic Tuning-Encoder for Fast Personalization of Text-To-Image Models paper page: https://
huggingface.co/papers/2307.06
925
… Text-to-image (T2I) personalization allows users to guide the creative image generation process by combining their own visual concepts in natural language -

T2I-CompBench: Comprehensive Benchmark for Compositional Text-to-Image Generation
By
–
T2I-CompBench: A Comprehensive Benchmark for Open-world Compositional Text-to-image Generation
— AK (@_akhaliq) 14 juillet 2023
paper page: https://t.co/K0AXSgLOuB
Despite the stunning ability to generate high-quality images by recent text-to-image models, current approaches often struggle to effectively… pic.twitter.com/EM1J4nLcORT2I-CompBench: A Comprehensive Benchmark for Open-world Compositional Text-to-image Generation paper page: https://
huggingface.co/papers/2307.06
350
… Despite the stunning ability to generate high-quality images by recent text-to-image models, current approaches often struggle to effectively -

In-context Autoencoder for Context Compression in Large Language Models
By
–
In-context Autoencoder for Context Compression in a Large Language Model paper page: https://
huggingface.co/papers/2307.06
945
… propose the In-context Autoencoder (ICAE) for context compression in a large language model (LLM). The ICAE has two modules: a learnable encoder adapted with LoRA from -

InternVid: Large-Scale Video-Text Dataset for Multimodal Learning
By
–
InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation paper page: https://
huggingface.co/papers/2307.06
942
… introduces InternVid, a large-scale video-centric multimodal dataset that enables learning powerful and transferable video-text representations for -

Generating Benchmarks for Factuality Evaluation of Language Models
By
–
Generating Benchmarks for Factuality Evaluation of Language Models paper page: https://
huggingface.co/papers/2307.06
908
… Before deploying a language model (LM) within a given domain, it is important to measure its tendency to generate factually incorrect information in that domain. Existing -

Distilling Large Language Models for Biomedical Knowledge Extraction
By
–
Distilling Large Language Models for Biomedical Knowledge Extraction: A Case Study on Adverse Drug Events paper page: https://
huggingface.co/papers/2307.06
439
… Large language models (LLMs), such as GPT-4, have demonstrated remarkable capabilities across a wide range of tasks, including health