.@huggingface passed 5 million users pic.twitter.com/TBjSy8ZvZ3
— AK (@_akhaliq) 19 août 2024
.
@huggingface passed 5 million users
By
–
.@huggingface passed 5 million users pic.twitter.com/TBjSy8ZvZ3
— AK (@_akhaliq) 19 août 2024
.
@huggingface passed 5 million users
By
–
.
@ManliShu and @Le_Xue01 will be presenting xGen-MM (BLIP-3) from Salesforce live today at 8 PM PST on X live broadcast
By
–
New Update, T2V results from CogVideoX + VEnhancerhttps://t.co/vTwpiJXvPw
— AK (@_akhaliq) 19 août 2024
Support enhancement for abitrary long videos (by spliting the videos into muliple chunks with overlaps); Fewer sampling steps (15) are enabled without obvious quality loss by setting –solver_mode 'fast'… pic.twitter.com/k32HTS3g0j
New Update, T2V results from CogVideoX + VEnhancer https://
github.com/Vchitect/VEnha
ncer?tab=readme-ov-file#-news
… Support enhancement for abitrary long videos (by spliting the videos into muliple chunks with overlaps); Fewer sampling steps (15) are enabled without obvious quality loss by setting –solver_mode 'fast'

By
–
Salesforce presents xGen-MM (BLIP-3) A Family of Open Large Multimodal Models discuss: https://
huggingface.co/papers/2408.08
872
… This report introduces xGen-MM (also known as BLIP-3), a framework for developing Large Multimodal Models (LMMs). The framework comprises meticulously curated

By
–
Automated Design of Agentic Systems discuss: https://
huggingface.co/papers/2408.08
435
… Researchers are investing substantial effort in developing powerful general-purpose agents, wherein Foundation Models are used as modules within agentic systems (e.g. Chain-of-Thought, Self-Reflection,
By
–
TurboEdit
— AK (@_akhaliq) 19 août 2024
Instant text-based image editing
discuss: https://t.co/PioX45KQQV
We address the challenges of precise image inversion and disentangled image editing in the context of few-step diffusion models. We introduce an encoder based iterative inversion technique. The… pic.twitter.com/morbXOR609
TurboEdit Instant text-based image editing discuss: https://
huggingface.co/papers/2408.08
332
… We address the challenges of precise image inversion and disentangled image editing in the context of few-step diffusion models. We introduce an encoder based iterative inversion technique. The

By
–
JPEG-LM LLMs as Image Generators with Canonical Codec Representations discuss: https://
huggingface.co/papers/2408.08
459
… Recent work in image and video generation has been adopting the autoregressive LLM architecture due to its generality and potentially easy integration into multi-modal
By
–
Has anyone tried cogvideo-2b + venhancer? pic.twitter.com/85aEip1SY5
— AK (@_akhaliq) 18 août 2024
Has anyone tried cogvideo-2b + venhancer?

By
–
3DGen-Arena : Benchmarking Text-to-3D generative models https://
huggingface.co/spaces/ZhangYu
han/3DGen-Arena
…