Denoising Reuse Exploiting Inter-frame Motion Consistency for Efficient Latent Generation discuss: https://
huggingface.co/papers/2409.12
532
… generation using diffusion-based models is constrained by high computational costs due to the frame-wise iterative diffusion process. This
@_akhaliq
-

Denoising Reuse: Efficient Video Generation Through Motion Consistency
By
–
-

GRADIO-CODER: Instruction Fine-Tuned StarCoder for Gradio Development
By
–
GRADIO-CODER model: https://
huggingface.co/matrixglitch/G
RADIO-CODER
… This language model is the version 0.0 of a Gradio Coding Assistant. It is an instruction fine-tuned version of StarCoder that is designed to provide assistance to developers who use gradio. -

3DTopia-XL: High-Quality 3D PBR Asset Generation via Primitive Diffusion
By
–
3DTopia-XL
— AK (@_akhaliq) 19 septembre 2024
High-Quality 3D PBR Asset Generation via Primitive Diffusion
demo: https://t.co/5Tbr3xJ6C9
model: https://t.co/j0pUloClEG
3DTopia-XL scales high-quality 3D asset generation using Diffusion Transformer (DiT) built upon an expressive and efficient 3D representation,… pic.twitter.com/nhFGrudqr23DTopia-XL High-Quality 3D PBR Asset Generation via Primitive Diffusion demo: https://
huggingface.co/spaces/FrozenB
urning/3DTopia-XL
…
model: https://
huggingface.co/FrozenBurning/
3DTopia-XL
… 3DTopia-XL scales high-quality 3D asset generation using Diffusion Transformer (DiT) built upon an expressive and efficient 3D representation, -

Updated Hacker News AI Papers Ranking Logic Using o1-mini
By
–
updated ranking logic for Hacker News AI papers using o1-mini
-

Chain-of-Thought Prompting: Benefits Limited to Math and Reasoning
By
–
To CoT or not to CoT? Chain-of-thought helps mainly on math and symbolic reasoning discuss: https://
huggingface.co/papers/2409.12
183
… Chain-of-thought (CoT) via prompting is the de facto method for eliciting reasoning capabilities from large language models (LLMs). But for what kinds of tasks -

LLMs with Persona-Plug for Personalized Language Models
By
–
LLMs + Persona-Plug = Personalized LLMs discuss: https://
huggingface.co/papers/2409.11
901
… Personalization plays a critical role in numerous language tasks and applications, since users with the same requirements may prefer diverse outputs based on their individual interests. This has led to the -

Qwen2-VL: Advanced Vision-Language Model with Dynamic Resolution
By
–
Qwen2-VL Enhancing Vision-Language Model's Perception of the World at Any Resolution discuss: https://
huggingface.co/papers/2409.12
191
… We present the Qwen2-VL Series, an advanced upgrade of the previous Qwen-VL models that redefines the conventional predetermined-resolution approach in visual -

Takin: Superior Quality Zero-shot Speech Generation Models
By
–
Takin A Cohort of Superior Quality Zero-shot Speech Generation Models discuss: https://
huggingface.co/papers/2409.12
139
… With the advent of the big data and large language model era, zero-shot personalized rapid customization has emerged as a significant trend. In this report, we introduce -

Microsoft Releases GRIN MoE Mixture of Experts Model
By
–
Microsoft releases GRIN MoE GRadient-INformed MoE demo: https://
huggingface.co/spaces/GRIN-Mo
E-Demo/GRIN-MoE
…
model: https://
huggingface.co/microsoft/GRIN
-MoE
…
github: https://
github.com/microsoft/GRIN
-MoE
… With only 6.6B activate parameters, GRIN MoE achieves exceptionally good performance across a diverse set of tasks, particularly in -
Paradigm Shift Demonstrated in Gradio Demo
By
–
paradigm shift shown in a @Gradio demo build with gradio