submit your paper: https://
huggingface.co/papers/submit
@_akhaliq
-

Submit Your Paper to Hugging Face Papers
By
–
-
TacSL: Library for Visuotactile Sensor Simulation and Learning
By
–
TacSL
— AK (@_akhaliq) 14 août 2024
A Library for Visuotactile Sensor Simulation and Learning
discuss: https://t.co/LznliHFNRs
For both humans and robots, the sense of touch, known as tactile sensing, is critical for performing contact-rich manipulation tasks. Three key challenges in robotic tactile… pic.twitter.com/yxeq1GDZ5MTacSL A Library for Visuotactile Sensor Simulation and Learning discuss: https://
huggingface.co/papers/2408.06
506
… For both humans and robots, the sense of touch, known as tactile sensing, is critical for performing contact-rich manipulation tasks. Three key challenges in robotic tactile -
LongWriter: Enabling 10,000+ Word Generation from Long Context LLMs
By
–
LongWriter
— AK (@_akhaliq) 14 août 2024
Unleashing 10,000+ Word Generation from Long Context LLMs
discuss: https://t.co/UeebckjbtH
Current long context large language models (LLMs) can process inputs up to 100,000 tokens, yet struggle to generate outputs exceeding even a modest length of 2,000 words.… pic.twitter.com/uXzdZsG4RVLongWriter Unleashing 10,000+ Word Generation from Long Context LLMs discuss: https://
huggingface.co/papers/2408.07
055
… Current long context large language models (LLMs) can process inputs up to 100,000 tokens, yet struggle to generate outputs exceeding even a modest length of 2,000 words. -

OpenResearcher: AI Solutions for Accelerated Scientific Research
By
–
OpenResearcher Unleashing AI for Accelerated Scientific Research discuss: https://
huggingface.co/papers/2408.06
941
… The rapid growth of scientific literature imposes significant challenges for researchers endeavoring to stay updated with the latest advancements in their fields and delve into -

Google announces Imagen 3 latent diffusion model for text-to-image generation
By
–
Google announces Imagen 3 discuss: https://
huggingface.co/papers/2408.07
009
… We introduce Imagen 3, a latent diffusion model that generates high quality images from text prompts. We describe our quality and responsibility evaluations. Imagen 3 is preferred over other state-of-the-art (SOTA) -

Submit Your Paper to Hugging Face Papers
By
–
if I missed your paper today, you can submit it here: https://
huggingface.co/papers/submit -

rStar: Self-Play Mutual Reasoning Strengthens Small Language Models
By
–
Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers discuss: https://
huggingface.co/papers/2408.06
195
… This paper introduces rStar, a self-play mutual reasoning approach that significantly improves reasoning capabilities of small language models (SLMs) without fine-tuning or superior -
FruitNeRF: Neural Radiance Field Framework for 3D Fruit Counting
By
–
FruitNeRF
— AK (@_akhaliq) 13 août 2024
A Unified Neural Radiance Field based Fruit Counting Framework
discuss: https://t.co/op97PCJqjl
We introduce FruitNeRF, a unified novel fruit counting framework that leverages state-of-the-art view synthesis methods to count any fruit type directly in 3D. Our… pic.twitter.com/NfUNAOd2RnFruitNeRF A Unified Neural Radiance Field based Fruit Counting Framework discuss: https://
huggingface.co/papers/2408.06
190
… We introduce FruitNeRF, a unified novel fruit counting framework that leverages state-of-the-art view synthesis methods to count any fruit type directly in 3D. Our -

Med42-v2: Clinical LLMs Suite for Healthcare Applications
By
–
Med42-v2 A Suite of Clinical LLMs discuss: https://
huggingface.co/papers/2408.06
142
… Med42-v2 introduces a suite of clinical large language models (LLMs) designed to address the limitations of generic models in healthcare settings. These models are built on Llama3 architecture and fine-tuned -

CogVideoX: Text-to-Video Diffusion Models with Expert Transformer
By
–
CogVideoX Text-to-Video Diffusion Models with An Expert Transformer discuss: https://
huggingface.co/papers/2408.06
072
… We introduce CogVideoX, a large-scale diffusion transformer model designed for generating videos based on text prompts. To efficently model video data, we propose to levearge a