featured in todays newsletter: https://
akhaliq.substack.com/p/trending-ai-
news-stories-papers-21e
…
@_akhaliq
-
Featured in today’s newsletter: Trending AI news stories and papers
By
–
-

Lamborghini Outdoor BBQ Grill Designed with Midjourney AI
By
–
Lamborghini products, midjourney AI 1. Outdoor BBQ Grill
-

Full Parameter Fine-tuning for Large Language Models with Limited Resources
By
–
Full Parameter Fine-tuning for Large Language Models with Limited Resources paper page: https://
huggingface.co/papers/2306.09
782
… Large Language Models (LLMs) have revolutionized Natural Language Processing (NLP) but demand massive GPU resources for training. Lowering the threshold for LLMs -

MagicBrush: Manually Annotated Dataset for Instruction-Guided Image Editing
By
–
MagicBrush: A Manually Annotated Dataset for Instruction-Guided Image Editing
— AK (@_akhaliq) 19 juin 2023
paper page: https://t.co/T6N8UmgEdz
Text-guided image editing is widely needed in daily life, ranging from personal use to professional applications such as Photoshop. However, existing methods are… pic.twitter.com/gu3PCBokeuMagicBrush: A Manually Annotated Dataset for Instruction-Guided Image Editing paper page: https://
huggingface.co/papers/2306.10
012
… Text-guided image editing is widely needed in daily life, ranging from personal use to professional applications such as Photoshop. However, existing methods are -

OCTScenes: Real-World Dataset for Object-Centric Learning
By
–
OCTScenes: A Versatile Real-World Dataset of Tabletop Scenes for Object-Centric Learning paper page: https://
huggingface.co/papers/2306.09
682
… Humans possess the cognitive ability to comprehend scenes in a compositional manner. To empower AI systems with similar abilities, object-centric -

CAJun: Continuous Adaptive Jumping Framework for Legged Robots
By
–
CAJun: Continuous Adaptive Jumping using a Learned Centroidal Controller
— AK (@_akhaliq) 19 juin 2023
paper page: https://t.co/9nMSBerv5w
present CAJun, a novel hierarchical learning and control framework that enables legged robots to jump continuously with adaptive jumping distances. CAJun consists of a… pic.twitter.com/5XMbsEFGhSCAJun: Continuous Adaptive Jumping using a Learned Centroidal Controller paper page: https://
huggingface.co/papers/2306.09
557
… present CAJun, a novel hierarchical learning and control framework that enables legged robots to jump continuously with adaptive jumping distances. CAJun consists of a -

Robot Learning with Sensorimotor Pre-training Using Transformer
By
–
Robot Learning with Sensorimotor Pre-training
— AK (@_akhaliq) 19 juin 2023
paper page: https://t.co/JNC4XuEu2f
present a self-supervised sensorimotor pre-training approach for robotics. Our model, called RPT, is a Transformer that operates on sequences of sensorimotor tokens. Given a sequence of camera… pic.twitter.com/DdW8qWTcDPRobot Learning with Sensorimotor Pre-training paper page: https://
huggingface.co/papers/2306.10
007
… present a self-supervised sensorimotor pre-training approach for robotics. Our model, called RPT, is a Transformer that operates on sequences of sensorimotor tokens. Given a sequence of camera -

AvatarBooth: High-Quality Customizable 3D Human Avatar Generation
By
–
AvatarBooth: High-Quality and Customizable 3D Human Avatar Generation
— AK (@_akhaliq) 19 juin 2023
paper page: https://t.co/5WIusjAJDA
introduce AvatarBooth, a novel method for generating high-quality 3D avatars using text prompts or specific images. Unlike previous approaches that can only synthesize… pic.twitter.com/zzfmvxsirbAvatarBooth: High-Quality and Customizable 3D Human Avatar Generation paper page: https://
huggingface.co/papers/2306.09
864
… introduce AvatarBooth, a novel method for generating high-quality 3D avatars using text prompts or specific images. Unlike previous approaches that can only synthesize -

CLIPSonic: Text-to-Audio Synthesis with Unlabeled Videos
By
–
CLIPSonic: Text-to-Audio Synthesis with Unlabeled Videos and Pretrained Language-Vision Models paper page: https://
huggingface.co/papers/2306.09
635
… Recent work has studied text-to-audio synthesis using large amounts of paired text-audio data. However, audio recordings with high-quality text -

Scaling Open-Vocabulary Object Detection with Vision-Language Models
By
–
Scaling Open-Vocabulary Object Detection paper page: https://
huggingface.co/papers/2306.09
683
… Open-vocabulary object detection has benefited greatly from pretrained vision-language models, but is still limited by the amount of available detection training data. While detection training data can