Best coke ad I have seen, text to video AI , pikalabs by @thedorbrothers
— AK (@_akhaliq) 11 juillet 2023
pic.twitter.com/ApDqP97mXC
Best coke ad I have seen, text to video AI , pikalabs by @thedorbrothers

By
–
Best coke ad I have seen, text to video AI , pikalabs by @thedorbrothers
— AK (@_akhaliq) 11 juillet 2023
pic.twitter.com/ApDqP97mXC
Best coke ad I have seen, text to video AI , pikalabs by @thedorbrothers
By
–
As AI gets more powerful, the bottleneck in human-AI collaboration can become the input channel between the human and the AI, rather than the capabilities of the AI itself. This bottleneck can get especially severe with longer context windows.

By
–
Elon Musk and Mark Zuckerberg have a Danceoff, text to video AI, pikalabs by u/311scottie pic.twitter.com/JIVKReIXUh
— AK (@_akhaliq) 11 juillet 2023
Elon Musk and Mark Zuckerberg have a Danceoff, text to video AI, pikalabs by u/311scottie

By
–
Vid2Player3D
https://bit.ly/3Orjmhd #AI #MachineLearning #DeepLearning #LLMs #DataScience

By
–
AnyTeleop: A General Vision-Based Dexterous Robot Arm-Hand Teleoperation System
— AK (@_akhaliq) 11 juillet 2023
paper page: https://t.co/sU1bZyoQ3p
Vision-based teleoperation offers the possibility to endow robots with human-level intelligence to physically interact with the environment, while only requiring… pic.twitter.com/h8tNLStaKP
AnyTeleop: A General Vision-Based Dexterous Robot Arm-Hand Teleoperation System paper page: https://
huggingface.co/papers/2307.04
577
… Vision-based teleoperation offers the possibility to endow robots with human-level intelligence to physically interact with the environment, while only requiring

By
–
Shelving, Stacking, Hanging: Relational Pose Diffusion for Multi-modal Rearrangement
— AK (@_akhaliq) 11 juillet 2023
paper page: https://t.co/iQq0qsd7Xb
propose a system for rearranging objects in a scene to achieve a desired object-scene placing relationship, such as a book inserted in an open slot of a… pic.twitter.com/Xq6P2M8BNH
Shelving, Stacking, Hanging: Relational Pose Diffusion for Multi-modal Rearrangement paper page: https://
huggingface.co/papers/2307.04
751
… propose a system for rearranging objects in a scene to achieve a desired object-scene placing relationship, such as a book inserted in an open slot of a

By
–
AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning
— AK (@_akhaliq) 11 juillet 2023
paper page: https://t.co/0GkcW9l1cP
With the advance of text-to-image models (e.g., Stable Diffusion) and corresponding personalization techniques such as DreamBooth and LoRA, everyone… pic.twitter.com/4nfCKyUXO3
AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning paper page: https://
huggingface.co/papers/2307.04
725
… With the advance of text-to-image models (e.g., Stable Diffusion) and corresponding personalization techniques such as DreamBooth and LoRA, everyone

By
–
Semantic-SAM: Segment and Recognize Anything at Any Granularity paper page: https://
huggingface.co/papers/2307.04
767
… introduce Semantic-SAM, a universal image segmentation model to enable segment and recognize anything at any desired granularity. Our model offers two key advantages:

By
–
VampNet: Music Generation via Masked Acoustic Token Modeling
— AK (@_akhaliq) 11 juillet 2023
paper page: https://t.co/FScor4pyFy
introduce VampNet, a masked acoustic token modeling approach to music synthesis, compression, inpainting, and variation. We use a variable masking schedule during training which… pic.twitter.com/pjxZlq5yWj
VampNet: Music Generation via Masked Acoustic Token Modeling paper page: https://
huggingface.co/papers/2307.04
686
… introduce VampNet, a masked acoustic token modeling approach to music synthesis, compression, inpainting, and variation. We use a variable masking schedule during training which

By
–
Sketch-A-Shape: Zero-Shot Sketch-to-3D Shape Generation paper page: https://
huggingface.co/papers/2307.03
869
… Significant progress has recently been made in creative applications of large pre-trained models for downstream tasks in 3D vision, such as text-to-shape generation. This motivates our