The new @apple Vision Pro has taken the internet by storm and opened new doors of endless possibilities. @jeevprabnivash #Apple #AppleVisionPro #INDIAai #MachineLearning #AugmentedReality
MULTIMODAL AI
-

Apple Vision Pro AI Framework and Privacy at WWDC 2023
By
–
What went behind the most anticipated event at Apple’s WWDC 2023 was indeed a spectacular advancement in AI and AR! We couldn’t help but dig deeper! Read more about Apple Vision Pro’s framework and privacy configurations here : https://
rb.gy/z7iqy -
Multimodal Technology Emerges as Top Prediction for Future
By
–
Hard to predict what’s next these days but multimodal is a top candidate for sure.
-

Blended-NeRF: Zero-Shot Object Generation and Blending in Neural Radiance Fields
By
–
Blended-NeRF: Zero-Shot Object Generation and Blending in Existing Neural Radiance Fields
— AK (@_akhaliq) 23 juin 2023
paper page: https://t.co/u9kLfEqYdR
Editing a local region or a specific object in a 3D scene represented by a NeRF is challenging, mainly due to the implicit nature of the scene… pic.twitter.com/4MlCPBJ9SVBlended-NeRF: Zero-Shot Object Generation and Blending in Existing Neural Radiance Fields paper page: https://
huggingface.co/papers/2306.12
760
… Editing a local region or a specific object in a 3D scene represented by a NeRF is challenging, mainly due to the implicit nature of the scene -

Continuous Layout Editing of Single Images with Diffusion Models
By
–
Continuous Layout Editing of Single Images with Diffusion Models
— AK (@_akhaliq) 23 juin 2023
paper page: https://t.co/Nld7ZzwFn5
Recent advancements in large-scale text-to-image diffusion models have enabled many applications in image editing. However, none of these methods have been able to edit the… pic.twitter.com/kofkDWJ3xbContinuous Layout Editing of Single Images with Diffusion Models paper page: https://
huggingface.co/papers/2306.13
078
… Recent advancements in large-scale text-to-image diffusion models have enabled many applications in image editing. However, none of these methods have been able to edit the -

Google Introduces AudioPaLM: Speech-Capable Large Language Model
By
–
Google presents AudioPaLM: A Large Language Model That Can Speak and Listen
— AK (@_akhaliq) 23 juin 2023
paper page: https://t.co/uLZwULDc94
introduce AudioPaLM, a large language model for speech understanding and generation. AudioPaLM fuses text-based and speech-based language models, PaLM-2 [Anil et al.,… pic.twitter.com/85fyv2B25RGoogle presents AudioPaLM: A Large Language Model That Can Speak and Listen paper page: https://
huggingface.co/papers/2306.12
925
… introduce AudioPaLM, a large language model for speech understanding and generation. AudioPaLM fuses text-based and speech-based language models, PaLM-2 [Anil et al., -
V5.2 Update: Enhanced Aesthetics, Text Understanding, and New Features
By
–
We're testing V5.2 today, it includes improved aesthetics, coherence, text understanding, sharper images, higher variation modes, zoom-out outpainting, and a new /shorten command for analyzing your prompt tokens. Enjoy!
-

HyperReel: High-Fidelity 6-DoF Video with Ray-Conditioned Sampling
By
–
HyperReel: High-Fidelity 6-DoF Video with Ray-Conditioned Sampling
— AK (@_akhaliq) 22 juin 2023
paper page: https://t.co/TkYI5ibRny
Volumetric scene representations enable photorealistic view synthesis for static scenes and form the basis of several existing 6-DoF video techniques. However, the volume… pic.twitter.com/7zIO8MXavDHyperReel: High-Fidelity 6-DoF with Ray-Conditioned Sampling paper page: https://
huggingface.co/papers/2301.02
238
… Volumetric scene representations enable photorealistic view synthesis for static scenes and form the basis of several existing 6-DoF video techniques. However, the volume -
SoundStorm: Parallel Decoding for Efficient Audio Generation
By
–
Many generative audio models rely on auto-regressive decoding, which produces tokens one by one and can be slow. Read all about SoundStorm, a new method tailored for audio tokens that uses parallel decoding for efficient and high-quality audio generation ↓
-

Google Imagen Editor: Text-Guided Image Editing Tool
By
–
Stop by the #CVPR2023 Google booth at 12:30pm today to speak with @wangsu_googleai about Imagen Editor, a novel tool for text-guided image editing, and to learn about EditBench, a benchmark for quality assessment of image-text alignment.