Real-time voice conversations with ChatGPT: This Chrome extension allows you to have a back-and-forth voice dialogue with ChatGPT. Soon we could see this with photorealistic avatars, and custom voice cloning using @elevenlabs
MULTIMODAL AI
-
Vision Transformers: Next Video on Visual Recognition Systems
By
–
The next video will be about Vision Transformers. (Vision) Transformers are an integral part of current large-scale visual recognition systems and have made it easy to build multimodal systems, language and vision in particular. Working on it slowly, but surely.
-
Solve Rubik’s Cube with Augmented Reality and AI
By
–
Learn how to solve your Rubik's Cube using Augmented Reality and AI#augmentedreality #artificialintelligence pic.twitter.com/rvqQ93DMVw
— Pascal Bornet (@pascal_bornet) 18 février 2023Learn how to solve your Rubik's Cube using Augmented Reality and AI #augmentedreality #artificialintelligence
-

ControlNet OpenPose Tracks Multiple Subjects in YMCA Video
By
–
https://
reddit.com/r/StableDiffus
ion/comments/112h3p9/ymca_controlnet_openpose_can_track_at_least_four/
… -

ControlNet Game Changer for Stable Diffusion Image Generation
By
–
https://
reddit.com/r/StableDiffus
ion/comments/113bgke/controlnet_is_such_a_game_changer_for_stable/
… -
Controlling Image Structure Generation: Latest Results
By
–
the way how this work enables us to control the structure from the images we generate is truly interesting. here are some of our favorite results so far:
-
Convolutional Architecture Encodes Condition Image in Cloned U-Net
By
–
it is worth mentioning that the authors used a convolutional architecture to encode the condition image before feeding it within the cloned U-Net.
-

ControlNet Architecture with Stable Diffusion Explained
By
–
the following is a representation of ControlNet's architecture when used with Stable Diffusion
-

ControlNet: Guide complet du fonctionnement et des applications
By
–
what is ControlNet and how does it work?
-

ControlNet: Conditioning Diffusion Models on Arbitrary Input Features
By
–
ControlNet is a method that can be used to condition diffusion models on arbitrary input features, such as image edges, segmentation maps, or human poses.