google using adobe firefly to generate images is really interesting… why not use its own technology tho? remember imagen?
MULTIMODAL AI
-
Curiosity about Bard’s image captioning capabilities and potential risks
By
–
ok honestly i'm v curious to see what kinds of captions bard comes up with for images… this could be useful but also it could go sideways in so many ways.
-
Meta AI’s Angela Fan Presents Text Generation Research in London
By
–
Welcoming @MetaAI 's Angela Fan at The Research and Applied AI Summit on June 23 in London! Angela focusses on research in text generation, including her recent work on No Language Left Behind and Universal Speech Translation for Unwritten Languages.
-

Extensible AI Tool with Multimodal Built-in Capabilities
By
–
It comes with built-in tools: • Document QA
• Speech-to-text and Text-to-speech
• Text {classification, summarization, translation, download, QA}
• Image {generation, transforms, captioning, segmentation, upscaling, QA}
• Text to video It is EXTENSIBLE by design. -

Transformers Agents: Democratizing ML with Multimodal Control
By
–
We just released Transformers' boldest feature: Transformers Agents. This removes the barrier of entry to machine learning Control 100,000+ HF models by talking to Transformers and Diffusers Fully multimodal agent: text, images, video, audio, docs… https://
huggingface.co/docs/transform
ers/transformers_agents
… -

Building AI Agents with LLMs for Multimodal Tasks
By
–
Create an agent using LLMs (OpenAssistant, StarCoder, OpenAI …) and start talking to transformers and diffusers It responds to complex queries and offers a chat mode. Create images using your words, have the agent read the summary of websites out loud, read through a PDF
-

Turn Action Figures into AI Friends with Raspberry Pi Pico
By
–
Raspberry Pi Pico (can access Wi-Fi) + GPT API w/ personality + voice input via Whisper and output via Coqui (and training data) + action figure = make your action figures your friends with a small ‘stick on’.
-

Hybrid Image Models: Combining Tabular Data with Vision Intelligence
By
–
Learn about Hybrid Image Models at https://
abacus.ai/vision_hybrid by @abacusai What is a Hybrid Model? Combines your tabular data with images to extract powerful insights. Improves model performance by overlaying structured data with unstructured vision data. Applies -

Abacus AI Deep Learning Computer Vision Solutions
By
–
@abacusai uses state-of-art #DeepLearning models to solve #ComputerVision problems: • Classification & Detection
• Hybrid Image Models
• Image Segmentation You bring your images, and they take care of the rest. http://
abacus.ai/vision
——
#AI #MachineLearning #DataScience -

JAX Diffusers Sprint Winners Showcase ControlNet Applications
By
–
Announcing the winners of JAX Diffusers sprint First place: https://
huggingface.co/spaces/control
net-interior-design/controlnet-seg
… Second place: https://
huggingface.co/spaces/ioclab/
brightness-controlnet
… Third Place: https://
huggingface.co/spaces/vllab/c
ontrolnet-hands
…