Expect OpenAI to reveal GPT-4V (image recognition) API and their own autonomous agents on November 6th.
MULTIMODAL AI
-
OpenAI November Developer Conference: Memory, Vision APIs Launch
By
–
OpenAI is rumoured to push major developer-friendly updates for cost-effective application building at their november 6th developer conference. Adding memory storage and vision capabilities, targeting ambitious revenue goals, launching stateful and vision APIs, focusing on
-
AI Visualizes Human Evolution and Future Predictions
By
–
La ia puede crear secuencias increíbles y en este vídeo resume se manera muy visual la evolución de la especie humana.
— Juan Merodio (@juanmerodio) 13 octobre 2023
La ia no se queda en el ser humano actual y realiza una previsión de cómo cree que será la evolución en los próximos años. Crees que en el futuro nos… pic.twitter.com/MpxZwCPwrcLa ia puede crear secuencias increíbles y en este vídeo resume se manera muy visual la evolución de la especie humana. La ia no se queda en el ser humano actual y realiza una previsión de cómo cree que será la evolución en los próximos años. Crees que en el futuro nos
-
Generative AI and Immersive Internet: Future Digital Worlds
By
–
Dive into the #future of the 'immersive internet' in 2024, exploring groundbreaking advancements from #generative #AI in #digitalworlds to the convergence of #AR and #VR.
-

GPT-4V for Mathematical Equation Understanding and Explanation
By
–
If you're like me and find it easier to read math than code, and you have access to @OpenAI GPT 4V, try pasting a image of an equation you wanna understand in there. It might just blow your mind.
-

GPT-4V identifies watch model but cannot read time
By
–
Interesting GPT-4V can correctly identify the make and model of a wristwatch but can’t read the time it’s displaying.
-
Llava Image Chat Support Added to Llama2.ai
By
–
https://t.co/3ZZemXeQjs pic.twitter.com/bAYkUniXwu
— Replicate (@replicate) 12 octobre 2023I just added Llava support to http://
llama2.ai — you can now chat with your images! It's fully open-source and pretty amazing. -
GROOT Builds on MineDojo for Open-Ended Minecraft Learning
By
–
Great to see further work building on MineDojo. Minecraft is the right benchmark to test for open-ended learning. GROOT uses reference videos for instruction following and composes them together. https://t.co/cDgAzp3O0X
— Prof. Anima Anandkumar (@AnimaAnandkumar) 12 octobre 2023Great to see further work building on MineDojo. Minecraft is the right benchmark to test for open-ended learning. GROOT uses reference videos for instruction following and composes them together.
-

GPT-4 Vision AMA: Multimodal AI Capabilities Discussion
By
–
Hell YEAH! I'm in! AmA. #GPT4V #GPT4-Vision
-

SDXL Adds Resend Button for Multiple Image Generations
By
–
New today: we added a resend button on web and iOS for SDXL and SDXL-based bots! Clicking resend will preserve the previous image and add a second generation from the model, so you don't have to copy and paste a prompt multiple times.