I'm using PhotoMaker to attempt synthetic photos of historical figures based on sculpture. Below is my artist's impression of Julius Caesar. I used 4 pics of a sculpture: https://
replicate.com/p/pfgr2jdbszgv
nfzfkiakoiodmy
… Based on this: https://
metmuseum.org/art/collection
/search/192717
… Upscaled with Magnific.
MULTIMODAL AI
-
PhotoMaker Creates Synthetic Portrait of Julius Caesar from Sculpture
By
–
-

GPT Vision Converses with ChatGPT: AI Self-Interaction
By
–
Watch GPT Vision with control over my OS talking to ChatGPT.
— Pietro Schirano (@skirano) 18 janvier 2024
The most fascinating part is that it's intrigued by having a conversation with another "similar."
"What do you think about human/AI interaction?" it asked.
Also, the superhuman speed at which it types, lol pic.twitter.com/ViffvDK5H9Watch GPT Vision with control over my OS talking to ChatGPT. The most fascinating part is that it's intrigued by having a conversation with another "similar." "What do you think about human/AI interaction?" it asked. Also, the superhuman speed at which it types, lol
-
Runway AI video editor to add advanced area tagging features
By
–
Runway AI video editor may get new features soon allowing you to tag specific areas on the image and have "more control" 👀
— 🚨 AI News | TestingCatalog (@testingcatalog) 17 janvier 2024
This feature is not available yet. https://t.co/OfIXY5cC5qRunway AI video editor may get new features soon allowing you to tag specific areas on the image and have "more control" This feature is not available yet.
-

Multimodal Deep Learning Models for Asynchronous Data Streams
By
–
A deep learning paper from our team on how to build multimodal models, where data for individual modalities arrives at different speeds, by combining powerful pretrained text, image and video models. Useful for robotics. Congrats @konradzolna @serkancabi @yutianc Eric, Claudio,… https://t.co/vPmZpDkve3
— Nando de Freitas (@NandoDF) 17 janvier 2024A deep learning paper from our team on how to build multimodal models, where data for individual modalities arrives at different speeds, by combining powerful pretrained text, image and video models. Useful for robotics. Congrats @konradzolna @serkancabi @yutianc Eric, Claudio,
-
Samsung Circle To Search: New AI Feature Raises Privacy Concerns
By
–
At #SamsungUnpacked, Samsung and Google just announced a "Circle To Search" tool that lets users search anything on their app by circling, highlighting or tapping on anything without switching apps. A big question I have: What do all these new AI features mean for user privacy?
-
Samsung Galaxy S24 Debuts Real-Time Call Translation AI
By
–
Happening now: Samsung’s #SamsungUnpacked event for the debut of the Galaxy S24. Expect lots of new AI features to be announced. One new feature: The S24 will have real-time language translations for calls & IRL convos with two-way text & voice translations.
-

GPT-4 Controls OS to Generate Images via Open Interpreter
By
–
This is wild. I gave OS control to GPT-4 via the latest update of Open Interpreter and now it's generating pictures it wants to see in @everartai 🤯
— Pietro Schirano (@skirano) 17 janvier 2024
GPT is controlling the mouse and adding text in the fields, I am not doing anything. pic.twitter.com/hGgML9epEcThis is wild. I gave OS control to GPT-4 via the latest update of Open Interpreter and now it's generating pictures it wants to see in @everartai GPT is controlling the mouse and adding text in the fields, I am not doing anything.
-

AlphaGeometry Neuro-Symbolic Architecture: System 1 and System 2
By
–
This is the neuro-symbolic architecture of #AlphaGeometry. Similar to System1 and System 2, in the book "Thinking, fast and slow", the symbolic engine will first take a crack at the problem mechanically; if it gets stuck it will ask the neural language model for suggestions of
