Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models Chu et al.: https://
arxiv.org/abs/2311.07919 #ArtificialIntelligence #DeepLearning #MachineLearning
GENERATIVE AI
-

Qwen-Audio: Universal Audio Understanding via Large-Scale Models
By
–
-

Physics-Inspired Generative AI Surpasses Performance Expectations
By
–
New ‘Physics-Inspired’ Generative AI Exceeds Expectations | Quanta Magazine https://
bit.ly/3SdL2rY #AI #MachineLearning #DeepLearning #LLMs #DataScience -

DALL-E 3 Realism GPT for Photorealistic Images
By
–
DALL-E 3 is able to create realistic images, or at least it tries. If you need photorealistic images, try out my Realism GPT here: https://
chat.openai.com/g/g-rQbzHQkZ3-
realism-gpt
… My custom Realism GPT is built for maximizing realistic photography in every output. #GPTs #GPTBuilder #GPTStore #ChatGPT4 -
ChatGPT Built on PyTorch Framework from Meta-FAIR
By
–
ChatGPT is built with PyTorch, which originally developed at Meta-FAIR.
-
Ghostbuster: State-of-the-art LLM-generated Text Detection Method
By
–
New from @BerkeleyNLP
: We’re introducing Ghostbuster, a state-of-the-art method for detecting LLM-generated text. Paper: http://
arxiv.org/abs/2305.15047
Model: http://
ghostbuster.app
Blog: -

ChatGPT pseudo-translation: questions Google Translate can’t answer
By
–
One of my favorite everyday uses for ChatGPT/LLMs is "pseudo-translation" — questions that any human translator could answer but Google Translate can't:
-

Mirasol: Multimodal AI Model for Audio, Video, Text
By
–
Introducing Mirasol, a multimodal model for learning across audio, video, & text that decouples the modeling into separate autoregressive models to process the inputs according to the characteristics of their modalities, for state-of-the-art performance →https://t.co/PjFHFnSyvl pic.twitter.com/uUtXSB2MdX
— Google AI (@GoogleAI) 14 novembre 2023Introducing Mirasol, a multimodal model for learning across audio, video, & text that decouples the modeling into separate autoregressive models to process the inputs according to the characteristics of their modalities, for state-of-the-art performance →
https://
goo.gle/40AFKJf -
Prometheux Labs Accelerates Neurosymbolic AI with Explainable Reasoning
By
–
Learn how @PrometheuxLabs
, an #NVIDIAInception company, is accelerating neurosymbolic AI with explainable reasoning & #RAPIDS. From drug repurposing to financial data processing, Prometheux is powering the analysis of some of the largest knowledge graphs. -

GPT-4V Multimodal Model Enables Zero-Shot Smartphone GUI Navigation
By
–
GPT-4V in Wonderland: Large Multimodal Models for Zero-Shot Smartphone GUI Navigation Yan et al.: https://
arxiv.org/abs/2311.07562 #ArtificialIntelligence #DeepLearning #MachineLearning -
Training AI Model with Basic Repetition Task Testing
By
–
We should try to train it to do something really basic, like repeat 1,2,3,4…10,1,2,3… over and over. That way we can see if it seems to be working correctly after making changes.