What we're asking is not whether an AI can solve this task, but whether there will be, by mid-2025, some technical advance whereby, in general, if the AI doesn't know how to do a thing, it will tell you so rather than making up a wrong answer.
AI
-
Skepticism about AI-generated advertisements using Stable Diffusion
By
–
Hey for all I know you just generated that ad with stable diffusion!
-

Generative AI Fabricates False McDonald’s Sesame Seeds History
By
–
Great example of how generative AI can make things up in a way that is easy to miss (and then pass on as fact). McDonald’s buns have, in fact, had sesame seeds since way before 1991. See this great @stephcliff story for proof. https://
nytimes.com/2008/07/17/bus
iness/media/17adco.html
… -
AGI Agent Must Excel at Math and String Processing
By
–
call me old fashioned, but IMHO if it is going to be pitched as a general purpose near AGI agent it oughta be pretty darn good at math and strings.
-

Pretraining Dominates Knowledge Acquisition in Large Language Models
By
–
Quoting:
"These results strongly suggest that almost all knowledge in large language models is learned during pretraining, and only limited instruction tuning data is necessary to teach models to produce high quality output." https://
bit.ly/3J7xWae -
LLM Hallucinations: When 95% Accuracy Is Not Enough
By
–
“largely eliminated” to me means “not really a problem anymore”; as noted, 5% error is tolerable in some domains, intolerable others. to take another example JPMPC can’t replace online banking powered by a database with chatbot that is 95% correct and 5% hallucinatory
-

IIoT ROI Depends on Data Strategy and Implementation
By
–
Achieving a return on investment (ROI) and deriving value from #IIoT depends on one critical factor: data. Download the white paper to find out more: http://
ow.ly/cz0G50OwTzT #sponsored #hivemq_iiot #digitaltransformation #data #mqtt @Datasciencectrl @jblefevre60 via @fogoros -
AI Progress Claims Not Yet Reliably Solved Despite Advances
By
–
note also that these problems are still not reliably solved despite the allegedly enormous progress since.
-
Physicians operating under greater uncertainty with AI
By
–
physicians are of course eg operating under greater uncertainty
-
Do AI Systems Hallucinate? Evaluating Truthfulness and Reliability
By
–
Let’s say you have a system that presents as if it can write biographies, summarize articles, interact w users, generate reports, etc. You can then ask two questions – Does the system invent stuff that isn’t true in the course of carrying out those functions? – Is the system