7/ Fine-Grained RLHF – trains LMs with fine-grained human feedback; instead of using overall preference, more explicit feedback is provided at the segment level which helps to improve efficacy on long-form question answering and reduces toxicity.
SAFETY
-
Detecting AI-Generated Content and Deepfakes for Digital Security
By
–
In the era of #AI-produced #content, from written pieces to #deepfake videos, understanding how to detect such creations is key to maintaining truth, integrity, and #security in our increasingly #digitized society. https://
linkedin.com/pulse/how-can-
you-detect-content-created-chatgpt-other-ais-bernard-marr
… -
LLM Hallucinations: Algorithm Problem Not Data Quality Issue
By
–
Even with perfect data LLMs would hallucinate. when an LLM tells you Elon Musk died a car crash, it’s not a problem with the data, it’s a problem w the algorithm, and its inherent lack of proper database records and proper methods for validating against formal knowledge stores.
-
Sewer Robots Identify and Destroy Dengue Mosquito Breeding Sites
By
–
#Sewer robots identify and destroy #dengue dens! https://
gavi.org/vaccineswork/s
ewer-robots-identify-and-destroy-dengue-dens
… #mosquitos #Robotics #robot #EthicalAI #OpenAI #bots #KillerRobot #roboticsainews #opensource #Python #tech #technology #cobot #humanoid #AI #ML #Algorithms #AIEthics #Robotech #RobotDesign #Engineering -
ChatGPT Exploited by Traditional Malware Attackers
By
–
Traditional malware increasingly takes advantage of #ChatGPT for attacks | CSO Online https://
csoonline.com/article/369851
8/traditional-malware-increasingly-takes-advantage-of-chatgpt-for-attacks.html
… via @csoonline #Malware #AI #GPT4 #RCE #ZeroTrust #ZeroDay #cybercrime #hacker #privacy #APT #bot #CISO #DDoS #hacking #phishing #CyberAttack #cybersecurity #Security -
Reddit thread about not expecting too much from AI
By
–
Reddit thread: https://
reddit.com/r/ProgrammerHu
mor/comments/145z7g6/do_not_pretend_too_much_from_ai/
… -
AI honesty: Will systems admit uncertainty by mid-2025?
By
–
What we're asking is not whether an AI can solve this task, but whether there will be, by mid-2025, some technical advance whereby, in general, if the AI doesn't know how to do a thing, it will tell you so rather than making up a wrong answer.
-

Generative AI Fabricates False McDonald’s Sesame Seeds History
By
–
Great example of how generative AI can make things up in a way that is easy to miss (and then pass on as fact). McDonald’s buns have, in fact, had sesame seeds since way before 1991. See this great @stephcliff story for proof. https://
nytimes.com/2008/07/17/bus
iness/media/17adco.html
… -
LLM Hallucinations: When 95% Accuracy Is Not Enough
By
–
“largely eliminated” to me means “not really a problem anymore”; as noted, 5% error is tolerable in some domains, intolerable others. to take another example JPMPC can’t replace online banking powered by a database with chatbot that is 95% correct and 5% hallucinatory
-
AI Progress Claims Not Yet Reliably Solved Despite Advances
By
–
note also that these problems are still not reliably solved despite the allegedly enormous progress since.
