Aligning Large Language Models with Human: A Survey by Yufei Wang, Wanjun Zhong, Liangyou Li, Fei Mi, Xingshan Zeng, Wenyong Huang, Lifeng Shang, Xin Jiang, Qun Liu
SAFETY
-
Open Problems and Fundamental Limitations of Reinforcement Learning from Human Feedback
By
–
Open Problems and Fundamental Limitations of Reinforcement Learning from Human Feedback, by @StephenLCasper @alxndrdavies @causalclaudia @sociotiose @jeremy_scheurer @javirandor @FreedmanRach @tomekkorbak @davlindner @5kovt @CRSegerie @MicahCarroll
, et al. -

Digital Data Guardrails: First Step in AI Regulation
By
–
Digital data guardrails are the first step in regulating AI
#AI #AIio #BigData #ML #NLU #Futureofwork @TopCyberNews @SpirosMargaris @MarshaCollier @MHiesboeck
@MHcommunicate @Fisher85M @MikeQuindazzi @NealSchaffer
http://
ow.ly/Y1Ge30sx0Te -

AI and Risk Management Discussion with EFIPM
By
–
Such a pleasure chatting up with @efipm on #AI and #RiskManagement. https://
youtube.com/watch?v=ZNvc2J
4bPwM
… @NACD #BoardDirector #CEOs -
Google Contractor Pay Disputes Impact Bard Quality Control
By
–
this sounds great- a Google subcontractor at Appen who works on Bard is not given enough time to check facts in the chatbot's answers, while Appen, which is in the middle of a turnaround, isn't paying the $15 an hour that Google requires for contractors https://
cnbc.com/2023/09/06/app
en-which-helps-amazon-and-google-train-ai-is-reeling.html
… -
Chinese AI Model Spreads COVID-19 Misinformation About US Origins
By
–
Chinese ChatGPT competitor "reckons that covid-19 originated among American vape-users in July 2019; later that year the virus was spread to the Chinese city of Wuhan, via American lobsters."
-
AI-Powered Phishing Threats Set to Escalate Dramatically
By
–
Eyal Benishti, Ironscale CEO: “Just imagine business email compromise and targeted phishing at the same volume as we experience spam, because that’s what will happen.” #AI #cybersecurity #generativeai #phishing
-
FACET: New Fairness Benchmark Dataset for Vision Models
By
–
Last week we released FACET, a new comprehensive benchmark dataset for evaluating the fairness of models across a number of different vision tasks, constructed of 32K images from SA-1B, labeled by expert annotators.
— AI at Meta (@AIatMeta) 5 septembre 2023
Read the paper ➡️ https://t.co/OoYV2eiSYX pic.twitter.com/9Q7uk59TvTLast week we released FACET, a new comprehensive benchmark dataset for evaluating the fairness of models across a number of different vision tasks, constructed of 32K images from SA-1B, labeled by expert annotators. Read the paper https://
bit.ly/3EotZvg -
AI Conversational Abilities Don’t Guarantee Safe Decision-Making
By
–
AI has demonstrated unprecedented capabilities such as engaging in natural conversations, but this doesn't mean we should trust these models to handle sensitive decision-making tasks. "It’s just not there yet,” says HAI faculty affiliate Sanmi Koyejo.
-
AI Autonomy and Future Warfare: Exploring Emerging Conflict Dynamics
By
–
I repeatedly encountered this issue while reporting a recent story that explores how autonomy and AI will change future conflicts: