However, when we further analyze model generations in this condition, we find that the model may rely on over-generalizations and country-specific stereotypes.
SAFETY
-
Evaluating Language Model Values: Frameworks for Global AI Alignment
By
–
Our preliminary findings show the need for rigorous evaluation frameworks to uncover whose values language models represent. We encourage using this methodology to assess interventions to align models with global, diverse perspectives. Paper:
-
Language Models Show Western-Centric Opinions and Steerability
By
–
We develop a method to test global opinions represented in language models. We find the opinions represented by the models are most similar to those of the participants in USA, Canada, and some European countries. We also show the responses are steerable in separate experiments. pic.twitter.com/QzHmRPNqSl
— Anthropic (@AnthropicAI) 29 juin 2023We develop a method to test global opinions represented in language models. We find the opinions represented by the models are most similar to those of the participants in USA, Canada, and some European countries. We also show the responses are steerable in separate experiments.
-
OpenAI Training Data Transparency: Copyright, Bias, and Generalization Issues
By
–
Agreeing and extending @sashamtl
, there are lots of reasons why OpenAI likely isn’t open about training data: copyright, bias, and also questions about generalization & data contamination. I too salute @abebab
’s great detective work, including an important new paper coming soon. -
Nature Editorial: Constructive Discussion on AI Existential Risks
By
–
This week's @Nature editorial "Fearmongering narratives about existential risks are not constructive. Serious discussion about actual risks and action to contain them, are." https://
nature.com/articles/d4158
6-023-02094-7
… -
AI Safety Needs More Resources Amid Rapid Development
By
–
“Right now there are 99 very smart people trying to make #AI better and one very smart person trying to figure out how to stop it taking over and maybe you want to be more balanced.”—
@geoffreyhinton https://
news.com.au/technology/inn
ovation/inventions/smarter-than-us-ai-godfathers-grim-warning-for-the-future/news-story/58684beaaa114b09d2a430dd08556818
… -

AI Industry Leaders Issue Extinction Risk Warning Statement
By
–
#AI industry and researchers sign statement warning of ‘extinction’ risk https://
cnn.it/3C0kqBB #ethics #FutureofWork -

Bias in AI: Exploring Myths of Neutral Artificial Intelligence
By
–
Join me TODAY for an insightful livestream on "When Bytes Bias: Unraveling the Myth of Neutral AI"! 12 Noon UK | ET 7AM | PT 4AM | CET 1PM | SGT 7PM Twitter > https://
lnkd.in/ehrXiUW #AIethics #BiasInTech #NeutralAI #Livestream #MeredithBroussard #AI #JoinTheDiscussion -
Digital Data Guardrails: First Step in AI Regulation
By
–
Digital data guardrails are the first step in regulating AI
#AI #AIio #BigData #ML #NLU #Futureofwork http://
ow.ly/wlm830svSeE -

Digital Data Guardrails: First Step in AI Regulation
By
–
Digital data guardrails are the first step in regulating AI
#AI #AIio #BigData #ML #NLU #Futureofwork @ahier @guzmand
@DavidBrin @denisegarth @dez_blanchfield @diioannid @DioFavatas @gerald_bader http://
ow.ly/hsBU30svSt1