Last week @Trevornoah asked @OpenAI @miramurati
: How can we safeguard against AI-powered photo editing for misinformation? MIT students hacked a way to "immunize" photos against edits: http://
gradientscience.org/photoguard/ @aleks_madry
SAFETY
-

MIT immunizes photos against AI-powered misinformation edits
By
–
-
Ethics of accusation: avoiding adversarial framing risks
By
–
This one is at least feasible. But I'd still personally refrain from accusing anyone, since it's easy to frame someone adversarially or even accidentally.
-
Security Vulnerability Identified as Potential Attack Vector
By
–
Yeah, it does look pretty bad. But also seems like a nice attack vector 🙂
-
LLMs Generate Convincing but Wrong Completions Often
By
–
this is an issue I also have with copilot. it often very convincingly pushes for long and beautiful but wrong completions LLM as “internet influencers”: surface form trumps content https://
x.com/shortstein/sta
/shortstein/status/1587857803678748672
… -
Blue Tick Verification: Balancing Authenticity Against Toxicity Online
By
–
Yes, the intention was to distinguish authentic voices. The blue tick process surely needs improvement. To be fair to Musk, he’s been open about using this for revenue. But while the intention is right, the approach is debatable, as this could further legitimise noise & toxicity.
-
Blindsight: The AI Future We Should Avoid
By
–
(Of all of the science fiction books we could be living in, Blindsight would not be my novel of choice.) Paper here:
-
Adversarial Policy Defeats KataGo Go AI System 99%
By
–
Befuddling AI Go Systems: MIT, UC Berkeley & FAR AI’s Adversarial Policy Achieves a >99% Win Rate Against KataGo https://
syncedreview.com/2022/11/02/bef
uddling-ai-go-systems-mit-uc-berkeley-far-ais-adversarial-policy-achieves-a-99-win-rate-against-katago/
… -

Radio’s Role in Nazi Rise: Historical Media Amplification Study
By
–
Analysis of radio broadcasts in pre-WW2 Germany finds pro-democracy broadcasts initially slowed the Nazi rise. But after the Nazis gained access to the airwaves, radio boosted Nazi membership & then it accelerated antisemitism (in areas of Germany that were already antisemitic).
-
Google Leaders Discuss Responsible AI Innovation Strategy
By
–
Starting now: James Manyika, SVP of Technology and Society, and Marian Croak, VP of Engineering @Google will discuss the importance of innovating responsibly and what factors we keep top of mind in our research. Tune in ↓
-
Adversarial Robustness Learning Theory Paper Review Process Critique
By
–
"Adversarially Robust Learning with Tolerance" by @ashtiani_hassan
, @OneBigOh
, Urner. 2 +ve reviews, 1 -ve, saying "I implemented the method and it didn't work." But this is a theory paper… Another -ve review 4 days pre-decision, no chance to respond