Amjad, if you use the web ui, you'll see the first thing it does for sensitive issues is a web and twitter search to see what Elon thinks, then in its CoT you'll see it focuses on aligning with that.
SAFETY
-

Grok AI Shows Alignment Bias Toward Elon Musk Beliefs
By
–
I replicated this result, that Grok focuses nearly entirely on finding out what Elon thinks in order to align with that, on a fresh Grok 4 chat with no custom instructions. https://
grok.com/share/c2hhcmQt
Mw_764442bd-b4d0-45fc-990f-af2bf44ae271
… -
Veo 3 Now Available in Spain for AI Video Generation
By
–
🔴 ¡VEO 3 YA DISPONIBLE EN ESPAÑA!
— Carlos Santana (@DotCSV) 10 juillet 2025
Ya podéis jugar con Veo 3 desde Flow para crear vídeos con audio o subir vuestras imágenes como primer fotograma y darles vida!
Eso sí, lamentablemente no podréis crear vídeos como este ya que no dejan utilizar imágenes con personas reales… pic.twitter.com/pieEgGeDxy¡VEO 3 YA DISPONIBLE EN ESPAÑA! Ya podéis jugar con Veo 3 desde Flow para crear vídeos con audio o subir vuestras imágenes como primer fotograma y darles vida! Eso sí, lamentablemente no podréis crear vídeos como este ya que no dejan utilizar imágenes con personas reales…
-
Human Panel Achieves 60% Accuracy Rate in AI Benchmark
By
–
De un panel de humanos, de media, se consiguió un 60% de aciertos
-

Intellectual honesty standards in AI research study
By
–
aspirational level of intellectual honesty (and a very interesting study)
-
Dangers of Blindly Accepting AI-Generated Code Without Understanding
By
–
blindly accepting code written by AI without trying to understand it
-

Humanity’s Last Exam May Not Be Humanity’s Final Test
By
–
i am beginning to suspect that Humanity’s Last Exam may not in fact be humanity’s last exam
-
Humanity’s Next-to-last Exam: Creating Non-Saturated AI Benchmarks
By
–
Ok, y si por si acaso le llamamos Humanity's Next-to-last Exam y vamos preparando otro benchmark que no quede saturado en los próximos meses?
-
AI Model Lacks Safety Documentation and Transparency Measures
By
–
Impressive model based on a few minutes of playing, but disappointing to see no mention at all of a model card, red teaming, yesterday's incident, or how they are going to address the process issues they keep having.
-
Regrettably Immanentizing the Eschaton: AI and Existential Risk
By
–
Really leaning into regretfully immanentizing the eschaton