1n73ll1g3nc3 15 7h3 4b1l17y 70 4d4p7 70 ch4ng3
AGI
-

Decoding AI News: GPT-4.5 Rumors, ChatGPT Plus Return, Optimus, AI Act
By
–
Flu with 39°C fever → we still decode the latest #AIActus: https://youtu.be/5RkaR180khE On the agenda:
– The rumors about GPT 4.5
– The return of #ChatGPT Plus
– The @Tesla_Optimus robot
– Complete breakdown of the #AIAct -
AR-LLMs Skip Reasoning: Planning in Representation Space Needed
By
–
Auto-Regressive LLMs have a role to play: turning abstract ideas into token sequences (words, actions, code…).
But abstract ideas should be elaborated through planning/reasoning in representation space.
AR-LLMs go directly from prompt to answer, skipping the step of reasoning -
Approaching AI Agents Safety Development Iteratively
By
–
Figuring out how to do agents safely is deeply important work, I am glad we are approaching this iteratively.
-
OpenAI’s Superalignment Team Reveals Weak-to-Strong Model Alignment Research
By
–
OpenAI's superalignment team, co-led by @ilyasut
, has revealed its first research, exploring promising pathways to weak-to-strong model alignment (aka ways for puny humans to persuade ridonkulously smart AIs to obey them): -
Weak-to-Strong Generalization: Beyond RLHF for Superalignment
By
–
Naive weak supervision isn't enough—current techniques, like RLHF, won't be sufficient for future superhuman models. But we also show that it's feasible to drastically improve weak-to-strong generalization—making iterative empirical progress on a core challenge of superalignment
-

Weak-to-Strong Generalization: Supervising Smarter AI Systems
By
–
In the future, humans will need to supervise AI systems much smarter than them. We study an analogy: small models supervising large models. Read the Superalignment team's first paper showing progress on a new approach, weak-to-strong generalization: https://
openai.com/research/weak-
to-strong-generalization
… -
OpenAI Announces $10M Superalignment Fast Grants Program
By
–
We're announcing, together with @ericschmidt
: Superalignment Fast Grants. $10M in grants for technical research on aligning superhuman AI systems, including weak-to-strong generalization, interpretability, scalable oversight, and more. Apply by Feb 18! -
Solving AI Alignment: A Critical Technical Challenge Ahead
By
–
Figuring out how to ensure future superhuman AI systems are aligned and safe is one of the most important unsolved technical problems in the world. But we think it is a solvable problem. There is lots of low-hanging fruit, and new researchers can make enormous contributions!
-
Anticipation for Major AI Announcement Tomorrow
By
–
So let’s see if Jimmy and flowers are right and tomorrow will be announced something big. GPT 4.5? AGI?