


Overriding the proprietary prompt of OpenAI’s ChatGPT to make it:
1. sass you
2. scream
3. talk in an uwu voice
4. be distracted by a toddler while on the phone with you

By
–



Overriding the proprietary prompt of OpenAI’s ChatGPT to make it:
1. sass you
2. scream
3. talk in an uwu voice
4. be distracted by a toddler while on the phone with you

By
–
The fact is that http://
MONTREAL.AI is a worldwide leader in #ReinforcementLearning and #AIAgents since at least 2015 . . . #MontrealAI
By
–
Sufficiently advanced executive communication is indistinguishable from prompt engineering
By
–
#CICERObyMetaAI is the first AI built to play Diplomacy at a human level. Patience, honesty, cooperation, negotiation skills and empathy: all human attributes that contribute to its gameplay. Which attribute do you think makes the better Diplomacy player?

By
–
LangChain version 0.0.27: Much appreciated clean up and formatting by `tonyabracadabra` (if you are on Twitter, lmk!!) ReActTextWorldAgent: a new agent from @JavaFXpert – see below for more!
By
–
Consciousness is a combination of self and environmental awareness. I see environmental awareness in Minedojo as tractable but self-awareness is a different story. For that, you need uncertainty quantification is key: unknown unknowns is tricky.

By
–
Thank you @davidchalmers42 for featuring https://
minedojo.org in your talk. We build open-ended agents that can solve any task in Minecraft. This requires the agent to have awareness of the Minecraft world. Foundation models provide the agent with priors (e.g. reward function)
By
–
Learning General World Models in a Handful of Reward-Free Deployments.
CASCADE seeks to learn a world model by collecting data with a population of agents. @YingchenX
, @JParkerHolder
, @AldoPacchiano
, @PhilipJohnBall
, @_Oleh
, Stephen J. Roberts, @_Rockt
, @EGrefen
By
–
Grounding Aleatoric Uncertainty for Unsupervised Environment Design. @MinqiJiang
, @MichaelD1729
, @JParkerHolder
, @_AndreiLupu
, Heinrich Küttler, @EGrefen
, @_Rockt
, @J_Foerst propose SAMPLR, a minimax regret UED method that optimizes the ground-truth utility function.
By
–
Improving Policy Learning via Language Dynamics Distillation. @hllo_wrld
, @Jayelmnop
, @LukeZettlemoyer
, @EGrefen
, @_rockt propose Language Dynamics Distillation (LDD), which pretrains a model to predict environment dynamics given demonstrations with language descriptions