A Clean Slate for Offline RL
Paper: https://
arxiv.org/pdf/2504.11453
@jiqizhixin
-

Clean Slate for Offline Reinforcement Learning Research
By
–
-

Oxford Unifloral Framework Advances Offline Reinforcement Learning
By
–
Offline RL is a mess: unclear goals, tangled code, and sneaky online tuning A team from University of Oxford fixs it with: Unifloral — one clean framework, shared hyperparams
Result? New SOTA algorithms: TD3-AWR & MoBRAC. -
ICLR Experiments with AI Agents for Meta-Reviewing Peer Feedback
By
–
ICLR 2025 is experimenting with an AI agent that leverages multiple LLMs to provide optional feedback on reviewer reports. It’s an interesting move—using LLMs to meta-review human reviews. Future of peer review?
-
OpenAI Builds AI-Powered Social Platform to Compete with X
By
–
According to The Verge, OpenAI is building its own version of 𝕏. A one-stop AI platform? Social + chat + tools? The lines are starting to blur…
-

Reasoning Models Effective Without Extended Thinking Process
By
–
Reasoning Models Can Be Effective Without Thinking https://
arxiv.org/pdf/2504.09858 -

Reasoning models achieve top performance without extended thinking
By
–
So apparently… Reasoning models can be effective without thinking! This flips the table on assumptions that long, detailed chains of thought are always necessary. Turns out, with the right training methods, even concise responses can deliver top-tier performance.
-

GPT-4.1 Prompting Guide: Strategies from OpenAI
By
–
If you're looking for the best way to prompt GPT-4.1, you should definitely check out the GPT-4.1 Prompting Guide—straight from OpenAI, the folks who built it. It’s packed with practical strategies to get the most out of the model. https://
cookbook.openai.com/examples/gpt4-
1_prompting_guide
… -
AI Model Enhanced: Better Instruction-Following and Cinematic Aesthetics
By
–
It now offers significantly better instruction-following, enhanced cinematic aesthetics, and richer artistic diversity. With support for over 60 stylized effect transfers, the model outputs are now more creative, imaginative, and visually compelling than ever. pic.twitter.com/0x7W6s4TEK
— 机器之心 JIQIZHIXIN (@jiqizhixin) 15 avril 2025It now offers significantly better instruction-following, enhanced cinematic aesthetics, and richer artistic diversity. With support for over 60 stylized effect transfers, the model outputs are now more creative, imaginative, and visually compelling than ever.
-
Kling 2.0 Launches with Enhanced Literary and Artistic AI Capabilities
By
–
Kling 2.0 is out!
— 机器之心 JIQIZHIXIN (@jiqizhixin) 15 avril 2025
Kling 2.0 is here with a full upgrade in literary and artistic capabilities! 🎨🎬
Try it now: https://t.co/gEBzZHIaWH pic.twitter.com/F0snBdHLzqKling 2.0 is out!
Kling 2.0 is here with a full upgrade in literary and artistic capabilities! Try it now: https://
app.klingai.com/global/ -

GRPO Simplified: Implementation Guide for Language Models
By
–
GRPO,simplified. https://
k-a.in/grpo.html
and
Implementing GRPO https://
k-a.in/grpo-1B.html
by @arjunkocher