Robin: A multi-agent system for automating scientific discovery Robin is the first multi-agent AI system to automate the entire scientific discovery loop—literature review, hypothesis generation, experiment design, data analysis, and hypothesis refinement—within a single
@askalphaxiv
-

GUI-G1 and Robin: High-Performance Agents with Multimodal Improvements
By
–
Major upgrades for high-performance agents with the rise of GUI-G1 and Robin, alongside the notable multimodal improvements of diffusion language models Check out the top 10 papers for the week – GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI
-
Claude 4 Sonnet Enables Advanced Research Paper Analysis
By
–
Introducing Claude 4 Sonnet for understanding arXiv papers 🚀
— alphaXiv (@askalphaxiv) 22 mai 2025
Highlight any section of a paper to ask questions and “@” other papers to quickly add to context and compare results, benchmarks, etc. pic.twitter.com/x3aZtaIT9rIntroducing Claude 4 Sonnet for understanding arXiv papers Highlight any section of a paper to ask questions and “@” other papers to quickly add to context and compare results, benchmarks, etc.
-

GenAI Advances Autonomous Driving to Level 5 Autonomy
By
–
How can Generative AI unlock Level 5 autonomy? This new survey explores how GenAI is reshaping the autonomous driving stack Covers VAEs, Diffusion, LLMs Applications in LiDAR, trajectory, digital twins Tackles long-tail generalization & safety bottlenecks
-

LLMs Enable More Precise Cyberattacks on Personal Data
By
–
LLMs are fundamentally changing cyberattacks. This new paper from Anthropic and Google shows that LLMs Extract passwords, credit cards, and SSNs with higher precision Mine sensitive personal info for tailored blackmail Lower the cost of targeting vulnerable software
-

Hidden Economic Knowledge in LLM Embeddings Revealed
By
–
Revealing economic facts: LLMs know more than they say This paper shows that hidden states (embeddings) of large language models (LLMs) contain rich economic information that can be used to estimate and impute economic and financial statistics more accurately than the LLMs’ text
-

LightLab: Diffusion Model Controls Light Sources in Images
By
–
LightLab: Controlling Light Sources in Images with Diffusion Models LightLab presents a diffusion-based method for precise, parametric control over light sources in a single image, enabling users to edit light intensity and color with photorealistic results. The approach
-

Maya: Open-Source Multilingual Vision Language Model
By
–
Behind Maya: Building a Multilingual Vision Language Model This paper introduces Maya, an open-source multilingual Vision-Language Model (VLM) designed to enhance performance on vision-language tasks across eight diverse languages. Maya addresses the underperformance of existing
-

Open-Source Edge Computing Simulators: Classification and Evaluation
By
–
A Survey on Open-Source Edge Computing Simulators and Emulators: The Computing and Networking Convergence Perspective This survey provides a comprehensive classification and evaluation of 40+ open-source edge computing simulators and emulators, emphasizing the convergence of
-

MiMo-7B: Reasoning-Optimized Language Model Outperforms Larger Models
By
–
MiMo: Unlocking the Reasoning Potential of Language Model — From Pretraining to Posttraining MiMo-7B is a reasoning-centric LLM series optimized from pretraining through posttraining, achieving state-of-the-art performance on math and code tasks—even outperforming models 4× its
