Thrilled to have the authors of Direct Preference Optimization answering questions on their paper! DPO offers a simple alternative to RLHF and has been hugely impactful, including being used to train Llama 3! Talk to the authors @archit_sharma97 and team directly to learn more!
@askalphaxiv
-
Open Access Science and Democratic Peer Review Culture
By
–
Absolutely agree here. We hope that we can help by fostering a culture of open access within science, as well as making peer review a true community service that is democratic and accessible by everyone.
-

AI Scientist Platform Enables Research Reproducibility Testing
By
–
Sharing an idea that offers an exciting use case of the AI Scientist from @SakanaAILabs
! Rather than generating new ideas, one can adapt the code to pick a paper and test its results. "Reproducibility as a platform." Thanks lakshaytalkstocomputer and @cong_ml for the insight! -

TaskGen: Open-Source LLM Agentic Framework for Task Automation
By
–
Introducing TaskGen, an open-sourced task-based, memory-infused LLM agentic framework that uses an Agent to break down a task into subtasks and assigns each subtask to either an equipped function or another Agent. Talk directly with @johntanchongmin and the TaskGen team here!
-

Code Data Improves LLM Generalization Beyond Coding Tasks
By
–
New from @CohereForAI and @cohere
! To Code, or Not To Code? This paper reveals that including code data in LLM pre-training is critical for generalization far beyond coding tasks. The authors @viraataryabumi @ahmetustun89 and @sarahookr are here to answer your questions! -
Community Comments on Preprints Improve Transparency and Quality
By
–
We absolutely agree here and hope that community comments on preprints will help improve transparency and quality.
-
Community engagement beyond trending papers in AI research
By
–
Thanks for bringing this up @lauriewired
! At this early stage of the community, we agree that the most trending papers are receiving the most comments, but there are many papers that are not necessarily trending but still receive a question or comment from a curious student or -
Vercel platform praised for development workflow
By
–
Thanks for sharing! @vercel has been great for us https://t.co/ruUO2irxop
— alphaXiv (@askalphaxiv) 10 septembre 2024Thanks for sharing! @vercel has been great for us
-
Platform adds subject filtering for AI research discussions
By
–
Thank you, Prof. Osmane! Apologies for the slow loading times, we are fixing this ASAP. As you suggested, we are also planning to add filtering based on subject/subfield to allow researchers to view discussion in their specific area. How we choose these subfields is an
-

Meta Waymo USC Develop Transfusion for Multimodal AI
By
–
Scientists from @AIatMeta
, @Waymo
, and @USC develop Transfusion, a recipe combining language modeling and diffusion to train models that generate both text and image. Talk to @violet_zct directly to discover possibilities for exciting multi-modal models offered by Transfusion!
