looking for feedback on an upcoming open-weight language model (our first since GPT-2!): https://
openai.com/open-model-fee
dback/
…
OPEN SOURCE
-
OpenAI releases first open-weight language model since GPT-2
By
–
-
OpenHands LM-32B: Strong Open Coding Agent Model
By
–
Blog post: https://
all-hands.dev/blog/introduci
ng-openhands-lm-32b—-a-strong-open-coding-agent-model
… Model: -

Open-Source LLMs for Code Make Strong Comeback
By
–
Wow – open-source LLMs for code are so back Just one week after the latest DeepSeek release which brought vibe-coding for free, here is the high profile @allhands_ai team dropping a 32B model with matching performance on software engineering tasks benchmarks like SWE-Bench
-

OpenAI to Release New Open-Weight Reasoning Model
By
–
BREAKING : OpenAI is aiming to fulfil its prophecy and release a new open-weight reasoning model soon.
-

Metadata Routing in Scikit-Learn Framework
By
–
Metadata routing in scikit-learn https://
buff.ly/p60umcA #AI #MachineLearning #DeepLearning #LLMs #DataScience -

CPPO Accelerates Group Relative Policy Optimization Training
By
–
CPPO: Accelerating the Training of Group Relative Policy Optimization-Based Reasoning Models
Paper: https://
arxiv.org/pdf/2503.22342
Code: https://
github.com/lzhxmu/CPPO -

Building SQL Layer for Distributed Database Course
By
–
GitHub – talent-plan/tinysql: A course to build the SQL layer of a distributed database. https://
buff.ly/oxXh9lO
#AI #MachineLearning #DeepLearning #LLMs #DataScience -
Llama 3 8B quantized to 4-bit reduces memory usage by 64%
By
–
We converted a full-precision (FP16) Llama 3 8B model into a 4‑bit version using Hugging Face’s bitsandbytes library with the “nf4” configuration. This reduced the model’s memory usage from ~15 GB to ~5.4 GB—a savings of roughly 64%.
-
AI Models Going Open Source Locally Within Short Timeline
By
–
I'm 100% confident this will be OSS and in our hands running locally in a short while. I do not believe there is long term benefit for the giants. Secondly, this is still fair use and afaik this is all public domain stuff. In every universe our math this was inevitable 🙁
-
DeepSeek V3 and Vibe-Coding LLMs Transform App Development
By
–
The new DeepSite space on Hugging Face is totally insane for vibe-codershttps://t.co/dKsMEDWOLN
— Thomas Wolf (@Thom_Wolf) 30 mars 2025
With the new wave of vibe-coding-optimized LLMs like the latest open-source DeepSeek model (version V3-0324), you can basically prompt out-of-the-box and create any app and game in… pic.twitter.com/JqBFQ74NsVThe new DeepSite space on Hugging Face is totally insane for vibe-coders https://
huggingface.co/spaces/enzostv
s/deepsite
… With the new wave of vibe-coding-optimized LLMs like the latest open-source DeepSeek model (version V3-0324), you can basically prompt out-of-the-box and create any app and game in