http://
H2O.ai h2oGPTe Agentic AI Webinar Recording is Now Live! Check it our here: https://
youtube.com/watch?v=m3I3ro
_ZAnE
… @ArnoCandel @pseudotensor #AI #GenAI #AgenticAI #AgenticRAG
LLMS
-
H2O.ai Releases h2oGPTe Agentic AI Webinar Recording
By
–
-
Cost comparison: AI vs human brain
By
–
—
Just so we're clear: Buy dual RTX 4090s and 128 GB RAM = $5k Download DeepSeek R1 70b weights, run offline – unlimited use You can buy a 130 IQ human brain for the cost of a used Civic.
— -
R1 Chain-of-Thought Reasoning Offers Superior LLM Responses
By
–
reading r1 chain-of-thoughts is quite hypnotizing hits different than your usual LLM straight answer
-
OpenAI o1 Think Deeper Now Free for All Copilot Users
By
–
Today we’ve made Think Deeper free and available for all users of Copilot.
— Mustafa Suleyman (@mustafasuleyman) 29 janvier 2025
This now gives everyone access to OpenAI’s world class o1 reasoning model in Copilot, everywhere at no cost.
I urge you to give it a try. It’s truly magical. Think Deeper helps you: pic.twitter.com/nzccl0tdhLToday we’ve made Think Deeper free and available for all users of Copilot. This now gives everyone access to OpenAI’s world class o1 reasoning model in Copilot, everywhere at no cost. I urge you to give it a try. It’s truly magical. Think Deeper helps you:
-
Crowdsourced Distributed Fine-Tuning: A Year-Long Advocacy
By
–
I've been arguing for something like this for over a year: crowdsourced distributed fine-tuning.
-
Groq’s Llama-3.3 Distilled from DeepSeek-R1 Open Source
By
–
Groq's Llama-3.3 distilled from DeepSeek-R1 would belong to the first line.
#OpenSourceMagic -
GSM8K Benchmark: Questioning AI Mathematical Reasoning Validity
By
–
Academics who aren't bought (at least yet) "The GSM8K benchmark is widely used to assess the mathematical reasoning of models…it remains unclear whether their mathematical reasoning capabilities have genuinely advanced, raising questions about the reliability of the
-
OpenAI o1 Model Enhances LLM Reasoning Through Inference Compute
By
–
HuggingFace: "OpenAI’s o1 model showed that when LLMs are trained to do the same—by using more compute during inference—they get significantly better at solving reasoning tasks like mathematics, coding, and logic."
-

TinyZero: Affordable RL Finetuning Under $30
By
–
TinyZero reproduction of R1-Zero
"experience the Ahah moment yourself for < $30" Given a base model, the RL finetuning can be relatively very cheap and quite accessible. -
Building Diverse RL Environments for LLM Cognitive Strategy Development
By
–
For friends of open source: imo the highest leverage thing you can do is help construct a high diversity of RL environments that help elicit LLM cognitive strategies. To build a gym of sorts. This is a highly parallelizable task, which favors a large community of collaborators.