“Flash Thinking (our first ‘thinking’ model, expect a lot more news on this soon – as many of you remember, we pioneered this type of model with AlphaGo, AlphaZero, AlphaProof…)”
LLMS
-
Comparing compute efficiency with o3 low-compute version
By
–
It's in the same ballpark as the low-compute version of the o3 entry — so way less than the high-compute version
-
Google Releases Imagen 3, Veo 2, Genie 2 and Gemini 2.0 Flash
By
–
It’s been an amazing last couple of weeks, hope you enjoyed our end of year extravaganza as much as we did! Just some of the things we shipped: state-of-the-art image, video, and interactive world models (Imagen 3, Veo 2 & Genie 2); Gemini 2.0 Flash (a highly performant and
-
Jeremy Berman open-sources ARC-AGI solution with LLM genetic programming
By
–
Jeremy Berman, former #1 on the ARC-AGI-Pub leaderboard, just open-sourced his solution as a template on Params (link in next tweet) It uses LLM-driven genetic programming, which has turned out to be way more powerful than anybody expected. You can book a consulting call with
-
Gemini models stand out in pro offerings
By
–
Gemini models become too good? They are def adding value to the pro offering purely cuz they are quite different from others (writing style, 2M context, etc)
-

Ollama Now Supports Private GGUF Models from Hugging Face
By
–
Not your weights, not your brain! Starting today, you can run your private GGUFs from the Hugging Face hub directly in @ollama! 🔥
— Vaibhav (VB) Srivastav (@reach_vb) 23 décembre 2024
You asked, we delivered!
Works out of the box, all you need to do is add your Ollama SSH key to your profile, and that's it! ⚡
Run private… pic.twitter.com/v832KNilLCNot your weights, not your brain! Starting today, you can run your private GGUFs from the Hugging Face hub directly in @ollama
! You asked, we delivered! Works out of the box, all you need to do is add your Ollama SSH key to your profile, and that's it! Run private -
Internet Access More Effective Than Hallucination Research for Models
By
–
Realization: the old style of “hallucinations research” via self-calibration is probably going to die down. I used to be very excited about it but now I am skeptical because giving models internet access (e.g., searchGPT, perplexity) is turning out to be way higher ROI. When
-
GroqCloud Unleashes AI Inference Performance with AppGen
By
–
1. Reduce the imagination gap.
— Groq Inc (@GroqInc) 23 décembre 2024
2. Unlock the potential of others.
3. Provide them the most performant, reliable, and affordable inference for their AI apps ever.
The GroqCloud™ team is cooking with chef @RickLamers 🧑🍳 Try it yourself: https://t.co/DNxCmmdtpN https://t.co/dF1JACX5ws1. Reduce the imagination gap.
2. Unlock the potential of others.
3. Provide them the most performant, reliable, and affordable inference for their AI apps ever.
The GroqCloud™ team is cooking with chef @RickLamers Try it yourself: https://
appgen.groqlabs.com -

Open Weight Model Deployment Alternative to Closed Source APIs
By
–
Open Science FTW! That's a fully open weight model, that you can deploy wherever you want, however you want mogging closed source APIs!