check out smallthinker preview on @apolloaiapp
. might be the first on a phone
LLMS
-
SmallThinker Preview Launches on Apollo AI App
By
–
-
Mobile Reasoning Models Now Available: AI Evolution Accelerates
By
–
Mobile reasoning models are already here. Things are moving so fast.
-
Significance of #8 grows; Grok used for daily thread summaries
By
–
The significance of #8 will grow a lot in 2025 I finally turned a point where I am asking Grok to summarise threads into a complete post almost every day now
-

RAG, Attention Mechanisms, and Mamba Architectures in LLMs
By
–
Databricks research scientist @shashank_r12 s shares approaches in LLMs:
– How RAG enhances accuracy – Evolution of attention mechanisms
– Practical applications & trade-offs of Mamba architectures -

Deploying LLMs with low latency and real-time responsiveness
By
–
Deploying LLMs with low latency and real-time responsiveness is no small task. But with @datarobot and Cerebras Inference, you’ll be ready to customize and deploy LLMs that deliver speed, precision, and real-time responsiveness. Dive in: https://
hubs.li/Q031lCn10 -
Scaling Test-Time Compute: New Frontier in AI Model Optimization
By
–
https://
huggingface.co/spaces/Hugging
FaceH4/blogpost-scaling-test-time-compute
… -
3B Reasoning Model Outperforms 70B: AI Intelligence Surge Expected
By
–
I keep coming back to this graph. A 3b reasoning model beating its 70b equivalent is insane. It’s still ridiculously early for reasoning too. Every model is going to 10X+ in intelligence this year.
-
User prefers o1 over Claude Pro for most tasks
By
–
100% I mostly use o1 and only use pro for really hard questions