3/5 Where most other tiny models choke at context lengths above 8K, Jamba Reasoning 3B stays steady, with a consistent 30-40 tokens/second on an M3 MacBook Pro, regardless of context size. This is up to an order of magnitude faster than other on-device models.
GENERATIVE AI
-

Jamba Reasoning 3B Excels in Knowledge and Instruction-Following
By
–
2/5 Jamba Reasoning 3B excels in general knowledge benchmarks (MMLU-Pro, HLE) and instruction-following (IFBench), making it reliable and performant
-

Jamba Reasoning 3B: Hybrid SSM-Transformer Apache 2.0 Release
By
–
1/5 Releasing Jamba Reasoning 3B under Apache 2.0: Hybrid SSM-Transformer architecture that tops accuracy & speed across record context lengths. e.g. 3-5X faster than Llama 3.2 3B and Qwen3 4B at 32K tokens.
-
GPT-Realtime and E2E Models in Production Projects
By
–
little survey: are you using gpt-realtime/ mini or any other E2E model in projects? if yes, what for?
-

Sam Altman, Gemini 2.5, and AI tools roundup
By
–
Top stories in AI today: – Sam Altman on Dev Day, AGI, and more
– Google releases Gemini 2.5 Computer Use
– Create LinkedIn carousels in ChatGPT with Canva
– Duke’s AI for smarter drug delivery
– 4 new AI tools, community workflows, and more Read more: https://
therundown.ai/p/exclusive-in
terview-sam-altman-on-dev-day-and-ais-future
… -
ChatGPT-5 Pro Generates Custom Dev Tool Configurations
By
–
pro tip: tell ChatGPT-5 Pro (or o3) “build me tmux/neovim/vscode configs based on my workflow from our conversations” you’ll be surprised how good it actually is
-
Gaussian Embeddings: How JEPAs Learn Data Density
By
–
Gaussian Embeddings: How JEPAs Secretly Learn Your Data Density Paper:
-
The Economics of Prompts: Why They Are New Purchase Buttons
By
–
He publicado un episodio en @ivoox
: "La Economía del Prompt: por qué son los Nuevos Botones de Compra #podcast -

Nvidia Fast-dLLM v2: Efficient Block-Diffusion LLM
By
–
Nvidia presents Fast-dLLM v2
— AK (@_akhaliq) 8 octobre 2025
Efficient Block-Diffusion LLM pic.twitter.com/wjqpvi1LoANvidia presents Fast-dLLM v2 Efficient Block-Diffusion LLM

