Over 625 T/sec. Priced at $0.11 input & $0.34 output per M tokens. Officially the lowest price high performance provider of Llama 4 Scout, without compromise on context window or tradeoffs on speed
LLMS
-

Llama4: 3 new Open Source AIs with 10 million tokens
By
–
Boom → @MetaAI just released #Llama4, 3 new Open Source AIs with 10 million tokens. What is it worth and how to use Llama 4 for free? I tell you everything here https://
youtu.be/-Ns2-CfgXC4 -

Llama 4 Scout Available on SambaNova Cloud at 697 Tokens/Second
By
–
Llama 4 Scout from @AIatMeta is now available on SambaNova Cloud! Fastest Inference clocked at 697+t/s. Llama 4 Maverick will be out next week, followed by higher context lengths up to 128K! Try it now on SambaNova Cloud
-
Llama 4 Scout Achieves Record 697 Tokens Per Second Speed
By
–
Don’t blink! You might miss just how fast we’re going⚡️
— SambaNova (@SambaNovaAI) 7 avril 2025
🚀 697 t/s on @AIatMeta's #Llama4, independently verified by @ArtificialAnlys
"…the fastest output speed we have measured yet for Llama 4 Scout.” — @_micah_h
Try it now 👇Don’t blink! You might miss just how fast we’re going 697 t/s on @AIatMeta
's #Llama4, independently verified by @ArtificialAnlys "…the fastest output speed we have measured yet for Llama 4 Scout.” — @_micah_h Try it now -
AI Agents Enable Interactive Stock Price Dashboards with Real-Time Analytics
By
–
With AI agents, you can have a simple interactive dashboard of stock prices. This was made on ChatLLM, and you can have access to easy analytics and real-time dashboards. pic.twitter.com/GA7Hwq1W33
— Abacus.AI (@abacusai) 7 avril 2025With AI agents, you can have a simple interactive dashboard of stock prices. This was made on ChatLLM, and you can have access to easy analytics and real-time dashboards.
-

Claude and Llama Struggle with Japanese Document Understanding
By
–
Claude 3.5/Llama 3.2 ace English DocVQA, but how do they fare in Japanese? New benchmark alert: JDocQA (curated by @NAIST_MAIN
, @RIKEN_RCCS
, ATR) exposes multilingual gaps in top VLMs. Read more -

AI Index 2025 Report: LLM Costs and Performance Insights
By
–
Curious about the current landscape of artificial intelligence? The #AIIndex2025 report includes fresh insights on the cost of large language models, who’s leading on AI performance, and more. Get a quick overview of the report with these 10 charts: https://
hai.stanford.edu/news/ai-index-
2025-state-of-ai-in-10-charts
… -

Google Gemini 2.5 Pro Public Preview Launches with Surging Demand
By
–
Google moved Gemini 2.5 Pro into public preview in its AI Studio, with higher rate limits
— Rowan Cheung (@rowancheung) 7 avril 2025
The model is seeing massive demand, leading to an 80%+ surge in active users in AI Studio + Gemini API.
2.5 Pro's I/O price per million tokens starts at $1.25/$10pic.twitter.com/HFUlC6p7G5Google moved Gemini 2.5 Pro into public preview in its AI Studio, with higher rate limits The model is seeing massive demand, leading to an 80%+ surge in active users in AI Studio + Gemini API. 2.5 Pro's I/O price per million tokens starts at $1.25/$10
-

OpenAI Delays GPT-5 Release, Plans o3 and o4-mini First
By
–
Sam Altman shared an update to OpenAI's model roadmap
— Rowan Cheung (@rowancheung) 7 avril 2025
He said the company will first release o3 and o4-mini and then follow up with a “much better than originally thought” GPT-5
Earlier, the plan was to launch GPT-5 right away, skipping the o models!pic.twitter.com/tJRIXFNPfqSam Altman shared an update to OpenAI's model roadmap He said the company will first release o3 and o4-mini and then follow up with a “much better than originally thought” GPT-5 Earlier, the plan was to launch GPT-5 right away, skipping the o models!
-
Meta AI Launches Llama 4 with 10M Token Context Window
By
–
TODAY'S AI NEWS: Meta AI just dropped Llama 4 AI with a 10M token context window Plus, more news from Midjourney, Microsoft, OpenAI, Google, and Kawasaki Here's everything you need to know:
