"The recent announcement of Aramco Digital partnering with Groq to deliver market-leading AI inference represents a key partnership in one of the most important and fastest-growing regions for AI investment and consumption." – @danielnewmanUV Read more:
AI HARDWARE
-
Exceptional performance-to-cost ratio achievement recognized
By
–
you guys did incredible work; the performance to cost ratio is so good!
-
Scaling LLM Training: Double Digit Batch Size for 70B Models
By
–
double digit batch size, eventual goal is one rack for 70B.
-

SambaNova Cloud Launches Fastest AI Inference Platform Free
By
–
The fastest API for #devs is here — for FREE. This week, we launched SambaNova Cloud, the only platform that offers #AI inference of over 570 t/s on @AIatMeta
’s Llama 3.1 70B & 100 t/s for Llama 3.1 405B. Read more https://
sambanova.ai/blog/fastest-i
nference-best-models
… -
Google’s Multimodal AI and Enterprise Solutions Lead Week’s Tech Advances
By
–
More Big News: Google’s testing a feature that understands your pictures & turns docs into discussions. Anthropic’s Workspaces offers more control over enterprise AI. HyperWrite’s Reflection AI model is creating buzz. Samsung updates Galaxy features. #AI #TechNews
-
Oracle Announces Zettascale OCI Supercluster for AI Training
By
–
At #OCW24, @Oracle announced the first zettascale OCI Supercluster accelerated by the #NVIDIABlackwell platform, allowing customers to train and deploy next-generation #AI models at scale. Learn more: https://
nvda.ws/4ehTXAq -
Groq partners with Aramco Digital for world’s largest inference data center
By
–
We are proud to announce our partnership with @Aramco Digital to establish the world's largest inference data center using Groq® LPU™ AI inference technology. Read the full press release here.
-

Cerebras Inference Speed Update: Llama 3.1 Performance Benchmarks
By
–
Cerebras Inference perf update:
Llama3.1-8B: 1,8001,927 tokens/s
Llama3.1-70B: 450481 tokens/s
Stillfor the most popular open model in the world. https://
inference.cerebras.ai -
Coding Inside Vision Pro: New Development Possibilities
By
–
Or you could try coding inside Vision Pro
-

Flux LoRA Training Space Updated with L40S GPU for Speed
By
–
Flux LoRA Training Space has now been updated to use the L40S GPU. Now its even faster and cheaper to train your own Flux LoRA Training Space: https://
huggingface.co/spaces/autotra
in-projects/train-flux-lora-ease
…