Here's details on Meta's 24k H100 Cluster Pods that we use for Llama3 training.
* Network: two versions RoCEv2 or Infiniband. * Llama3 trains on RoCEv2
* Storage: NFS/FUSE based on Tectonic/Hammerspace
* Stock PyTorch: no real modifications that aren't upstreamed
* NCCL with
HARDWARE
-
Meta’s 24k H100 Cluster Pods Infrastructure for Llama3
By
–
-

NVIDIA GTC 2026: Unify Your AI Stack with Hardware and Software
By
–
@NVIDIAGTC is almost here! Join us next week at booth #1603 to meet with our #appliedAI experts and learn how @nvidia and DataRobot can help you unify your AI stack with the hardware and software built for AI. Register today with our exclusive discount:
-

Pure Storage AI Platform at NVIDIA GTC 2024 Booth 1529
By
–
Visit @PureStorage and explore #AI at #NVIDIA #GTC2024 — Meet Pure staff at booth #1529 to learn how their data platform for AI can help your organization accelerate model training and inference, improve operational efficiency, & more. Read more details: https://
blog.purestorage.com/news-events/vi
sit-pure-storage-and-explore-ai-at-nvidia-gtc/
… -
South Korean Startup Launches $1,800 AI Companion Doll
By
–
South Korean startup Hydrol AI launched a $1,800 AI-powered companion doll.
— Rowan Cheung (@rowancheung) 12 mars 2024
The doll is aimed at alleviating loneliness among the country's rapidly aging population.https://t.co/soTpSSvJWgSouth Korean startup Hydrol AI launched a $1,800 AI-powered companion doll. The doll is aimed at alleviating loneliness among the country's rapidly aging population.
-

AI at the Edge: Embedded Machine Learning for Real-World Solutions
By
–
#AI at the #Edge — Solve Real-World Problems with Embedded #MachineLearning: http://
amzn.to/3GN70uC by @dansitu & @jennymplunkett
——
#IoT #IIoT #AIoT #EdgeAI #ML #DigitalTransformation #Industry40 #BigData #DataScience #DeepLearning #EdgeComputing #IoTSlam @IoTChannel -

AI Memory Chip Shortage Creates Investment Opportunities
By
–
Can we pick winners in AI’s memory obsession? The ever-increasing scale of artificial intelligence programs such as OpenAI’s ChatGPT is creating a crisis where the available memory chips can’t keep up with today’s fastest processors from Nvidia and others. Investors should look
-
Gaudi 2 Achieves 28% Faster Inference Speed Than A100
By
–
On Stable Beluga 2.5 70B, our fine-tuned version of Llama 2 70B, Gaudi 2 achieved 28% faster inference speed for tokens/second per accelerator versus the A100. Read the full analysis and learn how our findings underscore the need for alternatives in compute solutions here:
-
Gaudi 2 Outperforms H100 and A100 GPUs for Stable Diffusion 3
By
–
For Stable Diffusion 3, we measured the training throughput for the 2B Multimodal Diffusion Transformer (MMDiT) architecture model. Gaudi 2 trained images 1.5x faster than the H100-80GB, and 3x faster than A100-80GB GPU’s when scaled up to 32 nodes. (2/3)
-

Intel Gaudi 2 vs Nvidia A100 H100 Training Speed Comparison
By
–
In this installment of "Behind the Compute", a series dedicated to offering insights for others to harness the power of generative AI, we compared the training speed of @Intel Gaudi 2 accelerators versus @Nvidia
's A100 and H100 for two of our models. (1/3) -

Century-old tech firms: ZEISS optics leadership to modern semiconductors
By
–
Most of the oldest companies are breweries, banks, or quasi-governmental, but I would like to see more about firms that have been technological leaders for over a century. Like ZEISS was a leader in optics since 1843, and makes the UEV Bragg reflectors for ASML machines today.