“The demand for energy to power AI models will be infinite.” https://t.co/OxoZGRiONh
— Marek Rosa | European🇪🇺 | South African🇿🇦 (@marek_rosa) 6 octobre 2024
“The demand for energy to power AI models will be infinite.”
By
–
“The demand for energy to power AI models will be infinite.” https://t.co/OxoZGRiONh
— Marek Rosa | European🇪🇺 | South African🇿🇦 (@marek_rosa) 6 octobre 2024
“The demand for energy to power AI models will be infinite.”
By
–
Here’s an equation that represents I being powered by both T and A: (The cataclysmic future of AI, It's Nice Weather). [
I = frac{A}{T}
] In this equation:
– I is powered by T and A, implying that I is determined by the ratio of T and A. – This equation suggests that

By
–
SCADA systems have limitations. Learn how a #TimeSeries Database can unlock new possibilities for industrial operations! Explore more: https://
buff.ly/3zpCMhB #sponsored #influxdata_iiot #InfluxDB #IIoT #SmartIndustry #TechSolutions #SCADA @YvesMulkers via @fogoros

By
–
In today’s evolving AI landscape, businesses need solutions that not only drive innovation but also provide governance and security. “We need to bridge the AI value gap, and not just focus on does AI work in a demo… but think through what business processes AI is improving,

By
–
The Importance of Time Series Data in Manufacturing >> https://
buff.ly/3N8bVKl #sponsored #influxdata_iiot #InfluxDB #TimeSeries #manufacturing #Industry40 @DrHassanRashidi via @fogoros
By
–
RAS is blackwell+, so TBD. no single GPU isolation, whole host gets booted.
By
–
the scheduler side of things is a whole another level of complexity. we dont reprovision from scratch though.
By
–
"How to train a model on 10k H100 GPUs?"
has now been immortalized on my blog: https://
soumith.ch/blog/2024-10-0
2-training-10k-scale.md.html
…
By
–
one more thing to add.
at this scale, we also have to adjust the actual packet routing algorithms in our switches and NICs, to be able to load-balance well. Did you know switches have to have significant HBM memory as well (not just GPUs) because as the packets queue up, they
By
–
quick mini-post I wrote for @francoisfleuret broadly summarizing the things one needs to do to train models on 10k+ H100s.