torch.compile is cool but LLM compile: takes your .py repo as string and outputs a brand new, custom, from scratch, minimal code repository directly running your network in highly optimized CUDA
COMPUTING
-

Painting Robots: Art Meets Artificial Intelligence and Innovation
By
–
Painting #Robots
— Ronald van Loon (@Ronald_vanLoon) 12 avril 2024
by @anand_narang#Robotics #ArtificialIntelligence #Innovation #Autonomous #Technology #FutureTech
cc: @karpathy @ravikikan @patrickgunz_ch pic.twitter.com/TwwUchzbxoPainting #Robots
by @anand_narang #Robotics #ArtificialIntelligence #Innovation #Autonomous #Technology #FutureTech cc: @karpathy @ravikikan @patrickgunz_ch -
llm.c versus PyTorch: speed and simplicity comparison
By
–
…"in that llm.c will never be as fast nor as simple as pytorch"
-

Deep Learning Performance Optimization: Complexity and Resources
By
–
This post became popular; Few more thoughts / pointers on the topic for the interested reader. Example of the complexity involved: @cHHillee has a great post "Making Deep Learning Go Brrrr From First Principles" https://
horace.io/brrr_intro.html
I was always struck by this diagram from -

Low-Level Code Performance Optimization Zeitgeist
By
–
well that’s exactly what I said; anyway i’m referring to this tweet about performance optimizations, as well as the greater zeitgeist where everyone things it’s faster to write low-level code
-
Unity Catalog: Data Lineage and Feature Store Capabilities
By
–
The technology behind Unity Catalog supports a variety of business outcomes: faster innovation, cost reduction, compliance support & more. Dive into some of the capabilities that make this possible, such as our data lineage, Feature Store, & more!
-
NVIDIA Blackwell GPUs Power New Computing and Generative AI Era
By
–
Get an inside look at NVIDIA Blackwell GPUs and the latest data center technologies that are powering the new era of computing and #generativeAI. https://t.co/wjD43zOj3b #GTC24 pic.twitter.com/P2esW3Ej3c
— NVIDIA (@nvidia) 12 avril 2024Get an inside look at NVIDIA Blackwell GPUs and the latest data center technologies that are powering the new era of computing and #generativeAI. https://
nvda.ws/3xyL9Gl #GTC24 -
LLM.c Versus PyTorch: Practical Limitations of Low-Level Implementation
By
–
my take on all this llm.c stuff is that it’s very impressive, and karpathy is certainly brilliant, but it was sort of a futile exercise in that llm.c will never be as fast nor as simple as pytorch this whole movement to “write everything in the lowest language possible” is a tad
-
Intel Unveils Gaudi 3, Its New AI Acceleration Chip
By
–
[#Article] #GenAI: Intel Presents Gaudi 3, Its Latest AI Acceleration Chip https://actuia.com/actualite/genai-intel-presente-gaudi-3-sa-derniere-puce-dacceleration-dia/
… #AI #artificialintelligence -
Ripgrep UI for searching public source code repositories
By
–
As long as the total source code of all public Vals adds up to less than a few hundred MBs it may be worth brute forcing it with ripgrep – I built a UI for that for a bunch of my stuff here https://
ripgrep.datasette.io/-/ripgrep?patt
ern=hookimpl
… – see