Triton can generate anything, as long as you write a Triton backend. As of today, upstream Triton can generate AMD code too.
HARDWARE
-
Simplified Management of Serverless and Dedicated LLM Deployments
By
–
ICYI: Managing #serverelss and dedicated #LLM deployments has never been easier! 🚀
— Predibase by Rubrik (@predibase) 10 avril 2024
😎 see all #deployments and status in one place
🆕 create new dedicated deployments with just a few clicks
🛠️ select the right #GPU for the job and customize autoscalinghttps://t.co/W9hCTKaySr pic.twitter.com/zImeI45nNFICYI: Managing #serverelss and dedicated #LLM deployments has never been easier! see all #deployments and status in one place create new dedicated deployments with just a few clicks select the right #GPU for the job and customize autoscaling https://
pbase.ai/449UUXT -
Meta Launches MTIAv2 Inference Chip for AI Acceleration
By
–
Meta announces 2nd-gen inference chip MTIAv2.
* 708TF/s Int8 / 353TF/s BF16
* 256MB SRAM, 128GB memory
* 90W TDP. 24 chips per node, 3 nodes per rack.
* standard PyTorch stack (Dynamo, Inductor, Triton) for flexibility Fabbed on TSMC's 5nm process, its fully programmable via the -

Intel Modernizes Future Networks with Ecosystem Partners at MWC24
By
–
At #MWC24 @intel showed how they are working alongside their world-class ecosystem to modernize and monetize the #network of the future, today. @IntelEdge #AI #IoT #IntelAmbassador @pierrepinna @Hal_Good @gvalan @enilev @Analytics_699 @AlexMachicado
-

Robots Autonomously Assemble Fully Functional Mini Electric Vehicle
By
–
#Robots working together to make a fully functional mini EV Autonomously
— Ronald van Loon (@Ronald_vanLoon) 10 avril 2024
by @Ambots3D#MI #Robotics #RPA #Innovation #FutureOfWork #Automation
cc: @pascal_bornet @pbalakrishnarao @ravikikan pic.twitter.com/Ph57orSe8M#Robots working together to make a fully functional mini EV Autonomously
by @Ambots3D #MI #Robotics #RPA #Innovation #FutureOfWork #Automation cc: @pascal_bornet @pbalakrishnarao @ravikikan -

Revolutionary 3D Printed Hand Achieves Unprecedented Realism
By
–
#3Dprinted hand with unparalleled realism
— Ronald van Loon (@Ronald_vanLoon) 10 avril 2024
via @WevolverApp
Video by DASH 3D#TechForGood #Tech #Technology #FutureOfWork #Innovation #AI
cc: @pbalakrishnarao @kuriharan @chr1sa pic.twitter.com/cwKJHupZ98#3Dprinted hand with unparalleled realism via @WevolverApp by DASH 3D #TechForGood #Tech #Technology #FutureOfWork #Innovation #AI cc: @pbalakrishnarao @kuriharan @chr1sa
-

Zipline Unveils New Aerial Delivery Drone Technology
By
–
Zipline unveiled a new Aerial Delivery Drone
— Ronald van Loon (@Ronald_vanLoon) 10 avril 2024
by @CNET#ArtificialIntelligence #AI #MachineLearning #Drone #Tech #Innovation
cc: @bigdata @yvesmulkers @kuriharan pic.twitter.com/xH7I9kXhdjZipline unveiled a new Aerial Delivery Drone by @CNET #ArtificialIntelligence #AI #MachineLearning #Drone #Tech #Innovation cc: @bigdata @yvesmulkers @kuriharan
-

Edge and Cloud Integration: Beyond the Dichotomy to Applications
By
–
“We should stop talking about the edge vs the cloud, as without #edge there is no #cloud, and without the cloud there is no edge. We should focus on the applications and its use cases and how to enable these best” Our Flavio Devidé in the Future of #AI panel @embedded_world
-
Grok 2 and 3 need 20k and 100k H100 GPUs
By
–
More info on Grok 2 and 3: The training of Grok's version 2 model required as many as 20,000 Nvidia H100 GPUs, and Musk anticipates that future iterations will demand even greater resources, with the Grok 3 model needing around 100,000 Nvidia H100 chips to train.
-
Grok 3 to require 100,000 Nvidia H100 GPUs for training
By
–
4/ Elon Musk says the next-generation Grok 3 model will require 100,000 Nvidia H100 GPUs to train.
— God of Prompt (@godofprompt) 10 avril 2024
Check this to learn about Grok 3 and Grok 2 launch in detail. https://t.co/HRPnteBsky4/ Elon Musk says the next-generation Grok 3 model will require 100,000 Nvidia H100 GPUs to train. Check this to learn about Grok 3 and Grok 2 launch in detail.