We’re excited to join @NVIDIAGTC next month alongside #AI innovators, technologists and builders. Learn how DataRobot and @nvidia AI can accelerate your path to value with the hardware and software stack built for AI. Stop by booth #1603 to meet with our #appliedAI experts
HARDWARE
-
Nvidia Surges Beyond Relief Rally Amid GPU Shortage
By
–
Nvidia: More than a relief rally Nvidia’s having more than a relief rally, as the stock surges almost fifteen percent. GPUs are sold out, amplifying future returns, and China is almost not an issue anymore. The value of the TL20 group of stocks to consider surges with it. Mind
-

Metis AIPU: Edge AI Inference Hardware Innovation Unveiled
By
–
@ieee_isscc has wrapped up, and it was a rewarding experience for us. Thanks for all the engaging discussions about our Metis AI Processing Unit (AIPU), designed specifically for Edge inference applications.
-

NVIDIA Supercomputers Selene Eos Power Generative AI Applications
By
–
At #GTC24, hear from the architects of Selene and Eos about how they designed these supercomputers. Learn more about the next generation of #NVIDIADGX architecture to power your #generativeAI applications. Register now. https://
nvda.ws/49liuT9 -
Token Optimization and Custom AI Chips for Scaling
By
–
Tokens are a function of how optimized your tokenization system is, as well as your embeddings. Therefore, new breakthroughs can emerge that don't require an enormous amount of GPU power. Sam needs money to create his own chips and scale up more quickly. We've seen that Sora's
-

Nvidia beats earnings forecast with $24B revenue outlook
By
–
Nvidia beats, shares jumping 6% in after-hours. The revenue outlook number of $24 billion is easily above consensus for $22.2 billion, though below the buy-side whisper number. https://
thetechnologyletter.com/the-posts/nvid
ias-shares-up-slightly-as-forecast-beats
… $NVDA -
Google Gemma Launch Partner Delivers Optimized LLM Desktop Models
By
–
Announced today, we are collaborating as a launch partner with @Google in delivering Gemma, an optimized series of models that gives users the ability to develop with #LLMs using only a desktop #RTX GPU.
-

Google Gemma 2B 7B Models Optimized NVIDIA TensorRT-LLM
By
–
Google’s newly announced Gemma 2B and 7B models, optimized with NVIDIA TensorRT-LLM – allows developers the ability to optimize inference performance across NVIDIA AI platforms, from the datacenter to local PCs with RTX GPUs: https://
nvda.ws/48psSb5 -
Groq’s LPU Hardware Breakthrough Enables Near-Instantaneous LLM Response Times
By
–
Groq's recent hardware breakthroughs have been going viral on X.
— Rowan Cheung (@rowancheung) 21 février 2024
Groq (not Grok) uses LPUs instead of GPUs, allowing the chatbot to run LLMs at nearly instantaneous response times.
This unlocks a whole new world of potential AI and user experiences.pic.twitter.com/EnJnd3jEQmGroq's recent hardware breakthroughs have been going viral on X. Groq (not Grok) uses LPUs instead of GPUs, allowing the chatbot to run LLMs at nearly instantaneous response times. This unlocks a whole new world of potential AI and user experiences.
-
Nous-Hermes-2 Quantized to 4-bit for MLX Apple Silicon
By
–
I just quantized this amazing model to 4-bit, with support for the MLX platform so you can run it super fast on Apple Silicon https://
huggingface.co/mlx-community/
Nous-Hermes-2-Mistral-7B-DPO-4bit-MLX
…
