Google’s newly announced Gemma 2B and 7B models, optimized with NVIDIA TensorRT-LLM – allows developers the ability to optimize inference performance across NVIDIA AI platforms, from the datacenter to local PCs with RTX GPUs: https://
nvda.ws/48psSb5
AI HARDWARE
-

Google Gemma 2B 7B Models Optimized NVIDIA TensorRT-LLM
By
–
-
Live Twitch discussion on AI models, bias, and Groq hardware
By
–
LIVE TWITCH
– ChatGPT perd la tête. – Est ce Google Gemini est raciste ?… Le problème de la correction de biais.
-Nouveau modèle Gemma
– Groq qu'est ce truc dont tout les techos parlent ? L'actu est dense !! On parle de tout ca maintenant sur Twitch !! -
AI Funding Boom: Moonshot Raises $1B, Adobe and YouTube Lead Tech
By
–
Les sources de l'édition du jour du #FlashTweet sont ici https://
scmp.com/tech/big-tech/
article/3252574/chinese-start-moonshot-ai-raises-us1-billion-funding-round-led-alibaba-and-vc-hongshan-amid-strong?utm_
… https://
usine-digitale.fr/article/global
foundries-va-beneficier-d-une-aide-de-2-1-milliards-de-dollars-pour-la-production-de-semi-conducteurs.N2208484?utm_
… https://
theverge.com/2024/2/20/2407
7217/adobe-acrobat-generative-ai-assistant-chatbot-pdf-document
… https://
techcrunch.com/2024/02/20/you
tube-dominates-tv-streaming-in-u-s-per-nielsens-latest-report/?utm_
… https://
blogdumoderateur.com/sites-e-commer
ce-plus-visites-france-fevrier-2024/?utm_
… Visuel : @scmpnews -
Groq’s LPU Hardware Breakthrough Enables Near-Instantaneous LLM Response Times
By
–
Groq's recent hardware breakthroughs have been going viral on X.
— Rowan Cheung (@rowancheung) 21 février 2024
Groq (not Grok) uses LPUs instead of GPUs, allowing the chatbot to run LLMs at nearly instantaneous response times.
This unlocks a whole new world of potential AI and user experiences.pic.twitter.com/EnJnd3jEQmGroq's recent hardware breakthroughs have been going viral on X. Groq (not Grok) uses LPUs instead of GPUs, allowing the chatbot to run LLMs at nearly instantaneous response times. This unlocks a whole new world of potential AI and user experiences.
-
Nous-Hermes-2 Quantized to 4-bit for MLX Apple Silicon
By
–
I just quantized this amazing model to 4-bit, with support for the MLX platform so you can run it super fast on Apple Silicon https://
huggingface.co/mlx-community/
Nous-Hermes-2-Mistral-7B-DPO-4bit-MLX
… -
Canadian AI Chipmaker Untether Disrupts Nvidia Market
By
–
@BNNBloomberg
: Is it too much to say that you're a disruptor to #Nvidia? On the contrary, says Chris Walker, #disruption is exactly what @UntetherAI was created for. Learn more on his recent #TradingDay chat with Amber Kanwar. https://
bnnbloomberg.ca/technology/vid
eo/this-canadian-ai-chipmaker-wants-to-make-it-big2868796
… -

AI’s Growing Environmental Costs: Energy Crisis and Water Depletion
By
–
My latest for @Nature
: AI's environmental costs are soaring. The new energy-hungry models for video, text, and image could create an energy crisis – and impact drinking water reserves. We urgently need action from industry, researchers, and legislators. https://
nature.com/articles/d4158
6-024-00478-x
… -

NSF NAIRR Pilot Opens Cerebras AI Resources for Researchers
By
–
Attention AI Researchers: Apply now to access Cerebras AI resources as a part of the NSF NAIRR pilot! The National Artificial Intelligence Research Resource (NAIRR) Pilot is currently offering an opportunity for researchers to leverage cutting-edge computational resources,
-

Groq’s Vertical Integration Strategy in Token-as-Service Pricing
By
–
"We're very comfortable at our current pricing for Token as a Service. Very." — @JonathanRoss321
, re: Groq pricing. Model serving continues to be race to the bottom. The winners are going to be those who are fully vertically integrated – including hardware. #Groqspeed -

Groq AI processor leads breakthrough week in artificial intelligence
By
–
Top stories in AI today: -Blazing-fast Groq AI processor goes viral
-Telecomm giant to reveal futuristic AI phone
-How to utilize Gemini across Google apps
-AI chip uses light waves for processing
-6 new AI tools & 4 new AI jobs Read more: http://
therundown.ai/p/groq-creates
-fastest-ai
…