I just quantized this amazing model to 4-bit, with support for the MLX platform so you can run it super fast on Apple Silicon https://
huggingface.co/mlx-community/
Nous-Hermes-2-Mistral-7B-DPO-4bit-MLX
…
GENERATIVE AI
-
Nous-Hermes-2 Quantized to 4-bit for MLX Apple Silicon
By
–
-
AI Output Moderation and Hidden Policy Constraints
By
–
It randomly works, but it tends to stop generating once it notices that the output is conflicting with its policy somehow (it won't disclose the actual policy).
-

Speculative Streaming: Fast LLM Inference Without Auxiliary Models
By
–
Speculative Streaming: Fast LLM Inference without Auxiliary Models Bhendawade et al.: https://
arxiv.org/abs/2402.11131 #ArtificialIntelligence #DeepLearning #MachineLearning -
Multimodal models struggle with transfer learning to text despite image knowledge.
By
–
Also weird there’s no obvious transfer learning back to text. For all the ineffable, AGI-essential knowledge supposedly in images, multimodal models seem no better at spatial reasoning word problems, creating SVGs, designing web UI, or drawing ASCII art.
-
AI Prompt to Draw Portrait of 17th Century Physicist
By
–
"please draw a portrait of a famous physicist of the 17th century"
-

AI Agents Reasoning Code at Scale GTC24
By
–
Discover the next wave of #AI on Wednesday, March 20 at #GTC24. Join @imbue_ai CEO Kanjun Qiu and NVIDIA VP of Applied Deep Learning Research Bryan Catanzaro for a discussion on AI agents that reason code at scale. https://
nvda.ws/48jEAEm -
Deploying Long-Context LLMs: Practical Challenges at GTC24
By
–
Inference with long contexts has become increasingly important for building real-world applications with #LLMs. Join experts at #GTC24 to explore the practical challenges of deploying long-context LLMs.
— NVIDIA AI (@NVIDIAAI) 20 février 2024
Register here: https://t.co/0jCrgfPHEA pic.twitter.com/C0T5WcTNVhInference with long contexts has become increasingly important for building real-world applications with #LLMs. Join experts at #GTC24 to explore the practical challenges of deploying long-context LLMs. Register here: https://
nvda.ws/3UsuFsL -
Top Tech Leaders Share AI Breakthroughs at GTC24
By
–
Join leaders from top companies such as @Adobe
, @Amazon
, @Disney
, @GettyImages
, @GoogleDeepMind
, @MercedesBenz
, @Meta
, @Microsoft
, and more sharing the latest in #AI and tech breakthroughs. See the future at #GTC24, March 18-21 in San Jose, CA. -

Gemini 1.5 Pro Model Guide Released With 1M Token Context
By
–
The Gemini 1.5 Pro model guide is live! With support of up to 1 million tokens context length, you may be wondering what's possible with Gemini 1.5 Pro. My overall impression after our first round of testing is that Gemini 1.5 Pro is among the most powerful long context LLMs
-
Gemini-1.5 Pro Represents Most Significant LLM Advancement
By
–
— Oriol Vinyals (@OriolVinyalsML) 20 février 2024
Ok, I've been waiting a while before saying it confidently, but I now understand that the Gemini-1.5 Pro had its deserved spotlight stolen last week by Sora. After experimenting with it for a while, I believe it represents the most significant advancement in LLM capabilities