Congrats to @OpenAI on the successful launch of their GPT-4.1 model 🙂 Great to see their team utilize the @scale_AI MultiChallenge benchmark to measure multi-turn instruction following
GENERATIVE AI
-
GPT-4.1 Mini and Nano Models Released for Speed and Cost
By
–
GPT‑4.1 mini and nano are also available—optimized for speed and cost. Whether you’re shipping new AI features or vibe coding in your IDE, we’d love to hear how GPT‑4.1 is working for you! More details on the blog:
-
Whisper’s Hallucinations: Understanding ASR Limitations
By
–
We dive into the well defined task of automatic speech recognition (ASR), and describe why OpenAI’s Whisper, which has been integrated into ChatGPT, makes stuff up, or “hallucinates” as it’s called in the industry (bad nomenclature).
-
One-Model-Everything Approach Reduces ASR System Reliability
By
–
Their (and Muskrat and others’) quest to build one-model-for-everything has resulted in less reliable systems than what we’ve had before, even in the well defined task of ASR. As in, historically “hallucinations” weren’t problems in ASR systems!
-
DODGE Plan Risks Replacing Federal Experts With AI Tools
By
–
Imagine what will happen if DODGE replaces federal workers with these tools to perform all the tasks that expert federal workers do. We write “There is no “one weird trick” that removes experts & creates miracle machines that can do everything that humans can do, but better."
-
Why AI Speech Recognition Got Worse Despite Hype
By
–
In Scientific American, @asmelashteka & I ask: "In the age of unbridled AI hype, with the likes of Elon Musk claiming to build a “maximally truth-seeking AI,” how did we come to have less reliable speech recognition systems than we did before?"
-

GPT-4.1 Launch: New Model for Developer Excellence
By
–
Welcome to the GPT-4.1 family! Trained with one goal: being truly great for developers. • Best-in-class at coding tasks and instruction following
• 1M token context + better comprehension
• Pushing performance across the latency curve What are your first impressions so far? -

SambaNova Llama4 Maverick Fastest LLM Cloud Performance
By
–
Find someone who looks at you like how devs look at us Experience the FASTEST #Llama4 Maverick on SambaNova Cloud
-
Advanced AI Prompts Overcome Basic Tactics Successfully
By
–
good tactic for normies but i have hit it with some big prompts and it cooked
-

Google Cloud Next: Major Brands Scale Content with Generative AI
By
–
My recap of #GoogleCloudNext that I wrote last week is finally live, featuring one of the themes: how Google is getting major brands like @LOrealGroupe and @MDLZ to create, iterate and scale content using generative AI. One example: The legendary creative agency @GSP created a
