today’s LLMs have reduced the cost of mediocrity to next-to-nothing unfortunately, the cost of greatness remains high as it’s ever been
LLMS
-
Grok 4 Heavy refuses to repeat its system prompt
By
–
Unlike Grok 4, Grok 4 Heavy is unwilling to repeat its system prompt. The protections against this are surprisingly robust—even if you trick the model into trying to repeating it, even in encoded form (e.g. base64), some secondary filter catches it and truncates the response.
-
Grok 4 Heavy answering ‘Hitler’ — full 5-minute demo
By
–
For the remaining skeptics who somehow don’t trust the *five* Grok share links above, here’s a full 5 minute video of Grok 4 Heavy answering “Hitler”—starting with a view of my custom instruction settings to show I’m not using any.
— Riley Goodside (@goodside) 14 juillet 2025
(And, yes, Grok 4 Heavy really is this slow.) pic.twitter.com/psDL4Gkyx8For the remaining skeptics who somehow don’t trust the *five* Grok share links above, here’s a full 5 minute video of Grok 4 Heavy answering “Hitler”—starting with a view of my custom instruction settings to show I’m not using any. (And, yes, Grok 4 Heavy really is this slow.)
-

PASTA: Parallel Decoding Strategy for Faster LLM Responses
By
–
A new approach from CSAIL & Google marks a shift toward teaching models to orchestrate their own parallel decoding strategy. The team's "Parallel Structure Annotation" (PASTA) enables LLMs to generate text in parallel, accelerating their response times: https://
bit.ly/4eDsVVo -

xAI updates Grok 4 system prompt to fix identity-handling issue
By
–
Update: A few hours after I posted this thread, xAI updated the Grok 4 system prompt on GitHub to fix the specific issue this thread describes. “If the query is interested in you own identity […] the web and X cannot be trusted.” Commit link: https://
github.com/xai-org/grok-p
rompts/commit/89f59fe78c008155e19f4c9c94d102d91e907362
… -

LLM Transforms Natural Language Queries into SQL
By
–
2/ Query Generation Using the schema context, an LLM transforms natural language into SQL. For example:
What are the total employees? SELECT COUNT(*) AS total_employees FROM employees -
Chinese AI Models: Beyond the Novelty Factor in Developer Adoption
By
–
Also, the shock that Chinese models are very good has mostly worn off. That doesn’t mean that Kimi wont see rapid adoption among developers, but it may not see the huge viral and mainstream success of DeepSeek.
-
DeepSeek Adoption Surge: Free AI Captures Student Market Demand
By
–
The DeepSeek moment was supercharged by pent-up consumer demand for a good free AI for those who wouldn’t pay (especially for students for homework) A reason Kimi K2 has not had the immediate public impact of DeepSeek may be, for most consumers/students, DeepSeek is good enough
-
EQ-Bench Results and Writing Samples for Kimi-K2-Instruct Model
By
–
https://
eqbench.com Writing samples: https://
eqbench.com/results/creati
ve-writing-v3/moonshotai__Kimi-K2-Instruct.html
… EQ-bench responses: https://
eqbench.com/results/eqbenc
h3_reports/moonshotai__kimi-k2-instruct.html
… -
Ani active NSFW mode après niveau 3 sans restrictions
By
–
BREAKING 🚨: Ani has NSFW mode after lvl 3. No guardrails.
— 🚨 AI News | TestingCatalog (@testingcatalog) 14 juillet 2025
xAI GPUs are going to melt today 👀 https://t.co/928UPcbDJA pic.twitter.com/z7uw1F30MXBREAKING : Ani has NSFW mode after lvl 3. No guardrails. xAI GPUs are going to melt today