thanks man. not sure if the big guys see the value of a mad prompt scientist but I think I could make waves at Meta Superintelligence for instance.
PROMPT ENGINEERING
-

Go-to Prompt for UX Product Design Ideation
By
–
My go-to prompt for UX/product design ideation Crazy that this actually improves results
-
Separate Few-Shot Examples from Main Prompt for Better Performance
By
–
My main issue with this prompt is having it build few-shot examples in the main prompt. It is cognitively already doing a lot of heavy lifting and I think any examples should be generated in a separate pass. Otherwise it's not bad.
-
Framework for Recreating Grok Heavy Functionality Across Models
By
–
Releasing this tomorrow: A framework for recreating Grok Heavy functionality.
— Pietro Schirano (@skirano) 14 juillet 2025
In this video, I'm using Claude 4 Heavy, but it works with any model (GPT, Kimi, Gemini). pic.twitter.com/v5zLxTltN3Releasing this tomorrow: A framework for recreating Grok Heavy functionality. In this video, I'm using Claude 4 Heavy, but it works with any model (GPT, Kimi, Gemini).
-
LLMs can leverage external tools like chess engines for task performance
By
–
Y en realidad sí podría porque si se acepta usar software como el que usaría la Atari para ganar consistentemente, nada impide al LLM buscar en internet, instalarse un módulo de ajedrez en su sistema y usarlo como herramienta. Es el problema de creerse los argumentos de Marcus.
-
Model Refusal Rate Analysis and Prompt Engineering Support
By
–
Nope, it cannot be training because this behavior is tracked and the refusal rate is the lowest in its size category. I can help more if you send me the prompts.
-
Reinforcement Learning Field of View Limitations and Policy Optimization
By
–
Yep I think RL is misleading in that it restricts field of view. Eg like you mentioned you can imagine review/reflect doing a lot more – building tools for later use, or actively tuning the distribution for what to try next (instead of just sampling from policy independently as
-
Grok 4 Heavy refuses to repeat its system prompt
By
–
Unlike Grok 4, Grok 4 Heavy is unwilling to repeat its system prompt. The protections against this are surprisingly robust—even if you trick the model into trying to repeating it, even in encoded form (e.g. base64), some secondary filter catches it and truncates the response.
-
Grok 4 Heavy answering ‘Hitler’ — full 5-minute demo
By
–
For the remaining skeptics who somehow don’t trust the *five* Grok share links above, here’s a full 5 minute video of Grok 4 Heavy answering “Hitler”—starting with a view of my custom instruction settings to show I’m not using any.
— Riley Goodside (@goodside) 14 juillet 2025
(And, yes, Grok 4 Heavy really is this slow.) pic.twitter.com/psDL4Gkyx8For the remaining skeptics who somehow don’t trust the *five* Grok share links above, here’s a full 5 minute video of Grok 4 Heavy answering “Hitler”—starting with a view of my custom instruction settings to show I’m not using any. (And, yes, Grok 4 Heavy really is this slow.)