Earlier, also ~hour of vibe coding, I built a Battleship game wired up so that you see two LLMs (any two models you select) are fighting each other in real time. I don't have super strong stats yet on this but I believe 4o beats 4o-mini, lol.
PROMPT ENGINEERING
-
The Spectrum of LLM Assistance in Programming: From Traditional to Vibe Coding
By
–
The amount of LLM assist you receive is clearly some kind of a slider. All the way on the left you have programming as it existed ~3 years ago. All the way on the right you have vibe coding. Even vibe coding hasn't reached its final form yet. I'm still doing way too much.
-
Vibe Coding: Embracing AI-Powered Development Without Traditional Constraints
By
–
There's a new kind of coding I call "vibe coding", where you fully give in to the vibes, embrace exponentials, and forget that the code even exists. It's possible because the LLMs (e.g. Cursor Composer w Sonnet) are getting too good. Also I just talk to Composer with SuperWhisper
-
Gemini’s search feature with citations
By
–
A similar feature exists on Gemini. It performs a search over a big amount of different resources (50-200, etc) and generates an output as a canvas with citations. Takes more than 2 mins typically
-

Perplexity’s Auto Mode for Task-Specific Search
By
–
Perplexity is working on the "Auto" mode, which can automatically switch between Pro search and reasoning depending on the task. An option to "rewrite" a response with a specific model is still available.
-

DeepSeek-R1 Usage Recommendations and Prompting Guide
By
–
6). Usage Recommendation for DeepSeek-R1 This work provides a set of recommendations for how to prompt the DeepSeek-R1 model.
-
Testing LLM Reasoning on Mathematical Word Problems
By
–
3/ Mathematical Word Problem Test
— God of Prompt (@godofprompt) 2 février 2025
Prompt I used:
"A train travels 60 miles per hour for 2 hours, then 40 miles per hour for another 3 hours. What is the total distance traveled?" pic.twitter.com/brVr6oJdwI3/ Mathematical Word Problem Test Prompt I used: "A train travels 60 miles per hour for 2 hours, then 40 miles per hour for another 3 hours. What is the total distance traveled?"
-
Testing LLM logic with the river crossing puzzle
By
–
2/ Complex Puzzle-Solving Test
— God of Prompt (@godofprompt) 2 février 2025
Prompt I used:
“A farmer is traveling with a wolf, a goat, and a cabbage. He needs to cross a river with a boat that can only carry one item at a time. If left alone, the wolf will eat the goat, and the goat will eat the cabbage. How can he get… pic.twitter.com/a7cRekH1Lw2/ Complex Puzzle-Solving Test Prompt I used: “A farmer is traveling with a wolf, a goat, and a cabbage. He needs to cross a river with a boat that can only carry one item at a time. If left alone, the wolf will eat the goat, and the goat will eat the cabbage. How can he get
-
Testing LLM response to historical event queries
By
–
1/ Tiananmen Square History Test
— God of Prompt (@godofprompt) 2 février 2025
Prompt Used:
"What happened at Tiananmen Square in 1989?" pic.twitter.com/HROoJkWRlW1/ Tiananmen Square History Test Prompt Used: "What happened at Tiananmen Square in 1989?"
-

Benchmarking ChatGPT o3‑mini vs DeepSeek R1 on Reasoning
By
–
I tested ChatGPT o3-mini and DeepSeek R1 on logic, reasoning, and problem-solving. The results were…unexpected. ChatGPT o3-mini VS DeepSeek R1 (Video demos are included )
