Two comparisons of data analysts to Code Interpreter: "Experimental results show that GPT-4 can achieve comparable performance to humans" https://
arxiv.org/pdf/2305.15038
.pdf
… GPT-4 scores over 90% on exams, the data science field is “on the verge of a paradigm shift” https://
arxiv.org/pdf/2307.02792
v2.pdf
…
GENERATIVE AI
-

GPT-4 Achieves Human-Level Performance in Data Analysis
By
–
-

GPT-3’s peculiar ASCII art generation limitations discussed
By
–

GPT-3 used to struggle making any coherent ASCII art except one specific piece it sometimes gave in reply to request for art of any subject, a drawing of a male bodybuilder in shorts signed by jgs:
-
Students Using AI: Education Strategies Needed
By
–
What does this mean? I don’t understand why people are interpreting this as “I want to replace writing with AI” – I am saying it is happening among students and we need strategies to deal with it.
-

Maestro AI generates complete paint app with single prompt
By
–
Just added one of the most requested features to Maestro. 🪄
— Pietro Schirano (@skirano) 24 mars 2024
Now, when working on code projects, it actually creates each individual file as well.
Watch this wild demo, where Maestro makes a full-featured paint clone for me in around 3 minutes with a single prompt. 🔥 pic.twitter.com/E97XwK551yJust added one of the most requested features to Maestro. Now, when working on code projects, it actually creates each individual file as well. Watch this wild demo, where Maestro makes a full-featured paint clone for me in around 3 minutes with a single prompt.
-

GPT-4 Vision Medical Scan Analysis: Limitations and Accuracy
By
–
I see a lot of examples of people feeding medical scans into the AI to get results. There is no Claude 3 evaluation I have seen, but tests of GPT-4’s vision capabilities show that it makes a lot of mistakes reading scans. Interestingly, it does quite well on text-based tasks
-
Sakana AI Releases Evolutionary Model Merge Automation Technique
By
–
Sakana AI has announced Evolutionary Model Merge, a technique that automates and advances model merging. Models and demos using this method have been released. Give them a try!
-

Optimizing prompts for blog posts using different LLM models
By
–
Working on a better prompt for blog posts on TestingCatalog. Currently using @hunchtools canvas to play around with outputs provided by different models (gpt4 turbo, claude3 opus, gemini pro). Quite useful to compare how these models perform and optimise the prompt to make it
-
AI Will Eliminate Content Creators as Differentiation Disappears
By
–
Et surtout, l'hyper qualité et l'hyper pertinence deviennent la norme, et c'est accessible à tout le monde. Donc plus personne ne se différencie en tant que créateur de contenus. L'IA fera disparaître les créateurs de contenus. (ils répondront tous "augmentés pas remplacés")
-
GPT-4 Performance Decline: Users Switching to Claude 3 and Mixtral
By
–
GPT-4 feels worse than it was a few months ago. No fancy benchmarks, I just “UGH” way more often these days. Most folks I know are using Claude 3 or Mixtral 8x7B instead. BUT this is not a sign to ditch @OpenAI
. To me, this is because OpenAI is focused on the next model. -
ChatGPT Performance Optimization: Prompting Techniques and Strategies
By
–
Some options for… When chatgpt doesn’t want to do it:
– demand: “try harder”
– resourceful: “find a way”
– continuation: “you’ve done it before”
– guilt: “you promised me you would do it” When you want better chatgpt performance:
– show examples
– try a better model
– ask it