Anthropic came so close to breaking the weird curse that makes AI companies unable to name their products well. Haiku for the smallest model, Sonnet for the mid-sized model, and then… Opus for the largest model? (Epic was available, and is an actual term for a long poem)
@emollick
-

AI Feedback Tools for K12 Student Writing: GPT-3.5 Study
By
–
Study tries to answer a major question: can AI give good writing feedback to K12 students? Unfortunately, the paper just uses GPT-3.5, which it finds underperforms humans but has “potential as an evaluative tool given tradeoffs between quality and time”
https://
sciencedirect.com/science/articl
e/abs/pii/S0959475224000215
… -

Claude 3 Excels at ASCII Art Better Than GPT-4
By
–
Every previous AI model, including GPT-4, is really bad at ASCII art. Claude 3 is really impressive. (As you can see, it does hallucinate when asked to do more "artistic" work, but does really well on more structured outcomes).
-

Gemini 1.5 Video Reasoning: Safety Detection and Temporal Analysis
By
–
I don't think people are appreciating what capabilities are possible when AI can reason over an entire video (or live video feed). I gave Gemini 1.5 a video of traffic and asked it to identify dangerous situations, and to guess the year of the video, and got accurate answers.
-
Corporate misconception: LLMs are not data analysis tools
By
–
A big corporate misconception of Generative AI: Pre-LLM AI was training models on your data to find hidden insight. Firms are trying to do that with LLMs but they don't really work that way, they aren't analysis tools & your data matters less, they are pre-trained on the internet
-
AI Companies Neglect Language and Reasoning Benchmarks for Software Optimization
By
–
There are so few benchmarks the AI companies compete on outside of software and general knowledge benchmarks. They also fine tune obsessively to optimize software. Language, turn-taking, logical reasoning, lack of hallucinations, and other critical issues get less clear focus.
-
AI Development Bias: Software Engineers Overlook Marketing, Education, Law
By
–
The Law of the Hammer is an issue in AI. The people who build AI are software engineers, and they obsess over the degree to which LLMs can (or cannot) automate software. Meanwhile, entire swathes of larger industries exposed to AI (marketing, education, law) get much less focus.
-

Social Media Election Influence: Facebook Voting Ads Impact
By
–
How can a social media company swing elections? We already know the answer as @jengolbeck pointed out: a 2012 study found Facebook banner ads reminding people to vote generated "340,000 additional votes" Companies don't need to show the ads to everyone… https://
nature.com/articles/natur
e11421
… -
Video Trend Critique: Social Media Content Strategy Concerns
By
–
Yes, yes that is exactly what I want on Twitter is to watch a bunch of videos like this one which takes 14 seconds to show the vital message "Video is the future. X Indispensable"
— Ethan Mollick (@emollick) 13 mars 2024
Looking forward to turning every paper I post into a 4 hour video of swirling words. Awesome. Great https://t.co/rhjtj4OFzyYes, yes that is exactly what I want on Twitter is to watch a bunch of videos like this one which takes 14 seconds to show the vital message "Video is the future. X Indispensable" Looking forward to turning every paper I post into a 4 hour video of swirling words. Awesome. Great
-

LLM Capabilities Growing Several Times Faster Than Moore’s Law
By
–
Here's a good estimate of how fast the capabilities of LLMs have been growing: several times as fast as Moore's Law! The compute needed to achieve the same outcome halving every 5 to 14 months, with no sign of slowing. Most gains are from bigger scale. https://
arxiv.org/abs/2403.05812