Write long papers, so AI can shorten them in different ways for different people.
@pmddomingos
-
Scaling AI from 80s to 2000s: Compute Not Enough
By
–
Danny Hillis was scaling up AI with a massively parallel supercomputer in the 80s. In the 90s we had the data mining explosion, a.k.a. scaling up ML. In the 2000s we had the "big data" boom. And each time we noticed that no, compute etc. is not enough – you really need better
-
Professor says transformers paper would fail his student
By
–
If the transformers paper was written by one of my students, I wouldn’t let him graduate until he did a better job.
-
Deep Learning Papers Confusing Because Researchers Are Confused
By
–
Deep learning papers are confusing because deep learning researchers are confused.
-
AI enables cheating in education yet devalues skills for employers
By
–
AI lets you cheat your way to graduation, but then companies don't need your skills.
-
AI races: top contenders in models, data centers, chips
By
–
The three AI races and their top contenders:
Models: OpenAI, Anthropic, Google
Data centers: Amazon, Microsoft, Google
Chips: Nvidia, AMD, Google -
Google researchers: can’t imagine publishing transformer paper now, return to academia?
By
–
Google researchers say they can’t imagine being allowed to publish the transformer paper now. Time to return to academia?
-
Yoshua Bengio’s Large-Scale Graduate Student Descent Breakthrough
By
–
The key AI breakthrough of the last 20 years, due to Yoshua Bengio, was large-scale graduate student descent.
-
LLMs Memorization vs Overfitting Clarification
By
–
Memorization does not imply overfitting. Overfitting is strictly about what happens on non-training data. So, e.g., just because LLMs memorize data doesn’t make them stochastic parrots. What matters is what they do with it, and they typically paraphrase it quite appropriately.
-
Memorization vs Overfitting in Large Language Models
By
–
Memorization does not imply overfitting. Overfitting is strictly about what happens on non-training data. So, e.g., just because LLMs memorize data doesn’t make them stochastic parrots. What matters is what they do with it, and they typically paraphrase it quite appropriately.