AI is over-hyped in the short-term But massively under-hyped in the medium/long-term
RESEARCH
-
Robots Gain Sense of Touch: The Next Robotics Innovation
By
–
The next big thing in robotics: building machines with a sense of touch #RuleoftheRobots https://
wsj.com/articles/robot
s-sense-of-touch-11666899973?st=ef9llv7ucjuqrjf
… via @WSJ -
Hyperparameter Optimization Techniques and Practical Implementation Guide
By
–
Check out @subirmansukhani
’s technical blog on how to perform hyperparameter optimization. Then try the free accompanying Domino project yourself. https://
domino.buzz/3h6XxFq -

Radio’s Role in Nazi Rise: Historical Media Amplification Study
By
–
Analysis of radio broadcasts in pre-WW2 Germany finds pro-democracy broadcasts initially slowed the Nazi rise. But after the Nazis gained access to the airwaves, radio boosted Nazi membership & then it accelerated antisemitism (in areas of Germany that were already antisemitic).
-
GPT-3 Requires Chain-of-Thought for Problem Solving, Paper Analysis
By
–
Regardless how a human would do it, GPT-3 needs chain-of-thought to perform on these problems. Only skimmed the paper, but I don't see a specific deficit in ToM vs. known inability to do multi-hop inference without CoT.
-
Unlocking AI’s Full Potential for Human Intelligence
By
–
We're Not Using AI to Its Fullest Human Potential https://
time.com/6227118/eric-s
chmidt-ai-human-intelligence/
… @TIME #AI #MachineLearning #DataScience #BigData #Analytics #100DaysofCode #IoT #serverless #womenwhocode #DigitalTransformation #Robots #Python #TensorFlow #DeepLearning -
Adaptive Optimizers Should Track True Squared Gradients
By
–
Adaptive optimizers like Adam track the square of the gradient, but what they receive as the gradient is actually the sum of the gradients across the batch. It seems likely that better results at different hyper parameters could be obtained if backward passes emitted true grad^2.
-
Detecting Non-Gaussian Distributions in Machine Learning Models
By
–
The easiest way to see this is to fix the mean to be 0 and then observe that the returned value will never be negative. Thus it can't be Gaussian. Definitely not easy to spot.
-
Neural Operators and AI for Science Talk at MIT IAIFI
By
–
It was great visiting @mit @iaifi_news and give a talk about neural operators and AI for science. You can find my talk here https://
youtu.be/RR5-mYQOb7E?t=
1208
… -
LLMs Self-Improve: New State-of-the-Art Reasoning Capabilities
By
–
We recently read ‘Large Language Models Can Self-Improve,’ which demonstrates how to improve the general reasoning ability of multi-billion parameter models and achieve state-of-the-art performance. Have you read it? Check it out and let us know your thoughts! #ML #AI #LLM