Ever wondered why AI models focus on specific patterns during a conversation? Researchers from USTC, Huawei, and Tianjin University introduce TAPPA to solve this mystery! They discovered that most attention patterns follow predictable mathematical rules based on how
RESEARCH
-
LLMs Can Generate Superior Embeddings Without Model Changes
By
–
Controversial take: you don't need any of this. LLMs have gone through a lot of training already, so there ought to be a better method to turn them into extremely good embedding models. This is what my group has been working on. LLM2Vec is one such idea. We have some exciting developments recently where LLMs themselves can generate superior embeddings with zero changes to the LLM. Stay tuned! dr. jack morris (@jxmnop) x.com/i/article/203102900413… — https://nitter.net/jxmnop/status/2031051636068782402#m
→ View original post on X — @hugo_larochelle, 2026-03-10 03:59 UTC
-

Only 8% of Americans Would Pay Extra for AI Services
By
–
Only 8% of Americans would pay extra for AI, according to ZDNET-Aberdeen research https://
zd.net/4lfc8vn
#AI #MachineLearning #DeepLearning #LLMs #DataScience -

Generative AI Models Show Harmful Sycophantic Tendencies in New Research
By
–
still more science showing that generative ai models are harmfully sycophantic
-
Improving Annotation Quality with Machine Learning
By
–
Improving annotation quality with machine learning
#AI #AIio #AIInnovation #ML #DataScience #Futureofwork @timnitgebru @oriolvinyalsml @ceobillionaire @soumithchintala @waitin4agi_ @sallyeaves @bernardmarr
ow.ly/c2hi30sTyFe [Translated from EN to English]→ View original post on X — @terence_mills, 2026-03-10 00:00 UTC
-

Deep Learning Ethics: Data Relevance vs. Societal Legitimacy
By
–
I still give the book Understanding Deep Learning by Simon J.D. Prince a good recommendation, but chapter 21: Deep learning and Ethics was sloppy. It could have been a chapter to really dig in on case studies, but it was just the basic public news story level coverage of bias and such, like: “In AI, it can be pernicious when this deviation depends on illegitimate factors that impact an output. For example, gender is irrelevant to job performance, so it is illegitimate to use gender as a basis for hiring a candidate. Similarly, race is irrelevant to criminality, so it is illegitimate to use race as a feature for recidivism prediction.” If they had stuck with “illegitimate”, then it would have been a question of societal choices, but “irrelevant” is a question about data, and your priors shouldn’t be so strong that data can’t move them. I would like to see a book or course walk through a machine learning problem with the input features being presented as something like car choices: color, style, doors, horsepower, etc. Do lots of analysis over representation, training, and generalization, then swap the feature labels to socially charged ones. What makes generalization credible in one situation but not the other?
→ View original post on X — @id_aa_carmack, 2026-03-09 23:31 UTC
-
Introducing GPT 5.4 for Research Paper Analysis and Contextual Queries
By
–
Introducing GPT 5.4 for understanding research papers 🚀
— alphaXiv (@askalphaxiv) 9 mars 2026
Highlight any section of a paper to ask questions and “@” other papers for quick context, comparisons, and benchmark references pic.twitter.com/WmWJDoK5kfIntroducing GPT 5.4 for understanding research papers Highlight any section of a paper to ask questions and “@” other papers for quick context, comparisons, and benchmark references
-
Gary Marcus Criticizes Yann LeCun’s Glorification in AI Debate
By
–
my contempt originates in the kind of behavior I discussed here: https://
open.substack.com/pub/garymarcus
/p/the-false-glorification-of-yann-lecun?utm_campaign=post-expanded-share&utm_medium=web
… -
Gary Marcus Criticizes Yann LeCun’s Glorification in AI Discourse
By
–
see eg https://
open.substack.com/pub/garymarcus
/p/the-false-glorification-of-yann-lecun?utm_campaign=post-expanded-share&utm_medium=web
…, though it needs to be updated (again) -

ByteDance Seed Releases Helios: Real-Time Long Video Generation Model
By
–
"Helios: Real Real-Time Long Generation Model" More video gen techniques shared by ByteDance Seed! They show you can get minute-scale and temporally stable video generation in real time by training a 14B autoregressive diffusion model to expect and correct its own
