4/ A Survey of LLMs – a new 50 pages survey on large language models.
GENERATIVE AI
-
Segment Anything Model releases billion mask segmentation dataset
By
–
1/ Segment Anything Model – a set of resources for image segmentation; releases the largest segmentation dataset with over 1B masks on 11M licensed images; the model’s zero-shot performance is competitive with or superior to fully supervised results.https://t.co/sJvKhmeECe
— DAIR.AI (@dair_ai) 9 avril 20231/ Segment Anything Model – a set of resources for image segmentation; releases the largest segmentation dataset with over 1B masks on 11M licensed images; the model’s zero-shot performance is competitive with or superior to fully supervised results.
-

GPT-4 Instruction Tuning Dataset for LLaMA Models
By
–
2/ Instruction Tuning with GPT-4 – a "first attempt" to use GPT-4 to generate instruction-following data for LLM fine-tuning; includes 52K unique English & Chinese instruction-following data used to instruction-tune LLaMA models.
-
Top ML Papers Week April 3-9 Segment Anything Model
By
–
Top ML Papers of the Week (April 3 – 9): – Segment Anything Model
– SegGPT
– A Survey of LLMs
– Instruction Tuning with GPT-4
– 8 Things to Know about LLMs
– Summary of ChatGPT/GPT-4 Research
… -
ChatGPT Plugin Extracts Wisdom from Lex Fridman Podcast
By
–
A ChatGPT plugin for getting pearls of wisdom from the @lexfridman podcast: https://t.co/n2QeFs9MAO
— Greg Brockman (@gdb) 9 avril 2023A ChatGPT plugin for getting pearls of wisdom from the @lexfridman podcast:
-
GPT-4 excels as SQL JavaScript Python coding assistant
By
–
#GPT4 is the best #SQL/#JS/#Python writing assistant I've ever used
-
Solving Low-Frequency Language Understanding in AI Models
By
–
well, could be, but a lot of people told me a year ago it was *already* solved. and i wondered then/continue to wonder with low frequency/novel words, unusual phrases, etc.
-

Self-Attention Mechanism in Transformer Models for LLMs
By
–
Nice overview of self-attention mechanism used in Transformer models that underpin many Large Language Models (LLMs) & members of the GPT family.
-
Copyright commercialization restrictions for AI training data
By
–
Surface level solutions (which work today) would involve being more strict about commercialisation rights, e.g. in Copyright. If you train on web-scraped data, that's OK for research, personal use and non-profit only — but you *cannot* commercialize.
-
PEFT Democratizes State-of-the-Art LLM Research and Development
By
–
Very cool to see PEFT playing such an important role in democratising SoTA research. Checkout the mode weights at: https://
huggingface.co/andreabac3/Fau
no-Italian-LLM-7B
…
