pre-training data isn’t a commodity; we need like unseen books on topology
LLMS
-
Ilya Sutskever: Peak Data Reached, End of Scaling Era
By
–
https://t.co/D8MAjuh2h7 pic.twitter.com/o0MB17HvDM
— Alexandr Wang (@alexandr_wang) 14 décembre 2024Ilya Sutskever says that data is the fossil fuel of AI and "we have achieved peak data" because "we have but one internet", so the age of pre-training AI on bigger and bigger compute clusters "will unquestionably end"
-
Local LLM Support Added to Game Eliminating Cloud Costs
By
–
This will be possible eventually, that’s why we added support for local LLM – so that the game doesn’t need cloud LLM and monthly payments.
-

Jailbreaking frontier AI models still works in 2025
By
–
A year ago I thought jailbreaks would be fixed by now, or at least chased into newer modalities, but in 2025 you’ll still be able to jailbreak frontier models with a script that spams variations of tLAkIng Liek tHis
-

Free ChatGPT System Prompt to Learn Prompt Engineering
By
–
Prompt engineering is the most in-demand skill right now. But most people don't know how to write prompts. That's why I created this "ChatGPT System Prompt" to help you learn how to write the best prompts. It's free for 24 hours. Like + comment "Send" and I'll DM you the
-

LLM Engineer’s Handbook: Master Large Language Models in Production
By
–
LLM Engineer's Handbook — Master the art of engineering Large Language Models #LLMs from concept to production: http://
amzn.to/4dUQrv6 v/ @PacktPublishing ——
#DataScience #DataScientist #ML #GenAI #AI #GenerativeAI #MachineLearning ——
What you will learn: Implement robust -

Machine Learning Solutions Architect Handbook: ML Lifecycle and MLOps
By
–
#MachineLearning Solutions Architect Handbook — Practical Strategies and Best Practices in the #ML Lifecycle, System Design, #MLOps, and Generative AI: http://
amzn.to/4bx8t6b v/ @PacktPublishing ——————
#DataScience #DataScientist #AI #GenAI #GenerativeAI #LLMs #LLMOps -

Compute and Data: Post-Training Data as the Next Frontier
By
–
compute matters, but so does data. pre-training data wall -> post-training data boom @scale_AI is the data foundry that will accelerate us out of this rut
-
Seq2seq Architecture: A Decade of Evolution in Deep Learning
By
–
seq2seq 10 years later. 💙@ilyasut @quocleix. Talk 👇 https://t.co/fqZI0SXdN0 pic.twitter.com/Q5le02u1fE
— Oriol Vinyals (@OriolVinyalsML) 14 décembre 2024seq2seq 10 years later. @ilyasut @quocleix
. Talk -

Phi-4 Synthetic Data Generation Principles Released
By
–
Phi-4's principles for generating synthetic data remind me of something… It's a cool paper, I'm glad they released more stuff this time.