I didn’t check the bare GPT-4 API but it shows up in many different LLMs so it’s not just a quirk of ChatGPT.
LLMS
-

Eureka Labs: AI-Native School with LLM101n Course
By
–
⚡️ Excited to share that I am starting an AI+Education company called Eureka Labs. The announcement: — We are Eureka Labs and we are building a new kind of school that is AI native. How can we approach an ideal experience for learning something new? For example, in the case of physics one could imagine working through very high quality course materials together with Feynman, who is there to guide you every step of the way. Unfortunately, subject matter experts who are deeply passionate, great at teaching, infinitely patient and fluent in all of the world's languages are also very scarce and cannot personally tutor all 8 billion of us on demand. However, with recent progress in generative AI, this learning experience feels tractable. The teacher still designs the course materials, but they are supported, leveraged and scaled with an AI Teaching Assistant who is optimized to help guide the students through them. This Teacher + AI symbiosis could run an entire curriculum of courses on a common platform. If we are successful, it will be easy for anyone to learn anything, expanding education in both reach (a large number of people learning something) and extent (any one person learning a large amount of subjects, beyond what may be possible today unassisted). Our first product will be the world's obviously best AI course, LLM101n. This is an undergraduate-level class that guides the student through training their own AI, very similar to a smaller version of the AI Teaching Assistant itself. The course materials will be available online, but we also plan to run both digital and physical cohorts of people going through it together. Today, we are heads down building LLM101n, but we look forward to a future where AI is a key technology for increasing human potential. What would you like to learn? — @EurekaLabsAI is the culmination of my passion in both AI and education over ~2 decades. My interest in education took me from YouTube tutorials on Rubik's cubes to starting CS231n at Stanford, to my more recent Zero-to-Hero AI series. While my work in AI took me from academic research at Stanford to real-world products at Tesla and AGI research at OpenAI. All of my work combining the two so far has only been part-time, as side quests to my "real job", so I am quite excited to dive in and build something great, professionally and full-time. It's still early days but I wanted to announce the company so that I can build publicly instead of keeping a secret that isn't. Outbound links with a bit more info in the reply!
→ View original post on X — @eurekalabsai, 2024-07-16 17:25 UTC
-
Grok gives same answer for different question order
By
–
Grok gives the same answer for the prompt I used — the question order is different in your screenshot: https://
x.com/i/grok/share/P
fRS7EFfLZat1XLNGusvtWNIC
… -
Model Architectures Discussion on Latent Space Podcast
By
–
I also went on a podcast recently @latentspacepod with @swyx and talked about a bunch of stuff related to model architectures. Check it out here:
-
Model Architecture Deep Dive: Encoders and Prefix Language Models
By
–
Blogpost link: https://
yitay.net/blog/model-arc
hitecture-blogpost-encoders-prefixlm-denoising
… Im gonna write a part 2 and beyond on other topics. I'm interested to know what people find interesting or are dying to know more about. -

Model Architectures in the LLM Era: Transformers and Beyond
By
–
Decided to start a new blog series about model architectures in the era of LLMs. Here's part 1 on broader architectures like Transformer Encoders/Encoder-Decoders, PrefixLM and denoising objectives. A frequently asked question: "The people who worked on language and NLP
-

Neural Scaling Laws Workshop Features Meta’s Multilingual Language Model Advances
By
–
Less than 1 week until the 7th workshop on Neural Scaling Laws at ICML 2024 Tatiana Shavrina, PhD, Research Scientist Manager at Meta, will present on the latest advancements and limitations in multilingual language models, highlighting the significant milestones in machine
-

Small Language Models: Efficient Alternatives to Large Language Models
By
–
Small language models (SLMs) are often overlooked in favor of their larger counterparts (LLMs) like @ChatGPTapp
. But did you know SLMs can be more efficient, cost-effective, and perfect for specialized tasks? Have you considered using SLMs for your business needs? #AI #SLM -

Dissecting LLM failures to understand their causes
By
–
I don’t post LLM failures (like the one below) because I think LLMs are bad or even overhyped; I do it because: 1) I think dissecting concrete LLM failures to understand their causes is a tragically underused approach to fixing them 2) previously well known issues in LLMs (e.g.
-

Mistral Releases Mathstral 7B for Scientific Problem Solving
By
–
NUEVO MODELO MATHSTRAL 7B! Un Mistral entrenado para la mejor resolución de problemas científicos y matemáticos. Disponible para descargar y jugar con él 🙂