Look whos back on X. And look how he ignores o1 and CoT.
LLMS
-
LLMs Superior as Doctors, Tutors, and Life Coaches
By
–
LLMs are already better and more patient doctors, better tutors for school and university and life coaches for all matters. Sky is the limit.
-
Sutskever’s AI Scaling Predictions Prove Prescient Once Again
By
–
I don't wanna say "I told you so", but I told you so. Quote: "Ilya Sutskever, co-founder of AI labs Safe Superintelligence (SSI) and OpenAI, told Reuters recently that results from scaling up pre-training – the phase of training an AI model that uses a vast amount of unlabeled
-
LLMs Embodied Decision Making Capabilities Evaluation Study
By
–
I’m really excited by this collaborative work on evaluating the current LLMs capabilities for embodied decision making, publishing at #neurips2024 . Scroll down to see the fascinating key findings. 🦾🦿 https://t.co/xdAjMhtxlb
— Fei-Fei Li (@drfeifei) 13 novembre 2024I’m really excited by this collaborative work on evaluating the current LLMs capabilities for embodied decision making, publishing at #neurips2024 . Scroll down to see the fascinating key findings.
-

DeepSeek JanusFlow 1.3B Unified Multimodal LLM Released
By
–
DeepSeek is back! JanusFlow 1.3B – Unified multimodal LLM > Key Finding: Rectified flow can be trained within the large language model framework without complex modifications. > Base Model: Built on DeepSeek-LLM-1.3b-base. > Vision Encoder: SigLIP-L, supports 384 x 384
-
Ternary LLM and BERT Training: Starting with TernaryBERT Paper
By
–
I have never trained a ternary LLM or BERT language model. Looks like there are only few publicly available repos out there. I would maybe start with the "TernaryBERT: Distillation-aware Ultra-low Bit BERT" paper and go from there
-

Gen AI in Google Search Fails Simple Soccer Geometry Question
By
–
Gen AI in Google Search is unbelievably bad. Here is a simple question "What is the max angle from the center you can shoot in soccer from penalty kick position and still make a goal (assume ball travels in straight line)" Claude, Perplexity, and ChatGPT – all get the correct
-

Inference Time Scaling Advancement Surpasses o1 Preview
By
–
1) what? An advancement in inference time scaling that can be applied to any model? And that even excels o1 preview already?
-
Supermaven advances long-context understanding research capabilities
By
–
Supermaven brings research expertise in long-context understanding at a scale that even the big labs struggle with. (3/5)
