i think you will be very very happy with o3 pro!
LLMS
-
Claude Plus Tier Gets 100 Daily o3-mini Queries and Operator Access
By
–
ok we heard y’all. *plus tier will get 100 o3-mini queries per DAY (!)
*we will bring operator to plus tier as soon as we can
*our next agent will launch with availability in the plus tier enjoy -
Apollo Custom Backends Now Support Reasoning Tokens
By
–
Custom backends now support reasoning tokens in the latest version of Apollo. Just change the model type in settings.
-
DeepSeek R1 7B: Powerful Reasoning AI on Local Networks
By
–
Here’s deepseek r1 7b (qwen) beaming from my laptop to my phone at 60 tok/s.
— Aaron Ng (@localghost) 25 janvier 2025
Wild that you can serve such powerful reasoning AI models to your whole network with just a few apps. pic.twitter.com/RTHnrjmzRhHere’s deepseek r1 7b (qwen) beaming from my laptop to my phone at 60 tok/s. Wild that you can serve such powerful reasoning AI models to your whole network with just a few apps.
-
RL Model Achieves Superhuman Snark on Human Cope Chains
By
–
applying large-scale RL to chains of cope written by human seethers, we observe an “lmao moment” upon which the model spontaneously exhibits superhuman snark
-
Future Mobile AI Models: Shift Toward Specialized Smaller Systems
By
–
near term future for mobile models might be many small specialized small ones
-
LLMs Training Data Dependency: Stack Overflow Risk
By
–
@GergelyOrosz Given the LLMs' dependence on training data from Stack Overflow (and other Q&A forums), what happens when SO dies and that information source dies with it? How do LLMs replenish their sources?
-
Sakana AI Proposes Transformer² with Two-Stage Reasoning
By
–
東京発Sakana AIが新たな「Transformer」を提案 https://
xtech.nikkei.com/atcl/nxt/colum
n/18/02801/012200014/
…
Transformer²は2段階の推論プロセスを採用している。
1段階目では、プロンプトを試験的に実行してAIモデルの挙動を観察し、実行に必要なスキルを理解する。 -
DeepSeek AI breakthrough: genuine innovation or geopolitical psyop?
By
–
What's more likely? 1 – small group of AI engineers at @deepseek_ai figures out how to beat all of the top researchers in the world as a side project 2 – Chinese government has 100k GPUs they shouldn't have and releases open source models claiming $6m training cost as a psyop