I had a great conversation with Kara Swisher and Betsey Stevenson about #AI and the future of jobs. #RiseoftheRobots
GENERATIVE AI
-
GPT-5 Outperforms Sonnet for Idea Brainstorming Tasks
By
–
ChatGPT is the way to go! Specially for bouncing ideas – gpt-5 is way more steerable than sonnet
-
1B Parameter Model: 128K Context, Int4 Quantization, Llama 4 Distilled
By
–
chat is this real??? 128K context, int4 quantisation, 1B params, distilled from Llama 4
-
Rubric-Driven Evaluation Improves LLM Alignment and Data Quality
By
–
Rubric-driven evaluation doesn’t just improve human and LLMaJ alignment. It accelerates delivery, reduces rework, and improves the thing we care about most: data quality.
-
Trusted Scale: Snorkel’s Framework for Scaling Trust in AI
By
–
This is Trusted Scale — Snorkel’s framework for designing, validating, and scaling trust in AI data pipelines. Read the full post by @pham_derek →
https://
snorkel.ai/blog/scaling-t
rust-rubrics-in-snorkels-quality-process/
… #AI #LLM #DataQuality #SnorkelAI #MachineLearning -
Rubrics: Framework for Structured LLM Evaluation and Dataset Quality
By
–
Rubrics are the framework that make this possible. They connect human insight, LLM evaluation, and measurable outcomes — turning expert judgment into structured, repeatable metrics that drive confidence in every dataset.
-

GPT-5 o3.1 iteration represents next leap in AI reasoning capabilities
By
–
Lead OpenAI researcher: "GPT-5, in some way, can be considered o3.1 – iteration of the same concept"
— Peter Gostev (@petergostev) 16 octobre 2025
"What I'm after right now is something next, what would be a significant jump to how we interact with models, that are more capable, think for even longer and interact with even… https://t.co/bZOtxQYXHYLead OpenAI researcher: "GPT-5, in some way, can be considered o3.1 – iteration of the same concept" "What I'm after right now is something next, what would be a significant jump to how we interact with models, that are more capable, think for even longer and interact with even
-
Anthropic Releases Agent Skills for Extended Claude Capabilities
By
–
New on the Anthropic Engineering Blog: Our tips for developers on using Agent Skills, a new way to extend Claude's capabilities with instruction folders, scripts, and resources:
-
Poe Leaderboard Now Available for Desktop and Web
By
–
To view the current rankings, go to https://
poe.com/leaderboard (currently desktop and web only). (2/2) -

Poe Launches Daily Leaderboard for AI Models and Apps
By
–
New: Poe Leaderboard We launched a leaderboard that shows the most used AI models and apps across Poe's platform and API. Rankings are updated daily based on total token usage across the 200+ text, image, video, and audio bots on Poe. (1/2)