A glimpse of GPT-6 access for an hour of testing would be the best tbh
LLMS
-
xAI prepares support for Skills and Grok 4.3
By
–
xAI keeps working on SKILLs support 👀
— 🚨 AI News | TestingCatalog (@testingcatalog) 28 avril 2026
A new Skills tab is now available (but hidden), while Grok 4.3 now supports Skills creation.
With all the features being prepared and Google I/O approaching in May, I bet we should expect a Keynote from xAI quite soon. pic.twitter.com/vC8OwTbjXPxAI keeps working on SKILLs support A new Skills tab is now available (but hidden), while Grok 4.3 now supports Skills creation. With all the features being prepared and Google I/O approaching in May, I bet we should expect a Keynote from xAI quite soon.
-

OpenAI Agents Ask Better Questions Than Researchers
By
–
Sébastien Bubeck on the OpenAI Podcast: People think AI is only good at answering questions. OpenAI's internal agents are now asking questions so good that researchers are writing papers based on them.
— Chubby♨️ (@kimmonismus) 28 avril 2026
They're also finding and correcting mistakes in published work. His timeline… pic.twitter.com/637D1h2p3ySébastien Bubeck on the OpenAI Podcast: People think AI is only good at answering questions. OpenAI's internal agents are now asking questions so good that researchers are writing papers based on them. They're also finding and correcting mistakes in published work. His timeline
-
API tracking and coding models with recitation block solutions
By
–
We are working on tracking via API key, better coding models, and will track down these blocks on recitation.
-
DevTalks Ep. 3: Autonomous Coding with Cline
By
–
DevTalks Ep. 3 Autonomous coding with @cline
. Faster inference, smoother workflows, more control. April 30 | 11am PT | 2pm ET -
ICLR 2026: Five Papers on Real AI System Challenges
By
–
Research from ICLR 2026 From long-context limits to many-shot prompting and speculative prefill, our team presented 5 papers focused on real system challenges in AI.
What works, what doesn’t, and where things break. https://
arxiv.org/abs/2510.04618 https://
arxiv.org/abs/2602.16069 -

SambaNova Intel Partnership Advances Coding Agents Performance
By
–
April, you were great to us! Welcome to another edition of our Lightning Digest TLDR: SambaNova + Intel DevTalks with Cline New research at ICLR 2026 Upcoming events https://
sambanova.ai/sambanova-ligh
tning-digest-premium-inference-for-coding-agents?&utm_source=x&utm_medium=organic&utm_content=newsletter-research
… -

BigLaw Bench: Evaluating AI Legal Research Capabilities
By
–
We recently partnered with @harvey on BigLaw Bench: Research, evaluating how well models handle real end-to-end case law research, identifying the right authorities, applying legal reasoning, and delivering answers you can actually use. It also shows where things still break and
-

Skill Retrieval Augmentation for Agentic AI Systems
By
–
// Skill Retrieval Augmentation for Agentic AI // Great read for AI devs. (bookmark it) It's on finding efficient ways to incorporate skills for agents. The work introduces Skill Retrieval Augmentation (SRA) and SRA-Bench: 26,262 skills, 636 gold skills, 5,400
-
Autonomous Cars Don’t Rely on Large Language Models
By
–
For instance, autonomous cars don't hallucinate. But they don't drive on LLMs.