Top AI Papers of the Week (Nov 24 – 30): – LatentMAS
– INTELLECT-3
– OmniScientist
– Lightweight End-to-End OCR
– Evolution Strategies at Hyperscale
– Training LLMs with Reasoning Traces
– Cognitive Foundations for Reasoning in LLMs Read on for more:
MULTIMODAL AI
-
Top AI Papers of the Week: Latest Research Highlights
By
–
-

VQ-Insight: Progressive Learning for AI Video Quality Evaluation
By
–
How do we fix the messy state of AI video quality evaluation? Enter VQ-Insight: it uses progressive learning (starting with image quality, then adding temporal skills, and linking directly to generation models) plus multi-layered rewards to judge everything from frame
-
CT Nodule Segmentation: Generative AI Application in Healthcare
By
–
Not sure they're that differentiated, are they? E.g CT nodule segmentation is a GenAI task, isn't it?
-

9 AI Skills That Will Matter in 2026 Beyond ChatGPT
By
–
9 AI skills that will actually matter in 2026
(not just “use ChatGPT”) Prompt engineering RAG pipelines Agentic workflows Fine-tuning & distillation LLM evaluation & red-teaming Cost & latency optimization Tool calling + function calling Multimodal -

LabOS: AI-Powered Co-Scientist Integrating XR for Research
By
–
The world's first Co-Scientist integrating AI and XR (Extended Reality)! Meet LabOS. It uses multimodal perception, self-evolving agents, and XR tools to see what researchers see, grasp experimental context, and assist in real time. From cancer immunotherapy target discovery
-
Nano Banana Pro and Sora 2 Create Powerful Video Creation Combination
By
–
Nano Banana Pro combined with Sora 2 is a killer combination for videos! We totally recommend it
-
Gemini’s intelligent research capabilities
By
–
10. Research that actually feels intelligent
— God of Prompt (@godofprompt) 29 novembre 2025
Gemini isn’t hallucinating like old models.
It compares, verifies, synthesizes with context.
It’s like having a researcher who doesn’t get tired. pic.twitter.com/ua6c3zy9dB10. Research that actually feels intelligent Gemini isn’t hallucinating like old models. It compares, verifies, synthesizes with context. It’s like having a researcher who doesn’t get tired.
-

Gemini’s Image Model Outperforms Midjourney
By
–
1. Image Generation (Nano Banana Pro) Most people still think Gemini is “just a chatbot”.
Meanwhile the new image model makes Midjourney feel slow. You describe a scene. It gives you production-grade artwork. Here’s the image prompt I use daily: “Create a high-end studio -
Abacus AI DeepAgent Creates PowerPoints from Text Prompts Successfully
By
–
People have been using Abacus AI DeepAgent to make PowerPoints from prompts. The results have been solid.
-

PH-Reg: Post-hoc Method Fixes Vision Transformer Artifacts
By
–
Ever wondered how to fix the weird artifact tokens that mess up Vision Transformers’ fine-grained tasks—without retraining those massive models from scratch? HKU, Zhejiang & NTU introduce PH-Reg, a post-hoc method that adds register tokens to pre-trained ViTs via