ICYMI: Our newsletter is bringing the latest and greatest enterprise AI news to your inbox In the most recent edition, we break down our latest advancements in AI search and retrieval, highlighting how Embed 4, our most powerful embedding model yet, is revolutionizing the way
GENERATIVE AI
-
Portable ChatGPT Memory Profiles via JSON
By
–
Steal my ChatGPT prompt to generate your context profile from its memory – a portable JSON profile you can paste into any chat for personalized responses —————————
JSON PROFILE CREATOR
————————— You are an elite behavioral analyst with -
Gemini 2.0 Flash stability updates and model changes
By
–
2.0 flash image generation is not a stable model, so we make changes. But just “gemini-2.0-flash” is stable
-
Making AI Models More Token Efficient
By
–
We are working to make them much more token efficient, stay tuned 🙂
-
New pricing model for reasoning-enabled AI language models
By
–
As mentioned at the bottom, we chose to have a “thinking off” option with a much lower price specifically so that devs moving from 2.0 had a more clear migration path. But with reasoning is definitely more expensive, very different class of models.
-
Flash 2.0 Model Reaches Stable GA Release Status
By
–
what changed on 2.0 flash? the model is stable and GA
-
Qwen3 32B Delivers Lightning-Fast Performance on SambaNova Cloud
By
–
Speed Alert! @Alibaba_Qwen's Qwen3 32B is blazing fast on SambaNova Cloud! 🔥
— SambaNova (@SambaNovaAI) 8 mai 2025
🗣️ Juggle 100+ languages faster than you can say "Bonjour, 안녕하세요, 你好!"
🧠 Tackle complex agent apps like a caffeinated supercomputer.
Don’t miss out ➡️ https://t.co/zm6RCXXsaPSpeed Alert! @Alibaba_Qwen
's Qwen3 32B is blazing fast on SambaNova Cloud! Juggle 100+ languages faster than you can say "Bonjour, 안녕하세요, 你好!" Tackle complex agent apps like a caffeinated supercomputer. Don’t miss out http://
cloud.sambanova.ai -
Google Ships Implicit Caching for Gemini API Cost Savings
By
–
We just shipped implicit caching in the Gemini API, automatically enabling a 75% cost savings with the Gemini 2.5 models when your request hits a cache We also lowered the min token required to hit caches to 1K on 2.5 Flash and 2K on 2.5 Pro!
-
Microsoft introduces Pages for Copilot editing
By
–
Microsoft released Pages for Copilot 🔥
— 🚨 AI News | TestingCatalog (@testingcatalog) 8 mai 2025
"Meet Pages, your new editing BFF! No need to start from scratch – just hit "Edit this Response," highlight text, and tap "Ask" to tweak, expand, or polish your writing." pic.twitter.com/AZBiMdUsBnMicrosoft released Pages for Copilot "Meet Pages, your new editing BFF! No need to start from scratch – just hit "Edit this Response," highlight text, and tap "Ask" to tweak, expand, or polish your writing."
