Mistral has released a code-specific model! Codestral-22B seems to outperform LLaMA 3 70B while being more than 3x smaller. Very impressive!
@mattshumer_
-
HyperWriteAI Models Trained for Superior Writing Quality
By
–
@HyperWriteAI is what you’re looking for — models trained specifically to write really well
-
Golden Gate Claude Startup Idea: Building Bridges Humorously
By
–
Asking Golden Gate Claude for a startup idea — of course, it wants to build a bridge 🙂 https://t.co/uZiiht2ZZN pic.twitter.com/UOCnr8oS2C
— Matt Shumer (@mattshumer_) 23 mai 2024Asking Golden Gate Claude for a startup idea — of course, it wants to build a bridge 🙂
-
Raw Prompting Interface and Recording Options for AI Models
By
–
thanks nick! @swyx we offer both — recording definitely increases reliability, but if you want the raw prompting interface to the model, we offer it it's still very early, but feel free to try it out!
-
Model Evaluation and Reasoning-Focused Haystack Testing Improvements
By
–
I don’t always find the middle to be lost — really depends on the model/use-case. We definitely need better (more reasoning-focused) haystack tests!
-
Context Window Usability Metric for Language Models
By
–
Model providers should start publishing 'Usable Context' as a metric. Just because a model can support 10M tokens doesn't mean it can use all 10M effectively. Often, I find models >32K can only effectively use a smaller portion of their context window.
-
OpenAI’s Strategic Model Release Signals Imminent Advancement
By
–
I highly doubt OpenAI has hit a wall. Two strong acceleration signals: – ChatGPT makes the bulk of the money for OpenAI — they wouldn't put out a GPT-4-ish level model free for everyone if they didn't have a much better model coming very soon – If the exiting team members
-
Responsible Path to AGI: Beyond Black and White Thinking
By
–
Jan, Ilya, etc. are not doomers. Take a look at their previous work. They believe in the benefits of getting to AGI, but don’t want to be reckless in how we get there. Not everything is black and white.
-
OpenAI Executive Publicly States Capabilities Prioritized Over Safety
By
–
Wow. This is huge. The first time (I'm aware of) that an OpenAI exec has publicly stated that they believe OpenAI is clearly prioritizing capabilities over safety research. Massive implications, in many ways.
-
OpenAI Models: Prompt Strategies Matter Less Over Time
By
–
Most models within OpenAI's ecosystem respond well to similar prompting strategies. And over time, as models get better, adjusting prompts will matter less and less.
