Mixtral 8x22B Instruct is out. It significantly outperforms existing open models, and only uses 39B active parameters (making it significantly faster than 70B models during inference). 1/n
LLMS
-
Apache 2.0 Open Source AI Model Released Today
By
–
Official now, very proud of the team! Apache 2.0 and instructed versions for your pleasure, available today on la Plateforme
-

Microsoft invests in UAE AI as industry expands
By
–
Top stories in AI today: -Microsoft invests $1.5B in UAE AI firm
-AMD and Nvidia unveil new processors
-Access the best open-source LLM for free
-InstantMesh enables 3D generation in seconds
-6 new AI tools & 4 new AI jobs Read more: http://
therundown.ai/p/microsofts-u
ae-power-play
… -

MongoDB FireworksAI Llama Index AI Hackathon San Francisco
By
–
Join us this Saturday, April 20th at the @MongoDB with @FireworksAI_HQ @llama_index AI Hackathon in SF. Sigh up at https://
lnkd.in/gGcaDXaX and meet @upstageai #solarllm and full-stack LLM. -
20% of Americans flirt with chatbots: Curiosity and loneliness drive interaction
By
–
Already 20% of Americans have flirted with chatbots, according to this study [linked] Nearly half of them — 47.2% — did so out of curiosity while 23.9% said they were lonely and seeking interactions. It's only April 2024. The line on this graph only goes in one direction from
-
Anthropic Claude 3 Models Remain Unchanged Since Release
By
–
Three days ago Anthropic staff were saying that the Claude 3 models haven't changed a single byte since they were released
-
Fine-tuning risks and diminishing returns in AI model optimization
By
–
Fine tuning can help if you have a specific use-case for the model that you can optimize for, but even then it's a whole lot of expensive and risky work for something that may not help much and will likely be obsoleted by the next model release
-
Claude Model Degradation Claims: Perception vs Reality
By
–
I'm suspicious that many cases of models "degrading" are just people imagining things – see the recent Claude incident where people complained it had degraded when the model was entirely static since launch
-
AI Model Performance Degradation and Context Understanding Issues
By
–
It's making really dumb mistakes. Like, shockingly dumb. That wasn't happening before. Doesn't seem to use context as well as before.
-
AI Model Error Regarding Training Data Recitation and Bypass Method
By
–
The description in the docs is inaccurate. The error code means it detected recitation of training text — and there’s lots of false positives for this, like famous integer sequences. Anecdotally streaming the response seems to bypass it.
