Thanks for all this great info Stella. I'm not familiar with that tokenizer — what's special about it?
LLMS
-
LLMs Retrieval vs Reasoning: Understanding Fundamental Limitations
By
–
Don't confuse the approximate retrieval abilities of LLMs for actual reasoning abilities.
-
Elon Musk Unveils Grok AI Bot With Sarcasm and X Data Access
By
–
Elon Musk unveils new #AI bot ‘Grok,’ says it will have sarcasm and access to X information https://
wsj.com/tech/ai/elon-m
usk-says-his-new-ai-bot-grok-will-have-sarcasm-and-access-to-x-information-b4e169de?st=3xkdzh48i8h9zi9
… via @WSJ -
01.AI Launches Yi-34B Open-Source LLM Model
By
–
http://
01.AI is a bold but long-awaited endeavor for my pursuit of AI over 4 decades. Proud to introduce the world's top open-source model Yi-34B as our first release to the developers' community encouraging fantastic LLM projects with a moderate-size, high-performing -
Tackling AI Hallucinations: Solutions and Insights
By
–
Learn more in this newsletter iteration: https://
louisbouchard.substack.com/p/tackling-ai-
hallucinations
…. -
AI Hallucinations in Explainable AI: Understanding Confident Inaccurate Predictions
By
–
Explore the enigma of AI hallucinations in Explainable AI (XAI)! Understand why AI models often yield confident yet inaccurate predictions, a phenomenon coined as "AI hallucinations".
-

Data LLMs: Unlocking Hidden Insights from Your Data
By
–
Your data always has a story to tell, but most organizations don't know what it is. We built Data LLMs. It understands the patterns in your data and extracts hidden insights from it. Read our blog post: https://
blog.abacus.ai/blog/2023/08/2
4/data-llm-get-insights-from-your-data/
… -

DeepSpeed-FastGen: 2.3x Throughput Improvement for LLM Serving
By
–
Introducing DeepSpeed-FastGen V/ @MSFTDeepSpeed **************
Serve LLMs and generative AI models with
– 2.3x higher throughput
– 2x lower average latency – 4x lower tail latency
w. Dynamic SplitFuse batching Auto TP, load balancing w. perfect linear scaling, plus -
Scale AI Offers RLHF at Scale for Language Models
By
–
if only there was a way to get RLHF at scale for your language models, trusted by all of the leading LLMs (seriously, ping us @scale_AI if you want to build durable differentiation for your LLMs via RLHF…)
-
Model Innovation Beyond Profanity: Technical Merit Analysis
By
–
is there anything novel or technically interesting about this model besides the fact that it outputs swear words?