In every group I speak to, from business executives to scientists, including a group of very accomplished people in Silicon Valley last night, much less than 20% of the crowd has even tried a GPT-4 class model. Less than 5% has spent the required 10 hours to know how they tick.
LLMS
-
Speculation on the release order of GPT-5 and Sora
By
–
The big question is also what we will get first, GPT-5 or Sora
-

GenAI Payoff Event: Train and Build Your Own LLM
By
–
The #GenAI Payoff in 2024 virtual event with @matei_zaharia and @NaveenGRao is next week! Register to learn: How to train + build your own #LLM The latest research innovations to ensure LLM quality https://
bit.ly/4967zwu -

RAG Arena: Compare Retrieval Augmented Generation Strategies
By
–
RAG Arena There's so many RAG strategies out there – how do you decide which one is best? Run them in battle mode: the same question with 2 different strategies. Compare and vote for winner Awesome project by @mendableai
! Inspired by @lmsysorg http://
RAGArena.com -
Model comparison coding capabilities and context analysis
By
–
To me, it's mostly about how much better it is at coding, so of course, given that's what I care about the most, I am a little biased. But overall, it can just do things better than GPT-4, like analysis over large contexts, etc.
-
API Base Models Fine-tuning and Guardrail Instructions
By
–
As far as I know, you are not getting the base model through APIs, there is definitely fine-tuning and many guardrail instructions that are likely the result of prompting. Refusals to mess with copyright work, for example.
-
Claude 3 Performance: Design Over Actual Model Capability?
By
–
It is really hard to know how much of the Twitter reaction to the "smarts" of Claude 3 is due to the fact that Claude's system prompt/design is pushing the AI to act more human. I am not sure the model is actually better than GPT-4, but it more willing to play along with users.
-
Integrating LLMs into voice assistants for e-commerce
By
–
Why not "talking"? I am still not sure why my Alexa cannot just be a ChatGPT instead. It will be so much more convenient for e-commerce to build their own GPTs for ordering also (but ofc the adoption will take time)
-
Discussion on the stylistic differences between ChatGPT and Claude
By
–
"the final version of ChaGPT is not a chat" But regardless of that, I think Claude has its own style, I like this part
-

AI Predicts Neuroscience Experiment Outcomes Better Than Experts
By
–
Interesting new result on how AI can help advance scientific research by predicting in advanced which neuroscience experiments would yield positive findings better than human experts could And they only used GPT-3.5 class models & found fine-tuning helped https://
arxiv.org/abs/2403.03230