Note: this is only a small benchmark of 90 questions: 40 from gsm8k, 30 from HotpotQA, 20 from GAIA (cherrypicked to require only Search tool and Calculator).
So results may vary!
AGENTS
-
Small benchmark of 90 questions requiring only Search and Calculator
By
–
-

Llama3-70B-Instruct matches GPT4 in agent benchmark test
By
–
Preliminary testing on my agent benchmark (based on https://
github.com/aymeric-rouche
r/benchmark_agents
…): Llama3-70B-Instruct is on par with GPT4! cc @lvwerra -
Meta Rolls Out Free AI Assistant Across WhatsApp Instagram Facebook
By
–
Meta's free AI assistant is rolling out across WhatsApp, Instagram, Facebook, and Facebook, in the company's biggest push into consumer AI yet.
-
Making AI Girlfriends: The Next Frontier in Conversational AI
By
–
now we have to make it an AI girlfriend to really go full circle
-
Jason Exploits Variant of Shunyua Yao’s Koala Architecture
By
–
and jason is exploiting a variant of the @ShunyuYao12 Koala architecture
-

Open Interpreter 01 Light: Pocket AI Agent for Voice-Controlled Laptops
By
–
This New Pocket-Sized AI Agent Lets You Voice-Control Your Laptop: Open Interpreter introduces the 01 Light, a pioneering device in natural… https://
analyticsvidhya.com/blog/2024/03/o
pen-interpreter-light-pocket-sized-ai-agent-lets-you-voice-control-your-laptop/?utm_source=dlvr.it&utm_medium=twitter
… #DataAnalytics #DataScience #DataDriven #CIO #Blockchain #BusinessIntelligence #DeepLearning #MachineLearning -
M3 Max 8B model optimization for multi-agent orchestration systems
By
–
I have an M3 Max with 64GB. You can use the 8B model if you don't have a lot of VRAM. What I realize is that the 8B model works really well as a subagent. As long as the Orchestrator and the Refiner are the 70GB models, the whole thing works really well.
-
Maestro-Ollama: Local Llama 3 70B Agent Framework
By
–
Introducing Maestro-Ollama! 🦙
— Pietro Schirano (@skirano) 18 avril 2024
You can now harness the power of the Maestro framework entirely locally using Llama 3 70B via @ollama.
Let that sink in for a second, this is a model that outperforms Claude 3 Sonnet, operating as an agent, completely locally.
What a time! 🔥 pic.twitter.com/aC34Bd6F65Introducing Maestro-Ollama! You can now harness the power of the Maestro framework entirely locally using Llama 3 70B via @ollama
. Let that sink in for a second, this is a model that outperforms Claude 3 Sonnet, operating as an agent, completely locally. What a time! -

Flow Engineering with CodiumAI and LangChain LangGraph Webinar
By
–
Flow Engineering with CodiumAI & LangChain/LangGraph New Webinar Alert – tomorrow at 9 AM PT! https://
us06web.zoom.us/webinar/regist
er/WN_fVikSl9eQv68b3ZUdQgwzA#/registration
… "Flow Engineering" is a term that has been gaining in popularity recently. The first time it was mentioned as term was in @CodiumAI paper on AlphaCodium, -
AI Multi-Agent Cooperation and Collusion Dynamics
By
–
AI multi-agent cooperation is a big thing to watch. Especially as the first studies of how agents interact are coming out (it turns out they collude sometimes) I wrote about living in a world of agents a bit here: https://
oneusefulthing.org/p/an-ai-haunte
d-world
…