We published a new paper on arXiv yesterday, filling a gap that no one had tackled before: memory testing in multi-person, multi-group scenarios. A quick explainer on why this matters Previous benchmarks for testing AI memory capabilities were basically all "two people
New Memory Testing Benchmark for Multi-Agent AI Systems
By
–
