Example: GPT-4o on OdysseyBench-Neo: • Long-context: 51.99%, 6.7K tokens
• RAG + chunk summary: 56.29%, 1.36K tokens Semantic compression ≫ brute-force memory More tokens ≠ more understanding.
GPT-4o Performance Analysis on OdysseyBench-Neo
By
–
By
–
Example: GPT-4o on OdysseyBench-Neo: • Long-context: 51.99%, 6.7K tokens
• RAG + chunk summary: 56.29%, 1.36K tokens Semantic compression ≫ brute-force memory More tokens ≠ more understanding.