Research1 min read
GT Bench: Benchmark for LLM Graph Reasoning
Researchers introduced GT Bench, a benchmark with over 100,000 examples across four graph representations, to evaluate LLM performance on algorithmic graph problems. The Graph Theory Agent (GTA) system, combining a representation selector and a plan-and-decompose agent, achieved improved results on the GT Bench benchmark.
From arXiv cs.AI
