Urgent.News

What's breaking now, across thousands of outlets.

AI

When does an agent beat RAG? We benchmarked three pipelines on a TigerGraph knowledge graph

Agentic GraphRAG Hackathon by TigerGraph, Round 1. Code: https://github.com/anirudh12032008/agentic-graphrag-tigergraph Ask a RAG system "How many biathlon events at the 2018 Winter Olympics had more than 73 competitors?" and it will give you a confident number. It will also be wrong. The answer depends on 8 to 43 documents, and a top-8 vector search can't see all of them. We built three…

When faced with certain types of knowledge graph queries, traditional Retrieval-Augmented Generation (RAG) systems may struggle to provide accurate answers, while agent-based GraphRAG systems can surpass RAG in both accuracy and efficiency. Researchers from TigerGraph conducted a benchmark using three pipelines on the same Claude-sonnet-5 language model, embeddings, and TigerGraph Savanna instance.

The first pipeline, RAG, employed a top-8 vector search followed by a single LLM call, resulting in a 48% accuracy on 100 public questions. The second pipeline, GraphRAG, used a top-6 vector search, potentially up to three seed events, followed by a fixed graph expansion before another LLM call, achieving 69% accuracy. The third pipeline, Agentic GraphRAG, was an LLM orchestrator that performed tool-based steps such as entity linking, GSQL aggregation, graph traversal, venue and date lookup, vector search, and document reading, ultimately scoring 99% accuracy.

GraphRAG performed better than RAG on aggregation and superlative questions, while Agentic GraphRAG outperformed both on multi-hop questions. The agent utilized GSQL aggregation and graph traversal to solve specific question types, achieving 100% accuracy for lookups, 67% for aggregations, and 90% for superlatives. In contrast, GraphRAG only achieved 80% accuracy on aggregation questions.

The agent also outperformed RAG in multi-hop questions, scoring 100% against RAG's 21%. Agentic GraphRAG demonstrated an ability to flag real ambiguity, such as with multiple finals at the same venue on different dates, and reported all candidates with citations instead of guessing. It even caught a data tie in fencing competitions, reporting both results with citations.

Compared to RAG, Agentic GraphRAG required 1.24 times more tokens but delivered 2.06 times the accuracy, making it 40% cheaper per correct answer. Additionally, the agent ensures that all answers are backed by citations, preventing it from relying on memory alone. All components, including the graph schema, GSQL queries, pipelines, benchmarking tools, and a Streamlit dashboard, are open source and readily available for replication.

Written by urgent.news from Dev.to's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.

Read the original at dev.to →

More in AI

Where should an AI agent's idempotency key come from?

Your agent's worker sends a payout request, and then the process gets killed before the response comes back. The supervisor restarts it.

  • Idempotency key should be generated during action planning, not execution.
  • Persist the key with task state before first attempt, in databases, queues, or workflows.

Sierra and Meta propose a standard for personal AI agents

Sierra and Meta announced the Personal Agent Protocol on October 6, 2026, a proposed open standard for how a shopper's AI agent should deal with a business.

  • Sierra and Meta propose Personal Agent Protocol (PAP) as open standard for personal AI agents.
  • PAP aims to replace agents acting like human users on websites with clear rules.

Handoff patterns are your agent's worst enemy, unless you implement them this way

I spent six weeks debugging a handoff that looked perfect in the trace. The planner produced a clean, well-structured plan. The executor followed it step by step.

  • Handoff Tax causes information loss in agent chains
  • Structured handoff representation improves downstream feasibility
  • A2A protocol with stateful Task object mitigates Handoff Tax

More from Thursday 8 October →