ARXIV:2605.31031 · GRAPH REASONING · SUBMITTED 01 JUN · 20:23 UTC · FRESHNESS STALE

VerifiedSource: PDF linkedVerifiedPaperPack: citation fields availablePartialProof: unverified proof status

GraphARC: A Comprehensive Benchmark for Graph-Based Abstract Reasoning

Saku Peltonen · August Bøgh Rønberg · Andreas Plesner · Roger Wattenhofer · arXiv

GraphARC is a new benchmark for abstract reasoning on graph data, revealing limitations in current language models and paving the way for graph foundation models.

Ship in 2-4 weeks›Score7.0Evidence unverified

Opportunity summary

Pain GraphARC is a new benchmark for abstract reasoning on graph data, revealing limitations in current language models and paving the way for graph foundation models.

Evidence 0 refs | 3 sources | 50% coverage

Blocker Evidence unverified

Open Build Read PDF Signal Canvas Track

PROBLEM

GraphARC is a new benchmark for abstract reasoning on graph data, revealing limitations in current language models and paving the way for graph foundation models. We introduce GraphARC, a benchmark for abstract reasoning on…

METHOD

Full abstract

Relational reasoning lies at the heart of intelligence, but existing benchmarks are typically confined to formats such as grids or text. We introduce GraphARC, a benchmark for abstract reasoning on graph-structured data. GraphARC generalizes the few-shot transformation learning paradigm of the Abstraction and Reasoning Corpus (ARC). Each task requires inferring a transformation rule from a few input-output pairs and applying it to a new test graph, covering local, global, and hierarchical graph transformations. Unlike grid-based ARC, GraphARC instances can be generated at scale across diverse graph families and sizes, enabling systematic evaluation of generalization abilities. We evaluate state-of-the-art language models on GraphARC and observe clear limitations. Models can answer questions about graph properties but often fail to solve the full graph transformation task, revealing a comprehension-execution gap. Performance further degrades on larger instances, exposing scaling barriers. More broadly, by combining aspects of node classification, link prediction, and graph generation within a single framework, GraphARC provides a promising testbed for future graph foundation models.

RESULT

ScienceToStartup currently rates this 7.0/10 on the public viability pass. More broadly, by combining aspects of node classification, link prediction, and graph generation within a single framework, GraphARC provides a promising testbed for future…

WHY NOW

Graph Reasoning moved forward this cycle; last verified June 2026. Public score 7.0/10. Production flags indicate code availability.

Continue into Read for claims, analysis, references, and neighboring papers.

Opportunity summary

Score7.0

PainGraphARC is a new benchmark for abstract reasoning on graph data, revealing limitations in current language models and paving the way for graph foundation models.

Evidence0 refs | 3 sources | 50% coverage

Blockerno shell-level blocker reported

Analysis summary

GraphARC is a new benchmark for abstract reasoning on graph data, revealing limitations in current language models and paving the way for graph foundation models.

VerifiedSource: PDF linkedVerifiedPaperPack: citation fields availablePartialProof: unverified proof status

Competitive landscape

GraphARC is a new benchmark for abstract reasoning on graph data, revealing limitations in current language models and paving the way for graph foundation models.

Segment

Graph Reasoning

Adoption evidence

No public code link in the paper record yet

Commercial read

7.0/10 public viability

Direct

not classified

Adjacent

not classified

Substitute

not classified

Unknown

not classified

{ "contract_version": "paper-r2", "paper_id": "9bb91b6b-e58e-4444-b491-4062eab65564", "arxiv_id": "2605.31031", "canonical_route": "/paper/grapharc-a-comprehensive-benchmark-for-graph-based-abstract-reasoning", "active_tab": "synced from current hash by the drawer client", "selected_artifact": "grapharc-a-comprehensive-benchmark-for-graph-based-abstract-reasoning", "endpoints": { "paper_pack": "/api/v1/paper/grapharc-a-comprehensive-benchmark-for-graph-based-abstract-reasoning/paper-pack", "build_passport": "/api/v1/paper/grapharc-a-comprehensive-benchmark-for-graph-based-abstract-reasoning/build-passport", "mcp_resource": "sciencetostartup://surfaces/paper-workspace" } }

{ "surface": "paper", "mode": "paper", "query": "GraphARC: A Comprehensive Benchmark for Graph-Based Abstract Reasoning", "normalized_query": "2605.31031", "route": "/paper/grapharc-a-comprehensive-benchmark-for-graph-based-abstract-reasoning", "paper_ref": "grapharc-a-comprehensive-benchmark-for-graph-based-abstract-reasoning", "topic_slug": null, "benchmark_ref": null, "dataset_ref": null }

{ "@context": "https://schema.org", "@graph": [ { "@type": "WebPage", "@id": "https://sciencetostartup.com/paper/grapharc-a-comprehensive-benchmark-for-graph-based-abstract-reasoning#webpage", "url": "https://sciencetostartup.com/paper/grapharc-a-comprehensive-benchmark-for-graph-based-abstract-reasoning", "name": "GraphARC: A Comprehensive Benchmark for Graph-Based Abstract Reasoning", "description": "GraphARC is a new benchmark for abstract reasoning on graph data, revealing limitations in current language models and paving the way for graph foundation models.", "isPartOf": { "@id": "https://sciencetostartup.com/#website" } }, { "@type": "ScholarlyArticle", "@id": "https://sciencetostartup.com/paper/grapharc-a-comprehensive-benchmark-for-graph-based-abstract-reasoning#scholarlyArticle", "headline": "GraphARC: A Comprehensive Benchmark for Graph-Based Abstract Reasoning", "description": "GraphARC is a new benchmark for abstract reasoning on graph data, revealing limitations in current language models and paving the way for graph foundation models.", "url": "https://sciencetostartup.com/paper/grapharc-a-comprehensive-benchmark-for-graph-based-abstract-reasoning", "sameAs": "https://arxiv.org/abs/2605.31031", "identifier": { "@type": "PropertyValue", "propertyID": "arXiv", "value": "2605.31031" }, "isAccessibleForFree": true, "isPartOf": { "@id": "https://sciencetostartup.com/#website" }, "datePublished": "2026-05-29T09:03:30.000Z", "author": [ { "@type": "Person", "name": "Saku Peltonen" }, { "@type": "Person", "name": "August Bøgh Rønberg" }, { "@type": "Person", "name": "Andreas Plesner" }, { "@type": "Person", "name": "Roger Wattenhofer" } ], "additionalProperty": [ { "@type": "PropertyValue", "propertyID": "viabilityScore", "value": 7 }, { "@type": "PropertyValue", "propertyID": "researchDomain", "value": "Graph Reasoning" }, { "@type": "PropertyValue", "propertyID": "commercialReadiness", "value": "code" } ] }, { "@type": "BreadcrumbList", "itemListElement": [ { "@type": "ListItem", "position": 1, "name": "Home", "item": "https://sciencetostartup.com" }, { "@type": "ListItem", "position": 2, "name": "Graph Reasoning", "item": "https://sciencetostartup.com/topics" }, { "@type": "ListItem", "position": 3, "name": "GraphARC: A Comprehensive Benchmark for Graph-Based Abstract", "item": "https://sciencetostartup.com/paper/grapharc-a-comprehensive-benchmark-for-graph-based-abstract-reasoning" } ] } ] }

Competitive landscape

GraphARC is a new benchmark for abstract reasoning on graph data, revealing limitations in current language models and paving the way for graph foundation models.

Segment

Graph Reasoning

Adoption evidence

No public code link in the paper record yet

Commercial read

7.0/10 public viability

Direct

not classified

Adjacent

not classified

Substitute

not classified

Unknown

not classified

GraphARC: A Comprehensive Benchmark for Graph-Based Abstract Reasoning

GraphARC: A Comprehensive Benchmark for Graph-Based Abstract Reasoning

Claim map

Constellation map

Competitive landscape

Buzz

PDF

REFERENCES

Related Papers

Related Resources

Subscribe to the weekly brief

Build artifacts

Brief

Experiment plan

Validation checklist

Scientific founder

Translational engineer

Domain operator

GTM lead

Regulatory/clinical advisor

Timeline

Claim map

Constellation map

Competitive landscape

Buzz

PDF

REFERENCES

Related Papers

Related Resources

Subscribe to the weekly brief

Build artifacts

Brief

Experiment plan

Validation checklist

Scientific founder

Translational engineer

Domain operator

GTM lead

Regulatory/clinical advisor

Timeline