REVIEW 16 cited by
Let Your Graph Do the Talking: Encoding Structured Data for LLMs
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Let Your Graph Do the Talking: Encoding Structured Data for LLMs
read the original abstract
How can we best encode structured data into sequential form for use in large language models (LLMs)? In this work, we introduce a parameter-efficient method to explicitly represent structured data for LLMs. Our method, GraphToken, learns an encoding function to extend prompts with explicit structured information. Unlike other work which focuses on limited domains (e.g. knowledge graph representation), our work is the first effort focused on the general encoding of structured data to be used for various reasoning tasks. We show that explicitly representing the graph structure allows significant improvements to graph reasoning tasks. Specifically, we see across the board improvements - up to 73% points - on node, edge and, graph-level tasks from the GraphQA benchmark.
Forward citations
Cited by 16 Pith papers
-
GraphInfer-Bench: Benchmarking LLM's Inference Capability on Graphs
Presents GraphInfer-Bench to demonstrate that no evaluated LLM-based method family closes the performance gap on graph inference tasks requiring multi-node reasoning, with plain GNNs matching or exceeding them.
-
GARDRec: Decision-Level Graph Grounding for Large Language Model Recommendation
GARDRec improves LLM-based next-item ranking by grounding decisions in knowledge-graph embeddings, personalized graph contexts, and late-stage scoring rather than prompt text.
-
C-RE-ACT: Causal RE-ACTing Agent for O-RAN Forensic Triage
An agentic O-RAN triage system that ranks root causes via SAM causal discovery and graph soft-prompting claims 89% top-3 accuracy on 140 testbed experiments.
-
Mixture-of-Experts Knowledge Graph Retrieval-Augmented Generation for Multi-Agent LLM-based Recommendation
MixRAGRec is a multi-agent KG-RAG framework with an MoE retrieval agent for query-specific granularity, a knowledge alignment agent, and a contrastive recommendation agent trained jointly via MMAPO.
-
KoRe: Compact Knowledge Representations for Large Language Models
KoRe encodes 1-hop knowledge graph subgraphs as compact discrete tokens for injection into LLMs, achieving competitive benchmark performance with up to 10x token reduction.
-
KoRe: Compact Knowledge Representations for Large Language Models
KoRe compresses one-hop knowledge-graph subgraphs into 20 discrete tokens that, injected into Qwen3-8B, match or beat text-based knowledge injection on three QA benchmarks while using up to 10x fewer tokens.
-
A Unified Graph Language Model for Multi-Domain Multi-Task Graph Alignment Instruction Tuning
UniGraphLM uses a multi-domain multi-task GNN encoder and adaptive alignment to create unified graph tokens for LLMs across diverse domains and tasks.
-
SAGE: A Self-Evolving Agentic Graph-Memory Engine for Structure-Aware Associative Memory
SAGE is a self-evolving agentic graph-memory engine that dynamically constructs and refines structured memory graphs via writer-reader feedback, yielding performance gains on multi-hop QA, open-domain retrieval, and l...
-
Bridging Input Feature Spaces Towards Graph Foundation Models
ALL-IN projects node features to a random shared space and uses covariance operators to produce representations invariant to input feature permutations and orthogonal transformations, enabling transfer across graph datasets.
-
RelAgent: LLM Agents as Data Scientists for Relational Learning
RelAgent uses an LLM agent to autonomously generate SQL feature programs paired with classical models for interpretable relational learning predictions that execute efficiently on standard databases.
-
Heterogeneous Scientific Foundation Model Collaboration
Eywa enables language-based agentic AI systems to collaborate with specialized scientific foundation models for improved performance on structured data tasks.
-
Efficiently Learning Branching Networks for Multitask Algorithmic Reasoning
AutoBRANE learns tree-structured branching networks for multitask algorithmic reasoning via gradient-based task affinities and convex relaxation.
-
Exploring the In-Context Learning Capabilities of LLMs for Money Laundering Detection in Financial Graphs
LLMs using few-shot in-context learning on serialized k-hop subgraphs from synthetic AML scenarios can assess suspiciousness and generate natural-language justifications.
-
Query-Aware Learnable Graph Pooling Tokens as Prompt for Large Language Models
LGPT and Early Query Fusion create flexible graph representations for LLMs, achieving 4.13% improvement on GraphQA without training the model.
-
Are Large Language Models Suitable for Graph Computation? Progress and Prospects
A survey of LLMs for graph computation introduces a role-based taxonomy of executors versus planners and concludes that current models suit simple small-scale tasks but remain unreliable for large-scale exact computation.
-
Edge-Aware Curvature Modeling for Graph Understanding in Large Language Models
CureLLM adds curvature-aware edge modeling and prompt-based alignment to graph LLMs, claiming to resolve over-squashing from negative curvature and outperforming 20 baselines on 11 datasets.
discussion (0)
Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.