Pith. sign in

Span-Oriented Information Extraction -- A Unifying Perspective on Information Extraction

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

Information Extraction refers to a collection of tasks within Natural Language Processing (NLP) that identifies sub-sequences within text and their labels. These tasks have been used for many years to link extract relevant information and to link free text to structured data. However, the heterogeneity among information extraction tasks impedes progress in this area. We therefore offer a unifying perspective centered on what we define to be spans in text. We then re-orient these seemingly incongruous tasks into this unified perspective and then re-present the wide assortment of information extraction tasks as variants of the same basic Span-Oriented Information Extraction task.

citation-role summary

background 1

citation-polarity summary

fields

cs.IR 1

years

2025 1

verdicts

CONDITIONAL 1

roles

background 1

polarities

background 1

representative citing papers

AutoData: A Multi-Agent System for Open Web Data Collection

cs.IR · 2025-05-21 · conditional · novelty 6.0

AutoData, a multi-agent system with a hypergraph message cache, automates web dataset collection from a sentence instruction and outperforms general agent baselines on the new Instruct2DS benchmark.

citing papers explorer

Showing 1 of 1 citing paper.

  • AutoData: A Multi-Agent System for Open Web Data Collection cs.IR · 2025-05-21 · conditional · none · ref 66 · internal anchor

    AutoData, a multi-agent system with a hypergraph message cache, automates web dataset collection from a sentence instruction and outperforms general agent baselines on the new Instruct2DS benchmark.