Human reviewers stopped rewarding lexical complexity as LLMs made it cheap
A frozen-rater design separates preference drift from composition change across 124K reviews
Digital Libraries
Covers all aspects of the digital library design and document and text creation. Note that there will be some overlap with Information Retrieval (which is a separate subject area). Roughly includes material in ACM Subject Classes H.3.5, H.3.6, H.3.7, I.7.
sort pith recommended most recent
A frozen-rater design separates preference drift from composition change across 124K reviews
Citations spanning fully separate subfields decline while overlapping-field recombination grows
· “The conservative turn in science: The changing character of knowledge recombination”
Formal five-stage pipeline uses cheap filters before expensive AI, turning deduplication into data deepening
Systematic study across six formats gives practical advice for link rot analysis, knowledge graph building, and reproducibility
· “URL Extraction from Scholarly Documents: A Cross-Format Comparative Analysis”
X-DigCheck provides a domain-independent environment for continuous profile–graph alignment across regeneration cycles.
A domain-stripped computational skeleton surfaces same-problem papers in unrelated fields, lifting retrieval precision from 0.22 to 0.56.
An expert correcting machine pre-annotations for 33 hours produced a parser that outperforms all existing Latin treebank models.
· “Correction as Annotation: Bootstrapping a Dependency Parser for Documentary Medieval Latin”
Students, collaborators, and event organizers get a searchable, evidence-checked way to find diaspora academics.
· “VietProfs: A Public Directory of the Vietnamese Academic and Research Diaspora”
Multi‑agent system retrieves from multiple databases, re‑ranks by impact, and uses real peer‑review comments to revise drafts, achieving…
A synthetic control of 18 other arXiv archives isolates a math-specific surge tied to AI capability thresholds.
· “The Generative AI Gold Rush in Theoretical and Computational Research”
Citation networks reveal weakly connected power centers of nonprofits, social movements, and voluntary action
· “Constructing the Field of Philanthropic and Nonprofit Studies: Evidence from Citation Networks”
Machine-readable records captured at collection time let AI agents find and judge data too large to move.
New 928-book benchmark across six languages tests domain, author, and language transfer; Transformer detectors cross languages but with…
LLMs cover human strength points more reliably than weakness points when guided by an external score.
· “Guiding LLM Peer Reviewers: The Impact of Score Anchors on Review Evidence and Accuracy”
Five theory-guided lenses on 11,560 ICLR reviews show AI and human critique redistribute evaluative work.
· “Beyond Human-Likeness: Mapping the Scientific Critique Profiles of LLMs and Human Reviewers”
Six models consistently replace critical citations with support, cite older popular work, and ignore authors' collaborators.
· “Citing Less Critically: LLMs Reshape the Rhetoric and Reach of Scientific Citation”
A new open graph joins expert-curated math semantics to centuries of publications, beyond citation networks.
· “The zbMATH Open Knowledge Graph: Tracing Centuries of Mathematical Research”
In 73,489 health-science articles, two models consistently favoured molecular research over surveys and clinical studies.
Study of 1,975 Pharmakon Network papers shows cited outputs persist and authors keep funding after exposure.
· “Tracing high-profile attention to questionable research as a case for funder due diligence”
A new 1,378-lot benchmark compares commercial, institutional, and local VLM deployments for historical lot extraction.
· “Lot Machine: Multimodal Lot Extraction from Auction Catalogs”
SoniMet maps field-weighted citation counts to auditory parameters, adding an audio channel to bibliometric analysis for individual…
· “SoniMet - A tool for sonifying and visualizing the performance of single researchers”
ASKS system yields a source-traceable research portrait with zero splits and low membership churn in a longitudinal demonstration.
Persistent homology of concept embeddings reveals two asymmetries: gaps pay off in empirical fields, yet their number explodes while…
· “Filling holes in science draws collective attention, but most higher-order holes remain unexplored”
Across 191,375 Hugging Face repos and 2,214 papers, visibility far outpaces the evidence libraries need.
A global citation network analysis reveals modest domestic preference and convergence with US in disruptive research impact.
Measurement of 2,811 health dataset descriptions finds none meets all eight mandatory fields of HealthDCAT-AP Release 7
· “Measuring the Installed Base: Nordic Health Dataset Catalogues Against HealthDCAT-AP Release 7”
BLANC compares cluster associations before and after keyword filtering, flagging 'established globally, unexplored locally' at ~30%…
Two million papers show positive link to disruptive citations yet weaker cross-field recombination and narrower inputs.
· “AI-assisted writing and the reorganization of scientific knowledge”
Two-layer system assesses knowledge claims and human input levels to work within existing journal systems.
· “Rethinking Publication: A Certification Framework for AI-Enabled Research”
Patents cite hybrid models more often, but gold and diamond OA show stronger semantic links, especially inside patent bodies.
· “Discoverability matters: Open access models and the translation of science into patents”
Maps open-infrastructure principles onto a concrete hardware-to-application stack.
· “Towards a Definition of the Computational Architecture of Open Scholarly Infrastructures”
On 637 records it links 86% of metadata slots; the missing piece is human-verified correctness.
· “Automated Construction of FAIR Digital Object Knowledge Graphs from Flat Cultural Heritage Records”
Beyond data and metadata: how an instrument was used is itself preserve-worthy science.
· “Beyond FAIR Data: Instrument Traces for Active and Autonomous Scientific Experimentation”
A rooted taxonomy's height function fully determines a metric on topic profiles, computable in linear time.
· “Taxonomy-aware distances between scholarly topic profiles via an exact simplex embedding”
The new open dataset adds department-level detail and a validated parent-child hierarchy for universities.
Images, video, and audio become first-class nodes, so SPARQL can search and segment them directly.
· “MediaGraph: A Content-Aware Data Model and Query Framework for Multimodal Knowledge Graphs”
Tracing 64 Fields Medalists back 57 generations, the paper finds 84.4% of lineages converge on five medieval scholars.
An audit of 117 open-source projects finds their own metadata surfaces often contradict each other.
· “A Multi-Surface Consistency Audit of Software Citation Metadata”
Across three test areas, the knowledge base's modularity falls as the field forms and can later rebound.
· “Declining Modularity of Intellectual Bases During the Emergence of Research Areas”
Year-by-year list prices from 2019–2025 make open-access fee trends and publisher comparisons easy to compute.
· “A dataset of article processing charges from 14 scholarly publishers, 2019-2025”
Four expert-scored tasks move from spotting axes to judging evidence, with Bloom-informed questions scoring each level of understanding.
· “A Pathway to General-Purpose Scientific AI: Multimodal Comprehension of Scientific Images”
Ten iconic composers, 17,053 chords, 23,447 notes—enough to test what defines MPB's shared style.
Across 273,109 papers from 76 facilities, novelty rises when in-house scientists co-lead with external users.
· “Co-leading Teams Drive Scientific Novelty in Large-scale Research Infrastructures”
It consolidates models, accuracy ranges, and open challenges so new AI researchers can enter AD diagnosis.
Even the best model misses 40% of pinpoint-page errors in opinions, so AI citechecking cannot verify support.
A map tool lets any genealogist add sourced building records straight into the global open knowledge graph.
· “Domus: An Open-Data Web Platform for House History Research”
Across four models, the gain appears only where the grading key shares the injected resource.
A 24-year look at Canadian grant data finds medical sciences climbing while chemistry and space science fall.
· “Patterns of Research Funding Across Research Subjects: The Case of NSERC”
Keyword analysis of 113,690 SDG-16 papers finds welfare language in high-peace countries, security language in low-peace ones.
21-card protocol frozen before production made a No-Go boolean halt the work; outsiders can recompute it all.
A PageRank-style journal metric tracks impact while withstanding fake-citation attacks about ten times better than raw counts.
A review argues coordinated publishing fraud shows up in shared authors and citation rings that single-journal checks miss.
A word-frequency analysis of 1.2 million papers finds LLM editing in nearly all, led by Discussion sections.
· “Most biomedical publications show signs of LLM-assisted writing”
Fixation number plus total fixation duration gives the best F1 gains on Chinese academic abstracts.
· “Leveraging Human Reading Behavior for Keyphrase Extraction: A Webcam-based Eye-tracking Corpus”
HexEval scores intrinsic quality and external behavior separately, so each rating can be traced and checked.
· “HexEval: An Evidence-Driven Hexagonal Framework for Multidimensional Scholar Assessment”
Author names are withheld, yet a small male tilt remains across most UK research fields, an 89,744-article study finds.
· “Does ChatGPT score research quality differently by gender?”
Built from 62 journals, StatCite pairs paper metadata with co-citation, coupling, and journal-level graphs for public analysis.
· “StatCite: A Large-scale Citation Network Dataset for Statistics and Data Science”
Keyword clusters become named concepts covering 98.8% of OpenAlex topics, rated ~89% relevant by experts.
· “SCALE: Scientific Concept Aggregation via LLMs and Embeddings for Fine-Grained Taxonomy Extension”
A preregistered scan of 146,239 releases finds the punctuation shift came late and broad, with caveats.
Extreme heat and flood-hurricane mental health pairs recur far beyond chance, even after adjusting for term popularity.
Keeps the classic single-page style, adds per-entity RDF provenance, Markdown rendering, and static export.
Each dataset links to its source, citation guidance, conversion code, and fixed version snapshots for reproducible reuse.
· “A repository for discovery and reuse of higher-order network datasets”
Six resolution tiers on Florence's Loggia: native wins; AI detail fails multi-view checks.
· “Loggia dei Lanzi: AI Thermography Enhancement Comparisons through 3D Photogrammetry”
One distributed library spans influence, popularity, momentum, and field-normalised scores, running in minutes on 2.1 billion citations.
· “BIP! Ranker: A Software Library for Citation-Based Impact Indicators on Large-Scale Graphs”
Rate-matched random placebo still leaves a positive gap, and four designs that avoid the timing artifact agree.
Expert-finished field lists for biology, materials, imaging, physics, and psychology—ready for extraction and knowledge graphs.
A span-grounded AI extractor found 5,766 Analects reuses across 24 histories—wording drifted while practice held.
A 34-paper survey finds rapid growth, a heavy biomedical focus, and little direct evaluation of retrieval tools.
· “Scientific Knowledge Discovery in the Age of Large Language Models”
At equal reviewer scores, borderline papers without a top-25 author lose acceptance points, yet show no better downstream results.
· “Bias at the Borderline: Who Gets the Benefit of the Doubt in Peer Review?”
In a 200-article study, ChatGPT's averaged rankings matched or beat single human experts—but detailed PDF critiques did not improve scores.
Four bibliometric platforms classify by different mixes of experts, AI, citations, and granularity—know the scheme before you search.
Even preprints never peer-reviewed are cited as often as published ones — and the trend is exponential since 2014.
· “Preprints Without Curation Are Increasingly Cited by Journals”
Four to eight thousand raw Latin or Greek sentences turn token models into strong reuse detectors and retrievers.
Schema retrieval and Steiner join plans keep language-model answers tied to expert knowledge.
Executable Cypher plus OCR-variant expansion recovers counts and multi-hop facts standard retrieval misses.
A spatial text format and automatic traditional layout finally make Korean mensural notation editable data, not hand-built images.
Automated figure-and-text reading makes literature-embedded spectra machine-readable for training and cross-lab use.
· “Harnessing X-ray Absorption Spectroscopy Data through Multimodal Mining of Battery Literature”
Methods, DFT training data, and benchmark results reduce to a single SPARQL query instead of a manual literature read.
Adding impact factors of retracted papers gives a net-value score, h-z, that can go deeply negative.
Mainstream venues published almost nothing on how public agencies build software — a gap for evidence-based practice.