pith. sign in

ISBN 979-8-4007-0103-0

3 Pith papers cite this work. Polarity classification is still indexing.

3 Pith papers citing it

citation-role summary

method 1

citation-polarity summary

fields

cs.LG 3

years

2026 3

verdicts

UNVERDICTED 3

roles

method 1

polarities

use method 1

representative citing papers

Layer Collapse in Diffusion Language Models

cs.LG · 2026-05-07 · unverdicted · novelty 7.0

Diffusion language models develop early-layer collapse around an indispensable super-outlier due to overtraining, resulting in higher compressibility and reversed optimal sparsity patterns versus autoregressive models.

citing papers explorer

Showing 3 of 3 citing papers.

  • Layer Collapse in Diffusion Language Models cs.LG · 2026-05-07 · unverdicted · none · ref 20

    Diffusion language models develop early-layer collapse around an indispensable super-outlier due to overtraining, resulting in higher compressibility and reversed optimal sparsity patterns versus autoregressive models.

  • Arbitrarily Conditioned Hierarchical Flows for Spatiotemporal Events cs.LG · 2026-05-02 · unverdicted · none · ref 2

    ARCH is a hierarchical flow-based generative model that enables tractable conditional intensity computation and arbitrary conditioning for spatiotemporal event distributions.

  • Structured Neural Marked Point Processes for Interpretable Event Interaction Modeling cs.LG · 2026-05-17 · unverdicted · none · ref 19 · 2 links

    SNMPP builds a product-form neural influence kernel from a signed interaction network over event classes and a delay-aware monotonic temporal network to enable explicit discovery of inter-event relationships alongside strong prediction.