Pith. sign in

Title resolution pending

10 Pith papers cite this work. Polarity classification is still indexing.

10 Pith papers citing it

citation-role summary

background 2 method 1

citation-polarity summary

years

2026 9 2025 1

representative citing papers

LoopQ: Quantization for Recursive Transformers

cs.LG · 2026-05-08 · unverdicted · novelty 7.0

LoopQ provides a loop-aware PTQ framework for recursive Transformers that mitigates distribution shift, state reuse, and recursive error accumulation, yielding 68.8% higher average accuracy and 87.7% lower perplexity under W4A4 versus static baselines.

CoFrGeNet: Continued Fraction Architectures for Language Generation

cs.CL · 2026-01-29 · unverdicted · novelty 7.0 · 2 refs

CoFrGeNets implement a continued-fraction function class as plug-in replacements for transformer blocks, delivering competitive or superior downstream performance on GPT2-xl and Llama3-scale models with one-half to two-thirds the parameters.

Efficient Long-Horizon Learning for Learned Optimization

cs.LG · 2026-07-07 · conditional · novelty 6.0

A new meta-training algorithm, ELO, combines a failure-aware resume buffer with progressive expert supervision; its best learned optimizer, ELO-Celo2, outperforms AdamW on ImageNet and GPT-2 and matches Muon on language modeling.

citing papers explorer

Showing 10 of 10 citing papers.