Conditional Co-Ablation recovers self-repair backup heads in transformers by scoring conditional ablation growth, raising ROC-AUC from 0.33 to 0.91 on the IOI circuit and transferring to induction across models.
Think You Have Solved Question Answering? Try
4 Pith papers cite this work. Polarity classification is still indexing.
years
2026 4verdicts
UNVERDICTED 4representative citing papers
A parameter-neutral fuzzy-logic FFN augmented with self-forgetting quantifiers produces legible grammatical-licensing detectors while matching baseline perplexity on OpenWebText.
Coverage-aware pruning using per-corpus utility profiles on WikiText2 and C4 improves zero-shot accuracy and reduces perplexity degradation in two MoE models at 25-75% retention compared to baselines, without downstream data.
LuckyStar 111B adapts Cohere's Command A model with four scaling techniques to improve tool-use, math reasoning, and NL2SQL in Korean-English while preserving general instruction following.
citing papers explorer
-
Conditional Co-Ablation: Recovering Self-Repair Backups in Transformer Circuits
Conditional Co-Ablation recovers self-repair backup heads in transformers by scoring conditional ablation growth, raising ROC-AUC from 0.33 to 0.91 on the IOI circuit and transferring to induction across models.
-
Explicit Fuzzy Logic in the Feed-Forward Layer: Self-Forgetting Quantifiers Discover Legible Grammatical-Licensing Detectors
A parameter-neutral fuzzy-logic FFN augmented with self-forgetting quantifiers produces legible grammatical-licensing detectors while matching baseline perplexity on OpenWebText.
-
Generic Expert Coverage for Pruning SparseMixture-of-Experts Language Models
Coverage-aware pruning using per-corpus utility profiles on WikiText2 and C4 improves zero-shot accuracy and reduces perplexity degradation in two MoE models at 25-75% retention compared to baselines, without downstream data.
-
Think in English, Answer in Korean: Efficient Adaptation of Multilingual Tool-Using Agents
LuckyStar 111B adapts Cohere's Command A model with four scaling techniques to improve tool-use, math reasoning, and NL2SQL in Korean-English while preserving general instruction following.