A frozen BERT-style backbone with per-task LoRA adapters scores 27 PISA items at 60% lower GPU memory and 40% lower latency, with a 4.5% QWK drop.
Transformers: State-of-the-art natural language processing
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2024 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Efficient Multi-Task Inferencing with a Shared Backbone and Lightweight Task-Specific Adapters for Automatic Scoring
A frozen BERT-style backbone with per-task LoRA adapters scores 27 PISA items at 60% lower GPU memory and 40% lower latency, with a 4.5% QWK drop.