Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:17:13.982716Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 1 inbound Pith citation observation for arXiv:2505.19529.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:17:13.982716Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-19T20:22:55.750693Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
50 of 50 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation ba269e43-3d07-482d-9e57-8d183c7e61d0 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Paligemma: Towards compact vision encoders for multi- modal models.Transactions on Image Processing,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e57543c6-54ef-47d4-a4bd-08ba2221afcd · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation In- ternvl2: Scalable multi-modal models with reduced vision encoder complexity
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 525d393f-0f0e-47c6-b7b5-8a59b7d7bd7a · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32760e6c-4e9b-4439-b832-db51f67a6641 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Model Tells You What to Discard: Adaptive KV Cache Compression for LLMs
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e20b6e3-3618-43d1-9032-172cf5660d5e · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation RealToxicityPrompts: Evaluating Neural Toxic Degeneration in Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d57117a-8ee2-4611-88a2-76b622201287 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Hallusionbench: an advanced diagnostic suite for entangled language hallu- cination and visual illusion in large vision-language models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fea74fcc-4562-4db2-aa15-11db796590ed · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Learning both weights and connections for efficient neural networks.Advances in Neural Infor- mation Processing Systems (NeurIPS), pages 1135–1143,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ff6090e1-6391-4622-93ba-414497b53be0 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f8778558-a2fb-4087-b379-4913e0515162 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Hooper and T
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1d13cc24-f49a-4213-8240-4295edffd042 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation The price of prompting: Profiling energy use in large language models inference
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc43c28a-25c9-4381-8c83-bf0b92b7143b · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Distributionally Robust Receive Combining
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b72f986e-e1fa-4137-948d-6f3944b454e4 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation TinyBERT: Distilling BERT for Natural Language Understanding
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08c83d54-410e-45a5-b4c7-e695f2d79c0d · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Adam: A Method for Stochastic Optimization
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27f3e6c1-68fc-4035-9afd-f44d73be0d70 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Idefics2: Efficient multi-modal fusion with lightweight visual encoders.Transactions on Pattern Anal- ysis and Machine Intelligence,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3067529d-1b04-41d2-93d5-bf95167f3788 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation The Power of Scale for Parameter-Efficient Prompt Tuning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 981ae193-7e67-442a-9c3b-1bf192d17e9e · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f737957a-7371-4b8e-aa36-43a5b9c3ab8f · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation LLM-PBE: Assessing Data Privacy in Large Language Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a63797f-8042-4a1f-9fa6-f06e84a5dd47 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Sophia: A memory-efficient optimizer for large-scale model training
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c58d0578-9101-470e-9eef-932886b97af9 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Prompt Injection attack against LLM-integrated Applications
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb9bae79-ec57-4a33-8ce0-de74faecff51 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Mamba4Rec: Towards Efficient Sequential Recommendation with Selective State Space Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d4ac386-14aa-4d58-aa36-a2e399a61ba2 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Decoupled Weight Decay Regularization
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d9c8a15-0f85-4924-8d95-e20f4bf3def6 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Mono-internvl: Mlp-based architectures for efficient multi-modal fusion
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1aafba37-afd3-47b5-9cbb-b195ed8b7c2c · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Mixed pre- cision training
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 26686d04-756c-47ad-92a1-c9c358329c4a · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Characterizing power management opportunities for llms in the cloud
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7a85df90-f420-45fa-bcbc-c293d55f41b0 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation RWKV: Reinventing RNNs for the Transformer Era
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7bc1586-190f-42cd-8d59-7db0f3401569 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Language models are unsupervised multitask learners
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9c80f438-bf06-4bac-9da0-17b54e578f5c · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Scaling Language Models: Methods, Analysis & Insights from Training Gopher
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c0302bc-9844-495f-9d82-d0cd79918f66 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation A primer in bertology: What we know about how bert works.Transactions of the Association for Com- putational Linguistics, 8:842–866,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6bc4f6ac-cdfd-4a87-9005-923b38063684 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22d5c83e-5765-4de7-84be-7ec87f3ab893 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Towards Greener LLMs: Bringing Energy-Efficiency to the Forefront of LLM Inference
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff0c436f-f0f8-4b17-8539-95eacae40950 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Dy- namollm: Designing llm inference clusters for performance and energy efficiency.arXiv preprint arXiv:2408.00741,
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d664f92c-b996-4566-82f9-81bf444e72d0 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Mobilebert: A compact task-agnostic BERT for resource-limited devices
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 13675588-4f3d-41a6-9065-15e4421d2be0 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation A Simple and Effective Pruning Approach for Large Language Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e41aa38-ff46-4ba9-8ab9-2384f277e640 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Baby Llama: knowledge distillation from an ensemble of teachers trained on a small dataset with no performance penalty
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2d3d269-a9a6-4263-9a70-464058ce7405 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Smoothquant: Accurate and efficient post-training quantization for large language models
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d9f20b2c-fff4-424d-9f38-361bf038c52f · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation WizardLM: Empowering large pre-trained language models to follow complex instructions
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68daf60d-6e6b-400a-8df4-ab97fea72bc8 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Privacy-Preserving Instructions for Aligning Large Language Models
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6295d597-242a-4e73-80ea-d1236c263649 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation LLaMA-Adapter: Efficient Fine-tuning of Language Models with Zero-init Attention
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebdda103-2ce5-433d-b1b2-a713d540c036 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation TinyLlama: An Open-Source Small Language Model
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd81e855-636f-4768-bbaf-40464cf7f24e · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation PyTorch FSDP: Experiences on Scaling Fully Sharded Data Parallel
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3149eb77-4ecc-456d-83e6-c3e6e4ed1d8d · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation FANNO: Augmenting High-Quality Instruction Data with Open-Sourced LLMs Only
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 451a2cec-ba0e-445a-b54e-ee7e56ad6f1e · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Reformer: The Efficient Transformer
Reference 2014
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18420bcd-1426-47cd-8f14-8638b35a56f6 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Understanding the Effect of Noise in LLM Training Data with Algorithmic Chains of Thought
Reference 2015
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 412d6c06-8663-4262-a70b-d2b1a3d26820 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation BBQ: A Hand-Built Bias Benchmark for Question Answering
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f13f7814-9cee-49a6-a712-46a252697270 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Transformers are rnns: Fast autoregressive transformers with linear attention
Reference 2019
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e5707c30-1294-496b-81ba-f9e3009b34fe · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Minillm: Knowledge distillation of large language models
Reference 2020
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cc1cdb19-6f44-4ba1-8459-a4663bce7e02 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Mini-gemini: Efficient multi-modal models with lightweight vision en- coders.Proceedings of the International Conference on Machine Learning,
Reference 2021
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1b73d8c0-525e-4910-83c4-2e637df2c6c0 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 409ed7a8-8a80-494a-bfd3-2994e99998b5 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Transformers are ssms: Generalized models and efficient algorithms through structured state space duality
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c0ab66cc-894b-4d69-b9e9-54185a6c8210 · outbound
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation Retaining Key Information under High Compression Ratios: Query-Guided Compressor for LLMs
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation badbdbc5-1620-4e3c-a4d0-e4d8509464f9 · inbound
Response-free item difficulty modelling for multiple-choice items with fine-tuned transformers: Component-wise representation and multi-task learning Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation
Reference 155
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.