Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-20T13:58:55.899958Z
Paper Citation Record · LEDGER
As of 23 July 2026, this Paper Citation Record lists 38 of 38 outbound references and 0 inbound Pith citation observations for arXiv:2605.17653.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-20T13:58:55.899958Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-07-23T06:31:01.910684+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
38 of 38 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation f454a062-4362-4319-9d36-fb2482168c23 · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models Composer: A search framework for hybrid neural architecture design.arXiv preprint arXiv:2510.00379
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 245748d5-74d5-4bed-acde-5401e8f8a8ae · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 265c75c0-fe12-46f6-ab5a-447bf7c606b9 · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models SmolLM2: When Smol Goes Big -- Data-Centric Training of a Small Language Model
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 6a798b53-f60d-4ac7-821e-15873659d257 · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models Pythia: A suite for analyzing large language models across training and scaling
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 5752a7c0-9f73-44b7-bb53-621501972941 · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models Eyeriss v2: A flexible accelerator for emerging deep neural networks on mobile devices.IEEE Journal on Emerging and Selected Topics in Circuits and Systems, 9(2):292–308
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 699f9b74-2138-4ae5-9a4b-78d1d70e1324 · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models BoolQ: Exploring the surprising difficulty of natural yes/no questions
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation a6e6eb28-7048-4e65-ab03-5618a201d1ff · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 0738569c-ba2c-4615-ab2e-6dd38e770aa0 · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models and Pratap, A
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation b9fd3401-d6eb-464b-8225-b582a326ba57 · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation c62bd50b-8f30-4b14-8baa-e616ce2a0147 · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models The Pile: An 800GB Dataset of Diverse Text for Language Modeling
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 85520e9a-b14c-4937-bad8-ccb048a42cc0 · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models In: ACM/IEEE Design Automation Con- ference
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 3c51bfaf-0e7d-4185-af90-a4ca1c2af7fa · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models Jet-nemotron: Efficient language model with post neural architecture search
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 8f8aea15-1d14-4fde-b631-5fd2efe1f791 · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models Training Compute-Optimal Large Language Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 98073453-e083-40c0-afba-e9e15e0897d5 · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models The MiniPile Challenge for Data-Efficient Language Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation f00d71ef-4f62-4d27-a411-b6408ce28f75 · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models Emer, and Saman P
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 226e53a4-89e9-4e5e-8adf-2163eb5378a2 · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models MELTing Point: Mobile evaluation of language transformers
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation eb2689d7-2f6d-48f3-af2a-f5511c73072c · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models Transformers in Speech Processing: A Survey
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 528299c8-ee40-4e5c-b2ae-f9b979041b16 · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models Mobilellm: Optimizing sub-billion parameter language models for on-device use cases
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 24c938b5-d8a1-4bd5-a924-17a5ee49cb3d · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models MobileLLM: Optimizing Sub-billion Parameter Language Models for On-Device Use Cases
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 96a6c150-4b57-4b3c-b9c4-c4d85efb6ae9 · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models OpenELM: An Efficient Language Model Family with Open Training and Inference Framework
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation d6f035fe-a64a-4d01-9b9b-55ceb843e131 · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models Ying, Anurag Mukkara, Rangharajan Venkatesan, Brucek Khailany, Stephen W
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation c7a5d863-3547-4dd4-9818-ac887b45f09c · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models Parashar et al
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 19c52e5a-c2b9-4a5a-a79e-9efebffe1a61 · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models Hare, and Geoff V
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 30267bba-655b-4a12-a4ff-393bf88449a5 · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models The FineWeb Datasets: Decanting the Web for the Finest Text Data at Scale
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 947dfc33-844b-4adc-97af-099ff6d83087 · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models Fast Transformer Decoding: One Write-Head is All You Need
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 70276278-33ce-491f-9f6e-6854e5403ea5 · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models HW-GPT-Bench: Hardware-Aware Architecture Benchmark for Language Models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation fe3671b2-7049-4047-a71d-bd39ad74bb7a · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models An 11.16µj/token edge SLM decoder accelerator with scal- able ring-based configuration for token-level pipelining in 16 nm FinFET
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 55d2b65a-71c2-48a7-9a95-acf290cf1dec · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models Qwen2.5 Technical Report
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 13e9ea31-424f-4830-9936-a5afff4f09ce · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models Thomas, Rom N
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation b72d3640-b9cd-4966-a62f-d20ea2cf0a40 · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models Simultaneous planning and execution for quadro- tors flying through a narrow gap under disturbance
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation cb5ea27a-354b-4116-89ce-5afc2f4f6021 · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models Attention Is All You Need
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation bbf602cb-026f-49af-bef0-dc48d8b4bcf8 · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models HAT: Hardware-Aware Transformers for Efficient Natural Language Processing
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 5cacdc25-ad92-4f2d-9b66-a47112dea8fd · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models Crowdsourcing Multiple Choice Science Questions
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 0985eea9-059f-4604-ab3f-cc11231b0d6b · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models Conformer-based speech recognition on extreme edge-computing devices
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation b562e77b-c3ae-4b49-8c86-20b44dff5d3f · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models Zeus: Understanding and optimizing GPU energy consumption of DNN training
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 2cc53c94-ce99-4f45-9307-fff09a0f0ac0 · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models HellaSwag: Can a Machine Really Finish Your Sentence?
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 22d69211-ca82-44fe-a18e-d700c1bd060d · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 53ac1877-e907-420c-9363-ac986d39475b · outbound
LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models MAC precision
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
No inbound Pith citation observations are available.