Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-13T19:26:38.134505Z
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 97 of 97 outbound references and 4 inbound Pith citation observations for arXiv:2604.03044.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-13T19:26:38.134505Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-30T20:34:28.858967Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-01T21:36:14.566726Z
97 of 97 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e83c1682-8d36-4c0a-a63f-31753a58a7cc · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency OckBench: Measuring the Efficiency of LLM Reasoning
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4f66751f-7304-402c-ac92-e7fcbbf8b419 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 93e4d977-8e23-4a82-877d-cf763e0a796a · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b2eb55c2-bba2-45a9-9d74-28cb90a250d1 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Glm-4.5: Agentic, reasoning, and coding (arc) foundation models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ff979b09-f5cf-44db-8f4b-3bfb5f548407 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Qwen3-30b-a3b-instruct-2507, July 2026
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 98a17ecd-674d-4e54-99cd-b2226ee2fbf6 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Qwen3.5: Towards native multimodal agents, February 2026
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 69d54879-575a-48e6-a76b-ad5a51509966 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Step 3.5 flash: Open frontier-level intelligence with 11b active parameters
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fe974a28-39a1-414c-9941-0281b6f30ed6 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency DeepSeek-V3 Technical Report
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6b85af3a-39d8-43ac-8061-03f12decd517 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Kimi K2: Open Agentic Intelligence
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9cc083a4-e118-4998-8c1d-a883d4d251cb · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Root mean square layer normalization.Advances in neural information processing systems, 32
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 28e3f31a-c824-4132-ac1a-452525ff1316 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Roformer: Enhanced transformer with rotary position embedding.Neurocomputing, 568:127063
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bd6a736e-7a48-4edb-894a-f55baccda2a2 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Language modeling with gated convolutional networks
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 418fd39c-5fca-493a-aaf1-0dc8058b07f8 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Auxiliary-Loss-Free Load Balancing Strategy for Mixture-of-Experts
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9cad0653-cd53-4f05-8826-0542236f9436 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Muon: An optimizer for hidden layers in neural networks
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c7966360-1ac0-4758-af69-58adf84ceb18 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency MiMo-V2-Flash Technical Report
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 972da0ce-6073-4176-8935-4c2782f4a341 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8154921d-0e94-4d9c-b126-e4ee9cc36926 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Gpipe: Efficient training of giant neural networks using pipeline parallelism
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 42b0c08b-b80d-45f4-80e8-d1159d7256dd · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Pipedream: Fast and efficient pipeline parallel dnn training, 2018.URL https://arxiv
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bff34d02-a177-41bc-a1d2-26a718365025 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Breadth-first pipeline parallelism.Proceedings of Machine Learning and Systems, 5:48–67
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c9ff4410-ec78-4d76-a7cf-0529d22d8e50 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Hanayo: Harnessing wave-like pipeline parallelism for enhanced large model training efficiency
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 93f488b1-1467-4b3b-9a97-762ad8f6ec8f · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Efficient large-scale language model training on gpu clusters using megatron-lm
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9659b4c1-ea9a-4909-a3ac-c9bf427da449 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Zero Bubble Pipeline Parallelism
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 848e5051-58a3-4ece-bc3b-77d246cf5b7d · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Moe a2a interleaved 1f1b based computation and communication overlap
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 92c9ce82-5dc8-4a87-8955-66f78233af0d · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency GShard: Scaling Giant Models with Conditional Computation and Automatic Sharding
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5226da3e-05bd-477f-a5d5-f58fc148a094 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Zero: Memory optimizations toward training trillion parameter models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e89543ce-c11c-4ba6-b89a-8f94848dfbf8 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Flashattention-3: Fast and accurate attention with asynchrony and low-precision
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 379bddea-c353-4f4b-b2b6-88d285d5677e · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Datatrove: large scale data processing
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 90ec33bf-3237-40d2-9dd3-d8080a2b4634 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Approximate nearest neighbors: towards removing the curse of dimensionality
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1ab5c66c-91ee-4fc7-b2a5-fb69c48647fe · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Unresolved cited work
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ee0a8965-cc29-43a2-997c-804cb71dc952 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Starcoder 2 and the stack v2: The next generation
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f60544ac-7343-4dcb-8878-cd56db694050 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Qwen2.5 technical report
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9de5321a-0d54-4b08-9c7e-106ee08bbaa7 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Qwen2.5-Coder Technical Report
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation eb14207a-35af-46e8-b16e-aa7ed85fe724 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Rewriting pre-training data boosts llm performance in math and code
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1d036699-43c5-4e00-93b4-96875a8cf73b · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Olmo 3
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a567d985-2a20-450f-b0f7-5593ac4e1f32 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Nemotron 3 nano: Open, efficient mixture-of-experts hybrid mamba-transformer model for agentic reasoning
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 65ae559f-c648-4c24-815b-4b1fc0c925bb · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Deepseek-v3.2: Pushing the frontier of open large language models
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7dcd0ee4-b1ec-41c8-aa6d-a3db65e5e0b2 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Mineru2.5: A decoupled vision-language model for efficient high-resolution document parsing
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7c13c86d-926f-418e-b5e1-4ff8bef4bac0 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency DeepSeek-OCR: Contexts Optical Compression
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 29ee9997-a61d-412b-8f2b-21f9923124c5 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Reformulation for pretraining data augmentation
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c997e95f-e109-4df0-929b-214d5a3e4b14 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Nemotron-CC: Transforming Common Crawl into a Refined Long-Horizon Pretraining Dataset
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a0763103-0681-4bd7-ba9f-d5a0c6fd8a4b · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 60263677-369e-4d7a-89be-6700e60e9741 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Scaling Language Models: Methods, Analysis & Insights from Training Gopher
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 65d3dcef-7067-4a40-91ec-515fde9960a4 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Training Compute-Optimal Large Language Models
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1d1c8e8f-0a25-47b4-b9fe-61ead7bc5b30 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Scaling Laws for Neural Language Models
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f2990fce-e9b6-43c1-bea3-f114d2464e68 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Measuring Massive Multitask Language Understanding
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fa1c82da-88a2-4b60-a72b-ae8aceb694a8 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Mmlu-pro: A more robust and challenging multi-task language understanding benchmark.Advances in Neural Information Processing Systems, 37:95266–95290
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d353b196-0472-4e19-9c8d-a6aa33394dfc · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Cmmlu: Measuring massive multitask language understanding in chinese
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 15fb487e-e9a3-4ed7-a50c-e9e1e4e9cca9 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Training Verifiers to Solve Math Word Problems
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e3447d75-9edf-4314-9d23-18fc3982e822 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Measuring Mathematical Problem Solving With the MATH Dataset
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6cb59b89-fd52-4f34-a8be-958382cf77c4 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Evaluating Large Language Models Trained on Code
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 954bd46f-e309-41b4-8146-38c3c1dee172 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5044b67b-61d5-4d53-ac7a-f1f532a04fb4 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency RULER: What's the Real Context Size of Your Long-Context Language Models?
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 651e46dd-4fa7-4967-a19c-e6c86dc40db0 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency gpt-oss-120b & gpt-oss-20b model card
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e174b87f-b2cf-443c-96f5-d09727f6e3b1 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Omniforce: On human-centered, large model empowered and cloud-edge collaborative automl system.nature npj-ai
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4ef1e9a5-a8c2-472f-a3dd-19ed9535be14 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency SWE-smith: Scaling Data for Software Engineering Agents
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3c7b5a22-57df-4f3d-8299-dc24fe95488f · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency OpenHands: An Open Platform for AI Software Developers as Generalist Agents
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 706d6149-697f-4249-afaf-198ff7e4e76b · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency SWE-agent: Agent-computer interfaces enable automated software engineering
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 60494e77-c67c-4305-87c9-b2a9eaa4cee8 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Openr1-math-220k dataset
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8ff34400-00d6-450f-8c2f-fd7cf1c03826 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Nemotron-math: Efficient long-context distillation of mathematical reasoning from multi-mode supervision.arXiv preprint arXiv:2512.15489
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation dce4edec-6677-43b2-b2e4-b7ed1ddbc7e8 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Constructing a multi-hop qa dataset for comprehensive evaluation of reasoning steps
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3b942de5-657b-4581-adb8-c299083209df · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Musique: Multihop questions via single-hop question composition.Transactions of the Association for Computational Linguistics, 10:539–554
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bbb17914-00d1-4098-87cc-9644e2e3d536 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Measuring and narrowing the compositionality gap in language models
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4b1d43f9-c86c-4653-a431-9f3582fddad2 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Measuring short-form factuality in large language models
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 03034d27-23cc-424f-9028-902ac8ebde34 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Fact, fetch, and reason: A unified evaluation of retrieval-augmented generation
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 57875e66-e792-4ee4-a13c-0eee940264df · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency ScholarSearch: Benchmarking Scholar Searching Ability of LLMs
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 604c728f-b890-484a-a792-ca93e1f7bf0b · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Gaia: a benchmark for general ai assistants
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 41b4775f-0556-4a06-aa1d-cce40b586e82 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency TaskCraft: Automated Generation of Agentic Tasks
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5197e854-9ed4-4d59-a3f3-d62636f1b7e1 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Nemotron-Post-Training-Dataset-v1
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0b26cd51-d574-4a1a-a41e-6a8931e3f91c · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Training language models to follow instructions with human feedback.Advances in Neural Information Processing Systems, 35:27730–27744
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5f838149-68e7-4e35-83cd-fd550252606f · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Proximal Policy Optimization Algorithms
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation eb85d3bd-28e8-49a8-82e8-a47b56189951 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9676e8dc-918b-47fb-a07f-1c080e3295aa · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency DAPO: An Open-Source LLM Reinforcement Learning System at Scale
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a46fe3b2-18be-42b9-b7da-571d1f016f8f · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Fibration policy optimization
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a043c26b-7be5-4fee-99b0-0cd43b0ddde4 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Trust region policy optimiza- tion
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 39de28e9-a99b-4bfc-a586-cee283a09178 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Group Sequence Policy Optimization
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 84f17b70-ad9c-4475-a2d5-4cd220319b3c · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency HybridFlow: A Flexible and Efficient RLHF Framework
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f425af20-3c31-4a73-830b-4b7f56787d9b · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency HellaSwag: Can a Machine Really Finish Your Sentence?
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2ba9e8bf-1dae-499c-89c3-ad748e5bb852 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency C-eval: A multi-level multi-discipline chinese evaluation suite for foundation models
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b433831b-aa5b-41c6-b029-625c8c43419f · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Gpqa: A graduate-level google-proof q&a benchmark
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 230b77a3-c35c-4177-929a-7eb511be975b · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Supergpqa: Scaling llm evaluation across 285 graduate disciplines
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a798dcf7-71b2-46b4-bfc7-cb4426de9185 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency SWE-bench: Can language models resolve real-world github issues? InThe Twelfth International Conference on Learning Representations
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 05757393-257b-41f5-b47b-5a9d4a50e55f · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency AlignBench: Benchmarking Chinese alignment of large language models
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 26b2e4e7-0a2b-4777-9fdc-4c63d3f16a39 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Instruction-Following Evaluation for Large Language Models
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b0006e1a-090a-4177-85dc-508b41a5e606 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Livebench: A challenging, contamination-free LLM benchmark
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation da2d2657-f63d-4d4a-93bf-ebf0f73871ee · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency τ 2-bench: Evaluating conversa- tional agents in a dual-control environment
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a0f10731-a3dc-403d-abe3-f17d54838fa7 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Understanding Straight-Through Estimator in Training Activation Quantized Neural Nets
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f21ee827-939b-4cca-9a4a-23056b139123 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Kimi-k2-thinking
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a8a97c3a-f1b3-47ff-b32b-6a4d465b1eab · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Glm-5: from vibe coding to agentic engineering
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 61f9d045-218c-4aaf-baae-fac389ca0c0f · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Model optimizer quantization support matrix
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 85e45fc9-7078-4338-a1d8-4da8dc0d09ce · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency vllm: Easy, fast, and cheap llm serving
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation dd8a22d0-3eff-45d7-af6e-a56f949931d6 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Tensorrt-llm
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 299a89e9-73d7-47ab-8d3a-cfb38ab41e8d · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Program Synthesis with Large Language Models
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 59fd732c-c9f9-4529-a1e2-afe290e567af · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Readme: Gguf, 9 2025
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 060cb2ad-3c0b-42d5-ba9e-e3e0f0309882 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Unlocking efficiency in large language model inference: A comprehensive survey of speculative decoding
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 70a68d73-822a-4b10-99ac-9c8759fe61d9 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Glm-5: from vibe coding to agentic engineering
Reference 95
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bcc4465d-5219-47c3-811b-4b72d96a1340 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Offline optimization of your disaggregated dynamo graph
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2c786606-936f-41f7-9486-acba5244a115 · outbound
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency Mooncake: A KVCache-centric Disaggregated Architecture for LLM Serving
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c8c4e32a-6565-4346-a612-2693255b91d2 · inbound
Universally Empowering Zeroth-Order Optimization via Adaptive Layer-wise Sampling JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1a2d61d0-5b94-497a-a259-be904b1c18fd · inbound
Irminsul: MLA-Native Position-Independent Caching for Agentic LLM Serving JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0d40cc4a-4f5f-44ac-a5f5-2b7cf2996fcc · inbound
Leyline: KV Cache Directives for Agentic Inference JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c84704a2-18b0-48fb-a694-996c7be275fe · inbound
MRCoder: An Efficient Context Selecting Approach for Repository-Level Code Generation JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.