Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T20:27:18.375394Z
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 100 of 147 outbound references and 2 inbound Pith citation observations for arXiv:2412.06845.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T20:27:18.375394Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-20T07:28:20.248452Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-20T07:33:07.543757Z
100 of 147 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9cbc60f7-5b3e-44b4-8633-1a7323d305a5 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9025085-c100-493b-b23c-4e41eb1fffc7 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The claude 3 model family: Opus, sonnet, haiku
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d826f12-3d0b-4108-bc29-4d7e2e867748 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Gemini: A Family of Highly Capable Multimodal Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10ed27de-30c5-4417-88a0-505435db5fb8 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The Llama 3 Herd of Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59d72586-8095-46da-9571-cc457a7d749f · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Unresolved cited work
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d39dd68f-bab5-4011-b266-6092d5b56fb3 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Mistral 7B
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e7dc88a-ffcf-46c6-a8f3-42447b5e2d04 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The Foundation Model Transparency Index
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec3858cf-c5b1-48e2-b6a7-c71243e7a492 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement On the Societal Impact of Open Foundation Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba54148c-97a0-4644-892b-25d38fb8fca1 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The Model Openness Framework: Promoting Completeness and Openness for Reproducibility, Transparency, and Usability in Artificial Intelligence
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c83d35f-f71a-4a7c-82e8-248a2b1f7bcf · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Qwen Technical Report
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b66b9e42-7f90-4201-a427-1221c33f8bfd · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Qwen2 Technical Report
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b706cb5-2ab7-4e62-b10e-82f0109fbff9 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Federated Knowledge Distillation for Multi-Model Architectures Lithography Hotspot Detection
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 959ccb6b-a24b-49e1-b2ad-e26ec2abd7c5 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Advancing model pruning via bi-level optimization
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4326f0a1-abea-45fc-ae30-070bf34095b7 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement A generic layer pruning method for signal modulation recognition deep learning models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9dee28d-ff30-4c0e-a0fe-cc5122462b92 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Cross-layer graph knowledge distillation for image recognition
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a9bef99-2617-4551-90e3-a2535d9c6b07 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Peeling the onion: Hierarchical reduction of data redundancy for efficient vision transformer training
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bbca38c9-3d9b-4605-a7d5-db71898fbbef · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Pruning foundation models for high accuracy without retraining
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c30e468-be81-41a1-8e39-76f2099b05ec · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Sparse learning for state space models on mobile
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec0d0054-bb58-4d51-82b0-73d668d2164b · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Numerical pruning for efficient autoregressive models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ed4b091-0e04-4683-a625-60f728b84202 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Lazydit: Lazy learning for the acceleration of diffusion transformers
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d071bf78-c14c-4c64-93ae-0314752c9376 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Toward adaptive large language models structured pruning via hybrid-grained weight importance assessment
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db35236b-feb4-4bfb-876d-72d7da0919dd · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Search for efficient large language models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac3ec23f-a669-4a1c-aa0a-de67d80a8e05 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Pruning parameterization with bi-level optimization for efficient semantic segmentation on the edge
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a11a1879-0bbe-4ada-9530-a6752987ee55 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement COMAE: COMprehensive Attribute Exploration for Zero-shot Hashing
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 836a6ad3-4c45-4f51-ad9f-eb9fb609cf32 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Quartdepth: Post-training quantization for real-time depth estimation on the edge
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ca5c2dc-61b6-40b8-8eec-023d89028509 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Fast and memory-efficient video diffusion using streamlined inference
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da6348e0-ea44-47de-86f4-2c0607a1c85b · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Compiler-aware neural architecture search for on- mobile real-time super-resolution
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 929d0e0c-8283-478e-a49e-002c0d2ba333 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Achieving on-mobile real-time super- resolution with neural architecture and pruning search
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7264bf81-3b07-44dd-bf05-e01b7502d36d · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Towards real-time segmentation on the edge
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b344b1a0-2f93-4d9c-a25c-9ba8614f0daf · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Pruning-as-Search: Efficient Neural Architecture Search via Channel Pruning and Structural Reparameterization
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08e43dbf-ec4b-4690-940b-9eb3386b4d71 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Exploring token pruning in vision state space models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dacccc1f-49ac-4d78-afed-0c74e58f5741 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Rethinking token reduction for state space models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7bc951a-ed9f-4c89-b642-45ebfe351496 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Spvit: Enabling faster vision transformers via latency-aware soft token pruning
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab345c5d-4ba3-4cfe-8ce0-d16d64e0386c · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Efficient Reasoning with Hidden Thinking
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3abd59c6-13e2-419f-9512-6fc54eaa7876 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Enabling Flexible Multi-LLM Integration for Scalable Knowledge Aggregation
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d008724c-d0cd-40ac-82ac-18c98ba5a5d4 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement FastCar: Cache Attentive Replay for Fast Auto-Regressive Video Generation on the Edge
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ff82d57-1316-4469-8488-c44a09c98adc · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Taming Diffusion for Dataset Distillation with High Representativeness
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9f477e5-0d08-4c4b-aa69-b690855049c3 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Jamba: A Hybrid Transformer-Mamba Language Model
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7abe382-b2de-44fa-8b08-62d13f66f576 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Jamba-1.5: Hybrid Transformer-Mamba Models at Scale
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb18644c-6456-4998-95bb-4e0dba7e4456 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Neural Machine Translation of Rare Words with Subword Units
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed8036f4-53ab-4c7b-bc3e-bfabe73f3b56 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement tiktoken, 2022
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d2336e0-a065-4959-a9cf-b398bcb0f0f1 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement SentencePiece: A simple and language independent subword tokenizer and detokenizer for Neural Text Processing
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13db362d-04cd-4c7e-a853-44176c278e53 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement XLNet: Generalized Autoregressive Pretraining for Language Understanding
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0fe366e-4d9f-48f9-aac8-894e386fa23a · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Summary of the tokenizers
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62274b73-20d3-4c2c-9611-9e441fe72d59 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9de82e34-6ec3-4ffe-b736-97a01dcc7c2c · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2946d674-5f97-4698-b98d-4e23395c60de · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 966e4d78-04db-4a6a-b7e0-c018435e7efd · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement MM-LLMs: Recent Advances in MultiModal Large Language Models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5797e93d-968f-47a2-b402-e609358f855c · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Mixtral of Experts
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8dd8301c-3c20-493a-bd46-97f7816b8fd2 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Language Models are Few-Shot Learners
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cdb5b6d2-9fcb-493a-b2d7-ed5dc618b428 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19f4fd42-2c78-4445-80a8-eb53c669554f · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Scaling Language Models: Methods, Analysis & Insights from Training Gopher
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d74a08a5-22c9-4b86-8325-38d8a0183e85 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement CCNet: Extracting High Quality Monolingual Datasets from Web Crawl Data
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35e4e164-1993-4a38-9b69-bff2f948200c · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement mT5: A massively multilingual pre-trained text-to-text transformer
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ed5ce60-564d-46d9-9747-a4faf0895b00 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Exploring the limits of transfer learning with a unified text-to-text transformer
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56e9ca47-cc5f-482b-9cea-4abce246ec51 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Cross-lingual language model pretraining
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce2a3189-ff29-4dad-9d65-2d513948b7a8 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Evaluating Large Language Models Trained on Code
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f14ca8d5-1167-43f2-98e1-b7b94c07df9a · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The Pile: An 800GB Dataset of Diverse Text for Language Modeling
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec3b47f7-da73-41fe-a6ce-030286bc9153 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement How to Train Data-Efficient LLMs
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e2b32a2-7a26-473f-8553-26cf3fb73d34 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement A Pretrainer's Guide to Training Data: Measuring the Effects of Data Age, Domain Coverage, Quality, & Toxicity
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 752f4963-cd4a-4c94-9156-ca9497cac1a6 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Glam: Efficient scaling of language models with mixture-of-experts
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02b4b7f7-1d41-430d-a3ad-41511d2bedf1 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Deduplicating Training Data Makes Language Models Better
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0174190-2c29-4481-aaf5-c239ae7031a1 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Url normal- ization for de-duplication of web pages
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93160288-adc3-4018-b45a-9c52bc69cee3 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement DataComp-LM: In search of the next generation of training sets for language models
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2dd4150-d875-4f87-a0d3-53cf7d538c23 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Efficient online data mixing for language model pre-training
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2fdf681f-5d41-4609-807f-a96e95c0b2c1 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement SlimPajama-DC: Understanding Data Combinations for LLM Training
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf06d7aa-c143-4e16-8ff2-bd72429e5d1c · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Gemma: Open Models Based on Gemini Research and Technology
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb1cfb66-2300-4402-a56b-9d649ed14911 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Palm: Scaling language modeling with pathways
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37066c23-d6cc-420f-a195-ed135180da2a · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Dolma: an Open Corpus of Three Trillion Tokens for Language Model Pretraining Research
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06cd4c28-0fa0-4212-bf85-5c653126d384 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The FineWeb Datasets: Decanting the Web for the Finest Text Data at Scale
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0620b97a-ff06-4107-9b0f-07d1517fe1eb · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Llm-datasets: An open framework for pretraining datasets of large language models
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de883fd8-b005-422b-8a3f-4e9b7727f7cc · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement StarCoder 2 and The Stack v2: The Next Generation
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dff75c77-9380-4d3a-837e-7652698887c5 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Tulu 3: Pushing Frontiers in Open Language Model Post-Training
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f26dc05-8c6c-4aa6-a2fb-f4eb91b8330f · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Infinity Instruct
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa758c31-56f6-4f35-ac65-29d3e288cb1d · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 023814d8-389d-448b-84ee-33afc67d6a77 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Open Thoughts
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28bba9e2-7f7e-4f33-9e1f-ffaae0274b79 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement OpenR1-Math-220k
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ebb2631-8eb9-4d77-b7c2-40d75dfa6765 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Tang, Manan Roongta, Colin Cai, Jeffrey Luo, Tianjun Zhang, Li Erran Li, Raluca Ada Popa, and Ion Stoica
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc22fbdc-ae80-4eda-905d-b451cd60a763 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Areal: Ant reasoning rl.https://github.com/inclusionAI/AReaL, 2025
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23c34f15-0efb-47fb-978c-a4d69d38317a · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement ReaL: Efficient RLHF Training of Large Language Models with Parameter Reallocation
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d4264eb-14bd-4476-a0b6-568706d1e07b · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Gemini: A family of highly capable multimodal models, 2024
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1000de85-5982-4acc-b44d-f827120fe87d · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1079219-1573-4a92-b259-054ff5840111 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement DeepSeek-V3 Technical Report
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 393e8302-5398-4bde-9bba-66ca674bfbcd · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Baichuan 2: Open Large-scale Language Models
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77a95bc1-26f3-4888-8562-c411cbbd0479 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7ef42a6-fbdf-46e0-876e-e209d2d9e310 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Pythia: A suite for analyzing large language models across training and scaling
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 008db99f-adc3-4d05-a74f-15144c4ff43f · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement GPT-NeoX-20B: An Open-Source Autoregressive Language Model
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1c78cca-ad87-4ed1-8a20-8e592feeaa80 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement OLMo: Accelerating the Science of Language Models
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 716216dd-7521-4f6d-b796-b0750c31e976 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement LLM360: Towards Fully Transparent Open-Source LLMs
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 935359d0-c036-455e-8821-cb4cd580a93e · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Code Llama: Open Foundation Models for Code
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68d82c91-e340-4f6b-897e-961caa53a6da · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4718b4d8-001f-4af4-969e-3465b1320517 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Longformer: The Long-Document Transformer
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7cebbe5d-20f3-4220-970e-e17274678b59 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement SlimPajama: A 627B token cleaned and deduplicated version of RedPajama
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2d8eb90-4199-4e43-8dfc-40bd293a72de · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement RedPajama: an Open Dataset for Training Large Language Models
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8bdf83a3-2eb1-41cb-9444-fbdf07e88c10 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement SemDeDup: Data-efficient learning at web-scale through semantic deduplication
Reference 95
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 311a437a-1c19-4b49-b689-4fbcbc90f178 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Unresolved cited work
Reference 96
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2363aff-4900-4b7d-b92d-26ea6bcf49b0 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The Curious Case of Neural Text Degeneration
Reference 97
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35079eac-4071-48cf-93c1-6f310ea1a06c · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Mining of massive datasets, cambridge university press, cambridge, 2014
Reference 98
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a76da83-df64-4bd0-8ebb-1f1acb6abffb · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Introduction to common crawl datasets
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0125241-7206-4192-a342-6b63db95f694 · outbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement On the resemblance and containment of documents
Reference 100
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e2ddd1e-3c4d-49ab-bc90-28ed41baa203 · inbound
Human Cognition in Machines: A Unified Perspective of World Models 7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement
Reference 225
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation bde30324-447d-45fa-b538-054dd951d0a9 · inbound
PhyWorld: Physics-Faithful World Model for Video Generation 7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.