Pith. sign in

Paper Citation Record · LEDGER

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement

As of 12 August 2026, this Paper Citation Record lists 100 of 147 outbound references and 2 inbound Pith citation observations for arXiv:2412.06845.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.06845 v6

Coverage vector

measured 100 of 147 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T20:27:18.375394Z

measured 102 of 102 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-20T07:28:20.248452Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T07:33:07.543757Z

Reference resolution

100 of 147 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved99
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9cbc60f7-5b3e-44b4-8633-1a7323d305a5 · outbound

This paper cites GPT-4 Technical Report.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.908700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.908700Z digest=sha256:0627064eb0ceede2af2427e398f0a55635677bae9e2e392700b1835d4a68334d

Observation f9025085-c100-493b-b23c-4e41eb1fffc7 · outbound

This paper cites The claude 3 model family: Opus, sonnet, haiku.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The claude 3 model family: Opus, sonnet, haiku

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.913900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.913900Z digest=sha256:bea6e761b361a2e6190c98ac94c57effdb66043d824428251b3f28f93d03ce19

Observation 7d826f12-3d0b-4108-bc29-4d7e2e867748 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Gemini: A Family of Highly Capable Multimodal Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.917864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.917864Z digest=sha256:aee4186b0a4f3ba840ce3cd36f90e2eaa6dba20f50dae1bbc1f7ddbe32ea2009

Observation 10ed27de-30c5-4417-88a0-505435db5fb8 · outbound

This paper cites The Llama 3 Herd of Models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The Llama 3 Herd of Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.922529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.922529Z digest=sha256:34c6116fe829a3bc1b16f53ae7e622dd9b1a6f15dece5039a6604f3785529e31

Observation 59d72586-8095-46da-9571-cc457a7d749f · outbound

This paper cites an unresolved cited work.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.926903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.926903Z digest=sha256:13d45ce9c5c3d4c14265cba9b480bb5e6b5db53df9c8775872dbb823629d529d

Observation d39dd68f-bab5-4011-b266-6092d5b56fb3 · outbound

This paper cites Mistral 7B.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Mistral 7B

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.931276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.931276Z digest=sha256:a737de479ada9519d54027b563fc77fa12be3c66dc4d360498ea5ca93996313c

Observation 0e7dc88a-ffcf-46c6-a8f3-42447b5e2d04 · outbound

This paper cites The Foundation Model Transparency Index.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The Foundation Model Transparency Index

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.936174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.936174Z digest=sha256:c64876388779c5657418e6c5b7004d8437d7e7ab2ed64949fa17fb0ba60c7e20

Observation ec3858cf-c5b1-48e2-b6a7-c71243e7a492 · outbound

This paper cites On the Societal Impact of Open Foundation Models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement On the Societal Impact of Open Foundation Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.941292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.941292Z digest=sha256:ccccd11124317b8bfcbfc3e4d02636919a565b9734fad33a53f0ed2c7b3258b5

Observation ba54148c-97a0-4644-892b-25d38fb8fca1 · outbound

This paper cites The Model Openness Framework: Promoting Completeness and Openness for Reproducibility, Transparency, and Usability in Artificial Intelligence.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The Model Openness Framework: Promoting Completeness and Openness for Reproducibility, Transparency, and Usability in Artificial Intelligence

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.946278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.946278Z digest=sha256:25de98cbf94505742e4998079ade505f94aae95e06c600cb8f93e317f08b3ea7

Observation 9c83d35f-f71a-4a7c-82e8-248a2b1f7bcf · outbound

This paper cites Qwen Technical Report.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Qwen Technical Report

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.950661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.950661Z digest=sha256:d62f5f758e175b00618d738c77e70555b89dfeeb0b760d8564f979c6d61d0f54

Observation b66b9e42-7f90-4201-a427-1221c33f8bfd · outbound

This paper cites Qwen2 Technical Report.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Qwen2 Technical Report

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.955582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.955582Z digest=sha256:d8f8f29b6615a30274a8e379e7b31de3240e843a559091f0a69865b48ae40e65

Observation 4b706cb5-2ab7-4e62-b10e-82f0109fbff9 · outbound

This paper cites Federated Knowledge Distillation for Multi-Model Architectures Lithography Hotspot Detection.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Federated Knowledge Distillation for Multi-Model Architectures Lithography Hotspot Detection

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.960659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.960659Z digest=sha256:16335e70d6d546b21ea7195d5063c92378d654e114afd749a08c85ebd240d6f8

Observation 959ccb6b-a24b-49e1-b2ad-e26ec2abd7c5 · outbound

This paper cites Advancing model pruning via bi-level optimization.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Advancing model pruning via bi-level optimization

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.971474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.971474Z digest=sha256:eeb45b54de89bbfbeb358d787df9165bb69345320f8b860c877230d01d0b1187

Observation 4326f0a1-abea-45fc-ae30-070bf34095b7 · outbound

This paper cites A generic layer pruning method for signal modulation recognition deep learning models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement A generic layer pruning method for signal modulation recognition deep learning models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.977240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.977240Z digest=sha256:cde0849cd1aa64d9445cdd4d534379ad7112c81997648e44fda290e4fe863d5e

Observation c9dee28d-ff30-4c0e-a0fe-cc5122462b92 · outbound

This paper cites Cross-layer graph knowledge distillation for image recognition.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Cross-layer graph knowledge distillation for image recognition

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.983689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.983689Z digest=sha256:f09c2e735456b04be61cad9915b1959731771753efe321cf93b2daeba2f20cab

Observation 6a9bef99-2617-4551-90e3-a2535d9c6b07 · outbound

This paper cites Peeling the onion: Hierarchical reduction of data redundancy for efficient vision transformer training.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Peeling the onion: Hierarchical reduction of data redundancy for efficient vision transformer training

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.988351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.988351Z digest=sha256:126d4157f3e856e0e8dd630ccd54ee8a5a01a3755ecf766810eec4531e900e21

Observation bbca38c9-3d9b-4605-a7d5-db71898fbbef · outbound

This paper cites Pruning foundation models for high accuracy without retraining.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Pruning foundation models for high accuracy without retraining

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.992180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.992180Z digest=sha256:691a511c6b8f8aeededc58250411d49ab81f26750451d6620bd2a1d9f757ca47

Observation 8c30e468-be81-41a1-8e39-76f2099b05ec · outbound

This paper cites Sparse learning for state space models on mobile.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Sparse learning for state space models on mobile

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.996883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.996883Z digest=sha256:8acae013b97da101f97fb1e49cc7b5f08b4e01b901cb95a5c4e95b345e0d3f7f

Observation ec0d0054-bb58-4d51-82b0-73d668d2164b · outbound

This paper cites Numerical pruning for efficient autoregressive models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Numerical pruning for efficient autoregressive models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.001431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.001431Z digest=sha256:daafdbb852f5b9edbf811ade496f9e67ab8791c439cf44f8174f33dda31acefb

Observation 7ed4b091-0e04-4683-a625-60f728b84202 · outbound

This paper cites Lazydit: Lazy learning for the acceleration of diffusion transformers.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Lazydit: Lazy learning for the acceleration of diffusion transformers

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.006447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.006447Z digest=sha256:7835f47724b5286ee5392e07f205d10e788f287962a00fe773c5f28e3ba08eb0

Observation d071bf78-c14c-4c64-93ae-0314752c9376 · outbound

This paper cites Toward adaptive large language models structured pruning via hybrid-grained weight importance assessment.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Toward adaptive large language models structured pruning via hybrid-grained weight importance assessment

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.011095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.011095Z digest=sha256:6c8a8cd32b784659b9356b62a702f53ae686e94390144ec2481d2baf9732a10f

Observation db35236b-feb4-4bfb-876d-72d7da0919dd · outbound

This paper cites Search for efficient large language models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Search for efficient large language models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.015275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.015275Z digest=sha256:b8ff9fed0b7f9e70b918e6f7bb0407327b2f8bb698a04cc2e4b3b523b1ee207b

Observation ac3ec23f-a669-4a1c-aa0a-de67d80a8e05 · outbound

This paper cites Pruning parameterization with bi-level optimization for efficient semantic segmentation on the edge.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Pruning parameterization with bi-level optimization for efficient semantic segmentation on the edge

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.020380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.020380Z digest=sha256:08a8df617fb8cd9aa258d1c3ddc4ae79c98284aa5136466ff0348ae71980e0e9

Observation a11a1879-0bbe-4ada-9530-a6752987ee55 · outbound

This paper cites COMAE: COMprehensive Attribute Exploration for Zero-shot Hashing.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement COMAE: COMprehensive Attribute Exploration for Zero-shot Hashing

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.025591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.025591Z digest=sha256:bb212d0dae7ff52d86b2d8774e92b20133395a2065d0a8b8d32998a7d9b0a9a0

Observation 836a6ad3-4c45-4f51-ad9f-eb9fb609cf32 · outbound

This paper cites Quartdepth: Post-training quantization for real-time depth estimation on the edge.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Quartdepth: Post-training quantization for real-time depth estimation on the edge

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.030590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.030590Z digest=sha256:df99922add5926e1708467739ddcf8c31ef0d5048a6030610b96ad26931edf3d

Observation 1ca5c2dc-61b6-40b8-8eec-023d89028509 · outbound

This paper cites Fast and memory-efficient video diffusion using streamlined inference.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Fast and memory-efficient video diffusion using streamlined inference

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.036705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.036705Z digest=sha256:05e9113891ce94d5f85ade7adfb53accfbbb27e0751f97a8853f0b7158a3a85d

Observation da6348e0-ea44-47de-86f4-2c0607a1c85b · outbound

This paper cites Compiler-aware neural architecture search for on- mobile real-time super-resolution.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Compiler-aware neural architecture search for on- mobile real-time super-resolution

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.041450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.041450Z digest=sha256:67dbd0f7ea9e06a16dca68302c4a155c0233840c15a3232551e532f69e225bce

Observation 929d0e0c-8283-478e-a49e-002c0d2ba333 · outbound

This paper cites Achieving on-mobile real-time super- resolution with neural architecture and pruning search.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Achieving on-mobile real-time super- resolution with neural architecture and pruning search

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.045575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.045575Z digest=sha256:0c38666e4efb5922529f3eafd6a44f5009830442e25ba040a7a08408b193db03

Observation 7264bf81-3b07-44dd-bf05-e01b7502d36d · outbound

This paper cites Towards real-time segmentation on the edge.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Towards real-time segmentation on the edge

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.050052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.050052Z digest=sha256:e31c9036b8600cb336d969a28d25e37b4aab408d13e37768a32795dd0a87ba45

Observation b344b1a0-2f93-4d9c-a25c-9ba8614f0daf · outbound

This paper cites Pruning-as-Search: Efficient Neural Architecture Search via Channel Pruning and Structural Reparameterization.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Pruning-as-Search: Efficient Neural Architecture Search via Channel Pruning and Structural Reparameterization

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.054852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.054852Z digest=sha256:f495f20a35903d11fa361c0eabaa5c07915dc829ca4b2fbafe12c0a75da98514

Observation 08e43dbf-ec4b-4690-940b-9eb3386b4d71 · outbound

This paper cites Exploring token pruning in vision state space models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Exploring token pruning in vision state space models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.060294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.060294Z digest=sha256:a3b033a8f69527641ede7aac6a41a9c084d86d6ecfa52c4ea12d6d65cce13d84

Observation dacccc1f-49ac-4d78-afed-0c74e58f5741 · outbound

This paper cites Rethinking token reduction for state space models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Rethinking token reduction for state space models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.064480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.064480Z digest=sha256:28b3c1721445a0f24e31c83490f64e36803307ba77ccfe969da97b739133e8d3

Observation b7bc951a-ed9f-4c89-b642-45ebfe351496 · outbound

This paper cites Spvit: Enabling faster vision transformers via latency-aware soft token pruning.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Spvit: Enabling faster vision transformers via latency-aware soft token pruning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.068696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.068696Z digest=sha256:9592254fcfbf328026a3852c5bf35d4d419612a8f661894737520170bc7faff5

Observation ab345c5d-4ba3-4cfe-8ce0-d16d64e0386c · outbound

This paper cites Efficient Reasoning with Hidden Thinking.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Efficient Reasoning with Hidden Thinking

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.073656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.073656Z digest=sha256:23d84889340fc87ce6cf7332fb72d348eb30cdd8b41a18da61267f3de6a10de1

Observation 3abd59c6-13e2-419f-9512-6fc54eaa7876 · outbound

This paper cites Enabling Flexible Multi-LLM Integration for Scalable Knowledge Aggregation.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Enabling Flexible Multi-LLM Integration for Scalable Knowledge Aggregation

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-11T20:27:19.658885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T20:27:18.078272Z digest=sha256:9cdabc784094af3a98636171e25fa02f2aa42d4970985df4ce5fec760e29f55b

Observation d008724c-d0cd-40ac-82ac-18c98ba5a5d4 · outbound

This paper cites FastCar: Cache Attentive Replay for Fast Auto-Regressive Video Generation on the Edge.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement FastCar: Cache Attentive Replay for Fast Auto-Regressive Video Generation on the Edge

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.082776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.082776Z digest=sha256:85ae9f5755414ac2b0f7157fa6604de9a8b5b0496fb5c85485cbc47de4eba7af

Observation 1ff82d57-1316-4469-8488-c44a09c98adc · outbound

This paper cites Taming Diffusion for Dataset Distillation with High Representativeness.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Taming Diffusion for Dataset Distillation with High Representativeness

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.088114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.088114Z digest=sha256:72306db2d840892919faf81a42dde897859696c675df19de9614af174062950a

Observation c9f477e5-0d08-4c4b-aa69-b690855049c3 · outbound

This paper cites Jamba: A Hybrid Transformer-Mamba Language Model.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Jamba: A Hybrid Transformer-Mamba Language Model

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.092971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.092971Z digest=sha256:fd36096534fc410d9afbe75537db17615307806f56526c258d306b2fdd753f6d

Observation b7abe382-b2de-44fa-8b08-62d13f66f576 · outbound

This paper cites Jamba-1.5: Hybrid Transformer-Mamba Models at Scale.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Jamba-1.5: Hybrid Transformer-Mamba Models at Scale

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.097840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.097840Z digest=sha256:ddde7a2d33e7f8c8d6a5a9a3c174afbd085b95c51dd69bfd4402939279f53134

Observation fb18644c-6456-4998-95bb-4e0dba7e4456 · outbound

This paper cites Neural Machine Translation of Rare Words with Subword Units.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Neural Machine Translation of Rare Words with Subword Units

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.103285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.103285Z digest=sha256:d043bdcee8191fd4f0d8890180e26adb790e6ad070c65d9ec8b3ac2d158af3a0

Observation ed8036f4-53ab-4c7b-bc3e-bfabe73f3b56 · outbound

This paper cites tiktoken, 2022.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement tiktoken, 2022

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.107814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.107814Z digest=sha256:63260e3119318df83f8c3a8823af9d02aeb7cc70476bb257dc5ace887c050d39

Observation 7d2336e0-a065-4959-a9cf-b398bcb0f0f1 · outbound

This paper cites SentencePiece: A simple and language independent subword tokenizer and detokenizer for Neural Text Processing.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement SentencePiece: A simple and language independent subword tokenizer and detokenizer for Neural Text Processing

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.112010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.112010Z digest=sha256:7464593cdfe47e6ce860eb95284ced9aebce66277d82a33f2784e042c9e53ea2

Observation 13db362d-04cd-4c7e-a853-44176c278e53 · outbound

This paper cites XLNet: Generalized Autoregressive Pretraining for Language Understanding.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement XLNet: Generalized Autoregressive Pretraining for Language Understanding

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.116490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.116490Z digest=sha256:acaae6b898771435e550bb686f84f020c20207957bc720ba199f6f0e45dfd803

Observation b0fe366e-4d9f-48f9-aac8-894e386fa23a · outbound

This paper cites Summary of the tokenizers.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Summary of the tokenizers

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.120711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.120711Z digest=sha256:5ddebea68b0fa201e02c2c974a20aaf856ace173d7dc111e38fa71cae4ca8adb

Observation 62274b73-20d3-4c2c-9611-9e441fe72d59 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.126245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.126245Z digest=sha256:812ed1bb6318014577f264f8d38667924f1fedff7088f28441440a06927f3e3a

Observation 9de82e34-6ec3-4ffe-b736-97a01dcc7c2c · outbound

This paper cites Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.130591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.130591Z digest=sha256:4e7da921025460e9534867bb8c214a3cce9e7830a3a95550667ade5fa6aedbdc

Observation 2946d674-5f97-4698-b98d-4e23395c60de · outbound

This paper cites Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.135247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.135247Z digest=sha256:a665ae5b55dde3f64f140419a0f0f225ab759edcb7db0d9d572388c483d2e632

Observation 966e4d78-04db-4a6a-b7e0-c018435e7efd · outbound

This paper cites MM-LLMs: Recent Advances in MultiModal Large Language Models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement MM-LLMs: Recent Advances in MultiModal Large Language Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.140032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.140032Z digest=sha256:5ac8e27b63311d45cbcc7875a133ceb8b8367f6eb638c8883f3972e26d3f02e6

Observation 5797e93d-968f-47a2-b402-e609358f855c · outbound

This paper cites Mixtral of Experts.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Mixtral of Experts

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.144550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.144550Z digest=sha256:79d7b8b2edd01f4c0930e5a5a6cd1d0ed53e050f4988dacf9bf2383ac65c7c35

Observation 8dd8301c-3c20-493a-bd46-97f7816b8fd2 · outbound

This paper cites Language Models are Few-Shot Learners.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Language Models are Few-Shot Learners

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.148931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.148931Z digest=sha256:fc613eef43da2009eb2abb2ed0c26b32f6863abf6aab8e38e6b9ff1ee820a80d

Observation cdb5b6d2-9fcb-493a-b2d7-ed5dc618b428 · outbound

This paper cites The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.153116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.153116Z digest=sha256:bb41cfd6ccbb9e4ad3567ca09b2a7334b071bf64c356557ec8177cb97d3a1a0a

Observation 19f4fd42-2c78-4445-80a8-eb53c669554f · outbound

This paper cites Scaling Language Models: Methods, Analysis & Insights from Training Gopher.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Scaling Language Models: Methods, Analysis & Insights from Training Gopher

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.157473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.157473Z digest=sha256:2b779d3aab90c7d16fa0636c7e3601672d2933821de844c377dcd6e3e8bc275c

Observation d74a08a5-22c9-4b86-8325-38d8a0183e85 · outbound

This paper cites CCNet: Extracting High Quality Monolingual Datasets from Web Crawl Data.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement CCNet: Extracting High Quality Monolingual Datasets from Web Crawl Data

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.161664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.161664Z digest=sha256:771e7d790b7f526f66f5ee3e834456e6fed54875ad8f5d8a968b22967259ece6

Observation 35e4e164-1993-4a38-9b69-bff2f948200c · outbound

This paper cites mT5: A massively multilingual pre-trained text-to-text transformer.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement mT5: A massively multilingual pre-trained text-to-text transformer

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.166288Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.166288Z digest=sha256:228aae5eb185fad89561673d976043724db2f4aaa95073bf1f78c721a9128756

Observation 5ed5ce60-564d-46d9-9747-a4faf0895b00 · outbound

This paper cites Exploring the limits of transfer learning with a unified text-to-text transformer.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Exploring the limits of transfer learning with a unified text-to-text transformer

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.170671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.170671Z digest=sha256:0e7162b646b1ab007f46e581b89b5c4be9f5262b9a11a5080849142264d49057

Observation 56e9ca47-cc5f-482b-9cea-4abce246ec51 · outbound

This paper cites Cross-lingual language model pretraining.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Cross-lingual language model pretraining

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.175245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.175245Z digest=sha256:5a4c93afdaaa7eaf7d27455091e8edf76db4a6a9888f9c3fbda8e7d52f7e97f5

Observation ce2a3189-ff29-4dad-9d65-2d513948b7a8 · outbound

This paper cites Evaluating Large Language Models Trained on Code.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Evaluating Large Language Models Trained on Code

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.179395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.179395Z digest=sha256:cae768218a6ec36e14ab734519ce2a185bea31a298263f01363d2d4b976d12f1

Observation f14ca8d5-1167-43f2-98e1-b7b94c07df9a · outbound

This paper cites The Pile: An 800GB Dataset of Diverse Text for Language Modeling.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The Pile: An 800GB Dataset of Diverse Text for Language Modeling

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.183453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.183453Z digest=sha256:4993e55e74e83ac2b9da61b02033558005e377fe54f6cb254b362e7260ac15a9

Observation ec3b47f7-da73-41fe-a6ce-030286bc9153 · outbound

This paper cites How to Train Data-Efficient LLMs.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement How to Train Data-Efficient LLMs

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.187919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.187919Z digest=sha256:e33ff4d088e0b6fe4538d1d2c43a7c55676242f85fbfe7fcd9588391788649d8

Observation 8e2b32a2-7a26-473f-8553-26cf3fb73d34 · outbound

This paper cites A Pretrainer's Guide to Training Data: Measuring the Effects of Data Age, Domain Coverage, Quality, & Toxicity.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement A Pretrainer's Guide to Training Data: Measuring the Effects of Data Age, Domain Coverage, Quality, & Toxicity

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.192841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.192841Z digest=sha256:6242a5de9f3174d2d6d41b873e70e39b98bd4c8a5131ec18a72e5f82989ce9ca

Observation 752f4963-cd4a-4c94-9156-ca9497cac1a6 · outbound

This paper cites Glam: Efficient scaling of language models with mixture-of-experts.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Glam: Efficient scaling of language models with mixture-of-experts

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.197752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.197752Z digest=sha256:d77587b6de6ef2eb11c076bd23598c6f14cad78a78e3f08e2b513640a6dd1289

Observation 02b4b7f7-1d41-430d-a3ad-41511d2bedf1 · outbound

This paper cites Deduplicating Training Data Makes Language Models Better.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Deduplicating Training Data Makes Language Models Better

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.202214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.202214Z digest=sha256:e93c594109d3ce274801541f3e6e92d3cb9e339dc9c56e9dd6ac1c8c44338e16

Observation c0174190-2c29-4481-aaf5-c239ae7031a1 · outbound

This paper cites Url normal- ization for de-duplication of web pages.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Url normal- ization for de-duplication of web pages

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.207533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.207533Z digest=sha256:cf709953dbb13f86563c57f060a9d9ed24524812ae50b5b368063e8d866f7438

Observation 93160288-adc3-4018-b45a-9c52bc69cee3 · outbound

This paper cites DataComp-LM: In search of the next generation of training sets for language models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement DataComp-LM: In search of the next generation of training sets for language models

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.212836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.212836Z digest=sha256:bf5d2b941e105acb6097519969ecd4676507b9a2e4c8964ebce33fbbc6f3eaaa

Observation f2dd4150-d875-4f87-a0d3-53cf7d538c23 · outbound

This paper cites Efficient online data mixing for language model pre-training.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Efficient online data mixing for language model pre-training

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.217617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.217617Z digest=sha256:1b757bf0c318cb597f6e3d810caee59ed25e1219d0658ce7f4b3c1c0b05f6771

Observation 2fdf681f-5d41-4609-807f-a96e95c0b2c1 · outbound

This paper cites SlimPajama-DC: Understanding Data Combinations for LLM Training.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement SlimPajama-DC: Understanding Data Combinations for LLM Training

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.221452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.221452Z digest=sha256:5f4256e3f105b98af42ab98739443a7f96bba52bfba711144d0745fe604749af

Observation bf06d7aa-c143-4e16-8ff2-bd72429e5d1c · outbound

This paper cites Gemma: Open Models Based on Gemini Research and Technology.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Gemma: Open Models Based on Gemini Research and Technology

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.225753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.225753Z digest=sha256:730fd1df4150637531ce56e829277c4fb967057b9bc18e498d8bb9e270d8215e

Observation bb1cfb66-2300-4402-a56b-9d649ed14911 · outbound

This paper cites Palm: Scaling language modeling with pathways.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Palm: Scaling language modeling with pathways

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.230402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.230402Z digest=sha256:30d00382178cc0ce71007f5a659240ad78bc10d7ad0cc834d5cce7c28109009a

Observation 37066c23-d6cc-420f-a195-ed135180da2a · outbound

This paper cites Dolma: an Open Corpus of Three Trillion Tokens for Language Model Pretraining Research.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Dolma: an Open Corpus of Three Trillion Tokens for Language Model Pretraining Research

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.234748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.234748Z digest=sha256:fbcc003773f0c25b752249e1c602ec3dff134a9fc8e27d489e8d4a9fa8af8dfc

Observation 06cd4c28-0fa0-4212-bf85-5c653126d384 · outbound

This paper cites The FineWeb Datasets: Decanting the Web for the Finest Text Data at Scale.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The FineWeb Datasets: Decanting the Web for the Finest Text Data at Scale

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.243855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.243855Z digest=sha256:ca01ed1aa36b0ca3bfd3abcfd379019f7b2438a89bbd7a744badc377cf7d6c89

Observation 0620b97a-ff06-4107-9b0f-07d1517fe1eb · outbound

This paper cites Llm-datasets: An open framework for pretraining datasets of large language models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Llm-datasets: An open framework for pretraining datasets of large language models

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.249188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.249188Z digest=sha256:d8358f15182fdfb4be33f933be7232e4342778e2dcfef48518742e2bf0f876d0

Observation de883fd8-b005-422b-8a3f-4e9b7727f7cc · outbound

This paper cites StarCoder 2 and The Stack v2: The Next Generation.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement StarCoder 2 and The Stack v2: The Next Generation

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.254010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.254010Z digest=sha256:056adcf19cab3a1eef53d4d55d1ea0020cfada26a7748320972eb7558d243dd7

Observation dff75c77-9380-4d3a-837e-7652698887c5 · outbound

This paper cites Tulu 3: Pushing Frontiers in Open Language Model Post-Training.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Tulu 3: Pushing Frontiers in Open Language Model Post-Training

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.258465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.258465Z digest=sha256:21b1d08a3854bd6e0e6e342696ed075a66f7690f69477b2347b0213ffa5168f1

Observation 2f26dc05-8c6c-4aa6-a2fb-f4eb91b8330f · outbound

This paper cites Infinity Instruct.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Infinity Instruct

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.262501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.262501Z digest=sha256:10340a477294dbba081c2bb95f55c49e6bb1408101d579c947ce43c928258952

Observation fa758c31-56f6-4f35-ac65-29d3e288cb1d · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.266679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.266679Z digest=sha256:7122f361a390f226adf8e98e2539f9ea0f0fcd8daddce524d170607b72b099b4

Observation 023814d8-389d-448b-84ee-33afc67d6a77 · outbound

This paper cites Open Thoughts.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Open Thoughts

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.271271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.271271Z digest=sha256:6541275e0041a101c1fdde49201970699f379057134bea2e20f52bb44d644bc7

Observation 28bba9e2-7f7e-4f33-9e1f-ffaae0274b79 · outbound

This paper cites OpenR1-Math-220k.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement OpenR1-Math-220k

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.275117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.275117Z digest=sha256:ef3fc75193a9a9bd53a26d930413b1319e429d2349f6b9e2af2dc649c749c380

Observation 3ebb2631-8eb9-4d77-b7c2-40d75dfa6765 · outbound

This paper cites Tang, Manan Roongta, Colin Cai, Jeffrey Luo, Tianjun Zhang, Li Erran Li, Raluca Ada Popa, and Ion Stoica.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Tang, Manan Roongta, Colin Cai, Jeffrey Luo, Tianjun Zhang, Li Erran Li, Raluca Ada Popa, and Ion Stoica

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.279135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.279135Z digest=sha256:0ae7f1f7f7b48f210b27dd8907cf4a50090ef25ba5d2726e6c5d50d251dcc682

Observation bc22fbdc-ae80-4eda-905d-b451cd60a763 · outbound

This paper cites Areal: Ant reasoning rl.https://github.com/inclusionAI/AReaL, 2025.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Areal: Ant reasoning rl.https://github.com/inclusionAI/AReaL, 2025

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.283265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.283265Z digest=sha256:b9c83a758e4522ea492ca8ceb9fe331071254f64276fa8fbdfffd0c13c272326

Observation 23c34f15-0efb-47fb-978c-a4d69d38317a · outbound

This paper cites ReaL: Efficient RLHF Training of Large Language Models with Parameter Reallocation.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement ReaL: Efficient RLHF Training of Large Language Models with Parameter Reallocation

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.287868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.287868Z digest=sha256:32c1fb0dc11242a9d1596c2a6bca066a2227ae37c06b3265f25efbd6d62a58ea

Observation 3d4264eb-14bd-4476-a0b6-568706d1e07b · outbound

This paper cites Gemini: A family of highly capable multimodal models, 2024.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Gemini: A family of highly capable multimodal models, 2024

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.292130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.292130Z digest=sha256:c283c199c62aa06bed4aa572cd7830283982e9344f2831e38a8ff53249d47d59

Observation 1000de85-5982-4acc-b44d-f827120fe87d · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.295971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.295971Z digest=sha256:3392ea83df10458b991f58c3ed844a61ff59412f1a601b22ce748fb2e3d89a6f

Observation d1079219-1573-4a92-b259-054ff5840111 · outbound

This paper cites DeepSeek-V3 Technical Report.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement DeepSeek-V3 Technical Report

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.301141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.301141Z digest=sha256:4614bf5c3929aa0b360ee8cc9dbe646c8e705f5b9d24e143698aae7d6a163866

Observation 393e8302-5398-4bde-9bba-66ca674bfbcd · outbound

This paper cites Baichuan 2: Open Large-scale Language Models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Baichuan 2: Open Large-scale Language Models

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.305532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.305532Z digest=sha256:d8f35f321e6781c7c09e11db4212d638817d37250f251ac3945a666b7fd776bb

Observation 77a95bc1-26f3-4888-8562-c411cbbd0479 · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.310711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.310711Z digest=sha256:19ac14a1c71e3747c2dfceb6a04992ec68758d3722f75202a10395ffa6272873

Observation b7ef42a6-fbdf-46e0-876e-e209d2d9e310 · outbound

This paper cites Pythia: A suite for analyzing large language models across training and scaling.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Pythia: A suite for analyzing large language models across training and scaling

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.316066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.316066Z digest=sha256:255808ba59cb5be9ae479edf4686027986a49540f474111d44d1191b49953efc

Observation 008db99f-adc3-4d05-a74f-15144c4ff43f · outbound

This paper cites GPT-NeoX-20B: An Open-Source Autoregressive Language Model.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement GPT-NeoX-20B: An Open-Source Autoregressive Language Model

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.321064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.321064Z digest=sha256:d603023f759d24ef1406a257c43c58ac12be7354bc0efe410802fff9a702a8e2

Observation f1c78cca-ad87-4ed1-8a20-8e592feeaa80 · outbound

This paper cites OLMo: Accelerating the Science of Language Models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement OLMo: Accelerating the Science of Language Models

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.325609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.325609Z digest=sha256:aec3335a01c24b66709d538e4f2effa67c50dac41698f860c98e246ca19d7d23

Observation 716216dd-7521-4f6d-b796-b0750c31e976 · outbound

This paper cites LLM360: Towards Fully Transparent Open-Source LLMs.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement LLM360: Towards Fully Transparent Open-Source LLMs

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.330173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.330173Z digest=sha256:670939afaafe8ae52d3d0e36b39125d470560563f63a870a7d9a8d2c2a9c091b

Observation 935359d0-c036-455e-8821-cb4cd580a93e · outbound

This paper cites Code Llama: Open Foundation Models for Code.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Code Llama: Open Foundation Models for Code

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.334503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.334503Z digest=sha256:7479bc32f34d3b1eab8e3af44c9dc2119a1198b88bae96a188ce53d8a11771c6

Observation 68d82c91-e340-4f6b-897e-961caa53a6da · outbound

This paper cites GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.338879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.338879Z digest=sha256:52a8b5a50de9d484481547ae11b872a9c8fb39a6e3b07af3013cac625dc43a47

Observation 4718b4d8-001f-4af4-969e-3465b1320517 · outbound

This paper cites Longformer: The Long-Document Transformer.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Longformer: The Long-Document Transformer

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.342788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.342788Z digest=sha256:112120a4ce500a069ccf88b5b824e00fbb151a54867d1c8434dc37af4856d661

Observation 7cebbe5d-20f3-4220-970e-e17274678b59 · outbound

This paper cites SlimPajama: A 627B token cleaned and deduplicated version of RedPajama.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement SlimPajama: A 627B token cleaned and deduplicated version of RedPajama

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.347059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.347059Z digest=sha256:140d897c3229ec7ace804244263f6625e92b950e3fa0dffb14407834261f3768

Observation c2d8eb90-4199-4e43-8dfc-40bd293a72de · outbound

This paper cites RedPajama: an Open Dataset for Training Large Language Models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement RedPajama: an Open Dataset for Training Large Language Models

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.350792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.350792Z digest=sha256:0e05c50f11917cac8e147c7d86e1caa63847c3b60f8cb8675f435d72ccd09755

Observation 8bdf83a3-2eb1-41cb-9444-fbdf07e88c10 · outbound

This paper cites SemDeDup: Data-efficient learning at web-scale through semantic deduplication.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement SemDeDup: Data-efficient learning at web-scale through semantic deduplication

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.355433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.355433Z digest=sha256:dc6ae9f4fa34e1f6bf78e1c4a7d8c77b8076a63d5d950a9e5847482db70f1fb8

Observation 311a437a-1c19-4b49-b689-4fbcbc90f178 · outbound

This paper cites an unresolved cited work.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Unresolved cited work

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.359443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.359443Z digest=sha256:d266801e783ae87b7698ab1f3f683a97c6b841ee4dd2ee968598b59ebc136ae2

Observation f2363aff-4900-4b7d-b92d-26ea6bcf49b0 · outbound

This paper cites The Curious Case of Neural Text Degeneration.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The Curious Case of Neural Text Degeneration

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.363229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.363229Z digest=sha256:ead024e10f297dd230f99b6b60da28cfd31a202090579f450a5e73aa381fe5c1

Observation 35079eac-4071-48cf-93c1-6f310ea1a06c · outbound

This paper cites Mining of massive datasets, cambridge university press, cambridge, 2014.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Mining of massive datasets, cambridge university press, cambridge, 2014

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.367268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.367268Z digest=sha256:5916e16ed158dd3b649b4bcd1980215dd23c32af3607d3ef1287d96c8478a153

Observation 9a76da83-df64-4bd0-8ebb-1f1acb6abffb · outbound

This paper cites Introduction to common crawl datasets.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Introduction to common crawl datasets

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.371325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.371325Z digest=sha256:b3a5345aefa7c7403fbf0fca1db68388eabae13259d671e82a01dbb05380bbac

Observation e0125241-7206-4192-a342-6b63db95f694 · outbound

This paper cites On the resemblance and containment of documents.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement On the resemblance and containment of documents

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.375394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.375394Z digest=sha256:0844736b698ae21b5286a32240fc470bf4cc4daba85ecc1ea852a88265eb6c19

Pith citing papers

Observation 6e2ddd1e-3c4d-49ab-bc90-28ed41baa203 · inbound

Human Cognition in Machines: A Unified Perspective of World Models cites this paper.

Human Cognition in Machines: A Unified Perspective of World Models 7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement

Reference 225

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T08:12:26.076291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-10T08:12:15.663761Z digest=sha256:542c120ae46943bacd57b45b0c6b02ffb00138eb025bfe31fbdd79c87e68d144

Observation bde30324-447d-45fa-b538-054dd951d0a9 · inbound

PhyWorld: Physics-Faithful World Model for Video Generation cites this paper.

PhyWorld: Physics-Faithful World Model for Video Generation 7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:33:07.545335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:bc81ac043b156b651b1a8cbdcd315d0393c05d7d62828af2633a3f4f65023c1c