Pith. sign in

Paper Citation Record · LEDGER

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement

As of 12 August 2026, this Paper Citation Record lists 100 of 147 outbound references and 2 inbound Pith citation observations for arXiv:2412.06845.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.06845 v6

Coverage vector

measured 100 of 147 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T20:27:18.375394Z

measured 102 of 102 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-20T07:28:20.248452Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T07:33:07.543757Z

Reference resolution

100 of 147 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved99
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9cbc60f7-5b3e-44b4-8633-1a7323d305a5 · outbound

This paper cites GPT-4 Technical Report.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.908700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.908700Z digest=sha256:cf2fb4d95b096294e6293bff1afe3cb951e5107a4d23c442f91880ec36f6e878

Observation f9025085-c100-493b-b23c-4e41eb1fffc7 · outbound

This paper cites The claude 3 model family: Opus, sonnet, haiku.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The claude 3 model family: Opus, sonnet, haiku

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.913900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.913900Z digest=sha256:efa90d58552ea52a8da860fdc5fb4dc6ee2174ff0dc73e4ecb05f30021c17864

Observation 7d826f12-3d0b-4108-bc29-4d7e2e867748 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Gemini: A Family of Highly Capable Multimodal Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.917864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.917864Z digest=sha256:87a0ba6e8cbc837bfa4ccd43fc8db1835eca731c5ef4b42866c122642da13d8f

Observation 10ed27de-30c5-4417-88a0-505435db5fb8 · outbound

This paper cites The Llama 3 Herd of Models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The Llama 3 Herd of Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.922529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.922529Z digest=sha256:849fb80492c1f1e5366681db2855427195ed318ffda34b8fe1f01532336488c8

Observation 59d72586-8095-46da-9571-cc457a7d749f · outbound

This paper cites an unresolved cited work.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.926903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.926903Z digest=sha256:d650794a2a5014d4748bb4da2fb8845e66cba90e378d617a4d0c925e80514177

Observation d39dd68f-bab5-4011-b266-6092d5b56fb3 · outbound

This paper cites Mistral 7B.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Mistral 7B

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.931276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.931276Z digest=sha256:7546ee1bc1ab18c922415021a46ce4f5b36b1a9e76e13b8c6a163c4a382697cb

Observation 0e7dc88a-ffcf-46c6-a8f3-42447b5e2d04 · outbound

This paper cites The Foundation Model Transparency Index.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The Foundation Model Transparency Index

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.936174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.936174Z digest=sha256:ea3327ae6f280aba87a1c50674fca82589110bfd78c036a1a49c2b207915884c

Observation ec3858cf-c5b1-48e2-b6a7-c71243e7a492 · outbound

This paper cites On the Societal Impact of Open Foundation Models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement On the Societal Impact of Open Foundation Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.941292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.941292Z digest=sha256:b4d832fdea4dded314d14e56638e56f69a6981eac622afef8d249db53edb54e4

Observation ba54148c-97a0-4644-892b-25d38fb8fca1 · outbound

This paper cites The Model Openness Framework: Promoting Completeness and Openness for Reproducibility, Transparency, and Usability in Artificial Intelligence.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The Model Openness Framework: Promoting Completeness and Openness for Reproducibility, Transparency, and Usability in Artificial Intelligence

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.946278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.946278Z digest=sha256:5eb27640b7245f9583b7f6991b82f49e328a6ab5d4c5e52c10cb88f161bed312

Observation 9c83d35f-f71a-4a7c-82e8-248a2b1f7bcf · outbound

This paper cites Qwen Technical Report.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Qwen Technical Report

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.950661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.950661Z digest=sha256:2b898007a931a99b944133d71fa3475ea2e5fc76dd24bb8d942def232378e5f0

Observation b66b9e42-7f90-4201-a427-1221c33f8bfd · outbound

This paper cites Qwen2 Technical Report.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Qwen2 Technical Report

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.955582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.955582Z digest=sha256:1831820d8c3e71b3ebcfc6a275b32aa529cae35754181d5291dc88b981dcdff5

Observation 4b706cb5-2ab7-4e62-b10e-82f0109fbff9 · outbound

This paper cites Federated Knowledge Distillation for Multi-Model Architectures Lithography Hotspot Detection.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Federated Knowledge Distillation for Multi-Model Architectures Lithography Hotspot Detection

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.960659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.960659Z digest=sha256:8bac439373bc4474d30d55451f17dd39d2e42858ba1730a433e07b5c8897becb

Observation 959ccb6b-a24b-49e1-b2ad-e26ec2abd7c5 · outbound

This paper cites Advancing model pruning via bi-level optimization.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Advancing model pruning via bi-level optimization

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.971474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.971474Z digest=sha256:9d421f6da3b384f956193788828b9b78ac8971092ac0bcc7ad7765f41a7f0297

Observation 4326f0a1-abea-45fc-ae30-070bf34095b7 · outbound

This paper cites A generic layer pruning method for signal modulation recognition deep learning models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement A generic layer pruning method for signal modulation recognition deep learning models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.977240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.977240Z digest=sha256:96265e6776eb8de0bcb596f3543cdfa710ddabf1c207f6881695e39459d692d5

Observation c9dee28d-ff30-4c0e-a0fe-cc5122462b92 · outbound

This paper cites Cross-layer graph knowledge distillation for image recognition.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Cross-layer graph knowledge distillation for image recognition

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.983689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.983689Z digest=sha256:e4e9ef4eca1d90913d9622579ae841e6fddd7cb024ff64fab0f6893cb8342e55

Observation 6a9bef99-2617-4551-90e3-a2535d9c6b07 · outbound

This paper cites Peeling the onion: Hierarchical reduction of data redundancy for efficient vision transformer training.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Peeling the onion: Hierarchical reduction of data redundancy for efficient vision transformer training

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.988351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.988351Z digest=sha256:f026a489c9f7bbefbca443fe31ac143111b1c143e987268fb5a29684540c7de9

Observation bbca38c9-3d9b-4605-a7d5-db71898fbbef · outbound

This paper cites Pruning foundation models for high accuracy without retraining.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Pruning foundation models for high accuracy without retraining

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.992180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.992180Z digest=sha256:6c1718ee95a4e1cf9d6a0cf5a4ae2807c1410f278d0f6a462a41435fd7cfbc40

Observation 8c30e468-be81-41a1-8e39-76f2099b05ec · outbound

This paper cites Sparse learning for state space models on mobile.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Sparse learning for state space models on mobile

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.996883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.996883Z digest=sha256:1968b511ed90928c1a8812ee7fbc7a245d8d1afd4260b2294f9c20227bf1161b

Observation ec0d0054-bb58-4d51-82b0-73d668d2164b · outbound

This paper cites Numerical pruning for efficient autoregressive models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Numerical pruning for efficient autoregressive models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.001431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.001431Z digest=sha256:269bef792295abdc597c88c33fa70cdd58cb7dd986f1fe3f5b49c2cb4ead80ef

Observation 7ed4b091-0e04-4683-a625-60f728b84202 · outbound

This paper cites Lazydit: Lazy learning for the acceleration of diffusion transformers.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Lazydit: Lazy learning for the acceleration of diffusion transformers

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.006447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.006447Z digest=sha256:e35aa66ad0e1fb8d34834b6b4e7a3a84e1c388af8cac00ae9b9b0990cab6b6b6

Observation d071bf78-c14c-4c64-93ae-0314752c9376 · outbound

This paper cites Toward adaptive large language models structured pruning via hybrid-grained weight importance assessment.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Toward adaptive large language models structured pruning via hybrid-grained weight importance assessment

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.011095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.011095Z digest=sha256:d252fd1e9824f41571da1db8f7de34d273a7d54d8aa715319aab89c19b901b85

Observation db35236b-feb4-4bfb-876d-72d7da0919dd · outbound

This paper cites Search for efficient large language models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Search for efficient large language models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.015275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.015275Z digest=sha256:b2d769b5ccabf4bdbadfb67771b5737771711bd7543482b747b3acd4c807d336

Observation ac3ec23f-a669-4a1c-aa0a-de67d80a8e05 · outbound

This paper cites Pruning parameterization with bi-level optimization for efficient semantic segmentation on the edge.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Pruning parameterization with bi-level optimization for efficient semantic segmentation on the edge

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.020380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.020380Z digest=sha256:3e8fd61038ad442e34db05ecfe3d4fb2ceb6907037d4a28087a1ae1c61f64e9f

Observation a11a1879-0bbe-4ada-9530-a6752987ee55 · outbound

This paper cites COMAE: COMprehensive Attribute Exploration for Zero-shot Hashing.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement COMAE: COMprehensive Attribute Exploration for Zero-shot Hashing

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.025591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.025591Z digest=sha256:e5c48c3384de3c31024179ff442dd026e12d8b7edd8cab6f53f8569a1f2ae31d

Observation 836a6ad3-4c45-4f51-ad9f-eb9fb609cf32 · outbound

This paper cites Quartdepth: Post-training quantization for real-time depth estimation on the edge.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Quartdepth: Post-training quantization for real-time depth estimation on the edge

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.030590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.030590Z digest=sha256:3dcc13663faa0b9742c0dde5b0c7878cf64e42da5a8938f43f9bcbcd9973aaa8

Observation 1ca5c2dc-61b6-40b8-8eec-023d89028509 · outbound

This paper cites Fast and memory-efficient video diffusion using streamlined inference.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Fast and memory-efficient video diffusion using streamlined inference

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.036705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.036705Z digest=sha256:0705ce75efcf95ddf118d1a2e267aa2846f49c28238e3223dc529cbbd817f5d9

Observation da6348e0-ea44-47de-86f4-2c0607a1c85b · outbound

This paper cites Compiler-aware neural architecture search for on- mobile real-time super-resolution.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Compiler-aware neural architecture search for on- mobile real-time super-resolution

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.041450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.041450Z digest=sha256:b84eca87f4e94c77a779e75c4d2cf3a75e2524d9683707091ed2c3fb1be1a9dc

Observation 929d0e0c-8283-478e-a49e-002c0d2ba333 · outbound

This paper cites Achieving on-mobile real-time super- resolution with neural architecture and pruning search.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Achieving on-mobile real-time super- resolution with neural architecture and pruning search

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.045575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.045575Z digest=sha256:a267e96187d25886ca7aec3b39658a63b2bf84cf973851a0241e579b27ec8bcb

Observation 7264bf81-3b07-44dd-bf05-e01b7502d36d · outbound

This paper cites Towards real-time segmentation on the edge.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Towards real-time segmentation on the edge

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.050052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.050052Z digest=sha256:06bfebff7bcaf0a10af705c9347ed4de986c3ef3124a5f3db79d44435e4c6d6e

Observation b344b1a0-2f93-4d9c-a25c-9ba8614f0daf · outbound

This paper cites Pruning-as-Search: Efficient Neural Architecture Search via Channel Pruning and Structural Reparameterization.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Pruning-as-Search: Efficient Neural Architecture Search via Channel Pruning and Structural Reparameterization

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.054852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.054852Z digest=sha256:9e16eb6f0c6faceb2d991fdcc0e8ab9537d73cefe38a08d0457ce6655d5ca576

Observation 08e43dbf-ec4b-4690-940b-9eb3386b4d71 · outbound

This paper cites Exploring token pruning in vision state space models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Exploring token pruning in vision state space models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.060294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.060294Z digest=sha256:797ea8cd0ab2d06f6b814668ef635929b2eb0e8e1b6e195da4a515f3108b5254

Observation dacccc1f-49ac-4d78-afed-0c74e58f5741 · outbound

This paper cites Rethinking token reduction for state space models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Rethinking token reduction for state space models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.064480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.064480Z digest=sha256:b660bb1b207e00632a7596b3da37fe2c0586480a78836111887173362a9a91d2

Observation b7bc951a-ed9f-4c89-b642-45ebfe351496 · outbound

This paper cites Spvit: Enabling faster vision transformers via latency-aware soft token pruning.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Spvit: Enabling faster vision transformers via latency-aware soft token pruning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.068696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.068696Z digest=sha256:cf475e172c63b10523f860212c4a190fe7c24608735db9a1f1535ffeee7fa846

Observation ab345c5d-4ba3-4cfe-8ce0-d16d64e0386c · outbound

This paper cites Efficient Reasoning with Hidden Thinking.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Efficient Reasoning with Hidden Thinking

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.073656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.073656Z digest=sha256:44ed43483511c78c67a1bea5d61e0871eead0ffd20a13d3d9184012b6c7d6dda

Observation 3abd59c6-13e2-419f-9512-6fc54eaa7876 · outbound

This paper cites Enabling Flexible Multi-LLM Integration for Scalable Knowledge Aggregation.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Enabling Flexible Multi-LLM Integration for Scalable Knowledge Aggregation

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-11T20:27:19.658885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T20:27:18.078272Z digest=sha256:2dca8e437ec73c7a658c97dbef264f7b2a48db44912c32a267e77dfb495cfee1

Observation d008724c-d0cd-40ac-82ac-18c98ba5a5d4 · outbound

This paper cites FastCar: Cache Attentive Replay for Fast Auto-Regressive Video Generation on the Edge.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement FastCar: Cache Attentive Replay for Fast Auto-Regressive Video Generation on the Edge

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.082776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.082776Z digest=sha256:75133c2b4d59f06c4c3a439b053983c891f313e15d93636ab95668c22b319c16

Observation 1ff82d57-1316-4469-8488-c44a09c98adc · outbound

This paper cites Taming Diffusion for Dataset Distillation with High Representativeness.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Taming Diffusion for Dataset Distillation with High Representativeness

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.088114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.088114Z digest=sha256:718a788ef830ec2c252440684a290f171b5ee232321a4d8a6f4eef4acf0e79f7

Observation c9f477e5-0d08-4c4b-aa69-b690855049c3 · outbound

This paper cites Jamba: A Hybrid Transformer-Mamba Language Model.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Jamba: A Hybrid Transformer-Mamba Language Model

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.092971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.092971Z digest=sha256:5ee9af882d326994dfdde1fbb5f1a9c98d29c4279ab0077423c3acfaf8c1b360

Observation b7abe382-b2de-44fa-8b08-62d13f66f576 · outbound

This paper cites Jamba-1.5: Hybrid Transformer-Mamba Models at Scale.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Jamba-1.5: Hybrid Transformer-Mamba Models at Scale

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.097840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.097840Z digest=sha256:a8f9e54e9e9b51451df532334ffbbf9c007ca2c4ad0755f317a7c991ddc84218

Observation fb18644c-6456-4998-95bb-4e0dba7e4456 · outbound

This paper cites Neural Machine Translation of Rare Words with Subword Units.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Neural Machine Translation of Rare Words with Subword Units

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.103285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.103285Z digest=sha256:f284c17f7acc606b0f4e790c3f40a093993aba8defc01e2d52291a5908c435f9

Observation ed8036f4-53ab-4c7b-bc3e-bfabe73f3b56 · outbound

This paper cites tiktoken, 2022.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement tiktoken, 2022

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.107814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.107814Z digest=sha256:213c1f2cf13ecb61d29adb8bcc576863b61946773101e25443cbe4a896a70794

Observation 7d2336e0-a065-4959-a9cf-b398bcb0f0f1 · outbound

This paper cites SentencePiece: A simple and language independent subword tokenizer and detokenizer for Neural Text Processing.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement SentencePiece: A simple and language independent subword tokenizer and detokenizer for Neural Text Processing

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.112010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.112010Z digest=sha256:451e5e1d2193a258a9156b5eb56cf907c3f4c845cef1ab1bc94e8d2674c09855

Observation 13db362d-04cd-4c7e-a853-44176c278e53 · outbound

This paper cites XLNet: Generalized Autoregressive Pretraining for Language Understanding.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement XLNet: Generalized Autoregressive Pretraining for Language Understanding

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.116490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.116490Z digest=sha256:ea880b2ef9cb1d8bd61761ef7cd62eefb4330c7500c97e3fbfeed2fe5cab989d

Observation b0fe366e-4d9f-48f9-aac8-894e386fa23a · outbound

This paper cites Summary of the tokenizers.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Summary of the tokenizers

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.120711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.120711Z digest=sha256:9df770e8d0e9128cba50afb275c5954480c09248a6edc41fc91b8b7bdddb8901

Observation 62274b73-20d3-4c2c-9611-9e441fe72d59 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.126245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.126245Z digest=sha256:a10363929f3adf3f0947d65dbd387201b04add81d4c68226adf6880913bec947

Observation 9de82e34-6ec3-4ffe-b736-97a01dcc7c2c · outbound

This paper cites Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.130591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.130591Z digest=sha256:180e3c665579046c113dbb688037d268ae866ee853b3ba9f822e28eabab7fdd6

Observation 2946d674-5f97-4698-b98d-4e23395c60de · outbound

This paper cites Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.135247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.135247Z digest=sha256:58cc73592cc61648d765e558c2c3a134878409e3fca2a4ba176e26894886dee2

Observation 966e4d78-04db-4a6a-b7e0-c018435e7efd · outbound

This paper cites MM-LLMs: Recent Advances in MultiModal Large Language Models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement MM-LLMs: Recent Advances in MultiModal Large Language Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.140032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.140032Z digest=sha256:e25c4ad075387e0d055dc4105a6faf6f52cf865765d253e70549332ffa846220

Observation 5797e93d-968f-47a2-b402-e609358f855c · outbound

This paper cites Mixtral of Experts.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Mixtral of Experts

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.144550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.144550Z digest=sha256:1b9406a248a57f1e5e8180df01ad8cbc4c37156fbce24e1441d275387733bbf5

Observation 8dd8301c-3c20-493a-bd46-97f7816b8fd2 · outbound

This paper cites Language Models are Few-Shot Learners.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Language Models are Few-Shot Learners

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.148931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.148931Z digest=sha256:ccd0b5a913c25020d8eb16bd780a6b731f6ed3c2e7882223d0683412649a3304

Observation cdb5b6d2-9fcb-493a-b2d7-ed5dc618b428 · outbound

This paper cites The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.153116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.153116Z digest=sha256:8e2e924f14d79eb0bb8471aa8f0dda1390224f0cddf09c9101d8c3968d110a93

Observation 19f4fd42-2c78-4445-80a8-eb53c669554f · outbound

This paper cites Scaling Language Models: Methods, Analysis & Insights from Training Gopher.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Scaling Language Models: Methods, Analysis & Insights from Training Gopher

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.157473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.157473Z digest=sha256:320553c2d446ffec8fa92e4c24cd7ab9a0c4408aefb8e53fe805fb55bea22b4a

Observation d74a08a5-22c9-4b86-8325-38d8a0183e85 · outbound

This paper cites CCNet: Extracting High Quality Monolingual Datasets from Web Crawl Data.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement CCNet: Extracting High Quality Monolingual Datasets from Web Crawl Data

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.161664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.161664Z digest=sha256:84c1ecef1c52d674aab49937cbf2974d70b3dd4fe5a4bf970b1da6017f4332e5

Observation 35e4e164-1993-4a38-9b69-bff2f948200c · outbound

This paper cites mT5: A massively multilingual pre-trained text-to-text transformer.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement mT5: A massively multilingual pre-trained text-to-text transformer

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.166288Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.166288Z digest=sha256:f1fd9a2e4694c3a0b694cd4752c6ffc002da2da6fd8b0a0f3b8443931104aa2a

Observation 5ed5ce60-564d-46d9-9747-a4faf0895b00 · outbound

This paper cites Exploring the limits of transfer learning with a unified text-to-text transformer.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Exploring the limits of transfer learning with a unified text-to-text transformer

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.170671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.170671Z digest=sha256:8fa955aa1b2f0882f0d04ed66e2e3deef26dd3103af4477cc656119f383cd6d0

Observation 56e9ca47-cc5f-482b-9cea-4abce246ec51 · outbound

This paper cites Cross-lingual language model pretraining.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Cross-lingual language model pretraining

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.175245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.175245Z digest=sha256:0f32212e1c12698a29024c9d67790f4eac92483f3f42760b2da12a128028df4b

Observation ce2a3189-ff29-4dad-9d65-2d513948b7a8 · outbound

This paper cites Evaluating Large Language Models Trained on Code.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Evaluating Large Language Models Trained on Code

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.179395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.179395Z digest=sha256:6e8c0ded040b5457245e8f7a7bb6b3ec01d09e9f630da6383188168094765de5

Observation f14ca8d5-1167-43f2-98e1-b7b94c07df9a · outbound

This paper cites The Pile: An 800GB Dataset of Diverse Text for Language Modeling.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The Pile: An 800GB Dataset of Diverse Text for Language Modeling

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.183453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.183453Z digest=sha256:3f7c8e70f516859eefdf3f0f98444c86c3a94529828f9518a61fa4774c3d466c

Observation ec3b47f7-da73-41fe-a6ce-030286bc9153 · outbound

This paper cites How to Train Data-Efficient LLMs.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement How to Train Data-Efficient LLMs

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.187919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.187919Z digest=sha256:80fbb075ce4268e0c96d36ef0bbe30a032be8ae70d3ed472fd0c2373b2540910

Observation 8e2b32a2-7a26-473f-8553-26cf3fb73d34 · outbound

This paper cites A Pretrainer's Guide to Training Data: Measuring the Effects of Data Age, Domain Coverage, Quality, & Toxicity.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement A Pretrainer's Guide to Training Data: Measuring the Effects of Data Age, Domain Coverage, Quality, & Toxicity

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.192841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.192841Z digest=sha256:707337d3078bd8fa1b0d56a650c1087e7e06d82f20ce1893de5bba8f132ef83b

Observation 752f4963-cd4a-4c94-9156-ca9497cac1a6 · outbound

This paper cites Glam: Efficient scaling of language models with mixture-of-experts.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Glam: Efficient scaling of language models with mixture-of-experts

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.197752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.197752Z digest=sha256:d2caa1eaafa77ce8aecd87fb760d686f620276837f3cee93eee06b12bab38737

Observation 02b4b7f7-1d41-430d-a3ad-41511d2bedf1 · outbound

This paper cites Deduplicating Training Data Makes Language Models Better.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Deduplicating Training Data Makes Language Models Better

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.202214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.202214Z digest=sha256:838d02c085c3acce558852a09b81197f0e8285c3376f696b518a12c9bc90b2bb

Observation c0174190-2c29-4481-aaf5-c239ae7031a1 · outbound

This paper cites Url normal- ization for de-duplication of web pages.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Url normal- ization for de-duplication of web pages

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.207533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.207533Z digest=sha256:2f4d9a97cec5a6a5ac2c01f7bd1671dc9409ce5bf561b5465398a09b404833e7

Observation 93160288-adc3-4018-b45a-9c52bc69cee3 · outbound

This paper cites DataComp-LM: In search of the next generation of training sets for language models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement DataComp-LM: In search of the next generation of training sets for language models

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.212836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.212836Z digest=sha256:3a056836b631dd3795720d7a4df6cc259ea4dd303dff3af3c945843ef3a333e5

Observation f2dd4150-d875-4f87-a0d3-53cf7d538c23 · outbound

This paper cites Efficient online data mixing for language model pre-training.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Efficient online data mixing for language model pre-training

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.217617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.217617Z digest=sha256:a2e83ac3fa6b22457ad2b19d6fed4c88fc52b69b1ad0810ee048c129843a225c

Observation 2fdf681f-5d41-4609-807f-a96e95c0b2c1 · outbound

This paper cites SlimPajama-DC: Understanding Data Combinations for LLM Training.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement SlimPajama-DC: Understanding Data Combinations for LLM Training

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.221452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.221452Z digest=sha256:4030f8a5126906179a291ed0dcb92295c079411375a57c89089dd662dd01dfe0

Observation bf06d7aa-c143-4e16-8ff2-bd72429e5d1c · outbound

This paper cites Gemma: Open Models Based on Gemini Research and Technology.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Gemma: Open Models Based on Gemini Research and Technology

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.225753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.225753Z digest=sha256:468198d0081d46750333866ea8c5663bd50b62202534d0c4d00148a4ebc608aa

Observation bb1cfb66-2300-4402-a56b-9d649ed14911 · outbound

This paper cites Palm: Scaling language modeling with pathways.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Palm: Scaling language modeling with pathways

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.230402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.230402Z digest=sha256:dbd83dc2194ba1a0b57a4414f13cefca9b995fc245099ad10dc86c4cf08a9701

Observation 37066c23-d6cc-420f-a195-ed135180da2a · outbound

This paper cites Dolma: an Open Corpus of Three Trillion Tokens for Language Model Pretraining Research.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Dolma: an Open Corpus of Three Trillion Tokens for Language Model Pretraining Research

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.234748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.234748Z digest=sha256:69f197325eddde1a6f7629191939b1bad35df4b3d4ce0695ceb0b2c7d40fcc26

Observation 06cd4c28-0fa0-4212-bf85-5c653126d384 · outbound

This paper cites The FineWeb Datasets: Decanting the Web for the Finest Text Data at Scale.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The FineWeb Datasets: Decanting the Web for the Finest Text Data at Scale

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.243855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.243855Z digest=sha256:fa23f447b4a56d865bbb5201a04b052db858eb98a6fcabf0964daec5672197f4

Observation 0620b97a-ff06-4107-9b0f-07d1517fe1eb · outbound

This paper cites Llm-datasets: An open framework for pretraining datasets of large language models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Llm-datasets: An open framework for pretraining datasets of large language models

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.249188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.249188Z digest=sha256:3be12d861b0e137680f0d70440d95d5c0bb6ff45c7c89d40fb3a5555818d621a

Observation de883fd8-b005-422b-8a3f-4e9b7727f7cc · outbound

This paper cites StarCoder 2 and The Stack v2: The Next Generation.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement StarCoder 2 and The Stack v2: The Next Generation

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.254010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.254010Z digest=sha256:68ad12b1ced025de9d5519c42510b92ae86d8badfc0ba360b556a16f30d10822

Observation dff75c77-9380-4d3a-837e-7652698887c5 · outbound

This paper cites Tulu 3: Pushing Frontiers in Open Language Model Post-Training.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Tulu 3: Pushing Frontiers in Open Language Model Post-Training

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.258465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.258465Z digest=sha256:3d7506cf9f4c5145f22ea33d3690895ce0919a1bf931ec263bb3aa9e809ebdf8

Observation 2f26dc05-8c6c-4aa6-a2fb-f4eb91b8330f · outbound

This paper cites Infinity Instruct.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Infinity Instruct

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.262501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.262501Z digest=sha256:2043267268173209c152de93407d20dc0dcf73d04cd9b31708a900aad436c4af

Observation fa758c31-56f6-4f35-ac65-29d3e288cb1d · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.266679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.266679Z digest=sha256:981131a7f13503efdce8dda3be45d776c07078cc2aa5381091f7b4746d3f710a

Observation 023814d8-389d-448b-84ee-33afc67d6a77 · outbound

This paper cites Open Thoughts.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Open Thoughts

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.271271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.271271Z digest=sha256:b4e82ce7b1feb624f781c16820126fe3254e629a6bd846faa7d7cf3bc5f74878

Observation 28bba9e2-7f7e-4f33-9e1f-ffaae0274b79 · outbound

This paper cites OpenR1-Math-220k.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement OpenR1-Math-220k

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.275117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.275117Z digest=sha256:f6e6b697fd62395f22d0f881fdd49a54b83ac4583d5eddac4c8830ca40845951

Observation 3ebb2631-8eb9-4d77-b7c2-40d75dfa6765 · outbound

This paper cites Tang, Manan Roongta, Colin Cai, Jeffrey Luo, Tianjun Zhang, Li Erran Li, Raluca Ada Popa, and Ion Stoica.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Tang, Manan Roongta, Colin Cai, Jeffrey Luo, Tianjun Zhang, Li Erran Li, Raluca Ada Popa, and Ion Stoica

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.279135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.279135Z digest=sha256:f79f7240682229c66e5a29201dd715642dfb21d87aedf8f7546a648455dcdf98

Observation bc22fbdc-ae80-4eda-905d-b451cd60a763 · outbound

This paper cites Areal: Ant reasoning rl.https://github.com/inclusionAI/AReaL, 2025.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Areal: Ant reasoning rl.https://github.com/inclusionAI/AReaL, 2025

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.283265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.283265Z digest=sha256:7cc01a87ceca8279b2fc514878707d84979490cdb3c3e1a5093c91bd55a24105

Observation 23c34f15-0efb-47fb-978c-a4d69d38317a · outbound

This paper cites ReaL: Efficient RLHF Training of Large Language Models with Parameter Reallocation.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement ReaL: Efficient RLHF Training of Large Language Models with Parameter Reallocation

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.287868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.287868Z digest=sha256:9ad8ca1525b4b9fa413ba4a5855646f564d87e56c2e88831c38dca297c3e706c

Observation 3d4264eb-14bd-4476-a0b6-568706d1e07b · outbound

This paper cites Gemini: A family of highly capable multimodal models, 2024.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Gemini: A family of highly capable multimodal models, 2024

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.292130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.292130Z digest=sha256:9dfb20063721816690ce905d4950d0859231ae745df6a94f5f8b9012e44eae36

Observation 1000de85-5982-4acc-b44d-f827120fe87d · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.295971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.295971Z digest=sha256:4aa5ae6f03bbc043b54f382ecff1363cefba212c111151595f874bf561c276e5

Observation d1079219-1573-4a92-b259-054ff5840111 · outbound

This paper cites DeepSeek-V3 Technical Report.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement DeepSeek-V3 Technical Report

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.301141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.301141Z digest=sha256:d66f3838f9b3b82f8c88d7b71a89abf79244f1fbf71e69a620fdc2b7e8e3ddb0

Observation 393e8302-5398-4bde-9bba-66ca674bfbcd · outbound

This paper cites Baichuan 2: Open Large-scale Language Models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Baichuan 2: Open Large-scale Language Models

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.305532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.305532Z digest=sha256:a50a8149de707d09498c35f5c7c3ddba08f68b8b34f39a43fb6d7c6ababa3d37

Observation 77a95bc1-26f3-4888-8562-c411cbbd0479 · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.310711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.310711Z digest=sha256:fb508aca8fb3e57c1a21c5b1a8841928296439096f94c384eec8e52cb1da2569

Observation b7ef42a6-fbdf-46e0-876e-e209d2d9e310 · outbound

This paper cites Pythia: A suite for analyzing large language models across training and scaling.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Pythia: A suite for analyzing large language models across training and scaling

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.316066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.316066Z digest=sha256:86e36515a7be1dc91826c99b8a9e976dea9eb08230fa9339674fdacf186f9d7d

Observation 008db99f-adc3-4d05-a74f-15144c4ff43f · outbound

This paper cites GPT-NeoX-20B: An Open-Source Autoregressive Language Model.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement GPT-NeoX-20B: An Open-Source Autoregressive Language Model

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.321064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.321064Z digest=sha256:72e70f9411febb0bfce9e38fa11d1a4f75f6a984cb19373f05f025986d6a0728

Observation f1c78cca-ad87-4ed1-8a20-8e592feeaa80 · outbound

This paper cites OLMo: Accelerating the Science of Language Models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement OLMo: Accelerating the Science of Language Models

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.325609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.325609Z digest=sha256:3daa3fc5dfc30127252160ab0134cb2e689b912f366df01e9db88d1e409812d6

Observation 716216dd-7521-4f6d-b796-b0750c31e976 · outbound

This paper cites LLM360: Towards Fully Transparent Open-Source LLMs.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement LLM360: Towards Fully Transparent Open-Source LLMs

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.330173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.330173Z digest=sha256:6741d4ea0cfa38e731e7793f96419363ad434ec0245d4c67cae49b64417d7ec5

Observation 935359d0-c036-455e-8821-cb4cd580a93e · outbound

This paper cites Code Llama: Open Foundation Models for Code.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Code Llama: Open Foundation Models for Code

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.334503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.334503Z digest=sha256:30cbd8c662dd52c6e146f4d55ecebca4b9fb47d4130820c89449a5db732b10f3

Observation 68d82c91-e340-4f6b-897e-961caa53a6da · outbound

This paper cites GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.338879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.338879Z digest=sha256:37e2ba3191ea0dfc5bb8bc3d91c75939b88377130c77e68724a405dc0365f3e1

Observation 4718b4d8-001f-4af4-969e-3465b1320517 · outbound

This paper cites Longformer: The Long-Document Transformer.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Longformer: The Long-Document Transformer

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.342788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.342788Z digest=sha256:35e14aee486e400ee190de038ba4b510a8d5f6aa814dda0121d5d115e5608f5a

Observation 7cebbe5d-20f3-4220-970e-e17274678b59 · outbound

This paper cites SlimPajama: A 627B token cleaned and deduplicated version of RedPajama.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement SlimPajama: A 627B token cleaned and deduplicated version of RedPajama

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.347059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.347059Z digest=sha256:66945fb2c1701dce83281ce82cc8c481dc1abc387c5595564b827f6a3742020f

Observation c2d8eb90-4199-4e43-8dfc-40bd293a72de · outbound

This paper cites RedPajama: an Open Dataset for Training Large Language Models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement RedPajama: an Open Dataset for Training Large Language Models

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.350792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.350792Z digest=sha256:3ecb3f4d39ad227d6d8a4a2e6662f7c17578bc637c37cfacbb0feec60e5ed827

Observation 8bdf83a3-2eb1-41cb-9444-fbdf07e88c10 · outbound

This paper cites SemDeDup: Data-efficient learning at web-scale through semantic deduplication.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement SemDeDup: Data-efficient learning at web-scale through semantic deduplication

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.355433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.355433Z digest=sha256:7fdb6e994ac906c266ac15cbdca1c6d9afd1590e43369c5c4eabff6ae23bcc03

Observation 311a437a-1c19-4b49-b689-4fbcbc90f178 · outbound

This paper cites an unresolved cited work.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Unresolved cited work

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.359443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.359443Z digest=sha256:7b593f7af6b3e890fec78f5cac3c2bf02dcbce8541d7e850f3f35e89f1f7e06c

Observation f2363aff-4900-4b7d-b92d-26ea6bcf49b0 · outbound

This paper cites The Curious Case of Neural Text Degeneration.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The Curious Case of Neural Text Degeneration

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.363229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.363229Z digest=sha256:a0f8acbcc70b6aca9b109acd566a76f13c61f82a338efa5f6d0a5e80242ec038

Observation 35079eac-4071-48cf-93c1-6f310ea1a06c · outbound

This paper cites Mining of massive datasets, cambridge university press, cambridge, 2014.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Mining of massive datasets, cambridge university press, cambridge, 2014

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.367268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.367268Z digest=sha256:3c53e210d8f805c3f3cdaf724bab0401b1eac3870f260c3f2af09e2e6e7cb254

Observation 9a76da83-df64-4bd0-8ebb-1f1acb6abffb · outbound

This paper cites Introduction to common crawl datasets.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Introduction to common crawl datasets

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.371325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.371325Z digest=sha256:842303d63dcee478d967425db73e0e18d21f62749f1052cbf981a4b39e9e71d7

Observation e0125241-7206-4192-a342-6b63db95f694 · outbound

This paper cites On the resemblance and containment of documents.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement On the resemblance and containment of documents

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.375394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.375394Z digest=sha256:22368b73f4a683aa57b46f975116cc402def8f698b58ef105e6cb4cc0817db8d

Pith citing papers

Observation 6e2ddd1e-3c4d-49ab-bc90-28ed41baa203 · inbound

Human Cognition in Machines: A Unified Perspective of World Models cites this paper.

Human Cognition in Machines: A Unified Perspective of World Models 7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement

Reference 225

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T08:12:26.076291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-10T08:12:15.663761Z digest=sha256:0ebf43bfc37bc215fb3e97a4901468ddc958c5aaae297c75023689e2e7c131e4

Observation bde30324-447d-45fa-b538-054dd951d0a9 · inbound

PhyWorld: Physics-Faithful World Model for Video Generation cites this paper.

PhyWorld: Physics-Faithful World Model for Video Generation 7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:33:07.545335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:7d216c8b83c3fc31446fbaa4554cfb505a9eb52facae713b708d567c5611bce3