Pith. sign in

Paper Citation Record · LEDGER

NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 48 inbound Pith citation observations for arXiv:2508.14444.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.14444 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 48 of 48 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T16:32:54.589013Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 4654259f-d9e2-41d8-aa2b-14305740f5ea · inbound

Reinforcement Learning from Human Feedback cites this paper.

Reinforcement Learning from Human Feedback NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 180

Resolution
verified exact
local_arxiv, observed 2026-05-22T19:32:00.956328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T19:27:40.991325Z digest=sha256:5ffb71689bd208214950ae874e3d9f771ff47757d0deaa8996f09ddc0b7939c9

Observation 38e6e254-9dc3-44c9-ab1a-77991ae6acf8 · inbound

Hermes 4 Technical Report cites this paper.

Hermes 4 Technical Report NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T16:32:54.589013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:32:54.589013Z digest=sha256:cfca21e8622fc481a07ed901748f4fd3c6571a0a2f76d38cf1c0d5a6e792b5b8

Observation e2c2fb03-3dec-4539-8903-cc570cddf64c · inbound

Decision Potential Surface: A Theoretical and Practical Approximation of Large Language Model Decision Boundary cites this paper.

Decision Potential Surface: A Theoretical and Practical Approximation of Large Language Model Decision Boundary NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 15

Resolution
metadata mismatch
local_arxiv, observed 2026-05-22T13:24:53.337281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T13:22:37.107679Z digest=sha256:8049cf803e6ab697bb7efda79f59f46f72da8a79401cfdb12ab66c08576bd6e6

Observation 0ee8f9de-ee9f-4629-a551-5e3a89a7a585 · inbound

Don't Pass@k: A Bayesian Framework for Large Language Model Evaluation cites this paper.

Don't Pass@k: A Bayesian Framework for Large Language Model Evaluation NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 101

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:02:08.433616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T10:04:39.223895Z digest=sha256:d57fdca4337319a5274eccd01eafec201f5ad780787950f4646afaff66b7a4d3

Observation c946fc06-286f-4caf-b2ba-635ecd109e0e · inbound

Hybrid Architectures for Language Models: Systematic Analysis and Design Insights cites this paper.

Hybrid Architectures for Language Models: Systematic Analysis and Design Insights NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:02:08.433616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T10:18:04.431436Z digest=sha256:a72f7f5b68053914cacbc4a8a47959a63527784f962bce5f7c4dc6435851c2db

Observation ea39139f-c862-43e6-8735-ebef10df9c6b · inbound

Are Large Reasoning Models Interruptible? cites this paper.

Are Large Reasoning Models Interruptible? NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T10:07:52.588242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T10:07:52.588242Z digest=sha256:edf4d2bcfe4245a1a64cd48f6de5b7cad9e02fb507980b19d1320d4b78b34a78

Observation 1cd05432-0e8e-40b6-928a-cf6eadd7817a · inbound

MiroThinker: Pushing the Performance Boundaries of Open-Source Research Agents via Model, Context, and Interactive Scaling cites this paper.

MiroThinker: Pushing the Performance Boundaries of Open-Source Research Agents via Model, Context, and Interactive Scaling NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:02:08.433616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-17T21:44:18.744201Z digest=sha256:09976c211db55c9c700e086379987d8ea171119a62c1aceb1f21a805fd0783f7

Observation 9f214e3c-ab9c-47e3-9bc7-8acec09677b4 · inbound

IndustryNav: Exploring Spatial Reasoning of Embodied Agents in Dynamic Industrial Navigation cites this paper.

IndustryNav: Exploring Spatial Reasoning of Embodied Agents in Dynamic Industrial Navigation NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-03T20:59:18.876864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T20:59:18.876864Z digest=sha256:c87c8d1a49f014492c3ef2f68f23f3f1ea02421cf569fc9271f7848736798f45

Observation 39204971-65bc-471a-9576-6443205d4771 · inbound

Four Over Six: More Accurate NVFP4 Quantization with Adaptive Block Scaling cites this paper.

Four Over Six: More Accurate NVFP4 Quantization with Adaptive Block Scaling NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:02:08.433616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-17T02:23:01.845123Z digest=sha256:dc04821e0f5c5610e7323df2b3ec1115e3378b85bbe21f9c6075a5edcda971c2

Observation 98cf83be-a03a-4e5f-b194-5958f09aaa7a · inbound

Efficient-DLM: From Autoregressive to Diffusion Language Models, and Beyond in Speed cites this paper.

Efficient-DLM: From Autoregressive to Diffusion Language Models, and Beyond in Speed NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:02:08.433616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T22:29:08.669964Z digest=sha256:8bde3a43af2e36bf74b10b9c1257b659558392079f993ce396dcd06dcae853d9

Observation 99b553f2-8178-49fd-ba05-8e57239343bc · inbound

NVIDIA Nemotron 3: Efficient and Open Intelligence cites this paper.

NVIDIA Nemotron 3: Efficient and Open Intelligence NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 204

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:02:08.433616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T01:40:42.190369Z digest=sha256:cc52b1cbdfbc98d0fcc3a411595f30d687ce6f1bf93872c4f95bb69bf8706c67

Observation 24a7f68f-aebf-4318-9e37-21fd3684192f · inbound

LinMU: Multimodal Understanding Made Linear cites this paper.

LinMU: Multimodal Understanding Made Linear NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:02:08.433616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T18:32:23.951566Z digest=sha256:f49993623823f09f3944ae9e91c9c4492a8914fb3d1b5851b08ae386686cffa9

Observation 7318a73e-5827-4c07-823e-f53a9f6e8a28 · inbound

Mid-Think: Training-Free Intermediate-Budget Reasoning via Token-Level Triggers cites this paper.

Mid-Think: Training-Free Intermediate-Budget Reasoning via Token-Level Triggers NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T11:16:00.402950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T11:16:00.402950Z digest=sha256:954e256c03637d90acc92614fa5656d13cadab407b51c8394aaa088030fc9781

Observation 59c6336d-f3f9-4baa-8a0b-aa8dc9d95f56 · inbound

NSR-Boost: A Neuro-Symbolic Residual Boosting Framework for Industrial Legacy Models cites this paper.

NSR-Boost: A Neuro-Symbolic Residual Boosting Framework for Industrial Legacy Models NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-03T10:21:12.407476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:21:12.407476Z digest=sha256:51f30c8c9188ca6db933c841388524b95ec1a24c4937617cca83d06d41f437f6

Observation 22420771-f063-4dfc-b94d-42c6650cc472 · inbound

HEARTS: Benchmarking LLM Reasoning on Health Time Series cites this paper.

HEARTS: Benchmarking LLM Reasoning on Health Time Series NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T21:02:24.871813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:02:24.871813Z digest=sha256:6544dac7a5a69b3b1e6a3ef325372403cdc1b44a3a9f5d5fc5a7a850ccca9aaa

Observation b3fdaec7-5c4c-42dc-bd21-64a7a799a4cd · inbound

Ranking Reasoning LLMs under Test-Time Scaling cites this paper.

Ranking Reasoning LLMs under Test-Time Scaling NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:02:08.433616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T13:33:18.767999Z digest=sha256:e6fd95f7cc9a1c849440fae35817d187a7cbe0db6721f65b98d73d0626cf3204

Observation 93bca724-4431-433e-a342-a8519f94dbe0 · inbound

M$^2$RNN: Non-Linear RNNs with Matrix-Valued States for Scalable Language Modeling cites this paper.

M$^2$RNN: Non-Linear RNNs with Matrix-Valued States for Scalable Language Modeling NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:02:08.433616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T11:35:43.088803Z digest=sha256:389d2b1065b53582f0e49273934d71959a2f3a8daa3ee758d9d46ca5def33c7c

Observation 5594f315-8912-4b5e-84f1-cdb508fbca8b · inbound

The limits of bio-molecular modeling with large language models : a cross-scale evaluation cites this paper.

The limits of bio-molecular modeling with large language models : a cross-scale evaluation NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T11:02:08.433616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T20:09:30.460861Z digest=sha256:54227b94d1418e3f0b5e0a279090e472be38366416ac44aeb71f93bbb71124e3

Observation 75beaa44-4886-496f-ba71-63fc0fc3747e · inbound

Nexus: Same Pretraining Loss, Better Downstream Generalization via Common Minima cites this paper.

Nexus: Same Pretraining Loss, Better Downstream Generalization via Common Minima NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:02:08.433616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T17:53:39.356675Z digest=sha256:5576afb5c39139af393660565000974ed8423e18f4c8408d70d0976fbd4ce34b

Observation d5fdeeda-6c51-40a4-8198-939dfdd6b7db · inbound

Multilinguality at the Edge: Developing Language Models for the Global South cites this paper.

Multilinguality at the Edge: Developing Language Models for the Global South NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T11:02:08.433616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-09T21:32:37.544226Z digest=sha256:19eb4490cf6a7d9323ad15017a9f92c123e694b731d68e04b17a03fe08aa215a

Observation a17824ea-0758-4938-a708-8f4afa56a0eb · inbound

Stochastic KV Routing: Enabling Adaptive Depth-Wise Cache Sharing cites this paper.

Stochastic KV Routing: Enabling Adaptive Depth-Wise Cache Sharing NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:02:08.433616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T19:56:48.015363Z digest=sha256:67f8c1b93ba753499706de58e54b1020e66020da161100e02f9a52c5ffcd7888

Observation a4f69ef1-ecbe-492b-a7e5-0cee6a44a11c · inbound

Star Elastic: Many-in-One Reasoning LLMs with Efficient Budget Control cites this paper.

Star Elastic: Many-in-One Reasoning LLMs with Efficient Budget Control NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:02:08.433616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-11T01:03:28.231935Z digest=sha256:97454b3ed95838c37d94b7712ba136198de03179df083e7ea2d23b90bacb6f7c

Observation 6f681267-1086-4cd4-bc16-65167b3d0426 · inbound

Priming: Hybrid State Space Models From Pre-trained Transformers cites this paper.

Priming: Hybrid State Space Models From Pre-trained Transformers NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:02:08.433616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-12T01:14:01.584159Z digest=sha256:6f4e7a83dd0f89b480ab48cbd05ba24e42355d74f261841bcd0e101e2054dc5c

Observation 30879c91-ab55-4361-9600-5976ffcef276 · inbound

PARD-2: Target-Aligned Parallel Draft Model for Dual-Mode Speculative Decoding cites this paper.

PARD-2: Target-Aligned Parallel Draft Model for Dual-Mode Speculative Decoding NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:02:08.433616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-12T01:08:14.247342Z digest=sha256:75741cd25c1b4ed406d86c94088285a1e72cce121c84040fe0840ca1f94ee421

Observation a72e98cb-3a6a-48fa-b4e2-b0169b2ef778 · inbound

A Recipe for Long-Context Reasoning in Large Language Models via On-Policy Optimization and Distillation cites this paper.

A Recipe for Long-Context Reasoning in Large Language Models via On-Policy Optimization and Distillation NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T11:02:08.433616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-13T05:29:00.576006Z digest=sha256:51cbbbc8c692f3f8187a08c4fc76777d5bd4793211a9087c07c59b8180239b30

Observation 6121148a-8c27-4327-ba66-efe3990e90b4 · inbound

Tracing the ongoing emergence of human-like reasoning in Large Language Models cites this paper.

Tracing the ongoing emergence of human-like reasoning in Large Language Models NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 80

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:04:37.201302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T05:04:30.829386Z digest=sha256:5834690fb4ebeff49e02f8a9e72fd277a223be5cd187b0234f4d1e5e066f5e5a

Observation 5df7cd2c-c942-4cee-98c2-9fd201ffd2e5 · inbound

Dynamic Mixture of Latent Memories for Self-Evolving Agents cites this paper.

Dynamic Mixture of Latent Memories for Self-Evolving Agents NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-05-22T07:41:14.542617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T07:38:42.113849Z digest=sha256:85d2ac73a0bc26fb59531ba23562d851e01e7876930611963e8db04766365c32

Observation e835c4dc-b9f7-461c-8c7f-8dbe20322a3e · inbound

Do Language Models Need Sleep? Offline Recurrence for Improved Online Inference cites this paper.

Do Language Models Need Sleep? Offline Recurrence for Improved Online Inference NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-06-29T21:43:59.463375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T21:37:51.638904Z digest=sha256:aba4ae28d6de64109e8a113f1678d4a8ce55a71d4dc893138087d7ac3b9c6b77

Observation 76e76fe9-05a6-429f-99b3-e0f8f5634542 · inbound

Mellum2 Technical Report cites this paper.

Mellum2 Technical Report NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 54

Resolution
verified exact
local_arxiv, observed 2026-06-28T23:02:46.484941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T22:58:35.397914Z digest=sha256:7e0b7d47f39adec950411a7100580eefd4c6ecc9f1e5b4eabdf69e942598d6a6

Observation 93b6c9c9-5cfb-436e-b376-e50aaeccd57d · inbound

Scaling Parallel Sequence Models to Foundation-Scale Vision Encoders cites this paper.

Scaling Parallel Sequence Models to Foundation-Scale Vision Encoders NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 175

Resolution
metadata mismatch
local_arxiv, observed 2026-06-28T19:32:35.225565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-28T19:23:08.100056Z digest=sha256:bf55d8d78a1abb1010841d7cb4a86c54b6a344889d42b786992e4fe89abe571b

Observation 22087dba-e869-47bc-bb39-3fb19b0cd449 · inbound

Scaling Agentic Capabilities via Grounded Interaction Synthesis cites this paper.

Scaling Agentic Capabilities via Grounded Interaction Synthesis NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-07-01T22:46:19.244060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T15:04:54.247779Z digest=sha256:bcb6c95035f6e4c5231d70ec2fb28586e426c0e109e90834c2c53eb87d9ea012

Observation ce84be6c-1f62-4ca9-a62a-6977a54bc33a · inbound

Membrane: A Self-Evolving Contrastive Safety Memory for LLM Agent Defense cites this paper.

Membrane: A Self-Evolving Contrastive Safety Memory for LLM Agent Defense NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 55

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T13:46:59.925020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-28T00:55:42.688093Z digest=sha256:93cce9245f592fc7614d9e3aa60c3bfac573f210d93edfb2cf5c5c48d0c22f5a

Observation 1adf6c37-b8f2-449a-890a-21d94f19020a · inbound

Data-Efficient Autoregressive-to-Diffusion Language Models via On-Policy Distillation cites this paper.

Data-Efficient Autoregressive-to-Diffusion Language Models via On-Policy Distillation NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-07-02T13:26:58.867105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T01:21:52.821813Z digest=sha256:903fdeb9842b36d078cbdbf8fc4ea533bc5b115dc250f7b2c74d7404efbb3113

Observation 19b42d67-3cc8-446b-8b91-218a19741887 · inbound

End-to-End Context Compression at Scale cites this paper.

End-to-End Context Compression at Scale NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 65

Resolution
verified exact
local_arxiv, observed 2026-07-03T01:17:31.619156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T16:36:54.699174Z digest=sha256:ea49880040c17ba8bf961fc082b09900bcaccfd6eac0086a2258137e0ec40561

Observation 42442150-9b91-4fec-b97a-d2c43be3edfe · inbound

Architecture-Aware Reinforcement Learning Makes Sliding-Window Attention Competitive in Math Reasoning cites this paper.

Architecture-Aware Reinforcement Learning Makes Sliding-Window Attention Competitive in Math Reasoning NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 41

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T09:47:59.919051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T10:18:54.163862Z digest=sha256:f75a514441d5ef54734b8d8f4f27d2eba22f4cb861c54f317ab24ac26b111dab

Observation 517c9560-6379-4aff-bf3b-9107df5e39a1 · inbound

When Does Generating More Help? Disentangling Fixed-Source Synthesis from Source Expansion in Synthetic Data Scaling cites this paper.

When Does Generating More Help? Disentangling Fixed-Source Synthesis from Source Expansion in Synthetic Data Scaling NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 26

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T15:28:33.790022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-07-03T15:20:51.474398Z digest=sha256:fec51767288201f2ce1ac6ad30387873a3d1d71eb824d344a6b6779fb6ecc9a6

Observation 804dbc10-94b1-439c-a10b-d3979b0f1d76 · inbound

RADIO1D: Elastic Representations for Condensed Vision Modeling cites this paper.

RADIO1D: Elastic Representations for Condensed Vision Modeling NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-12T01:07:20.766474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:07:20.766474Z digest=sha256:a5f8cd94ef611ba0b44cfad8f05019f7caace8f410e4a86c053d9ecf1d681613

Observation 933fef17-26ff-4e69-97d7-030e7e2ccaac · inbound

Full-Stack FP4: Stable LLM Pretraining with Quantized Projections, Optimizers, and Attention cites this paper.

Full-Stack FP4: Stable LLM Pretraining with Quantized Projections, Optimizers, and Attention NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-11T19:17:59.044982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:17:59.044982Z digest=sha256:aa8caa436c014ba70d7d8fe36364a593fbd8695c14fdbebb85779a5639aa405c

Observation 63117d4b-6445-4e67-9660-7a94f40a75d7 · inbound

Nemotron-Labs-Diffusion: A Tri-Mode Language Model Unifying Autoregressive, Diffusion, and Self-Speculation Decoding cites this paper.

Nemotron-Labs-Diffusion: A Tri-Mode Language Model Unifying Autoregressive, Diffusion, and Self-Speculation Decoding NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-07-11T03:07:51.292757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-11T03:04:12.500342Z digest=sha256:bcbdf624490841cf5855f7054c68985b27392fb1faa31aa727421b090d725e0d

Observation 6cd1a953-717b-4d4a-b5d6-dea0927ea9e4 · inbound

Accurate, Interdisciplinary and Transparent Structure-property Understanding with Deep Native Structural Reasoning cites this paper.

Accurate, Interdisciplinary and Transparent Structure-property Understanding with Deep Native Structural Reasoning NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 79

Resolution
metadata mismatch
local_arxiv, observed 2026-07-09T02:05:52.227738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-09T01:58:52.324533Z digest=sha256:35458be17faa4851adaebdec15a6700dd022df3e7674b679555cc2de55a587c6

Observation ea6df32d-a4a7-4c8b-a6ce-bbeaf37fd116 · inbound

Jet-Long: Efficient Long-Context Extension with Dynamic Bifocal RoPE cites this paper.

Jet-Long: Efficient Long-Context Extension with Dynamic Bifocal RoPE NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 50

Resolution
malformed identifier
local_arxiv, observed 2026-07-10T20:07:33.529785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-10T19:59:46.713277Z digest=sha256:0761c3cf318b46fccb7a22a2474610b3e4e2ce5cd4a148136d008f81737d7060

Observation 66aaeaef-bd8a-4b9f-a7cc-8c12d87137ac · inbound

Jet-Long: Efficient Long-Context Extension with Dynamic Bifocal RoPE cites this paper.

Jet-Long: Efficient Long-Context Extension with Dynamic Bifocal RoPE NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 53

Resolution
unresolved
no resolver link, observed 2026-07-13T06:47:09.927626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T06:47:09.927626Z digest=sha256:ccd65f121228e6f6262203de537cc7bc7706acb7489c76c945996893373384a5

Observation a7d854bc-43f0-4d96-851e-6ecd6459d118 · inbound

Length Penalties Make Chain-of-Thought Less Monitorable cites this paper.

Length Penalties Make Chain-of-Thought Less Monitorable NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 72

Resolution
unresolved
no resolver link, observed 2026-07-14T15:45:54.532529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T15:45:54.532529Z digest=sha256:442758efb36aec5adf571757eec2ab5542e35477dc96ff17a194bd209f30501a

Observation 92468da2-e18a-42fb-b8b8-77b8c3419d85 · inbound

Length Penalties Make Chain-of-Thought Less Monitorable cites this paper.

Length Penalties Make Chain-of-Thought Less Monitorable NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-02T08:06:10.832885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T08:06:10.832885Z digest=sha256:e2a7f4397f0bb6e20ad6ce063a071f581b66eb108987097d526652cf00a63da9

Observation 611fe834-32c6-4fbe-811d-fefe8563d370 · inbound

Length Penalties Make Chain-of-Thought Less Monitorable cites this paper.

Length Penalties Make Chain-of-Thought Less Monitorable NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T04:30:29.444171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T04:30:29.444171Z digest=sha256:764f52e9d71cb0c468e49f81dd0b70c88edbf008a25a865b0b1f784e25bd327f

Observation 804ba7b9-efb4-461f-9048-8380c1aa4aed · inbound

JOR-Bench: Japanese Operations Research Benchmarks for Large Language Models cites this paper.

JOR-Bench: Japanese Operations Research Benchmarks for Large Language Models NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T20:02:17.164217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T20:02:17.164217Z digest=sha256:510f6f3acf38253f7d0c630c808153c75fd3d7ed029b8a11278cd2c4c6179ee3

Observation 3abac5ac-1298-4f6a-97f1-9a1f3cf88d49 · inbound

Learning to Prepare Molecular Ground States with Transformer Models cites this paper.

Learning to Prepare Molecular Ground States with Transformer Models NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T04:45:22.713677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:45:22.713677Z digest=sha256:28bfefc66d0174d0dbe4d50d16c656c4877c761d47f7428382a2acac62eade91

Observation da4e1aa0-da80-45c2-827e-94467bb1bb68 · inbound

LoopMTP: A looped transformer guided by latent multi-token prediction cites this paper.

LoopMTP: A looped transformer guided by latent multi-token prediction NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T15:54:11.896996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:54:11.896996Z digest=sha256:2a7df0f8f66486c78e5d36a9ecd541f57a71509bf566bd286e2d98703b16df34