Pith. sign in

Paper Citation Record · LEDGER

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context

As of 8 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 23 inbound Pith citation observations for arXiv:2506.21277.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.21277 v1

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:36:16.371915Z

measured 72 of 72 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 23 of 23 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T12:00:10.926459Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T19:20:06.286106Z

Reference resolution

49 of 49 outbound references displayed

  • verified exact0
  • verified fuzzy6
  • unresolved43
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d9bc3846-ed5c-434e-b476-f42861c1cec7 · outbound

This paper cites Qwen2.5-Omni Technical Report.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context Qwen2.5-Omni Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:13.100076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:13.100076Z digest=sha256:81945894b4dd749245c8b018ab8512889d34f4491ca547c1afac543dd45b9d7a

Observation 02c0e5a7-4ff0-45df-83de-54ac2fc6f8d4 · outbound

This paper cites Ocean-omni: To understand the world with omni-modality,.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context Ocean-omni: To understand the world with omni-modality,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:13.189958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:13.189958Z digest=sha256:6665c168d4e15eda6206134495b51092728eb1a8ee16ef144eee794529e8388c

Observation b21fc665-8563-4dac-84d1-1b4f15f5c7a2 · outbound

This paper cites Ola: Pushing the Frontiers of Omni-Modal Language Model.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context Ola: Pushing the Frontiers of Omni-Modal Language Model

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:13.250680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:13.250680Z digest=sha256:31e4e66c26921915388beb72699e3e5e7d891374c895197a837e18b422c3bec9

Observation d5cdd6a8-e1dd-4942-a1e9-e3d1f4c900d9 · outbound

This paper cites VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:13.311095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:13.311095Z digest=sha256:942668bea2468a94b31fbe19279c9e17772283d12845edcee6d736af1bab76cc

Observation cc1a1c1b-483e-4412-9a66-b0bac0933d76 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:13.354539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:13.354539Z digest=sha256:1115d0456e16191aba8262b66aaf19d3b5245ec32233807594d25335d2e4be85

Observation fccec2a2-899d-4fb8-9ec7-dcb3dea65492 · outbound

This paper cites MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:13.414790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:13.414790Z digest=sha256:de373bcd28e9689f9189358687db4bf22c1e9d6bf4bf6e4c20ae2570a606850d

Observation 2017c838-81a4-459c-89f9-da640f51f34e · outbound

This paper cites Visual-RFT: Visual Reinforcement Fine-Tuning.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context Visual-RFT: Visual Reinforcement Fine-Tuning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:13.542217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:13.542217Z digest=sha256:8e140d79bb711db470529c4f97859781fe5d16d102f9d3f46c4a2d109a067577

Observation 0a69b6e7-73fe-43b2-a617-8f53fb280752 · outbound

This paper cites VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:13.605014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:13.605014Z digest=sha256:2fa148f9836b7545334f984559ec567291af7d132b8b46c28ac3da7f694c8efd

Observation 83c73b5c-953f-4b72-b537-44a7165a64ef · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:13.642332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:13.642332Z digest=sha256:0a224778e2d922939d18aef457ddf82bc22781e6fb2871ff60b7d6675c176c9b

Observation 2e8db14b-11e6-41f0-95b6-b26a32bac4a4 · outbound

This paper cites Let’s verify step by step,.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context Let’s verify step by step,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:13.716590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:13.716590Z digest=sha256:ddf2f4def49320fd8eec2fb194f430b0d95c507c70368c5ef93b919382fdc3bf

Observation 00094bd8-5090-4b3d-a269-be2a6a281e40 · outbound

This paper cites Deep reinforcement learning from human preferences,.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context Deep reinforcement learning from human preferences,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:13.791285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:13.791285Z digest=sha256:5eb3c2f84d8e9671c68752181bf7529a6b0b517a4394f2897118c854c16f4319

Observation c2c5ed5d-25b5-49d3-b8da-6b4715ff9d5c · outbound

This paper cites Training language models to follow instructions with human feedback,.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context Training language models to follow instructions with human feedback,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:13.854051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:13.854051Z digest=sha256:0f4db05303fbe652567276dbfcebaccbdf5c05197a0b737b11f2aeda230c8574

Observation cadeb560-c2f2-400e-90c2-e440629f71fa · outbound

This paper cites Daily-omni: Towards audio-visual reasoning with temporal alignment across modalities,.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context Daily-omni: Towards audio-visual reasoning with temporal alignment across modalities,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:13.927417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:13.927417Z digest=sha256:99815f7eeb57be231c24f1fb35ea39554ac308dbab3fa25edf998a22a12802bc

Observation 4e3c89a2-e02d-4033-8cbf-0a45e2aebf3e · outbound

This paper cites WorldSense: Evaluating Real-world Omnimodal Understanding for Multimodal LLMs.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context WorldSense: Evaluating Real-world Omnimodal Understanding for Multimodal LLMs

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:14.008552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:14.008552Z digest=sha256:1d9da3bee1e10968a9c05175d939e1fec99a7d2fbc84ab528e43fcf20ea97f9a

Observation b6abf601-e5a5-40b6-998a-9d42fbc6d159 · outbound

This paper cites HumanOmni: A Large Vision-Speech Language Model for Human-Centric Video Understanding.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context HumanOmni: A Large Vision-Speech Language Model for Human-Centric Video Understanding

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:14.070930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:14.070930Z digest=sha256:11e452e32ed9e6231cd668d8484c4918caaff1e38f58ec446fa772145ebf0a84

Observation 8fd7646a-2680-4dfc-ad4e-60c889f0c778 · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:14.132504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:14.132504Z digest=sha256:2391e65517843a011fe495201afc98038f1e6a4558e35dde28091c524ff8c9ad

Observation 3ec31ea1-a90e-4493-a56b-c09be0b7b9bb · outbound

This paper cites InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:14.198672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:14.198672Z digest=sha256:ae962524eef8f48af1ddc856a3f76aa00e0f3208edc97e09948df76440f5ee03

Observation dac0beed-5947-41bd-acb0-9d2fb423d933 · outbound

This paper cites ViSpeak: Visual Instruction Feedback in Streaming Videos.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context ViSpeak: Visual Instruction Feedback in Streaming Videos

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:14.252623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:14.252623Z digest=sha256:925a9287c4a56f588388385ac4ab26d41040f369f4825d6d7381e68ccde801de

Observation eb3617e0-da3f-4432-a830-7d9fd7b958eb · outbound

This paper cites Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:14.298315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:14.298315Z digest=sha256:f90ef37097e10d1ab93e863d7e17145d5c686e231443c4877975de5e2ec69db9

Observation e1e885b5-8e18-4d07-801f-3ce16ed229f2 · outbound

This paper cites Video-mme: The first-ever comprehensive evaluation benchmark of multi-modal llms in video analysis,.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context Video-mme: The first-ever comprehensive evaluation benchmark of multi-modal llms in video analysis,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:36:18.497671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:36:14.371699Z digest=sha256:f6e55f206910224d32b48066e297ec1e5735781a2b8f0867c7b83958664d79bf

Observation da4ef8b4-87b9-44b4-bb5c-c0544257c32a · outbound

This paper cites ActionArt: Advancing Multimodal Large Models for Fine-Grained Human-Centric Video Understanding.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context ActionArt: Advancing Multimodal Large Models for Fine-Grained Human-Centric Video Understanding

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:14.442972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:14.442972Z digest=sha256:5e79d44793a0699a16586c6366f3c1c7173d3cbb444cacd96810df9b318eb4f4

Observation 868b216d-5934-4f50-84e1-78e86dcae62b · outbound

This paper cites Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi,.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:36:18.346314Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:36:14.506749Z digest=sha256:7d7a185974b21405b7a42f690de7f39a4159f2597eb418b2d9845b4a82f91d6a

Observation 0ccc3c08-c875-420a-b814-3aef9656fe1c · outbound

This paper cites Omnibench: Towards the future of universal omni-language models,.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context Omnibench: Towards the future of universal omni-language models,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:14.573744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:14.573744Z digest=sha256:b3fd607a6a5daefdd2d9af0a5fd7d7a990a0faf621ab751d319f80bc7c2786d9

Observation 7762b41f-f762-4dd8-8538-925bb04c6bbb · outbound

This paper cites Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:14.648069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:14.648069Z digest=sha256:1c69d32dfbbd49907e2ae4c8d0aeea10d50361de3a7f47acb2220f2175f8eaa4

Observation be188088-53f3-405c-8485-d3c128398b7b · outbound

This paper cites R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:14.709508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:14.709508Z digest=sha256:b00ab6978e18a70041eaab05a200763865e4e2e621eea0e33a1a4f7098e50715

Observation e3a3aa44-7bc6-4e3b-b410-5bbf810d6588 · outbound

This paper cites Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:14.773345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:14.773345Z digest=sha256:89d9eaacd7303a0d837427be08df3631804134c1db5f0d4fdcb0f8c232fc56e8

Observation 7acb12a4-1f90-432c-a59d-a77bcbe0afa5 · outbound

This paper cites Visionary-r1: Mitigating shortcuts in visual reasoning with reinforcement learning,.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context Visionary-r1: Mitigating shortcuts in visual reasoning with reinforcement learning,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:14.841838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:14.841838Z digest=sha256:98a0bdf29b59d936b00cf27058b4905581458dc692791aad0086b038ef5cb192

Observation ccc0b7ab-2e2b-4db4-ad07-1df3b15e635b · outbound

This paper cites Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:14.896841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:14.896841Z digest=sha256:47ded081d7ee0a9dfcee42c33588a9a2ad2fff0eb1ce756ae6949956d1ad649a

Observation b5e06ef8-e5bd-4aa1-a082-be15c5a55d81 · outbound

This paper cites R1-Omni: Explainable Omni-Multimodal Emotion Recognition with Reinforcement Learning.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context R1-Omni: Explainable Omni-Multimodal Emotion Recognition with Reinforcement Learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:14.965189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:14.965189Z digest=sha256:5677cca861b761357ca38c5eb958648ca41c2bdb913d7f61f9afea1c7b75d211

Observation 2caf0f52-2c2f-4cb8-bcf5-c365bbbc7f9b · outbound

This paper cites EchoInk-R1: Exploring Audio-Visual Reasoning in Multimodal LLMs via Reinforcement Learning.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context EchoInk-R1: Exploring Audio-Visual Reasoning in Multimodal LLMs via Reinforcement Learning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:15.038514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:15.038514Z digest=sha256:8884ec8ff8fc7e9f91d427d2626a8e95831a4f39f002efd7c4dda860296176e8

Observation 8f50cf77-265a-42b5-a02d-2589dfdb4edc · outbound

This paper cites MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:15.103448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:15.103448Z digest=sha256:51aff256e70d85cc5b361e8903ed6ab5ef0cf57088049cd99985edd6e0c709cf

Observation 41b50db4-1a75-499d-8957-a94752de91bb · outbound

This paper cites MMVU: Measuring Expert-Level Multi-Discipline Video Understanding.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context MMVU: Measuring Expert-Level Multi-Discipline Video Understanding

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:15.170351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:15.170351Z digest=sha256:80c99d00c164b1c1e8ec9fa119f5e5aa697cfe2cd1e6a4e44141564d9943c59b

Observation 976133f1-7b17-43e2-9abe-947893c2214f · outbound

This paper cites Social-iq: A question answering benchmark for artificial social intelligence,.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context Social-iq: A question answering benchmark for artificial social intelligence,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:36:18.184139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:36:15.224232Z digest=sha256:a5edd75242b8abd6ead948f1cf21575c6037504f2d8994d3650e134ff6d890a6

Observation 6f381da0-0ce4-4880-bcce-f2d97caca2c4 · outbound

This paper cites Social-iq 2.0 challenge: Benchmarking multimodal social understanding,.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context Social-iq 2.0 challenge: Benchmarking multimodal social understanding,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:36:18.047248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:36:15.295001Z digest=sha256:f239cfd2545e4c233c865cedca71ed6025b89eabc63541271c6fb85c0978057a

Observation 6fb568fc-47de-4fac-9f85-d4df837e52ca · outbound

This paper cites Explainable Multimodal Emotion Recognition.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context Explainable Multimodal Emotion Recognition

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:15.352601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:15.352601Z digest=sha256:4f45103369ac3725ca25918b0425be8502d9b2e8e384c2834c7a0e47f96ae1a1

Observation eff13216-134e-4527-abcc-8b7c9b4f2b5c · outbound

This paper cites MDPE: A Multimodal Deception Dataset with Personality and Emotional Characteristics.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context MDPE: A Multimodal Deception Dataset with Personality and Emotional Characteristics

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:15.404271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:15.404271Z digest=sha256:a7b9a4fad7a92be2ae65f5b87af8a50b3ffa4795483ed0fd23c4dadada9c3713

Observation 0ac7f971-756b-4df0-a07a-adb9a148390e · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:15.530089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:15.530089Z digest=sha256:356affad02fcdd3095ac386669689c4b80f2ce25a932cdb97bf3e1114a0ad0b2

Observation 6e4ade0f-7057-44cf-9251-b15672f059d8 · outbound

This paper cites Understanding R1-Zero-Like Training: A Critical Perspective.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context Understanding R1-Zero-Like Training: A Critical Perspective

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:15.612633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:15.612633Z digest=sha256:416d675d2354491d8a988569267bc255735ff33035444d95e2cd40d7363a188a

Observation 737c5beb-ed58-4d15-bef9-3c84af483d50 · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation,.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context Bleu: a method for automatic evaluation of machine translation,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:15.669533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:15.669533Z digest=sha256:ee089165e20c9f181160014d0ba971963238c848640dfc86493e453437732130

Observation 65645993-1ed6-42cb-9e0d-07efdf6be16a · outbound

This paper cites Rouge: A package for automatic evaluation of summaries,.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context Rouge: A package for automatic evaluation of summaries,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:15.737016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:15.737016Z digest=sha256:00c8028970db53db513083a928505568802816864f9dfc2016bf4b57164383b9

Observation dab78a1e-8eb2-4f93-a1f6-d2595de50d6e · outbound

This paper cites Video-R1: Reinforcing Video Reasoning in MLLMs.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context Video-R1: Reinforcing Video Reasoning in MLLMs

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:15.835979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:15.835979Z digest=sha256:7e5672b26f5752a6dc71080f6c35e49ca5df50d2b2aebe8c7c632e9e43da73d2

Observation cc07c2b5-094e-49ea-8069-d1d9f188f4d7 · outbound

This paper cites Gemini 2.5 pro,.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context Gemini 2.5 pro,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:36:17.887554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:36:15.883030Z digest=sha256:6a27211c50171ea0baff1b99fffa041e2d554f15ed8f6924b66d952d3879b03d

Observation 361cf6a6-743f-43ea-9829-93aa50d84b9a · outbound

This paper cites Unified-io 2: Scaling autoregressive multimodal models with vision language audio and action,.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context Unified-io 2: Scaling autoregressive multimodal models with vision language audio and action,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:15.960063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:15.960063Z digest=sha256:5f2abaea7e2958cde145aafc02c41442d355bf7a00f43cffb355e3fcba1f6d52

Observation dd9cc03c-bc9a-42bf-9f5a-a2e6bf2ac43f · outbound

This paper cites VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:16.021781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:16.021781Z digest=sha256:e294674d535bb204bf2c3eb933ee46cbdb354bb67d53796e8818c00b7d46474d

Observation a0a0d3ee-7351-42c2-8670-9c7d5b0507d0 · outbound

This paper cites Introducing the next generation of Claude,.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context Introducing the next generation of Claude,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:36:17.754014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:36:16.124771Z digest=sha256:4fd193df980adeca62c1994e1f3b57ec12109f0384147965231bc8c20196493f

Observation 41a5225d-a5b5-4634-887a-413f92aba628 · outbound

This paper cites GPT-4o System Card.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context GPT-4o System Card

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:16.173667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:16.173667Z digest=sha256:da1926a32e9c8cdbad7644d263bf61c3f988243caeb05d1a7321ffad48d16a65

Observation 942a1024-803c-49a2-93fa-b668823c8534 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:16.270052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:16.270052Z digest=sha256:83b2d9e89a7f7859a3554aaf42250576dbfc112576e6349759436d9a21d10b84

Observation 64396dc5-da23-4b78-9b50-770932dda3d7 · outbound

This paper cites OpenAI o1 System Card.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context OpenAI o1 System Card

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:16.328116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:16.328116Z digest=sha256:eb317faa103eb1ea3e0aa6cb8276307df25bb0b8c2b227ee7bbd65f8e40d03d8

Observation c104b7a9-1c1f-41f6-97fa-cccf779cf544 · outbound

This paper cites AffectGPT: Dataset and Framework for Explainable Multimodal Emotion Recognition.

HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context AffectGPT: Dataset and Framework for Explainable Multimodal Emotion Recognition

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:16.371915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:36:16.371915Z digest=sha256:e92e86f49f0e1ce1635e963fb04e791957adbe9ee9ccc88b55ccdea26f0b1566

Pith citing papers

Observation 5254e4da-853c-4ac8-8643-6b2e024fa21e · inbound

OmniZip: Audio-Guided Dynamic Token Compression for Fast Omnimodal Large Language Models cites this paper.

OmniZip: Audio-Guided Dynamic Token Compression for Fast Omnimodal Large Language Models HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:50:15.048217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T20:45:37.418493Z digest=sha256:733dd5bab5d3186a6217ecb2dcddeb7e7dacc47afe790088c14de23b0af50f81

Observation c9b2aa38-e488-4e17-83a9-43891b92c31f · inbound

EchoingPixels: Aliasing-Resistant Joint Token Reduction for Audio-Visual LLMs cites this paper.

EchoingPixels: Aliasing-Resistant Joint Token Reduction for Audio-Visual LLMs HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-03T17:16:42.542703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T17:16:42.542703Z digest=sha256:ebc0b8d9c42dcac30eff7dcee0528ffd02af4aca14fe9892f070a1712a74d77e

Observation 95218483-f9a3-4ebb-b2e0-62ea836414ee · inbound

TimeChat-Captioner: Scripting Multi-Scene Videos with Time-Aware and Structural Audio-Visual Captions cites this paper.

TimeChat-Captioner: Scripting Multi-Scene Videos with Time-Aware and Structural Audio-Visual Captions HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:11.377213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:11.377213Z digest=sha256:61a877af242a2a144b585505a9a64c680b7f335bd99617d349d59ac265604514

Observation 48f64fe9-9131-45e2-8b4d-3a57bd8dfc5c · inbound

OmniJigsaw: Enhancing Omni-Modal Reasoning via Modality-Orchestrated Reordering cites this paper.

OmniJigsaw: Enhancing Omni-Modal Reasoning via Modality-Orchestrated Reordering HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context

Reference 47

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T06:11:01.853846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T17:45:51.528645Z digest=sha256:742bc4e49d57d620b37dba47bad6109d44855fbec6c611287f4cda595b94d8af

Observation ad84e811-f1ba-43c7-ba5f-a6cfadc8f22f · inbound

Script-a-Video: Deep Structured Audio-visual Captions via Factorized Streams and Relational Grounding cites this paper.

Script-a-Video: Deep Structured Audio-visual Captions via Factorized Streams and Relational Grounding HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:11:03.635645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T15:07:45.595260Z digest=sha256:97685e5c7a1141108ac5a70bb493615b4319191473eef37c93d254e9f9a75b87

Observation 2d070b2f-fe4b-43aa-8871-d9c741d48a2e · inbound

Chain of Modality: From Static Fusion to Dynamic Orchestration in Omni-MLLMs cites this paper.

Chain of Modality: From Static Fusion to Dynamic Orchestration in Omni-MLLMs HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-10T12:10:22.096457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T12:05:54.551728Z digest=sha256:722153575ea66a86984f91d9e1302e9cad0a1c7b974fb19718ff0d92cf367563

Observation 9c31247b-557b-4232-b4a8-1d464eac4536 · inbound

AVRT: Audio-Visual Reasoning Transfer through Single-Modality Teachers cites this paper.

AVRT: Audio-Visual Reasoning Transfer through Single-Modality Teachers HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T09:18:32.195070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T08:02:53.574120Z digest=sha256:37e28e0f835fc81820ecfe46a1bb377e76b5023ea03bec8d5ca09eb4c3a22080

Observation f0ea401e-48d2-42cb-97e1-5bb6bd5864e1 · inbound

Omni-Persona: Systematic Benchmarking and Improving Omnimodal Personalization cites this paper.

Omni-Persona: Systematic Benchmarking and Improving Omnimodal Personalization HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:46:24.087743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T04:01:23.060562Z digest=sha256:1ed22a929effbfb896159397af945212b0d87428de11fac088902abaf91613ee

Observation 6fec6123-dd93-45a8-be08-fed71f6a3750 · inbound

Boosting Omni-Modal Language Models: Staged Post-Training with Visually Debiased Evaluation cites this paper.

Boosting Omni-Modal Language Models: Staged Post-Training with Visually Debiased Evaluation HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-13T03:52:12.705531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T03:49:58.240883Z digest=sha256:f785c1e058b523739424a4298087b3935aa6ac908075fc7366918c2517b53b8e

Observation 5e34008f-04ec-4570-9bde-2018ab456caa · inbound

Boosting Omni-Modal Language Models: Staged Post-Training with Visually Debiased Evaluation cites this paper.

Boosting Omni-Modal Language Models: Staged Post-Training with Visually Debiased Evaluation HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-15T06:09:50.263174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T06:06:20.658030Z digest=sha256:5fcc9c4f1b3575dd919e1560a8689b4e378058e04a570df6ab713da6bc0b3137

Observation f2b4c485-182d-433f-a46b-7206a4820b9c · inbound

OmniRefine: Alignment-Aware Cooperative Compression for Efficient Omnimodal Large Language Models cites this paper.

OmniRefine: Alignment-Aware Cooperative Compression for Efficient Omnimodal Large Language Models HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-13T04:52:16.258450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T04:52:03.076788Z digest=sha256:164e8002546332222da655754aca18a8c7418cab33d830b909c53c6fde9c95e6

Observation f4bb1498-5f94-4c4e-9d23-241b5b30ffcb · inbound

See What I Mean: Aligning Vision and Language Representations for Video Fine-grained Object Understanding cites this paper.

See What I Mean: Aligning Vision and Language Representations for Video Fine-grained Object Understanding HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context

Reference 81

Resolution
verified exact
arxiv_id, observed 2026-05-20T12:13:16.164422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T12:10:54.874012Z digest=sha256:ec62347a6b5ddd107c9bd000fc39365439c10148976deab9f81918efd59fd0b9

Observation 98d9eaa6-52e2-4e35-ae19-f7b6495b6242 · inbound

LatentOmni: Rethinking Omni-Modal Understanding via Unified Audio-Visual Latent Reasoning cites this paper.

LatentOmni: Rethinking Omni-Modal Understanding via Unified Audio-Visual Latent Reasoning HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-22T06:36:10.573414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T06:34:57.483234Z digest=sha256:12a7dee3fd9ebecdb43c374c4691a6e8fcc9c407a4466d74116142d0d44cd987

Observation b44a9268-54eb-4b16-9f84-0587a19943c4 · inbound

MODF-SIR: A Multi-agent Omni-modal Distilled Framework for Social Intelligence Reasoning cites this paper.

MODF-SIR: A Multi-agent Omni-modal Distilled Framework for Social Intelligence Reasoning HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:58:03.460392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T09:43:56.299302Z digest=sha256:02066298460f9ff1251fd413bc526118c4b826dce020a3002acdf728b2a938e8

Observation db542710-299f-4a10-8c70-18144ce331a2 · inbound

CogniRoute: Learning to Route Social Evidence in Omni-Modal Models cites this paper.

CogniRoute: Learning to Route Social Evidence in Omni-Modal Models HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context

Reference 91

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T03:49:30.332893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T17:37:11.371892Z digest=sha256:c8bda857212802fdf429acc3dcf174bd97faf6d6d8a002e4f39b1702a8176754

Observation be526fec-bd7e-4d07-86f0-23097d16a4f7 · inbound

Omni-Perception Policy Optimization for Multimodal Emotion Reasoning cites this paper.

Omni-Perception Policy Optimization for Multimodal Emotion Reasoning HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T19:20:06.288028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-25T21:31:38.450382Z digest=sha256:d63397c9814b6be71d4df18f594929cc801dfe28d89794f82410793a2f18c2df

Observation db441dfa-501f-404a-b6d7-55eb6216d76b · inbound

MER-R1: Multimodal Emotion Reasoning via Slow-Fast Thinking Synergy cites this paper.

MER-R1: Multimodal Emotion Reasoning via Slow-Fast Thinking Synergy HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:23:51.450720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T05:06:09.216428Z digest=sha256:377e7b9d8fb3f6bb668e6553ce17e5e9933df60509801d3610654509d6a80367

Observation 4f5011fd-6082-4ad2-83be-cf2feec23b41 · inbound

Temporal and Cross-Modal Alignment for Enhanced Audiovisual Video Captioning cites this paper.

Temporal and Cross-Modal Alignment for Enhanced Audiovisual Video Captioning HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context

Reference 64

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T16:38:39.614890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-03T16:37:06.384435Z digest=sha256:9c54b93979b77440227008275c87caf4ba54112483d769b8edf15a835e0ef23f

Observation 8890fdbd-95f2-49ea-9cb8-00dcbccba527 · inbound

Empowering Long-form Omni-modal Understanding with Robust Audio Perception cites this paper.

Empowering Long-form Omni-modal Understanding with Robust Audio Perception HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context

Reference 56

Resolution
unresolved
no resolver link, observed 2026-07-14T12:48:58.688011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:48:58.688011Z digest=sha256:2b904c8e0ffcc8978b39551b04a018b7680f002097d26cd1e0ac52bef91db53d

Observation 6dedb69d-cfeb-400f-85af-d9d6873b62f4 · inbound

LenGuard-GPC: Length Guarding with Guided-Prompt Consistency for Spatial Reasoning Reinforce Learning cites this paper.

LenGuard-GPC: Length Guarding with Guided-Prompt Consistency for Spatial Reasoning Reinforce Learning HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T18:40:15.585890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:40:15.585890Z digest=sha256:b49f766ae95c98e0b6d1f9ab40e27790b6c6bba83341e9747640d3f4e2fbbb18

Observation 7602b7d7-0cd3-4c99-ad6b-1853d8bd6901 · inbound

OmniScope: Modality-Decoupled Token Compression for Omnimodal Large Language Models cites this paper.

OmniScope: Modality-Decoupled Token Compression for Omnimodal Large Language Models HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-01T03:25:31.638356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:25:31.638356Z digest=sha256:9823d15982a1da00fed1b41b6ff547c03c8431f53982776c096d3248167a9703

Observation 2bcccdc2-9feb-42fb-acce-f629ddb82464 · inbound

OmniScope: Modality-Decoupled Token Compression for Omnimodal Large Language Models cites this paper.

OmniScope: Modality-Decoupled Token Compression for Omnimodal Large Language Models HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-04T04:02:54.515743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:02:54.515743Z digest=sha256:ba612c3db9bc198f703ef9d877e4119bfe5fbbdaef2d13ed73ee8a0b3ba0ceb1

Observation 5ff0a44c-6bed-496d-a2cd-b373b08cf750 · inbound

OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models cites this paper.

OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-05T12:00:10.926459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:00:10.926459Z digest=sha256:458ab58524c210ebb51092471dbfb7620bf9d551c63b74daa915029dd583b734