Pith. sign in

Paper Citation Record · LEDGER

LongCat-Video-Avatar 1.5 Technical Report

As of 22 July 2026, this Paper Citation Record lists 17 of 17 outbound references and 1 inbound Pith citation observation for arXiv:2605.26486.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.26486 v1

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-29T18:28:16.968339Z

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-07-22T06:31:00.163083+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-10T01:43:20.551939Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T01:46:41.055922Z

Reference resolution

17 of 17 outbound references displayed

  • verified exact13
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch4

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 54833615-f7d3-425a-9ff9-2a9524fd7069 · outbound

This paper cites In this video, does {positional description} {template}? Answer Yes or No.

LongCat-Video-Avatar 1.5 Technical Report In this video, does {positional description} {template}? Answer Yes or No

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T18:33:50.642525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-06-29T18:28:16.968339Z digest=sha256:000e1f2e174f726f92109edd194ddeaf14fa42d3bc396bec7a9590827972d376

Observation 0e763e58-9599-4309-a7f2-73e958f7da37 · outbound

This paper cites OmniHuman-1.5: Instilling an Active Mind in Avatars via Cognitive Simulation.

LongCat-Video-Avatar 1.5 Technical Report OmniHuman-1.5: Instilling an Active Mind in Avatars via Cognitive Simulation

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:33:50.663643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-06-29T18:28:16.968339Z digest=sha256:bbca14472cbde0cec1ee751069e11c828e2a3b6de4a86c75e7fabaa9f0f68cbc

Observation 7fb0aff1-5cb3-463f-9259-9ffaa2df951f · outbound

This paper cites Longcat-video technical report.arXiv preprint arXiv:2510.22200.

LongCat-Video-Avatar 1.5 Technical Report Longcat-video technical report.arXiv preprint arXiv:2510.22200

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:33:50.660854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-06-29T18:28:16.968339Z digest=sha256:344ce641422adaa6408cc50373a51783c33f8bee297d258fb30c0f390bb84dc0

Observation 39e56e96-9933-4318-b211-5bfd45fd9511 · outbound

This paper cites Wan-S2V: Audio-Driven Cinematic Video Generation.

LongCat-Video-Avatar 1.5 Technical Report Wan-S2V: Audio-Driven Cinematic Video Generation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:33:50.607869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-06-29T18:28:16.968339Z digest=sha256:0c442388ee79b5a22e0089c96b5982d0b93cee7f78cd4a8f0b3feb8542b90bf4

Observation 03f8bafa-4f21-4467-a1c6-69091ca72bab · outbound

This paper cites InfiniteTalk: Audio-driven Video Generation for Sparse-Frame Video Dubbing.

LongCat-Video-Avatar 1.5 Technical Report InfiniteTalk: Audio-driven Video Generation for Sparse-Frame Video Dubbing

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:33:50.661119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-06-29T18:28:16.968339Z digest=sha256:fa35cff6dda0732538e2b2444c83362fe547d7e554ea707eeef02296ae4110ac

Observation 7f083a95-514a-49f4-84fa-b92b464bbf32 · outbound

This paper cites HunyuanVideo-Avatar: High-Fidelity Audio-Driven Human Animation for Multiple Characters.

LongCat-Video-Avatar 1.5 Technical Report HunyuanVideo-Avatar: High-Fidelity Audio-Driven Human Animation for Multiple Characters

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:33:50.666638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-06-29T18:28:16.968339Z digest=sha256:33c61388abb0567e647f9145865a60d1d9b5dbab4d15ab095f75fb7c05faaf21

Observation 7f4c1ccb-9ef3-41c9-a383-cbd4e1f21d97 · outbound

This paper cites arXiv preprint arXiv:2512.11423 (2025) 30.

LongCat-Video-Avatar 1.5 Technical Report arXiv preprint arXiv:2512.11423 (2025) 30

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:33:50.654895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-06-29T18:28:16.968339Z digest=sha256:ea015b5ff40dd3aca24d4b01543b206daf33e8876d775a9e8a591b36b53a333c

Observation 95cdb6c0-d47b-47d3-9245-f7093c86a5ef · outbound

This paper cites Soulx-livetalk technical report.

LongCat-Video-Avatar 1.5 Technical Report Soulx-livetalk technical report

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T18:33:50.617285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-06-29T18:28:16.968339Z digest=sha256:a6adf3978dcf1e10aa96ccafc33eae05a56ac3407e3a14f40786a62c6c94eaca

Observation 780d28e6-b29b-4f7d-80f0-b2b84fa524c2 · outbound

This paper cites Robust Speech Recognition via Large-Scale Weak Supervision.

LongCat-Video-Avatar 1.5 Technical Report Robust Speech Recognition via Large-Scale Weak Supervision

Reference 9

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T18:33:50.663647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-06-29T18:28:16.968339Z digest=sha256:af117e4b315fb18d4eef6c2626c1e0ed1402fa475e94496f3474c12403600e8a

Observation 930ff083-3975-4ea4-a8fc-6df493d7cf18 · outbound

This paper cites Revisiting Active Speaker Detection: An In-the-Wild Benchmark for Generalization and Robustness.

LongCat-Video-Avatar 1.5 Technical Report Revisiting Active Speaker Detection: An In-the-Wild Benchmark for Generalization and Robustness

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:33:50.645459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-06-29T18:28:16.968339Z digest=sha256:efdf0f19baf501c809079849aeb76c8353b628920f804dcdcc9ee03599ac0217

Observation 74f07768-04d8-4f0a-bf6c-15099e8605c2 · outbound

This paper cites YOLOv6: A Single-Stage Object Detection Framework for Industrial Applications.

LongCat-Video-Avatar 1.5 Technical Report YOLOv6: A Single-Stage Object Detection Framework for Industrial Applications

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:33:50.642374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-06-29T18:28:16.968339Z digest=sha256:e96f74499116967bd44a2b708beabce27d3c6bea183b87b5a2ca36272a3ba637

Observation 47276ab3-eac1-4581-8fb7-93a526dc6f12 · outbound

This paper cites Qwen3-Omni Technical Report.

LongCat-Video-Avatar 1.5 Technical Report Qwen3-Omni Technical Report

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:33:50.651492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-06-29T18:28:16.968339Z digest=sha256:626ea3650c9c00370af603aa83099cd95974d967442bcd3de8d2865c4a3a6d7c

Observation d6273a1b-97be-42be-9056-05c2fed91795 · outbound

This paper cites Qwen3-VL Technical Report.

LongCat-Video-Avatar 1.5 Technical Report Qwen3-VL Technical Report

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:33:50.666546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-06-29T18:28:16.968339Z digest=sha256:d71e1b3bc9d96ae0d1bd48d0008eb377eb2ee32cecf1cd080a8651b7bcadf688

Observation b3cb5746-2fbf-4d8a-9bfd-8afa6b38967f · outbound

This paper cites HuMo: Human-Centric Video Generation via Collaborative Multi-Modal Conditioning.

LongCat-Video-Avatar 1.5 Technical Report HuMo: Human-Centric Video Generation via Collaborative Multi-Modal Conditioning

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:33:50.636837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-06-29T18:28:16.968339Z digest=sha256:fa3005e9954fad4d048f77a3d34a95f00b64a9070216458377cf932e2261ef5c

Observation f0af46ae-4b1a-4e37-b75c-0364ae4fdbd0 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

LongCat-Video-Avatar 1.5 Technical Report DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:33:50.657818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-06-29T18:28:16.968339Z digest=sha256:ac1ef322c842eed216b900cac6c6f378e55dd6835313dff468af10ba98d4d39a

Observation c245d17f-f489-40d5-9a60-6403570517ba · outbound

This paper cites Flow Matching for Generative Modeling.

LongCat-Video-Avatar 1.5 Technical Report Flow Matching for Generative Modeling

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:33:50.639608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-06-29T18:28:16.968339Z digest=sha256:aeecf560b7eed01ba99346d8f42f238ec11a230bc5f73d7bfc71fae60e08ac26

Observation 248f2157-4eaa-4040-a110-a691f1e14df4 · outbound

This paper cites 10•Jimin Tang et al.

LongCat-Video-Avatar 1.5 Technical Report 10•Jimin Tang et al

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T18:33:50.669690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-06-29T18:28:16.968339Z digest=sha256:9aaf0b12cd01ae21b59b202cdc8a4b5371dd1d31b4f4886ef94f77a219f70dc6

Pith citing papers

Observation 0b23d11b-2430-4e79-923d-68d38a35a64d · inbound

OPSD-V: On-Policy Self-Distillation for Post-Training Few-Step Autoregressive Video Generators cites this paper.

OPSD-V: On-Policy Self-Distillation for Post-Training Few-Step Autoregressive Video Generators LongCat-Video-Avatar 1.5 Technical Report

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-07-10T01:46:41.057375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-22T06:31:00.163083+00:00.

source=pdf_text observed=2026-07-10T01:43:20.551939Z digest=sha256:40c69b698dc1a8fefc45cf5035e9e5638576ff7b50a5c44ac3d32ede59fe11ee