Pith. sign in

Paper Citation Record · LEDGER

Gesture-Aware Pretraining and Token Fusion for 3D Hand Pose Estimation

As of 22 July 2026, this Paper Citation Record lists 25 of 25 outbound references and 0 inbound Pith citation observations for arXiv:2603.17396.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2603.17396 v3

Coverage vector

measured 25 of 25 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-15T10:09:08.788131Z

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-07-21T06:31:05.380196+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

25 of 25 outbound references displayed

  • verified exact2
  • verified fuzzy23
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 170b51a1-8c57-460f-a13d-b15fdf53f3e5 · outbound

This paper cites 3d hand shape and pose from images in the wild.

Gesture-Aware Pretraining and Token Fusion for 3D Hand Pose Estimation 3d hand shape and pose from images in the wild

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T10:09:55.960044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-05-15T10:09:08.788131Z digest=sha256:1f66a5d6e600f214c8edffd4708f112677bac4302d79d562a00a32666aff8128

Observation 56c5bb53-1b06-497c-b0a5-1c1ec954da17 · outbound

This paper cites Weakly-supervised 3d hand pose estimation from monocu- lar rgb images.

Gesture-Aware Pretraining and Token Fusion for 3D Hand Pose Estimation Weakly-supervised 3d hand pose estimation from monocu- lar rgb images

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T10:09:55.945734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-05-15T10:09:08.788131Z digest=sha256:b32e2239e83487af576f22360401df6cd6d4fb755050bc56954765e20edceac3

Observation 308bbab6-b7f3-4af8-bf7a-9753160da935 · outbound

This paper cites Subunets: End-to-end hand shape and continuous sign language recognition.

Gesture-Aware Pretraining and Token Fusion for 3D Hand Pose Estimation Subunets: End-to-end hand shape and continuous sign language recognition

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T10:09:55.936267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-05-15T10:09:08.788131Z digest=sha256:c0191693a24e1ef5e44ec6dc8c02bff233716d02c73c642b429513dbad7ef526

Observation 4779bea9-bbda-4eb4-b3e0-052cd963a4ca · outbound

This paper cites Model- based 3d hand reconstruction via self-supervised learning.

Gesture-Aware Pretraining and Token Fusion for 3D Hand Pose Estimation Model- based 3d hand reconstruction via self-supervised learning

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T10:09:55.951303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-05-15T10:09:08.788131Z digest=sha256:e9a9d375114aa69138a7706a84352a84d35b8ad152fd5a9c235d93e467697796

Observation ac6df174-ac59-4b8a-8940-ad650fd2f48a · outbound

This paper cites 3d hand shape and pose estimation from a single rgb image.

Gesture-Aware Pretraining and Token Fusion for 3D Hand Pose Estimation 3d hand shape and pose estimation from a single rgb image

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T10:09:55.954692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-05-15T10:09:08.788131Z digest=sha256:93ae3755b269a88b2b525ded8d3b53c232931b08ae0fd52c5fb18565fc7ecc42

Observation b46ecf5f-5472-40e4-b8f7-4e2aaca38f01 · outbound

This paper cites an unresolved cited work.

Gesture-Aware Pretraining and Token Fusion for 3D Hand Pose Estimation Unresolved cited work

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T10:09:55.933200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-05-15T10:09:08.788131Z digest=sha256:2508525d983e146fb2572baefdf44e3ca1ad51e98142dd4dd57af075fe45dfaa

Observation b5b90498-394d-4145-8a29-c5fff990bb34 · outbound

This paper cites FineHand: Learning hand shapes for american sign language recogni- tion.

Gesture-Aware Pretraining and Token Fusion for 3D Hand Pose Estimation FineHand: Learning hand shapes for american sign language recogni- tion

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T10:09:55.931379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-05-15T10:09:08.788131Z digest=sha256:f2b43cfd78ddf7f28e3b4d07697b0ae961b42894c58c8f06ecfe33f9d8c5faa0

Observation f6aad9fd-6980-47b1-a4a0-d85ad308b27f · outbound

This paper cites Black, David W.

Gesture-Aware Pretraining and Token Fusion for 3D Hand Pose Estimation Black, David W

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T10:09:55.939491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-05-15T10:09:08.788131Z digest=sha256:4cd7cce47d35558c0039aa175d734af2aca9ad0b06cd9212ee5d1a189b9137b5

Observation da98a552-ee9b-48e8-84d4-184c000d3738 · outbound

This paper cites Deep hand: How to train a cnn on 1 million hand images when your data is continuous and weakly labelled.

Gesture-Aware Pretraining and Token Fusion for 3D Hand Pose Estimation Deep hand: How to train a cnn on 1 million hand images when your data is continuous and weakly labelled

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T10:09:55.956261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-05-15T10:09:08.788131Z digest=sha256:8c7dd1bd2acc205ff20a26f49078f61c33ad4a57abc524b3c2f01b2b1bbf9d52

Observation 745a307d-92c3-4607-b54d-358d99af6509 · outbound

This paper cites Im2hands: Learning attentive implicit representa- tion of interacting two-hand shapes.

Gesture-Aware Pretraining and Token Fusion for 3D Hand Pose Estimation Im2hands: Learning attentive implicit representa- tion of interacting two-hand shapes

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T10:09:55.959033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-05-15T10:09:08.788131Z digest=sha256:a3b0fa3c302427e324ae46cd4517478ec2deabe9e06688e6c1b49124b67d1a61

Observation 17bb5d8f-76f2-4e5c-80d8-496a443334bc · outbound

This paper cites Interacting attention graph for single image two-hand reconstruction.

Gesture-Aware Pretraining and Token Fusion for 3D Hand Pose Estimation Interacting attention graph for single image two-hand reconstruction

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T10:09:55.927951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-05-15T10:09:08.788131Z digest=sha256:8d1787167792ebd991685b18421ac5e2c9923ba4c79833b0140c39908960d683

Observation 9c242acd-1ce8-46d6-9956-b782f3aa81e6 · outbound

This paper cites Interhand2.

Gesture-Aware Pretraining and Token Fusion for 3D Hand Pose Estimation Interhand2

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T10:09:55.948284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-05-15T10:09:08.788131Z digest=sha256:744d9f97391d697f483e1e63cb169c4c9b548be587e918a38b7cd14f2af818a5

Observation b9609050-e55e-4e54-ad6e-b84c13efb5ee · outbound

This paper cites Extract-and-adaptation network for 3d interacting hand mesh recovery.

Gesture-Aware Pretraining and Token Fusion for 3D Hand Pose Estimation Extract-and-adaptation network for 3d interacting hand mesh recovery

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T10:09:55.930646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-05-15T10:09:08.788131Z digest=sha256:1a0e3e1d7aa8c08389803ffdd5144e21699bf74ccd25a8838a1e2421009d94c5

Observation 70249d40-d82d-449c-93a0-db387fb8fb9f · outbound

This paper cites Recon- structing hands in 3d with transformers.

Gesture-Aware Pretraining and Token Fusion for 3D Hand Pose Estimation Recon- structing hands in 3d with transformers

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T10:09:55.992110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-05-15T10:09:08.788131Z digest=sha256:1aa1992614e29e3b9a261fe9db3d9acb5e77dda2cca4334f467b393dd8628587

Observation 41961a09-5d10-4470-b7a2-60b006ae50a6 · outbound

This paper cites Wilor: End-to-end 3d hand localization and reconstruction in-the-wild.

Gesture-Aware Pretraining and Token Fusion for 3D Hand Pose Estimation Wilor: End-to-end 3d hand localization and reconstruction in-the-wild

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T10:09:55.981239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-05-15T10:09:08.788131Z digest=sha256:91447779977bdc34967a5793cff0179de1775f37d8cc033faf22ff52868a3b56

Observation 1fa789ba-baad-4b41-a281-4c054f0ea301 · outbound

This paper cites Em- bodied hands: Modeling and capturing hands and bodies to- gether.

Gesture-Aware Pretraining and Token Fusion for 3D Hand Pose Estimation Em- bodied hands: Modeling and capturing hands and bodies to- gether

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T10:09:55.968478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-05-15T10:09:08.788131Z digest=sha256:6f321b4f8397637f3337d931523335c99761e2d6b768ec0198218a57e612f422

Observation cd6c8106-3372-4294-b1dc-b6541dd54cd1 · outbound

This paper cites Frankmocap: A monocular 3d whole-body pose estimation system via re- gression and integration.

Gesture-Aware Pretraining and Token Fusion for 3D Hand Pose Estimation Frankmocap: A monocular 3d whole-body pose estimation system via re- gression and integration

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T10:09:55.971683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-05-15T10:09:08.788131Z digest=sha256:4792dc2a83bc01770a1ed5f8098a0eb8f70dff57a442597c4988c0af265a9d3a

Observation 61bea5fb-e3e3-452c-9869-025bf2dd4c7c · outbound

This paper cites Deep high-resolution representation learning for human pose es- timation.

Gesture-Aware Pretraining and Token Fusion for 3D Hand Pose Estimation Deep high-resolution representation learning for human pose es- timation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T10:09:55.987349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-05-15T10:09:08.788131Z digest=sha256:9253ad4830c05657a38e86d104b683b139c23f7cba020c7fe5417ef2a510b436

Observation 9a5bf11b-bb65-448b-87d5-3b5a833fd2fc · outbound

This paper cites High-Resolution Representations for Labeling Pixels and Regions.

Gesture-Aware Pretraining and Token Fusion for 3D Hand Pose Estimation High-Resolution Representations for Labeling Pixels and Regions

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-05-15T10:09:55.374303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-05-15T10:09:08.788131Z digest=sha256:5c327e7baee96cd4b0a7e4ffcb98c735423cf8dbb309a82a5de4aff700874d46

Observation e36968a3-417a-40fb-802d-5e93df4bdb91 · outbound

This paper cites Deep high-resolution repre- sentation learning for visual recognition.IEEE transactions on pattern analysis and machine intelligence, 43(10):3349– 3364.

Gesture-Aware Pretraining and Token Fusion for 3D Hand Pose Estimation Deep high-resolution repre- sentation learning for visual recognition.IEEE transactions on pattern analysis and machine intelligence, 43(10):3349– 3364

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T10:09:55.974920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-05-15T10:09:08.788131Z digest=sha256:cb287214165c0a8ee099c3bc7a04088ed6974d5a636f3d8489684c84ec517404

Observation 5f75b008-5df5-47bd-a253-634456106509 · outbound

This paper cites BiHand: Recovering Hand Mesh with Multi-stage Bisected Hourglass Networks.

Gesture-Aware Pretraining and Token Fusion for 3D Hand Pose Estimation BiHand: Recovering Hand Mesh with Multi-stage Bisected Hourglass Networks

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-15T10:09:55.372954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-05-15T10:09:08.788131Z digest=sha256:b00d9d6f8df2f1be48be9831055940acaa561ef400f4ff0aa33d09d6277c9bdc

Observation bb34265d-a893-40d4-a77d-14f514cec13b · outbound

This paper cites Acr: Attention collaboration-based regres- sor for arbitrary two-hand reconstruction.

Gesture-Aware Pretraining and Token Fusion for 3D Hand Pose Estimation Acr: Attention collaboration-based regres- sor for arbitrary two-hand reconstruction

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T10:09:55.965736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-05-15T10:09:08.788131Z digest=sha256:b7484eddd13091b927d3cface3dc11b8cb0bc54970c0c6567309ba63d06eb537

Observation a865a8db-3dec-473f-a747-822ef0a517aa · outbound

This paper cites On the continuity of rotation representations in neural networks.

Gesture-Aware Pretraining and Token Fusion for 3D Hand Pose Estimation On the continuity of rotation representations in neural networks

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T10:09:55.962955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-05-15T10:09:08.788131Z digest=sha256:cd92563e700c1bce0f9b15fbd97070dcf1609cd3175872355ac2385a9343eaee

Observation bdeb5adb-f73c-41da-b8d1-7f2c62fbd5ff · outbound

This paper cites Monocular real- time hand shape and motion capture using multi-modal data.

Gesture-Aware Pretraining and Token Fusion for 3D Hand Pose Estimation Monocular real- time hand shape and motion capture using multi-modal data

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T10:09:55.978207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-05-15T10:09:08.788131Z digest=sha256:cc8d641425f90d47078e907efe659929892e25964181f2afb00c11b19f904fad

Observation 75bb7e56-a00e-49af-9717-409064a9047d · outbound

This paper cites Freihand: A dataset for markerless capture of hand pose and shape from single rgb images.

Gesture-Aware Pretraining and Token Fusion for 3D Hand Pose Estimation Freihand: A dataset for markerless capture of hand pose and shape from single rgb images

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T10:09:55.984325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-05-15T10:09:08.788131Z digest=sha256:83e9c0a5f135de2c974fbaedb9bb8b5701ffd6bb3627d0fb0fce587fae7c571f

Pith citing papers

No inbound Pith citation observations are available.