Pith. sign in

Paper Citation Record · LEDGER

Waterfall Transformer for Multi-person Pose Estimation

As of 16 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2411.18944.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.18944 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T10:45:40.882265Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

28 of 28 outbound references displayed

  • verified exact1
  • verified fuzzy20
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 49b2c69f-a853-4185-98c1-55aa3cb81cc3 · outbound

This paper cites OmniPose: A Multi-Scale Framework for Multi-Person Pose Estimation.

Waterfall Transformer for Multi-person Pose Estimation OmniPose: A Multi-Scale Framework for Multi-Person Pose Estimation

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-12T10:45:41.014544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T10:45:40.754930Z digest=sha256:0886012940aaf04fa792552a94b732877873942d3fd250782ecf6f60367b087f

Observation 3c1b5c4e-a886-4c86-a2ab-22003cca9fc9 · outbound

This paper cites Unipose+: A unified framework for 2d and 3d human pose es- timation in images and videos.

Waterfall Transformer for Multi-person Pose Estimation Unipose+: A unified framework for 2d and 3d human pose es- timation in images and videos

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:45:41.319831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T10:45:40.761083Z digest=sha256:36e18d82be942e0a9d7ebb4a2955fb858942f5e101044d621b7e042d0604c0d8

Observation 1b14281e-ec90-4655-9db5-e876def5fd98 · outbound

This paper cites BAPose: Bottom-up pose estimation with disentangled water- fall representations.

Waterfall Transformer for Multi-person Pose Estimation BAPose: Bottom-up pose estimation with disentangled water- fall representations

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:45:41.305921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T10:45:40.765786Z digest=sha256:389102dda0594bf6a574d96b7981a0f975e5a58a8854e2a4f183b6e4b8757d90

Observation d239611b-80e3-46f0-92db-5cd14179a6aa · outbound

This paper cites Full-BAPose: Bottom up framework for full body pose estimation.

Waterfall Transformer for Multi-person Pose Estimation Full-BAPose: Bottom up framework for full body pose estimation

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:45:41.290399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T10:45:40.771217Z digest=sha256:5400148ad111ea5d73f754174c99a9a61757ca3c5e10a2ff02929c7fbfa72e69

Observation 1c42810f-6717-44a8-bbad-538185bc1cef · outbound

This paper cites Realtime multi-person 2d pose estimation us- ing part affinity fields.

Waterfall Transformer for Multi-person Pose Estimation Realtime multi-person 2d pose estimation us- ing part affinity fields

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:45:41.277295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T10:45:40.776181Z digest=sha256:458bd799d09aaf0a7d84d17f202038dfc04eff0d92e113fc20dc2577ab1693ef

Observation 8bf586b1-2653-4038-848a-11750dc2696f · outbound

This paper cites Yuille, and Xiaogang Wang.

Waterfall Transformer for Multi-person Pose Estimation Yuille, and Xiaogang Wang

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:45:41.263225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T10:45:40.780889Z digest=sha256:1ab68ffb7fbcfca05bda8b9d8ebacfacf675455f0a01e64302ef04f7254f09d2

Observation c2146e82-d257-474f-b143-50aff9440978 · outbound

This paper cites Openmmlab pose estimation toolbox and benchmark.

Waterfall Transformer for Multi-person Pose Estimation Openmmlab pose estimation toolbox and benchmark

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:45:41.246717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T10:45:40.786489Z digest=sha256:14528347e3f1caf9a9eb2e189000b47f6507b7cf021c6ec65e044a9957fcd0be

Observation cedd7a66-f382-4208-bc40-5c12a5f30c62 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Waterfall Transformer for Multi-person Pose Estimation An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T10:45:40.791075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:45:40.791075Z digest=sha256:a46df87c5cc4a9ee6a6ed886a6dd8d6a833fc6101ee504f0199f8424b91eeb55

Observation 6f70adb9-6405-4ea1-93da-b13d67a8cbab · outbound

This paper cites Dilated Neighborhood Attention Transformer.

Waterfall Transformer for Multi-person Pose Estimation Dilated Neighborhood Attention Transformer

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T10:45:40.795479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:45:40.795479Z digest=sha256:5ebe209987688c5f57fba1c2d24aed24fbc064ac9c0603d67ef4a252ae3e1e9a

Observation 6d842eda-a9d8-4aa0-965a-cf1d7799f921 · outbound

This paper cites Deep residual learning for image recognition.

Waterfall Transformer for Multi-person Pose Estimation Deep residual learning for image recognition

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:45:41.232533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T10:45:40.800136Z digest=sha256:3f99984bfddb1b1859cffe81e4da6388c01b0daa10bfefa356ae7ab77be725dd

Observation 946f6bbe-7478-4852-9fd0-061826f8e1e1 · outbound

This paper cites Rethinking on Multi-Stage Networks for Human Pose Estimation.

Waterfall Transformer for Multi-person Pose Estimation Rethinking on Multi-Stage Networks for Human Pose Estimation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T10:45:40.804871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:45:40.804871Z digest=sha256:9b250125d650aa072ac6601a3fd2ea41e6d61ee9ff1eda29f08b6356861d0883

Observation 44929b41-d367-4730-8037-b72c42a43728 · outbound

This paper cites Token- Pose: Learning keypoint tokens for human pose es- timation.

Waterfall Transformer for Multi-person Pose Estimation Token- Pose: Learning keypoint tokens for human pose es- timation

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:45:41.218227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T10:45:40.809858Z digest=sha256:e0f2d765a6a2adaa4273bcf87a27aad90f4eb057b8ed9c14f6f96cf3bd393f45

Observation 0428547c-f1ef-4ccb-889c-ccd4a12be50c · outbound

This paper cites Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Doll´ar, and C.

Waterfall Transformer for Multi-person Pose Estimation Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Doll´ar, and C

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:45:41.204328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T10:45:40.814490Z digest=sha256:a205cb9e165ef15c8ab163d35f822f7239b5c539099f241cdaa949cc84c30089

Observation 51547b5c-ead1-4f7c-892f-42e806023fad · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows.

Waterfall Transformer for Multi-person Pose Estimation Swin transformer: Hierarchical vision transformer using shifted windows

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:45:41.188848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T10:45:40.819136Z digest=sha256:10d258d60286fdceaef3b02db3b6ce3a5559016131cee44e63fa5526b0754c92

Observation e14903af-0820-4d31-9532-4f5f4a3d2d0b · outbound

This paper cites Stacked hourglass networks for human pose estimation.

Waterfall Transformer for Multi-person Pose Estimation Stacked hourglass networks for human pose estimation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:45:41.172765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T10:45:40.823482Z digest=sha256:2cee65c2c87df551fcd96bf742ac49f49bb9f40cf8853921500b1328c8bb8f1d

Observation a8dec8df-d24b-4455-8348-d01eeae7c86d · outbound

This paper cites On the Convergence of Adam and Beyond.

Waterfall Transformer for Multi-person Pose Estimation On the Convergence of Adam and Beyond

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T10:45:40.827758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:45:40.827758Z digest=sha256:468adb205cd6112e5b9d78b6bfa520288c8c1ff10f4d8bf9a33923aa4a8be894

Observation 93cef3a9-45e3-4152-8840-d8d7348ef792 · outbound

This paper cites 15 keypoints is all you need.

Waterfall Transformer for Multi-person Pose Estimation 15 keypoints is all you need

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:45:41.158898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T10:45:40.832587Z digest=sha256:8ba688ef440fdd1c576bef7ac4bb76bcf46b206baacec2886f566717167121d3

Observation ef167870-978c-41a5-9cc3-17b8960760a3 · outbound

This paper cites Deep high-resolution representation learning for hu- man pose estimation.

Waterfall Transformer for Multi-person Pose Estimation Deep high-resolution representation learning for hu- man pose estimation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:45:41.144722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T10:45:40.837133Z digest=sha256:b9d96295c49e16cd5269ee85b9236e9537aa7f766b937f91d471f6558100c9d2

Observation 850a7f89-1548-4479-a50a-d22121947c87 · outbound

This paper cites Deeply learned com- positional models for human pose estimation.

Waterfall Transformer for Multi-person Pose Estimation Deeply learned com- positional models for human pose estimation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:45:41.129878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T10:45:40.841675Z digest=sha256:0749db31eb76aae75a414b213e407c2818f5341c0616343189608afd1ccb256f

Observation f00b7b4f-58f2-4949-aa41-eba39dbcf6d8 · outbound

This paper cites DeepPose: Human pose estimation via deep neural networks.

Waterfall Transformer for Multi-person Pose Estimation DeepPose: Human pose estimation via deep neural networks

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:45:41.115506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T10:45:40.846447Z digest=sha256:ed3d965e6ff10fddb39d81a9b29c40db37b47bb2e14fe90c116b9caf5e1e571e

Observation b6d9ea08-9453-47e2-9a25-da2b23b31e43 · outbound

This paper cites Attention is all you need.

Waterfall Transformer for Multi-person Pose Estimation Attention is all you need

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T10:45:40.851188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:45:40.851188Z digest=sha256:f6fec4b03cd3812238a06e0b2f82ed27a0f6ffac5f38aa9b78061f21860ff53b

Observation df18c7d0-eed9-413a-ab73-a4f08c1d7a0e · outbound

This paper cites Deep high- resolution representation learning for visual recogni- tion.

Waterfall Transformer for Multi-person Pose Estimation Deep high- resolution representation learning for visual recogni- tion

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:45:41.089839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T10:45:40.855807Z digest=sha256:9c4b76f0c373cf45379839710c3288e9926ac1c4c7dda0a2cf554f82bbf5ff8a

Observation 106afd55-cb4b-4ac3-ba3a-149ba594317e · outbound

This paper cites Convolutional pose machines.

Waterfall Transformer for Multi-person Pose Estimation Convolutional pose machines

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:45:41.074381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T10:45:40.860167Z digest=sha256:46b26d1feb9a1cf8a6eecb8058c1a308383a6b25808d00ad04b0da13b8c96775

Observation a2362cc5-7c70-4a12-b54e-ec1b8cacb17c · outbound

This paper cites ViTPose: Simple Vision Transformer Baselines for Human Pose Estimation.

Waterfall Transformer for Multi-person Pose Estimation ViTPose: Simple Vision Transformer Baselines for Human Pose Estimation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T10:45:40.864432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:45:40.864432Z digest=sha256:b6cad987705562d87f408427b33c2eb729b98c554c25afe0f4a68b11038418ad

Observation 1d8b589b-af4d-4921-8095-9f669c78bed3 · outbound

This paper cites TransPose: Keypoint localization via transformer.

Waterfall Transformer for Multi-person Pose Estimation TransPose: Keypoint localization via transformer

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:45:41.060053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T10:45:40.869277Z digest=sha256:bff33cb56b6a54809a9f707878a19945f4702c3fcd9f6ed315deac74251498a5

Observation d48e2c71-5285-4b0b-9c18-d769ce02cad0 · outbound

This paper cites HRFormer: High-resolution vision transformer for dense predict.

Waterfall Transformer for Multi-person Pose Estimation HRFormer: High-resolution vision transformer for dense predict

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:45:41.044908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T10:45:40.873458Z digest=sha256:36faf620ff7380fd3a4759ae95ece3a139a1e0601bf2e391e91bfb57e05af002

Observation a88c9097-a826-491e-9fbd-0f3d5403951a · outbound

This paper cites Human Pose Estimation with Spatial Contextual Information.

Waterfall Transformer for Multi-person Pose Estimation Human Pose Estimation with Spatial Contextual Information

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T10:45:40.877764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:45:40.877764Z digest=sha256:8b84683c3973729a3e03fca43b9814750c09fc641d9957fdd568e7d238772f05

Observation 52b4b00b-bf0e-453b-b94e-51f4350897ed · outbound

This paper cites 3d human pose estimation with spatial and temporal transform- ers.

Waterfall Transformer for Multi-person Pose Estimation 3d human pose estimation with spatial and temporal transform- ers

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:45:41.030614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T10:45:40.882265Z digest=sha256:48a8b445ca0f3177482b9dbe224e75ab2a4282fbaf4edf22d0e1b0be03ebdb57

Pith citing papers

No inbound Pith citation observations are available.