Pith. sign in

Paper Citation Record · LEDGER

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes

As of 19 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 0 inbound Pith citation observations for arXiv:2608.10886.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.10886 v1

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T15:03:04.897476Z

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

40 of 40 outbound references displayed

  • verified exact2
  • verified fuzzy29
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 609006d0-bda2-466f-bc87-51031f7d2d2f · outbound

This paper cites Describe anything anywhere at any moment,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Describe anything anywhere at any moment,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T15:03:04.660370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:03:04.660370Z digest=sha256:2c163f437ddd4356c1fe642e698d68db05e6060457c5b4f26fa36de5587f7a44

Observation edaa7b90-29e3-4669-b5a3-d161cc246c38 · outbound

This paper cites Towards spatio-temporal world scene graph generation from monocular videos,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Towards spatio-temporal world scene graph generation from monocular videos,

Reference 2

Resolution
verified exact
raw_fallback, observed 2026-08-12T15:03:05.246521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.667213Z digest=sha256:bedd5e51b4ada8c57383cccf3f4a6da8b8625475ea9e2364e313e1fc1b059f2c

Observation 3eee583b-93d6-4e6b-8c3c-93a3b26b22c4 · outbound

This paper cites Worth Remembering: Surprise-Gated Robot Episodic Memory.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Worth Remembering: Surprise-Gated Robot Episodic Memory

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T15:03:04.672635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:03:04.672635Z digest=sha256:bad8364740275baa2fe042ae705ecc835582a9a699618383411eb282fdfa1d90

Observation 7a870ba7-e522-47ce-9c04-67e384ccaa49 · outbound

This paper cites A V A: A video dataset of spatio-temporally localized atomic visual actions,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes A V A: A video dataset of spatio-temporally localized atomic visual actions,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:03:05.903178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.679755Z digest=sha256:671519b6c5c13620365e029076943bb32e7773fb8570d4a09a86932cff1b999b

Observation c091b86b-8a03-464e-b8e5-a399f3c1214f · outbound

This paper cites Action genome: Ac- tions as compositions of spatio-temporal scene graphs,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Action genome: Ac- tions as compositions of spatio-temporal scene graphs,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:03:05.885225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.685091Z digest=sha256:9e3cb8e9fa1bedd4f432b9cdfa18a2ee0c81dc4b87f519fec642b9008a67261c

Observation 8982389d-8e2d-4bfe-af93-822c135fa6aa · outbound

This paper cites Moma: Multi-object multi-actor activity parsing,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Moma: Multi-object multi-actor activity parsing,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:03:05.865386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.690592Z digest=sha256:2ce182755c74643b1939390b8edfdf3e0795116689d622382aa81f6dc055a0ff

Observation a874839e-cc7f-4d11-ae9f-8582d4f1da5c · outbound

This paper cites Ego4d goal-step: Toward hierarchical understanding of procedu- ral activities,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Ego4d goal-step: Toward hierarchical understanding of procedu- ral activities,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:03:05.847370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.697023Z digest=sha256:55aac580780190ad473d51249217ba2f1be3d91ad1ac5a308998eb7f2b44af38

Observation 356aa0b3-f757-4d72-a1bc-ddfacc9c2d61 · outbound

This paper cites Event-grounding graph: Uni- fied spatio-temporal scene graph from robotic observations,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Event-grounding graph: Uni- fied spatio-temporal scene graph from robotic observations,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:03:05.828813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.702508Z digest=sha256:12835b43af4df68826af62096f5c1b4387f114084126d9312b72101751fdde1e

Observation 5bef97e0-45fb-4c91-9ed0-8b5af750825d · outbound

This paper cites 3d scene graph: A structure for unified semantics, 3D space, and camera,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes 3d scene graph: A structure for unified semantics, 3D space, and camera,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:03:05.811590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.708359Z digest=sha256:36c024370a7f8c986a5ac21fab09e71bd58a574a388ac234636d94dcecf6e297

Observation b94c1532-2771-4303-93b7-423a142bc8c8 · outbound

This paper cites Hier- archical Open-V ocabulary 3D Scene Graphs for Language-Grounded Robot Navigation,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Hier- archical Open-V ocabulary 3D Scene Graphs for Language-Grounded Robot Navigation,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:03:05.794354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.714270Z digest=sha256:6c94022c52be000295287d7ecc3515780575c1868d4267e441edc20bafab1df8

Observation 124d26ed-7556-467b-aa29-890dcbafafeb · outbound

This paper cites Foundations of spatial perception for robotics: Hierarchical representations and real-time systems,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Foundations of spatial perception for robotics: Hierarchical representations and real-time systems,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:03:05.777204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.725424Z digest=sha256:64fb1cff89e0887211a591843c3daa1b8d57c223f4ef6cc7bd41fbcc19e429a0

Observation 36a65fd1-6a52-459c-bd91-9e0d83d7154d · outbound

This paper cites Keysg: Hierar- chical keyframe-based 3d scene graphs,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Keysg: Hierar- chical keyframe-based 3d scene graphs,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:03:05.759729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.731230Z digest=sha256:f3555f6e793baf2abb33027800fe0c3c17d0e5591d4268bc734983ac8ac5073b

Observation 6fee0534-6b4e-4b96-aae2-4198e99c6ca5 · outbound

This paper cites 3D Scene Graphs: Open Challenges and Future Directions.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes 3D Scene Graphs: Open Challenges and Future Directions

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T15:03:04.737115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:03:04.737115Z digest=sha256:c83db2917049b95c99ec3801ab56ec24b2e3588fe41b8af1c582d112eae89acb

Observation caf0d940-6f01-4224-9d0f-ab55b4c4feac · outbound

This paper cites Hydra: A Real-time Spatial Perception System for 3D Scene Graph Construction and Optimization.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Hydra: A Real-time Spatial Perception System for 3D Scene Graph Construction and Optimization

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T15:03:04.744184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:03:04.744184Z digest=sha256:4bda544592afb19680f889c6bf1e86a4ed5601edc614db44fbbb9c6ac52ff12e

Observation bd696861-8ddb-4ebd-8231-bc4d72e0ad2a · outbound

This paper cites Conceptgraphs: Open-vocabulary 3D scene graphs for perception and planning,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Conceptgraphs: Open-vocabulary 3D scene graphs for perception and planning,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:03:05.741263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.750680Z digest=sha256:fe3fe27c70bfcd07315a8847cf5929a859137eeaaab8d98f15a47ed84422cd98

Observation 70e24170-5f17-4db7-9f0e-c165e553e2d4 · outbound

This paper cites Hi- erarchical open-vocabulary 3D scene graphs for language-grounded robot navigation,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Hi- erarchical open-vocabulary 3D scene graphs for language-grounded robot navigation,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:03:05.719997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.756398Z digest=sha256:85858eba0e66fed2792a24d06a13ecacb52eaf5e31d90a0da24036d21ab4a09e

Observation 7a844798-92ff-4749-ba09-3a8bf163324e · outbound

This paper cites Clio: Real-time task-driven open-set 3D scene graphs,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Clio: Real-time task-driven open-set 3D scene graphs,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:03:05.700047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.762061Z digest=sha256:94b49dc3a8b17436154ed83249a679c592ccbdc49b8fcf983637d3eb152e173a

Observation 356ecfc7-7906-4d3d-9854-593c16b043ac · outbound

This paper cites Open-vocabulary functional 3D scene graphs for real- world indoor spaces,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Open-vocabulary functional 3D scene graphs for real- world indoor spaces,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:03:05.680840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.768062Z digest=sha256:2719f0bcc7f9c8364dedd775e1ff436b5b8d243e613bef95e2dcb5fe7cf6da1e

Observation c58c59b2-e3ae-44d5-b570-f671e445db54 · outbound

This paper cites Fungraph: Functionality aware 3D scene graphs for language-prompted scene interaction,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Fungraph: Functionality aware 3D scene graphs for language-prompted scene interaction,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:03:05.661045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.773876Z digest=sha256:063447c2e83d339638d9f0de2390482bfcfb1bb3423d8f769a0fa384f91f96e1

Observation 6bab9d3e-6c6f-450d-a7e2-2ddb4378841e · outbound

This paper cites Ashita: Automatic scene-grounded hierarchical task analysis,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Ashita: Automatic scene-grounded hierarchical task analysis,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:03:05.639662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.779756Z digest=sha256:c0d3fd9e8363a35f552cda2afa25f43fb35009fcb7d45bbd4845c9596808de27

Observation cdb6a8cc-6e97-40ad-b04d-729b1f6638e0 · outbound

This paper cites 3d dynamic scene graphs: Actionable spatial perception with places, objects, and humans,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes 3d dynamic scene graphs: Actionable spatial perception with places, objects, and humans,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:03:05.617433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.786724Z digest=sha256:cbfdd584510bfd686140fcf9bce0c34b7cd69e24cea2d7b997f8f174c552e535

Observation 7bc32eb3-5495-45e6-ab30-04921c4fe262 · outbound

This paper cites Kimera: From SLAM to spatial perception with 3D dynamic scene graphs,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Kimera: From SLAM to spatial perception with 3D dynamic scene graphs,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:03:05.595863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.792293Z digest=sha256:651ceaf4b11959fccd61f789f0a77fbff20292882bebfb41ae4984b18cc4d627

Observation 8b71e6bf-3da0-4109-9ce8-56f6f19f44f7 · outbound

This paper cites Khronos: A unified approach for spatio-temporal metric-semantic SLAM in dynamic envi- ronments,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Khronos: A unified approach for spatio-temporal metric-semantic SLAM in dynamic envi- ronments,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:03:05.576433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.797763Z digest=sha256:44078b53469fddbc9b829b3840d736bd35e03765a2112f66dbc23558660db5d7

Observation 8073c7dd-1f0d-4356-8181-9ac3d3fbd8ee · outbound

This paper cites Aion: Towards hierarchical 4D scene graphs with temporal flow dynamics,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Aion: Towards hierarchical 4D scene graphs with temporal flow dynamics,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:03:05.555807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.803377Z digest=sha256:9310a27dc23590fb3ed430fa43413124bb59fffcb7e5e9d089632e468b19067e

Observation 9aa7f07a-3948-4bf9-8ad3-03e1c73144ef · outbound

This paper cites Remembr: Building and reasoning over long-horizon spatio-temporal memory for robot navigation,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Remembr: Building and reasoning over long-horizon spatio-temporal memory for robot navigation,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T15:03:04.808870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:03:04.808870Z digest=sha256:0a833e1018332751ee5609e86793115caa61a8a3d268f7a97f9a1194b0bd8117

Observation e5f51535-7c3e-47c8-9cfe-702697a39021 · outbound

This paper cites Embodied-RAG: General Non-parametric Embodied Memory for Retrieval and Generation.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Embodied-RAG: General Non-parametric Embodied Memory for Retrieval and Generation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T15:03:04.814128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:03:04.814128Z digest=sha256:fd3e438c10e3fc088ef767d0b68075aa4d804e4f478430ab0b5a29fce59d5bcb

Observation 0a7afc51-3f4c-4b15-b074-e39e6d666be5 · outbound

This paper cites Long-Term Planning Around Humans in Domestic Environments with 3D Scene Graphs.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Long-Term Planning Around Humans in Domestic Environments with 3D Scene Graphs

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-08-12T15:03:05.023775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.819586Z digest=sha256:261bc50086d09a6f8a6280b8209ee85cefbf3ddae3ac8a7dd9d9c4e4e060f28b

Observation cbe90bb7-239e-4f7d-acd6-ba0a0ebee589 · outbound

This paper cites Social 3d scene graphs: Modeling human actions and relations for interactive service robots,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Social 3d scene graphs: Modeling human actions and relations for interactive service robots,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:03:05.522355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.824985Z digest=sha256:9268aca895bfa39b48a6882a56e41608d064adff2a9b638c98ce14c14dca59eb

Observation 19b27015-0332-4e61-9690-1c206fdab684 · outbound

This paper cites Event segmentation,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Event segmentation,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:03:05.503322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.830926Z digest=sha256:a4182206dc9b6667a710af1adb28fa95fab3b0714fa2664b33b15de98ffb5218

Observation 0d5aef68-62f2-4aa8-b444-fd23570c2e58 · outbound

This paper cites The language of actions: Re- covering the syntax and semantics of goal-directed human activities,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes The language of actions: Re- covering the syntax and semantics of goal-directed human activities,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:03:05.484667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.836274Z digest=sha256:a55a2a05b0fa9e532e91b9ce484d4fd83d2aaad7ca1ac4efb5b47df071d34ce8

Observation 922550e5-2c38-4621-97a7-04aa3a9f5731 · outbound

This paper cites 4D panoptic scene graph generation,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes 4D panoptic scene graph generation,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:03:05.465959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.841426Z digest=sha256:a6b17f644c4224c235da044401112e6cff50c13134de2389d7f2ea7da033d60a

Observation 597700f0-6387-4321-8397-ecddc790cd98 · outbound

This paper cites Action scene graphs for long-form understanding of egocentric videos,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Action scene graphs for long-form understanding of egocentric videos,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:03:05.446128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.847064Z digest=sha256:83b11a9bc46af1e839d018740a9ba601db616afe5f7a36e679cd875886c14384

Observation db2bdb32-5953-4ce2-b903-74b66cb1edb1 · outbound

This paper cites Building a mind palace: Structuring environment-grounded semantic graphs for effective long video analysis with LLMs,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Building a mind palace: Structuring environment-grounded semantic graphs for effective long video analysis with LLMs,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:03:05.427826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.852709Z digest=sha256:cc0ec486300517a6cda52b930e3f747c6d31cd0b35ac5727045ff37119aa127e

Observation cab0803a-1a3d-4c37-b43a-30757d7753cb · outbound

This paper cites Ego4d: Around the world in 3,000 hours of egocentric video,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Ego4d: Around the world in 3,000 hours of egocentric video,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:03:05.406403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.859207Z digest=sha256:6391e34ff65266353bce07a72e68a94ac289f421941488541e8c3c1fbaa0ed7b

Observation de8ab2f1-2a15-4b63-8291-4765f6b3fdb3 · outbound

This paper cites The epic-kitchens dataset: Collection, challenges and baselines,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes The epic-kitchens dataset: Collection, challenges and baselines,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:03:05.387884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.864554Z digest=sha256:f7be4194f9bd54ff399faf6fb4bb21cb79dc76ae5a7a30b805688d0ce0a43268

Observation 6cdcc482-030d-4983-8d1d-fe281b52494a · outbound

This paper cites EgoLive: A Large-Scale Egocentric Dataset from Real-World Human Tasks.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes EgoLive: A Large-Scale Egocentric Dataset from Real-World Human Tasks

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T15:03:04.870411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:03:04.870411Z digest=sha256:de900c077edcdd4188e5aaa9e687399483110a3f2d9dc779e418368fa6fd9eb8

Observation 014c149e-ccba-42dd-8363-35b7f8198fdd · outbound

This paper cites Cosmos-Reason2-8B: Physical ai common sense and embodied reasoning model,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Cosmos-Reason2-8B: Physical ai common sense and embodied reasoning model,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:03:05.369643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.877863Z digest=sha256:6563686d2c3fce66ea84480ced4f60c20f2fd047229a0ecc1b77f45361e6f603

Observation aa835902-7575-4184-9de6-b5698407eb22 · outbound

This paper cites Fast Segment Anything.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Fast Segment Anything

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-12T15:03:04.883808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:03:04.883808Z digest=sha256:e47513e5c8bbdb6161c52e3f240f60a0e4880310beb2e55bec02f77d65da5c9e

Observation bbc49c2b-ec06-471f-be6b-6aaf4fe206bc · outbound

This paper cites Sentence-t5: Scalable sentence encoders from pre-trained text-to-text models,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Sentence-t5: Scalable sentence encoders from pre-trained text-to-text models,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T15:03:04.890991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:03:04.890991Z digest=sha256:fb56e4e9910fbcbfd9e079ac66dbce3ca0b788b025a539a4517bdde6a6434193

Observation 772c1934-9ab1-45c2-a26b-cb1c640df169 · outbound

This paper cites Hoi4d: A 4d egocentric dataset for category- level human-object interaction,.

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes Hoi4d: A 4d egocentric dataset for category- level human-object interaction,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:03:05.338395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T15:03:04.897476Z digest=sha256:6498b4487dbb5dfe0ff459f8683fc0c1ca5e8ffbb39958e7464b6edb78cce52f

Pith citing papers

No inbound Pith citation observations are available.