Pith. sign in

Paper Citation Record · LEDGER

Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning

As of 15 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 1 inbound Pith citation observation for arXiv:2411.19732.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.19732 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T05:58:28.147883Z

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:45:12.118469Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

28 of 28 outbound references displayed

  • verified exact2
  • verified fuzzy12
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e5215204-8d73-4759-8c12-ca645677c05c · outbound

This paper cites Learning agile and dynamic motor skills for legged robots.

Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning Learning agile and dynamic motor skills for legged robots

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T05:58:27.822299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:58:27.822299Z digest=sha256:e58d6c545569b2f9a9e8d853c02a4459ba9850766fd68a77d55003c46e2557c2

Observation b5acf4bf-a2e4-4ea8-839c-9794a77f651a · outbound

This paper cites Learning dexterous in-hand manipulation,.

Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning Learning dexterous in-hand manipulation,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:58:29.322014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:58:27.828199Z digest=sha256:01228fd2c19f1cc92198891bcb28fa28b6deabf37617ec59e35417ffcfc8fc2f

Observation 0cf07051-7ab1-4b72-9430-8c2decbcbed4 · outbound

This paper cites Sim-to-Real: Learning Agile Locomotion For Quadruped Robots.

Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning Sim-to-Real: Learning Agile Locomotion For Quadruped Robots

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T05:58:27.838369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:58:27.838369Z digest=sha256:fafbb8fb72a4c6879fa826e0b219c7a601c42bf9da1c9f500b7cc6ce22eba2d6

Observation 7cab7f34-cf34-40ca-92cb-ecc3f83400c6 · outbound

This paper cites Domain Randomization for Transferring Deep Neural Networks from Simulation to the Real World.

Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning Domain Randomization for Transferring Deep Neural Networks from Simulation to the Real World

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T05:58:27.857376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:58:27.857376Z digest=sha256:0376f8d54bf156898febbf04a42ec75ca350dc140e5ae3c2c0bd95dde36aa982

Observation 93fffa95-15c9-4f45-bf94-6de1761eca25 · outbound

This paper cites Pods: Policy optimization via differentiable simulation,.

Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning Pods: Policy optimization via differentiable simulation,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:58:29.307024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:58:27.880528Z digest=sha256:445372657b0fde9d5cd7db9ae2fe1f5079c17986c39bcb72f081cc861400019b

Observation 5fae4561-3444-40b2-9626-2c4430c9b7e5 · outbound

This paper cites Contact Models in Robotics: a Comparative Analysis.

Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning Contact Models in Robotics: a Comparative Analysis

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-12T05:58:28.356098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:58:27.915241Z digest=sha256:88d86fc16142e1f3f17508941ef88f1ff5c0fa5ea054dc484053a2e0c8ad9e1d

Observation d1e928db-488d-4818-ba65-b66fa8bb294e · outbound

This paper cites A comparative analysis of contact models in trajectory optimization for manipulation,.

Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning A comparative analysis of contact models in trajectory optimization for manipulation,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:58:29.103455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:58:27.982861Z digest=sha256:47c4b365ab3aafc3508b90865dd32bb8d37bd5f954076bf4d01bb2192e49fa87

Observation 7e1a0f29-e96e-4b8b-b2f6-2e3d97c438bf · outbound

This paper cites Rethinking Optimization with Differentiable Simulation from a Global Perspective.

Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning Rethinking Optimization with Differentiable Simulation from a Global Perspective

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-12T05:58:28.332885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:58:27.988542Z digest=sha256:86ad23001ea51a5fa9febd81d01c9b3098d671bf082a8997a0ca2f6f807df7ec

Observation 545d37bf-9109-48b9-bfdb-a0e333718917 · outbound

This paper cites Do differentiable simulators give better policy gradients?.

Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning Do differentiable simulators give better policy gradients?

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:58:29.087617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:58:27.993425Z digest=sha256:9ad5676510b36d60d810de59987402b314053686d68a0901b0e0dcb7d261242d

Observation f5a0c475-cb60-4796-b2ef-5f44034b9299 · outbound

This paper cites Brax -- A Differentiable Physics Engine for Large Scale Rigid Body Simulation.

Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning Brax -- A Differentiable Physics Engine for Large Scale Rigid Body Simulation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T05:58:27.998425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:58:27.998425Z digest=sha256:f2f44916299f096e71dcc81693e091f6f499b111c1ee7e87fa225e563cefb036

Observation 52593a4e-99e8-423d-96f9-e1fb800cbb63 · outbound

This paper cites DiffPD: Differentiable Projective Dynamics.

Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning DiffPD: Differentiable Projective Dynamics

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T05:58:28.003416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:58:28.003416Z digest=sha256:b599ac8fae837de368c318101a65c893c1d8554ad95f5a413df458b436628dd6

Observation 0d58ba7f-5de9-4a0d-8a83-fc4721ff6437 · outbound

This paper cites Dojo: A differentiable physics engine for robotics,.

Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning Dojo: A differentiable physics engine for robotics,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:58:29.063345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:58:28.008505Z digest=sha256:db462cf0d2a23855df51cdf23841bd19675a9968545096e382e8a1f33e41434d

Observation a76f7bf1-5a96-4569-8834-8c9f328ea81e · outbound

This paper cites ADD: Analytically Differentiable Dynamics for Multi-Body Systems with Frictional Contact.

Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning ADD: Analytically Differentiable Dynamics for Multi-Body Systems with Frictional Contact

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T05:58:28.012979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:58:28.012979Z digest=sha256:e9cdf2620968e4554e09271efddd6f4334bdb7be7e614570ffb981ee91ff3da0

Observation 67505436-490c-4a8b-a1ce-3caa20701c1c · outbound

This paper cites Acceler- ated policy learning with parallel differentiable simulation,.

Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning Acceler- ated policy learning with parallel differentiable simulation,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:58:28.876086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:58:28.018312Z digest=sha256:598ae33dfdeb9627818ad4dfd8dc51af7e3d52fea65337b925bfaffa75f99602

Observation cc7979b7-fc1c-4642-ba61-9f96ad7cc931 · outbound

This paper cites Adaptive Horizon Actor-Critic for Policy Learning in Contact-Rich Differentiable Simulation.

Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning Adaptive Horizon Actor-Critic for Policy Learning in Contact-Rich Differentiable Simulation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T05:58:28.022995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:58:28.022995Z digest=sha256:2891e3dc32677406f3295c2ae890922d098abc7611e6845df335852aca3c66e5

Observation e5083952-ada2-4da2-89b8-d78accd094e4 · outbound

This paper cites Barkour: Benchmarking Animal-level Agility with Quadruped Robots.

Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning Barkour: Benchmarking Animal-level Agility with Quadruped Robots

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T05:58:28.027984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:58:28.027984Z digest=sha256:3ca70a38611cf6df5a51cf539faf89dca1bff0b2996fa1066266de2e981cc7c8

Observation 30478da4-1c69-4209-bf8f-fe49c1353420 · outbound

This paper cites Data-efficient domain randomization with bayesian optimization,.

Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning Data-efficient domain randomization with bayesian optimization,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:58:28.852093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:58:28.032605Z digest=sha256:893e13d6c2caa8dfe24b703c5fe19ea19670d08e88325d8d6578f5acbe7e2b4b

Observation a97b0479-9778-409d-8c4f-299b26705e71 · outbound

This paper cites Long short-term memory,.

Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning Long short-term memory,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T05:58:28.072125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:58:28.072125Z digest=sha256:a2f1112545ce71d6bec1cb22bf0c30a6ca289442b9327eba9767abf01854fd88

Observation 417528ba-57ad-49e8-9961-6a557148a9ad · outbound

This paper cites On Large-Batch Training for Deep Learning: Generalization Gap and Sharp Minima.

Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning On Large-Batch Training for Deep Learning: Generalization Gap and Sharp Minima

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T05:58:28.102397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:58:28.102397Z digest=sha256:68a6d0a3a6f24c13356305aaba67950cd599162d34d0958065b4e2ffaba1d853

Observation b2226780-1ba7-4fd0-9def-dd730d305d0c · outbound

This paper cites Sharpness-Aware Minimization for Efficiently Improving Generalization.

Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning Sharpness-Aware Minimization for Efficiently Improving Generalization

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T05:58:28.116815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:58:28.116815Z digest=sha256:f144b5d9435d593b6ceedcad7403b93c03d17f8b1736cec60652bc782d5ab71e

Observation b9ff14db-ce32-451a-b4e5-afd32c38ba06 · outbound

This paper cites Asam: Adaptive sharpness-aware minimization for scale-invariant learning of deep neural networks,.

Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning Asam: Adaptive sharpness-aware minimization for scale-invariant learning of deep neural networks,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:58:28.837181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:58:28.121395Z digest=sha256:dea1380fcbe3567690397107a054825c403b63465ae6be0be8e6fd16df7fc239

Observation 7df25e25-6b9e-4bc8-a327-c49cb45f6773 · outbound

This paper cites Sharpness-aware training for free,.

Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning Sharpness-aware training for free,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:58:28.788876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:58:28.125847Z digest=sha256:51857017dff083cb5ff370d06283d4c77bb1eb67f3ca383c6aac890498c1d0b7

Observation dc355f70-1ac2-425f-9bd0-6e88467df202 · outbound

This paper cites Efficient sharpness-aware minimization for improved training of neural networks,.

Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning Efficient sharpness-aware minimization for improved training of neural networks,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:58:28.624031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:58:28.130111Z digest=sha256:1dc8dbba5dba0b2dca30aed971a77c44f9e239e82934ae3abb13119928b7f3e4

Observation 45495eb9-c3f3-4d75-b9cd-e0d35593eb99 · outbound

This paper cites Proximal policy optimization algorithms,.

Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning Proximal policy optimization algorithms,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T05:58:28.134396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:58:28.134396Z digest=sha256:43051c10f68a00f60c37af5c4cd1f27fcd4e9c5ec5470f707a3df52468a8c046

Observation 6ad1927a-be49-4a6d-bb25-2b7896d6638f · outbound

This paper cites sam: Sharpness-aware minimization for efficiently improving generalization,.

Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning sam: Sharpness-aware minimization for efficiently improving generalization,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:58:28.596644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:58:28.138793Z digest=sha256:aeabbeeb5500bf9aba29d5cf2aab4b8f97cab85a5162b88da791f32514aa73c1

Observation b1b42a46-7d0f-4d76-a421-ac623338addc · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning Adam: A Method for Stochastic Optimization

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T05:58:28.143359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:58:28.143359Z digest=sha256:e4cb9c3bf798114544a5bfa8583ce85cf615f479fe7caf58b970636b22960a3d

Observation e3985710-d03e-4907-8284-f7e5dc21ccaa · outbound

This paper cites Mujoco: A physics engine for model-based control,.

Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning Mujoco: A physics engine for model-based control,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:58:28.484568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T05:58:28.147883Z digest=sha256:80be29ad6394d1c802448bb4c5bd612b4ec88b4ca1ee258d222fd22cd9f864b6

Observation 7fc0b2b3-f7c3-4e2a-b348-7a0afe578416 · outbound

This paper cites Learning Dexterous In-Hand Manipulation.

Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning Learning Dexterous In-Hand Manipulation

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-12T05:58:27.833097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:58:27.833097Z digest=sha256:a73330ccf423de96578e1d49b57addaf518ae62f8b4b882c542be654c8b57a46

Pith citing papers

Observation c54ccec5-d083-4a8f-91d1-1281e659ab9b · inbound

AI4Research: A Survey of Artificial Intelligence for Scientific Research cites this paper.

AI4Research: A Survey of Artificial Intelligence for Scientific Research Improving generalization of robot locomotion policies via Sharpness-Aware Reinforcement Learning

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T20:45:12.118469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:45:12.118469Z digest=sha256:14efee3f89662495571cbd0f10e71cef1abda413d6f3a0a04fcd6ca3992e70cd