Pith. sign in

Paper Citation Record · LEDGER

Composite Reward Design in PPO-Driven Adaptive Filtering

As of 8 August 2026, this Paper Citation Record lists 11 of 11 outbound references and 0 inbound Pith citation observations for arXiv:2506.06323.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.06323 v2

Coverage vector

measured 11 of 11 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:40:50.234790Z

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

11 of 11 outbound references displayed

  • verified exact0
  • verified fuzzy10
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3c8aee21-0035-4f9a-8356-6061eed19e58 · outbound

This paper cites Widrow and S.

Composite Reward Design in PPO-Driven Adaptive Filtering Widrow and S

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:53.254426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:40:49.366247Z digest=sha256:692930b99e98de8ea81530d74a9520a70c609f0ada5aac183a7d367fd65d1f00

Observation 0c8fe8a1-8059-47f9-a4e8-989c34c684cb · outbound

This paper cites Haykin, Adaptive Filter Theory , 5th ed.

Composite Reward Design in PPO-Driven Adaptive Filtering Haykin, Adaptive Filter Theory , 5th ed

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:53.009927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:40:49.452380Z digest=sha256:fd1f88e80a005fa38e6b8ef659b3a349d6e8d050c3e1e06b45d9912a502dceb8

Observation 6e5e1a67-48b4-4a92-b379-4d6401d8d055 · outbound

This paper cites A new approach to linear filtering and prediction problems,.

Composite Reward Design in PPO-Driven Adaptive Filtering A new approach to linear filtering and prediction problems,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:51.701720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:40:49.514475Z digest=sha256:6f4890d1377ce0422525b3908e15138ccd6b15eff615c3b0fe45197e318eb673

Observation 8a71db63-14b9-4900-ab64-bf8015c37014 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Composite Reward Design in PPO-Driven Adaptive Filtering Proximal Policy Optimization Algorithms

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:49.608574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:49.608574Z digest=sha256:bf732e82304c5c483c701e3b9fbe592469d1a25186056b56957a6c6d359609bc

Observation 2b08d94c-328d-4b3b-be5b-da2341566ea1 · outbound

This paper cites Channel estimation via successive denoising in MIMO- OFDM systems: a reinforcement learning approach,.

Composite Reward Design in PPO-Driven Adaptive Filtering Channel estimation via successive denoising in MIMO- OFDM systems: a reinforcement learning approach,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:51.376705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:40:49.665086Z digest=sha256:ff092865e33d74142f4c87ae421ddbc7318ccd41a1b098cd350b34d074235fec

Observation 403fe8ea-26ea-4d55-9620-35e71d7ccf8b · outbound

This paper cites Adaptive filtering algorithm based on reinforcement learning,.

Composite Reward Design in PPO-Driven Adaptive Filtering Adaptive filtering algorithm based on reinforcement learning,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:51.206030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:40:49.763547Z digest=sha256:9e73f3c830bbafda9c7354d87d63e9ede7310ccebf0826ae8aee3cf7562e4630

Observation cdced8ad-8335-4e11-8999-9ad71a5a3bf6 · outbound

This paper cites A fractional filter based on reinforcement learning for effective tracking under impulsive noise,.

Composite Reward Design in PPO-Driven Adaptive Filtering A fractional filter based on reinforcement learning for effective tracking under impulsive noise,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:51.028859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:40:49.870251Z digest=sha256:72a1a24db2ec6b4cb450bec4336c40f15a97e4dfac46ed13835441785494fb25

Observation 2eeaa10a-75f4-4a6a-aae9-d82215b318d8 · outbound

This paper cites Reinforcement learning adaptive Kalman filter for AE signal’s AR-mode denoise,.

Composite Reward Design in PPO-Driven Adaptive Filtering Reinforcement learning adaptive Kalman filter for AE signal’s AR-mode denoise,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:50.879560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:40:49.971626Z digest=sha256:c015317dc86eac08f20ce47582573ecbc9b0b704da24df151c4202c3102ca52f

Observation b0715b7e-f8ed-446f-a8b9-1af84787da06 · outbound

This paper cites Beyond static obstacles: integrating Kalman filter with reinforcement learning for drone navigation,.

Composite Reward Design in PPO-Driven Adaptive Filtering Beyond static obstacles: integrating Kalman filter with reinforcement learning for drone navigation,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:50.737417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:40:50.077465Z digest=sha256:cf85a052119e0c2d22060fde2ba3c0dc5d376e9b72b541d77dc821195c42b0e2

Observation 279bae49-1e4c-4a45-ba93-e1eaefdec071 · outbound

This paper cites GFANC-RL: reinforcement learning-based generative fixed-filter active noise control,.

Composite Reward Design in PPO-Driven Adaptive Filtering GFANC-RL: reinforcement learning-based generative fixed-filter active noise control,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:50.573462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:40:50.165670Z digest=sha256:29dbb5b5845ed96ad6b010e7ac46d42268aab5a55f0d4bb3d4ba12464fc7b4bf

Observation 81efcc73-f3a0-4682-ad7b-d1498e7fd065 · outbound

This paper cites Exploration by random network distillation,.

Composite Reward Design in PPO-Driven Adaptive Filtering Exploration by random network distillation,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:50.451669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:40:50.234790Z digest=sha256:808c59e4e4903438cf19c7f24d0dd29fc5bc226a550d4c63ec576e98c75029f2

Pith citing papers

No inbound Pith citation observations are available.