Pith. sign in

Paper Citation Record · LEDGER

DMRL: Data- and Model-aware Reward Learning for Data Extraction

As of 17 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 0 inbound Pith citation observations for arXiv:2505.06284.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.06284 v1

Coverage vector

measured 52 of 52 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:39:04.641425Z

measured 52 of 52 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

52 of 52 outbound references displayed

  • verified exact1
  • verified fuzzy1
  • unresolved50
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e1d7eeb1-b463-48fc-9c01-d22917479c94 · outbound

This paper cites online" 'onlinestring :=.

DMRL: Data- and Model-aware Reward Learning for Data Extraction online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.457313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.457313Z digest=sha256:76eaf698d35bed544abd75301d19bd97f33ab15102ee0f366bc09c4e6faae5a5

Observation 4e7f6344-02ad-4cab-807f-548d5b6bd192 · outbound

This paper cites write newline.

DMRL: Data- and Model-aware Reward Learning for Data Extraction write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.461585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.461585Z digest=sha256:d25a7856c58b7cdccdf7c14cf1092938190777a81ab31f2989128169b3360137

Observation 29e127dd-6ea0-448f-a466-f15275bc9c70 · outbound

This paper cites Yelp Dataset Challenge: Review Rating Prediction.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Yelp Dataset Challenge: Review Rating Prediction

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.465762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.465762Z digest=sha256:236a527aefe0370334b08639341e46fb9cf9b435259a555b9b96a62775a4a044

Observation d88d3cb3-e4bc-464a-abcf-66dd8bb0a939 · outbound

This paper cites Qwen Technical Report.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Qwen Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.469796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.469796Z digest=sha256:43d1fbdeaab8bafe612c7086e7bc3769cc65ac327540387f48e3374ea902a38d

Observation 4932c4dd-7e71-404d-8cdb-e266a645fad5 · outbound

This paper cites Baichuan 2: Open Large-scale Language Models.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Baichuan 2: Open Large-scale Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.473604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.473604Z digest=sha256:a020315b71f95fda08ef205086f97791988d7a496ee1eff6198908a932a449f0

Observation 50793b11-d42d-4dd6-a124-7243f9100ae6 · outbound

This paper cites an unresolved cited work.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-15T23:39:05.149866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T23:39:04.477128Z digest=sha256:5cb37be49a185570bf904bec5e8222281b6a74bbe7893d3cea1accbaea5a9e53

Observation 1ef7cd50-4d48-4e3d-bef8-e7327f50adba · outbound

This paper cites an unresolved cited work.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.480564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.480564Z digest=sha256:62f2c8a98151639e28de2910ba03c34734462c876f4d70dcdb94309043bf09e9

Observation ef7ad25e-7668-47ea-b82f-cf774fdd06b8 · outbound

This paper cites an unresolved cited work.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.483889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.483889Z digest=sha256:989f1c57634556e0ab31972d95ff8c1f26b772e03db265d9c424001ff67b181c

Observation 34eed6f2-825d-4e7f-b923-04336f9b3846 · outbound

This paper cites Neural Legal Judgment Prediction in English.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Neural Legal Judgment Prediction in English

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.487300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.487300Z digest=sha256:28e38608b9624a3f83d833c6ff459fc2d57451f3525465d862b9108e61f2f0da

Observation 35f351b9-ae24-4542-8a2a-78f126e4c45c · outbound

This paper cites an unresolved cited work.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-15T23:39:05.128137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T23:39:04.494110Z digest=sha256:09849c0f753b672e319dd07416466de88df5db1e061b4a57fcaa6908a8168e22

Observation b768bc67-380c-44e9-b288-60e2377e64f4 · outbound

This paper cites Gibberish is All You Need for Membership Inference Detection in Contrastive Language-Audio Pretraining.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Gibberish is All You Need for Membership Inference Detection in Contrastive Language-Audio Pretraining

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.497355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.497355Z digest=sha256:9456b8b85e5c5af93e6fb305eb5f0c8bab58b571cb5f3687f18c81d576201d1a

Observation 034bfd2e-9c15-4bc4-a705-19dcf4e7f151 · outbound

This paper cites Reinforcement Learning from Multi-role Debates as Feedback for Bias Mitigation in LLMs.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Reinforcement Learning from Multi-role Debates as Feedback for Bias Mitigation in LLMs

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.501096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.501096Z digest=sha256:bd3da4f33b5ff4c9398f5192f39441c8b50eea0f5b4286e647c4e810175db685

Observation 4b873466-2852-4ffe-b4f0-879ba74078d5 · outbound

This paper cites an unresolved cited work.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.504533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.504533Z digest=sha256:fc7025390a5f3bbe7c054448fd98cbbf3d64ae99414018587047b1fa14f23fde

Observation d3dcbfd1-b99a-44f6-afa9-d36a2907aabd · outbound

This paper cites an unresolved cited work.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.507504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.507504Z digest=sha256:2aa7ce2fe970698549cbad1f3f56bcf9e0bbd47985757b25befb03e244e9255b

Observation 232640c4-2297-45f0-af11-d329903fd17e · outbound

This paper cites ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools.

DMRL: Data- and Model-aware Reward Learning for Data Extraction ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.510607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.510607Z digest=sha256:70dd20964850f41aac561abb5d9210e2f76ad0a3e8a5b5899d4f5a9940f8ffa5

Observation 2bbdf3bb-9a31-4673-901c-443605bc29d5 · outbound

This paper cites The Llama 3 Herd of Models.

DMRL: Data- and Model-aware Reward Learning for Data Extraction The Llama 3 Herd of Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.514067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.514067Z digest=sha256:5ca25fb22e3030e42ebbc1a8d89e1186b63d64bb5feaa507d31f8bbc00d7b04e

Observation 63394824-2a46-4a70-9655-97948711b3c3 · outbound

This paper cites Supervised Contrastive Learning for Pre-trained Language Model Fine-tuning.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Supervised Contrastive Learning for Pre-trained Language Model Fine-tuning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.517664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.517664Z digest=sha256:8267e6ea857e0731d9eb3f02dd62c7b38e889f54a98c10b6008361a3d933f727

Observation 1a73e93b-3a55-494c-a66a-5f2563fa23f1 · outbound

This paper cites Training Compute-Optimal Large Language Models.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Training Compute-Optimal Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.522083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.522083Z digest=sha256:540a7eb961357b001f92f4937c4e4cf617566a76e5379e65fd942f940562818b

Observation de95eea4-198e-4f93-a001-a4c26cf28bf1 · outbound

This paper cites an unresolved cited work.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-15T23:39:05.111253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T23:39:04.525578Z digest=sha256:5c034e98dea9f4c0fa9baf7fe6b9189228165fe6a4b575f361679ebf38539fdc

Observation 6f712e05-036e-4e1d-a116-a46589878d2c · outbound

This paper cites Training Data Leakage Analysis in Language Models.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Training Data Leakage Analysis in Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.528843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.528843Z digest=sha256:6b08a53972c15a28f976cd41a1d93c922f47e257f2cae1a367fd12549bd4e4da

Observation 8f45686f-7a91-44f8-a0b6-6f73bbd1e5ca · outbound

This paper cites Measuring Forgetting of Memorized Training Examples.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Measuring Forgetting of Memorized Training Examples

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.533270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.533270Z digest=sha256:d95805945405220dd47f590f4911ca3a5d4a81bff206b84d1b53757efd0967f1

Observation d7099288-ef38-4037-8fc1-ffd689207ac4 · outbound

This paper cites an unresolved cited work.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-15T23:39:05.101337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T23:39:04.537285Z digest=sha256:4cbdd1cc01564f1426f3dd50adc9f9301149ce934d666a97be363392702a68a8

Observation c56af638-3299-40fd-80de-656b4f89dca3 · outbound

This paper cites an unresolved cited work.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.540249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.540249Z digest=sha256:08cd4c171a931846b35bfad92d74b398aba9fc6643b814e795a754f92b688bdf

Observation 2b603f1e-ed03-41fe-aa9e-cd52bf9a0bd3 · outbound

This paper cites an unresolved cited work.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.543291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.543291Z digest=sha256:31bd2d01e462f65bccb92967f4bb65d8d8e3533e85634c35a7c18a8ffe79daa9

Observation 206378ca-cd9d-4f03-9d93-fb6493062bdd · outbound

This paper cites an unresolved cited work.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.546265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.546265Z digest=sha256:2f54b9657e7f809214a3e392c69a5ce13d119854703f2d892afa18bfd3f843a3

Observation 32592992-01cb-4497-a287-6d56e90f44a0 · outbound

This paper cites Models of human preference for learning reward functions.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Models of human preference for learning reward functions

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.549420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.549420Z digest=sha256:8288c1e47e9660c654f0249ebe44a00f1a89b75f542413755a7fbe87d07f0aae

Observation 0e70d11f-103d-4448-ac26-2bf7cf4802e8 · outbound

This paper cites Does BERT Pretrained on Clinical Notes Reveal Sensitive Data?.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Does BERT Pretrained on Clinical Notes Reveal Sensitive Data?

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.552704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.552704Z digest=sha256:fd324d26dfd36166a47b134d1d446b72ed1e0467911c3dbae4e524c421b06ed1

Observation 06eb2465-f36d-4ff4-89aa-66242f99055d · outbound

This paper cites an unresolved cited work.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-15T23:39:05.072149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T23:39:04.555888Z digest=sha256:5231cbbb0a84ee0afcfdf035a7a2115dd6d195a4e75c6d83308637d62f3802a8

Observation 4d62aedb-c229-472f-97f1-672273eb6c61 · outbound

This paper cites an unresolved cited work.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-15T23:39:05.062072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T23:39:04.558899Z digest=sha256:7f1ee317b81f3b3efd757fc434dc6b83e164c8886d060562f37e2570ef56bc20

Observation 9665581d-ecd3-450f-943f-53670a5a7ea8 · outbound

This paper cites TUNI: A Textual Unimodal Detector for Identity Inference in CLIP Models.

DMRL: Data- and Model-aware Reward Learning for Data Extraction TUNI: A Textual Unimodal Detector for Identity Inference in CLIP Models

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-08-15T23:39:04.768336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T23:39:04.561946Z digest=sha256:b6c68102b73d04f84f4efc5308d793e287a866bb5cdab7329c3093ba52a0c2a7

Observation 5dbc892d-a385-4f60-a3d4-176ca4a7344b · outbound

This paper cites note bloat.

DMRL: Data- and Model-aware Reward Learning for Data Extraction note bloat

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:39:05.051949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T23:39:04.565309Z digest=sha256:e810078cc0f75850ca076900190743b6f595a83b47e7ce8ab0345ec058c80a0a

Observation 519ad9db-0f9d-4aea-a96e-b5d503063391 · outbound

This paper cites an unresolved cited work.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.568305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.568305Z digest=sha256:4412cfa4b3fa47b1142216675ed56f6218389daa9e2014d52a5a9f4418017105

Observation bbdce563-d3dd-4b5d-bdff-9d1455f41857 · outbound

This paper cites Scalable Extraction of Training Data from (Production) Language Models.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Scalable Extraction of Training Data from (Production) Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.571444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.571444Z digest=sha256:5805ad8104962b0e97ce9130be267196bf988c1b1ec4c60934c7bd16c130d05d

Observation 2a4d3b75-120b-41a8-bd8d-f688c5ee8e4e · outbound

This paper cites an unresolved cited work.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-15T23:39:05.036278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T23:39:04.575094Z digest=sha256:ec4d5ab6247b43165244aebb87973f33158932b63f9bbfdfe442665bc59033c7

Observation 9c18248a-c9b9-43c6-8403-269d70772a45 · outbound

This paper cites LeakAgent: RL-based Red-teaming Agent for LLM Privacy Leakage.

DMRL: Data- and Model-aware Reward Learning for Data Extraction LeakAgent: RL-based Red-teaming Agent for LLM Privacy Leakage

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.578072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.578072Z digest=sha256:860a2c5e8ca6a2a10cbc5676df65b49c171f226ef7afbef7949fef9526bdffd7

Observation cd7861f2-49a3-4ecc-93f4-aa614d15a2fa · outbound

This paper cites an unresolved cited work.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.581556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.581556Z digest=sha256:8cfa0e439166c4a4065ab719a2a2e83f694adf0883bf5a1832b695917aebbfc8

Observation 944a99ab-8976-4dc0-b89f-61bbc2178f63 · outbound

This paper cites SelfPrompt: Autonomously Evaluating LLM Robustness via Domain-Constrained Knowledge Guidelines and Refined Adversarial Prompts.

DMRL: Data- and Model-aware Reward Learning for Data Extraction SelfPrompt: Autonomously Evaluating LLM Robustness via Domain-Constrained Knowledge Guidelines and Refined Adversarial Prompts

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.589952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.589952Z digest=sha256:56aed1a8d2cab3f802ca71de096cea4114a8e2d442ea2a67192ee212c564135a

Observation 5d1af099-603a-4257-8fbc-1b813e669d0b · outbound

This paper cites an unresolved cited work.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.594330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.594330Z digest=sha256:34ff69925ebc635dc4e03cee44c6f2ef5a439476c1a254c32c4b44b468a6272c

Observation 2adee9b4-b7e3-4521-a674-f34024fb777a · outbound

This paper cites an unresolved cited work.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.597532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.597532Z digest=sha256:b067ff6d46c3f3ff7b22eccfacf4908a799b164d5ac2c9682af087cd70346dc7

Observation c9481b93-da40-4f44-852f-f17a2ef9f004 · outbound

This paper cites an unresolved cited work.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.600881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.600881Z digest=sha256:756f26480f65e428739059ce870875f9a9353f6a82fdbfae6d701d350fbb7599

Observation 11374d0e-88e1-4338-a95f-61be6617e568 · outbound

This paper cites an unresolved cited work.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-15T23:39:05.001079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T23:39:04.604160Z digest=sha256:7bf5331a8df7a83800f96e160be405d6028cd84c2edd893b19879d69e3e08866

Observation e03fc48e-bba8-4f5a-b8f5-5ef2716f2aa4 · outbound

This paper cites Neural Machine Translation of Rare Words with Subword Units.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Neural Machine Translation of Rare Words with Subword Units

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.607490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.607490Z digest=sha256:3216b4a8aabf50600d08a33c9dba02478e7e8c563258f25419f33044597cd4cb

Observation 7c897ac8-3555-44c3-b84f-c057702da972 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

DMRL: Data- and Model-aware Reward Learning for Data Extraction DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.611030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.611030Z digest=sha256:8199b6ca9dbda3cc572d711c1d730f1c231b0475c9f6f0f851fa8fc52711ac7f

Observation da9d2a0a-bed2-4e99-a730-9f7e5de6e995 · outbound

This paper cites an unresolved cited work.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.614370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.614370Z digest=sha256:aac1df36c505ff88779036ee9ff62d8736fc3f79f5cd1342ab9d7149eafc971d

Observation b19bc193-932d-4eb5-8407-f99fd1bd8ab1 · outbound

This paper cites Aligning Large Multimodal Models with Factually Augmented RLHF.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Aligning Large Multimodal Models with Factually Augmented RLHF

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.617630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.617630Z digest=sha256:f23d8324f626eb04bf3dfeb9e01151f60f585d2d79e12ebb56232782c4ca9c67

Observation 074983fb-27c1-496c-8869-2a8d55bc88dd · outbound

This paper cites an unresolved cited work.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-15T23:39:04.985070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T23:39:04.621043Z digest=sha256:412dffb8dab59bad93feaa03da29f4465f4d8f27fdc59c17279ebc2477a90c01

Observation 9008bd6f-6915-482c-8bec-acf8ee915d9a · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.624178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.624178Z digest=sha256:463df106a34817098b9e37bc0e6daa5572335b3315c1ed7470549b5d70c070e9

Observation c73e3805-2381-4811-9ad3-1c1406b5dd60 · outbound

This paper cites Differentially Private Fine-tuning of Language Models.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Differentially Private Fine-tuning of Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.627594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.627594Z digest=sha256:afbe13407ed75ade8f58695640459ec997b427efcb26aabb0675c63ce3269b02

Observation 36a57df2-7c08-49db-96cb-4718ccc7ad37 · outbound

This paper cites Assessing Prompt Injection Risks in 200+ Custom GPTs.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Assessing Prompt Injection Risks in 200+ Custom GPTs

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.631255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.631255Z digest=sha256:4cc197d713922760eb52f695247f21ffc29e9d9f15c6cbb4e09a72ea15bce74f

Observation 94687acb-66e8-4171-bed0-1544a715a9f0 · outbound

This paper cites an unresolved cited work.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-15T23:39:04.974969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T23:39:04.634916Z digest=sha256:300b7d3a160bfa565148f2655ebae3c0352f7d6703deb05a0a45ad23b6d55c04

Observation 2d982d1a-a949-4b68-afce-71f9f0510975 · outbound

This paper cites an unresolved cited work.

DMRL: Data- and Model-aware Reward Learning for Data Extraction Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-15T23:39:04.965018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T23:39:04.638270Z digest=sha256:1b38778ef75733bb7282f5e5b517278625c018ab25aba1da49b00c0e425f5afd

Observation f956bd2d-24cd-455b-99d4-92ce6b352209 · outbound

This paper cites From Demonstrations to Rewards: Alignment Without Explicit Human Preferences.

DMRL: Data- and Model-aware Reward Learning for Data Extraction From Demonstrations to Rewards: Alignment Without Explicit Human Preferences

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:04.641425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:04.641425Z digest=sha256:dc65e29f826af5a571b34db24ef7c013176eb761e1ce04c0281da87e25145939

Pith citing papers

No inbound Pith citation observations are available.