Pith. sign in

Paper Citation Record · LEDGER

Atla Selene Mini: A General Purpose Evaluation Model

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2501.17195.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.17195 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:34:50.298845Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T06:49:38.206304Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 10d6e181-a23c-4326-a1ea-b681e34819c8 · inbound

Reward Reasoning Model cites this paper.

Reward Reasoning Model Atla Selene Mini: A General Purpose Evaluation Model

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T15:34:50.298845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:34:50.298845Z digest=sha256:ef7be815a331f91f9f517a6cb4d67711d84c873192017de70e46f14623b6d7ba

Observation 473fb600-3bd8-4f37-939b-d734da71ea0f · inbound

ReliableEval: A Recipe for Stochastic LLM Evaluation via Method of Moments cites this paper.

ReliableEval: A Recipe for Stochastic LLM Evaluation via Method of Moments Atla Selene Mini: A General Purpose Evaluation Model

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:16.293260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:20:16.293260Z digest=sha256:751b6688c3f57eb5bbf8dace4530f5c11c92ee6c52db85920d7070408917d6c6

Observation b21ac2cc-0103-47d7-a5f5-69cad4632d49 · inbound

PaTaRM: Bridging Pairwise and Pointwise Signals via Preference-Aware Task-Adaptive Reward Modeling cites this paper.

PaTaRM: Bridging Pairwise and Pointwise Signals via Preference-Aware Task-Adaptive Reward Modeling Atla Selene Mini: A General Purpose Evaluation Model

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-18T03:20:49.325679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T03:15:57.744706Z digest=sha256:604fdf2d714af3405c9303d6964a8d452c77aaf244cb9a30d2d7338483a2886d

Observation 186743ae-0d1a-452e-8e87-9a8674819b32 · inbound

Reward Modeling for Reinforcement Learning-Based LLM Reasoning: Design, Challenges, and Evaluation cites this paper.

Reward Modeling for Reinforcement Learning-Based LLM Reasoning: Design, Challenges, and Evaluation Atla Selene Mini: A General Purpose Evaluation Model

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-03T03:04:43.002417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:04:43.002417Z digest=sha256:1d8a02a12bfcb932672e68f19c767742021b309f776af8b50bfc3593233c6afe

Observation 633b275c-4d7e-406b-9fe7-66f5b8e3a17a · inbound

Toward Robust LLM-Based Judges: Taxonomic Bias Evaluation and Debiasing Optimization cites this paper.

Toward Robust LLM-Based Judges: Taxonomic Bias Evaluation and Debiasing Optimization Atla Selene Mini: A General Purpose Evaluation Model

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T05:56:17.811091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:56:17.811091Z digest=sha256:b4fe49856e6ba0cdcf3c49044a9223ba76d389ee7eec94c96eef7a1e6fc5e314

Observation 8646f0f6-5d56-4f6d-8d75-6de409b7d501 · inbound

CompliBench: Benchmarking LLM Judges for Compliance Violation Detection in Dialogue Systems cites this paper.

CompliBench: Benchmarking LLM Judges for Compliance Violation Detection in Dialogue Systems Atla Selene Mini: A General Purpose Evaluation Model

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:26:02.445512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T14:56:00.449776Z digest=sha256:2327eab467be264d8806b6997b8f4161d5edf8170aac21bf271474d139061152

Observation 2c34bab5-fb6a-464f-8d10-984958b188a9 · inbound

VERDI: Single-Call Confidence Estimation for Verification-Based LLM Judges via Decomposed Inference cites this paper.

VERDI: Single-Call Confidence Estimation for Verification-Based LLM Judges via Decomposed Inference Atla Selene Mini: A General Purpose Evaluation Model

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T01:52:05.633574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T01:48:29.633350Z digest=sha256:3fc61f97152eb54ba3eae61ef6fdb57a8a94eafc419520c4e0411fe0df5b6a20

Observation 2d4c3f1a-13ab-4a46-ba7c-0ad48e2eb519 · inbound

A Finite-Calibration Regime Map for LLM Judge Panels cites this paper.

A Finite-Calibration Regime Map for LLM Judge Panels Atla Selene Mini: A General Purpose Evaluation Model

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T21:06:13.157300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T17:34:28.258222Z digest=sha256:e1d49f11d175ec8bae2b5155cd1cb84a57e1d621849b4519d79078b6901f6f7a

Observation b61684d5-7b12-4020-adf8-8846be308705 · inbound

Decoupled Smart Contract Audits: Lightweight LLM Framework via Distillation and Aggregation cites this paper.

Decoupled Smart Contract Audits: Lightweight LLM Framework via Distillation and Aggregation Atla Selene Mini: A General Purpose Evaluation Model

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T03:26:29.997093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T09:58:30.611774Z digest=sha256:cf4834418079c7984d157c224ac25ebd3158ac000480330edb4b676bd27d3c2f

Observation a097e6de-8e70-4f5b-b512-32b3998d4134 · inbound

Counsel: A Meta-Evaluation Dataset for Agentic Tasks cites this paper.

Counsel: A Meta-Evaluation Dataset for Agentic Tasks Atla Selene Mini: A General Purpose Evaluation Model

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-04T06:49:38.208141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T14:07:59.446478Z digest=sha256:88caed37d464b1eda33bf7b8d48d618dedfbbc3310cd3b570df4a3e858bdc584

Observation 69427740-b102-4ced-8a95-6852b925c18c · inbound

Persistent Sparse Autoencoders: Learning Feature Timescales in Language Models cites this paper.

Persistent Sparse Autoencoders: Learning Feature Timescales in Language Models Atla Selene Mini: A General Purpose Evaluation Model

Reference 167

Resolution
unresolved
no resolver link, observed 2026-08-01T19:02:46.756858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T19:02:46.756858Z digest=sha256:601c16aa3cbe0abda5d1e34f8b4c933534a90ae9c521f776e4771cd4daa99645

Observation 1072ece7-ec15-493e-843c-e6c2c3ebaf38 · inbound

Reflection or Re-Generation? Why LLM Revision Fails Where Human Revision Succeeds cites this paper.

Reflection or Re-Generation? Why LLM Revision Fails Where Human Revision Succeeds Atla Selene Mini: A General Purpose Evaluation Model

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T17:29:12.599582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T17:29:12.599582Z digest=sha256:099958293f799aefff72bc921846490c2a55e1584fcd79a53eeb5a3e7d8c192f