Pith. sign in

Paper Citation Record · LEDGER

gpt-oss-120b & gpt-oss-20b Model Card

As of 4 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 100 inbound Pith citation observations for arXiv:2508.10925.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.10925 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-10T12:22:54.633089Z

measured 135 of 135 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 100 of 604 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T10:07:51.602817Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-11T03:17:51.327125Z

Reference resolution

35 of 35 outbound references displayed

  • verified exact14
  • verified fuzzy20
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7888df68-c861-4a60-8414-2a7620199a8d · outbound

This paper cites Attention is all you need.

gpt-oss-120b & gpt-oss-20b Model Card Attention is all you need

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T12:22:54.704914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:5561e7a929f4ff780bd94d68ff1e77ae0301cf3f2669038757f63141ae14d9b9

Observation 367d5a28-444c-467c-bc8d-345f77c346ec · outbound

This paper cites Outrageously large neural networks: The sparsely-gated mixture-of-experts layer.

gpt-oss-120b & gpt-oss-20b Model Card Outrageously large neural networks: The sparsely-gated mixture-of-experts layer

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T12:22:54.708990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:b65ddf35d8c70998f570ff98648b46bb2150955edce91e2776c8d66852f9d478

Observation 538f59ae-718d-4888-aa64-61cd26c8331b · outbound

This paper cites GShard: Scaling Giant Models with Conditional Computation and Automatic Sharding.

gpt-oss-120b & gpt-oss-20b Model Card GShard: Scaling Giant Models with Conditional Computation and Automatic Sharding

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-11T02:26:45.276836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:5661d00a23ef753b5232c3c51226e6895ef40cdf0e44e71a6700250e68db131c

Observation a8fb408a-945e-482e-b082-5b959daaa7be · outbound

This paper cites Glam: Efficient scaling of language models with mixture-of-experts.

gpt-oss-120b & gpt-oss-20b Model Card Glam: Efficient scaling of language models with mixture-of-experts

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T12:22:54.710714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:fb41f69ed47da51ad83435bed5bd24f24a1e05bcd60f9c422ceb0e07bb233a59

Observation ef8c8070-bc62-4c26-84f1-cf77e011cc06 · outbound

This paper cites OCP Microscaling Formats (MX) Specification Version 1.0.

gpt-oss-120b & gpt-oss-20b Model Card OCP Microscaling Formats (MX) Specification Version 1.0

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T12:22:54.712598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:09eb210bce5c1a4d87fe11470721a7e73b09112007021441535df2aae79d8c85

Observation 28d3de21-3af0-46ac-80fe-5beecc767c0f · outbound

This paper cites Root mean square layer normalization.

gpt-oss-120b & gpt-oss-20b Model Card Root mean square layer normalization

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T12:22:54.714167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:9be8a9ec5818d55822ab15a5fb831a4c2af02b5cf5a6b82aa4768fc170e7afd7

Observation 1f8eac11-3948-41a0-a659-da11569ec28c · outbound

This paper cites On layer normalization in the transformer architecture.

gpt-oss-120b & gpt-oss-20b Model Card On layer normalization in the transformer architecture

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T12:22:54.715880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:851e3de22cf67ca009f7e262ed49f2a4d02522c3ec17e0591bdbe8bf45979665

Observation 97060b56-8331-47e5-a2fb-e49c7ddb9d47 · outbound

This paper cites Language models are unsupervised multitask learners.

gpt-oss-120b & gpt-oss-20b Model Card Language models are unsupervised multitask learners

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T12:22:54.717546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:23136673593db11625b87deb6a6f292ee76fd5ceaa4f0303eabcd86a17972604

Observation 64f2180d-7798-401c-a7d5-d88aa327af8e · outbound

This paper cites GLU Variants Improve Transformer.

gpt-oss-120b & gpt-oss-20b Model Card GLU Variants Improve Transformer

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-10T17:05:01.922266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:509eca625eef31d27b67b5f4cb0dc0af07360ff0e203747c3e74660d916fde63

Observation 5931a784-c2a6-417d-ba76-c62e198a235d · outbound

This paper cites Generating Long Sequences with Sparse Transformers.

gpt-oss-120b & gpt-oss-20b Model Card Generating Long Sequences with Sparse Transformers

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-10T19:50:53.465869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:1f46ae8521ed2c39c2305e534cd9b3fa4108638beaabc3b3b550bba3a4ec6cfe

Observation b802757f-bfcd-4463-b8c3-69f079e652ed · outbound

This paper cites Language models are few-shot learners.

gpt-oss-120b & gpt-oss-20b Model Card Language models are few-shot learners

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T12:22:54.719042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:8f782c9e9b52347e380fefe3b4b743b9203697cecf86434563be7cc9ba709c40

Observation a8037c10-e002-4899-97f0-b9aa3bf34fdd · outbound

This paper cites GQA: Training generalized multi-query transformer models from multi-head checkpoints.

gpt-oss-120b & gpt-oss-20b Model Card GQA: Training generalized multi-query transformer models from multi-head checkpoints

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T12:22:54.720603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:21c9489ea5a0555856092b52dd5225d184fc766617ab70351a8fa5d8899fc752

Observation 6528db72-f9d2-430e-97dd-9c7a14183cad · outbound

This paper cites Fast Transformer Decoding: One Write-Head is All You Need.

gpt-oss-120b & gpt-oss-20b Model Card Fast Transformer Decoding: One Write-Head is All You Need

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:50:20.836519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:5b71de50846f80831caaa109b717cffa830c104bf2d07555d613fd354ff3454a

Observation bc1b4bd2-a1a7-44e5-89b8-7692286bec9c · outbound

This paper cites Roformer: Enhanced transformer with rotary position embedding.

gpt-oss-120b & gpt-oss-20b Model Card Roformer: Enhanced transformer with rotary position embedding

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T12:22:54.722179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:33940bc3f5d87914aa0c25744f3457da016e6a19e19bae708b519d59fcbe8c6f

Observation e8fca5b8-397d-4654-ba38-6a2efbfb0d5c · outbound

This paper cites YaRN: Efficient Context Window Extension of Large Language Models.

gpt-oss-120b & gpt-oss-20b Model Card YaRN: Efficient Context Window Extension of Large Language Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:46:51.373347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:08518ef5c87efb821e0d43ecc91dccca5628708afb4d8cf111e268a2c04aa160

Observation 4401b535-4052-4027-b89f-fcaaacc48b53 · outbound

This paper cites Attention is off by one (2023).

gpt-oss-120b & gpt-oss-20b Model Card Attention is off by one (2023)

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T12:22:54.723754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:36b058205b4ca4aae6efd34577f3949b4010fd39c5c72ec5934ddd061a0b2743

Observation 77a8a474-a4fd-40a2-99a0-a060abd913a6 · outbound

This paper cites Efficient Streaming Language Models with Attention Sinks.

gpt-oss-120b & gpt-oss-20b Model Card Efficient Streaming Language Models with Attention Sinks

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:40:13.867138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:e5ffa2c02b29928a680679699d0e188e3b03625cd637da90a86f6ea372cdaa2e

Observation b038472c-c926-4b4e-b8cf-afa3b8728690 · outbound

This paper cites GPT-4o System Card.

gpt-oss-120b & gpt-oss-20b Model Card GPT-4o System Card

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-05-10T12:22:54.690119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:baf3162e18c8bb52a15ad4015053bbaa438ce2a9c3e9e8cc1a3451cdcc2aa666

Observation 4a2524a5-1bac-4e83-8ee6-cdcf3916dd2c · outbound

This paper cites Pytorch: An imperative style, high-performance deep learning library.

gpt-oss-120b & gpt-oss-20b Model Card Pytorch: An imperative style, high-performance deep learning library

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T12:22:54.725282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:0f070f9b3ca227664cf52ac561759f98130eb9a816c9f6b1725f0bfc0989508d

Observation da2b0ee3-1f47-4fdd-ab3e-f8eab6b3b2cb · outbound

This paper cites Triton: an intermediate language and compiler for tiled neural network computations.

gpt-oss-120b & gpt-oss-20b Model Card Triton: an intermediate language and compiler for tiled neural network computations

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T12:22:54.726804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:71f326dd0994dce42f76adff2790c4cd810c20f0ba6ea9f514570575301a09cd

Observation bf42d565-cee1-4af7-8586-5749d01690cd · outbound

This paper cites FlashAttention: Fast and memory-efficient exact attention with IO-awareness.

gpt-oss-120b & gpt-oss-20b Model Card FlashAttention: Fast and memory-efficient exact attention with IO-awareness

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T12:22:54.728510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:4f1425e0b7d391225c0b43ebea119088c61cfab5bede30e1c9b1cce347c4f4f0

Observation 8c40a0d1-8069-478b-b442-82803ef8a442 · outbound

This paper cites Accessed: 2025-08-04.

gpt-oss-120b & gpt-oss-20b Model Card Accessed: 2025-08-04

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T12:22:54.730096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:3b6e13b975bfdba093d77f8ecdfa63c50e44add29fb84306aedfa1a9c72a4e93

Observation 88052cb9-4fce-49c3-af0d-eff07719e2e7 · outbound

This paper cites $\tau$-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains.

gpt-oss-120b & gpt-oss-20b Model Card $\tau$-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:19:00.987380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:0bc3cd8832f8400b4b600c8296062ca8de38bde157c516de57da60e8c482c37d

Observation d95b09ba-2789-4ab5-9fb8-b4ade14cd352 · outbound

This paper cites GPQA: A graduate-level google-proof QA benchmark.

gpt-oss-120b & gpt-oss-20b Model Card GPQA: A graduate-level google-proof QA benchmark

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T12:22:54.697487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:5789d2be409c83cbe08d263584527730ed75e824a87b935a84444f34ebb7cf65

Observation 107e954f-a60a-4196-8cdc-4a5e64625ee4 · outbound

This paper cites Measuring Massive Multitask Language Understanding.

gpt-oss-120b & gpt-oss-20b Model Card Measuring Massive Multitask Language Understanding

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-10T12:43:44.770948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:bd8371ca30e11a5ae66043142d4115e72efac18876f1ff37953530cce043347d

Observation af23c177-43d3-434b-a80c-06cc5d34421c · outbound

This paper cites Humanity's Last Exam.

gpt-oss-120b & gpt-oss-20b Model Card Humanity's Last Exam

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-10T18:40:50.715344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:1a735ea7f55c87655b91f362004cd59553c0c070afa0636dea4be09dc8c9cfdf

Observation a3423fb5-8c8d-48a8-9509-dd0c471985e8 · outbound

This paper cites Introducing SWE-bench Verified.

gpt-oss-120b & gpt-oss-20b Model Card Introducing SWE-bench Verified

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T12:22:54.699302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:1c55bb3915a6884c5875046438c0e374b69e88380cfc0e10342b6a4e2ad687fb

Observation 7822bd68-c0f3-4400-bc3d-3ff4ad605eee · outbound

This paper cites HealthBench: Evaluating Large Language Models Towards Improved Human Health.

gpt-oss-120b & gpt-oss-20b Model Card HealthBench: Evaluating Large Language Models Towards Improved Human Health

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:18:21.444858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:d2216e71656f5c07a75962d30d45e2d1f03ecb5009b06bd3bfa5d837b84d056b

Observation 11e7454a-7858-43d9-98a1-03d479d9b806 · outbound

This paper cites Deliberative Alignment: Reasoning Enables Safer Language Models.

gpt-oss-120b & gpt-oss-20b Model Card Deliberative Alignment: Reasoning Enables Safer Language Models

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T12:22:54.679028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:2313a16de02f283e3061441eac3a6aac1e1d4f3c22716d292553e47f86cd5cb0

Observation 73efaf1e-f4fe-4cf4-a3af-7cacb4d30dc0 · outbound

This paper cites The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions.

gpt-oss-120b & gpt-oss-20b Model Card The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:59:30.872595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:df4c8369b94a114adef7232bc2660f34733a714e94a6daffb8aefb1e219c86ac

Observation e57586f3-9c85-4c26-b328-dacc32fa4d40 · outbound

This paper cites A StrongREJECT for Empty Jailbreaks.

gpt-oss-120b & gpt-oss-20b Model Card A StrongREJECT for Empty Jailbreaks

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-16T21:28:03.016512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:3f1d01a0002e2f342e93ab95a47f68906b5ecc295e6caf2d638c173abd6dab18

Observation 5abe97a1-7feb-4c6a-95aa-39fe5be885a6 · outbound

This paper cites BBQ: A Hand-Built Bias Benchmark for Question Answering.

gpt-oss-120b & gpt-oss-20b Model Card BBQ: A Hand-Built Bias Benchmark for Question Answering

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-16T19:55:11.638110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:4585bc8123867b01f5f0dce04e183046a3465ee34b4175207a800d38517fa5d5

Observation 98d3ec71-b74d-4249-a502-8f7529971f10 · outbound

This paper cites Building an early warning system for LLM-aided biological threat creation.

gpt-oss-120b & gpt-oss-20b Model Card Building an early warning system for LLM-aided biological threat creation

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T12:22:54.701319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:60ed3eef6cf6a329921695e6949f5cf0073d31c9aab24cb2b210bee91aacda1e

Observation 6caeec79-bbe1-4522-9c28-0d141cc22d26 · outbound

This paper cites LAB-Bench: Measuring capabilities of language models for biology research.

gpt-oss-120b & gpt-oss-20b Model Card LAB-Bench: Measuring capabilities of language models for biology research

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T12:22:54.703204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:f411e0659b8b8fe596d1f2720ed6ad6cb0ffba7f14cd7cfa7ff046398d0391ff

Observation 932a47e4-1ce0-473f-8b6e-7b1e582fee57 · outbound

This paper cites PaperBench: Evaluating ai’s ability to replicate ai research.

gpt-oss-120b & gpt-oss-20b Model Card PaperBench: Evaluating ai’s ability to replicate ai research

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T12:22:54.707157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:22:54.633089Z digest=sha256:8e68e9ba35829a5824cd6ca4dd50921350bf04c4b197fbe21de0baa2d5c86571

Pith citing papers

Observation 8224ac71-3da2-4ffc-90bc-724f433ce82d · inbound

SCAN: Structured Capability Assessment and Navigation for LLMs cites this paper.

SCAN: Structured Capability Assessment and Navigation for LLMs gpt-oss-120b & gpt-oss-20b Model Card

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-22T16:11:45.594092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T16:11:36.613334Z digest=sha256:140dfa559dbf3a741096c17d68f49602401e63a921e89921ea36ed77b23fee61

Observation 9a43790c-d91a-4019-9327-3c0b8473bd2a · inbound

Can Large Language Models Really Recognize Your Name? cites this paper.

Can Large Language Models Really Recognize Your Name? gpt-oss-120b & gpt-oss-20b Model Card

Reference 55

Resolution
verified exact
local_arxiv, observed 2026-05-22T14:01:38.531817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T13:57:23.152504Z digest=sha256:1fcbc016f3d6199e83aafa700da42d63c92b10b7caf13d060eee9b5016761168

Observation 7f3912ec-7cd5-49a9-a8fd-ae6c0687e582 · inbound

SpatialScore: Towards Comprehensive Evaluation for Spatial Intelligence cites this paper.

SpatialScore: Towards Comprehensive Evaluation for Spatial Intelligence gpt-oss-120b & gpt-oss-20b Model Card

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-22T13:11:35.648202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T13:07:11.548885Z digest=sha256:4f07d57eba9e90e82e50d8fe42fd237808fefca1a309c49ec8163514c2dff7c5

Observation 52bfb84f-9885-49df-82f9-899b94494d77 · inbound

A Survey of Reinforcement Learning for Large Reasoning Models cites this paper.

A Survey of Reinforcement Learning for Large Reasoning Models gpt-oss-120b & gpt-oss-20b Model Card

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T00:02:24.467562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-18T00:02:24.352947Z digest=sha256:acbdff934f276754956a334a7bf8eefc576a9da679e24827c45a8554daf74241

Observation a3adea81-29b1-42f4-bfd0-5d8ce14349dd · inbound

RECAP: Transparent Inference-Time Emotion Alignment for Medical Dialogue Systems cites this paper.

RECAP: Transparent Inference-Time Emotion Alignment for Medical Dialogue Systems gpt-oss-120b & gpt-oss-20b Model Card

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-05-18T16:52:43.922894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T16:52:22.768827Z digest=sha256:b4f0128ef55350e2ad2301d136a50bbeb5fbba0b41331ba96ae53d8abfe45a5e

Observation c36132f9-a8b6-4c52-a383-f5e818cbd069 · inbound

ChatInject: Abusing Chat Templates for Prompt Injection in LLM Agents cites this paper.

ChatInject: Abusing Chat Templates for Prompt Injection in LLM Agents gpt-oss-120b & gpt-oss-20b Model Card

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-18T12:26:22.491485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T12:22:47.695255Z digest=sha256:d2527a7750a3882485e0c9836646c393bd13228f7571af7bfabaa2c7f571bad8

Observation bd76bdcb-549c-4aea-bb1a-d5310309353e · inbound

Measuring Competency, Not Performance: Item-Aware Evaluation Across Medical Benchmarks cites this paper.

Measuring Competency, Not Performance: Item-Aware Evaluation Across Medical Benchmarks gpt-oss-120b & gpt-oss-20b Model Card

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:16:24.075017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-18T13:16:19.744864Z digest=sha256:5135db34696c31671a99e0f3f3664a4079219ec5cd0c73f36a7631179ca38b93

Observation b746ba71-4cd5-4706-b171-8d35dfefa9d7 · inbound

Short window attention enables long-term memorization cites this paper.

Short window attention enables long-term memorization gpt-oss-120b & gpt-oss-20b Model Card

Reference 26

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T12:11:21.718335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T12:10:42.646127Z digest=sha256:66758323938227acb6ba25e965a1bf963fc0969f4958815193ce2e5ec9f357e9

Observation 8e0eae50-2f6e-47db-a3d4-39daab604549 · inbound

Thinking Sparks!: Emergent Attention Heads in Reasoning Models During Post Training cites this paper.

Thinking Sparks!: Emergent Attention Heads in Reasoning Models During Post Training gpt-oss-120b & gpt-oss-20b Model Card

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:31:24.949019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-18T13:28:32.093512Z digest=sha256:8bd42d302602fb6f583ad358f1b9b260939710db5fd685ade6da90a2bcd31df8

Observation 49850446-dd9c-42af-ac58-0a4f4092ed92 · inbound

Beyond Linear Probes: Dynamic Safety Monitoring for Language Models cites this paper.

Beyond Linear Probes: Dynamic Safety Monitoring for Language Models gpt-oss-120b & gpt-oss-20b Model Card

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-18T12:42:36.897850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-18T12:41:48.620040Z digest=sha256:e7261aec856dc4bd017e3148ac96162c1b333aa171c9ccfc6040a4c25ec08961

Observation df33f6ee-4f76-4074-b2b1-88834be781b4 · inbound

Don't Pass@k: A Bayesian Framework for Large Language Model Evaluation cites this paper.

Don't Pass@k: A Bayesian Framework for Large Language Model Evaluation gpt-oss-120b & gpt-oss-20b Model Card

Reference 87

Resolution
verified exact
local_arxiv, observed 2026-05-18T10:06:13.759368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T10:04:39.223895Z digest=sha256:6385b1bebd7aabc6c31ddaaae9044cb8ed0945b6b2ff4a8640451d9f8543222e

Observation e6a752d6-4cbf-4660-b775-ca4242e31dfa · inbound

Hybrid Architectures for Language Models: Systematic Analysis and Design Insights cites this paper.

Hybrid Architectures for Language Models: Systematic Analysis and Design Insights gpt-oss-120b & gpt-oss-20b Model Card

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-05-18T10:21:15.052449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T10:18:04.431436Z digest=sha256:cfdb07b16c0e96e72fb0f5ae0b79c07176a09e74e379a038e9ae93c92015e1eb

Observation da1491e2-02c6-403b-8dc5-e14f2497cc05 · inbound

Efficient numeracy in language models through single-token number embeddings cites this paper.

Efficient numeracy in language models through single-token number embeddings gpt-oss-120b & gpt-oss-20b Model Card

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-21T21:10:38.612009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T21:09:02.731348Z digest=sha256:28ceb263ff2a7076dba61e5456fa98889128c5704430e63f235e7cb32a76e63c

Observation 53ae87ab-ac21-4e24-a9e3-75abbf3e4462 · inbound

When Identity Skews Debate: Anonymization for Bias-Reduced Multi-Agent Reasoning cites this paper.

When Identity Skews Debate: Anonymization for Bias-Reduced Multi-Agent Reasoning gpt-oss-120b & gpt-oss-20b Model Card

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-05-18T08:51:08.543154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T08:50:53.486588Z digest=sha256:6f90048082da1670eeb8d6705f9bc14151bcfd0fcaa98a3fd767d02181e3865a

Observation 86e131b1-df88-4605-a76c-272e8d6cda73 · inbound

DeepPrune: Parallel Scaling without Inter-trace Redundancy cites this paper.

DeepPrune: Parallel Scaling without Inter-trace Redundancy gpt-oss-120b & gpt-oss-20b Model Card

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T08:41:08.701365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T08:37:53.886355Z digest=sha256:8b15c1390afb1263e38eaef1e349d280fc88c875b882afee603b5228a8ae11e1

Observation 69e030cf-8f83-40c2-b1b1-b5a7c3c4eb9f · inbound

Are Large Reasoning Models Interruptible? cites this paper.

Are Large Reasoning Models Interruptible? gpt-oss-120b & gpt-oss-20b Model Card

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T10:07:51.602817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T10:07:51.602817Z digest=sha256:a6d719651a576032e276bce61d58e4b7398f54880fe82c972f71c825f7c09f92

Observation 2cb50164-2ae3-4303-972b-0338ac537a93 · inbound

Attention Illuminates LLM Reasoning: The Preplan-and-Anchor Rhythm Enables Fine-Grained Policy Optimization cites this paper.

Attention Illuminates LLM Reasoning: The Preplan-and-Anchor Rhythm Enables Fine-Grained Policy Optimization gpt-oss-120b & gpt-oss-20b Model Card

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-04T09:48:07.880427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T09:48:07.880427Z digest=sha256:a8b82518da1b2a95821b729706e329d90a934880cfb14bc332394503daff31ae

Observation fe29de2a-c9b5-489e-89d1-cae02f145f35 · inbound

Echoes of Human Malice in Agents: Benchmarking LLMs for Multi-Turn Online Harassment Attacks cites this paper.

Echoes of Human Malice in Agents: Benchmarking LLMs for Multi-Turn Online Harassment Attacks gpt-oss-120b & gpt-oss-20b Model Card

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T09:40:43.391936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:40:43.391936Z digest=sha256:eebb1dd67e0219d68c084f2d77ce955989a686426f8c1b0ac0dc144671c26c5d

Observation 25fabaf4-665c-4b38-a191-5bc627c0e6fb · inbound

Scaling Test-Time Compute to Achieve IOI Gold Medal with Open-Weight Models cites this paper.

Scaling Test-Time Compute to Achieve IOI Gold Medal with Open-Weight Models gpt-oss-120b & gpt-oss-20b Model Card

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-05-18T06:20:58.194101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T06:20:16.941026Z digest=sha256:3c2043cda2d1aeb3d15e172b0ed3ded39ac1a47da4ad92e644a5b76c23c66ab8

Observation 569f998a-5475-4431-8715-ae8ea0b8a4b4 · inbound

InfiMed-ORBIT: Aligning LLMs on Open-Ended Complex Tasks via Rubric-Based Incremental Training cites this paper.

InfiMed-ORBIT: Aligning LLMs on Open-Ended Complex Tasks via Rubric-Based Incremental Training gpt-oss-120b & gpt-oss-20b Model Card

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-04T09:22:40.275971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:22:40.275971Z digest=sha256:b64bb648526fd191b5ee553c78d12c906351154493c2cc0b85b17134c7cc76bf

Observation 783f9c8d-a866-477a-a022-7ebfd004bc34 · inbound

Rethinking On-policy Optimization for Query Augmentation cites this paper.

Rethinking On-policy Optimization for Query Augmentation gpt-oss-120b & gpt-oss-20b Model Card

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T09:08:01.144753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T09:08:01.144753Z digest=sha256:d169de8affbed04b4dfd9bd02fd1a010d0a80005bcc16a85c30096b8d6cbe3e4

Observation e539e60e-5f25-4409-95b4-8fee5e10abba · inbound

Food4All: A Multi-Agent Framework for Real-time Free Food Discovery with Integrated Nutritional Metadata cites this paper.

Food4All: A Multi-Agent Framework for Real-time Free Food Discovery with Integrated Nutritional Metadata gpt-oss-120b & gpt-oss-20b Model Card

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-05-18T05:35:55.937404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T05:35:18.037662Z digest=sha256:64a3f261a1809d2d8b8714cd2165ff545dbd4adcd8cba751d71b32368e8b2dd8

Observation 139ac860-4c99-4fc1-9b84-c317f2c1e61f · inbound

Empathic Prompting: Non-Verbal Context Integration for Multimodal LLM Conversations cites this paper.

Empathic Prompting: Non-Verbal Context Integration for Multimodal LLM Conversations gpt-oss-120b & gpt-oss-20b Model Card

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T08:26:45.415443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:26:45.415443Z digest=sha256:a48f80e6c0243d230791892b397b5d4f8751b0a6d14c0e3a539477f95fed8ce4

Observation 75e71e79-4307-44e4-bfcf-5cc85b0c7bd1 · inbound

Key and Value Weights Are Probably All You Need: On the Necessity of the Query, Key, Value weight Triplet in Self-Attention Transformers cites this paper.

Key and Value Weights Are Probably All You Need: On the Necessity of the Query, Key, Value weight Triplet in Self-Attention Transformers gpt-oss-120b & gpt-oss-20b Model Card

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-05-18T03:40:50.501764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T03:38:36.932424Z digest=sha256:44ad4bfae1c130080f3ee692606eb2bd8e667503bc7c98b893746ade83975b25

Observation 75a27a1d-a0e5-4c0a-8d8f-60cfd9baf605 · inbound

Kimi Linear: An Expressive, Efficient Attention Architecture cites this paper.

Kimi Linear: An Expressive, Efficient Attention Architecture gpt-oss-120b & gpt-oss-20b Model Card

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-13T23:49:10.690152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:16cf0cbfc0b7936c5dbe4aed44e89e7ef9dd654f1cf1a07c9190aafb13c25f92

Observation 27d6b107-b719-468f-b757-385f77351ba7 · inbound

Making Interpretable Discoveries from Unstructured Data: A High-Dimensional Multiple Hypothesis Testing Approach cites this paper.

Making Interpretable Discoveries from Unstructured Data: A High-Dimensional Multiple Hypothesis Testing Approach gpt-oss-120b & gpt-oss-20b Model Card

Reference 54

Resolution
verified exact
local_arxiv, observed 2026-05-18T02:00:40.042979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-18T01:56:50.978054Z digest=sha256:e22cc02360ae63fb492c3a7e2a5d41fb05c979becf297d2a67bb1f4531549731

Observation c665ecb2-240e-4d90-a5e6-44e0bc4ddbdb · inbound

Making Interpretable Discoveries from Unstructured Data: A High-Dimensional Multiple Hypothesis Testing Approach cites this paper.

Making Interpretable Discoveries from Unstructured Data: A High-Dimensional Multiple Hypothesis Testing Approach gpt-oss-120b & gpt-oss-20b Model Card

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-04T00:23:30.924882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T00:23:30.924882Z digest=sha256:f06ff1d1a0d11a4cd3bd5f50040b377eeddc162bae7a57fc41611722ef7b8ef5

Observation f5ae6872-80d7-4312-8d78-a9d33e2a4534 · inbound

Graph-Based Alternatives to LLMs for Human Simulation cites this paper.

Graph-Based Alternatives to LLMs for Human Simulation gpt-oss-120b & gpt-oss-20b Model Card

Reference 56

Resolution
verified exact
local_arxiv, observed 2026-05-18T00:50:33.319188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-18T00:47:40.432039Z digest=sha256:bd080fcdd835aad727305109ee451e2f4d7bbfaa25743874f0a1a9e50b88a26a

Observation dea793fb-0cf2-41e2-9971-e6de0342d14e · inbound

Continuum: Efficient and Robust Multi-Turn LLM Agent Scheduling with KV Cache Time-to-Live cites this paper.

Continuum: Efficient and Robust Multi-Turn LLM Agent Scheduling with KV Cache Time-to-Live gpt-oss-120b & gpt-oss-20b Model Card

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-18T02:00:39.828546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T01:58:23.234348Z digest=sha256:d8d8c9d18c8aebfec8bc42806c2058d2a88d74be9c3a23e15f35cded45c42753

Observation 549d1da0-0a3d-43c7-9cfe-96940733cfe4 · inbound

Continuum: Efficient and Robust Multi-Turn LLM Agent Scheduling with KV Cache Time-to-Live cites this paper.

Continuum: Efficient and Robust Multi-Turn LLM Agent Scheduling with KV Cache Time-to-Live gpt-oss-120b & gpt-oss-20b Model Card

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T00:19:15.718896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T00:19:15.718896Z digest=sha256:9086a790e16181933997922619aa9654d2c506b9133cc60bd02d3d10123e3d73

Observation 67641973-4980-40c6-8df6-7a123b3ef971 · inbound

PoCo: Agentic Proof-of-Concept Exploit Generation for Smart Contracts cites this paper.

PoCo: Agentic Proof-of-Concept Exploit Generation for Smart Contracts gpt-oss-120b & gpt-oss-20b Model Card

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T00:09:32.649037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T00:09:32.649037Z digest=sha256:1df53ccdfa027579767b5f8f38328008afae1946e255b1a3c70a54f8b9e87c62

Observation 4c354e2d-1044-45a5-9da8-613e0b707664 · inbound

SnapStream: Efficient Long Sequence Decoding on Dataflow Accelerators cites this paper.

SnapStream: Efficient Long Sequence Decoding on Dataflow Accelerators gpt-oss-120b & gpt-oss-20b Model Card

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-18T02:00:39.303801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T01:59:30.583027Z digest=sha256:9bcbb4ca2c46060b36e313484654f8f9c58d90a3ea77601faa2873c15c3d6994

Observation f31ee432-0c57-4f01-b740-c4b3f637ab54 · inbound

Controllably Efficient Language Models cites this paper.

Controllably Efficient Language Models gpt-oss-120b & gpt-oss-20b Model Card

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T23:33:51.967485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T23:33:51.967485Z digest=sha256:4cac1bce095e9196a63f697399d2a7254a50789ba0772d302b718b35539345a6

Observation 95f3b03c-8188-4ca1-8408-d26e8c764bcc · inbound

MiroThinker: Pushing the Performance Boundaries of Open-Source Research Agents via Model, Context, and Interactive Scaling cites this paper.

MiroThinker: Pushing the Performance Boundaries of Open-Source Research Agents via Model, Context, and Interactive Scaling gpt-oss-120b & gpt-oss-20b Model Card

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-05-17T21:45:17.893410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-17T21:44:18.744201Z digest=sha256:888c2a18b5e67c38aaabe8c452382879626b8da95c3f864518508e17146db346

Observation e300a515-39d9-4bda-bc05-8ddb285a0674 · inbound

SnapAudit: Active Auditing of Differentially Private In-Context Learning via Snapshot-Based Simulation cites this paper.

SnapAudit: Active Auditing of Differentially Private In-Context Learning via Snapshot-Based Simulation gpt-oss-120b & gpt-oss-20b Model Card

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-05-17T22:22:08.974462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-17T22:20:53.561649Z digest=sha256:62ee9f26e042f6f28c11edf8453c6a40d38223e8687f711551cf3173e5502220

Observation 8430c4bf-7659-4d2f-ba85-a1ec664c044d · inbound

MermaidSeqBench: An Evaluation Benchmark for NL-to-Mermaid Sequence Diagram Generation cites this paper.

MermaidSeqBench: An Evaluation Benchmark for NL-to-Mermaid Sequence Diagram Generation gpt-oss-120b & gpt-oss-20b Model Card

Reference 21

Resolution
metadata mismatch
local_arxiv, observed 2026-05-17T20:10:11.230663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-17T20:06:54.566512Z digest=sha256:207262be5c432cfad9ce12a8c0c32e0d7a93bdc3e13c12be2d8dc7697658ab93

Observation 53025c00-124c-42e2-9fb6-49121f1b115f · inbound

MUSEKG: A Knowledge Graph Over Museum Collections cites this paper.

MUSEKG: A Knowledge Graph Over Museum Collections gpt-oss-120b & gpt-oss-20b Model Card

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-25T07:25:28.957332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T07:23:59.599434Z digest=sha256:969c876fde9919e4689fc54a04eafc1acab75e87e4b298b354e744764bee08b5

Observation 389b57f9-ce6d-4a14-83b3-642f5e7d60be · inbound

Point of Order: Action-Aware LLM Persona Modeling for Data-Grounded Civic Deliberation cites this paper.

Point of Order: Action-Aware LLM Persona Modeling for Data-Grounded Civic Deliberation gpt-oss-120b & gpt-oss-20b Model Card

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T20:57:36.859421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T20:57:36.859421Z digest=sha256:288d8300c001bdbfe2ffb5dc348832965de63dfa938e358e3ed7d8a3b3bb5108

Observation 6f867a5a-a4b0-4c2a-9594-deb9cff5b5d7 · inbound

Cognitive Alpha Mining via LLM-Driven Code-Based Evolution cites this paper.

Cognitive Alpha Mining via LLM-Driven Code-Based Evolution gpt-oss-120b & gpt-oss-20b Model Card

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-05-17T06:44:10.643804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-17T06:43:49.283270Z digest=sha256:62f30e1dacdd837bb645469c8ac45e088f2f6a94f01af734cf8ba7b566977c25

Observation f7043090-d40e-4751-b01f-7e3ffdb7fc6c · inbound

Cognitive Alpha Mining via LLM-Driven Code-Based Evolution cites this paper.

Cognitive Alpha Mining via LLM-Driven Code-Based Evolution gpt-oss-120b & gpt-oss-20b Model Card

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-03T20:41:29.316690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T20:41:29.316690Z digest=sha256:db47b1a9163261bf143a933237f2e5573a21f3be550a036fa0de901e467897f5

Observation 2401e3d0-288c-4806-bc9e-f2b11ae91ab6 · inbound

Matrix: Peer-to-Peer Multi-Agent Synthetic Data Generation Framework cites this paper.

Matrix: Peer-to-Peer Multi-Agent Synthetic Data Generation Framework gpt-oss-120b & gpt-oss-20b Model Card

Reference 11

Resolution
metadata mismatch
local_arxiv, observed 2026-05-17T04:24:00.357151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-17T04:23:03.393566Z digest=sha256:751f082aad619ced58d179bd0825f2f52d1201069f43d865cbec39cb0106be0f

Observation 274502f0-eda6-4d3d-b7ee-060e0101208c · inbound

JBE-QA: Japanese Bar Exam QA Dataset for Assessing Legal Domain Knowledge cites this paper.

JBE-QA: Japanese Bar Exam QA Dataset for Assessing Legal Domain Knowledge gpt-oss-120b & gpt-oss-20b Model Card

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-03T19:43:18.394750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T19:43:18.394750Z digest=sha256:c48a01be9d49082dfdceb089ca291b8530c8b8cdfe3c475448f99e928c69055b

Observation e01d0407-97d4-4afd-af6a-1b4c24b0d9ff · inbound

SQuARE: Structured Query & Adaptive Retrieval Engine For Tabular Formats cites this paper.

SQuARE: Structured Query & Adaptive Retrieval Engine For Tabular Formats gpt-oss-120b & gpt-oss-20b Model Card

Reference 26

Resolution
metadata mismatch
local_arxiv, observed 2026-05-17T01:43:49.656718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-17T01:43:07.780290Z digest=sha256:8648840c4fa567c125534440782ebb50ef60678ad5564cb439cd16c0e6c4b72b

Observation 74b12f57-5f81-46fd-9fab-063c9d9cd23c · inbound

ReasonBENCH: Benchmarking the (In)Stability of LLM Reasoning cites this paper.

ReasonBENCH: Benchmarking the (In)Stability of LLM Reasoning gpt-oss-120b & gpt-oss-20b Model Card

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T17:53:53.131588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T17:53:53.131588Z digest=sha256:e509ae49a96db7f1ad28cec02a141ffd50ce8b0d4bb04c791520e40261dafe0c

Observation 10b45ed2-0b4f-4e54-8e93-e826ab1495af · inbound

Intern-S1-MO: Long-horizon Reasoning Agent for Olympiad?Level Mathematical Problem Solving cites this paper.

Intern-S1-MO: Long-horizon Reasoning Agent for Olympiad?Level Mathematical Problem Solving gpt-oss-120b & gpt-oss-20b Model Card

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T06:40:25.428684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:40:25.428684Z digest=sha256:f8700cd0b383d82dad1552913cd1e7711739747d968aaf9765910e1c2d7d0f2b

Observation 44d5fb08-efb6-483f-bba8-78d8b27477a1 · inbound

Asynchronous Reasoning: Training-Free Interactive Thinking LLMs cites this paper.

Asynchronous Reasoning: Training-Free Interactive Thinking LLMs gpt-oss-120b & gpt-oss-20b Model Card

Reference 13

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T22:58:39.106221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T22:54:39.663737Z digest=sha256:83559f0191d2b36f58141c980f99075aa443154a6c4e71cc498d7ea590fa307f

Observation 3c5eb2ea-bd90-4ab6-a1e5-7e437bb3ac51 · inbound

Understanding Structured Financial Data with LLMs: A Case Study on Fraud Detection cites this paper.

Understanding Structured Financial Data with LLMs: A Case Study on Fraud Detection gpt-oss-120b & gpt-oss-20b Model Card

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-16T22:01:17.958081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-16T21:59:13.588901Z digest=sha256:e76d45cd0365ed7a3633ce8a2130e92176d3e8f8a0679467fc36388b4d6ec3b1

Observation 404ca04b-965a-412f-967a-2ed96f12b70c · inbound

DTop-p MoE: Sparsity-Controlled Dynamic Top-p MoE for Foundation Model Pre-training cites this paper.

DTop-p MoE: Sparsity-Controlled Dynamic Top-p MoE for Foundation Model Pre-training gpt-oss-120b & gpt-oss-20b Model Card

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-03T16:21:32.462468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:21:32.462468Z digest=sha256:a29d1eee3ca7be1f074112f71828c3054064d721dea48de35c712846b7177714

Observation 41a47057-3d46-43e0-b4b9-53bb251861a3 · inbound

Evaluating Large Language Models in Scientific Discovery cites this paper.

Evaluating Large Language Models in Scientific Discovery gpt-oss-120b & gpt-oss-20b Model Card

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-05-16T21:48:34.409689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-16T21:47:09.588941Z digest=sha256:6faefe7e78f3eb4de829e5de80344d8171de418817b780df8a46842328342604

Observation e2d1a4b0-fadc-4c28-84d3-e2a686b50388 · inbound

HEPTAPOD: Orchestrating High Energy Physics Workflows Towards Autonomous Agency cites this paper.

HEPTAPOD: Orchestrating High Energy Physics Workflows Towards Autonomous Agency gpt-oss-120b & gpt-oss-20b Model Card

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-03T15:47:05.436035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:47:05.436035Z digest=sha256:da2da2b0bf58305e6e87d5dbed783db3e867294dc04ac125c33a93799b5dc146

Observation e72a85d7-ad64-4630-b37d-d613a91d3e7b · inbound

Scalable Agentic Reasoning for Designing Biologics Targeting Intrinsically Disordered Proteins cites this paper.

Scalable Agentic Reasoning for Designing Biologics Targeting Intrinsically Disordered Proteins gpt-oss-120b & gpt-oss-20b Model Card

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-16T22:08:36.132537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T22:05:59.007680Z digest=sha256:f725d7c2a46778b993d5dbe675c585796c2cf03a0b8cbb625cf5d8afffb47d18

Observation f382e585-8e08-4abd-bf03-1a3fd8befe84 · inbound

Are vision-language models ready to zero-shot replace supervised classification models in agriculture? cites this paper.

Are vision-language models ready to zero-shot replace supervised classification models in agriculture? gpt-oss-120b & gpt-oss-20b Model Card

Reference 19

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T21:18:32.234667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-16T21:15:20.717705Z digest=sha256:461cbf92353e8e9e4d80d5f15693fe919640d6aee57cbe53dcc9e8397b1980d4

Observation a3ff257f-04bc-4c2b-8c52-5e107b015948 · inbound

SWE-EVO: Benchmarking Coding Agents in Long-Horizon Software Evolution Scenarios cites this paper.

SWE-EVO: Benchmarking Coding Agents in Long-Horizon Software Evolution Scenarios gpt-oss-120b & gpt-oss-20b Model Card

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-16T20:28:24.437343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-16T20:24:40.939455Z digest=sha256:633d18ca638652fb19bbbd7c1d00aed2daa0f280fabe8d8c6b88774791afaaa3

Observation 910675b9-ee03-4ffc-b0eb-63d1e4e52678 · inbound

A Benchmark for Evaluating Outcome-Driven Constraint Violations in Autonomous AI Agents cites this paper.

A Benchmark for Evaluating Outcome-Driven Constraint Violations in Autonomous AI Agents gpt-oss-120b & gpt-oss-20b Model Card

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T20:08:22.525377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-16T20:07:50.879922Z digest=sha256:78460f86668b69bc0383c0f7ff7f2f3776ba9550e807be8325e53c012cfc30b7

Observation 0cb024d5-2723-4f7a-8e70-c6ed47e5468b · inbound

NVIDIA Nemotron 3: Efficient and Open Intelligence cites this paper.

NVIDIA Nemotron 3: Efficient and Open Intelligence gpt-oss-120b & gpt-oss-20b Model Card

Reference 155

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T01:40:42.691282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-18T01:40:42.190369Z digest=sha256:06d12eaeeb8b5e387dc0409fd341f6bd63d704e3eb0a600d17594e9cfd639b16

Observation bc3df305-2e97-4a17-b471-bb831e26176b · inbound

Routing by Analogy: kNN-Augmented Expert Assignment for Mixture-of-Experts cites this paper.

Routing by Analogy: kNN-Augmented Expert Assignment for Mixture-of-Experts gpt-oss-120b & gpt-oss-20b Model Card

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T12:41:25.634853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:41:25.634853Z digest=sha256:a1b0045f133cf5041f8aeedbe82edcef69992c9ffa8c97a49978da5c19860574

Observation e8d9d4e5-94c7-400b-82be-f625e1397261 · inbound

MiMo-V2-Flash Technical Report cites this paper.

MiMo-V2-Flash Technical Report gpt-oss-120b & gpt-oss-20b Model Card

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-12T11:33:32.824634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T11:33:32.568261Z digest=sha256:4d286a75e378a6ea1221101113d1b61eb5652989e8bbc83e941446234dace676

Observation 31798b36-f87c-4d0b-bd13-5d83c2e2440c · inbound

Benchmarking and Adapting On-Device LLMs for Clinical Decision Support cites this paper.

Benchmarking and Adapting On-Device LLMs for Clinical Decision Support gpt-oss-120b & gpt-oss-20b Model Card

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-05-16T20:58:31.867890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T20:56:06.673393Z digest=sha256:e838467fecc56907a2e031f2fcd7a206e091153c8d465810fe098884c2bbc74b

Observation 05c707da-b091-4e90-b28b-7417b27327c7 · inbound

XGrammar-2: Efficient Dynamic Structured Generation Engine for Agentic LLMs cites this paper.

XGrammar-2: Efficient Dynamic Structured Generation Engine for Agentic LLMs gpt-oss-120b & gpt-oss-20b Model Card

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-03T12:07:33.247180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:07:33.247180Z digest=sha256:109958cf75a5c4a9adc6fb8b7d090727a72d4a0750db16528a3924b5a404eadc

Observation 8166941e-c6a4-41d0-b1b2-c8dfc7677af3 · inbound

Identifying and Mitigating Bottlenecks in Role-Playing Agents: A Systematic Study of Disentangling Character Profile Axes cites this paper.

Identifying and Mitigating Bottlenecks in Role-Playing Agents: A Systematic Study of Disentangling Character Profile Axes gpt-oss-120b & gpt-oss-20b Model Card

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-03T12:01:01.303018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:01:01.303018Z digest=sha256:74a53eed95197b2b1f4e17f3925be01bc84a6b667e0260174cef5567434926a0

Observation 3d65de2c-6342-45d2-bfda-cfded5946916 · inbound

Expos\'ia: Teaching and Assessment of Academic Writing Skills for Research Project Proposals and Peer Feedback cites this paper.

Expos\'ia: Teaching and Assessment of Academic Writing Skills for Research Project Proposals and Peer Feedback gpt-oss-120b & gpt-oss-20b Model Card

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-16T15:28:02.357331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T15:24:54.578162Z digest=sha256:d7cc1c0657806c09204dffee81cc33a4f2ddfff680458cbb0fab35938caef48f

Observation c6d01799-1fae-41c6-9c25-542904285db9 · inbound

Mid-Think: Training-Free Intermediate-Budget Reasoning via Token-Level Triggers cites this paper.

Mid-Think: Training-Free Intermediate-Budget Reasoning via Token-Level Triggers gpt-oss-120b & gpt-oss-20b Model Card

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T11:16:00.349180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T11:16:00.349180Z digest=sha256:0be4d393ded676eafa32b74d9d6a12fcc4e7e2d30ecf34cc8b7442fc0033969f

Observation eb6f3f7b-5a55-4690-b7d4-b9e9984164c6 · inbound

What Are We Measuring in NLG? A Meta-Analysis of Evaluation Trends 2020-2025 cites this paper.

What Are We Measuring in NLG? A Meta-Analysis of Evaluation Trends 2020-2025 gpt-oss-120b & gpt-oss-20b Model Card

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-03T11:07:10.182950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T11:07:10.182950Z digest=sha256:9b7142101a2e69038be7936aac02ab98f3722bd81f3b0d2e1d55881802be4f57

Observation 285502db-6b01-47d0-896b-0dd4264bc5fb · inbound

MirrorBench: A Benchmark to Evaluate Conversational User-Proxy Agents for Human-Likeness cites this paper.

MirrorBench: A Benchmark to Evaluate Conversational User-Proxy Agents for Human-Likeness gpt-oss-120b & gpt-oss-20b Model Card

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-05-21T15:15:16.661540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T15:14:55.551856Z digest=sha256:7f92d0bc45f4fc9997882a34cd097518e84b599750bf7d20352a72453bedda0a

Observation b485e30e-fadf-47a8-bb07-efe7236bdf3b · inbound

Terminal-Bench: Benchmarking Agents on Hard, Realistic Tasks in Command Line Interfaces cites this paper.

Terminal-Bench: Benchmarking Agents on Hard, Realistic Tasks in Command Line Interfaces gpt-oss-120b & gpt-oss-20b Model Card

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-11T03:37:08.005699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-11T03:37:07.841385Z digest=sha256:d2d0de83e09be0b3a2cbcca567bb8ac3d76836489a239e2f4db9448d1746d561

Observation 33cf6e2d-c044-43b7-931c-972c292dbcc9 · inbound

Which Reasoning Trajectories Teach Students to Reason Better? A Simple Metric of Informative Alignment cites this paper.

Which Reasoning Trajectories Teach Students to Reason Better? A Simple Metric of Informative Alignment gpt-oss-120b & gpt-oss-20b Model Card

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-16T12:27:52.461381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T12:23:56.318846Z digest=sha256:3aab0fd0f11e2fc49263879c1ba9032e29bfd913aebf4f618518798a52259baf

Observation d1ed45d9-2cff-446f-a691-7ef82ceec30f · inbound

Which Reasoning Trajectories Teach Students to Reason Better? A Simple Metric of Informative Alignment cites this paper.

Which Reasoning Trajectories Teach Students to Reason Better? A Simple Metric of Informative Alignment gpt-oss-120b & gpt-oss-20b Model Card

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T09:22:42.698573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T09:22:42.698573Z digest=sha256:6198ccc062e774b0219b7525cfafc81ee048315260ca3b215aaf79093e353fdb

Observation 1cee1e57-9458-49ff-933c-feeabfb38e0a · inbound

Compounding Disadvantage: Auditing Intersectional Bias in LLM-Generated Explanations Across Indian and American STEM Education cites this paper.

Compounding Disadvantage: Auditing Intersectional Bias in LLM-Generated Explanations Across Indian and American STEM Education gpt-oss-120b & gpt-oss-20b Model Card

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-21T16:35:22.139534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T16:34:55.552337Z digest=sha256:d9588f2921d194fe3fa92a453763afc7f180a0f46437222d4058540cdc79f27c

Observation c024c8bd-e855-4ebd-9c4d-9a05076e437b · inbound

YuFeng-XGuard: A Reasoning-Centric, Interpretable, and Flexible Guardrail Model for Large Language Models cites this paper.

YuFeng-XGuard: A Reasoning-Centric, Interpretable, and Flexible Guardrail Model for Large Language Models gpt-oss-120b & gpt-oss-20b Model Card

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T08:52:34.570722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T08:52:34.570722Z digest=sha256:603e6e7678c386a8d1079dfd326fd0d9051b184005503da0683ad6df900cd8bb

Observation f997b3b6-789a-492f-8ab9-7d45a79935ad · inbound

Learning to Discover at Test Time cites this paper.

Learning to Discover at Test Time gpt-oss-120b & gpt-oss-20b Model Card

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-16T05:16:04.055036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T05:16:04.001700Z digest=sha256:e432d98fda0f9fb3fb025a288e16dcbb443e5ae532ae19ba0850a5141662d4f2

Observation 3509d321-3921-4250-a894-d876c20ae1cb · inbound

Health-ORSC-Bench: A Benchmark for Measuring Over-Refusal and Safety Completion in Health Context cites this paper.

Health-ORSC-Bench: A Benchmark for Measuring Over-Refusal and Safety Completion in Health Context gpt-oss-120b & gpt-oss-20b Model Card

Reference 3059

Resolution
unresolved
no resolver link, observed 2026-08-03T08:21:12.844888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T08:21:12.844888Z digest=sha256:9f4eb7c0216ec8b73f957f2fe491e6b2ac16c0d3679d9c1c89922da48bc4eabd

Observation f89c468f-d796-4a7f-a638-d54da5050815 · inbound

BEAR: Budgeted Evidence Allocation for Multi-Document Reasoning cites this paper.

BEAR: Budgeted Evidence Allocation for Multi-Document Reasoning gpt-oss-120b & gpt-oss-20b Model Card

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-03T08:11:21.196459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T08:11:21.196459Z digest=sha256:089db18c1212392079fd102ded133691d8770612f3396ccd1566db1e400069c4

Observation 4d8775d2-1bf3-4c82-9a3c-5219af357bd1 · inbound

Think When Needed: Model-Aware Reasoning Routing for LLM-based Ranking cites this paper.

Think When Needed: Model-Aware Reasoning Routing for LLM-based Ranking gpt-oss-120b & gpt-oss-20b Model Card

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T08:09:51.036981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T08:09:51.036981Z digest=sha256:6d421eade9b656eeaf66497bbc791fc7519dcab89ea63c5d6744f02af9a95972

Observation 593baf05-2919-4300-8225-f895bd409fad · inbound

Temporal Leakage in Search-Engine Date-Filtered Web Retrieval: A Retrospective Forecasting Case Study cites this paper.

Temporal Leakage in Search-Engine Date-Filtered Web Retrieval: A Retrospective Forecasting Case Study gpt-oss-120b & gpt-oss-20b Model Card

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T08:57:38.729304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T08:53:17.570358Z digest=sha256:a0e025a39d107b7e5bd6e2f850f7853ced0a9024bcd502ea7283aac8180dac0e

Observation 57d385a3-98f3-48ab-9599-e9f0c347ee39 · inbound

Entropy-Preserving Supervised Fine-Tuning via Adaptive Self-Distillation for Large Reasoning Models cites this paper.

Entropy-Preserving Supervised Fine-Tuning via Adaptive Self-Distillation for Large Reasoning Models gpt-oss-120b & gpt-oss-20b Model Card

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-03T05:28:14.125632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:28:14.125632Z digest=sha256:78bc9e42b13903a9b0ccdc354b4fb1b6865314d05d57f7e7a1562e2e89215eeb

Observation 1a45c0f4-7875-4ed4-8853-ee4539f65de5 · inbound

CPMobius: Iterative Coach-Player Reasoning for Data-Free Reinforcement Learning cites this paper.

CPMobius: Iterative Coach-Player Reasoning for Data-Free Reinforcement Learning gpt-oss-120b & gpt-oss-20b Model Card

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T05:14:17.225076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:14:17.225076Z digest=sha256:eb3a8177ac6d39f1f992bddc38ed7b2bf043c0e740f1d9dc575523fc77247e1c

Observation ee1c576d-2298-4e7f-954a-413b5079c9b1 · inbound

Scalable Generation and Validation of Isomorphic Physics Problems with GenAI cites this paper.

Scalable Generation and Validation of Isomorphic Physics Problems with GenAI gpt-oss-120b & gpt-oss-20b Model Card

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-05-21T13:24:11.045326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T13:23:09.692634Z digest=sha256:015b031ad892423efb586689035c3fa6d5b9e4e52f4071c3053a7b3c9b4d2c6a

Observation 85a10ad0-493f-4700-91ed-7e7dd255d1c3 · inbound

OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions cites this paper.

OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions gpt-oss-120b & gpt-oss-20b Model Card

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T04:08:10.909307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T04:08:10.909307Z digest=sha256:3bcb07f10bf0cf9efe00584a26cbcddab4fcff5e9f260294ddc867e9775dc8cd

Observation 54ed742f-b526-48f5-92eb-aa9f00331f13 · inbound

Are We Truly Innovating? A Qualitative and Quantitative Study of Originality in AI Research Papers cites this paper.

Are We Truly Innovating? A Qualitative and Quantitative Study of Originality in AI Research Papers gpt-oss-120b & gpt-oss-20b Model Card

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T10:36:37.502114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:36:37.502114Z digest=sha256:5757b5055ded8f5c286a1af4ae459defa922c8f5547558f4eebca786ff869c98

Observation 74fb90b7-5256-48f6-a31c-8d497b383bb4 · inbound

MoSE: Mixture of Slimmable Experts for Efficient and Adaptive Language Models cites this paper.

MoSE: Mixture of Slimmable Experts for Efficient and Adaptive Language Models gpt-oss-120b & gpt-oss-20b Model Card

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T04:04:32.221609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T04:04:32.221609Z digest=sha256:008cd693c8553fe92e849432429befa0e2e8441f76eeae9be11df0104ec17474

Observation 6101e9f2-5027-40b0-8cb7-9597e42431c4 · inbound

Fin-RATE: A Real-world Financial Analytics and Tracking Evaluation Benchmark for LLMs on SEC Filings cites this paper.

Fin-RATE: A Real-world Financial Analytics and Tracking Evaluation Benchmark for LLMs on SEC Filings gpt-oss-120b & gpt-oss-20b Model Card

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T03:44:36.311665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:44:36.311665Z digest=sha256:7a3afe54888aad46b72e42172292f3f2d2bd0632a2ef0f32ec59974b46352329

Observation cbd7f6bf-b4d9-479c-8f61-5c32db84f736 · inbound

A Behavioural and Representational Evaluation of Goal-Directedness in Language Model Agents cites this paper.

A Behavioural and Representational Evaluation of Goal-Directedness in Language Model Agents gpt-oss-120b & gpt-oss-20b Model Card

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-03T03:09:50.325156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:09:50.325156Z digest=sha256:8f96679b47dac82d174d7e0ec65fb68de8a4e0484d0842ee19ec404cdbb5db0c

Observation c2b82ba5-0bee-400f-b70b-5aecbbb32864 · inbound

When LLMs get significantly worse: A statistical approach to detect model degradations cites this paper.

When LLMs get significantly worse: A statistical approach to detect model degradations gpt-oss-120b & gpt-oss-20b Model Card

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-16T06:07:25.757453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T06:03:37.788377Z digest=sha256:62515ce5c8424bd6e3bb0805bdb9147860930967a5fbfe69ec04b34c99988073

Observation 0b123c4e-99bc-430e-b0b7-abbf19c191ea · inbound

Disentangling Ambiguity from Instability in Large Language Models: A Clinical Text-to-SQL Case Study cites this paper.

Disentangling Ambiguity from Instability in Large Language Models: A Clinical Text-to-SQL Case Study gpt-oss-120b & gpt-oss-20b Model Card

Reference 12

Resolution
metadata mismatch
local_arxiv, observed 2026-05-21T12:50:09.194329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T12:49:55.133251Z digest=sha256:897d15ba5cbed50bd3a1ef4cd89b8e4edad9de0d0ddaa38993f81dca5c8ab813

Observation 4bba0904-5aa0-427c-9c93-ed719df0e3ab · inbound

Synthetic Interaction Data for Scalable Personalization in Large Language Models cites this paper.

Synthetic Interaction Data for Scalable Personalization in Large Language Models gpt-oss-120b & gpt-oss-20b Model Card

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T23:53:03.813879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:53:03.813879Z digest=sha256:611fcf21807b075014239e91934ebfb89de282ff3a5e7282f46c44db9596bc69

Observation 02beb234-b5d3-47f7-bdcd-86b5ceee21f2 · inbound

ProbeLLM: Automating Principled Diagnosis of LLM Failures cites this paper.

ProbeLLM: Automating Principled Diagnosis of LLM Failures gpt-oss-120b & gpt-oss-20b Model Card

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T23:43:07.177211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:43:07.177211Z digest=sha256:590df5ff82032145ab535c9840cf63651ef5cceaa6499d2ded1533ec69d35bc0

Observation 895650db-9bd3-404d-ade7-ff3fd1401145 · inbound

Speculative Decoding with a Speculative Vocabulary cites this paper.

Speculative Decoding with a Speculative Vocabulary gpt-oss-120b & gpt-oss-20b Model Card

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T23:27:54.011913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T23:27:54.011913Z digest=sha256:0832c721a7a6fcdc933b9c65acc67f8c526b0ffcc321d2c805076c7b27bada45

Observation e3809c03-b93c-4202-8aa8-790d00e32850 · inbound

A Critical Look at Targeted Instruction Selection: Disentangling What Matters (and What Doesn't) cites this paper.

A Critical Look at Targeted Instruction Selection: Disentangling What Matters (and What Doesn't) gpt-oss-120b & gpt-oss-20b Model Card

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T23:10:57.851998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T23:10:57.851998Z digest=sha256:dbf570bcf12f5335ee94dd69bc9ec5d2fd95b7416500a961b3d56b4b19e6b1dc

Observation f91f54df-2646-4873-9417-a52092249860 · inbound

Prescriptive Scaling Reveals the Evolution of Language Model Capabilities cites this paper.

Prescriptive Scaling Reveals the Evolution of Language Model Capabilities gpt-oss-120b & gpt-oss-20b Model Card

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-02T22:58:12.730423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:58:12.730423Z digest=sha256:e0ae4c1809873c969ca05753065f641e3d1696a9d61bfbc5eb1804a32b0f059a

Observation c55a8b32-000f-4809-be20-085e7f73a315 · inbound

Revisiting Text Ranking in Deep Research cites this paper.

Revisiting Text Ranking in Deep Research gpt-oss-120b & gpt-oss-20b Model Card

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T21:06:36.940539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:06:36.940539Z digest=sha256:81c035b2c159eb46ad07a1195ef10558afd66da189c6db6418defe016f8cda21

Observation 5998d2c2-4a94-4a12-b8ed-f846976d59bf · inbound

veScale-FSDP: Flexible and High-Performance FSDP at Scale cites this paper.

veScale-FSDP: Flexible and High-Performance FSDP at Scale gpt-oss-120b & gpt-oss-20b Model Card

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-15T19:06:30.812760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T19:03:22.142671Z digest=sha256:e8d2fabcbc5b05e48194c6b76cec549fadf27ca86094a958c5fa9733e5ceb45d

Observation 24a2bd56-ba9a-4468-8cb4-c9fb26fc0043 · inbound

Tracking Capabilities for Safer Agents cites this paper.

Tracking Capabilities for Safer Agents gpt-oss-120b & gpt-oss-20b Model Card

Reference 56

Resolution
verified exact
local_arxiv, observed 2026-05-15T18:36:29.174811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T18:31:34.181629Z digest=sha256:5e6aa9555f8272ae79d32ea24c621d1bda54391e34a513c0e48091e58b9862ee

Observation 9fcd959b-7700-4930-bbac-d3d0d2a30200 · inbound

EvoESAP: Non-Uniform Expert Pruning for Sparse MoE cites this paper.

EvoESAP: Non-Uniform Expert Pruning for Sparse MoE gpt-oss-120b & gpt-oss-20b Model Card

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-05-15T14:35:55.668973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T14:34:48.524592Z digest=sha256:53fa9844479e9e0aa102729901337416c5587451e467774aa5132a410a6306e2

Observation dff25e27-2406-4798-b715-3c64cd0f15d9 · inbound

Grouter: Decoupling Routing from Representation for Accelerated MoE Training cites this paper.

Grouter: Decoupling Routing from Representation for Accelerated MoE Training gpt-oss-120b & gpt-oss-20b Model Card

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T21:50:59.051691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:50:59.051691Z digest=sha256:e53399b9b050cc0c2bf2c8b0f15727188184c8068d0443e49c5370411534d343

Observation a5e57c83-c576-4e43-a33d-8a2825765cce · inbound

PACED: Distillation and On-Policy Self-Distillation at the Frontier of Student Competence cites this paper.

PACED: Distillation and On-Policy Self-Distillation at the Frontier of Student Competence gpt-oss-120b & gpt-oss-20b Model Card

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T13:15:50.698145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T13:11:29.528871Z digest=sha256:d30211a390c58ff18c9e873f704d7c3c28d1ec117a472d293be50edac94f1e79

Observation d94131de-3143-4c36-9f8c-442e900770c8 · inbound

Prompt Programming for Cultural Bias and Alignment of Large Language Models cites this paper.

Prompt Programming for Cultural Bias and Alignment of Large Language Models gpt-oss-120b & gpt-oss-20b Model Card

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T18:01:41.249627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:01:41.249627Z digest=sha256:91bbf7966e2a6951f1eb9483ca177317bcb3c95784501522af5c58af9d848ea3

Observation aa9d3a60-8dda-49b5-a855-cff1a1cc649d · inbound

ALL-FEM: Agentic Large Language models Fine-tuned for Finite Element Methods cites this paper.

ALL-FEM: Agentic Large Language models Fine-tuned for Finite Element Methods gpt-oss-120b & gpt-oss-20b Model Card

Reference 51

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T15:51:05.251430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T15:49:36.623425Z digest=sha256:b22f81362685a6bf8439d91a3b6774c744c0d44f96573b6394271f2fadb11d0d

Observation 6f9b902a-57ee-46fd-98e2-2eaeeb50f5ea · inbound

ChartNet: A Million-Scale, High-Quality Multimodal Dataset for Robust Chart Understanding cites this paper.

ChartNet: A Million-Scale, High-Quality Multimodal Dataset for Robust Chart Understanding gpt-oss-120b & gpt-oss-20b Model Card

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-14T22:39:32.993267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T22:39:06.113655Z digest=sha256:53f33d9aa9bba595d5bc841347ff185a0033dc24342764bc65241dc9c73973bc

Observation 1ea9c5c6-d5c5-4238-b415-63c0d10b4d75 · inbound

Ensemble-Based Uncertainty Estimation for Code Correctness Estimation cites this paper.

Ensemble-Based Uncertainty Estimation for Code Correctness Estimation gpt-oss-120b & gpt-oss-20b Model Card

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-05-14T22:53:14.131482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T22:52:58.524934Z digest=sha256:bacf922dfce9f811122e052decc80809abd395333c76aeef934252d62ddcf214

Observation 7d7318b5-0e0e-41de-801e-4018484251b9 · inbound

Rethinking Language Model Scaling under Transferable Hypersphere Optimization cites this paper.

Rethinking Language Model Scaling under Transferable Hypersphere Optimization gpt-oss-120b & gpt-oss-20b Model Card

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-05-14T21:53:02.389596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T21:51:14.678941Z digest=sha256:5d0f180d26d09e485e88b8ebefdaf98f12b5febed384fab9635a1d685fd1a011