Pith. sign in

Paper Citation Record · LEDGER

DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

As of 27 July 2026, this Paper Citation Record lists 17 of 17 outbound references and 100 inbound Pith citation observations for arXiv:2512.02556.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2512.02556 v1

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-10T13:05:26.667750Z

measured 117 of 117 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-07-27T06:30:09.085275+00:00

measured 100 of 391 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-15T15:03:20.637041Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-11T01:57:51.613379Z

Reference resolution

17 of 17 outbound references displayed

  • verified exact11
  • verified fuzzy3
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6d5ec91b-27e0-4642-8aa5-1a9541b16c1d · outbound

This paper cites $\tau^2$-Bench: Evaluating Conversational Agents in a Dual-Control Environment.

DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models $\tau^2$-Bench: Evaluating Conversational Agents in a Dual-Control Environment

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:52:17.902869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T13:05:26.667750Z digest=sha256:dfb4d03ba501f5b9c072bed1b4ecff71b2aa360171bf6cd0dae512133edbcdac

Observation 7cd66631-132a-43f4-9123-1dfc682e8e99 · outbound

This paper cites DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model.

DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T05:36:27.707517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T13:05:26.667750Z digest=sha256:90f4716f248ec6f00cfda4fa9a51e4ab0bdd2ddf3e6758325c84dda08f009184

Observation bc4d3366-d61f-4ea8-918c-bcb3a4416d18 · outbound

This paper cites 10 Guoqing Ma, Jia Zhu, Hanghui Guo, Weijie Shi, Yue Cui, Jiawei Shen, Zilong Li, and Yidan Liang.

DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models 10 Guoqing Ma, Jia Zhu, Hanghui Guo, Weijie Shi, Yue Cui, Jiawei Shen, Zilong Li, and Yidan Liang

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:05:26.754096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T13:05:26.667750Z digest=sha256:25b6761d917c6b2f6f0576054834a3080c12dd243cd722cf4c84810520188ba1

Observation 69125550-6611-4428-b14e-e6d3a9f12807 · outbound

This paper cites MCP-Universe: Benchmarking Large Language Models with Real-World Model Context Protocol Servers.

DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models MCP-Universe: Benchmarking Large Language Models with Real-World Model Context Protocol Servers

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:05:26.757169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T13:05:26.667750Z digest=sha256:ad0aefb2e3d191e188270db50787f8459a38f64054622c590079e8257b2bb169

Observation 7c389788-4e3b-4a21-ae29-9db6fdc73da3 · outbound

This paper cites Luong, D.

DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models Luong, D

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T13:05:26.775691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T13:05:26.667750Z digest=sha256:6f5dd48b53d1283787db976969a104b8f1a32f905d4fbd5e83a9b47d83c1044d

Observation 7c9de2e6-7884-4322-bc83-985851b13582 · outbound

This paper cites 18 MiniMax.

DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models 18 MiniMax

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T13:05:26.773841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T13:05:26.667750Z digest=sha256:e518862d0b95a16347f62527fe4ab69e4b7b72c20ea5474e09daaf2f4892a42c

Observation ef2627d0-f6a3-4118-b287-6c71d3802a02 · outbound

This paper cites Humanity's Last Exam.

DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models Humanity's Last Exam

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-10T18:40:50.715344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T13:05:26.667750Z digest=sha256:7cce351b0410ea9851573157580673b60bc1c99ed63de5e43063a92dc61e22a7

Observation 10e021ad-e2d0-4eb6-aab4-1a2e5fd2d8a5 · outbound

This paper cites Qwen3 Technical Report.

DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models Qwen3 Technical Report

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-10T13:05:26.744711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T13:05:26.667750Z digest=sha256:1d394cb2052e31c2f1046d8b60d279ebddcc16fea6b55a2d24486dccc7492d5e

Observation ad2e1c8d-7782-4f5a-aa2a-f13657ae03b2 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-05-10T13:05:26.747838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T13:05:26.667750Z digest=sha256:1f73dab94f20be6f1bcb05e06b57fa6680057fe45209178535f537281cc5d4ac

Observation eecac730-9ca6-4fd3-9adf-75b6f3948c0a · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-05-10T13:05:26.687674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T13:05:26.667750Z digest=sha256:805e60bd96431b65354842de998d976d2a65ba88642b7fc901c05862d91e571e

Observation f65269c7-def7-4a3d-b93f-b1d38396155d · outbound

This paper cites Fast Transformer Decoding: One Write-Head is All You Need.

DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models Fast Transformer Decoding: One Write-Head is All You Need

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:50:20.836519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T13:05:26.667750Z digest=sha256:7aecb868640f8245b41540fb2e113d4d74ed024a094ce406abec741b3c8993fc

Observation ca3ea0aa-fbef-42cf-a952-6f954ac057f4 · outbound

This paper cites MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark.

DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:51:14.588885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T13:05:26.667750Z digest=sha256:f08e1522f463bf2dbc8815461acee5273b7f5ba759c79097ccea6720c117fdde

Observation d56cf372-e911-493c-b755-c56914d404c9 · outbound

This paper cites MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark.

DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:51:14.588885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T13:05:26.667750Z digest=sha256:0e981fd376b368ca9dcfc4aea048f99d595b7e4deda780bdca33fc48fb2fbeb9

Observation b36c2ba3-b3d7-423c-a8b7-2713c4f43f24 · outbound

This paper cites SWE-smith: Scaling Data for Software Engineering Agents.

DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models SWE-smith: Scaling Data for Software Engineering Agents

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-15T10:22:07.061654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T13:05:26.667750Z digest=sha256:674df9cef41f5fc7de80e25959c85b7c69a8e7ba4c649715ecc918c708c83503

Observation 6a6d8cd1-5eea-4c57-b795-cd742f2313b7 · outbound

This paper cites GLM-4.5: Agentic, Reasoning, and Coding (ARC) Foundation Models.

DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models GLM-4.5: Agentic, Reasoning, and Coding (ARC) Foundation Models

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:50:08.654972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T13:05:26.667750Z digest=sha256:34fefeb356ca6bd15fc20be94dc9aa5ef4ef7b50608588f3d400ea0f3f984264

Observation f004d965-9645-4905-8ee7-3c6b5b6c91dd · outbound

This paper cites BrowseComp-ZH: Benchmarking Web Browsing Ability of Large Language Models in Chinese.

DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models BrowseComp-ZH: Benchmarking Web Browsing Ability of Large Language Models in Chinese

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-17T22:04:50.009871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T13:05:26.667750Z digest=sha256:93b78d3beeb81d697cd9e944702ebaed1861411b1e7266c9154f896ba066e97f

Observation 11d831bc-c519-43c4-bab4-ec75bc44642b · outbound

This paper cites an unresolved cited work.

DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models Unresolved cited work

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T13:05:26.777420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T13:05:26.667750Z digest=sha256:dbab4a567795dbebd5c5bd5f3adce7d908902c1dc22ffae3340a381384554940

Pith citing papers

Observation 13b287bc-8737-47a1-818f-0ebd4ad20387 · inbound

Reinforcement Learning from Human Feedback cites this paper.

Reinforcement Learning from Human Feedback DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 185

Resolution
verified exact
local_arxiv, observed 2026-05-22T19:32:01.083881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-22T19:27:40.991325Z digest=sha256:b9dc3d6537ee2ef5b1d64936e67e2e856260ca7797d93b0371567a3eee3447df

Observation 7d7790e9-0003-43c0-a01e-646ee76b3410 · inbound

Evaluating Code Reasoning Abilities of Large Language Models Under Real-World Settings cites this paper.

Evaluating Code Reasoning Abilities of Large Language Models Under Real-World Settings DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-05-16T21:28:34.198311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-16T21:23:44.762007Z digest=sha256:46570b36c87994b79cc5ad3c23298f10f5f1f781f91d1e8d2e50ee6042a02d2c

Observation 79ff48ad-c1f0-4814-9c0c-5f361711fcba · inbound

Toward Training Superintelligent Software Agents through Self-Play SWE-RL cites this paper.

Toward Training Superintelligent Software Agents through Self-Play SWE-RL DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-05-21T16:10:20.183526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-21T16:07:48.570995Z digest=sha256:2ad8d5845a400998a344f13cdb78f45ff682fd897cfe5363026ffe26a2ef09d9

Observation c7408991-06f0-4ea2-9c13-6e477be21545 · inbound

MiMo-V2-Flash Technical Report cites this paper.

MiMo-V2-Flash Technical Report DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-05-12T11:33:32.843842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-12T11:33:32.568261Z digest=sha256:a99b6f28d9a62e2bd828663e3347da4bf7fa17cc6296f713811f43772b84fc3d

Observation bfb105ca-d9b9-4b53-8e48-547149e74a6b · inbound

DiffCoT: Diffusion-styled Chain-of-Thought Reasoning in LLMs cites this paper.

DiffCoT: Diffusion-styled Chain-of-Thought Reasoning in LLMs DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T17:21:07.905175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-16T17:19:58.334098Z digest=sha256:00302ec8624d6790ef18b50dd5f0091b0dff28eb668c85021108d7c06a3af1d4

Observation 6a391917-9599-4436-b4e9-948ba9e4c19c · inbound

GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization cites this paper.

GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-15T05:31:56.042488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-15T05:31:55.864438Z digest=sha256:6131181ce9bb141eca72c3ba8ad7f90881bf8f838acebe4aedefaa3aff27d2ab

Observation 2f5eb001-1b51-4ec0-97fd-7c938abff947 · inbound

OPT-Engine: Benchmarking the Limits of LLMs in Optimization Modeling via Complexity Scaling cites this paper.

OPT-Engine: Benchmarking the Limits of LLMs in Optimization Modeling via Complexity Scaling DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-05-16T16:18:05.375328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-16T16:17:43.055124Z digest=sha256:09e7cb04328e956a3dde82982f4ff6877720a02bd0267a975d2e85264634481c

Observation 011f576a-749d-40b2-982e-cb43dfb7e2d3 · inbound

ETS: Energy-Guided Test-Time Scaling for Training-Free RL Alignment cites this paper.

ETS: Energy-Guided Test-Time Scaling for Training-Free RL Alignment DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-16T09:40:49.039750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-16T09:37:57.120779Z digest=sha256:b09e90f002b56b42ef4d33e7545c716299e5b1f8c8f558db0eb4770e137b6040

Observation f0c8d3af-fff1-4936-9936-13fb9537a5c6 · inbound

ETS: Energy-Guided Test-Time Scaling for Training-Free RL Alignment cites this paper.

ETS: Energy-Guided Test-Time Scaling for Training-Free RL Alignment DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-21T14:10:13.433825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-21T14:05:37.120262Z digest=sha256:5b0beee3657f7ec46da76f7f948f2835e2a6efe481d9ffadff0b828cbb79e7e8

Observation 5952d2ac-0938-4ea3-ad65-63b65a5af5dc · inbound

Learning from Self-Debate: Preparing Reasoning Models for Multi-Agent Debate cites this paper.

Learning from Self-Debate: Preparing Reasoning Models for Multi-Agent Debate DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-21T14:30:13.844752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-21T14:29:15.751497Z digest=sha256:b64098671d96ef3b7bf3b57f926af7ad10e7626095a0ee0ad011f0ef756db4f4

Observation 4a8b1526-2862-4d0d-af63-1fc1bcf6c7c0 · inbound

Mixture-of-Top-k Attention: Efficient Attention via Scalable Fast Weights cites this paper.

Mixture-of-Top-k Attention: Efficient Attention via Scalable Fast Weights DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-05-16T08:52:37.657663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-16T08:51:25.375847Z digest=sha256:65b0933b7b74726694a354b8d166ef99c37f9e206c54122a967d782bc7ba472b

Observation 073f598e-5868-4c56-be21-9240dc020f49 · inbound

Reasoning in a Combinatorial and Constrained World: Benchmarking LLMs on Natural-Language Combinatorial Optimization cites this paper.

Reasoning in a Combinatorial and Constrained World: Benchmarking LLMs on Natural-Language Combinatorial Optimization DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-16T08:20:46.036461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-16T08:18:52.540487Z digest=sha256:0aa2e619a18ef53e738d499d68f0633fbc00c25d2ecaa0788819282220713f89

Observation 42a2c601-c2cb-4c01-bebb-0b925f9a6db0 · inbound

Kimi K2.5: Visual Agentic Intelligence cites this paper.

Kimi K2.5: Visual Agentic Intelligence DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-05-10T16:09:05.290142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T16:09:05.225767Z digest=sha256:ca22dabfce099b57f98cdd77094bfccac7a470fffdd41e64745ddf6381067662

Observation 89f5751e-3973-4d74-b41a-4bec561847b9 · inbound

AgenticSZZ: Temporal Knowledge Graph-Guided Agentic Bug-Inducing Commit Identification cites this paper.

AgenticSZZ: Temporal Knowledge Graph-Guided Agentic Bug-Inducing Commit Identification DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T08:27:36.401445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-16T08:25:59.552687Z digest=sha256:0ff0a0045259c9c117715e6895dd020be10e66c084166bf96bb3f5ef0180c98d

Observation 9b525396-13a6-46b7-abb7-576053392d43 · inbound

Rethinking the Value of Agent-Generated Tests for LLM-Based Software Engineering Agents cites this paper.

Rethinking the Value of Agent-Generated Tests for LLM-Based Software Engineering Agents DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-05-16T06:37:28.480502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-16T06:36:06.214753Z digest=sha256:8c86bf411fa7ece59df1fc69445c435a218b38431e83343fc96dbdad1f15cdbe

Observation 67ec9ade-b324-49f3-9711-dccc8bf99866 · inbound

AceGRPO: Adaptive Curriculum Enhanced Group Relative Policy Optimization for Autonomous Machine Learning Engineering cites this paper.

AceGRPO: Adaptive Curriculum Enhanced Group Relative Policy Optimization for Autonomous Machine Learning Engineering DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-05-16T06:32:27.293930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-16T06:32:22.038300Z digest=sha256:c576ea7762354ed06d8bae672bcb57362e797130c9e644a654081ec46f97dd00

Observation ff04e2b7-85ac-4e7e-834c-bf1ce329ac26 · inbound

SnapMLA: Efficient Long-Context MLA Decoding via Hardware-Aware FP8 Quantized Pipelining cites this paper.

SnapMLA: Efficient Long-Context MLA Decoding via Hardware-Aware FP8 Quantized Pipelining DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-16T06:00:40.740867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-16T05:58:03.113220Z digest=sha256:140bd1fc0fb5a45614f14a45f13984fad58e03df0ae3676b1e64bb1466d23fec

Observation 59a735ec-5cb8-489b-aa64-5f5a60f9c4be · inbound

GLM-5: from Vibe Coding to Agentic Engineering cites this paper.

GLM-5: from Vibe Coding to Agentic Engineering DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-05-11T05:46:41.022884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-11T05:46:40.836161Z digest=sha256:21bd269e0a87e781434723312ee6e6122830e925243f86ceb7d595a192e67d8c

Observation da751b51-9317-4fad-891e-a396799cc1ea · inbound

RAT+: Train Dense, Infer Sparse -- Recurrence Augmented Attention for Dilated Inference cites this paper.

RAT+: Train Dense, Infer Sparse -- Recurrence Augmented Attention for Dilated Inference DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-05-15T21:00:17.865951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-15T20:59:33.902420Z digest=sha256:6b80c405af5b55a865e0537801164a31cf63655dbae0c95fb86c49bc105b7d57

Observation ca58d92f-44ce-4f5f-87fb-06321a515d2b · inbound

RAT+: Train Dense, Infer Sparse -- Recurrence Augmented Attention for Dilated Inference cites this paper.

RAT+: Train Dense, Infer Sparse -- Recurrence Augmented Attention for Dilated Inference DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-05-21T12:50:09.468445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-21T12:45:27.150368Z digest=sha256:b5eeb90179c28c216692e157372b4b3ef9e6562ceacf53b0d03f1240d32e2e07

Observation 6d2d7069-c0ee-4dfd-afac-e5b490f72895 · inbound

Tracking Capabilities for Safer Agents cites this paper.

Tracking Capabilities for Safer Agents DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-05-15T18:36:29.234504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-15T18:31:34.181629Z digest=sha256:9432105cf2ee791a4f25af9310be4962f08dcbdf2c879742efa67f6435793ba4

Observation 00739758-af1d-4e03-b837-36d3eb117991 · inbound

GroupGPT: A Token-efficient and Privacy-preserving Agentic Framework for Multi-User Chat Assistant cites this paper.

GroupGPT: A Token-efficient and Privacy-preserving Agentic Framework for Multi-User Chat Assistant DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-05-15T18:30:14.465308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-15T18:27:16.527361Z digest=sha256:286b9233e1bcedc526dae3869d7f6cc855d65d4c3accdb495ce184a12e131b43

Observation 3761befd-14f4-42c0-9164-9be2da645156 · inbound

Multistage Stochastic Programming for Rare Event Risk Mitigation in Power Systems Management cites this paper.

Multistage Stochastic Programming for Rare Event Risk Mitigation in Power Systems Management DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-15T15:03:20.637041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T15:03:20.637041Z digest=sha256:4f82af0d4d874a9688cdc50fa4d9cd90d1de68e312c653ba5b61f452180c84e0

Observation 20f79b34-1c8c-49f1-bb20-4ef06c279669 · inbound

MICA: Multi-granularity Intertemporal Credit Assignment for Long-Horizon Emotional Support Dialogue cites this paper.

MICA: Multi-granularity Intertemporal Credit Assignment for Long-Horizon Emotional Support Dialogue DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 64

Resolution
verified exact
local_arxiv, observed 2026-05-15T15:26:10.736564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-15T15:25:51.138776Z digest=sha256:4bd73f376d4dcac64f24e7d01bcca739c012bd974932c493fc53826a9bd0bfc4

Observation 1cf6b1c9-aeb6-4e22-b3db-3d77137fbe44 · inbound

Story Point Estimation Using Large Language Models cites this paper.

Story Point Estimation Using Large Language Models DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-05-15T15:30:07.602476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-15T15:27:47.957035Z digest=sha256:2b71e24dd64b8a5b9f773df53b98856a5c27250b0220122865ac3f8182118269

Observation 48906e6d-a64f-428f-86dc-fc179467c923 · inbound

CODA: Difficulty-Aware Compute Allocation for Adaptive Reasoning cites this paper.

CODA: Difficulty-Aware Compute Allocation for Adaptive Reasoning DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-05-15T14:25:55.449930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-15T14:23:00.793443Z digest=sha256:d03a261763d89d9f16f0ce5db9f4b9a9a85f5435c356781c01ca2926d438e26b

Observation 5b217549-1d76-4c52-b8e4-c50f12c4db75 · inbound

When Does Sparsity Mitigate the Curse of Depth in LLMs cites this paper.

When Does Sparsity Mitigate the Curse of Depth in LLMs DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-14T20:29:33.439034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T20:29:33.439034Z digest=sha256:81f240288da57856b2d0b24c9c734992a6517219281a3e9c56c48c292b9d0298

Observation a71671bd-d4c7-45df-a4c7-f855cfc63b63 · inbound

NeSy-Route: A Neuro-Symbolic Benchmark for Constrained Route Planning in Remote Sensing cites this paper.

NeSy-Route: A Neuro-Symbolic Benchmark for Constrained Route Planning in Remote Sensing DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-15T11:59:49.461107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T11:59:49.461107Z digest=sha256:3910c9d3e811adedafbfe2212e9e7cd44d66635787c8841a69b7a44c0cee77e7

Observation c287aeac-973c-4da6-84a7-2b3863938e9a · inbound

Towards Safer Large Reasoning Models by Promoting Safety Decision-Making before Chain-of-Thought Generation cites this paper.

Towards Safer Large Reasoning Models by Promoting Safety Decision-Making before Chain-of-Thought Generation DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-15T09:35:22.397588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-15T09:33:47.702580Z digest=sha256:2a91aa568adefbfa8dd9ad8971c20b287c8eb6d0f8cc9c5e17f126f8b61bd829

Observation 17e64d24-aec0-4341-9fcc-c75d0244fed8 · inbound

VIGIL: Part-Grounded Structured Reasoning for Generalizable Deepfake Detection cites this paper.

VIGIL: Part-Grounded Structured Reasoning for Generalizable Deepfake Detection DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-13T20:48:11.937716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T20:48:11.937716Z digest=sha256:029e34f8f22665c8903781ce3e710c7ecc5f4aa9314dcd3c5662daf3b3cd84eb

Observation 58b88682-0a97-47b4-97ff-7c46f8f6b546 · inbound

MSA: Memory Sparse Attention for Efficient End-to-End Memory Model Scaling to 100M Tokens cites this paper.

MSA: Memory Sparse Attention for Efficient End-to-End Memory Model Scaling to 100M Tokens DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-05-15T16:00:10.048061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-15T16:00:03.578685Z digest=sha256:5ae5edde3bfe0be791858920c65e3d7607479527ddcced5e1c25cdb3cc28829c

Observation e097af8d-6057-4214-a9a0-b7230172519b · inbound

From Pixels to Digital Agents: An Empirical Study on the Taxonomy and Technological Trends of Reinforcement Learning Environments cites this paper.

From Pixels to Digital Agents: An Empirical Study on the Taxonomy and Technological Trends of Reinforcement Learning Environments DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 214

Resolution
verified exact
local_arxiv, observed 2026-05-15T01:23:27.195363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-15T01:20:03.181903Z digest=sha256:9198e6ba0f1327f013616d1d9ecbb0b9133a5129b5daff5a729c1c79a8aa0fe0

Observation 26cdfeaf-d3bf-44c0-b8bc-c3793b1494ef · inbound

YingMusic-Singer: Controllable Singing Voice Synthesis with Flexible Lyric Manipulation and Annotation-free Melody Guidance cites this paper.

YingMusic-Singer: Controllable Singing Voice Synthesis with Flexible Lyric Manipulation and Annotation-free Melody Guidance DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-15T00:23:22.743224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-15T00:21:26.758816Z digest=sha256:941dcfd2b4548bb100e5c7c440df0540e069434bc6e9022d01b3bdfed67bd8c3

Observation 64c9240b-d80c-4d27-9436-ba83fbb8f295 · inbound

YingMusic-Singer: Controllable Singing Voice Synthesis with Flexible Lyric Manipulation and Annotation-free Melody Guidance cites this paper.

YingMusic-Singer: Controllable Singing Voice Synthesis with Flexible Lyric Manipulation and Annotation-free Melody Guidance DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-13T18:44:48.222670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T18:44:48.222670Z digest=sha256:26c6707d0145d0e184ab85be2dbd6d2b6996ae08b91b1816a233166761b18e08

Observation 0f7490f4-6736-4acc-a2c2-44d4bfbfa56f · inbound

AIRA_2: Overcoming Bottlenecks in AI Research Agents cites this paper.

AIRA_2: Overcoming Bottlenecks in AI Research Agents DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-14T23:08:14.542424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-14T23:06:39.545404Z digest=sha256:72402541921d1162afc7521835cfc9e8373b5ebd0989a5b143741287cb4239ad

Observation 0aec5961-d9ef-40a2-9eb1-738cccf3ee58 · inbound

AlpsBench: An LLM Personalization Benchmark for Real-Dialogue Memorization and Preference Alignment cites this paper.

AlpsBench: An LLM Personalization Benchmark for Real-Dialogue Memorization and Preference Alignment DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-15T14:51:08.369181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-15T14:50:23.587894Z digest=sha256:faa4cafda096dc4bf78d2cc0e5d957479ec1d8fe7817be3f21cc9905594989f7

Observation 5d857eb3-6c95-4746-9d07-b33b0d7086ad · inbound

Evaluating Privilege Usage of Agents with Real-World Tools cites this paper.

Evaluating Privilege Usage of Agents with Real-World Tools DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-05-14T22:18:04.333729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-14T22:16:56.848520Z digest=sha256:7a38ba3fd9719129df7489b258633710af8a88968c45a90c450517edff3dede8

Observation 6f4d5bf2-3a3c-4ec4-8269-48ec766ead44 · inbound

Kernel-Smith: A Unified Recipe for Evolutionary Kernel Optimization cites this paper.

Kernel-Smith: A Unified Recipe for Evolutionary Kernel Optimization DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-05-14T22:18:04.314611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-14T22:17:13.431164Z digest=sha256:24e0749a1e50179861a345b6a3bb1c109da7faf6d21e8e6ed76ff4125ffd219c

Observation acf269c7-6b1f-4e74-aa92-a78512e25f17 · inbound

Three non-Hermitian random matrix universality classes of complex edge statistics: Spacing ratios and distributions cites this paper.

Three non-Hermitian random matrix universality classes of complex edge statistics: Spacing ratios and distributions DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-13T16:19:49.419805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T16:19:49.419805Z digest=sha256:c948034b9df2634deb86a369bb39b0ed2e2fb45e602d78f9bd3f7e1ad3e9588b

Observation cc750233-6e61-455a-a1d5-31ab643d3690 · inbound

HISA: Efficient Hierarchical Indexing for Fine-Grained Sparse Attention cites this paper.

HISA: Efficient Hierarchical Indexing for Fine-Grained Sparse Attention DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-05-14T21:43:00.746243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-14T21:40:28.854082Z digest=sha256:5eddc9ac48cb16a7514d72009491ce4574932ac8cd71db2ee671246fb0b4b2b4

Observation a5d1997a-2c65-4ace-937a-2c0f6a38ec18 · inbound

Understand and Accelerate Memory Processing Pipeline for Large Language Model Inference cites this paper.

Understand and Accelerate Memory Processing Pipeline for Large Language Model Inference DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-14T00:28:29.873880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-14T00:25:49.807277Z digest=sha256:20d1184a821982ff98093e792b51ea6c1b4e9ec2ee5a68cc99578e5065251c56

Observation 4aa441cc-d63b-4d34-8347-e36346bf7010 · inbound

From SWE-ZERO to SWE-HERO: Execution-free to Execution-based Fine-tuning for Software Engineering Agents cites this paper.

From SWE-ZERO to SWE-HERO: Execution-free to Execution-based Fine-tuning for Software Engineering Agents DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-13T21:48:19.302532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=arxiv_source observed=2026-05-13T21:44:07.775936Z digest=sha256:bd4e7cbe0e0c8760c1d998df1672443ea0dee51aa5883439b360999c59704bfc

Observation f976a9e2-3fdb-4825-8400-d61596c45f75 · inbound

SkVM: Revisiting Language VM for Skills across Heterogenous LLMs and Harnesses cites this paper.

SkVM: Revisiting Language VM for Skills across Heterogenous LLMs and Harnesses DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-05-13T19:28:09.732726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-13T19:26:42.499046Z digest=sha256:0fd9346acdb25cdd0a84d440205a6a909d1ea0b1fda47542553002b30ccfe034

Observation d3a861aa-ca40-44dc-834b-8b864afd0680 · inbound

InCoder-32B-Thinking: Industrial Code World Model for Thinking cites this paper.

InCoder-32B-Thinking: Industrial Code World Model for Thinking DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-13T18:58:08.708184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-13T18:56:31.975626Z digest=sha256:c68755c441348641c87b9196575e07aa56eb40af8f03cff556ba9333ea879b54

Observation 8067f823-5ae2-497d-9dcb-5c394dfe1c93 · inbound

BAS: A Decision-Theoretic Approach to Evaluating Large Language Model Confidence cites this paper.

BAS: A Decision-Theoretic Approach to Evaluating Large Language Model Confidence DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-05-13T19:53:11.974126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=arxiv_source observed=2026-05-13T19:48:13.133733Z digest=sha256:c2c7a612e789ad70b812778b1f9db2620535a328e7715c37d22e6a9d0ee9e62d

Observation d9792ae3-b2c0-4694-8e75-26feb7da099a · inbound

Scaling Teams or Scaling Time? Memory Enabled Lifelong Learning in LLM Multi-Agent Systems cites this paper.

Scaling Teams or Scaling Time? Memory Enabled Lifelong Learning in LLM Multi-Agent Systems DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-14T22:43:12.002256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-14T22:42:43.070265Z digest=sha256:182191bbc0c6e4954640105b85174bb2c4bce0e17eaf9056d7f5c078173a92fd

Observation 6e1157fd-d02d-4033-b92e-d483c397010c · inbound

Hierarchical SVG Tokenization: Learning Compact Visual Programs for Scalable Vector Graphics Modeling cites this paper.

Hierarchical SVG Tokenization: Learning Compact Visual Programs for Scalable Vector Graphics Modeling DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-05-10T23:45:52.530761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T18:53:33.951884Z digest=sha256:ac9e81df219ea8c076a312f281cf6de99bf5d4a0613e27c58837bc36016c6d23

Observation 1334830d-fba5-4aeb-9448-b0a046cb54f3 · inbound

Watch Before You Answer: Learning from Visually Grounded Post-Training cites this paper.

Watch Before You Answer: Learning from Visually Grounded Post-Training DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-05-10T22:15:51.711356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T20:01:13.305374Z digest=sha256:a7b542681094d4fcf5eb06106f44b854684290e7961651ec70ec48d2c193033d

Observation ed383789-9c0e-4fad-9dae-960d73aa21ec · inbound

SpecRL: Reinforcement Learning with Test-Based Completeness Rewards for Formal Specification Synthesis cites this paper.

SpecRL: Reinforcement Learning with Test-Based Completeness Rewards for Formal Specification Synthesis DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 12

Resolution
malformed identifier
local_arxiv, observed 2026-05-10T23:30:50.817514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T19:06:10.718673Z digest=sha256:049ca1af8d12a52e79bdc98bcb01af2c481ba3faa7dc65d61fe9ba961d066321

Observation 50dadd83-1403-4265-a1ed-621708688e36 · inbound

REAgent: Requirement-Driven LLM Agents for Software Issue Resolution cites this paper.

REAgent: Requirement-Driven LLM Agents for Software Issue Resolution DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-05-11T05:50:58.577354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T17:56:32.201591Z digest=sha256:8404335b15709e3881dd8dac47767f2accf4c6c7c517d2514a7ff748142fad5a

Observation 8bdc3a86-13af-4453-92b1-7d7c60c9aa7d · inbound

Program Analysis Guided LLM Agent for Proof-of-Concept Generation cites this paper.

Program Analysis Guided LLM Agent for Proof-of-Concept Generation DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-05-11T07:36:00.209357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T17:05:01.812948Z digest=sha256:b57d4fe851920225611dc31e07c35dcbb728af7d10a951983b55748f91f9c3b7

Observation 7b67e5cd-5ccb-461f-aebf-1c53d331ff8f · inbound

Large Language Model Post-Training: A Unified View of Off-Policy and On-Policy Learning cites this paper.

Large Language Model Post-Training: A Unified View of Off-Policy and On-Policy Learning DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-05-11T00:30:53.710658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T18:28:58.515666Z digest=sha256:b4a7af7030d31e8c70dcf214a34dc8b1f7b306af875854070dcc83e8fb76c5d6

Observation 1b621726-91db-4533-bd3b-78915e2fbf99 · inbound

A Decomposition Perspective to Long-context Reasoning for LLMs cites this paper.

A Decomposition Perspective to Long-context Reasoning for LLMs DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-05-11T05:30:59.921211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T18:05:34.666937Z digest=sha256:ad88c8657d23286fdd5c17f350933cc9e144f0efe3078e397640e57cba6eb6f8

Observation b7147a38-937d-42db-917f-c52807dd3359 · inbound

PASK: Toward Intent-Aware Proactive Agents with Long-Term Memory cites this paper.

PASK: Toward Intent-Aware Proactive Agents with Long-Term Memory DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-11T05:30:59.695736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T18:05:40.242925Z digest=sha256:572cbfce94828a004d5382e21f24fc9c5d58ec08b95853bf398e79619522001c

Observation 7b017469-cb3b-4365-bf37-ff275c96ffe1 · inbound

Distributed Multi-Layer Editing for Rule-Level Knowledge in Large Language Models cites this paper.

Distributed Multi-Layer Editing for Rule-Level Knowledge in Large Language Models DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-11T00:20:54.363033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T18:34:55.124964Z digest=sha256:5dffeaf8ee354f2c53e379d82b0686b5ca37140c6e4382fe7245e4c1eb6a760a

Observation 65bb6f64-2c2b-4119-a4e1-b2ec558d85a1 · inbound

AssemLM: A Spatial Reasoning Multimodal Large Language Model for Robotic Assembly cites this paper.

AssemLM: A Spatial Reasoning Multimodal Large Language Model for Robotic Assembly DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-11T00:41:02.799435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T18:24:42.800233Z digest=sha256:03577f90587fda6dbabb137b7e196425d5d1177845756257b82d33bd2aec16b9

Observation bf0fbfd8-833f-440e-9bad-e68bbd2d8993 · inbound

AssemLM: A Spatial Reasoning Multimodal Large Language Model for Robotic Assembly cites this paper.

AssemLM: A Spatial Reasoning Multimodal Large Language Model for Robotic Assembly DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-07-12T23:37:01.875768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T23:37:01.875768Z digest=sha256:44ff76e525a5f03747b2e1d48355a0f5ccca3e14c0825b56fb7178ca8aafccbf

Observation 042b559c-d17b-4c4e-b231-ebf44eb4b579 · inbound

MAG-3D: Multi-Agent Grounded Reasoning for 3D Understanding cites this paper.

MAG-3D: Multi-Agent Grounded Reasoning for 3D Understanding DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-05-11T06:51:19.650353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T17:25:31.097385Z digest=sha256:b17c9651fe84bc2b7a3cef18e093837cdac443da8f04697a12544063e58e2a3f

Observation 298b9050-d7eb-4e5f-b062-7e3a740fa38c · inbound

A Minimal Model of Representation Collapse: Frustration, Stop-Gradient, and Dynamics cites this paper.

A Minimal Model of Representation Collapse: Frustration, Stop-Gradient, and Dynamics DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-05-11T08:30:57.509727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T16:36:05.941649Z digest=sha256:1da5e2a1445036639ca4e86b2aa16243a0dcb567691d0d96781dd8218dedb189

Observation a272958f-626b-4385-acec-5d1b58e3142b · inbound

AdverMCTS: Combating Pseudo-Correctness in Code Generation via Adversarial Monte Carlo Tree Search cites this paper.

AdverMCTS: Combating Pseudo-Correctness in Code Generation via Adversarial Monte Carlo Tree Search DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-05-11T08:40:57.457265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=arxiv_source observed=2026-05-10T16:35:16.056397Z digest=sha256:fbdf96f26958d6f7042abc91ffc692a72bf91dd0ece94e44f8a2fae5e353bce7

Observation 984a2ee0-8de0-4935-bdfa-7bcbfa87bb56 · inbound

AIM-Bench: Benchmarking and Improving Affective Image Manipulation via Fine-Grained Hierarchical Control cites this paper.

AIM-Bench: Benchmarking and Improving Affective Image Manipulation via Fine-Grained Hierarchical Control DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T10:11:04.164915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T15:35:49.575589Z digest=sha256:4600c76727109ab6d5d47d439ef3770223df3d20c7688ac700eb690aa42ef4ac

Observation e8ce957f-ee88-4df2-9491-cbf6cbcfe418 · inbound

Beyond Compliance: A Resistance-Informed Motivation Reasoning Framework for Challenging Psychological Client Simulation cites this paper.

Beyond Compliance: A Resistance-Informed Motivation Reasoning Framework for Challenging Psychological Client Simulation DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T09:46:08.222445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T15:50:19.740170Z digest=sha256:30e4e931b7d85d7e4b9bb2717d5b1a0f31b1decee75d82e323964245318254ad

Observation ad692bc9-4e1f-47c1-837f-a33883dd0011 · inbound

LLMs for Qualitative Data Analysis Fail on Security-specificComments in Human Experiments cites this paper.

LLMs for Qualitative Data Analysis Fail on Security-specificComments in Human Experiments DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-11T11:06:02.326352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T15:10:06.042478Z digest=sha256:acd7582c5fe1ff0d90f0c8d64d2a354c0c2dbd58d0cd61e3853a3c8c9d6a1dce

Observation c44ffc03-2ed3-4aea-bab0-c0f43cefca63 · inbound

OccuBench: Evaluating AI Agents on Real-World Professional Tasks via Language Environment Simulation cites this paper.

OccuBench: Evaluating AI Agents on Real-World Professional Tasks via Language Environment Simulation DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-11T08:20:57.872101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T16:43:27.037355Z digest=sha256:351ab33ba8bd5b0984b6b04856ffe8324768d3c227a976df4a3cd47ac395d93d

Observation 0f49b1ae-996b-417a-b919-22b21b97e154 · inbound

MAFIG: Multi-agent Driven Formal Instruction Generation Framework cites this paper.

MAFIG: Multi-agent Driven Formal Instruction Generation Framework DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-05-11T09:36:08.356350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T15:55:09.238676Z digest=sha256:95c39a2907af76361038bd21b610f3f9a0b45a56d1eeba14d2e0ae8ddbf1140d

Observation 03ec45f3-47f1-436a-a5bc-a79facb8c5c5 · inbound

From Context to Rules: Toward Unified Detection Rule Generation cites this paper.

From Context to Rules: Toward Unified Detection Rule Generation DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-05-11T08:55:59.980006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T16:25:59.609473Z digest=sha256:761647bb4f33ef05575239493d878d4ad9ae5ca90489a0545fc6f45b13334a51

Observation 04d5f922-529f-40a1-8c22-3110c41f5c46 · inbound

Exploring Knowledge Conflicts for Faithful LLM Reasoning: Benchmark and Method cites this paper.

Exploring Knowledge Conflicts for Faithful LLM Reasoning: Benchmark and Method DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-11T11:11:05.825915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T15:06:39.045269Z digest=sha256:bbdb81dd8ad243c7877956d6b3e6471ce55e5dcc802f85642102fb620fb81c40

Observation 6d35105f-3116-4f3f-b8be-9d4ef9f3c1d5 · inbound

Policy Split: Incentivizing Dual-Mode Exploration in LLM Reinforcement with Dual-Mode Entropy Regularization cites this paper.

Policy Split: Incentivizing Dual-Mode Exploration in LLM Reinforcement with Dual-Mode Entropy Regularization DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-11T11:21:04.446427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T14:59:49.488127Z digest=sha256:44618550dbde1ad48986328a0de1592c61fc45940dc4b84fd465590b943dda69

Observation 85ff35b0-69c9-4dfb-a0ce-7004c183cbb1 · inbound

CodeSpecBench: Benchmarking LLMs for Executable Behavioral Specification Generation cites this paper.

CodeSpecBench: Benchmarking LLMs for Executable Behavioral Specification Generation DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-11T11:26:03.628018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=arxiv_source observed=2026-05-10T14:54:51.051959Z digest=sha256:8ff6d55934c30b6111e23d9807138137f80507a115c84e279d07bbdc955fa26e

Observation 385ef909-07e9-4842-b8a9-50f2d593d6c1 · inbound

Every Picture Tells a Dangerous Story: Memory-Augmented Multi-Agent Jailbreak Attacks on VLMs cites this paper.

Every Picture Tells a Dangerous Story: Memory-Augmented Multi-Agent Jailbreak Attacks on VLMs DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-11T11:11:05.356850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T15:06:43.140580Z digest=sha256:aa444bc1be9955024f857b92bb206358c4404e12f6ef105ff28eb616ccb79255

Observation cfa79e9b-ea5c-4bdb-940b-ff0f86c3c748 · inbound

OSC: Hardware Efficient W4A4 Quantization via Outlier Separation in Channel Dimension cites this paper.

OSC: Hardware Efficient W4A4 Quantization via Outlier Separation in Channel Dimension DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-11T09:05:59.512684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T16:15:59.622461Z digest=sha256:484e5a4f14c24e7705608374218c2b38e51ab5e967424e13a338502d83cf50af

Observation 8bb0316c-d9d6-4de3-abf4-9ad2808c59e6 · inbound

VFA: Relieving Vector Operations in Flash Attention with Global Maximum Pre-computation cites this paper.

VFA: Relieving Vector Operations in Flash Attention with Global Maximum Pre-computation DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-11T09:16:00.125017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T16:10:40.858525Z digest=sha256:b3efbaa909889170fdf8d0920027732446a97c66f4f7f511494b084556182105

Observation 7e36aa95-5a6c-4717-9586-9bcfff9d2d7e · inbound

Towards Long-horizon Agentic Multimodal Search cites this paper.

Towards Long-horizon Agentic Multimodal Search DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 64

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T10:01:04.560729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T15:40:32.137708Z digest=sha256:9f02e192519d541e6e786ce4ed65419faff90af6ffcc0ae5b14b7ac54717a410

Observation 279996ff-7c3d-4756-b6e1-24f93ed7be70 · inbound

Round-Trip Translation Reveals What Frontier Multilingual Benchmarks Miss cites this paper.

Round-Trip Translation Reveals What Frontier Multilingual Benchmarks Miss DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-05-10T16:30:35.059816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-07-11T11:50:26.030339Z digest=sha256:799188f009d72c33bd6908789980ab57157b07c0cdc651494c8d8ccea06c1c73

Observation 76b6f5e4-9898-4c4a-b3ee-4da0b5def3bb · inbound

Empirical Evidence of Complexity-Induced Limits in Large Language Models on Finite Discrete State-Space Problems with Explicit Validity Constraints cites this paper.

Empirical Evidence of Complexity-Induced Limits in Large Language Models on Finite Discrete State-Space Problems with Explicit Validity Constraints DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-05-10T14:15:29.446584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=arxiv_source observed=2026-05-10T14:12:45.438246Z digest=sha256:d61223f3e733a6cad93c94bd7d62e154ac2f023526623c9cbff22222d16ecca2

Observation 970d04dc-baa6-478a-a6ed-e76a28633e65 · inbound

Training-Free Test-Time Contrastive Learning for Large Language Models cites this paper.

Training-Free Test-Time Contrastive Learning for Large Language Models DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-05-10T14:05:29.466800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T14:02:10.277246Z digest=sha256:0b6b7eceebc5ff0c684071d4707e87739bdf7fd8ffa2a0dd99625fce5c373960

Observation ab0ddab6-c4e4-44ff-a596-25d1603a6881 · inbound

SparseBalance: Load-Balanced Long Context Training with Dynamic Sparse Attention cites this paper.

SparseBalance: Load-Balanced Long Context Training with Dynamic Sparse Attention DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 6

Resolution
metadata mismatch
local_arxiv, observed 2026-05-10T13:55:28.054554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T13:51:03.591878Z digest=sha256:da0d969408c3c8d5812c52cd805cd969d8da222aa29ff2fd4a389535cbd8518e

Observation 255b7f7a-e94d-4176-a773-7119dee936d1 · inbound

POINTS-Seeker: Towards Training a Multimodal Agentic Search Model from Scratch cites this paper.

POINTS-Seeker: Towards Training a Multimodal Agentic Search Model from Scratch DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-05-10T13:10:26.383253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T13:09:24.304696Z digest=sha256:e3f1dd5d8d671be3751a2a015deda6ecf3cb873a5aaff038b170a012df418436

Observation 5d8a187a-8c76-44a4-bb73-fc978dadfa8d · inbound

Evaluation of Agents under Simulated AI Marketplace Dynamics cites this paper.

Evaluation of Agents under Simulated AI Marketplace Dynamics DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 64

Resolution
verified exact
local_arxiv, observed 2026-05-11T11:51:00.738264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T12:33:04.984953Z digest=sha256:1cc62fbbefb9d231cf4d53702f9a0b81501c4ca51122d904c65a694cc3a852dd

Observation d75c90ba-faa7-4f8c-ad2a-4252c0452481 · inbound

HWE-Bench: Benchmarking LLM Agents on Real-World Hardware Bug Repair Tasks cites this paper.

HWE-Bench: Benchmarking LLM Agents on Real-World Hardware Bug Repair Tasks DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:05:26.778195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T11:27:11.257817Z digest=sha256:990a48bc4ef175e1078c3db1d892d224fd3d8f475758c6ea7e67609836a17b8e

Observation cdcd8b51-5481-456c-bb96-3b7b0d9badec · inbound

AdaSplash-2: Faster Differentiable Sparse Attention cites this paper.

AdaSplash-2: Faster Differentiable Sparse Attention DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T13:05:26.778195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T11:19:13.934756Z digest=sha256:ae4935b12daa05e28a9d7b5e5193f2e24ca0249c776753e2d367f837ea532d54

Observation c2f5eada-66bc-447e-9d7e-308ddf3c4614 · inbound

HarmfulSkillBench: How Do Harmful Skills Weaponize Your Agents? cites this paper.

HarmfulSkillBench: How Do Harmful Skills Weaponize Your Agents? DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:05:26.778195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T10:32:37.967401Z digest=sha256:28413c28caeebbddeb0eb5c043ffcc645b39feb904ac4548389af60cef0d1e0d

Observation f085f3d7-58ed-4fc0-ad20-b3b120d9cd1b · inbound

GTA-2: Benchmarking General Tool Agents from Atomic Tool-Use to Open-Ended Workflows cites this paper.

GTA-2: Benchmarking General Tool Agents from Atomic Tool-Use to Open-Ended Workflows DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:05:26.778195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T08:42:57.035226Z digest=sha256:c6c05496e66556679bf02043f5f2e8510a8258d50f810ab8f848cf616a9fdb81

Observation f387ef77-4855-4fe0-bfeb-926ac978b70f · inbound

Bridging the Gap between User Intent and LLM: A Requirement Alignment Approach for Code Generation cites this paper.

Bridging the Gap between User Intent and LLM: A Requirement Alignment Approach for Code Generation DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:05:26.778195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T08:13:43.804756Z digest=sha256:564eb1db05f8c99dae32c5a1ac36a5c28333e3a7f381d6215fb88352036d7a14

Observation 6942799a-8cf2-42e9-88e1-f03b6331033d · inbound

Analyzing Process Data from Computer-Based Assessments: A Tutorial on Preprocessing, Feature Extraction, and Model-Based Inference cites this paper.

Analyzing Process Data from Computer-Based Assessments: A Tutorial on Preprocessing, Feature Extraction, and Model-Based Inference DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:05:26.778195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T07:16:09.962300Z digest=sha256:4c12d1a4508e4815da208ea185869cef1590496533e0bf79756e11beeeaed508

Observation f35622dd-6a67-44cd-95be-f8bc84d86799 · inbound

RoTRAG: Rule of Thumb Reasoning for Conversation Harm Detection with Retrieval-Augmented Generation cites this paper.

RoTRAG: Rule of Thumb Reasoning for Conversation Harm Detection with Retrieval-Augmented Generation DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:05:26.778195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T06:56:24.757369Z digest=sha256:924468927435235b2b5883ef0cef5937376830976266971a8dc882557cff1df8

Observation b37fa6f7-5993-4e4d-990e-80537f8855e7 · inbound

Hive: A Multi-Agent Infrastructure for Algorithm- and Task-Level Scaling cites this paper.

Hive: A Multi-Agent Infrastructure for Algorithm- and Task-Level Scaling DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:05:26.778195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T06:31:36.776819Z digest=sha256:9249c0324756515e0a1714d445ef74eec6346f40603e082d63fa241a0c4a270a

Observation 06450ae4-6604-45f9-a2a4-66e51efd95b7 · inbound

STRIDE: Strategic Iterative Decision-Making for Retrieval-Augmented Multi-Hop Question Answering cites this paper.

STRIDE: Strategic Iterative Decision-Making for Retrieval-Augmented Multi-Hop Question Answering DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:05:26.778195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T06:03:51.049997Z digest=sha256:f85e2007618570be37cd9bb42174d61fcb197de7a440fdd25fed58969f149743

Observation 9a4f5b73-431f-4c61-86ac-007fc5698fb9 · inbound

TrafficClaw: A Generalizable LLM Agent in the Unified Physical Environment for Urban Traffic Control cites this paper.

TrafficClaw: A Generalizable LLM Agent in the Unified Physical Environment for Urban Traffic Control DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:05:26.778195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T05:47:45.651391Z digest=sha256:bd470372f663ddf7ecc086d07b4ebbce493f11f3d0f0d3e2c776a7b8d4824463

Observation 754bad7d-3d62-4951-99a2-3ac73b39eb6a · inbound

TrafficClaw: A Generalizable LLM Agent in the Unified Physical Environment for Urban Traffic Control cites this paper.

TrafficClaw: A Generalizable LLM Agent in the Unified Physical Environment for Urban Traffic Control DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-12T19:02:56.493440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T19:02:56.493440Z digest=sha256:7551af7d71509f8ddeeea915413127a5380ec80d3370e048b913b7916be9bd09

Observation 4ed46c93-a32e-48ec-9d25-a3df915ab42f · inbound

Matlas: A Semantic Search Engine for Mathematics cites this paper.

Matlas: A Semantic Search Engine for Mathematics DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:05:26.778195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T05:32:30.216388Z digest=sha256:09a3b2607d8a735e996f55c22274fb40fa3585d586212dd5d1317754267833db

Observation 6574306d-fa96-4e0c-a821-1d64ff85d36f · inbound

Single-Language Evidence Is Insufficient for Automated Logging: A Multilingual Benchmark and Empirical Study with LLMs cites this paper.

Single-Language Evidence Is Insufficient for Automated Logging: A Multilingual Benchmark and Empirical Study with LLMs DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T13:05:26.778195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T05:18:59.656997Z digest=sha256:d8da525866776c6a805acb7d04ef93e35cd7508c8b269fc9f627589e397abbc2

Observation f56f7d1a-8ec9-4245-be0f-0587b3fbd4dc · inbound

Co-evolving Agent Architectures and Interpretable Reasoning for Automated Optimization cites this paper.

Co-evolving Agent Architectures and Interpretable Reasoning for Automated Optimization DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:05:26.778195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=arxiv_source observed=2026-05-10T05:21:51.915690Z digest=sha256:ac5f4c5177cb2f140b8672108cd1152c72213a9ac50fba8693af8e7fc02fa73e

Observation d34ea390-4133-4b99-a1bd-bf9a4122d463 · inbound

Contrastive Attribution in the Wild: An Interpretability Analysis of LLM Failures on Realistic Benchmarks cites this paper.

Contrastive Attribution in the Wild: An Interpretability Analysis of LLM Failures on Realistic Benchmarks DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:05:26.778195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T05:12:31.218055Z digest=sha256:ce29a145ece6800d08142cbd2d19ac1ca8f114ef130d1e427b8b2d26c50be37e

Observation 2b18eb95-b4f8-4731-aee8-f5fe235c72ea · inbound

Neural Garbage Collection: Learning to Forget while Learning to Reason cites this paper.

Neural Garbage Collection: Learning to Forget while Learning to Reason DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 9

Resolution
malformed identifier
arxiv_id, observed 2026-05-10T13:05:26.778195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T04:57:12.923258Z digest=sha256:86c7490a3b87baf6c4a6ff29d6b5e03f7d9aa8a56976d8efc5fa88ababb7804f

Observation 2b1cc089-6564-4381-b24e-813f947d3488 · inbound

Model in Distress: Sentiment Analysis on French Synthetic Social Media cites this paper.

Model in Distress: Sentiment Analysis on French Synthetic Social Media DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-11T12:01:05.920313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T04:20:02.018560Z digest=sha256:61554454ace79646c45f47df1927cd216ef266a38a0aae6a3ddf2479ceab0e30

Observation 42547a3c-5c98-487d-b16a-fc74a4eff51c · inbound

MFMDQwen: Multilingual Financial Misinformation Detection Based on Large Language Model cites this paper.

MFMDQwen: Multilingual Financial Misinformation Detection Based on Large Language Model DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-05-11T12:31:07.794522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=arxiv_source observed=2026-05-10T03:21:32.522414Z digest=sha256:096193f5b70ebdcdc11c19647256f608c763f9e5abfef8eace719d5a18051839

Observation 05293769-6d34-4005-b0cd-d790c9ccf720 · inbound

Agent-World: Scaling Real-World Environment Synthesis for Evolving General Agent Intelligence cites this paper.

Agent-World: Scaling Real-World Environment Synthesis for Evolving General Agent Intelligence DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:05:26.778195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T05:24:00.503836Z digest=sha256:c5d1b1b4e501323e4c5a09e3b99f1a7906b60fb73371f61b9671a210d417b8a7

Observation 3d853668-9ad2-4e5f-b01e-94eec4a500bf · inbound

ComPASS: Towards Personalized Agentic Social Support via Tool-Augmented Companionship cites this paper.

ComPASS: Towards Personalized Agentic Social Support via Tool-Augmented Companionship DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:05:26.778195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T04:54:54.317884Z digest=sha256:9dcfe504f42faf6539de4bef4cd929a9a8822287fe3c8255111cea0251a10690

Observation 48ebe058-6d54-41dd-8d3a-d328754bb8ea · inbound

Using large language models for embodied planning introduces systematic safety risks cites this paper.

Using large language models for embodied planning introduces systematic safety risks DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:05:26.778195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-27T06:30:09.085275+00:00.

source=pdf_text observed=2026-05-10T04:45:46.483867Z digest=sha256:5fee73f33096c9c97953bef9b0bd95ec586b738ad0a66cbd069655d8d5c1a4b3