Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T21:16:42.794738Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 5 inbound Pith citation observations for arXiv:2507.00699.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T21:16:42.794738Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-01T08:17:10.481202Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
46 of 46 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation 98959035-344d-4dd5-b524-8cc2a7ddf776 · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback Soen-101: Code generation by emulating software process models using large language model agents,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d930fc1b-9079-4376-9294-e8d94752079b · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback ROCODE: Integrating Backtracking Mechanism and Program Analysis in Large Language Models for Code Generation
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 815bdd69-f192-49d1-a45f-c0730df2b1fd · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback Skcoder: A sketch-based approach for automatic code generation,
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05f51ad5-7b73-48d2-af70-60563fdb60ed · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback Enhancing Code Generation via Bidirectional Comment-Level Mutual Grounding
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0d3c8c6-b55a-49ef-ad5e-effa4a22987b · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback Fixing Large Language Models' Specification Misunderstanding for Better Code Generation
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b61f1b99-39fe-4665-b59b-aa73c79502a7 · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback Evaluating Large Language Models Trained on Code
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b47c095f-2342-44b3-9824-95727e3a39bc · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback Program Synthesis with Large Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3bc9bdaa-f0df-4836-91cb-e64f7b267922 · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback A Survey on Evaluating Large Language Models in Code Generation Tasks
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43c39c9e-b5aa-4659-bec4-0e0753fedd37 · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback Instruction-following evaluation for large language models,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da751ba6-c1d1-47df-985f-5a1f713dda1c · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback InFoBench: Evaluating Instruction Following Ability in Large Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1774027-b0c4-4f17-81a5-df9a12cc59c7 · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback CodeIF: Benchmarking the Instruction-Following Capabilities of Large Language Models for Code Generation
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1193c980-a560-4bd8-85e2-898adf58dbbe · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback Codeif-bench: Evaluating instruction-following capabilities of large language models in interactive code generation,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79701b5d-baf1-4604-9754-541bf883a75a · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback A hierarchical and evolvable benchmark for fine-grained code instruction following with multi-turn feedback,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ea3e0791-063d-489c-a370-71143a52ca3b · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback Beyond Functional Correctness: Investigating Coding Style Inconsistencies in Large Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cea143ab-1230-4175-9893-9bb48f9b74f8 · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback FollowBench: A Multi-level Fine-grained Constraints Following Benchmark for Large Language Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb171d85-fe5a-43b1-837c-f0361ca7f398 · outbound
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f6eabef0-6a5c-4206-bbcc-a25731a628f8 · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback Prompt-Based Cost-Effective Evaluation and Operation of ChatGPT as a Computer Programming Teaching Assistant
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1e647956-e148-494d-8d08-dc3e467a846e · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback tree-sitter/tree-sitter: v0.25.5,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a19eca15-2d4a-4f3a-9595-400190df6f7e · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback GPT-4 Technical Report
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6c825af-d137-4486-9936-ea7107ce2302 · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback ROUGE: A package for automatic evaluation of summaries,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5a0c9c0b-bf7f-4272-9ae4-e7c2665ef132 · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback GPT-4o System Card
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26e3e459-d30c-460a-a7ed-295320b0c912 · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback Claude 3.7 sonnet and claude code,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 84baed67-510b-43ca-b183-a206c5ce0c6f · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8abd2f57-2fd9-40f2-8480-f908485b89dd · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback DeepSeek-V3 Technical Report
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a297632-837c-4105-8ad7-12f0e53f9b6d · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback Qwen3 Technical Report
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1927e71d-37ba-4f26-8e8d-3d54415b000e · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback The Llama 3 Herd of Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37067ce1-3533-4b5b-ac4f-c3abc6105d35 · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback Efficient memory management for large language model serving with pagedattention,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 984bcaa4-5042-4ed1-8926-f2ea96675d1d · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback Guesslang: A neural network to guess the programming language from code snippet,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 154ca5db-694d-4241-b4e1-f0df76995d2e · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback DS-1000: A Natural and Reliable Benchmark for Data Science Code Generation
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a938fd60-4e67-407d-864f-5b65fe7c3f08 · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback ClassEval: A Manually-Crafted Benchmark for Evaluating LLMs on Class-level Code Generation
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72cdc08a-3578-47b8-8024-0aced3a8923c · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback Learning-based widget matching for migrating gui test cases,
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6008267-b874-4237-a2fc-97efca23c204 · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback DevEval: A Manually-Annotated Code Generation Benchmark Aligned with Real-World Code Repositories
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 459eb448-1417-48ee-bb16-61b7931bcb71 · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback RustEvo^2: An Evolving Benchmark for API Evolution in LLM-based Rust Code Generation
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0b49b14-e3ac-4ed8-9abc-f6418ac8b6db · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback Feedbackeval: A benchmark for evaluating large language models in feedback-driven code repair tasks,
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a95a8c12-ebd9-42f2-bcf9-72e7ba49ac5e · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback Benchmarking complex instruction-following with multiple constraints composition,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 472da2bc-29fc-4846-9f68-71cf6927d5ff · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback Generating Equivalent Representations of Code By A Self-Reflection Approach
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 674a015e-81df-4cc0-8811-24c3c0e410d4 · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback Codescore: Evaluating code generation by learning code execution,
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0639506-78c1-4147-a23d-9a7d2faeb7e1 · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback Benchmarking Complex Instruction-Following with Multiple Constraints Composition
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40aa013e-3916-4b50-84d9-62400e7e56f0 · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback WizardLM: Empowering large pre-trained language models to follow complex instructions
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca44739c-0b3f-4e91-aadd-61e8e4bc0971 · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback Magicoder: Empowering Code Generation with OSS-Instruct
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d317827-9bbe-47f6-a99c-de1dfb3fb9b5 · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback Self-Instruct: Aligning Language Models with Self-Generated Instructions
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57834f6e-a9ec-4a0c-8657-fbc15a330194 · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback Genetic instruct: Scaling up synthetic generation of coding instructions for large language models,
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9b693b4d-0478-4898-8651-a21d4f48dba9 · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback WaveCoder: Widespread And Versatile Enhancement For Code Large Language Models By Instruction Tuning
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54853ea3-73e6-4c99-80e1-4982df7be0bd · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback Instruction-Following Evaluation for Large Language Models
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 918c7b8c-6c40-4878-8711-7b74779eb127 · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback SOEN-101: Code Generation by Emulating Software Process Models Using Large Language Model Agents
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eabb4b14-24c4-49af-a418-bedcec7e29c1 · outbound
A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback Genetic Instruct: Scaling up Synthetic Generation of Coding Instructions for Large Language Models
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39c86638-d467-4dad-9eba-bfbb7c127715 · inbound
Bridging Generation and Training: A Systematic Review of Quality Issues in LLMs for Code A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 290c59ff-805d-4373-ba60-59b11e8a9b3e · inbound
When Should Models Change Their Minds? Contextual Belief Management in Large Language Models A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 57e95625-d60e-4a7b-a963-890c4f78eb3d · inbound
Prompt Governance? On Governing Technologies Governed by Natural Language A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c805b868-143f-4aaf-89c6-f46868efae58 · inbound
CodeChat-Eval: Evaluating Large Language Models in Multi-Turn Code Refinement Dialogues A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9f50b92b-d06f-40a4-a0b3-2d24dcfef8ae · inbound
CodeChat-Eval: Evaluating Large Language Models in Multi-Turn Code Refinement Dialogues A Hierarchical and Evolvable Benchmark for Fine-Grained Code Instruction Following with Multi-Turn Feedback
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.