Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-10T18:29:09.202017Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 0 inbound Pith citation observations for arXiv:2604.06683.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-10T18:29:09.202017Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
35 of 35 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 82741301-51c7-42dc-aa9a-70212ffc191e · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation Mathqa: Towards interpretable math word problem solving with operation-based formalisms
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8b2ecd7d-23f2-4028-b36b-623bd33ec61d · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation Program Synthesis with Large Language Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e8f42a16-a396-4752-b0b7-8c2f24543b01 · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation Software architecture documentation in practice: Documenting architectural layers
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a1eb5317-b036-4790-b7bf-b609390b4492 · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation Assessing the suitability of large language models in generating uml class diagrams as conceptual models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 08a4946e-01ff-4797-870a-86ea7f3bfbf0 · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation On the assessment of generative ai in modeling tasks: an experience report with chatgpt and uml
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 14a9cace-8f07-4bc0-9da5-692b80e7f91a · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation Evaluating Large Language Models Trained on Code
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation de570f53-e4db-454a-ad3c-17013dcae278 · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation Can llms generate architectural design decisions?-an exploratory empirical study
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c5c24daa-1288-4ca6-ba1d-6d139e1ca80c · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation CrossCodeEval: A Diverse and Multilingual Benchmark for Cross-File Code Completion
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c6cdb260-851c-441f-af74-1efa4759c57f · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation ClassEval: A Manually-Crafted Benchmark for Evaluating LLMs on Class-level Code Generation
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2093ceb7-6e83-4e33-8950-c90976c7c51d · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation Erni and C
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 67dc89eb-c377-4cbf-9f32-f949bd3345fa · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 19e60196-bab4-4112-9814-f85aa389a6d6 · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation A survey on llm-as-a- judge.The Innovation
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 864208b2-b7c8-484a-90e4-0355648f2f74 · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation MetaGPT: Meta programming for a multi-agent collaborative framework
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 04b7dddc-bda8-47ab-a31d-86aa83496536 · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation Large language models for software engineering: A systematic literature review.ACM Transactions on Software Engineering and Methodology, 33(8):1–79
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0319fbe5-e359-4dbd-a0bb-767511f6469a · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation Role of ai in requirements engineering
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e6e759d7-9800-43e2-a6be-6d2279fe4791 · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation The unified modeling language reference manual
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f561b1c9-a849-4f95-a97b-443185be38ed · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation Testgeneval: A real world unit test generation and test completion benchmark
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f9463452-975f-4659-905c-a9c440bb1d46 · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation Survey of hallucination in natural language generation.ACM computing surveys, 55(12):1–38
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d5b74280-14a8-4aa0-ad23-a283df4b56fa · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation Prompting large language models to tackle the full software development lifecycle: A case study (devbench)
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 71a01ba6-55e0-4628-8dbf-1bdcb945c2ee · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation Rouge: A package for automatic evaluation of summaries
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6da977d5-731e-4e0b-ab3f-4a838a1c59cf · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation C4 model: a research guide for designing software architectures
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f6c957cb-3acb-45ff-b1c0-bb24c0c36003 · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation Rec- ommended practice for architectural description of software intensive systems
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bb23b425-4288-411a-b143-b7b846754eca · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation Bleu: a method for automatic evaluation of machine translation
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9162af84-f866-4733-9947-c930167df4d0 · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation Sys- tematic literature reviews in software engineering—enhancement of the study selection process using cohen’s kappa statistic.Journal of Systems and Software, 168:110657
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 06a649f3-b440-4879-8271-36d2b5dd7774 · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation Software Architecture Meets LLMs: A Systematic Literature Review
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4af3e750-4c75-4789-84a8-2aa44fd6d5a9 · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation MermaidSeqBench: An Evaluation Benchmark for NL-to-Mermaid Sequence Diagram Generation
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0cacff00-a0b5-4b3c-86be-7e00ffac72d5 · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation Application of the tree-of-thoughts framework to llm-enabled domain modeling
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d8a70f91-457a-4af0-99d7-cb4000cfe1e7 · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation Collaborative llm agents for c4 software architecture design automation
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 46c1c2dc-a90d-43a1-b69c-dd8133990c65 · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation Contest: A unit test comple- tion benchmark featuring context
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 636e62ab-10e8-4820-8303-bbc44aa03ccd · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation Testeval: Benchmarking large language models for test case generation
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c48d842a-8a82-40a1-bbb6-c1333ce61a16 · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation OpenHands: An Open Platform for AI Software Developers as Generalist Agents
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 74898c61-df5d-4224-8563-710129465bea · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation Repocoder: Repository-level code completion through iterative retrieval and generation
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9f6baf82-1bc8-46de-baa5-fb76f067c1ed · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation Towards realistic project-level code generation via multi-agent collaboration and semantic architecture modeling
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 85d692d5-015c-446c-a72e-0c1982abf198 · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation Judging llm-as-a-judge with mt-bench and chatbot arena.Advances in neural information processing systems, 36:46595–46623
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 70757e43-0783-4fb0-ba10-7d1c729123f9 · outbound
Benchmarking Requirement-to-Architecture Generation with Hybrid Evaluation Codegeex: A pre-trained model for code generation with multilingual benchmarking on humaneval-x
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.