Pith. sign in

Paper Citation Record · LEDGER

LIFEBench: Evaluating Length Instruction Following in Large Language Models

As of 7 August 2026, this Paper Citation Record lists 100 of 129 outbound references and 4 inbound Pith citation observations for arXiv:2505.16234.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.16234 v2

Coverage vector

measured 100 of 129 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:08:08.453311Z

measured 104 of 104 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T15:28:14.401581Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-15T00:39:35.853384Z

Reference resolution

100 of 129 outbound references displayed

  • verified exact1
  • verified fuzzy13
  • unresolved86
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d03fcbf4-d152-4577-8b0c-e084be42cdbe · outbound

This paper cites Abedi Firouzjaei.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Abedi Firouzjaei

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:59.915175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:59.915175Z digest=sha256:45e8b9d9509bc7bfc9ffadc5213713a446430ad945621ed24362d67e0bebc5ce

Observation 52722d46-9a3d-4a13-9c58-5dec7b1fdc4a · outbound

This paper cites Alzantot, Y.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Alzantot, Y

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:59.980923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:59.980923Z digest=sha256:d65525d20f9d9e1bc5040357095a2abb759e58b0712e05fb08964726d1ceb0fd

Observation d62c8ae7-3dfe-454f-bb53-6b6025bae52f · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:00.100922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:00.100922Z digest=sha256:76ddbecb3ae02a8077285071d33032a8eef8b05b1fe878097300ff01769199b8

Observation 12609bb8-f46c-4f25-88c3-f655ffd5fb18 · outbound

This paper cites Claude 3.7 Sonnet and Claude Code.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Claude 3.7 Sonnet and Claude Code

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:00.243342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:00.243342Z digest=sha256:23b07440083c7d8fe29fff8511e3b501c59a17ae823baa4298a7515f58bdf380

Observation 75e7b2b1-f57d-478d-8d66-009a1e343575 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:00.370074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:00.370074Z digest=sha256:77c42257db05695f1fcab07d9841629e5b52f65d811113f71fb0d8ff652d7a21

Observation 76e6cf04-00d6-4840-bdfa-52c5db7dd6ca · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:00.482208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:00.482208Z digest=sha256:0cc3b7bca5c7c702eafa3f9a6f165d5d27be7e2231149cc503d2d01092c502ef

Observation f427873a-8f62-4dc0-91be-8d9997e3035f · outbound

This paper cites LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks.

LIFEBench: Evaluating Length Instruction Following in Large Language Models LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:00.616769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:00.616769Z digest=sha256:5a3062c4789117433a5e860cd2d9f8972c4c8a0c12152b5324406338c890c17f

Observation c8073588-1d7d-4018-a79e-6ede488102e6 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:00.759419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:00.759419Z digest=sha256:a934f4d21c93fe777631ffa58af23cecb15bb662bf4405fa2f62ca81590d2202

Observation 85558c24-4962-4f40-9ab3-7f563662fba8 · outbound

This paper cites Bordes, Y .-L.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Bordes, Y .-L

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:00.923158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:00.923158Z digest=sha256:ea671e025b753286e3f49a346290c0eca5758204b4f6de7c6c413483e2d27403

Observation 8a000c4b-be89-4539-878b-9cbd7123646c · outbound

This paper cites Bosselut, A.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Bosselut, A

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.022186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.022186Z digest=sha256:9ef3afbf448ac282c6d5240a64df9fedcf91a7b1310f65de3e7ae6fbe2f51d4f

Observation dee3ee11-3c31-4520-b503-0d57e76e40e9 · outbound

This paper cites Butcher, M.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Butcher, M

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.144208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.144208Z digest=sha256:f7770153c400b75ec9d8a0e013f894baef26fb9c9ffb771c4edf8623b35505df

Observation 0cec495e-48db-4aa0-aa49-185bd1bb0555 · outbound

This paper cites Doubao-1.5-Pro.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Doubao-1.5-Pro

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.233200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.233200Z digest=sha256:14ba3da5495ca322ebab640821516a799508f30f2b4305539d395ec9cd0b1379

Observation d83dcc59-f853-4d12-ad7f-9fb3c1bf25dc · outbound

This paper cites Chang, X.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Chang, X

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.312710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.312710Z digest=sha256:54e15fd3f63d177a328ce08ee4143ba785101fde0e4430293575497a52c1d54e

Observation 2845bc06-2ac8-469d-a358-83afe998016c · outbound

This paper cites Mistral 7B.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Mistral 7B

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.389348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.389348Z digest=sha256:b927ef25afb9d2efbb7f4572d97ca683d03e240fb85910e91a48a846643fad6b

Observation 3a2e6bd6-f44e-4228-b0da-817dd6009ed0 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.475747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.475747Z digest=sha256:714ddda0a9cd5aef9e3189fdb3b4281191cc664233e5acce58bf8e44447f0d8a

Observation 1b9c43a9-c31c-4586-b9f3-67543687f7d8 · outbound

This paper cites Chen and C.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Chen and C

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.576145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.576145Z digest=sha256:2a29b232aaa0e7c2ed001d241119cc6fb27993f3500defe182179908d0343f00

Observation 7b252921-98f3-457a-b125-ee69f5ac761e · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.643294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.643294Z digest=sha256:6359b09a8bf47b35ffaf7d9de221d113bafc9c2d30d8fc8da23cf30db2088c16

Observation f5bcecf5-7db7-4429-9b1b-7a332a67e84a · outbound

This paper cites Chiang, L.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Chiang, L

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.723680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.723680Z digest=sha256:91c76a908b94512cf188309a791c7752104426e3050191a15e63230f896debdc

Observation 9f64694b-b587-4229-9e0e-87e2b99537da · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.821582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.821582Z digest=sha256:b4d9b32fd4f178f233886fdf0410a4e331cdc5fed09a537a3417fac35a40bcc0

Observation 9ee681a9-1b69-4f20-84b7-9353dea5e07a · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Training Verifiers to Solve Math Word Problems

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.907072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.907072Z digest=sha256:a5da7202b8d6f276b6e967a00884d94d6cf06401cf99c7cdcb3fbb44c061d7f4

Observation 5eef2f6c-3b9a-4e35-bcc9-6219b1c0f537 · outbound

This paper cites Cohan, F.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Cohan, F

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:02.090683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:02.090683Z digest=sha256:9c8f655f22dc7ade46cb844b9d7e6d5c6a7b636ba75bc3948864d0d18d1fd814

Observation b8a754b4-d26d-4973-8e81-433936305d43 · outbound

This paper cites Collobert, J.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Collobert, J

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:02.249374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:02.249374Z digest=sha256:52a070ab209ba8685c3429b494e91965cba55ac3e83c3bcb2ec661633a0ecb58

Observation 642495c2-3a0b-4a7a-9b68-31d725882625 · outbound

This paper cites LCFO: Long Context and Long Form Output Dataset and Benchmarking.

LIFEBench: Evaluating Length Instruction Following in Large Language Models LCFO: Long Context and Long Form Output Dataset and Benchmarking

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:08:12.116541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:02.336424Z digest=sha256:6d84c414b8be211f41efd26aa077079960b55d2bf5d7370a2d855a16c6351e20

Observation 8575674f-54ee-4641-8746-f565c54fe5dc · outbound

This paper cites Davidson, D.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Davidson, D

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:02.496551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:02.496551Z digest=sha256:36d3a9bb5788d649398617b66120033ae43876eef1f9e71cd311dbf94fe14801

Observation 264207e3-a4f7-49e7-9fea-48f19a70c967 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:02.698130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:02.698130Z digest=sha256:22f717a3dd93d6034fc5181e665fe6f1a1a241a02b3180858e43c54929a33119

Observation 207f2d8c-d40f-4223-8f31-f95caddd5bb1 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:02.802654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:02.802654Z digest=sha256:b127535f27e730c22b96c6e445da300f8f0255d533e754957f36defda474516c

Observation 348202ab-4617-46ca-b372-c9179b218782 · outbound

This paper cites Dubois, C.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Dubois, C

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:02.954068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:02.954068Z digest=sha256:d6048c068c19d31363f32d98002c5ab0524f4ae3adfcf4d388f2afdfe44021ea

Observation 74eced30-2b84-4219-8b65-3311a6304d66 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.106742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.106742Z digest=sha256:f9b868ca884226d63a9c4acaef722189671bc707fd25fe05763e322af5ce0a40

Observation 29979e30-9658-4795-aae9-fb2f03def193 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.246808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.246808Z digest=sha256:8ace752789abc6c5182add6267c07915e2f469cc3d78a306c4ecc0cfb5b3a1e7

Observation 02d05053-243d-4d2a-9694-2ef27cfb62d8 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.356120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.356120Z digest=sha256:e258b752061cae42d54fb9402ac400770c84b31ce5ececb2a9fab0a5f9c109c8

Observation 7520112c-1822-41e5-aeb0-119846f48bbb · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.507635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.507635Z digest=sha256:678472c147f5dbc95b6d3582951f6cb19be334bd2e9081893a8645931f59cbd2

Observation 1f012298-b49c-4271-9aa4-65d5d67af0c8 · outbound

This paper cites Foundation.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Foundation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.619331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.619331Z digest=sha256:3cc95c306e4a4ce9303ca66718c0248b1d504168f04e0cb6227d6cee9f80be34

Observation c4b3fe12-b0fc-4900-b9c7-ee56d1e7a38c · outbound

This paper cites ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools.

LIFEBench: Evaluating Length Instruction Following in Large Language Models ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.761610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.761610Z digest=sha256:aa9babed7c2db7ba866be21432d421ec4b9d1dbab6fb2150fbb3e117eb92b7cb

Observation 31377b83-3942-4967-b76e-687b8a78e71f · outbound

This paper cites Gemini 2.0 Flash.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Gemini 2.0 Flash

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.821979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.821979Z digest=sha256:0adbebe535e1919ca3ae9c09c2354a9b9f86c079848362e3bffd93709bee8686

Observation ed19b53f-3d7b-4290-8c8b-879e4119f478 · outbound

This paper cites Gemini 2.5 Pro.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Gemini 2.5 Pro

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.877823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.877823Z digest=sha256:6f6cb84b004c15d901687dbc748c8591232e30f8a516b1522fa8cb886768d77d

Observation 9e937a04-a837-4446-a22f-80a85be20f3b · outbound

This paper cites The Llama 3 Herd of Models.

LIFEBench: Evaluating Length Instruction Following in Large Language Models The Llama 3 Herd of Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.923252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.923252Z digest=sha256:22c51b1308e6675cd1a591e6b52903a9d82616e6f02a389680804d88e78230e9

Observation f19c070b-17d7-4cfa-9a32-0a7e8985ddd5 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.998056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.998056Z digest=sha256:16f8dadd9cdcea23621874fa62b450b0607d6bf7f573ec372e5091b1b4335fed

Observation caef362d-1e6e-4407-81dc-227fe3ba992a · outbound

This paper cites A Survey on LLM-as-a-Judge.

LIFEBench: Evaluating Length Instruction Following in Large Language Models A Survey on LLM-as-a-Judge

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.044070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.044070Z digest=sha256:4b9f66f6d7c311effa4e0f6d9190efec4eeec9d7cbe6ae62507e48704c22851b

Observation 0a1bbbcf-23ea-4ed9-9ec2-9ab31402e39f · outbound

This paper cites Length Controlled Generation for Black-box LLMs.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Length Controlled Generation for Black-box LLMs

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.110426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.110426Z digest=sha256:414a50cfc321ae46fd971ae373af3e36182fdc45009cf75b7658af38e752a2e0

Observation 6e458285-5368-4628-a9a7-9cfa31d92340 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

LIFEBench: Evaluating Length Instruction Following in Large Language Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.157989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.157989Z digest=sha256:bd01b5cdb3142ccc672a6c8e68256d2f11a9f2e383921ff818c4ec0cbdcbb93a

Observation f3ab0426-6758-4aac-bd61-241b9d68ca63 · outbound

This paper cites Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.207512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.207512Z digest=sha256:f85b1366fb2e026de1464ca3d07786a2f16c979b6743c872b162b557866fff16

Observation fd5c61b4-e6f6-45aa-8607-9abb2af7ac50 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.244676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.244676Z digest=sha256:a7ae3bda1652db7fcfffb30d77b67087f49d4bf04b9a9a550428e5bea0ff1f9c

Observation 48bc6929-6b5d-479b-b24e-28ae7ef193cd · outbound

This paper cites Hsieh, S.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Hsieh, S

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.314289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.314289Z digest=sha256:8839bebca006f1468c30ff67e2d019f820b4a24e41af5c09ead99b30ac9009b1

Observation 0e770906-3b15-4109-8d83-d733cff41d74 · outbound

This paper cites Huang and K.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Huang and K

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.379241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.379241Z digest=sha256:dfafad0543266a1ea54a69070ed662a88ac18ed8e7711b6b3cbea9a39f1db340

Observation eaee9ae1-cf63-48cb-84a5-8bfc7eedc978 · outbound

This paper cites Huang, X.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Huang, X

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.438908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.438908Z digest=sha256:adf05df4046460532bb26ed4c11389d26ca5c0e79cc90a72db7fb547803ccda3

Observation 5bcac5cd-9a07-4a6b-92fc-abaa0f1faeae · outbound

This paper cites A Comprehensive Survey on Evaluating Large Language Model Applications in the Medical Industry.

LIFEBench: Evaluating Length Instruction Following in Large Language Models A Comprehensive Survey on Evaluating Large Language Model Applications in the Medical Industry

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.497560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.497560Z digest=sha256:902480a4fccdce557214121ee104e75bcf6eed1fafaedf897d666fdc570f9748

Observation 9b0ac765-2a37-4dd1-9ad1-3c987da4f752 · outbound

This paper cites The FACTS Grounding Leaderboard: Benchmarking LLMs' Ability to Ground Responses to Long-Form Input.

LIFEBench: Evaluating Length Instruction Following in Large Language Models The FACTS Grounding Leaderboard: Benchmarking LLMs' Ability to Ground Responses to Long-Form Input

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.558589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.558589Z digest=sha256:5d7200db192ea5f5587ed90833304c295d56c4b0d06473216131911f4239316e

Observation 5f1ff18e-cd0c-4fee-8eae-5582c8694fa2 · outbound

This paper cites OpenAI o1 System Card.

LIFEBench: Evaluating Length Instruction Following in Large Language Models OpenAI o1 System Card

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.604933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.604933Z digest=sha256:2fbe9f23d203c7168db6aea48bb3721bdec0f2a23248dd46d8a4ab8cfcc5d2ea

Observation 67340472-1885-4a2d-8bcb-ab51bfa796ae · outbound

This paper cites Jhamtani, V.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Jhamtani, V

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.656826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.656826Z digest=sha256:6652714f691b59d96a866172be4088f706843b3fb5bea9f78963329213ba3185

Observation 4201875f-84b9-4362-abea-85aa730b65fb · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.700834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.700834Z digest=sha256:8f2c0e3a48e6a40fede86bc09e811cb06192eb06cd111ab2799f57cb7f5666d7

Observation f8f43d67-01fe-4215-b405-3477faba6465 · outbound

This paper cites webnovel_cn (revision 745338c), 2023.

LIFEBench: Evaluating Length Instruction Following in Large Language Models webnovel_cn (revision 745338c), 2023

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.773018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.773018Z digest=sha256:f8f5dbd78e0f57beb38544e36b278855570f646872e5454e45cb6fb71d69c6ae

Observation 01265994-1f64-48b2-b6cb-afa6fad5df13 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.828592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.828592Z digest=sha256:d10446ef9ed66cf1180e8bc6201df25c7c26ef4ded6e067733e19ba555b60810

Observation 57ab9db9-b0ff-480c-b1f8-5985cdbab959 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.891231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.891231Z digest=sha256:d5f2fa10a2bd09adac1ed873a84c415a73fb271df2e17a523792eb7d07ef819b

Observation b05c97b9-1322-40df-8b17-02e580e2fcc5 · outbound

This paper cites Koupaee and W.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Koupaee and W

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.954876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.954876Z digest=sha256:47424528fa56f088d55088644cd353dcc20b701b5986d8bf469b3da7db86d91d

Observation 0a4ae45a-d174-4dbe-82ed-da489038619a · outbound

This paper cites Kry´sci´nski, N.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Kry´sci´nski, N

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.023199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.023199Z digest=sha256:7b1b8db4b57152431bff64c2367eeca22f21bf4bd12ea31f9db0a34a77d8816e

Observation 2ce96cb9-3ce3-4822-8fd9-f5373e7b8f16 · outbound

This paper cites Kuratov, A.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Kuratov, A

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.089583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.089583Z digest=sha256:88c144e7a59c911a074a305810d5b97fe0893bcc5cbd9a159d074f79b8b929c5

Observation 06ebd402-d64f-402d-a331-cf01a447796f · outbound

This paper cites Lample, M.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Lample, M

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.137728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.137728Z digest=sha256:3c47e1559701b9b4331884d616345feefa1e8eaefa4ee170fb954c1ce5cce166

Observation f5181dd5-d25f-48fb-83d8-73d647c30973 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.197847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.197847Z digest=sha256:55538ce9b23ce4ff64797a46f7c2fb35f191beb8de1c28bc0615225c6d7dca5b

Observation 49645e82-b652-4441-9933-cf3cd5d6f730 · outbound

This paper cites Long-context LLMs Struggle with Long In-context Learning.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Long-context LLMs Struggle with Long In-context Learning

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.259340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.259340Z digest=sha256:aa79c4abb9118a379131e4c298367746eaec8cc2bb71a41aff3382590fc4c62a

Observation 9555014a-566b-40ee-927f-818b14f6d62b · outbound

This paper cites AI Awareness.

LIFEBench: Evaluating Length Instruction Following in Large Language Models AI Awareness

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.339186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.339186Z digest=sha256:d338f5522dbc4d48d11bd4e3f82514acec0eaa4fb9f62f57e9350e9a397eebec

Observation 24f626ea-adfb-44ea-94c5-37195056c79d · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.453724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.453724Z digest=sha256:cc6203b486c01adcfe144ef168b1c6698485cadbb9d0489e9cc67e3a84932c5d

Observation 4d97b2c8-357a-470f-8ada-2521b0871e72 · outbound

This paper cites Controllable Text Generation for Large Language Models: A Survey.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Controllable Text Generation for Large Language Models: A Survey

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.544649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.544649Z digest=sha256:df1090a148e1bc853cae8e31fe54171b3a45f2a487d3a390b65f1293a6d02bf6

Observation 730f358e-1ce1-4107-bdd1-0a7bb8fc9a1c · outbound

This paper cites Lightman, V.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Lightman, V

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.644559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.644559Z digest=sha256:7e95cf2a8866d5e1681a916dcc90b2e26c82f3cec82cbb06d613e99000c1ee69

Observation bda9e038-2d63-4e11-90c4-3171e0cd5aa6 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.718550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.718550Z digest=sha256:10e0e518842f45a92d2101c81158e1de54947078d66f08670a25f1a0f6c293d4

Observation 2fe3fbff-a7f4-4f7c-a39c-f29a07718948 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.827379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.827379Z digest=sha256:2f11fda2797e3d45b94b2aaea578f31bfb742759c92920777c9d4098380984bd

Observation 9f8e5921-e614-4fe2-9239-c1d4d2eca7e8 · outbound

This paper cites DeepSeek-V3 Technical Report.

LIFEBench: Evaluating Length Instruction Following in Large Language Models DeepSeek-V3 Technical Report

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.916363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.916363Z digest=sha256:76cdf4b6be30ca2cbaf57108c150e94d072f598c255e576b641e6ca59c84ed4d

Observation cd80fcd9-43ab-461a-9333-a5737480603a · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.968830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.968830Z digest=sha256:dfd477712204483970f49a55a4539fc7b1ae8a380a107e6a6e6fb51e2233ba5a

Observation a8ce4bf1-95c6-4c87-aaa5-dc73ae8255f3 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:06.040972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:06.040972Z digest=sha256:77456d5a9bef000b2a4a0ff949807062f5cecf013985fa4ebbf2ea2669e43b8f

Observation f44b67bb-10ee-43a8-ac99-29f04933732a · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 69

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:20.965757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:06.138421Z digest=sha256:713f3eec3f6868de92813abad8212d4aa4abae4f8fd3b5f62d1a2e5d7dd2b88d

Observation fe4da9eb-02ca-4ea7-9620-ebff6fd0f76e · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:06.237974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:06.237974Z digest=sha256:174fc966dda48ec16cda4e9b87f077f6d5c2605e8628b2a5ccbfcae60386953b

Observation 5380913f-d36a-42ef-b2c7-db46a4ade9d3 · outbound

This paper cites ExpertQA: Expert-Curated Questions and Attributed Answers.

LIFEBench: Evaluating Length Instruction Following in Large Language Models ExpertQA: Expert-Curated Questions and Attributed Answers

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:06.311393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:06.311393Z digest=sha256:9829b0dfb4d1b3e0fc3f0d041c0ba05a81992606d334619533b9ff1c09fb3c22

Observation 76fa2654-0be0-4af4-8c2c-89fa5fe9345d · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 72

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:20.823549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:06.397371Z digest=sha256:b0ce4488e61301c2954f441ee35d54349dff0bb71fd73178758e258e7bcaf24c

Observation 69b59cad-1a91-4d25-a19e-dc830b707bc7 · outbound

This paper cites Chinesenlpcorpus.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Chinesenlpcorpus

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:20.743188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:06.513962Z digest=sha256:2878e3aa24cbeeef85fe76ee8829114334b51503ac7c0eb8bac4fe10c1dfc91d

Observation 24b62dbc-691b-4772-9348-0ed39bdcd484 · outbound

This paper cites Mnbvc: Massive never-ending bt vast chinese corpus.https://github.com/esbatmop/MNBVC, 2023.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Mnbvc: Massive never-ending bt vast chinese corpus.https://github.com/esbatmop/MNBVC, 2023

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:20.639920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:06.674885Z digest=sha256:7a2155747f7e7a43778aad76b01c2eb84e0300d0a75c9f5a645c92e80e522bc1

Observation 1173264d-9c54-422a-93b9-da8a23ce010e · outbound

This paper cites Mostafazadeh, N.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Mostafazadeh, N

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:20.537826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:06.786681Z digest=sha256:e8d682c0fa33fcf5608c91050301a62d2640b5db22b48ba5756c12de6c12b265

Observation 85b58e44-50d7-4733-8ca1-173a4139db81 · outbound

This paper cites Nallapati, B.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Nallapati, B

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:20.369749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:06.882313Z digest=sha256:90a1f8dced6f4b2f14f85b915955386906269d934929c322671f26f5543c498c

Observation f6c223a5-af30-4fcf-a99f-2bbb5af790b9 · outbound

This paper cites GPT-4o mini: advancing cost-efficient intelligence.

LIFEBench: Evaluating Length Instruction Following in Large Language Models GPT-4o mini: advancing cost-efficient intelligence

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:20.180851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:06.933239Z digest=sha256:d22f1ebc3811a966b926cd9e0894851aca8641c1cd72ce75272465cacbde156a

Observation 0dd40658-09f2-4836-9104-014091f5039e · outbound

This paper cites Hello GPT-4o.https://openai.com/index/hello-gpt-4o/, 2024.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Hello GPT-4o.https://openai.com/index/hello-gpt-4o/, 2024

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:07.032956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:07.032956Z digest=sha256:f9e6aef8f7b9e79c4bab8bfb156d6c1eb21e9b0c80592e733a328957b7eaca25

Observation c04d3ce7-0fc1-4c56-890c-8de0610a4e57 · outbound

This paper cites OpenAI o1-mini: Advancing cost-efficient reasoning.

LIFEBench: Evaluating Length Instruction Following in Large Language Models OpenAI o1-mini: Advancing cost-efficient reasoning

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:20.018525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:07.136231Z digest=sha256:6caa6afa78848ee180437d9a64f4d3c7e2c03bdd49d9c45a716a5ac70f2e4c4b

Observation a5e6597d-4497-42ea-8a4e-8e43c9b6123b · outbound

This paper cites OpenAI o3-mini: Pushing the frontier of cost-effective reasoning.

LIFEBench: Evaluating Length Instruction Following in Large Language Models OpenAI o3-mini: Pushing the frontier of cost-effective reasoning

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:19.781683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:07.212823Z digest=sha256:658d0bf6575b9b87e9e77f7507a081e729d9b152574f6f5a7146ddaa9726085a

Observation e2b1c336-0ac1-4c52-8c46-6131af0bed4a · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 81

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:19.563974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:07.302010Z digest=sha256:06538331f938523218c77d75dddad9f3b44f359fe8a48d2ccdb458ef7320dec0

Observation 5f1e47e0-eb8d-4b31-9ccb-2af426340b81 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 82

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:19.417697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:07.421996Z digest=sha256:4071a03f59e7107b0197b7d1c911a173845d2e243b060b39111738fb6e583eaa

Observation 84f16566-6a20-4804-866e-4ec8307b499e · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 83

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:19.248048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:07.521225Z digest=sha256:36106f1de66f1f5c75db2f32a7483c87603842598031bfac82e6250ee15f71d5

Observation 5a92f655-6357-4763-aa38-2c13fc4c6f1e · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 84

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:19.112561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:07.583818Z digest=sha256:609caa85e611ab1b2829f0b30eaa80a46b508f7e79bd96c0a8f66a6af4859bae

Observation bb0e9050-1346-4443-b9b1-b22197d61123 · outbound

This paper cites Language Models can Self-Lengthen to Generate Long Texts.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Language Models can Self-Lengthen to Generate Long Texts

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:07.651762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:07.651762Z digest=sha256:f1330ba2b64c368b0cac5d9873aedcde655aa46acff7b407e89c6b0b28213fcd

Observation 27da7725-b122-4352-8b43-795015fecb70 · outbound

This paper cites HelloBench: Evaluating Long Text Generation Capabilities of Large Language Models.

LIFEBench: Evaluating Length Instruction Following in Large Language Models HelloBench: Evaluating Long Text Generation Capabilities of Large Language Models

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:07.698689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:07.698689Z digest=sha256:8fb7aef2fdf9475a8b7576c47fc5d9d9c6f28091dd91652f94a16dd4e50975e3

Observation cfa6af53-cd98-4c87-8776-4bde43e9b3ee · outbound

This paper cites Radford and K.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Radford and K

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:18.997721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:07.792526Z digest=sha256:1160b92b45fdfa38e8a6ea22ea5d308b1603beff152ecf8802adc4386b13bc50

Observation ae6662e7-ee89-48b2-be05-5437d31a5453 · outbound

This paper cites Rafailov, A.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Rafailov, A

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:18.862390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:07.833724Z digest=sha256:fae142e9bcb3297e7545c4716f96d5a4b090db00ad5ef4e81fbccaa6392d573e

Observation 91c869c8-9537-4c21-a830-09f45af784a0 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 89

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:18.746199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:07.872975Z digest=sha256:3075dece36f9c9401b548f6fa1e6dc3252618083392c4a97bf9fd756099ef8fc

Observation dd45e7ff-fbd0-4514-b49d-f0ff703a619c · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 90

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:18.635380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:07.912824Z digest=sha256:7fa9048c3d2abf96bf680ac51efd0eb9e1c92e45fb31b8404e087c00380f8685

Observation 84318624-9748-4bb5-86af-f06cfb3c21e0 · outbound

This paper cites Sennrich, B.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Sennrich, B

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:18.512717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:07.978742Z digest=sha256:2ad7a969f5039a3682b8438542c1e25b0f3bcc42ad797d5014f4a1e5328f307b

Observation 75dad7df-64c9-439f-ad22-0c65def8afee · outbound

This paper cites Shaham, M.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Shaham, M

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:18.362523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:08.027546Z digest=sha256:68519cbe7741a00c031ac03b6dd19622f5c031caa351d3a03b8aa084f456a4f3

Observation bc065065-870c-4664-b9d5-a5dc81738cb4 · outbound

This paper cites Socher, A.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Socher, A

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:18.204207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:08.073224Z digest=sha256:db4812c29481302cc56030fb42ea2bebd06e050bb73e4414526d567c2a0d19fc

Observation d8ffbb68-2cad-46e2-b7a1-28e802e14dff · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 94

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:18.065546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:08.114897Z digest=sha256:59fe15e48657b275cd8787f193a3249e67727216563f5ae644c09f372a7e7e86

Observation c2cd78d9-fbae-46fa-9bd9-170b91d2a54a · outbound

This paper cites Sutskever, O.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Sutskever, O

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:08.152779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:08.152779Z digest=sha256:12f65ac510160ac94285a431952a84df03fe437b5572637c516536137df41d65

Observation d8b7c369-252c-4771-ba80-ece7033421e9 · outbound

This paper cites Talmor, J.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Talmor, J

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:17.933821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:08.199068Z digest=sha256:4344f0f7933e8e50492417602c800eb93a921cc2e6441739088fc508c0068e81

Observation aeea2fa9-9cbd-49b9-8125-dfbc067f22bc · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 97

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:17.705498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:08.252432Z digest=sha256:0943f7938667bd2195ef22f8ef086b6d56cc8bfff3cc253985ceb5898fe92a8d

Observation 4b274721-20f3-4893-9441-212289a8c5cb · outbound

This paper cites CollabStory: Multi-LLM Collaborative Story Generation and Authorship Analysis.

LIFEBench: Evaluating Length Instruction Following in Large Language Models CollabStory: Multi-LLM Collaborative Story Generation and Authorship Analysis

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:08.312320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:08.312320Z digest=sha256:354072b54d7890fe00507930208b02c0b9c1dca4560785e1afb8ef16578c11af

Observation fc9de4cb-0846-4cd2-a62e-5f5ab5bb5382 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 99

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:17.405637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:08.393222Z digest=sha256:554d4b54ff3547151dfe709d155e4716a664ae7466a9898197d4e7b652e4a464

Observation 617e84f1-45a8-486b-bc09-a595734d39da · outbound

This paper cites A Comprehensive Survey in LLM(-Agent) Full Stack Safety: Data, Training and Deployment.

LIFEBench: Evaluating Length Instruction Following in Large Language Models A Comprehensive Survey in LLM(-Agent) Full Stack Safety: Data, Training and Deployment

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:08.453311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:08.453311Z digest=sha256:80cfe47ffb8a3df1b25032196889b60138a80965e83ac8ea2795e2973afce5d8

Pith citing papers

Observation 30905453-4f45-4e4b-b4e7-d3288f6a18c7 · inbound

TiCo: Time-Controllable Spoken Dialogue Model cites this paper.

TiCo: Time-Controllable Spoken Dialogue Model LIFEBench: Evaluating Length Instruction Following in Large Language Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:39:35.854869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T00:38:52.182973Z digest=sha256:1efbe65389bad39a398c64fc60408607e2522e71cabaf709a8c70d809ab2bbcd

Observation eacdae65-3aa9-48a7-84d2-fe4f586cb22c · inbound

Length Value Model: Scalable Value Pretraining for Token-Level Length Modeling cites this paper.

Length Value Model: Scalable Value Pretraining for Token-Level Length Modeling LIFEBench: Evaluating Length Instruction Following in Large Language Models

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:31:26.140879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-07T10:42:27.644514Z digest=sha256:4830ecb6cee28b427d71e2f643c017ffcdfbc6f94f854e965259229a72c54595

Observation 463c613c-2a32-44e3-bfe1-3cbabdb90551 · inbound

Length Value Model: Scalable Value Pretraining for Token-Level Length Modeling cites this paper.

Length Value Model: Scalable Value Pretraining for Token-Level Length Modeling LIFEBench: Evaluating Length Instruction Following in Large Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T15:28:14.401581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T15:28:14.401581Z digest=sha256:f4500d229c48a325cbbf6a056afb2750fa20e4b5fc720e929b0e264101021fda

Observation 9003ed95-f141-480b-8560-443b02ae4816 · inbound

The Librarian Who Refused to Code: Model-Dependent Identity Enactment in LLM Code Generation cites this paper.

The Librarian Who Refused to Code: Model-Dependent Identity Enactment in LLM Code Generation LIFEBench: Evaluating Length Instruction Following in Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T18:03:36.857585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:03:36.857585Z digest=sha256:d200ddeeac17c7da8c1227554afa9c3fc3c6923dc0afcc7ce705d3f64d260047