Pith. sign in

Paper Citation Record · LEDGER

Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following

As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2410.15553.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.15553 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T11:54:19.499279Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-10T06:15:00.866473Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 47c23202-94f2-4fd8-a96f-18cc631a6762 · inbound

Process Reinforcement through Implicit Rewards cites this paper.

Process Reinforcement through Implicit Rewards Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following

Reference 125

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T20:23:31.002560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-11T20:23:30.763794Z digest=sha256:05bfd01b133cd82c90b0cba47d75f4cf9c1b8da8771b3db4ad1e48e184564ee4

Observation 23eff24a-7695-4818-97ec-809f5d827dfb · inbound

Qwen3 Technical Report cites this paper.

Qwen3 Technical Report Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:35:28.517996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T06:35:27.813995Z digest=sha256:53dd5bf090a5c880a1a6487728c281089d817370d6890858c39c0c64d6f2a2db

Observation c9dcc368-e1f4-4179-abc3-6f483870ce25 · inbound

ImgEdit: A Unified Image Editing Dataset and Benchmark cites this paper.

ImgEdit: A Unified Image Editing Dataset and Benchmark Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-12T18:17:45.234132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T18:17:45.123690Z digest=sha256:af560087b0b30bc6dad86d89dd5269896995ec891e226da8975f47249cdcae5a

Observation 34cafd27-05fe-4369-9257-7813e9389d4b · inbound

Qwen3-Omni Technical Report cites this paper.

Qwen3-Omni Technical Report Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:20:37.803106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-11T00:20:37.406351Z digest=sha256:d7bba1ecc69a00cdaccedbe2106ba0cb2d4393f72978525424bcbaf88406def6

Observation e9fa95bc-041b-4cf2-9c9c-3dbc0706a646 · inbound

Qwen3-Omni Technical Report cites this paper.

Qwen3-Omni Technical Report Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T00:20:37.526718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-11T00:20:37.406351Z digest=sha256:1f514e353f2167fde779d9c42fb35ff2266a7da4de6a73002efd428eecff5ff6

Observation 5fa6c9ae-3293-4511-94e0-478007e2e4c7 · inbound

Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory cites this paper.

Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following

Reference 130

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T23:13:16.097093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-14T23:13:15.016486Z digest=sha256:c2a578ef1a0828c57545a58063482d47e0ed7f7be22f6b16e607acb30bc0c572

Observation a1bb53e8-6162-40b9-82f2-661310aa9f41 · inbound

Efficient Evaluation of LLM Performance with Statistical Guarantees cites this paper.

Efficient Evaluation of LLM Performance with Statistical Guarantees Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:00:51.716840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T10:58:40.958435Z digest=sha256:3b071606b93a59ffc18e4596b3fde18453e2945983c0b5d452a57b6bd5370276

Observation c65110b4-8ebf-4fd2-977e-ec4c63dd4b63 · inbound

SAGE: A Service Agent Graph-guided Evaluation Benchmark cites this paper.

SAGE: A Service Agent Graph-guided Evaluation Benchmark Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:21:00.857352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T16:41:23.956104Z digest=sha256:a99cd9cf22eed7789d192c2623dc543be32a8c76a41167b518f732ea29af12e3

Observation 85dfff82-621f-44da-864a-afc68d82b734 · inbound

Alignment has a Fantasia Problem cites this paper.

Alignment has a Fantasia Problem Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following

Reference 68

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T14:31:07.561994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-09T21:37:54.545865Z digest=sha256:c72da4d1700490dfbfa882e669d39dc0b6b34a976c6520f956c564e6a39fccec

Observation 75faee17-ffac-4d3d-aa77-d67d838e6eb6 · inbound

SEIF: Self-Evolving Reinforcement Learning for Instruction Following cites this paper.

SEIF: Self-Evolving Reinforcement Learning for Instruction Following Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:20:55.925346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-11T01:51:09.514927Z digest=sha256:45ea777fe4cdc9926ab42ba4e17a588bf0f53c3f1c8a9a466829018e4648376a

Observation 4ebe96a1-4ec0-4f61-b5cb-d078b3596d3b · inbound

Benchmarking EngGPT2-16B-A3B against Comparable Italian and International Open-source LLMs cites this paper.

Benchmarking EngGPT2-16B-A3B against Comparable Italian and International Open-source LLMs Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following

Reference 70

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:40:53.469699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-11T03:40:04.692279Z digest=sha256:9b736408e0c125a698eb6ec6581a0696bfe730ce22a108b602b8445da41279e2

Observation 4aa78386-0c4f-48fb-b006-cefa5d0a4cfd · inbound

Benchmarking EngGPT2-16B-A3B against Comparable Italian and International Open-source LLMs cites this paper.

Benchmarking EngGPT2-16B-A3B against Comparable Italian and International Open-source LLMs Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following

Reference 70

Resolution
verified exact
arxiv_id, observed 2026-05-21T08:19:52.895325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T08:14:55.858466Z digest=sha256:53e8eb431ffa5142b4be0448a5cd12d49459ac7ebf03ebd15227710dad3889ed

Observation c9639bb0-a7b6-4a1d-86e8-14e563ea8a0a · inbound

When Attention Closes: How LLMs Lose the Thread in Multi-Turn Interaction cites this paper.

When Attention Closes: How LLMs Lose the Thread in Multi-Turn Interaction Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T20:22:55.474903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-14T20:20:27.212052Z digest=sha256:ec0fe5ca41677c268bab12ef5cfc9184b611ac571b993a1e1ba7862cfe6226b2

Observation 9f0cbc76-5b08-40ae-b50e-c43bdbde0bf2 · inbound

Hy-MT2: A Family of Fast, Efficient and Powerful Multilingual Translation Models in the Wild cites this paper.

Hy-MT2: A Family of Fast, Efficient and Powerful Multilingual Translation Models in the Wild Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:04:41.500899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-22T07:03:00.655352Z digest=sha256:da24e25171619d9bc5cf889c0c771f5e8b1f2fd624348a6b931d835f22725f84

Observation 53af5766-6f74-4442-aa83-e7be4c537bf6 · inbound

Hy-MT2: A Family of Fast, Efficient and Powerful Multilingual Translation Models in the Wild cites this paper.

Hy-MT2: A Family of Fast, Efficient and Powerful Multilingual Translation Models in the Wild Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-30T17:44:57.876165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-30T17:38:21.754380Z digest=sha256:5b4a437f460f4202dd2f6fe6ba7930579df1529a1650936a7893d5ed81d16e46

Observation a4b14268-9d3e-4385-aa8a-b0213eac01bf · inbound

IFMTBench: A Comprehensive Benchmark for Multilingual Translation Instruction Following cites this paper.

IFMTBench: A Comprehensive Benchmark for Multilingual Translation Instruction Following Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T12:33:24.273265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-29T12:32:00.055500Z digest=sha256:99816a9233fbf289efaf34a01e0e62f41006621981d9f7dac22f007ddb8b0023

Observation 9ead1e01-90ec-43a3-a4b2-4732219508ef · inbound

Ling and Ring 2.6 Technical Report: Efficient and Instant Agentic Intelligence at Trillion-Parameter Scale cites this paper.

Ling and Ring 2.6 Technical Report: Efficient and Instant Agentic Intelligence at Trillion-Parameter Scale Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following

Reference 149

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T22:17:25.642579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-07-02T22:10:59.568675Z digest=sha256:1199ff82c68c1b5ae055259fd9fcbc27689670d096216772f932f1a3fe6b8be9

Observation cab3315e-dfe5-43a9-830d-37610feac956 · inbound

In-Place Tokenizer Expansion for Pre-trained LLMs cites this paper.

In-Place Tokenizer Expansion for Pre-trained LLMs Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T23:50:15.426936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:50:15.426936Z digest=sha256:c41d019ba7e2d0c25d96a8a35be5e8ef99f5d21c9bff1ce1569089ebc88e9133

Observation e8d3030b-e021-4b7e-a711-47af8e9d0fd1 · inbound

HANDBOOK.md: A Benchmark for Long-Context Agentic Instruction Following cites this paper.

HANDBOOK.md: A Benchmark for Long-Context Agentic Instruction Following Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T02:36:07.636136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:36:07.636136Z digest=sha256:b19e1b780c20165b599178900e7d5a60872e6dacab625a5d930aa9961a568995

Observation 3ce27f70-24a4-4c64-89b2-5b147d3a0374 · inbound

Hy-MultiTurn: A Six-Dimensional Benchmark for Deep Multi-Turn Dialogue Understanding cites this paper.

Hy-MultiTurn: A Six-Dimensional Benchmark for Deep Multi-Turn Dialogue Understanding Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-03T11:54:19.499279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T11:54:19.499279Z digest=sha256:377bac16889a79d39ab5fe6e2446a23e5a0d4e142e534d1e12d717b98ad4753e