Pith. sign in

Paper Citation Record · LEDGER

EGM: Efficient Visual Grounding Language Models

As of 4 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 1 inbound Pith citation observation for arXiv:2601.13633.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2601.13633 v3

Coverage vector

measured 59 of 59 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-16T13:07:55.655699Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T23:38:51.488742Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

59 of 59 outbound references displayed

  • verified exact6
  • verified fuzzy29
  • unresolved10
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch13

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a6173652-6af4-4e58-9664-5d89c0a99f58 · outbound

This paper cites GPT-4 Technical Report.

EGM: Efficient Visual Grounding Language Models GPT-4 Technical Report

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T13:10:56.734420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:c5f6c2314a35bdeb0490bffcdf90a73ebe07b0089b9a748ebafbe3a3bb9f752e

Observation 17d41e28-cfab-479a-b096-2881c783bb00 · outbound

This paper cites an unresolved cited work.

EGM: Efficient Visual Grounding Language Models Unresolved cited work

Reference 2

Resolution
parse uncertain
raw_fallback, observed 2026-05-16T13:10:57.606712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:d1da236d788db4f69a3aedac208f29752d65e209d3a03ba337831316299c9cd8

Observation e3fee928-080c-4cc0-a3fa-6803df9e2f3c · outbound

This paper cites an unresolved cited work.

EGM: Efficient Visual Grounding Language Models Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-05-16T13:10:57.571737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:e39eabd165a2c7500d5dd06acf2838d4a156f00db25f993a5ca2b80e75ab1f33

Observation 5202c00d-aabc-4b9c-8091-f04881ccb8cc · outbound

This paper cites Qwen2.5-VL Technical Report.

EGM: Efficient Visual Grounding Language Models Qwen2.5-VL Technical Report

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-05-16T13:10:56.749395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:d6a09f012e5761748d50a8d7fbe5510d8820d9931995c86ef89bc37b9b7282d8

Observation bdb57217-bae4-4d4d-a551-d3b7f0b7a0f6 · outbound

This paper cites an unresolved cited work.

EGM: Efficient Visual Grounding Language Models Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-05-16T13:10:57.601763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:372080fce57bad0e25130ea15657da90ccfd023930a839f90311b454b808363e

Observation 9374484c-766c-47a3-8528-315e8d0bfa21 · outbound

This paper cites In: Proceedings of the Computer Vision and Pattern Recognition Conference.

EGM: Efficient Visual Grounding Language Models In: Proceedings of the Computer Vision and Pattern Recognition Conference

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.583624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:d3f8bf4a95de14a0b0dfa16fa68dc8ec675e3b3c77373a9814ddcc6a17b3684a

Observation f557b503-347e-4997-8c4d-7c825f3316bf · outbound

This paper cites In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

EGM: Efficient Visual Grounding Language Models In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.588840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:19d690c37bd7fa837b784c2572d3a4da7c61a50ffc4c1c24bc0d8dc87f72dc44

Observation 73c063f4-5930-464d-98eb-6750b130d49f · outbound

This paper cites Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities.

EGM: Efficient Visual Grounding Language Models Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T13:10:56.731528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:16b7e322749c5ea8ee8982211294c9c25b8e08db95d020250ea7b1b693643aba

Observation 9145da1d-2f48-4990-93b3-964499d31458 · outbound

This paper cites arXiv e-prints pp.

EGM: Efficient Visual Grounding Language Models arXiv e-prints pp

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.616917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:9d0d0a664354e789566f197ce8f913658ad95c10919497c909ed98ae7b9fef64

Observation e871a1c0-4d5f-4179-afeb-a96f3ce206c4 · outbound

This paper cites arXiv e-prints pp.

EGM: Efficient Visual Grounding Language Models arXiv e-prints pp

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.596770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:eda7d81afd11235e612c1c926818aa28f2903632579d52b951ebf53aff9c5998

Observation 2033dc5d-a4ba-41ff-8e0d-01fa538263d6 · outbound

This paper cites Proceedings of the Ad- vances in Neural Information Processing Systems (NeurIPS) (2024).

EGM: Efficient Visual Grounding Language Models Proceedings of the Ad- vances in Neural Information Processing Systems (NeurIPS) (2024)

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.568872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:e77a360f952e3b373cd15f3c95499d92c4da5c015b6cefd27bb548ed705ece6f

Observation ce7bc1b4-cd7e-4b0f-9a93-7db533d80b9a · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

EGM: Efficient Visual Grounding Language Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 12

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T13:10:56.719054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:9d28850892e581dae07f7dfec1d4365f5c3def1d5d0efa1a1cbc035e50376970

Observation ab5e14a9-a4d5-4f02-aad9-42cfe5919bcf · outbound

This paper cites TAO-Amodal: A Benchmark for Tracking Any Object Amodally.

EGM: Efficient Visual Grounding Language Models TAO-Amodal: A Benchmark for Tracking Any Object Amodally

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-16T13:10:56.714909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:7a29b4516cd398f183481a3198793a51ff1b205ad550ffcfd0def67d9f2aeb31

Observation ea035e0a-1f92-4a7d-98e4-9ff029116b21 · outbound

This paper cites GPT-4o System Card.

EGM: Efficient Visual Grounding Language Models GPT-4o System Card

Reference 14

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T13:10:56.725695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:280bb1833f54604c1d693be41138af298d31f1e0fd2ee67a7e5340aab46a2673

Observation 3aba5604-d2d8-4a51-a8a3-0161ebcdfe5e · outbound

This paper cites Psychological Research88(2), 307–337 (2024) EGM 17.

EGM: Efficient Visual Grounding Language Models Psychological Research88(2), 307–337 (2024) EGM 17

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.614041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:6d7eb58dff2bc1d50b12aff4f01b78626832e319727fdeb9d584805c715565c8

Observation 4b004b11-4843-495d-bac8-1d1f29523eea · outbound

This paper cites In: Proceedings of the 2014 conference onempiricalmethodsinnaturallanguageprocessing(EMNLP).pp.787–798(2014).

EGM: Efficient Visual Grounding Language Models In: Proceedings of the 2014 conference onempiricalmethodsinnaturallanguageprocessing(EMNLP).pp.787–798(2014)

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.604327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:0aec32a2670ecd229c273a4d13020bd7c2f990493b4a27a8a1c5cc36c41c3127

Observation 27bf4157-3794-4325-9f68-d34c245c1a6a · outbound

This paper cites an unresolved cited work.

EGM: Efficient Visual Grounding Language Models Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-05-16T13:10:57.610286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:e78a19562de00d039ebb706ebea26536fc5d8079714d48aa3b81cc850116b7e2

Observation 06dab8fa-903d-4831-8a6d-40dd5ef4b1a7 · outbound

This paper cites In: Proceedings of the ACM SIGOPS 29th Symposium on Operating Systems Principles (2023).

EGM: Efficient Visual Grounding Language Models In: Proceedings of the ACM SIGOPS 29th Symposium on Operating Systems Principles (2023)

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.611850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:c1df588bd67466a2beb5be6a1e7df32295fdc58f0303bb4eb2f94dd253ecffb2

Observation 69071cf1-ea2a-41f0-b636-c806130a04da · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

EGM: Efficient Visual Grounding Language Models LLaVA-OneVision: Easy Visual Task Transfer

Reference 19

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T13:10:56.737621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:9060f833dbaadb44895ad590575e98bfbda5ae3dfc21bb4c7d4a1d8f9eac1ced

Observation a0e41b03-23e4-44f4-9025-879840fce802 · outbound

This paper cites In: Proceedings of the European Conference on Computer Vision (ECCV).

EGM: Efficient Visual Grounding Language Models In: Proceedings of the European Conference on Computer Vision (ECCV)

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.619094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:f67ca34bd792b8b72ff26257641eabea2b8f097b41801d7f99a44673f90a6ef1

Observation a42441e3-423f-4393-8f34-56d2fa134908 · outbound

This paper cites In: Proceedings of the IEEE/CVF Interna- tional Conference on Computer Vision.

EGM: Efficient Visual Grounding Language Models In: Proceedings of the IEEE/CVF Interna- tional Conference on Computer Vision

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.621580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:a6a5f4cf507e9b2fc925ef9d81935d03453919936c841e1cf705d3fafa09fdd7

Observation cd6dc917-a95e-438b-a24c-b981569f3c02 · outbound

This paper cites In: Proceedings of the European Conference on Computer Vision (ECCV).

EGM: Efficient Visual Grounding Language Models In: Proceedings of the European Conference on Computer Vision (ECCV)

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.591576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:1e204d71209a4238a65c253d898f14d68f774c33f4223f1dc0617d8c71c63263

Observation fff80151-937a-4af2-bd96-e9fe936685d6 · outbound

This paper cites IEEE Transactions on Multimedia (MM) (2023).

EGM: Efficient Visual Grounding Language Models IEEE Transactions on Multimedia (MM) (2023)

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.623745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:c1babc492aa4c652c2bc2e49047233ad39ac64ca6b80891fb0562d3eb1ec7552

Observation 1d09fc3b-caed-4dd3-8d68-f08324997b17 · outbound

This paper cites In: Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV).

EGM: Efficient Visual Grounding Language Models In: Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV)

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.625998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:8b5f7d40d5117562a3c34f99c4257fd3cdf0bd224c9577c37d4942d091f294da

Observation 8e0eb2c0-3a9e-48f4-b9aa-3da62be6e593 · outbound

This paper cites In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

EGM: Efficient Visual Grounding Language Models In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.586577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:4794b893e5d02bc28b5e2f1f013c84c7c5eceaa387df7e11123a96c43c1dd550

Observation 16d83e56-4db8-439d-99a5-6d4e170db180 · outbound

This paper cites In: Proceedings of the Computer Vision and Pattern Recognition Conference.

EGM: Efficient Visual Grounding Language Models In: Proceedings of the Computer Vision and Pattern Recognition Conference

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.587743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:87dafc089194db404ad5a609926937676870dfd34b52606bbd77ce4c0513c99f

Observation db26aa49-bdb0-4431-a281-a926174af312 · outbound

This paper cites In: Proceedings of the Computer Vision and Pattern Recognition Conference.

EGM: Efficient Visual Grounding Language Models In: Proceedings of the Computer Vision and Pattern Recognition Conference

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.593569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:f5940320b7b0f338c87543bc3496a705bc1480e0c164df938be3e3960126223f

Observation ed0eae98-da38-42de-be17-6a7d87ad7f8d · outbound

This paper cites In: Proceedings of the IEEE/CVF International Conference on Computer Vision.

EGM: Efficient Visual Grounding Language Models In: Proceedings of the IEEE/CVF International Conference on Computer Vision

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.609352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:5d09f404dc893aa5d662f59744c099190d0ff625830bf806aca5b72033bb0f97

Observation 20af6d65-1e89-472a-acd5-e58e16de1f38 · outbound

This paper cites In: Proceedings of the IEEE conference on computer vision and pattern recognition.

EGM: Efficient Visual Grounding Language Models In: Proceedings of the IEEE conference on computer vision and pattern recognition

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.591288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:c7bc73cd588d8d99e99366dd0bfebd7f4f4ac73565843b39374e9ba7d8c5700d

Observation 788a646b-dc6b-4146-8462-0d85452d4fd6 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

EGM: Efficient Visual Grounding Language Models DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-05-16T13:10:56.706873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:10fec4386f9ebb48cd6326722ca60d3061e610cef536e5380b89a76755febf33

Observation b8b534b6-04cd-42bf-b771-754bf3d6eb09 · outbound

This paper cites HybridFlow: A Flexible and Efficient RLHF Framework.

EGM: Efficient Visual Grounding Language Models HybridFlow: A Flexible and Efficient RLHF Framework

Reference 31

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T13:10:56.740725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:911adc1e536cc0b021f3ebaf34449ed73ebd109120c87af7f57c7c4ddd02bf61

Observation a39eb763-a5eb-4593-9e5d-205bccb1c8a0 · outbound

This paper cites Gtpo and grpo-s: Token and sequence-level reward shaping with policy entropy.arXiv preprint arXiv:2508.04349.

EGM: Efficient Visual Grounding Language Models Gtpo and grpo-s: Token and sequence-level reward shaping with policy entropy.arXiv preprint arXiv:2508.04349

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-16T13:10:56.722822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:c257e3ddadb01e2a6e0a41e291e63fa375648f6bf01a2eb8addca84137174bf3

Observation 80558958-e0dc-4f5d-848f-b8ac3e96cea6 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

EGM: Efficient Visual Grounding Language Models Gemini: A Family of Highly Capable Multimodal Models

Reference 33

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T13:10:56.728483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:652a77dd463e6fa5b79700e1732235a6d4991009a37e5111474bb0c30553a9ef

Observation 650fd6b2-170b-4c43-ab65-78ef70da1454 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

EGM: Efficient Visual Grounding Language Models Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 34

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T13:10:56.746512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:bcf55f5689389aba6a6f905da9a4a796a1f5fa541bfcec8734166848ffeffd6d

Observation 803bb090-7124-412a-ab9f-4a258a95aebc · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

EGM: Efficient Visual Grounding Language Models LLaMA: Open and Efficient Foundation Language Models

Reference 35

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T13:10:56.743529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:fee6be13f34f76bd6d5fb39781ac2db0445831ea6086dab76b42211c8fe12991

Observation 6c83cf21-5767-4c75-b8ee-c9aae01b99e8 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

EGM: Efficient Visual Grounding Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 36

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T13:10:56.702965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:e5e7dc179a20badcd931b3459324fcdac35eab4d8c497b477911ef85ad238583

Observation 6a1c0a8c-c229-4107-baab-02f1846fb20c · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

EGM: Efficient Visual Grounding Language Models Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 37

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T13:10:56.689386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:07ddbdf4e8d783665c5679b7057f446fa3e97596a5149220d996e8a2d0d2fa14

Observation 0af3fd9f-14c7-4f89-8c2d-17653cc9f45c · outbound

This paper cites InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency.

EGM: Efficient Visual Grounding Language Models InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-05-16T13:10:56.692513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:805cb53af41c8032e44236cfa5e2ddd0770382cc60d1efc1af43536c9bb4523a

Observation 83fe5f8f-d0e0-4606-8235-a429a6b25ece · outbound

This paper cites Amodal3R: Amodal 3D Reconstruction from Occluded 2D Images.

EGM: Efficient Visual Grounding Language Models Amodal3R: Amodal 3D Reconstruction from Occluded 2D Images

Reference 39

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T13:10:56.707028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:fc1872c0d644a9229233afb7257ae03a8ba2e8c2bb40f5382d21559d22af80dd

Observation 8b77253e-ecc1-4f49-b04e-4492587c300c · outbound

This paper cites International Journal of Computer Vision (IJCV) (2025).

EGM: Efficient Visual Grounding Language Models International Journal of Computer Vision (IJCV) (2025)

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.602387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:b3494d669c708599ef23c72aeaaea190ec8dda4b390ca092bdc64b88d5fb1f2b

Observation 2ff12b36-d25e-415f-ab4b-1d8e48e5c6ad · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2024).

EGM: Efficient Visual Grounding Language Models Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2024)

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.571957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:98b3100f23fa1eedccaaa8d6244a53c1bf683f9afa36c1c667a1b165cd31026e

Observation 50020c9f-6b63-4e53-82b8-2848c2951df9 · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

EGM: Efficient Visual Grounding Language Models DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-05-16T13:10:56.699161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:9ee9f4b51f6f18b93ea562be0d566d7a8cdcce15e74f4965c2eecb1d7af3d550

Observation 37434e43-ec97-4df6-bc82-056928281abe · outbound

This paper cites Proceedings of the IEEE International Con- ference on Content-Based Multimedia Indexing (CBMI) (2025).

EGM: Efficient Visual Grounding Language Models Proceedings of the IEEE International Con- ference on Content-Based Multimedia Indexing (CBMI) (2025)

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.574891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:ab46ce5db86dd69f04dcf2830f82e633eb585a592d34287f9c6c50b60a910284

Observation f473c917-b39a-41e4-b6e5-69eecd34e16f · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2024) EGM 19.

EGM: Efficient Visual Grounding Language Models Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2024) EGM 19

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.577825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:c7ca85b4d6626cd3047ab483083127b0fc57e8a1a4092bb94c4902275c4d1e93

Observation 7ccb2ba9-ea63-445c-9ee6-61086bb3cab9 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR).

EGM: Efficient Visual Grounding Language Models In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.566424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:020ec83b051d1907475ff5ad2068cdc3c1550dc036ed63436800de237cbbc3b5

Observation baba472e-15bf-4990-85a9-1516a91fb7fa · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

EGM: Efficient Visual Grounding Language Models InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 46

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T13:10:56.685410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:efc8f1febda6464b96db6e67e5cefe2f75294cdb313a842ee45ddc1b6fb9a2a9

Observation cd3f772e-aa3d-4404-a03d-94a822547520 · outbound

This paper cites {question}.

EGM: Efficient Visual Grounding Language Models {question}

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.637127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:74fe644f39f6785143bf92370a91f7e406e64e7e8c5337da286124d7db1a621a

Observation 6547ba3f-a7f1-41d0-a6a7-0e4923cb688d · outbound

This paper cites an unresolved cited work.

EGM: Efficient Visual Grounding Language Models Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-05-16T13:10:57.639128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:6e3357b4b4e2980ebe02cafe8be548aa1a8d5b1ff887ec59bcfc38941e9fc608

Observation 8561e2fc-8969-4e6b-8df1-291603d19f4b · outbound

This paper cites the second/third/fourth xxx from left/right/top/bottom.

EGM: Efficient Visual Grounding Language Models the second/third/fourth xxx from left/right/top/bottom

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.563719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:58e0efeccfaee977411053d74a4715f0d49feb22db2b94590267f534600c02bc

Observation 1a4f842c-8da4-4839-858b-6b76bc1b02dd · outbound

This paper cites an unresolved cited work.

EGM: Efficient Visual Grounding Language Models Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-05-16T13:10:57.628153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:75cf4d59e0a7d86d69df7832140df2e02e644565d0356ae252a45e241c922f08

Observation 9b543d76-e20b-4fb0-ab13-39313e2e9c80 · outbound

This paper cites the man in yellow coat.

EGM: Efficient Visual Grounding Language Models the man in yellow coat

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.630277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:02871c3ffcf6a0521f57e32d548ba8342aec37bb581033f6a53c3545e4950adc

Observation 55f4ddcb-abfc-4e38-8cc2-385938014c2b · outbound

This paper cites an unresolved cited work.

EGM: Efficient Visual Grounding Language Models Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-05-16T13:10:57.632503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:f3ae131550fa6583931e361373c1c2386bc3f9b1c5ff44aa41f13ee7da7d6e8e

Observation bc65e258-d994-48fa-9b7f-79a6ae81e0da · outbound

This paper cites an unresolved cited work.

EGM: Efficient Visual Grounding Language Models Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-05-16T13:10:57.607436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:acbf878cf68fccd17aa4f7f1895b519c363c7b7cdb231cc830e702efab048e58

Observation dfd2ddf3-682f-4c5c-acea-c1095bbb4296 · outbound

This paper cites YES" if the description uniquely and accurately identifies the TARGET region -.

EGM: Efficient Visual Grounding Language Models YES" if the description uniquely and accurately identifies the TARGET region -

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.634900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:c95574be472b2b16e5aa883d5ff03efc827875c5e45089a0c2bde8ae5cf6eeb0

Observation d165cc06-722f-4a60-9f3f-a43aa7955017 · outbound

This paper cites an unresolved cited work.

EGM: Efficient Visual Grounding Language Models Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-05-16T13:10:57.550405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:e73654556d85a8bdb67422b6f5b6eea0f7d25cd8d22bbfcb90b0a1f69fc7c972

Observation eee4c9f9-958f-4392-a0cf-4f7d002a1a8d · outbound

This paper cites an unresolved cited work.

EGM: Efficient Visual Grounding Language Models Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-05-16T13:10:57.552967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:b16e96e4ba28075e30f05a2e2503fef222b0a55da283bba626325507c01a1bdd

Observation 57899d1d-061d-4301-93af-3c2633546d67 · outbound

This paper cites an unresolved cited work.

EGM: Efficient Visual Grounding Language Models Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-05-16T13:10:57.540048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:e9971a2042350bd5cb809a1366e733099b0e42deb17691d479ee4cbff3830d09

Observation 2021a286-609b-48c4-8ad6-d0ee04cf3821 · outbound

This paper cites slightly.

EGM: Efficient Visual Grounding Language Models slightly

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.542633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:e541a0def00dc6e896bdca99da7d39c4690187672e648f9c2679fea114015cb8

Observation 4c51ddc9-50be-4398-a342-23277cb24cd0 · outbound

This paper cites sofa against the wall.

EGM: Efficient Visual Grounding Language Models sofa against the wall

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.537439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:60b1ef8894534725622ba4a360f875c737bd5d82662a509b7123eb0af20b0454

Pith citing papers

Observation 69f25467-88b4-4b13-bac1-282fa739d11b · inbound

Reasoning-Guided Part-Level Visual Grounding via Reinforcement Learning cites this paper.

Reasoning-Guided Part-Level Visual Grounding via Reinforcement Learning EGM: Efficient Visual Grounding Language Models

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-01T23:38:51.488742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:38:51.488742Z digest=sha256:282118e1672abe54358d254bf2c3cab271ca4aac5e9bccc1aece1fb9b8067800