Pith. sign in

Paper Citation Record · LEDGER

Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2404.12387.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.12387 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T19:25:38.135221Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T10:09:45.386126Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6b7d89b6-f32c-4a5d-a4cc-bf10d1d785b8 · inbound

How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites cites this paper.

How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

Reference 89

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T20:58:59.171594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-12T20:58:58.849040Z digest=sha256:cf58210197c47a3f30551442b796ce58f0d45f41c418531f52bb487a6050776a

Observation e93eb41e-81e8-443e-a612-465ee0902820 · inbound

VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs cites this paper.

VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

Reference 39

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T02:44:53.616921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-11T02:44:53.284345Z digest=sha256:9ec79055ba13eb242e7187203b2d0aa22acbaae329886897f80095004df77dc6

Observation f429b08d-65b1-41f1-9c39-98f38a7d72e6 · inbound

Long Context Transfer from Language to Vision cites this paper.

Long Context Transfer from Language to Vision Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

Reference 57

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:08:36.171513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-12T07:08:35.946669Z digest=sha256:49042afe42272f0303e95e79097b801b20187b0891f901f6d462a9e578b30433

Observation 5b9e54a3-21fd-4971-8b71-00a3a6b2bfea · inbound

mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models cites this paper.

mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

Reference 175

Resolution
verified exact
arxiv_id, observed 2026-05-20T06:20:36.412811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-05-20T06:20:36.235304Z digest=sha256:eac7cf914f39b4e518de2730e6b2726c71ee29141fd4343bbf43d704cecb8a64

Observation 15bb612a-1020-41bf-bbff-89b0aba88ed6 · inbound

Perception Test 2024: Challenge Summary and a Novel Hour-Long VideoQA Benchmark cites this paper.

Perception Test 2024: Challenge Summary and a Novel Hour-Long VideoQA Benchmark Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T05:56:46.424916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T05:56:46.424916Z digest=sha256:60e0645cb7e3949a56dc62a8977faa1ffec1b5fc5308d2ed53c2bb8fab77c6cf

Observation 00c1e656-c5c4-43d1-8c7d-b27b4072fc8b · inbound

AV-Odyssey Bench: Can Your Multimodal LLMs Really Understand Audio-Visual Information? cites this paper.

AV-Odyssey Bench: Can Your Multimodal LLMs Really Understand Audio-Visual Information? Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-11T23:19:11.651540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:19:11.651540Z digest=sha256:00abc5456f24919916db9902010bedb43ed260c347946f79f9cbafc64d9f9e9f

Observation d5aa68bd-f30b-4e65-858d-7977993ceef8 · inbound

From Specific-MLLMs to Omni-MLLMs: A Survey on MLLMs Aligned with Multi-modalities cites this paper.

From Specific-MLLMs to Omni-MLLMs: A Survey on MLLMs Aligned with Multi-modalities Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T14:43:41.818671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:43:41.818671Z digest=sha256:85762e7f83cedcf6061cf52df2a858ed1ebab18a271d1c1dae0b9e6deca8ac03

Observation 9b117915-a4df-4d8e-9b6c-4336376c7e33 · inbound

CL-Attack: Textual Backdoor Attacks via Cross-Lingual Triggers cites this paper.

CL-Attack: Textual Backdoor Attacks via Cross-Lingual Triggers Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T01:03:12.205474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T01:03:12.205474Z digest=sha256:5fa117a3981bc12a4b19c59953944fbcc4f39587c6d855fbf87cc0f75b2f71d2

Observation f7a866fc-febe-41f6-8dce-43297f6a59de · inbound

LLM360 K2: Building a 65B 360-Open-Source Large Language Model from Scratch cites this paper.

LLM360 K2: Building a 65B 360-Open-Source Large Language Model from Scratch Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-10T20:54:02.313497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T20:54:02.313497Z digest=sha256:a423b5616f664072359b112204fec4d2c768e389d036395a6223c24c171fe21f

Observation aa8832c5-e30d-471c-a17d-7504612c1fe3 · inbound

AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment cites this paper.

AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-10T00:02:26.260452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T00:02:26.260452Z digest=sha256:bb1fb2726779aaaad05bff39d108f478f139d8a6ea57daa5e545261d13bb3b8e

Observation c5ae3c7c-3872-4190-835f-b79be9d1e161 · inbound

WorldSense: Evaluating Real-world Omnimodal Understanding for Multimodal LLMs cites this paper.

WorldSense: Evaluating Real-world Omnimodal Understanding for Multimodal LLMs Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:53:26.140705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-17T05:53:26.066674Z digest=sha256:92bb45950fb2c74ac6132bcd1045b10828187f2e420434d1dfb1dfc4882bb209

Observation 72595cfc-cc34-4337-9f94-be24665ba15d · inbound

ZeroBench: An Impossible Visual Benchmark for Contemporary Large Multimodal Models cites this paper.

ZeroBench: An Impossible Visual Benchmark for Contemporary Large Multimodal Models Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T20:55:05.932609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T20:55:05.932609Z digest=sha256:ad084da867aa97459734493df4debd14653bc5a0a4e6e75e869faad96a023697

Observation 7117dadc-598f-44f6-81c8-edfdafcf3256 · inbound

Relational Deep Learning: Challenges, Foundations and Next-Generation Architectures cites this paper.

Relational Deep Learning: Challenges, Foundations and Next-Generation Architectures Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-15T19:25:38.135221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:25:38.135221Z digest=sha256:158c05ebcbbfe1185526501d08478b73707c6242ac196b98a0810c693d07f256

Observation 3bf026fa-d614-4338-bdd4-9ba4f8da7982 · inbound

MMCircuitEval: A Comprehensive Multimodal Circuit-Focused Benchmark for Evaluating LLMs cites this paper.

MMCircuitEval: A Comprehensive Multimodal Circuit-Focused Benchmark for Evaluating LLMs Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T15:51:17.148602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:51:17.148602Z digest=sha256:97264e5ac064b89aa8d545a25faf392184774623ed5870bd1abcb43b11c200e5

Observation e28eb4f0-7cfa-45e4-880c-b2695d6e4ce8 · inbound

GraphThinker: Reinforcing Temporally Grounded Video Reasoning with Event Graph Thinking cites this paper.

GraphThinker: Reinforcing Temporally Grounded Video Reasoning with Event Graph Thinking Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:50:17.299732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-15T20:48:44.933542Z digest=sha256:c1cce99aa86968cbccbffbfb355e24c13dca8aed7a16899893c7e8892f0f4b60

Observation 9459c624-962d-4ac2-9a33-916af85596d5 · inbound

Each Judge Its Own Yardstick: Discovering Per-VLM Taxonomies for Physical Video Evaluation cites this paper.

Each Judge Its Own Yardstick: Discovering Per-VLM Taxonomies for Physical Video Evaluation Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-07-04T10:09:45.387734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-06-26T09:04:23.965554Z digest=sha256:e94f71a97c2d566a4ecbb7ee9f5cfc161ac281238a8dd699ef4cd4c91fbbf918