Pith. sign in

Paper Citation Record · LEDGER

Scalable Object Detection in the Car Interior With Vision Foundation Models

As of 21 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 0 inbound Pith citation observations for arXiv:2508.19651.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.19651 v3

Coverage vector

measured 23 of 23 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-18T21:07:31.589011Z

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

23 of 23 outbound references displayed

  • verified exact5
  • verified fuzzy17
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3a101111-01ce-4b6e-8f84-41df0e57e8d4 · outbound

This paper cites Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks.

Scalable Object Detection in the Car Interior With Vision Foundation Models Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T21:11:51.866349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T21:07:31.589011Z digest=sha256:8743f09b9c2228631fb0ebd183c545e7508091824d4261174d7551dff490a184

Observation c448229a-f97e-48b1-9559-b050a1f9f0b5 · outbound

This paper cites You Only Look Once: Unified, Real-Time Object Detection.

Scalable Object Detection in the Car Interior With Vision Foundation Models You Only Look Once: Unified, Real-Time Object Detection

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T21:11:51.906815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T21:07:31.589011Z digest=sha256:d0feb550c0e48bc09bece9174e86272ecc8d91b91ec5f1cfb22490d2e06d74f3

Observation 6eb0b139-8d96-4479-b385-25062cb04170 · outbound

This paper cites End-to-End Object Detection with Transformers.

Scalable Object Detection in the Car Interior With Vision Foundation Models End-to-End Object Detection with Transformers

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T21:11:51.887673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T21:07:31.589011Z digest=sha256:c5e2c23b9704acb7e92e93ac55cd1c6689262ceff2f6a7646ae505c28127bfff

Observation d650fc28-86b1-49dc-9534-b7df95b75b9d · outbound

This paper cites FCOS: Fully Convolutional One-Stage Object Detection.

Scalable Object Detection in the Car Interior With Vision Foundation Models FCOS: Fully Convolutional One-Stage Object Detection

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T21:11:51.910094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T21:07:31.589011Z digest=sha256:f7cc772e6714db87fb919f1b06a7d792e8f46d0c713fc5fa2e6f53c99cf3f6a6

Observation 0d13a0c3-9f3b-4a85-86e7-f83f41fb2e43 · outbound

This paper cites Object Detection with Deep Learning: A Review.

Scalable Object Detection in the Car Interior With Vision Foundation Models Object Detection with Deep Learning: A Review

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T21:11:51.896925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T21:07:31.589011Z digest=sha256:4e327198baeeecde8173fdb8f432dfdc931478d3b72f10307ef89a05c88eb1bb

Observation 748e12ce-db6d-4f28-bc71-64e4efd88072 · outbound

This paper cites Learning Transferable Visual Models From Natural Language Supervision.

Scalable Object Detection in the Car Interior With Vision Foundation Models Learning Transferable Visual Models From Natural Language Supervision

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T21:11:51.899979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T21:07:31.589011Z digest=sha256:e6ec53b1562530a1801f54a9382ca02fc7759b95c56f34ba98a58b090e2c7943

Observation 62b52d54-8266-4c37-882a-debd25975f88 · outbound

This paper cites On the Opportunities and Risks of Foundation Models.

Scalable Object Detection in the Car Interior With Vision Foundation Models On the Opportunities and Risks of Foundation Models

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-18T21:11:50.836611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T21:07:31.589011Z digest=sha256:588f1950be0a3f92b5daafa9c263c763698efe48a506cd6fe4b217ed49a7e37f

Observation 7124073d-5987-42cc-ba9d-583d8517ab8d · outbound

This paper cites Flamingo: a Visual Language Model for Few- Shot Learning.

Scalable Object Detection in the Car Interior With Vision Foundation Models Flamingo: a Visual Language Model for Few- Shot Learning

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T21:11:51.894044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T21:07:31.589011Z digest=sha256:c01b48bd398b0c9c5ae0645d86dedd98f6f8fc858d14e8f4fe86315205bec635

Observation 78493b22-4cb1-4c8b-9bb0-ab0ab2ea3747 · outbound

This paper cites InstructBLIP: Towards General-Purpose Vision- Language Models with Instruction Tuning.

Scalable Object Detection in the Car Interior With Vision Foundation Models InstructBLIP: Towards General-Purpose Vision- Language Models with Instruction Tuning

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T21:11:51.916751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T21:07:31.589011Z digest=sha256:c92c52f94a73443db0413ea5073e39dd3dcf55ae49800a0471d2ce2e2ae155a5

Observation ba10b948-499c-4d21-9975-910fa5c85a0a · outbound

This paper cites CenterNet: Keypoint Triplets for Object Detection.

Scalable Object Detection in the Car Interior With Vision Foundation Models CenterNet: Keypoint Triplets for Object Detection

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T21:11:51.903449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T21:07:31.589011Z digest=sha256:2c9c588722acbf1195bbbe3f276d480e00a92f826d725b87fabf2b9ccf103237

Observation 4993291d-4ae3-4ca6-9c4b-9a12a837c6c1 · outbound

This paper cites Visual Instruction Tuning.

Scalable Object Detection in the Car Interior With Vision Foundation Models Visual Instruction Tuning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T21:11:51.913329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T21:07:31.589011Z digest=sha256:b4509fa011693ec30dae15d35f6f6f27999113d8a24fd390e8790707cb8cbe5a

Observation 9e1a16c8-b164-42a0-8ffb-ed3912fca433 · outbound

This paper cites Improved Baselines with Visual Instruction Tuning.

Scalable Object Detection in the Car Interior With Vision Foundation Models Improved Baselines with Visual Instruction Tuning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T21:11:51.890958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T21:07:31.589011Z digest=sha256:0094bf7bf566a1bad78415a5ddec503129e1222fe97658039e5526585849bef9

Observation 414963db-4e4a-45cb-8f6a-ef2ad93dad8d · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Scalable Object Detection in the Car Interior With Vision Foundation Models LoRA: Low-Rank Adaptation of Large Language Models

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-05-18T21:11:50.848073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T21:07:31.589011Z digest=sha256:355490b470e993cc90d1f7843f5df223bd0e9e86f90f6219e580bd6931230c51

Observation d38b0cd0-40aa-4a9d-a794-0a3d11f0e51d · outbound

This paper cites Scaling Vision Transformers to 22 Billion Parameters.

Scalable Object Detection in the Car Interior With Vision Foundation Models Scaling Vision Transformers to 22 Billion Parameters

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T21:11:50.831055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T21:07:31.589011Z digest=sha256:e3b363eb679960a4c70dd53b9db630f506a3f1bd4c6391e33ef0782b6215d4fb

Observation 6a048251-cdec-4df6-b4f5-0069e62057b0 · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

Scalable Object Detection in the Car Interior With Vision Foundation Models Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-05-18T21:11:50.842683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T21:07:31.589011Z digest=sha256:61dfa6cfe31c2baba00400d3f3dacada2d68a9e1d8c97d2e1ff625a0992d32a8

Observation bfd24fad-a06c-4300-bc09-912f8d7fa934 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Scalable Object Detection in the Car Interior With Vision Foundation Models LLaMA: Open and Efficient Foundation Language Models

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-18T21:11:50.808632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T21:07:31.589011Z digest=sha256:aaaf2281a16e007be9b2dd43cee27714bce169f9eb0025f70e7eea6004632ecd

Observation f7341699-e70d-4d5d-b462-b1e639527892 · outbound

This paper cites Vicuna: An Open-Source Chatbot Impressing GPT-4 with 90%* ChatGPT Quality.

Scalable Object Detection in the Car Interior With Vision Foundation Models Vicuna: An Open-Source Chatbot Impressing GPT-4 with 90%* ChatGPT Quality

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T21:11:51.846423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T21:07:31.589011Z digest=sha256:0c956a100254279df5f4e42672b5e52c06ec8016316ce246b521e69e6315030d

Observation a9238ccd-9a82-4a9b-8490-ca1f1ffd0c12 · outbound

This paper cites Learning Transferable Visual Models From Natural Language Super- vision.

Scalable Object Detection in the Car Interior With Vision Foundation Models Learning Transferable Visual Models From Natural Language Super- vision

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T21:11:51.849972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T21:07:31.589011Z digest=sha256:f18ae0eb79aab89b02b4826fff1d51538db2671b5c0511766bef3e06073b4c37

Observation 8c5f44b1-98fc-4f01-91cc-f4eb9dad8e37 · outbound

This paper cites Open-Set Recognition in the Age of Vision-Language Models.

Scalable Object Detection in the Car Interior With Vision Foundation Models Open-Set Recognition in the Age of Vision-Language Models

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T21:11:51.862990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T21:07:31.589011Z digest=sha256:bbbe8214eb80a183bb84387d24c31073befc575e9d3938c47e3971d794dfe8a0

Observation 874991cd-731b-4e5f-88db-bff4dd03d179 · outbound

This paper cites Renovating Names in Open-Vocabulary Segmentation Benchmarks.

Scalable Object Detection in the Car Interior With Vision Foundation Models Renovating Names in Open-Vocabulary Segmentation Benchmarks

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-18T21:11:50.824529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T21:07:31.589011Z digest=sha256:085fd76a8fe50e12149de7186d301b0d4fd574ac55ff32fb70ffe91833f871f2

Observation 8b72aa70-f6dd-446d-aa31-a1c97fe449e4 · outbound

This paper cites Transformers: State-of-the-Art Natural Language Pro- cessing.

Scalable Object Detection in the Car Interior With Vision Foundation Models Transformers: State-of-the-Art Natural Language Pro- cessing

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T21:11:51.856897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T21:07:31.589011Z digest=sha256:d752b3b281cca4bf70287c7544dc74291f0e9d76e903295cd05feb6d8441d52d

Observation a253fe9c-6e70-4f90-9bca-850a26f9ebdf · outbound

This paper cites TRL: Transformer Reinforcement Learning.

Scalable Object Detection in the Car Interior With Vision Foundation Models TRL: Transformer Reinforcement Learning

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T21:11:51.859916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T21:07:31.589011Z digest=sha256:051976389c617e35f5077cb75d8fbf54c5a32833a87a8b27f5c7535f1556826a

Observation d6260fc6-aa88-4bf0-aae0-471cf40d5938 · outbound

This paper cites PEFT: State-of-the-art Parameter-Efficient Fine- Tuning methods.

Scalable Object Detection in the Car Interior With Vision Foundation Models PEFT: State-of-the-art Parameter-Efficient Fine- Tuning methods

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T21:11:51.853483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T21:07:31.589011Z digest=sha256:5d7eb68fbc4981c678212372e6f4e4056d8e1176103ff7f478d59881f20f7ea0

Pith citing papers

No inbound Pith citation observations are available.