Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-23T22:20:21.427717Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 100 of 129 outbound references and 45 inbound Pith citation observations for arXiv:2402.11684.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-23T22:20:21.427717Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T23:28:02.771576Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-10T17:07:25.737454Z
100 of 129 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9d2d683a-73e5-49f9-aecc-1b27beacc083 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c2537725-e74f-4c91-8305-e37a9d132f3a · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 491845b6-90d9-44ed-b7b6-947560ba8a38 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e7294995-6413-44a0-ac14-6c1f788230f3 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e6b9af04-0852-4d05-8783-832f9e0234f0 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c7096c37-3649-4320-af67-c54aba6502b0 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2b775ada-2493-4478-9f65-54978b561c9a · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Introducing our Multimodal Models , url =
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8a57a259-5d04-4df7-8fe9-a9a1f15f1fd3 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , howpublished =
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 4041ad40-fa85-4071-92f4-61c4401e36b7 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 81503d65-50e5-4020-b37f-944af054eb09 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Vision-Flan:Scaling Visual Instruction Tuning , url =
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 44388c28-940f-4b21-a68e-ae5388db61f2 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models LIMA: Less Is More for Alignment
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b2b65ead-b4b2-4f62-92f4-f60724531658 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2022 , eprint=
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 653a7255-9f02-4f2c-a643-a9efa08ca86d · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2022 , eprint=
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ac03eb89-2e31-478d-b1de-444eba8ecece · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2bb80d9b-327a-430c-b4f4-cc186303cd69 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5529c882-0365-40af-b32a-025c0c678ceb · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2015 , eprint=
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2761e794-4eee-4be3-8331-a62df4cd818a · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2016 , eprint=
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9a364230-b94b-4b0c-824b-194c81bb0575 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2019 international conference on document analysis and recognition (ICDAR) , pages=
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f7de95bf-1e31-4817-9aa5-654de6c00192 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , pages=
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f314c186-052f-4e68-b298-b7de2e46e4c9 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2019 , eprint=
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b52e2ce0-e127-4caf-a304-db723afd055f · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2024 , eprint=
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e66aa8fc-b5a2-4309-9f2b-451c59ba5630 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2021 , eprint=
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f2d30a15-bf42-4802-8f6a-bf172d9396f8 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 342c2df9-f52d-4545-b04a-326a3135fb98 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2021 , eprint=
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b15b2755-52d0-4faf-9afd-3eeea75888bf · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Advances in neural information processing systems , volume=
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 31dbce92-e27c-4df3-95d2-a976e0054382 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ec4803c7-c36e-40f4-8e98-0077178dccf5 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2024 , eprint=
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6c41fa27-360f-4d73-a9e6-ac8b34d64a29 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d678331f-1226-4d5f-ab12-1497f4f11b06 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2024 , eprint=
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9c763088-936e-4117-846d-c9005c092b69 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 64bb3308-17a0-4d28-a27d-974286800dde · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Unresolved cited work
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 03275804-3167-4c0a-b26e-f08b8cd5cde8 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 640d4cdc-74d0-4c4a-ad7f-02a9c05bccd6 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 77284240-dda2-46c6-9465-5f2bcf952931 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 20e6a517-f5e9-496e-a7ea-2e883cc060cc · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ea12edbe-a67f-4603-917b-44961e0e5a4c · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation fe31deec-7e16-4f6e-ab73-c43d7230412d · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 08637f11-df65-4b36-96e4-e25aee153dca · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2021 , eprint=
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation cccad769-eb23-4b6a-ae05-73e822704d42 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2021 , eprint=
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2867262c-0806-431e-a96e-d69978d76030 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2021 , eprint=
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 81396910-be46-42da-a8bd-6562720adabe · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models and Stoica, Ion and Xing, Eric P
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2c88cf11-40d6-4f60-b450-75541e1df7d2 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2022 , eprint=
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d8512e99-c20c-4292-8635-f9e2904e77ad · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2a90798a-c33c-488b-95ab-a06655f2fa64 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 42b7c44c-d132-4065-8939-01ec6b74359d · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 95697ec4-af41-44c1-a16f-2cf994274dea · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 130973af-461b-44ad-9635-7c731621ff65 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e985cbda-1e97-4a7b-8112-bf29668133b1 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d5dde8f3-90df-43a0-bc4d-767f24ea125e · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation dcd1b7e0-0f8d-4a22-80f1-8d1e31b630ec · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c97d7154-6b70-4f32-bcb3-ab6216e28512 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2024 , eprint=
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f4384be8-2529-4326-856f-7c6c2200c731 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2009 , publisher=
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 868d901a-f915-4ae0-9030-ffb1fa32bd3b · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Proceedings of the IEEE/CVF international conference on computer vision , pages=
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1a780d89-ddf8-4c5f-a88e-b4c7b7e40652 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Transactions of the Association for Computational Linguistics , volume=
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 53b17f81-28c1-4ed7-9d11-dcb53d741470 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 08752f88-a987-485f-8c0d-d5ad64c66248 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 85893a63-bbdf-4e22-be26-8255eba7e608 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2022 , eprint=
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5b28d85a-173f-48f6-94bd-92dfa9fadbb2 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation def1f119-3f35-45eb-9521-c7fe23ebc60f · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2022 , eprint=
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 50e4f4d1-5156-47d7-a703-82349913dca7 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2021 , eprint=
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 889f30d6-66ff-46e3-a8cd-ac864de3c0a6 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1d1b284d-f935-45ae-b303-4fc392fec446 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2022 , eprint=
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b23d89ee-492c-4e7a-a013-8dba0bbd83cb · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Ilharco, M
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3c0f7ea7-d416-4146-8cf0-75fb7c87e0c3 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0e9966ce-37d1-44f0-93b7-2048ccca50f7 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Journal of machine Learning research , volume=
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation be2d446c-6f5f-4912-ad96-f1a291d8d32b · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models MALLET: A Machine Learning for Language Toolkit
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 10cf12b4-4978-4ef0-9fcc-fdc6bd908b85 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 2023 , eprint=
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 32d13e77-f61a-4c37-a4a6-e532db40460c · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models author=
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7b8f70a5-b063-4cdb-aef6-48d3396a4cce · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Chandra and Dexter C
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation dcac1eed-0ebb-491a-b960-166c4efa90cf · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Scalable training of
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1ede3114-024a-4da3-927d-086b40a14301 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Unresolved cited work
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3b1fd3fb-9745-42ea-aa2f-0e51e5f14d2e · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Tetreault , title =
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1cb50b23-0e7d-42c6-9dc1-56bd9a7a53d5 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models A Framework for Learning Predictive Structures from Multiple Tasks and Unlabeled Data , Volume =
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e5e25d00-0079-4f9b-ae9e-13f995e26361 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models and Tukey, John W
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 66ed7113-b106-4834-9797-bbc5901cf189 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Aho and Jeffrey D
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ee221ca3-c181-47f2-8079-9466e7e6e69e · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Unresolved cited work
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3817f296-48d9-494c-953d-d93ed57dcd4c · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Hewett, Jamie Huynh, Mojan Javaheripi, Xin Jin, Piero Kauffmann, Nikos Karampatziakis, Dongwoo Kim, Mahoud Khademi, Lev Kurilenko, James R
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1ccf5454-28e5-461a-a525-038fe49142fb · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Nocaps: Novel object captioning at scale
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6d274f27-fba9-4712-b910-459621df9cb5 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Qwen-vl: A versatile vision-language model for understanding, localization, text reading, and beyond, 2023 a
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 20ec7aa3-2814-4e87-9e78-c87026f775e7 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Touchstone: Evaluating vision-language models by language models, 2023 b
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9e0f0d3b-1789-4b9e-8e98-1e974de303a3 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Latent dirichlet allocation
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5eb180f7-2442-4c7a-9999-de2859a1ce72 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Conceptual 12m: Pushing web-scale image-text pre-training to recognize long-tail visual concepts
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1356a2e4-adec-4f61-a46a-02785d6caf6b · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models ShareGPT4V: Improving Large Multi-Modal Models with Better Captions
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 12d96111-2720-40b2-9b76-e3cbc86953f5 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Phoenix: Democratizing chatgpt across languages, 2023 b
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f88c942c-3803-4db3-ab38-982550ec80c8 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Gonzalez, Ion Stoica, and Eric P
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ced2eb04-bf6e-4629-bf64-b210f31dadbd · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Mobilevlm : A fast, strong and open vision language assistant for mobile devices
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 62c3fc62-8603-4219-8272-b505e84d3f81 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Instructblip: Towards general-purpose vision-language models with instruction tuning
Reference 95
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6702dafa-4bf8-4cc9-b834-a48acdd1b6c4 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Mme: A comprehensive evaluation benchmark for multimodal large language models
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0515903f-faf0-488a-aa9e-39721797804d · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Mllm-bench, evaluating multi-modal llms using gpt-4v
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 66c33652-3752-4a31-bdee-ff5ab1ee64e4 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Hallusionbench: An advanced diagnostic suite for entangled language hallucination & visual illusion in large vision-language models
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation bacb8e41-b547-4573-b237-c2c20e2998e5 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Hudson and Christopher D
Reference 99
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 53b20947-e357-44ba-a96b-8a9779ef5b46 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Introducing idefics: An open reproduction of state-of-the-art visual language model
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 901b63b4-68e7-4a75-9ce4-8859f4c7549d · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models TinyLLaVA Factory: A Modularized Codebase for Small-scale Large Multimodal Models
Reference 101
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a981e278-73ff-494c-81c0-d73501d91d8e · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Shamma, Michael S
Reference 102
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation cd9df9c1-8a33-4183-950a-690d86f4bbe4 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Seed-bench: Benchmarking multimodal llms with generative comprehension, 2023 a
Reference 103
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 26b3f7ae-e8f5-4f3e-a94c-4e7c50e5880a · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation
Reference 104
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 319d4a34-5fb3-42cf-beb0-dfa1ba506e61 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models, 2023 b
Reference 105
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2c650b40-4f4a-4d41-bb7a-cc3b06e5423c · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Silkie: Preference Distillation for Large Visual Language Models
Reference 106
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a78c0d9a-9103-471e-83d7-a84e24bb6b72 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Lawrence Zitnick, and Piotr Dollár
Reference 107
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 71ab9a6d-b6f1-4d9d-80df-8bff52901a14 · outbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Improved baselines with visual instruction tuning, 2023 a
Reference 108
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7f484300-e510-404c-b8f2-cf66eb5029db · inbound
A Survey on Multimodal Large Language Models ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 64a706b9-3249-4f6c-aa99-61d74a1960ea · inbound
Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0843e0c7-7283-4bf8-b8e7-4ace62ffb8b5 · inbound
How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e750d7f9-9365-4719-9f7d-3889b382ab3b · inbound
Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 57684e2e-a9c1-4d60-ab39-4700ec59b21a · inbound
InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 56e4d9ec-0f37-4b91-a8f7-79e175ca21f7 · inbound
MiniCPM-V: A GPT-4V Level MLLM on Your Phone ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 19822b22-57b7-44b1-9d38-ca6665b1f39d · inbound
LLaVA-OneVision: Easy Visual Task Transfer ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 188228ea-7320-4e61-9843-0f905066800c · inbound
mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 201
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1b9acf5c-c89a-4316-912f-b40cbaa05b3a · inbound
Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7c203caa-1a8f-46c5-b985-4fe3afd5b520 · inbound
VisRAG: Vision-based Retrieval-augmented Generation on Multi-modality Documents ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 12465fea-7643-4fdb-a10f-7759182e5100 · inbound
NVILA: Efficient Frontier Visual Language Models ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 95
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 60562b2e-a27c-4d56-b8c2-6b73887b1e0f · inbound
Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8a85821c-0631-48c4-9033-01459026a816 · inbound
DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f0646d9d-8115-40e0-9c7f-0ae1c7abd54f · inbound
MetaMorph: Multimodal Understanding and Generation via Instruction Tuning ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7311a77c-640e-4ac4-9681-d6e43ef63cfe · inbound
VideoChat-Flash: Hierarchical Compression for Long-Context Video Modeling ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation be7b2f50-db56-46cf-a2b5-3f3ee2c0b42b · inbound
InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 29948f7d-4168-40ed-9555-073fad16b4fd · inbound
Qwen2.5-VL Technical Report ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b8c5d58f-23a4-4a3e-9389-833e49223ddb · inbound
FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 09d2fd83-eeef-4386-bd3f-8c938035f13b · inbound
From Standalone LLMs to Integrated Intelligence: A Survey of Compound Al Systems ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation eac4d61d-d032-4427-a682-7161337aadda · inbound
ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f09fa8bd-6a5a-4514-a61f-37094c21544b · inbound
DenseWorld-1M: Towards Detailed Dense Grounded Caption in the Real World ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c74c8ca2-29e0-459b-8ba7-1b55b8ab2e6c · inbound
Advancing Multimodal LLMs by Large-Scale 3D Visual Instruction Dataset Generation ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 244696a9-e355-4869-9b6a-18fa1cf73075 · inbound
Mono-InternVL-1.5: Towards Cheaper and Faster Monolithic Multimodal Large Language Models ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 97
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae659049-9bdc-4c7c-a9a9-ae018ab13db5 · inbound
Being-H0: Vision-Language-Action Pretraining from Large-Scale Human Videos ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc5c1853-1c5a-4793-94e1-21bd4be5d566 · inbound
MAGE: Multimodal Alignment and Generation Enhancement via Bridging Visual and Semantic Spaces ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70656234-ccfe-4790-b2aa-9f3d34e43bcd · inbound
See Different, Think Better: Visual Variations Mitigating Hallucinations in LVLMs ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c68df4a-2ce6-4e09-9029-597b53eeb8ae · inbound
Concepts or Skills? Rethinking Instruction Selection for Multi-modal Models ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f65a0b2-c7f3-4021-9f79-277a0b1b0540 · inbound
UItron: Foundational GUI Agent with Advanced Perception and Planning ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11029495-1e23-44bf-945f-aa1297ea12f3 · inbound
InterAct: Advancing Large-Scale Versatile 3D Human-Object Interaction Generation ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3142315b-10aa-4957-acab-53d14926c21f · inbound
Generative Semantic Multi-Object Tracking: A Large-Scale Benchmark and an MLLM-Driven Reasoning Framework ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1b0a231-d878-4ea6-b605-325350dca75e · inbound
LLaVA-CKD: Bottom-Up Cascaded Knowledge Distillation for Vision-Language Models ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ca2f9252-a0b9-4afb-acd6-f99770275511 · inbound
Mitigating Mask Prior Drift and Positional Attention Collapse in Large Diffusion Vision-Language Models ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f32c4e1f-9cd0-415c-9197-f90f2246059a · inbound
Mitigating Mask Prior Drift and Positional Attention Collapse in Large Diffusion Vision-Language Models ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0345c5d3-e620-41ad-85ae-c1afd9c177e1 · inbound
A Nash Equilibrium Framework For Training-Free Multimodal Step Verification ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6b6c3cf9-6943-450e-be68-f3ad21c67230 · inbound
VCap: Hypergeometric Rewards for Weak-to-Strong Visual Captioning ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation bf538334-bcce-42a9-92c9-b16a4252d183 · inbound
CapRL++: Unified Reinforcement Learning with Verifiable Rewards for Dense Image and Video Captioning ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0b6e4a42-02b3-4b5c-98df-dfa2c8b94edc · inbound
DataComp-VLM: Improved Open Datasets for Vision-Language Models ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6d30b0ad-a591-435d-953b-7ea19ac1832f · inbound
DataComp-VLM: Improved Open Datasets for Vision-Language Models ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9bf5078f-b318-4c00-bc21-b63c3a4982a1 · inbound
Infinity-Parser2 Technical Report ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d9ce6021-0736-4b19-8f31-9f19cda1160a · inbound
Infinity-Parser2 Technical Report ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f52ac92f-c83b-4782-94bf-78228e7463b5 · inbound
RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 277
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0da804b-bc58-427f-b0a9-b877b11f34f1 · inbound
Twins: Learn to Predict Unified Representations with Focal Loss ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 292
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc030710-7f3d-4c6a-9c50-c2a7651aeed0 · inbound
StepX-Edge: An On-Device UI Vision-Language Model via Architecture-Training-Deployment Co-Design ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c14cd281-c78b-4d79-8b87-cbae69fe064e · inbound
ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 104
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee0a4277-ee5e-4f3e-ba16-6697ec990925 · inbound
Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.