Pith. sign in

Paper Citation Record · LEDGER

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models

As of 22 August 2026, this Paper Citation Record lists 70 of 70 outbound references and 2 inbound Pith citation observations for arXiv:2602.23802.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2602.23802 v2

Coverage vector

measured 70 of 70 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T20:14:04.208238Z

measured 72 of 72 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-26T20:58:30.855338Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T00:49:18.237054Z

Reference resolution

70 of 70 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved70
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4cd56d96-2b52-4103-a29d-74bb6f807d40 · outbound

This paper cites Qwen2.5-VL Technical Report.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Qwen2.5-VL Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:03.926761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:03.926761Z digest=sha256:e58661aa2643c17f76baaa4583bcdb7f04adb8a8ce78690e52f25b6db4eaf524

Observation 3e6a2209-868a-4065-a677-5949b44dea94 · outbound

This paper cites Chat-based person retrieval via dialogue-refined cross- modal alignment.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Chat-based person retrieval via dialogue-refined cross- modal alignment

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:03.931853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:03.931853Z digest=sha256:f4b9528c8190904b8f646c451faae7c0193ecf94a2e34af07d86bda0677ffc43

Observation f5f82134-7950-47e0-b50e-59a4701f04db · outbound

This paper cites LLaVA Steering: Visual Instruction Tuning with 500x Fewer Parameters through Modality Linear Representation-Steering.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models LLaVA Steering: Visual Instruction Tuning with 500x Fewer Parameters through Modality Linear Representation-Steering

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:03.936058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:03.936058Z digest=sha256:6df475b27ab15f1cb8392c7bb4dc855a9ec82474cae511c7636e174ef68cb2fa

Observation 4b031acc-1e60-4a7e-9592-a5d988094944 · outbound

This paper cites SEED-GRPO: Semantic Entropy Enhanced GRPO for Uncertainty-Aware Policy Optimization.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models SEED-GRPO: Semantic Entropy Enhanced GRPO for Uncertainty-Aware Policy Optimization

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:03.940645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:03.940645Z digest=sha256:c69faa2715a658d9cbebe3f26460b68cfdca975ba7903af04178004b9326d3f7

Observation b75a2b15-288a-46cc-bfc9-084c113ae010 · outbound

This paper cites Internvl: Scaling up vision foundation mod- els and aligning for generic visual-linguistic tasks.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Internvl: Scaling up vision foundation mod- els and aligning for generic visual-linguistic tasks

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:03.944872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:03.944872Z digest=sha256:557c352543cc3343ed045d8fcbce4d6bfef1b7807fb21ba7339fee201edbfab5

Observation 867c155f-91dd-45cd-84de-92652adef16f · outbound

This paper cites Emotion-llama: Multimodal emo- tion recognition and reasoning with instruction tuning.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Emotion-llama: Multimodal emo- tion recognition and reasoning with instruction tuning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:03.948806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:03.948806Z digest=sha256:679f97a07ec0c08fa3f41e592d22c7d1549f7d0b53eea21d1d7c6a3caf5c9d15

Observation 2c596fe0-263a-40e7-ae4c-fc5c78ff0345 · outbound

This paper cites Emoe: Modality-specific enhanced dynamic emotion experts.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Emoe: Modality-specific enhanced dynamic emotion experts

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:03.953378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:03.953378Z digest=sha256:e45da23956a6ca68c1b1f28cb8e803b27fdc26afdda45f487e2c0636195683f6

Observation e3a43277-9b65-48cc-8d6c-d0ace6ab562c · outbound

This paper cites Catch your emotion: Sharpening emotion perception in multimodal large language models.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Catch your emotion: Sharpening emotion perception in multimodal large language models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:03.957598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:03.957598Z digest=sha256:71c76db5aa9d2ae13a95160f4f28d2b35b902cacaa2394d5f274d9a7d0a34798

Observation 69e9830d-e0f9-489f-9055-05e1339ef576 · outbound

This paper cites Video-R1: Reinforcing Video Reasoning in MLLMs.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Video-R1: Reinforcing Video Reasoning in MLLMs

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:03.962128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:03.962128Z digest=sha256:197036e5b48e33fcb61af228beaa38327264c20cfcde1e1f9812186b06d7e2ef

Observation 4c969bbd-7ccb-46ba-894c-6d7ce7317ced · outbound

This paper cites On Designing Effective RL Reward at Training Time for LLM Reasoning.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models On Designing Effective RL Reward at Training Time for LLM Reasoning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:03.966604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:03.966604Z digest=sha256:805f2a26b0b23019dee78a6c53433d4b2f12603c2e68c0b1cef3eb729eef0119

Observation 169d64ab-9bc3-4aaa-a3d1-e31d2e091adb · outbound

This paper cites Making the V in VQA matter: Ele- vating the role of image understanding in Visual Question Answering.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Making the V in VQA matter: Ele- vating the role of image understanding in Visual Question Answering

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:03.970645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:03.970645Z digest=sha256:b5cf6b682e460b19370f56a54f7c752fe42674cf8e969994de0854fd0e9f156c

Observation 0e5c2c49-9e68-49f4-821b-c86b13211d03 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:03.975221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:03.975221Z digest=sha256:337645bd6b9e687f46259580224077c3e63ed249f95f7c7dd005f86f815575c7

Observation c884e79d-b998-4036-9c26-39a9e3a79689 · outbound

This paper cites Onellm: One framework to align all modalities with language.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Onellm: One framework to align all modalities with language

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:03.979604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:03.979604Z digest=sha256:b9c89b72b3b25bb7ec7f2c8b8e5bdc98893cb5a345ab9b85856798be54d95227

Observation 88669c8e-ff9e-460f-a9bc-475363cf071b · outbound

This paper cites Boosting MLLM Reasoning with Text-Debiased Hint-GRPO.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Boosting MLLM Reasoning with Text-Debiased Hint-GRPO

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:03.983641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:03.983641Z digest=sha256:99acfc5aab6642f26a470f0ae07cb1979afcb07dbd56aac6562583387688b439

Observation 09b70d92-c221-4c5d-bb18-351c279dd5fc · outbound

This paper cites Keeping Yourself is Important in Downstream Tuning Multimodal Large Language Model.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Keeping Yourself is Important in Downstream Tuning Multimodal Large Language Model

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:03.988119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:03.988119Z digest=sha256:751d8812b651b88a6e141480fbf67cdc0c2d1fcbcca0f92b9495d6181e88840d

Observation 166182ab-f88f-4a23-94be-1b58b69e413f · outbound

This paper cites Learn from downstream and be yourself in multimodal large language model fine-tuning.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Learn from downstream and be yourself in multimodal large language model fine-tuning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:03.992089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:03.992089Z digest=sha256:f5e70ac26cbb9b178e88f0e6774d7b166f1ba3766b4ec7a1fb1787e13e80264e

Observation b49e719c-bbed-4059-8cbf-2654b980bcec · outbound

This paper cites Be confident: Uncovering overfitting in mllm multi-task tuning.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Be confident: Uncovering overfitting in mllm multi-task tuning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:03.995576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:03.995576Z digest=sha256:b2d7b6e9358c8a1ac02da187ab6415a3777a351e081c1d1e9e22cda69ae3b9a8

Observation f978c6f9-0930-4f9e-9c52-2fafc4da15a5 · outbound

This paper cites Mapo: Mixed advantage policy optimization.arXiv preprint arXiv:2509.18849, 2025.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Mapo: Mixed advantage policy optimization.arXiv preprint arXiv:2509.18849, 2025

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.000130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.000130Z digest=sha256:f9869eb5da47c6b5edf46134f38145443d0ac510d74cd5d60d4125d3e0880a76

Observation 8d0ee372-0197-4789-b5fd-33a3547311bf · outbound

This paper cites Gqa: A new dataset for real-world visual reasoning and compositional question answering.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Gqa: A new dataset for real-world visual reasoning and compositional question answering

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.004215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.004215Z digest=sha256:e7771d4f3eab1b5d6a833bdf6551a89b2f5f43fbf28b0b180984a83ec5c4477a

Observation 6ba12ad3-7345-43f1-b9d5-bf74065a11d0 · outbound

This paper cites Learning from teaching reg- ularization: Generalizable correlations should be easy to im- itate.NeurIPS, 37:966–994, 2024.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Learning from teaching reg- ularization: Generalizable correlations should be easy to im- itate.NeurIPS, 37:966–994, 2024

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.014431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.014431Z digest=sha256:652382527930290e1d6c1f8c4df2c5d538e0cb7b373ec3bd70d71e1eefccd13c

Observation b643d058-c196-4ab9-8cf6-775d461e7f7e · outbound

This paper cites Two Heads are Better Than One: Test-time Scaling of Multi-agent Collaborative Reasoning.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Two Heads are Better Than One: Test-time Scaling of Multi-agent Collaborative Reasoning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.018075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.018075Z digest=sha256:71ca3d0fa791dae552836ceb8a4c66f05c9c55d5239b1742a32410906ea6d98f

Observation 4f545ca4-f4e5-46f0-8ab6-9fc7b8480e2a · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models LLaVA-OneVision: Easy Visual Task Transfer

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.022097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.022097Z digest=sha256:13802b37c17bcc9bacda55d01074921d15709d26929ab72bedb37c193b23e2ec

Observation e9ed3b7d-bb1c-48e3-8590-3c8bb8682315 · outbound

This paper cites VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.026500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.026500Z digest=sha256:6a4dfeedb31f4f9d5b8f351a5d2ca995ee92a29bb09d3ce359ad6c731152d74d

Observation 52a03b49-545a-4add-8801-f2ef62ba0e70 · outbound

This paper cites Mon- key: Image resolution and text label are important things for large multi-modal models.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Mon- key: Image resolution and text label are important things for large multi-modal models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.031632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.031632Z digest=sha256:8432676efc02a050c025303f82a1fc170ab71422314eea5ae0e6dff245f70e44

Observation 59cc902c-ffc9-4312-b13e-d044579918a7 · outbound

This paper cites Explainable Multimodal Emotion Recognition.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Explainable Multimodal Emotion Recognition

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.035428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.035428Z digest=sha256:9405aa42d94f8fa077aedcdbbeca393e7574c536777325bd1a8531f273496d4b

Observation 6259af68-65e9-4987-87fb-aeb8c558bf51 · outbound

This paper cites Ex- plainable multimodal emotion reasoning.CoRR, 2023.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Ex- plainable multimodal emotion reasoning.CoRR, 2023

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.040180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.040180Z digest=sha256:f0aca11876b5dc9f9ac2f6f9eabc8b81047ef018e0a77d968a87e257e34534ee

Observation 53e94391-8752-4c87-91f6-da012278d3c6 · outbound

This paper cites AffectGPT: A New Dataset, Model, and Benchmark for Emotion Understanding with Multimodal Large Language Models.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models AffectGPT: A New Dataset, Model, and Benchmark for Emotion Understanding with Multimodal Large Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.044246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.044246Z digest=sha256:39b5a46b1ae8b421f0db3dde31356c19772994e32ea23f3ae1bc844cb7831ac6

Observation 27510ae0-731b-4a09-ba3b-d1d3fa7dfc4a · outbound

This paper cites Lorasculpt: Sculpting lora for harmonizing gen- eral and specialized knowledge in multimodal large language models.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Lorasculpt: Sculpting lora for harmonizing gen- eral and specialized knowledge in multimodal large language models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.049463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.049463Z digest=sha256:4f41b4230aebf442c4e31b889d735a35cb41b8b3003252b6964a03a8dc05bf9d

Observation 5198f810-e98c-4f43-a03b-29e392bb71b6 · outbound

This paper cites Microsoft coco: Common objects in context.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Microsoft coco: Common objects in context

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.053566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.053566Z digest=sha256:5e659e5467f7d0cf1d2b9b3d4ba384071dedd6647d0682c7cba1d734c292b018

Observation f2c09ffd-9984-4cc6-8f12-dffbfb00adc5 · outbound

This paper cites Improved baselines with visual instruction tuning.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Improved baselines with visual instruction tuning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.057327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.057327Z digest=sha256:afb29aaed7db23429b9b1be740b461ab27241af6effdc1da4d2dbfcb1bbaf5d4

Observation 0e13ce9d-1ac6-4f46-ba5e-ece760393791 · outbound

This paper cites Reinforcement Learning Meets Large Language Models: A Survey of Advancements and Applications Across the LLM Lifecycle.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Reinforcement Learning Meets Large Language Models: A Survey of Advancements and Applications Across the LLM Lifecycle

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.060992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.060992Z digest=sha256:89e47a303a157a590f388e559308c8941478807691e4c17e12e0c41e03c7a624

Observation ca3dee2f-ea98-4697-b24c-e2cc7392bf62 · outbound

This paper cites DoRA: Weight-Decomposed Low-Rank Adaptation.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models DoRA: Weight-Decomposed Low-Rank Adaptation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.064395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.064395Z digest=sha256:8d1846d40a6a4063d9057fc674eeafb3550a86b2184008aa1189b03086d73667

Observation 3d6b8d6d-4372-4639-b0c1-1623d1bf4361 · outbound

This paper cites GuardReasoner-VL: Safeguarding VLMs via Reinforced Reasoning.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models GuardReasoner-VL: Safeguarding VLMs via Reinforced Reasoning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.067928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.067928Z digest=sha256:a5552ca58a66581bffadc2bfbbab2af56e392d959bb69aa7f7f86f8d6d252357

Observation 962da5aa-3069-4019-b7c8-bbc4b4b49f14 · outbound

This paper cites Visual-RFT: Visual Reinforcement Fine-Tuning.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Visual-RFT: Visual Reinforcement Fine-Tuning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.072542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.072542Z digest=sha256:2c81e1c2bfba99d256dfee45187705915ca330bbc901f6fc4eac3ccfe7b3428e

Observation 99ecd203-8520-4069-aa2e-076ab866789e · outbound

This paper cites Learn to explain: Multimodal reasoning via thought chains for science question answering.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Learn to explain: Multimodal reasoning via thought chains for science question answering

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.076915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.076915Z digest=sha256:aa5522f70c9a92c7fd51897a70fc963c994396de85f080b3da8f6711cedb6b59

Observation 92ee911f-cf4c-47d1-9d89-8c804cb215b8 · outbound

This paper cites ChartQA: A benchmark for question answering about charts with visual and logical reasoning.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models ChartQA: A benchmark for question answering about charts with visual and logical reasoning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.080359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.080359Z digest=sha256:1d4e0da1839a528066cddfd0ab56d721717bee151d2e2a6bfb5af94ff7337578

Observation aff71ea1-2ca9-4465-803d-97ee918da007 · outbound

This paper cites Con- templating visual emotions: Understanding and overcoming dataset bias.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Con- templating visual emotions: Understanding and overcoming dataset bias

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.083661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.083661Z digest=sha256:38ee7388fe5f7dded97564b117794da81c294af16570da9e7359f5f68983568c

Observation 27520b90-5628-4f04-b25c-a61d3aa78a1f · outbound

This paper cites A mixed bag of emotions: Model, predict, and transfer emotion distributions.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models A mixed bag of emotions: Model, predict, and transfer emotion distributions

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.087354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.087354Z digest=sha256:7086f2cc006b6bdd72f5e41744a6aede8080a592310230518aabd4be17d523ca

Observation 0bf45be1-2d87-43d4-819e-b047bcce5e4d · outbound

This paper cites Direct preference optimization: Your language model is secretly a reward model.NeurIPS, 36:53728–53741, 2023.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Direct preference optimization: Your language model is secretly a reward model.NeurIPS, 36:53728–53741, 2023

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.090762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.090762Z digest=sha256:25a775ebff25a35d94e3cac74b5a7afdf96f2f40aa9982a9c52f0173664ba239

Observation 6a028c1d-a1f0-43e3-8bb6-79290508172e · outbound

This paper cites Scalpel vs. Hammer: GRPO Amplifies Existing Capabilities, SFT Replaces Them.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Scalpel vs. Hammer: GRPO Amplifies Existing Capabilities, SFT Replaces Them

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.093885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.093885Z digest=sha256:bdac2a903d1a785f05d7302bd2eda46b61c68e3d9ea43e565fbde0f3ed04bac7

Observation 4728e5c8-49e0-4840-8f3d-9bea881b4156 · outbound

This paper cites Group robust preference optimization in reward- free rlhf.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Group robust preference optimization in reward- free rlhf

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.097359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.097359Z digest=sha256:3a24387801ecb04f9b2be8a92e3f40f69042f6a0fe8d87e79e9b29ddba14bcd7

Observation c4a76982-bab2-4716-b3a1-911f1a06ad74 · outbound

This paper cites Improving LLM-Generated Code Quality with GRPO.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Improving LLM-Generated Code Quality with GRPO

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.101113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.101113Z digest=sha256:3f05877679bf9493c94e656e8ec590e93e9e1c87c69c0f3d503253f9b95f4501

Observation c032eb3f-1a85-4ae7-80d9-82f2ad050429 · outbound

This paper cites Backdoor Cleaning without External Guidance in MLLM Fine-tuning.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Backdoor Cleaning without External Guidance in MLLM Fine-tuning

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.105502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.105502Z digest=sha256:cdbb48e23f47bcd457dd4399e8d9c8c9320dc460f4231f5dc0094b4c42adfb8d

Observation 46c5102c-3f2c-401e-89c3-e8042cc4199c · outbound

This paper cites Safegrpo: Self-rewarded mul- timodal safety alignment via rule-governed policy optimiza- tion.arXiv preprint arXiv:2511.12982, 2025.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Safegrpo: Self-rewarded mul- timodal safety alignment via rule-governed policy optimiza- tion.arXiv preprint arXiv:2511.12982, 2025

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.109291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.109291Z digest=sha256:c2f7f1052a4453843d008bcb44367763da98d06e154a52ef542fda20e616a261

Observation 69886bcf-14df-4f05-8660-906311129984 · outbound

This paper cites Proximal Policy Optimization Algorithms.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Proximal Policy Optimization Algorithms

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.113356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.113356Z digest=sha256:267d72c094470fd6ca6454fbd23d3b136d6a4a01ebf655b731d9b0c2e20bc8d8

Observation 1e57f08a-bbc3-4180-9778-995074537ebb · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.117269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.117269Z digest=sha256:90691fb12a741ed2d3e8a1431941005b85f5bc6b20f1bad70cc7936b4b2ebf86

Observation ecce399c-cda8-4b52-b950-650d78ca25d5 · outbound

This paper cites Towards vqa models that can read.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Towards vqa models that can read

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.120825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.120825Z digest=sha256:f0a8c457dec1b808794d6c65f019e47f166305a7037e97cefa03cfada8d0886a

Observation f372d45e-5e7b-40e5-94aa-87ed35887192 · outbound

This paper cites Delving into RL for Image Generation with CoT: A Study on DPO vs. GRPO.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Delving into RL for Image Generation with CoT: A Study on DPO vs. GRPO

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.125091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.125091Z digest=sha256:956e9e3342eb5c3078cb8c62c01dbf7edb845778ab3c11833ad8d3089535152e

Observation 7a3c0ec5-ac01-4374-90f1-de67f8959d2e · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models LLaMA: Open and Efficient Foundation Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.128711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.128711Z digest=sha256:409872299c16430290a14ea2834431666d4d1a01ce2dc773978cb5a97da5bb66

Observation 522417f4-1a78-4044-a568-0fd3d45fca66 · outbound

This paper cites Safety in Large Reasoning Models: A Survey.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Safety in Large Reasoning Models: A Survey

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.132863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.132863Z digest=sha256:83284a97a5cb2caca70d141055ee3f9b042d818b6dd78f6077acea765800bf49

Observation 22e2dcba-707d-4f81-b689-e8ad3995e6b8 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.136705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.136705Z digest=sha256:d284b1e8befb2ff12ecf2ce8656a150138734fedb87591583e10b732c77bcef8

Observation 59060050-f419-4957-89b1-643d4c823cba · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large lan- guage models.NeurIPS, 35:24824–24837, 2022.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Chain-of-thought prompting elicits reasoning in large lan- guage models.NeurIPS, 35:24824–24837, 2022

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.140872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.140872Z digest=sha256:8e422509245d4a280302db2c4848428076017625f04f749ae13d8872467a91b7

Observation c1e93403-6ca3-4643-ac57-c4e73cf20b0e · outbound

This paper cites Emovit: Revolutionizing emotion insights with vi- sual instruction tuning.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Emovit: Revolutionizing emotion insights with vi- sual instruction tuning

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.144878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.144878Z digest=sha256:8f390a508b35b0760b91abc525fb7670a9418ea13a1938e8014be35ca022d2d6

Observation 87fc2543-5052-495c-859b-030fe9d8634d · outbound

This paper cites EMO-LLaMA: Enhancing Facial Emotion Understanding with Instruction Tuning.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models EMO-LLaMA: Enhancing Facial Emotion Understanding with Instruction Tuning

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.148188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.148188Z digest=sha256:a965c8d9f77a7275ab0848f3f387426f450343c721e8aefdbac750f96f773eae

Observation 3df01c42-bc3c-42a3-bcce-7636b9fec44e · outbound

This paper cites Is DPO Superior to PPO for LLM Alignment? A Comprehensive Study.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Is DPO Superior to PPO for LLM Alignment? A Comprehensive Study

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.152731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.152731Z digest=sha256:027f71a707184be7aeedd3859d8516c14b816337802a540729ce12cacf1da14c

Observation 692916be-1ac3-489f-a15d-be3cdbf6720c · outbound

This paper cites Context de-confounded emo- tion recognition.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Context de-confounded emo- tion recognition

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.156268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.156268Z digest=sha256:3356df137cbf3cc52e25d651402336d9a571bf973098e4ffa5f1e03b6da35919

Observation af7a6d69-2908-4c83-93d8-f171cf81bf40 · outbound

This paper cites Emoset: A large-scale visual emotion dataset with rich attributes.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Emoset: A large-scale visual emotion dataset with rich attributes

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.160499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.160499Z digest=sha256:8adc4ce959081437d99687c0e382c4f89b6ab6210e93dad70357ed9e695b9e06

Observation f35fdd9d-b03b-4a5d-aa92-085b7533807b · outbound

This paper cites EmoLLM: Multimodal Emotional Understanding Meets Large Language Models.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models EmoLLM: Multimodal Emotional Understanding Meets Large Language Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.163905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.163905Z digest=sha256:d7d36ef681aeb3790603e8d8e9279e1ea9b97a9cbf41e5b1dd9aab747a3e904a

Observation 556aad84-e02d-4eb8-89cc-e5b46b98d3bf · outbound

This paper cites Treerpo: Tree relative policy optimization.arXiv preprint arXiv:2506.05183, 2025.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Treerpo: Tree relative policy optimization.arXiv preprint arXiv:2506.05183, 2025

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.167966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.167966Z digest=sha256:92ce50395b54eee453dee29526ef63cdebb87a95cf136034fa048d9dce3df81b

Observation 8f57803f-09a7-4de7-acb0-196233095552 · outbound

This paper cites R1-ShareVL: Incentivizing Reasoning Capability of Multimodal Large Language Models via Share-GRPO.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models R1-ShareVL: Incentivizing Reasoning Capability of Multimodal Large Language Models via Share-GRPO

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.172212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.172212Z digest=sha256:45dce259435f561a5c78acdc9aedde7fac95297ad789ac9479b189ab265f8967

Observation 62ccc93d-2d7e-438a-95c0-4eb85a2c60fa · outbound

This paper cites A Survey of Safety on Large Vision-Language Models: Attacks, Defenses and Evaluations.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models A Survey of Safety on Large Vision-Language Models: Attacks, Defenses and Evaluations

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.175590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.175590Z digest=sha256:678423573b66de55e9f196718e5fde332026950a8614fb1e1f9b27bf3fa368c0

Observation 3dcfe257-3936-4fa6-b59a-5fe1adcdfa63 · outbound

This paper cites From image descriptions to visual denotations: New similarity metrics for semantic inference over event descrip- tions.TACL, 2:67–78, 2014.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models From image descriptions to visual denotations: New similarity metrics for semantic inference over event descrip- tions.TACL, 2:67–78, 2014

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.179183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.179183Z digest=sha256:79185d29b7c80b1787fd2ebd3467612e6316b259b950897f91fdf6b95cabc1c0

Observation 6d08763c-cd4f-4ed1-8cd7-fb6ff35ccfcf · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.182502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.182502Z digest=sha256:b514757471110614723d813ef5c764c45e3f0ed1fdc3cd982dc1b72eeb742eda

Observation cbd7b248-e55a-4389-b6b7-990c45a2056c · outbound

This paper cites R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.186690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.186690Z digest=sha256:00a9d1007d83c9c070aad36f190f34e058d7f529da6279057cce6f065fb1e687

Observation 66cb5cf7-9f92-45a8-8ccb-796b1d8260de · outbound

This paper cites Microemo: Time-sensitive multimodal emotion recognition with subtle clue dynamics in video dialogues.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Microemo: Time-sensitive multimodal emotion recognition with subtle clue dynamics in video dialogues

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.190105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.190105Z digest=sha256:dadf15cfd3277c812d8be27532285acfb34b7a18af887358dd14344d181cd5ef

Observation 8d8e9f2e-b3b5-446e-9ad1-4a64d29f34d9 · outbound

This paper cites How Can LLM Guide RL? A Value-Based Approach.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models How Can LLM Guide RL? A Value-Based Approach

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.194029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.194029Z digest=sha256:94afec23d8dde558b0c9c219faf4c2805773eece35a83813411c266bfbff165f

Observation 52018fda-f74d-4068-8447-68a61f8ef8f0 · outbound

This paper cites Facephi: Lightweight multimodal large language model for facial landmark emotion recogni- tion.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Facephi: Lightweight multimodal large language model for facial landmark emotion recogni- tion

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.198029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.198029Z digest=sha256:aec2c2c8c158dd02aee21d97ee047981ac726815a57220cbd24e9fb80e06d0bb

Observation 76d0ea6c-b4c6-4acf-a1aa-53d26834d7d3 · outbound

This paper cites GaLore: Memory-Efficient LLM Training by Gradient Low-Rank Projection.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models GaLore: Memory-Efficient LLM Training by Gradient Low-Rank Projection

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.201189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.201189Z digest=sha256:e49857b7f8594313148b95bbb14f1a0e16db4a81f7813d96b1b33a8211854dde

Observation da0ed51d-46e0-4298-a61d-1cb6ba46f34e · outbound

This paper cites R1-Omni: Explainable Omni-Multimodal Emotion Recognition with Reinforcement Learning.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models R1-Omni: Explainable Omni-Multimodal Emotion Recognition with Reinforcement Learning

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.204686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.204686Z digest=sha256:1efe9cfa72d9af8af6c2c3066b94b4a2b58dc25022223ff4c6d02507e1b8e0cd

Observation 2c6fc92a-3f10-48c5-9a5d-e03dc2aaa7f9 · outbound

This paper cites Reinforced MLLM: A Survey on RL-Based Reasoning in Multimodal Large Language Models.

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Reinforced MLLM: A Survey on RL-Based Reasoning in Multimodal Large Language Models

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-02T20:14:04.208238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:14:04.208238Z digest=sha256:3424cabd16afff2334f88730da7d23573faf4a2a5d0c740e55c7247295d113ac

Pith citing papers

Observation fa78e801-308f-4fc0-9c07-e87da5f93f55 · inbound

EmoTrans: A Benchmark for Understanding, Reasoning, and Predicting Emotion Transitions in Multimodal LLMs cites this paper.

EmoTrans: A Benchmark for Understanding, Reasoning, and Predicting Emotion Transitions in Multimodal LLMs EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-09T02:19:51.159867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-08T08:34:14.297331Z digest=sha256:a8611caa2c352b0573365c15441f738e98f400edd4344f8080e228b7e74f6218

Observation 69fc91f4-6db2-471b-96ed-195c024bd24c · inbound

ThinkDeception: A Progressive Reinforcement Learning Framework for Interpretable Multimodal Deception Detection cites this paper.

ThinkDeception: A Progressive Reinforcement Learning Framework for Interpretable Multimodal Deception Detection EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-09T02:19:51.159867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-26T20:58:30.855338Z digest=sha256:858fbd6ace2ef4844e2680ff08e0fafc3cdc18582f667bda023e7e4750caacef