Pith. sign in

Paper Citation Record · LEDGER

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding

As of 9 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 2 inbound Pith citation observations for arXiv:2507.08031.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.08031 v2

Coverage vector

measured 26 of 26 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:08:59.002718Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T22:36:30.735420Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T23:07:26.843285Z

Reference resolution

26 of 26 outbound references displayed

  • verified exact1
  • verified fuzzy19
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 340f2f81-bc9b-4e8e-8c9f-16712fde00ca · outbound

This paper cites The global economic burden of noncommunicable diseases,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding The global economic burden of noncommunicable diseases,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:05.516187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:55.374946Z digest=sha256:9191824bb585b5d28d242c810d939a9ee5aa686a1ba35ae63322d1e2a0430594

Observation ac2ee850-762c-4aa8-8611-02cb2e2cbe99 · outbound

This paper cites The State of Mental Health in America 2024,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding The State of Mental Health in America 2024,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:05.220131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:55.536436Z digest=sha256:be62cab0b183c5156bf91042aa859e812d3cb88374ab1fb8a1d28b8cc91721de

Observation ac6d70c3-97d2-46e4-9bc7-c3dfe9647294 · outbound

This paper cites How mental health care should change as a consequence of the COVID-19 pandemic,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding How mental health care should change as a consequence of the COVID-19 pandemic,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:04.823250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:55.674744Z digest=sha256:2f01fc85a828dd02844b210230e559433d52dad89a29dcfe7d948f41d2d4a99e

Observation cf8a9203-3ef4-42f9-ac90-1cfefefd974b · outbound

This paper cites Social media and mental health: benefits, risks, and opportunities for research and practice,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Social media and mental health: benefits, risks, and opportunities for research and practice,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:04.443059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:55.855024Z digest=sha256:ae327c1aae27dac394d2042d1d2f60d5bbc25f41f76d2a70101b151f3a85b13d

Observation ed8ba62f-76f1-460d-8fdf-8bf111c002cd · outbound

This paper cites Dang, et al.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Dang, et al

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:04.117800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:55.986628Z digest=sha256:4d03a824682a6e415ade88453f0d98f8cff9b3e4627f8ea635d455f75a477800

Observation bbcaa6df-4ea7-40b5-b83f-f99f87c2a3e1 · outbound

This paper cites Mental-llm: Leveraging large language models for mental health prediction via online text data,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Mental-llm: Leveraging large language models for mental health prediction via online text data,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:03.759548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:56.164960Z digest=sha256:98c73aac6e5a21877808e3e6b91913e0e5ccbb9f607bd87fa3cb14320af0f030

Observation df586b4e-7273-459e-baa1-82a4c6c9d3b0 · outbound

This paper cites A taxonomy of ethical tensions in inferring mental health states from social media,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding A taxonomy of ethical tensions in inferring mental health states from social media,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:03.410345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:56.349071Z digest=sha256:c0d3907b216e19578f74936a0b5b7a3636275cfd418a7c43a25f72dabf1ac189

Observation 232ef0ce-4eb6-4ad8-a40e-6b0ddb21574f · outbound

This paper cites Beyond LDA: exploring supervised topic modeling for depression-related language in Twitter,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Beyond LDA: exploring supervised topic modeling for depression-related language in Twitter,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:03.038047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:56.489214Z digest=sha256:1f044660add797dd24627dd5cfaf2bdf818df966e73fba53c7a7830a535daf78

Observation 6a88d9a5-b0a4-4b26-a13b-a651dbc3ec61 · outbound

This paper cites AER-LLM: Ambiguity-aware emotion recognition leveraging large language models,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding AER-LLM: Ambiguity-aware emotion recognition leveraging large language models,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:02.718421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:56.606960Z digest=sha256:2d75922313c5e3907f634052c3d4028570030b260d76614619d775a37539996e

Observation d0cf2ed4-002f-405d-88a0-61037b8c5166 · outbound

This paper cites Token-Level Logits Matter: A Closer Look at Speech Foundation Models for Ambiguous Emotion Recognition.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Token-Level Logits Matter: A Closer Look at Speech Foundation Models for Ambiguous Emotion Recognition

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-06T19:08:59.364163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:56.780027Z digest=sha256:9749d76b1ec5dcb87a22235d53e941bb47a3f21a3d9394004812f1164d6b321e

Observation a30306b3-31d2-4bd9-9826-d96e1df3cc58 · outbound

This paper cites Privacy-preserving deep learning,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Privacy-preserving deep learning,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:02.377071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:56.960705Z digest=sha256:13bdc7790119e9530fd4f34c954a99b0cc89798a93233e15f574ac6ee19bbec9

Observation a42e6388-c2e0-436c-ae8a-ee4d146ba5ed · outbound

This paper cites ”Efficient and personalized mobile health event prediction via small language models.” Proceedings of the 30th Annual International Conference on Mobile Computing and Networking.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding ”Efficient and personalized mobile health event prediction via small language models.” Proceedings of the 30th Annual International Conference on Mobile Computing and Networking

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:02.076532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:57.110786Z digest=sha256:ddbd60426fcbf420f478644b6823aead8ac62c3fd987765989a845327001360f

Observation 1f9d8cff-568d-4c06-b2ae-4bd1ed3793b9 · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T19:08:57.240115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:08:57.240115Z digest=sha256:a25e387fdf812fc3974306db947a9b8ee1bfafe5f5fd572b80b0d542de02446c

Observation a65a9797-5452-4623-b3fe-cbf952fb2d17 · outbound

This paper cites Gemma: Open Models Based on Gemini Research and Technology.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Gemma: Open Models Based on Gemini Research and Technology

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T19:08:57.400446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:08:57.400446Z digest=sha256:0cfb27c55586db8c7d662ecdce8dd0e40c71076962a0b04d82637d7e2bcbbb10

Observation 8166ceba-ab53-44dd-b316-1fe5fb4bfdb3 · outbound

This paper cites Federated Learning for Mobile Keyboard Prediction.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Federated Learning for Mobile Keyboard Prediction

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T19:08:57.534782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:08:57.534782Z digest=sha256:0b2358a7d9c46cba5b8fe7d5607f2777cd47563f329b50de7b406c802d2e2acb

Observation 0b7885b5-bb66-4003-804b-31905044fb24 · outbound

This paper cites A primer on neural network models for natural language processing,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding A primer on neural network models for natural language processing,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:01.809452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:57.664397Z digest=sha256:50b32cd21839b7b7c5b23389cf95c8bdf6d82853392cba783151886000def448

Observation 7687ee98-4b31-467c-b682-31924f681c16 · outbound

This paper cites On the dangers of stochastic parrots: Can language models be too big?,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding On the dangers of stochastic parrots: Can language models be too big?,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:01.492345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:57.830414Z digest=sha256:18cffb5b9dbf997cce151b7b1fb464ceba18164837a33d340d8b2999958cc16d

Observation d5a478d2-ff6e-4e7b-aa3e-80cdef78c9c5 · outbound

This paper cites A discourse-aware attention model for abstractive summarization of long documents,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding A discourse-aware attention model for abstractive summarization of long documents,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:01.162022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:57.987186Z digest=sha256:d56dd315b0d7aee1bb16bcdb8299a2c147c34a3ff6bdbd455d7db3dc542a902e

Observation ffc4567f-fd2d-442e-a9c4-3989fa3b3b58 · outbound

This paper cites Language models are few-shot learners,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Language models are few-shot learners,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:00.910111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:58.082008Z digest=sha256:1650532a6afec8bb0df9f0dc5afc66cd8a0ca0960f04540247273afd88abc54e

Observation 7e54fe16-3fee-41c6-9104-bde0ff0e3067 · outbound

This paper cites CLPsych 2019 shared task: Predicting the degree of suicide risk in Reddit posts,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding CLPsych 2019 shared task: Predicting the degree of suicide risk in Reddit posts,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:00.546999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:58.196456Z digest=sha256:9af06a0523baec1188c149f02213416794537ee49d59fddf7b52805432598448

Observation 801508b4-45de-4c0b-856c-f66ca3c56fe2 · outbound

This paper cites Dreaddit: A Reddit Dataset for Stress Analysis in Social Media.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Dreaddit: A Reddit Dataset for Stress Analysis in Social Media

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T19:08:58.333008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:08:58.333008Z digest=sha256:f177ba39d147ca80a206fccc90449d3f666d8bf5e9e16bd5fe0e5689ed0ce846

Observation 18cfa444-353b-4a67-a39f-244094fce3d6 · outbound

This paper cites Early Detection of Depression Severity Levels on Reddit using a Multi-aspect-based Model,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Early Detection of Depression Severity Levels on Reddit using a Multi-aspect-based Model,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:00.292798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:58.452346Z digest=sha256:1fc214bd33e980b56eb36e2f09de555cb9d90a2aa814a4a8e583df07ddca6fb9

Observation c1393aab-2ac7-419d-9a35-fa20ba1fa548 · outbound

This paper cites Deep Learning for Suicidal Ideation Detection and Classification from Social Media Texts,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Deep Learning for Suicidal Ideation Detection and Classification from Social Media Texts,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:00.017728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:58.561269Z digest=sha256:6a8a3c85c85b942af2cf5d8a1a2609c546b728c020974f407ec6e0dc173920e0

Observation 5717e3a2-0cdd-4bb5-aea2-794249a99618 · outbound

This paper cites Knowledge-aware assessment of severity of suicide risk for early intervention,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Knowledge-aware assessment of severity of suicide risk for early intervention,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:08:59.651302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:58.739711Z digest=sha256:6d308f308d98663664fff561a251002f36b6e85692038b4db1b0ac2e47081900

Observation 291d6569-3c46-4c92-81aa-2470f7abd395 · outbound

This paper cites Qwen2.5-Omni Technical Report.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Qwen2.5-Omni Technical Report

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T19:08:58.873146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:08:58.873146Z digest=sha256:8b297275775ae080bd98eb8b1c98840666455d2ce4536c333a4a1260ffda4179

Observation 676a3f63-a633-4aa0-bf8d-354b810fa809 · outbound

This paper cites The Llama 3 Herd of Models.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding The Llama 3 Herd of Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T19:08:59.002718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:08:59.002718Z digest=sha256:da3709be085ca55c12e80b712ef444168a75e69718b17def6edb357b34e40e96

Pith citing papers

Observation fc558154-3f8c-421c-8323-3cdfaa4de00e · inbound

HealthSLM-Bench: Benchmarking Small Language Models for Mobile and Wearable Healthcare Monitoring cites this paper.

HealthSLM-Bench: Benchmarking Small Language Models for Mobile and Wearable Healthcare Monitoring Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-04T22:36:30.735420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:36:30.735420Z digest=sha256:abf8a617d1762ec66fead0204eb050d3b078034fd758a5f875364eaf794afd01

Observation 938cb96a-f2ed-49b2-807f-6d2308bc1532 · inbound

Titans-as-a-Layer: Test-Time Memory for Conversational Speech Emotion Recognition cites this paper.

Titans-as-a-Layer: Test-Time Memory for Conversational Speech Emotion Recognition Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-02T23:07:26.845155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T18:28:44.652813Z digest=sha256:2312836d9f435b2f24eb298bb98f5f10f8436180273645cb4ea002eebd8ae589