Pith. sign in

Paper Citation Record · LEDGER

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding

As of 18 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 2 inbound Pith citation observations for arXiv:2507.08031.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.08031 v2

Coverage vector

measured 26 of 26 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:08:59.002718Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T22:36:30.735420Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T23:07:26.843285Z

Reference resolution

26 of 26 outbound references displayed

  • verified exact1
  • verified fuzzy19
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 340f2f81-bc9b-4e8e-8c9f-16712fde00ca · outbound

This paper cites The global economic burden of noncommunicable diseases,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding The global economic burden of noncommunicable diseases,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:05.516187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:08:55.374946Z digest=sha256:5494b651d4fecb7386f88b9ac80f548004e04641e621f9e7557c2cb404bfedca

Observation ac2ee850-762c-4aa8-8611-02cb2e2cbe99 · outbound

This paper cites The State of Mental Health in America 2024,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding The State of Mental Health in America 2024,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:05.220131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:08:55.536436Z digest=sha256:97e49c25865b9d98c6c475a61b466e757312a76f86a0342348986dd971f8a4aa

Observation ac6d70c3-97d2-46e4-9bc7-c3dfe9647294 · outbound

This paper cites How mental health care should change as a consequence of the COVID-19 pandemic,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding How mental health care should change as a consequence of the COVID-19 pandemic,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:04.823250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:08:55.674744Z digest=sha256:676d1b7290ece78ada0cbecf3393564caf3a79a7dec64742114061803b58d985

Observation cf8a9203-3ef4-42f9-ac90-1cfefefd974b · outbound

This paper cites Social media and mental health: benefits, risks, and opportunities for research and practice,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Social media and mental health: benefits, risks, and opportunities for research and practice,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:04.443059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:08:55.855024Z digest=sha256:3e17bec25abf76ea203374feb7052504f9b89150be306873bf7d8a094b09e2a9

Observation ed8ba62f-76f1-460d-8fdf-8bf111c002cd · outbound

This paper cites Dang, et al.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Dang, et al

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:04.117800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:08:55.986628Z digest=sha256:bad111f09536a7a14c916c676b503052eb76e6ee39754bb6fb98ead038820aed

Observation bbcaa6df-4ea7-40b5-b83f-f99f87c2a3e1 · outbound

This paper cites Mental-llm: Leveraging large language models for mental health prediction via online text data,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Mental-llm: Leveraging large language models for mental health prediction via online text data,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:03.759548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:08:56.164960Z digest=sha256:0d5be8cb5a11da5ffb5576df802bb8be2a079e0a1ba5cb95df3b95f8ca66133f

Observation df586b4e-7273-459e-baa1-82a4c6c9d3b0 · outbound

This paper cites A taxonomy of ethical tensions in inferring mental health states from social media,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding A taxonomy of ethical tensions in inferring mental health states from social media,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:03.410345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:08:56.349071Z digest=sha256:f4b66252437059a1044f99f9bc3700d15f3cf8bde88465ca675e3e5ec0f074f1

Observation 232ef0ce-4eb6-4ad8-a40e-6b0ddb21574f · outbound

This paper cites Beyond LDA: exploring supervised topic modeling for depression-related language in Twitter,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Beyond LDA: exploring supervised topic modeling for depression-related language in Twitter,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:03.038047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:08:56.489214Z digest=sha256:b031578fd07b6e36c98e0179c265cedd1c6928aad0b6fb48b489e8e7fc449f64

Observation 6a88d9a5-b0a4-4b26-a13b-a651dbc3ec61 · outbound

This paper cites AER-LLM: Ambiguity-aware emotion recognition leveraging large language models,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding AER-LLM: Ambiguity-aware emotion recognition leveraging large language models,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:02.718421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:08:56.606960Z digest=sha256:cd002d1a773d7bbc12b4359bc650a81be6b19ca5a84c33bac1565fee7d3e41ff

Observation d0cf2ed4-002f-405d-88a0-61037b8c5166 · outbound

This paper cites Token-Level Logits Matter: A Closer Look at Speech Foundation Models for Ambiguous Emotion Recognition.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Token-Level Logits Matter: A Closer Look at Speech Foundation Models for Ambiguous Emotion Recognition

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-06T19:08:59.364163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:08:56.780027Z digest=sha256:7faeb584bc3b1fbc2ee8f98b4b916ae940e62b11a98cee0bb56a4cbfd72a9b3c

Observation a30306b3-31d2-4bd9-9826-d96e1df3cc58 · outbound

This paper cites Privacy-preserving deep learning,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Privacy-preserving deep learning,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:02.377071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:08:56.960705Z digest=sha256:b9051df6997998283f612cf9ff22b93f9f3b55268ff4c8373355c271182c8c30

Observation a42e6388-c2e0-436c-ae8a-ee4d146ba5ed · outbound

This paper cites ”Efficient and personalized mobile health event prediction via small language models.” Proceedings of the 30th Annual International Conference on Mobile Computing and Networking.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding ”Efficient and personalized mobile health event prediction via small language models.” Proceedings of the 30th Annual International Conference on Mobile Computing and Networking

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:02.076532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:08:57.110786Z digest=sha256:238be44bf600303295c7de154ac6845e141d95afc18a2beea22b8b4ad16049d7

Observation 1f9d8cff-568d-4c06-b2ae-4bd1ed3793b9 · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T19:08:57.240115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:08:57.240115Z digest=sha256:64ab50c8dc8d4d374798e353c7d30464643c12a55bcd849d6bac445d72bd4e3d

Observation a65a9797-5452-4623-b3fe-cbf952fb2d17 · outbound

This paper cites Gemma: Open Models Based on Gemini Research and Technology.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Gemma: Open Models Based on Gemini Research and Technology

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T19:08:57.400446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:08:57.400446Z digest=sha256:a868e466f3243b5422512c69a44720e6496f73017d8629164668d97b3bae6d94

Observation 8166ceba-ab53-44dd-b316-1fe5fb4bfdb3 · outbound

This paper cites Federated Learning for Mobile Keyboard Prediction.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Federated Learning for Mobile Keyboard Prediction

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T19:08:57.534782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:08:57.534782Z digest=sha256:29e2ec1d81b5ae2c0837d625e6c22135c69976cd6047e206dfe5e141dbd5849b

Observation 0b7885b5-bb66-4003-804b-31905044fb24 · outbound

This paper cites A primer on neural network models for natural language processing,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding A primer on neural network models for natural language processing,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:01.809452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:08:57.664397Z digest=sha256:7cf7ad15877ca2cac42854e4d650bb094029ed34c92c1c4a1bb8ecfe268260ae

Observation 7687ee98-4b31-467c-b682-31924f681c16 · outbound

This paper cites On the dangers of stochastic parrots: Can language models be too big?,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding On the dangers of stochastic parrots: Can language models be too big?,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:01.492345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:08:57.830414Z digest=sha256:c7567a8bd602381d86cfda53a3713339f963ce3190867d15a638caae561401ce

Observation d5a478d2-ff6e-4e7b-aa3e-80cdef78c9c5 · outbound

This paper cites A discourse-aware attention model for abstractive summarization of long documents,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding A discourse-aware attention model for abstractive summarization of long documents,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:01.162022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:08:57.987186Z digest=sha256:9b5d140dd30e5fdbc5a96c522ffc0f19222cac59c4364d0e84fd458994a6d860

Observation ffc4567f-fd2d-442e-a9c4-3989fa3b3b58 · outbound

This paper cites Language models are few-shot learners,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Language models are few-shot learners,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:00.910111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:08:58.082008Z digest=sha256:e92e4a17d6663127a3417c1938ed2678f24f1bb86b379afd5cb696cc99c12294

Observation 7e54fe16-3fee-41c6-9104-bde0ff0e3067 · outbound

This paper cites CLPsych 2019 shared task: Predicting the degree of suicide risk in Reddit posts,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding CLPsych 2019 shared task: Predicting the degree of suicide risk in Reddit posts,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:00.546999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:08:58.196456Z digest=sha256:e83de0ad932e56f33b1538fd39fa553cc202ffc660a6ef7c885ace6a304f3e8d

Observation 801508b4-45de-4c0b-856c-f66ca3c56fe2 · outbound

This paper cites Dreaddit: A Reddit Dataset for Stress Analysis in Social Media.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Dreaddit: A Reddit Dataset for Stress Analysis in Social Media

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T19:08:58.333008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:08:58.333008Z digest=sha256:3b8747fdea6f53d34ffd5bf950b9e6c27855a4cab72ee1a4afa7c2159bb0475a

Observation 18cfa444-353b-4a67-a39f-244094fce3d6 · outbound

This paper cites Early Detection of Depression Severity Levels on Reddit using a Multi-aspect-based Model,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Early Detection of Depression Severity Levels on Reddit using a Multi-aspect-based Model,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:00.292798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:08:58.452346Z digest=sha256:b0f4cf99058c6177c25415c0fb51e31352228108af909c97bbd56c5306c188ea

Observation c1393aab-2ac7-419d-9a35-fa20ba1fa548 · outbound

This paper cites Deep Learning for Suicidal Ideation Detection and Classification from Social Media Texts,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Deep Learning for Suicidal Ideation Detection and Classification from Social Media Texts,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:00.017728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:08:58.561269Z digest=sha256:bf1b89c4eb468e0b6e934baafe00c1625207a2c820e6890af2aff98fe46eb2de

Observation 5717e3a2-0cdd-4bb5-aea2-794249a99618 · outbound

This paper cites Knowledge-aware assessment of severity of suicide risk for early intervention,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Knowledge-aware assessment of severity of suicide risk for early intervention,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:08:59.651302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T19:08:58.739711Z digest=sha256:d40103a39bd2fdc0d2e7a362a7a9c683548cd3ba3dc357ab8a8326b028bfdf9e

Observation 291d6569-3c46-4c92-81aa-2470f7abd395 · outbound

This paper cites Qwen2.5-Omni Technical Report.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Qwen2.5-Omni Technical Report

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T19:08:58.873146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:08:58.873146Z digest=sha256:9bbfa63191748c2762d8a35be5e8fbb3a64d3a43c73cdc4ed33be9bf00b2ee72

Observation 676a3f63-a633-4aa0-bf8d-354b810fa809 · outbound

This paper cites The Llama 3 Herd of Models.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding The Llama 3 Herd of Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T19:08:59.002718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:08:59.002718Z digest=sha256:95f89d41ac66395e57e4318a7a2cc328adf4a15c1d9cbb2df46b5df016c74593

Pith citing papers

Observation fc558154-3f8c-421c-8323-3cdfaa4de00e · inbound

HealthSLM-Bench: Benchmarking Small Language Models for Mobile and Wearable Healthcare Monitoring cites this paper.

HealthSLM-Bench: Benchmarking Small Language Models for Mobile and Wearable Healthcare Monitoring Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-04T22:36:30.735420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:36:30.735420Z digest=sha256:8fa159f1c035d562bdf2d06a72cac0cc8b442b0c8e91b7a0a9ef2c928cd2d28e

Observation 938cb96a-f2ed-49b2-807f-6d2308bc1532 · inbound

Titans-as-a-Layer: Test-Time Memory for Conversational Speech Emotion Recognition cites this paper.

Titans-as-a-Layer: Test-Time Memory for Conversational Speech Emotion Recognition Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-02T23:07:26.845155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-27T18:28:44.652813Z digest=sha256:5a7f672f75d5f88a8d447f34200cfc65537c46d357b8096262271394fe6817e0