Pith. sign in

Paper Citation Record · LEDGER

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding

As of 9 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 2 inbound Pith citation observations for arXiv:2507.08031.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.08031 v2

Coverage vector

measured 26 of 26 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:08:59.002718Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T22:36:30.735420Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T23:07:26.843285Z

Reference resolution

26 of 26 outbound references displayed

  • verified exact1
  • verified fuzzy19
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 340f2f81-bc9b-4e8e-8c9f-16712fde00ca · outbound

This paper cites The global economic burden of noncommunicable diseases,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding The global economic burden of noncommunicable diseases,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:05.516187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:55.374946Z digest=sha256:e99a7f6cb3cd0759fee0caa1d0f93d12c204b42bfbe4dbd007c3890cac834b18

Observation ac2ee850-762c-4aa8-8611-02cb2e2cbe99 · outbound

This paper cites The State of Mental Health in America 2024,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding The State of Mental Health in America 2024,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:05.220131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:55.536436Z digest=sha256:9a7b79e1704fbc4353672cd3591f79c2fa8357aebc5ce39c88a2e6d4996073a6

Observation ac6d70c3-97d2-46e4-9bc7-c3dfe9647294 · outbound

This paper cites How mental health care should change as a consequence of the COVID-19 pandemic,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding How mental health care should change as a consequence of the COVID-19 pandemic,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:04.823250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:55.674744Z digest=sha256:03675150e26cb27c089d5dd9bc6d6277dcc15496efa8731253f5db47e58c97f3

Observation cf8a9203-3ef4-42f9-ac90-1cfefefd974b · outbound

This paper cites Social media and mental health: benefits, risks, and opportunities for research and practice,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Social media and mental health: benefits, risks, and opportunities for research and practice,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:04.443059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:55.855024Z digest=sha256:00bfda1ac4a07b98560faed7da496d71210928b39274517a445e1dea73639047

Observation ed8ba62f-76f1-460d-8fdf-8bf111c002cd · outbound

This paper cites Dang, et al.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Dang, et al

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:04.117800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:55.986628Z digest=sha256:f6f2c725b9252e6c6c431ad35cfdc4c402ed105dceedc8a7b8dd9d72d5087423

Observation bbcaa6df-4ea7-40b5-b83f-f99f87c2a3e1 · outbound

This paper cites Mental-llm: Leveraging large language models for mental health prediction via online text data,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Mental-llm: Leveraging large language models for mental health prediction via online text data,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:03.759548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:56.164960Z digest=sha256:1e35378b786ed03a8e12f9495ee112a4fec9a4c08375c3d836d355277389ff2f

Observation df586b4e-7273-459e-baa1-82a4c6c9d3b0 · outbound

This paper cites A taxonomy of ethical tensions in inferring mental health states from social media,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding A taxonomy of ethical tensions in inferring mental health states from social media,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:03.410345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:56.349071Z digest=sha256:df5182fa0fb0389e84754264ed479c9b8d2b58f468686e8f03d09073a91d2dd6

Observation 232ef0ce-4eb6-4ad8-a40e-6b0ddb21574f · outbound

This paper cites Beyond LDA: exploring supervised topic modeling for depression-related language in Twitter,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Beyond LDA: exploring supervised topic modeling for depression-related language in Twitter,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:03.038047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:56.489214Z digest=sha256:185f94b32fec0356183717b6cbc6a8c6b5a9c575e80ffa95664a9391e9af80d1

Observation 6a88d9a5-b0a4-4b26-a13b-a651dbc3ec61 · outbound

This paper cites AER-LLM: Ambiguity-aware emotion recognition leveraging large language models,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding AER-LLM: Ambiguity-aware emotion recognition leveraging large language models,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:02.718421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:56.606960Z digest=sha256:9899d231c2c10200fb19808702eb6e90fa4a95d954a846aee4a5c6b476670288

Observation d0cf2ed4-002f-405d-88a0-61037b8c5166 · outbound

This paper cites Token-Level Logits Matter: A Closer Look at Speech Foundation Models for Ambiguous Emotion Recognition.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Token-Level Logits Matter: A Closer Look at Speech Foundation Models for Ambiguous Emotion Recognition

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-06T19:08:59.364163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:56.780027Z digest=sha256:44a57c374d2d0e316f99492f4e0a698678ddfcee332b83629f729288990fd150

Observation a30306b3-31d2-4bd9-9826-d96e1df3cc58 · outbound

This paper cites Privacy-preserving deep learning,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Privacy-preserving deep learning,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:02.377071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:56.960705Z digest=sha256:22bfa55ea0c6e1349f62ce80c4d675a531b00abccf03439643eb30f3db1a6c23

Observation a42e6388-c2e0-436c-ae8a-ee4d146ba5ed · outbound

This paper cites ”Efficient and personalized mobile health event prediction via small language models.” Proceedings of the 30th Annual International Conference on Mobile Computing and Networking.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding ”Efficient and personalized mobile health event prediction via small language models.” Proceedings of the 30th Annual International Conference on Mobile Computing and Networking

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:02.076532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:57.110786Z digest=sha256:08673dfafe2d8cf8b1bc24149b8354a227663920e08c90827155dc58315db750

Observation 1f9d8cff-568d-4c06-b2ae-4bd1ed3793b9 · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T19:08:57.240115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:08:57.240115Z digest=sha256:06ab9eb1aa71751249c8d4fa604bb68487f76031c0746a5845e44082d0d36a8e

Observation a65a9797-5452-4623-b3fe-cbf952fb2d17 · outbound

This paper cites Gemma: Open Models Based on Gemini Research and Technology.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Gemma: Open Models Based on Gemini Research and Technology

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T19:08:57.400446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:08:57.400446Z digest=sha256:f6ecb48a1e5ed939deb0f350d35cead419c057d236641a9bc1fce04383c56f19

Observation 8166ceba-ab53-44dd-b316-1fe5fb4bfdb3 · outbound

This paper cites Federated Learning for Mobile Keyboard Prediction.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Federated Learning for Mobile Keyboard Prediction

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T19:08:57.534782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:08:57.534782Z digest=sha256:d8b1b1706b300bf957c9ba4322061efdfafa42d66ffcca532ac7cf270281048a

Observation 0b7885b5-bb66-4003-804b-31905044fb24 · outbound

This paper cites A primer on neural network models for natural language processing,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding A primer on neural network models for natural language processing,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:01.809452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:57.664397Z digest=sha256:90dc366ce41f471fd1de46bc00af6bd7df19d7791dec8e0392099cc3439605f8

Observation 7687ee98-4b31-467c-b682-31924f681c16 · outbound

This paper cites On the dangers of stochastic parrots: Can language models be too big?,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding On the dangers of stochastic parrots: Can language models be too big?,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:01.492345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:57.830414Z digest=sha256:d0eeb63a92dd272f9c20626896d5c112077012a7a119952146ec0943839cebee

Observation d5a478d2-ff6e-4e7b-aa3e-80cdef78c9c5 · outbound

This paper cites A discourse-aware attention model for abstractive summarization of long documents,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding A discourse-aware attention model for abstractive summarization of long documents,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:01.162022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:57.987186Z digest=sha256:11c65191745e499c8540ad86852fa142214120565827d2bfc46ae416df26e15a

Observation ffc4567f-fd2d-442e-a9c4-3989fa3b3b58 · outbound

This paper cites Language models are few-shot learners,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Language models are few-shot learners,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:00.910111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:58.082008Z digest=sha256:69635a9bc0de906d72bde787f7f157cd2f59bd62ddd7e82c27411f172743a47d

Observation 7e54fe16-3fee-41c6-9104-bde0ff0e3067 · outbound

This paper cites CLPsych 2019 shared task: Predicting the degree of suicide risk in Reddit posts,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding CLPsych 2019 shared task: Predicting the degree of suicide risk in Reddit posts,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:00.546999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:58.196456Z digest=sha256:8af11ac87033cdb3b5f57c69ef9444efab8b1791184b42eed676fb08e5f9837a

Observation 801508b4-45de-4c0b-856c-f66ca3c56fe2 · outbound

This paper cites Dreaddit: A Reddit Dataset for Stress Analysis in Social Media.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Dreaddit: A Reddit Dataset for Stress Analysis in Social Media

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T19:08:58.333008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:08:58.333008Z digest=sha256:771c72ce9a5144f518efd42d548365a127c6dcf3f11668db4b2ee49a368a651e

Observation 18cfa444-353b-4a67-a39f-244094fce3d6 · outbound

This paper cites Early Detection of Depression Severity Levels on Reddit using a Multi-aspect-based Model,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Early Detection of Depression Severity Levels on Reddit using a Multi-aspect-based Model,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:00.292798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:58.452346Z digest=sha256:733fb991daddf717ceec5a94ced983a4d00515af49881cdedb388235e059d537

Observation c1393aab-2ac7-419d-9a35-fa20ba1fa548 · outbound

This paper cites Deep Learning for Suicidal Ideation Detection and Classification from Social Media Texts,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Deep Learning for Suicidal Ideation Detection and Classification from Social Media Texts,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:09:00.017728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:58.561269Z digest=sha256:6449377761a6d96b4bb7ef6bacd270d0dd1cc1a69f612a41178f2b8e0bfd81e5

Observation 5717e3a2-0cdd-4bb5-aea2-794249a99618 · outbound

This paper cites Knowledge-aware assessment of severity of suicide risk for early intervention,.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Knowledge-aware assessment of severity of suicide risk for early intervention,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:08:59.651302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:08:58.739711Z digest=sha256:fedcf53f308945e790a294d4c3d8c0379157dd0c0dd436607746909552d87dd9

Observation 291d6569-3c46-4c92-81aa-2470f7abd395 · outbound

This paper cites Qwen2.5-Omni Technical Report.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding Qwen2.5-Omni Technical Report

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T19:08:58.873146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:08:58.873146Z digest=sha256:18a105a3c4cfc45a5c1f85c4a66289fb002a6aa32c6b543488410fd48b1061fb

Observation 676a3f63-a633-4aa0-bf8d-354b810fa809 · outbound

This paper cites The Llama 3 Herd of Models.

Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding The Llama 3 Herd of Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T19:08:59.002718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:08:59.002718Z digest=sha256:8b6de6ef9511e54701d9393a4e8c19eeee373abd461f9e0b474d5a79e1523c07

Pith citing papers

Observation fc558154-3f8c-421c-8323-3cdfaa4de00e · inbound

HealthSLM-Bench: Benchmarking Small Language Models for Mobile and Wearable Healthcare Monitoring cites this paper.

HealthSLM-Bench: Benchmarking Small Language Models for Mobile and Wearable Healthcare Monitoring Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-04T22:36:30.735420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:36:30.735420Z digest=sha256:a3e7324b2305814dfea531f69d45461622ea7116ced4e387c2d6ac092e385402

Observation 938cb96a-f2ed-49b2-807f-6d2308bc1532 · inbound

Titans-as-a-Layer: Test-Time Memory for Conversational Speech Emotion Recognition cites this paper.

Titans-as-a-Layer: Test-Time Memory for Conversational Speech Emotion Recognition Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-02T23:07:26.845155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T18:28:44.652813Z digest=sha256:4565a30ce34ae3101231df6fc979e9de779043d586eddde7e937b914862ca0f0