Pith. sign in

Paper Citation Record · LEDGER

Less is More: Accurate Speech Recognition & Translation without Web-Scale Data

As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2406.19674.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.19674 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T22:05:08.007118Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T08:54:29.724934Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c19323f2-420b-4a33-94e5-9a010a26fdd0 · inbound

OWLS: Scaling Laws for Multilingual Speech Recognition and Translation Models cites this paper.

OWLS: Scaling Laws for Multilingual Speech Recognition and Translation Models Less is More: Accurate Speech Recognition & Translation without Web-Scale Data

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T18:32:14.838590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T18:32:14.838590Z digest=sha256:1d246850d379b73283bda36d33a63ea507c51ce9353906cd98fa59fbd3d69f36

Observation 62578be3-561a-4009-a532-f413992788b2 · inbound

Unveiling the Best Practices for Applying Speech Foundation Models to Speech Intelligibility Prediction for Hearing-Impaired People cites this paper.

Unveiling the Best Practices for Applying Speech Foundation Models to Speech Intelligibility Prediction for Hearing-Impaired People Less is More: Accurate Speech Recognition & Translation without Web-Scale Data

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T22:05:08.007118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:05:08.007118Z digest=sha256:3eee392793fc14f956122bc9a39ba73baecfbb5fe381e0a030cf4099aad0ff5f

Observation 87f9833d-f831-4172-bac3-58b98ddb43d3 · inbound

Granary: Speech Recognition and Translation Dataset in 25 European Languages cites this paper.

Granary: Speech Recognition and Translation Dataset in 25 European Languages Less is More: Accurate Speech Recognition & Translation without Web-Scale Data

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T20:19:15.045498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:19:15.045498Z digest=sha256:7a4db8304e41c6d0904182f237bbd6c2c8be7747ea07beb0e531bc153bc90692

Observation c2246641-265a-4b2e-9ca7-b4fa9137736c · inbound

Word Level Timestamp Generation for Automatic Speech Recognition and Translation cites this paper.

Word Level Timestamp Generation for Automatic Speech Recognition and Translation Less is More: Accurate Speech Recognition & Translation without Web-Scale Data

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:34.078536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:34.078536Z digest=sha256:9aa5f70a7023ab7595aef11e47d9d443761ac233d3807dceeddbc0b16baaa3ff

Observation 99b7891b-7077-4c39-b587-7e6c1e7f1550 · inbound

From Tens of Hours to Tens of Thousands: Scaling Back-Translation for Speech Recognition cites this paper.

From Tens of Hours to Tens of Thousands: Scaling Back-Translation for Speech Recognition Less is More: Accurate Speech Recognition & Translation without Web-Scale Data

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:56.665186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:55:56.665186Z digest=sha256:cb176ffc81c0013ed3ff637e38016066767ee2169855704c7d21016e41fbc2d0

Observation 7e660ff5-0a7d-4bfe-af5a-21dc603fa87a · inbound

VietASR: Achieving Industry-level Vietnamese ASR with 50-hour labeled data and Large-Scale Speech Pretraining cites this paper.

VietASR: Achieving Industry-level Vietnamese ASR with 50-hour labeled data and Large-Scale Speech Pretraining Less is More: Accurate Speech Recognition & Translation without Web-Scale Data

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:40:56.020123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:40:56.020123Z digest=sha256:33510f6c75c4deea3348bfa9e233e24044b77252bef81abd2c56c630981a4793

Observation 088fb169-20b6-4e00-8d5c-f6aff80069df · inbound

SwitchLingua: The First Large-Scale Multilingual and Multi-Ethnic Code-Switching Dataset cites this paper.

SwitchLingua: The First Large-Scale Multilingual and Multi-Ethnic Code-Switching Dataset Less is More: Accurate Speech Recognition & Translation without Web-Scale Data

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:38:01.599022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:38:01.599022Z digest=sha256:811a4ab2de641d0be047337983f0d04f65de0941ca2c8cd0345d6f2a360dd8c2

Observation 86e970ca-86a7-4fdb-8db1-89a9f125a60d · inbound

SC-SOT: Conditioning the Decoder on Diarized Speaker Information for End-to-End Overlapped Speech Recognition cites this paper.

SC-SOT: Conditioning the Decoder on Diarized Speaker Information for End-to-End Overlapped Speech Recognition Less is More: Accurate Speech Recognition & Translation without Web-Scale Data

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T00:50:17.135316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:50:17.135316Z digest=sha256:d6ed1899d27d0f7532a25c48bb5e338e9c78409617af382ac8d98c9cbfecf129

Observation ee875b2f-4b94-4e81-95e7-236554a03d50 · inbound

LCS-CTC: Leveraging Soft Alignments to Enhance Phonetic Transcription Robustness cites this paper.

LCS-CTC: Leveraging Soft Alignments to Enhance Phonetic Transcription Robustness Less is More: Accurate Speech Recognition & Translation without Web-Scale Data

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T01:10:15.812317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T01:10:15.812317Z digest=sha256:934f2b689a3a9ee095e5a9c9481076549328423b7027573ff46610348e2c1519

Observation c9266033-efc4-4172-be3f-a11987ce6352 · inbound

Flavors of Moonshine: Tiny Specialized ASR Models for Edge Devices cites this paper.

Flavors of Moonshine: Tiny Specialized ASR Models for Edge Devices Less is More: Accurate Speech Recognition & Translation without Web-Scale Data

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T16:40:38.089835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:40:38.089835Z digest=sha256:196eccd9909c2999eebf54adb7c78aaca6da0504df092fd61968322788700679

Observation 65006ca9-d904-489e-a04b-4d48d134aebc · inbound

BlasBench: An Open Benchmark for Irish Speech Recognition cites this paper.

BlasBench: An Open Benchmark for Irish Speech Recognition Less is More: Accurate Speech Recognition & Translation without Web-Scale Data

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:41:01.629583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-05-10T15:53:54.092426Z digest=sha256:908ef32c2773b7081e7f815483f62a3c78e8dd4692499db243b5190290dbafd7

Observation ffe59950-8bbc-4965-b663-cf6402d36cfd · inbound

Frame-Aligned Fusion of Canary and WavLM for Non-Intrusive Intelligibility Prediction of Hearing-Aid-Processed Speech cites this paper.

Frame-Aligned Fusion of Canary and WavLM for Non-Intrusive Intelligibility Prediction of Hearing-Aid-Processed Speech Less is More: Accurate Speech Recognition & Translation without Web-Scale Data

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-25T02:35:14.579945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-25T02:34:11.297579Z digest=sha256:7735ff758d8cd37b025da7d4fa377eedbf4ca4651a1f1083f6d5fc29c559a4cb

Observation 78eaefbf-eff5-479a-be1b-f83ce363f2d6 · inbound

CTC-Seeded Token Edit Refinement for Non-Autoregressive Speech Recognition cites this paper.

CTC-Seeded Token Edit Refinement for Non-Autoregressive Speech Recognition Less is More: Accurate Speech Recognition & Translation without Web-Scale Data

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-06-30T08:54:29.728249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-30T08:52:12.463360Z digest=sha256:f51fa790cc9bdafe37eace368d0ecc48d381d5a365f9fc51cdcb1c462e696794