Pith. sign in

Scaling speech technology to 1,000+ languages

11 Pith papers cite this work, alongside 116 external citations. Polarity classification is still indexing.

11 Pith papers citing it
116 external citations · external index

representative citing papers

A framework for analyzing concept representations in neural models

cs.CL · 2026-05-02 · unverdicted · novelty 7.0

A new framework shows concept subspaces are not unique, estimator choice affects containment and disentanglement, LEACE works well but generalizes poorly, and HuBERT encodes phone info as contained and disentangled from speaker info while speaker info resists compact containment.

Tadabur: A Large-Scale Quran Audio Dataset

cs.SD · 2026-04-21 · unverdicted · novelty 7.0

Tadabur is a large-scale Quran audio dataset with over 1400 hours from 600+ reciters to support speech research and benchmarks.

BlasBench: An Open Benchmark for Irish Speech Recognition

cs.CL · 2026-04-12 · conditional · novelty 6.0

BlasBench supplies an Irish-aware normalizer and scoring harness that enables reproducible ASR comparisons and exposes a 33-43 point generalization gap for fine-tuned models versus 7-10 points for massively multilingual ones.

Coherence in the brain unfolds across separable temporal regimes

q-bio.NC · 2025-12-23 · conditional · novelty 6.0

Language coherence arises from slow contextual integration in default-mode cortex and rapid event-driven reconfiguration in auditory and language areas, captured by LLM-derived signals in single-subject fMRI.

MLAAD: The Multi-Language Audio Anti-Spoofing Dataset

cs.SD · 2024-01-17 · unverdicted · novelty 6.0

MLAAD provides a large-scale multi-language synthetic audio dataset for training and evaluating audio anti-spoofing models, showing better training performance than InTheWild and FakeOrReal and alternating superiority with ASVspoof 2019 across eight test sets.

citing papers explorer

Showing 11 of 11 citing papers.