Pith. sign in

Paper Citation Record · LEDGER

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts

As of 7 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 1 inbound Pith citation observation for arXiv:2508.10009.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.10009 v1

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T01:05:46.447560Z

measured 32 of 32 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T01:05:43.913842Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T01:05:46.601121Z

Reference resolution

31 of 31 outbound references displayed

  • verified exact0
  • verified fuzzy21
  • unresolved8
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 586c0fd9-9f91-47e6-b24d-e0a6ec76ebec · outbound

This paper cites Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T01:05:46.690100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T01:05:43.913842Z digest=sha256:bcbfd64ddad60c5ba4094f44b91b90de144a72fbe9ba54834675db785b9aed9b

Observation fbb2e7e8-46e7-41d7-b8d2-f8a1c289cea2 · outbound

This paper cites Further details are presented in the following subsections.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Further details are presented in the following subsections

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:50.422775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T01:05:43.958842Z digest=sha256:fc6a29b0df2d2ae2dc4b41c398a8c7101a0d5b93c6336580e4a123eebc9209ff

Observation e6ec2d64-78eb-446e-a424-3b027d711362 · outbound

This paper cites Datasets 3.1.1.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Datasets 3.1.1

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:50.305824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T01:05:44.019470Z digest=sha256:4a8626b6e4810a2e30a3a55f2eb3c80a2c7fa5016fe73bdfa78172d1500a275d

Observation 0aeff68f-136c-4610-8c23-4e739c4276f6 · outbound

This paper cites Decoder S-MoE To validate the effectiveness of S-MoE applied to the decoder blocks, we conducted experiments on two tasks: Korean-to- English (ko2en) ST and Korean ASR (ko-ASR).

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Decoder S-MoE To validate the effectiveness of S-MoE applied to the decoder blocks, we conducted experiments on two tasks: Korean-to- English (ko2en) ST and Korean ASR (ko-ASR)

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:50.012876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T01:05:44.179893Z digest=sha256:130a2dd3bf836c4ae2130361cdd0fced01be53aa389e34e7f3327229ca4eecc1

Observation 8aa07d1d-aefc-43ca-bda8-c453b431af92 · outbound

This paper cites By us- ing guiding tokens instead of dynamic gating functions, S-MoE ensures efficient training and inference while improving per- formance across various tasks.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts By us- ing guiding tokens instead of dynamic gating functions, S-MoE ensures efficient training and inference while improving per- formance across various tasks

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:49.696543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T01:05:44.304450Z digest=sha256:a029633ccdc46463834eb2d99b2145a0468e164031bd61df33cf739ed4712363

Observation d9e30dc1-fdae-407f-aa2d-6e8400a5ed5a · outbound

This paper cites Sources of degradation of speech recognition in the telephone network,.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Sources of degradation of speech recognition in the telephone network,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:49.372762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T01:05:44.870745Z digest=sha256:bd188861c6529078bb3f9d533aef11623cdaeed9b5d077aece7ded281631829c

Observation 52a8f881-d60d-4db8-ad6b-d52e9d955011 · outbound

This paper cites Training wideband acoustic mod- els using mixed-bandwidth training data for speech recognition,.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Training wideband acoustic mod- els using mixed-bandwidth training data for speech recognition,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:49.194971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T01:05:44.974020Z digest=sha256:b2b1f6150b112b06988de9ce598b50f850c043d6df55ef79c447781bd0f2eb57

Observation ad7de39c-ef70-43ce-85c2-34c2e3cdf340 · outbound

This paper cites Multi-Task Learning with Deep Neural Networks: A Survey.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Multi-Task Learning with Deep Neural Networks: A Survey

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T01:05:44.392830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T01:05:44.392830Z digest=sha256:c87cbaec2ab21ae0d26c945b97fe1652416a18b95af3b63294a1cfef6c41d089

Observation 8b4dfa92-13ca-4b43-b161-c2d33b5e5357 · outbound

This paper cites An Overview of Multi-Task Learning in Deep Neural Networks.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts An Overview of Multi-Task Learning in Deep Neural Networks

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T01:05:44.487656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T01:05:44.487656Z digest=sha256:4c417373b7e608f2935c9a8bc6b7989a9391a739214e737e0a860b3477668772

Observation 3df0ff5c-f5c7-4ec8-9e47-2ad73d12ec50 · outbound

This paper cites A survey on multi-task learning ieee transactions on knowledge and data engineering,.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts A survey on multi-task learning ieee transactions on knowledge and data engineering,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:49.508010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T01:05:44.579694Z digest=sha256:5d6c453d908ca66a89b4cf289cfa46ce334c7ae82817e5035e250ade48eea7c8

Observation 13ecb848-b814-4af7-853b-f9a0e5895d14 · outbound

This paper cites Sparsely Activated Mixture-of-Experts are Robust Multi-Task Learners.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Sparsely Activated Mixture-of-Experts are Robust Multi-Task Learners

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T01:05:44.679323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T01:05:44.679323Z digest=sha256:f10ac6b4799ad6ea3c7adba3f918045a9ed697287db8aee6a99aa90d019eb14d

Observation b14f26aa-863c-4f9f-8fb7-d577a9f976d5 · outbound

This paper cites Outrageously Large Neural Networks: The Sparsely-Gated Mixture-of-Experts Layer.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Outrageously Large Neural Networks: The Sparsely-Gated Mixture-of-Experts Layer

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T01:05:44.775645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T01:05:44.775645Z digest=sha256:74f33da68abf856a6934462921fc2d6f246fe1dd5cc26a58f1807fa70073dbed

Observation acd40c53-6442-45f1-b63d-2d5296943386 · outbound

This paper cites [Online].

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts [Online]

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:48.012664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T01:05:45.460528Z digest=sha256:43c54d80e1b974a046d59c8b2c2552a3d45cef29c7c00444ebfe3ef0d7aca44c

Observation 1ffa1fea-0693-407b-934a-1c9022565ecd · outbound

This paper cites [Online].

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts [Online]

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:47.877416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T01:05:45.556818Z digest=sha256:052da603aaf15a4c1040710d2c312f9f68569fb4b0c5ff3c829f16fc91777bab

Observation eaeb04fd-c4b1-4509-9e5f-27669a752e2d · outbound

This paper cites [Online].

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts [Online]

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:49.060196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T01:05:45.073418Z digest=sha256:9a451ddf06ebc2b59c2a359247e0739cff8677af81fda1171a1eff6b0aaf716d

Observation bed821b7-7dc2-48d8-8567-32de69b8f85a · outbound

This paper cites [Online].

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts [Online]

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:48.848835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T01:05:45.170383Z digest=sha256:6b3e5026b97b0943bd21659edb07669484baba81fc8e552b61d4eb551685576d

Observation f46f5ba2-6f3d-4384-b665-a628fc6e8ef8 · outbound

This paper cites [Online].

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts [Online]

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:48.642320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T01:05:45.264475Z digest=sha256:ee7ce2b11514c9e643a439eaed5903a0a0e3303c7cf233576ed6816a0c1866b2

Observation d5367c58-a785-4998-84df-386c5f577384 · outbound

This paper cites [Online].

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts [Online]

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:48.440198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T01:05:45.321574Z digest=sha256:7beccaa9af5210a0a981e184e41b4b6101667be9dcb0bbbb3987c455e771a3e6

Observation db6ba20b-e15a-44fb-8a18-ec018b8b582e · outbound

This paper cites [Online].

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts [Online]

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:48.164816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T01:05:45.394744Z digest=sha256:cf94a5c64d7ba98c24f9851b8be3691de7d66144c73017376fffed7e80595374

Observation 072708d6-4e41-4654-a0ed-d3ab503c465a · outbound

This paper cites Pulse code modulation (pcm) of voice fre- quencies,.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Pulse code modulation (pcm) of voice fre- quencies,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:46.982149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T01:05:46.071817Z digest=sha256:26aa7acb2b09d4637ebb06cfbd0efcb6956c2c6649f2f5180d00ae7484664e56

Observation 33a791ee-bc8a-4d32-8f6d-292134b25192 · outbound

This paper cites Our in-house test sets consist of 1,000 samples of male and female speech from daily conversations, with refer- ence translations curated by professional translators.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Our in-house test sets consist of 1,000 samples of male and female speech from daily conversations, with refer- ence translations curated by professional translators

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:50.162275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T01:05:44.122596Z digest=sha256:9c39a1ff133f7edd5cf03e6b66dc2cae6dbb6ab103af7ce69cf9d767d4ba676f

Observation aa6474cd-aa62-4796-975f-820f642c40e4 · outbound

This paper cites [Online].

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts [Online]

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:47.654187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T01:05:45.618663Z digest=sha256:0ad64b06a73dda5855bc2e04c94f0bc0ff691d9bb78209cf849b44d1a202cab4

Observation a1f6a03c-ea88-490c-be24-e1e607f20541 · outbound

This paper cites [Online].

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts [Online]

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:47.474698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T01:05:45.683072Z digest=sha256:a490ddc351903ccab7b4fde135d86272c01541d3b9fb37f2bf6455aa2872fa6f

Observation 38a3220f-17f0-4c27-8d6f-3a9caec67139 · outbound

This paper cites Whisper is a widely used mul- tilingual speech model trained with diverse language pairs.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Whisper is a widely used mul- tilingual speech model trained with diverse language pairs

Reference 24

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T01:05:49.880492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T01:05:44.239321Z digest=sha256:c96ce3ff1827bcb35903632a29be40b2b8bbb4745e25fe1a7048e947f117766e

Observation 49206302-ec76-4c0c-8fc2-e0053d1bb54b · outbound

This paper cites [Online].

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts [Online]

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:47.280919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T01:05:45.783547Z digest=sha256:93cb93ebe619ea4a1176e6314713ea7bbb96813f482e00ac9066d598d44bab67

Observation c4cad7bc-fb68-44ec-99b9-fde756a834a2 · outbound

This paper cites [Online].

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts [Online]

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:47.116778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T01:05:45.863666Z digest=sha256:86b2df0d74be8cef2722eeac65264d310b66dbf3d463bb80fcbfc4bb0fef9cfc

Observation e2fefd12-a43e-47c1-9e49-927204004dd2 · outbound

This paper cites The adaptive multirate wideband speech codec (amr-wb),.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts The adaptive multirate wideband speech codec (amr-wb),

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T01:05:45.952971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T01:05:45.952971Z digest=sha256:423d29929945bf8fdc3aef5e9de5ca0b53505665a3f66c0c76dbe9501b12e53b

Observation ad761f03-d506-4a96-bd8f-686fc5e454ef · outbound

This paper cites Fleurs: Few-shot learning evaluation of universal representations of speech,.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Fleurs: Few-shot learning evaluation of universal representations of speech,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T01:05:46.175496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T01:05:46.175496Z digest=sha256:9c45806a2c70de6f8293ef1e2ebddec38ec6fe6842fd89cbb6a9531107eff4a4

Observation 29b34952-b47b-4dee-9584-9a971a3bf21d · outbound

This paper cites Attention is all you need,.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Attention is all you need,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:46.876936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T01:05:46.281915Z digest=sha256:3fe7fd2548bf164791b91a2967d9db2cba3e1b8bc390396aecbd2af1972545e2

Observation bd80abc3-b9aa-49ce-9548-76222386d705 · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation,.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Bleu: a method for automatic evaluation of machine translation,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T01:05:46.369512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T01:05:46.369512Z digest=sha256:c8f730464d02b60c4f941409653d913ba4ef145d6572689e4db2b5f9e87c1f97

Observation 2be105f0-24c8-4d44-b516-90f8e82b548b · outbound

This paper cites Robust speech recognition via large-scale weak supervision,.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Robust speech recognition via large-scale weak supervision,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T01:05:46.447560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T01:05:46.447560Z digest=sha256:01fbd51c5e6b4aa6a296eada34c7ee3158d0cdc89d04cd2bfb9225b5ade9f224

Pith citing papers

Observation 586c0fd9-9f91-47e6-b24d-e0a6ec76ebec · inbound

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts cites this paper.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T01:05:46.690100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T01:05:43.913842Z digest=sha256:bcbfd64ddad60c5ba4094f44b91b90de144a72fbe9ba54834675db785b9aed9b