Pith. sign in

Paper Citation Record · LEDGER

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts

As of 7 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 1 inbound Pith citation observation for arXiv:2508.10009.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.10009 v1

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T01:05:46.447560Z

measured 32 of 32 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T01:05:43.913842Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T01:05:46.601121Z

Reference resolution

31 of 31 outbound references displayed

  • verified exact0
  • verified fuzzy21
  • unresolved8
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 586c0fd9-9f91-47e6-b24d-e0a6ec76ebec · outbound

This paper cites Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T01:05:46.690100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T01:05:43.913842Z digest=sha256:77c0efebd6ad45e9b5e4fd6a0aba6ba866c965b69a38f0bbfea9d77203798c54

Observation fbb2e7e8-46e7-41d7-b8d2-f8a1c289cea2 · outbound

This paper cites Further details are presented in the following subsections.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Further details are presented in the following subsections

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:50.422775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T01:05:43.958842Z digest=sha256:05169ef847a5cddddba1191006d76eda888c7a851ca4b0336b6f75ee0780805b

Observation e6ec2d64-78eb-446e-a424-3b027d711362 · outbound

This paper cites Datasets 3.1.1.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Datasets 3.1.1

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:50.305824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T01:05:44.019470Z digest=sha256:362a33b4f9390c1d035deb7361298ce583c477da8fe31aa19b14e1bce6034bc2

Observation 0aeff68f-136c-4610-8c23-4e739c4276f6 · outbound

This paper cites Decoder S-MoE To validate the effectiveness of S-MoE applied to the decoder blocks, we conducted experiments on two tasks: Korean-to- English (ko2en) ST and Korean ASR (ko-ASR).

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Decoder S-MoE To validate the effectiveness of S-MoE applied to the decoder blocks, we conducted experiments on two tasks: Korean-to- English (ko2en) ST and Korean ASR (ko-ASR)

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:50.012876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T01:05:44.179893Z digest=sha256:5aa9fc79e912a6e5a20fbc9e1979841b7c52841bbe459b4db7879a99a3145453

Observation 8aa07d1d-aefc-43ca-bda8-c453b431af92 · outbound

This paper cites By us- ing guiding tokens instead of dynamic gating functions, S-MoE ensures efficient training and inference while improving per- formance across various tasks.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts By us- ing guiding tokens instead of dynamic gating functions, S-MoE ensures efficient training and inference while improving per- formance across various tasks

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:49.696543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T01:05:44.304450Z digest=sha256:1bb1ecee4b9419133342ab36d35dbec79628e1bd34458801a2775bb18f3030a5

Observation d9e30dc1-fdae-407f-aa2d-6e8400a5ed5a · outbound

This paper cites Sources of degradation of speech recognition in the telephone network,.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Sources of degradation of speech recognition in the telephone network,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:49.372762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T01:05:44.870745Z digest=sha256:db70eaecaa75e4765430d21857f818dd161908ab6c46f80953bd42719637b984

Observation 52a8f881-d60d-4db8-ad6b-d52e9d955011 · outbound

This paper cites Training wideband acoustic mod- els using mixed-bandwidth training data for speech recognition,.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Training wideband acoustic mod- els using mixed-bandwidth training data for speech recognition,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:49.194971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T01:05:44.974020Z digest=sha256:2560c898da5c5f1842c8b8236065f05751d1eb39f501f4a5f33e8c13ea1f4b7d

Observation ad7de39c-ef70-43ce-85c2-34c2e3cdf340 · outbound

This paper cites Multi-Task Learning with Deep Neural Networks: A Survey.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Multi-Task Learning with Deep Neural Networks: A Survey

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T01:05:44.392830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T01:05:44.392830Z digest=sha256:c87cbaec2ab21ae0d26c945b97fe1652416a18b95af3b63294a1cfef6c41d089

Observation 8b4dfa92-13ca-4b43-b161-c2d33b5e5357 · outbound

This paper cites An Overview of Multi-Task Learning in Deep Neural Networks.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts An Overview of Multi-Task Learning in Deep Neural Networks

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T01:05:44.487656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T01:05:44.487656Z digest=sha256:4c417373b7e608f2935c9a8bc6b7989a9391a739214e737e0a860b3477668772

Observation 3df0ff5c-f5c7-4ec8-9e47-2ad73d12ec50 · outbound

This paper cites A survey on multi-task learning ieee transactions on knowledge and data engineering,.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts A survey on multi-task learning ieee transactions on knowledge and data engineering,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:49.508010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T01:05:44.579694Z digest=sha256:d25199d295e9489cc910c0392a4c556b38f91504fdf0d4964fbc8134757638e2

Observation 13ecb848-b814-4af7-853b-f9a0e5895d14 · outbound

This paper cites Sparsely Activated Mixture-of-Experts are Robust Multi-Task Learners.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Sparsely Activated Mixture-of-Experts are Robust Multi-Task Learners

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T01:05:44.679323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T01:05:44.679323Z digest=sha256:f10ac6b4799ad6ea3c7adba3f918045a9ed697287db8aee6a99aa90d019eb14d

Observation b14f26aa-863c-4f9f-8fb7-d577a9f976d5 · outbound

This paper cites Outrageously Large Neural Networks: The Sparsely-Gated Mixture-of-Experts Layer.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Outrageously Large Neural Networks: The Sparsely-Gated Mixture-of-Experts Layer

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T01:05:44.775645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T01:05:44.775645Z digest=sha256:74f33da68abf856a6934462921fc2d6f246fe1dd5cc26a58f1807fa70073dbed

Observation acd40c53-6442-45f1-b63d-2d5296943386 · outbound

This paper cites [Online].

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts [Online]

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:48.012664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T01:05:45.460528Z digest=sha256:cba09d6912ba3b82ce5902282f3907128c542c8c7d808e18379f64284745318f

Observation 1ffa1fea-0693-407b-934a-1c9022565ecd · outbound

This paper cites [Online].

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts [Online]

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:47.877416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T01:05:45.556818Z digest=sha256:b75a508f3c59685a3c6d5f360280c473e965994e88633f41085ed2ac8b5f23e5

Observation eaeb04fd-c4b1-4509-9e5f-27669a752e2d · outbound

This paper cites [Online].

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts [Online]

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:49.060196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T01:05:45.073418Z digest=sha256:3d7158fb20d28d0735ca5e396b318d2395721a4674cb89582f91e2d9615346e9

Observation bed821b7-7dc2-48d8-8567-32de69b8f85a · outbound

This paper cites [Online].

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts [Online]

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:48.848835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T01:05:45.170383Z digest=sha256:557b8974e62869989795700569ef264f52cad63de16b5aa099ba06ed6cba4671

Observation f46f5ba2-6f3d-4384-b665-a628fc6e8ef8 · outbound

This paper cites [Online].

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts [Online]

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:48.642320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T01:05:45.264475Z digest=sha256:4e48849fe60a87aa5cae8051284a09cb317fc6f34272cfdc7b017a91cb6cde54

Observation d5367c58-a785-4998-84df-386c5f577384 · outbound

This paper cites [Online].

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts [Online]

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:48.440198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T01:05:45.321574Z digest=sha256:1dd2ffac9d5f73601dc58f6bb3896f6dd513de8aeb6a20490f2cd63c4600707e

Observation db6ba20b-e15a-44fb-8a18-ec018b8b582e · outbound

This paper cites [Online].

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts [Online]

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:48.164816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T01:05:45.394744Z digest=sha256:4df64cca20eae173733f8e716e45da841f182098f08933fd3c058434a4c7d938

Observation 072708d6-4e41-4654-a0ed-d3ab503c465a · outbound

This paper cites Pulse code modulation (pcm) of voice fre- quencies,.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Pulse code modulation (pcm) of voice fre- quencies,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:46.982149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T01:05:46.071817Z digest=sha256:b64336cc5bd1f086247a30f1ed10ec43b5332fffd6fedfc5590329431bb848a9

Observation 33a791ee-bc8a-4d32-8f6d-292134b25192 · outbound

This paper cites Our in-house test sets consist of 1,000 samples of male and female speech from daily conversations, with refer- ence translations curated by professional translators.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Our in-house test sets consist of 1,000 samples of male and female speech from daily conversations, with refer- ence translations curated by professional translators

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:50.162275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T01:05:44.122596Z digest=sha256:d63c11cb2b4642214c0162d939030595cf15354d7f6a98cddef40caa604770bd

Observation aa6474cd-aa62-4796-975f-820f642c40e4 · outbound

This paper cites [Online].

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts [Online]

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:47.654187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T01:05:45.618663Z digest=sha256:c517ec56c20a50bb67295849277fd9f3f9f2d99a2c25009856a12d9b02a7a12a

Observation a1f6a03c-ea88-490c-be24-e1e607f20541 · outbound

This paper cites [Online].

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts [Online]

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:47.474698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T01:05:45.683072Z digest=sha256:1c742c086f7c4cf4d2a5c4872d546117acf2b8af8dc124249b87375a23c866f1

Observation 38a3220f-17f0-4c27-8d6f-3a9caec67139 · outbound

This paper cites Whisper is a widely used mul- tilingual speech model trained with diverse language pairs.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Whisper is a widely used mul- tilingual speech model trained with diverse language pairs

Reference 24

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T01:05:49.880492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T01:05:44.239321Z digest=sha256:9c6eb7319826bc832a48d78863b83859a12374806c49d8cc7eced6d35ece66bd

Observation 49206302-ec76-4c0c-8fc2-e0053d1bb54b · outbound

This paper cites [Online].

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts [Online]

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:47.280919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T01:05:45.783547Z digest=sha256:7456bc88749d00e3533861612da352c3500d1ace5d651118f015d8035975a3bd

Observation c4cad7bc-fb68-44ec-99b9-fde756a834a2 · outbound

This paper cites [Online].

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts [Online]

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:47.116778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T01:05:45.863666Z digest=sha256:f34840479e15d364878134578fcc3449f8b9c67a25257b83b2d7b6e19e1a1305

Observation e2fefd12-a43e-47c1-9e49-927204004dd2 · outbound

This paper cites The adaptive multirate wideband speech codec (amr-wb),.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts The adaptive multirate wideband speech codec (amr-wb),

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T01:05:45.952971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T01:05:45.952971Z digest=sha256:423d29929945bf8fdc3aef5e9de5ca0b53505665a3f66c0c76dbe9501b12e53b

Observation ad761f03-d506-4a96-bd8f-686fc5e454ef · outbound

This paper cites Fleurs: Few-shot learning evaluation of universal representations of speech,.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Fleurs: Few-shot learning evaluation of universal representations of speech,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T01:05:46.175496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T01:05:46.175496Z digest=sha256:9c45806a2c70de6f8293ef1e2ebddec38ec6fe6842fd89cbb6a9531107eff4a4

Observation 29b34952-b47b-4dee-9584-9a971a3bf21d · outbound

This paper cites Attention is all you need,.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Attention is all you need,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T01:05:46.876936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T01:05:46.281915Z digest=sha256:eaa16fc0275e22e84e5a14f2d6883d5b7a34bc3d63d2543e8b48445ad240fd29

Observation bd80abc3-b9aa-49ce-9548-76222386d705 · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation,.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Bleu: a method for automatic evaluation of machine translation,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T01:05:46.369512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T01:05:46.369512Z digest=sha256:c8f730464d02b60c4f941409653d913ba4ef145d6572689e4db2b5f9e87c1f97

Observation 2be105f0-24c8-4d44-b516-90f8e82b548b · outbound

This paper cites Robust speech recognition via large-scale weak supervision,.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Robust speech recognition via large-scale weak supervision,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T01:05:46.447560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T01:05:46.447560Z digest=sha256:01fbd51c5e6b4aa6a296eada34c7ee3158d0cdc89d04cd2bfb9225b5ade9f224

Pith citing papers

Observation 586c0fd9-9f91-47e6-b24d-e0a6ec76ebec · inbound

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts cites this paper.

Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts Beyond Hard Sharing: Efficient Multi-Task Speech-to-Text Modeling with Supervised Mixture of Experts

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T01:05:46.690100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T01:05:43.913842Z digest=sha256:77c0efebd6ad45e9b5e4fd6a0aba6ba866c965b69a38f0bbfea9d77203798c54