Pith. sign in

Paper Citation Record · LEDGER

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning

As of 13 August 2026, this Paper Citation Record lists 33 of 33 outbound references and 0 inbound Pith citation observations for arXiv:2502.03884.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.03884 v1

Coverage vector

measured 33 of 33 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T00:26:05.310741Z

measured 33 of 33 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

33 of 33 outbound references displayed

  • verified exact0
  • verified fuzzy24
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b862c3bc-00b8-4c60-9541-8ca000df6a8e · outbound

This paper cites A survey on evaluation of large language models.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning A survey on evaluation of large language models

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:26:05.618170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T00:26:05.206525Z digest=sha256:a465f674eb6a49780abe460aac4a970c5930c26488df4925df73c2831c2c48df

Observation 5f67a3fb-5490-419c-b9f6-b6f4bddaef38 · outbound

This paper cites Qlora: Efficient fine- tuning of quantized llms.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning Qlora: Efficient fine- tuning of quantized llms

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:26:05.600642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T00:26:05.213503Z digest=sha256:7d8aaa9b1717220b0c2233f068212ad4d1dfaeed56f2f3ca6a1b295e8abdcd7c

Observation 35d77e59-1004-4f6d-b6f5-5ef3294885b7 · outbound

This paper cites Loramoe: Alleviating world knowledge forgetting in large language models via moe-style plugin.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning Loramoe: Alleviating world knowledge forgetting in large language models via moe-style plugin

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:26:05.591718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T00:26:05.217045Z digest=sha256:27e8dc50b9eed53a484811cf59cb09b95d1adca90f685c5348515b651ca6ec11

Observation 96dae0cf-a698-444a-a428-87e8ba1e363f · outbound

This paper cites Glam: Efficient scaling of language models with mixture- of-experts.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning Glam: Efficient scaling of language models with mixture- of-experts

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:26:05.583043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T00:26:05.220317Z digest=sha256:906dc0ca6009c49d77a43ba86dd5eee2a1858ae64d454a4635d63a167cb24fa3

Observation 5dd8a9dd-b79d-434c-9eaa-d518eef1847f · outbound

This paper cites Higher Layers Need More LoRA Experts.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning Higher Layers Need More LoRA Experts

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-09T00:26:05.227041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:26:05.227041Z digest=sha256:02d6ea8f1dc26ee16dea8b53fca28456638331e1bfba04b73543c6ca58ea5a88

Observation 31edfba0-dbb4-404a-bf4c-c46d75abf908 · outbound

This paper cites Parameter-Efficient Fine-Tuning for Large Models: A Comprehensive Survey.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning Parameter-Efficient Fine-Tuning for Large Models: A Comprehensive Survey

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T00:26:05.230497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:26:05.230497Z digest=sha256:26b69c1b2becc263042f4d6f9a131c1b288cc78a1cea0407d179cf22dab3316c

Observation 1d0cc5e4-843b-41ec-822d-2e8b20a45221 · outbound

This paper cites Lora+: Efficient low rank adaptation of large models.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning Lora+: Efficient low rank adaptation of large models

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:26:05.569066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T00:26:05.233857Z digest=sha256:804760cdacd1a63dac3546241a3a197deb4d19d4927be5ec7d21e5ba514a6f61

Observation 7e7fa0a2-6eda-4973-8b48-856e02ef725d · outbound

This paper cites Parameter-efficient transfer learning for nlp.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning Parameter-efficient transfer learning for nlp

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:26:05.560832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T00:26:05.236940Z digest=sha256:9e5ce734404117653ff8098757a73c3c4004407cfb54f0faa044c52300b34a5c

Observation 8d36c390-b455-4bd0-8148-d23742f39920 · outbound

This paper cites Harder tasks need more experts: Dynamic routing in moe mod- els.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning Harder tasks need more experts: Dynamic routing in moe mod- els

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:26:05.542500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T00:26:05.243594Z digest=sha256:d3eadc6a08ee5eada1ae27f93ddf831d37a223ad3292649fd0ac2d53f0a5cbd8

Observation d6f04330-20f2-465c-ac58-ef062d5a9ff7 · outbound

This paper cites Adaptive mix- tures of local experts.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning Adaptive mix- tures of local experts

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:26:05.532936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T00:26:05.246724Z digest=sha256:94d726adf0ebeecb33a174feac861825f970d8ab3c1792895011fd44864d6e36

Observation 3be8d3ee-24ce-4fea-84f5-217c2fa72aad · outbound

This paper cites Gshard: Scaling giant models with conditional computa- tion and automatic sharding.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning Gshard: Scaling giant models with conditional computa- tion and automatic sharding

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:26:05.514589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T00:26:05.252873Z digest=sha256:8cccf531085c635494ce79aa815c5c808230ec703dc277c11ec1e0c31a4a8468

Observation 88492b77-2f3a-48e3-94a1-188bb4436969 · outbound

This paper cites Prefix- tuning: Optimizing continuous prompts for generation.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning Prefix- tuning: Optimizing continuous prompts for generation

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:26:05.496140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T00:26:05.259314Z digest=sha256:917f2c63dc6ecb2b4f773cd5ca8e7f1c324572b860074ba7ca87fa9a95f9a7b8

Observation 9cb577ec-e8ce-4346-a649-a7084e949904 · outbound

This paper cites MixLoRA: Enhancing Large Language Models Fine-Tuning with LoRA-based Mixture of Experts.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning MixLoRA: Enhancing Large Language Models Fine-Tuning with LoRA-based Mixture of Experts

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-09T00:26:05.262430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:26:05.262430Z digest=sha256:adab1359f9471e91773ce58406077faa91cad8e0d42a89e0fe8db38a7d6332ee

Observation f9c4d36b-7e27-48d6-aa18-9a705494613b · outbound

This paper cites Learn to ex- plain: Multimodal reasoning via thought chains for sci- ence question answering.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning Learn to ex- plain: Multimodal reasoning via thought chains for sci- ence question answering

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:26:05.486504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T00:26:05.265680Z digest=sha256:9a09729f954199db1c10918161e8941b5c98f20365c057d132a071a937e7f6f6

Observation 0f3d2e9c-8792-4b45-8d72-cee32c5e99c4 · outbound

This paper cites Can a suit of armor con- duct electricity? a new dataset for open book question an- swering.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning Can a suit of armor con- duct electricity? a new dataset for open book question an- swering

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:26:05.476142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T00:26:05.268990Z digest=sha256:722b3ef0deeb7e46246733d1f27741433b7f0cc572ccb38fb8bb7be32b482098

Observation 41587668-7476-456c-b892-dcc884480144 · outbound

This paper cites Alphalora: Assigning lora experts based on layer training quality.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning Alphalora: Assigning lora experts based on layer training quality

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:26:05.456912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T00:26:05.274643Z digest=sha256:b21a757c8f6c68df2f5deaa093b42c9188b6b98d43dcc11df91ca313a3a7ff5e

Observation ac3c672e-5152-4bbc-936f-060f344b019d · outbound

This paper cites Scaling vision with sparse mixture of experts.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning Scaling vision with sparse mixture of experts

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-09T00:26:05.277593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:26:05.277593Z digest=sha256:74228e0de07870bbaa8a5997ee3138935f334a6de7a2b2e9c34f40bd7afb7b2b

Observation 895c8a30-f7df-487d-b9e9-8969837719c7 · outbound

This paper cites Hash layers for large sparse mod- els.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning Hash layers for large sparse mod- els

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:26:05.442901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T00:26:05.280812Z digest=sha256:486199f58bc188819f1df3a63e597912eed8bfe7e6047029f88abed2d51dcb76

Observation b83c8b17-d7f7-4207-9f8d-e2f1dc9fa41f · outbound

This paper cites Commonsenseqa: A question answering challenge targeting commonsense knowledge.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning Commonsenseqa: A question answering challenge targeting commonsense knowledge

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:26:05.433804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T00:26:05.283795Z digest=sha256:639cea649b4c55edd985869a33af49e26647b1ad987f6b36ebd0aade5f3cc85b

Observation a13da25f-642e-42ca-acd6-721fc7a9b1a9 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T00:26:05.286874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:26:05.286874Z digest=sha256:af1224f3a8589bb9abb09651dc8df17a72061d83bf3b5b55028ba7177733e0d3

Observation e32b3fa7-b9e7-4485-9bb8-363f58235f74 · outbound

This paper cites Parameter-Efficient Sparsity Crafting from Dense to Mixture-of-Experts for Instruction Tuning on General Tasks.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning Parameter-Efficient Sparsity Crafting from Dense to Mixture-of-Experts for Instruction Tuning on General Tasks

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-09T00:26:05.296940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:26:05.296940Z digest=sha256:b9086a7b51b1cc68484fc9f26a1b886257b7a06f1f72c9a1d24e647bfaa58efe

Observation 3db82f3a-efc2-4ede-98c4-1e601e4bf402 · outbound

This paper cites MoRAL: MoE Augmented LoRA for LLMs' Lifelong Learning.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning MoRAL: MoE Augmented LoRA for LLMs' Lifelong Learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-09T00:26:05.300812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:26:05.300812Z digest=sha256:997203c4dcabc4280943fd5d55abb25f66b4e9466a4aeb3e7f1790dc9fd4396a

Observation 94cc8bc0-d0de-4f9c-8b49-f8a6bc49ca4f · outbound

This paper cites Xmoe: Sparse models with fine-grained and adaptive expert se- lection.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning Xmoe: Sparse models with fine-grained and adaptive expert se- lection

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:26:05.409859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T00:26:05.304503Z digest=sha256:64a37a7b58f4e6e9e7bc24ec76fa62af3a61e612bc7f5fe547ea57f10ecf10af

Observation f0dea8cb-847c-486d-89f8-cb17503e4b1e · outbound

This paper cites Adamoe: Token- adaptive routing with null experts for mixture-of-experts language models.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning Adamoe: Token- adaptive routing with null experts for mixture-of-experts language models

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:26:05.400436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T00:26:05.307620Z digest=sha256:aa8485eef409371642bb92779ad2b6eb0fe10771e38b521bfb772b48b6cd496c

Observation 1f6bf3ae-769a-4653-81dd-23d6822969ca · outbound

This paper cites Mixture-of-experts with expert choice routing.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning Mixture-of-experts with expert choice routing

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:26:05.390357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T00:26:05.310741Z digest=sha256:8dda54c4b808ec9df811bd5d5a2a2dbc8a23d9126fe8e0eca388bd7679f11e46

Observation ebec409b-746f-4247-a1a7-098a29718c3e · outbound

This paper cites Vera: Vector-based random matrix adaptation.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning Vera: Vector-based random matrix adaptation

Reference 1991

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:26:05.523562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T00:26:05.249763Z digest=sha256:d8451b7d6d9f07efcf5a70298bf4ca6e016c6113c315bddfd6b56eba78a96348

Observation cac6bcff-8588-455c-955c-8556a90ec3f4 · outbound

This paper cites Glue: A multi-task benchmark and analysis platform for natural language understanding.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning Glue: A multi-task benchmark and analysis platform for natural language understanding

Reference 2017

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:26:05.419326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T00:26:05.293824Z digest=sha256:88513dd00a11830737be489ff55df791d1a4be6bb6eaef826fa5924a51caa3f4

Observation b4304389-fac6-48c8-a1cf-4da1303a624c · outbound

This paper cites Mul- timodal contrastive learning with limoe: the language- image mixture of experts.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning Mul- timodal contrastive learning with limoe: the language- image mixture of experts

Reference 2018

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:26:05.466505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T00:26:05.271874Z digest=sha256:5f2a66cae6ea5cb0ca5b609b9d3d396d90f7c5a0244c867d95640548765a1a3c

Observation 079040a4-e268-4004-b6a7-44b7d53d6499 · outbound

This paper cites Lora: Low-rank adaptation of large language mod- els.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning Lora: Low-rank adaptation of large language mod- els

Reference 2019

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:26:05.551804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T00:26:05.240397Z digest=sha256:9723c542c77ebebae6558ae50a72d331f2efa6aef54213025d55a933f7033535

Observation cb8743ae-9c25-492f-b201-204851d03ec8 · outbound

This paper cites The power of scale for parameter-efficient prompt tuning.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning The power of scale for parameter-efficient prompt tuning

Reference 2021

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:26:05.505038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T00:26:05.256133Z digest=sha256:e2e4d5534d4e357a13b6446ed295b81e2fca3e960d52fba945bc9edffff86403

Observation 2e061d58-8c4a-4ed8-965e-2e0a19b3470b · outbound

This paper cites Switch transformers: Scaling to trillion param- eter models with simple and efficient sparsity.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning Switch transformers: Scaling to trillion param- eter models with simple and efficient sparsity

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-09T00:26:05.223668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:26:05.223668Z digest=sha256:1fef64754c0f06ff7d4888ecfc2e553a1dd8ee339590a4d35098843f38e1ff94

Observation 49c68872-daee-4356-bd72-109067ede51a · outbound

This paper cites Attention is all you need.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning Attention is all you need

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-09T00:26:05.290362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:26:05.290362Z digest=sha256:ebf3350287d74c37de16cc1a8b209762242e87cead26f1745ca864f0cf8c3e71

Observation 19c7e156-59f2-4ce9-bcd0-9c6802229211 · outbound

This paper cites Deepseek- moe: Towards ultimate expert specialization in mixture- of-experts language models.

Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning Deepseek- moe: Towards ultimate expert specialization in mixture- of-experts language models

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:26:05.609529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-09T00:26:05.210228Z digest=sha256:0ae597327f7a5c73b7a586ca6ec4c695d836caa6d3286e992345b2095c612e15

Pith citing papers

No inbound Pith citation observations are available.