Pith. sign in

Paper Citation Record · LEDGER

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs

As of 11 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 0 inbound Pith citation observations for arXiv:2605.24681.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.24681 v1

Coverage vector

measured 59 of 59 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-30T13:22:40.017922Z

measured 59 of 59 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

59 of 59 outbound references displayed

  • verified exact14
  • verified fuzzy44
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0c468522-896f-4dc7-83d3-e88935379fde · outbound

This paper cites Multilingual mix: Example interpolation improves multilingual neural machine translation,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Multilingual mix: Example interpolation improves multilingual neural machine translation,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.702346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:25c82fa7366dfe420bb4f22fe147b7cb6ef8d30b15d8de9a2054825d5e482cf3

Observation c2f4e7ea-6ea4-4cf1-b685-54700ac7fd85 · outbound

This paper cites Towards higher pareto frontier in multilingual machine translation,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Towards higher pareto frontier in multilingual machine translation,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.708005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:299034ff5636c4fad8cf09c4d7b3b20c0f00f8c69fccf0fe1394573a0449b38c

Observation 4d52d4fb-cd7e-43c9-86d3-297b3599e121 · outbound

This paper cites Neural machine translation by jointly learning to align and translate,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Neural machine translation by jointly learning to align and translate,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.724660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:a7d7e9e640b06384b417a829fa3b26f0c5317099980dcddcad6ab78306bf8ad2

Observation 86d73fc4-849a-416d-bc47-1d3a9a389b4d · outbound

This paper cites A Paradigm Shift: The Future of Machine Translation Lies with Large Language Models.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs A Paradigm Shift: The Future of Machine Translation Lies with Large Language Models

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-30T13:24:40.052918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:154c564faba06ac8088aac43401f7392e104179eea7f467db0d824fd61af0307

Observation a8bdbb2f-6e1b-445b-863d-fe089945ecb1 · outbound

This paper cites Multilingual machine translation with large language models: Empirical results and analysis,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Multilingual machine translation with large language models: Empirical results and analysis,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.760100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:6844dbe91b09abd59137b0cec08a43d3908fa87b0f8b13b16d1906c1c21c543b

Observation 11a5d4af-1f2c-44c0-b7a4-4917809711a2 · outbound

This paper cites Revolutionising translation with ai: Unravelling neural machine translation and generative pre-trained large language models,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Revolutionising translation with ai: Unravelling neural machine translation and generative pre-trained large language models,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.693154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:82f951747a485c3d8e5f7ffbc1f42bed564599e655ee1a0b4d0b6dfcb0ccf5c4

Observation 03915e2e-6e39-442f-a504-4867067db930 · outbound

This paper cites Continual learning with semi-supervised contrastive distillation for incremental neural machine translation,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Continual learning with semi-supervised contrastive distillation for incremental neural machine translation,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.691265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:c3d65dce64cc406f11fe9fceb219f727e7eee47b5fbdb70b26e113411b9aaea9

Observation 6c9c10a7-2c89-4329-a2f6-e7c5a5253ebb · outbound

This paper cites An Empirical Study of Catastrophic Forgetting in Large Language Models During Continual Fine-tuning.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs An Empirical Study of Catastrophic Forgetting in Large Language Models During Continual Fine-tuning

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-06-30T13:24:40.057926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:7775a196fa1f8c906a3cb09197014d5e68e8492a5250e92f0c83e6e42a044774

Observation 9e7cf721-bffd-42e9-8fbd-9ff1c1eb2d65 · outbound

This paper cites Simple and scalable strategies to continually pre-train large language models,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Simple and scalable strategies to continually pre-train large language models,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.694956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:0271d2e1555a9c92b1bc83d96ffdfc6b4543efd7d1c90fe4680d7f45f4a54591

Observation f9563875-34c0-47d8-a256-a29fe363d2cd · outbound

This paper cites Breaking the script barrier in multilingual pre-trained language models with transliteration-based post- training alignment,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Breaking the script barrier in multilingual pre-trained language models with transliteration-based post- training alignment,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.696817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:a7a0383a76ec30a79e26fe945c5c92e9a7d4adefbad1597c6c05157d9e28bac6

Observation b3d51a11-47be-498a-bd35-ef25ce5b80ea · outbound

This paper cites Towards Incremental Learning in Large Language Models: A Critical Review.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Towards Incremental Learning in Large Language Models: A Critical Review

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-06-30T13:24:40.028942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:57ad02aa9724d35958b750f116a5b5ed3628b2a2b9ec0879814401f8c9505f61

Observation 58e7acd9-9b86-44f3-b3a1-1ddc654cc474 · outbound

This paper cites MoE-LPR: Multilingual Extension of Large Language Models through Mixture-of-Experts with Language Priors Routing.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs MoE-LPR: Multilingual Extension of Large Language Models through Mixture-of-Experts with Language Priors Routing

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-06-30T13:24:40.065937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:09a3da52b1de38aa62e7b296c28e359331465cf983c685e4aaf0f639163ea922

Observation f4339d35-43a1-4d1f-b98f-838af3c084e3 · outbound

This paper cites Overcoming language barriers via machine translation with sparse mixture-of-experts fusion of large language models,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Overcoming language barriers via machine translation with sparse mixture-of-experts fusion of large language models,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.776476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:172c7a6b7a0a1f9a9733447d15e65dd1a976e22d200e5143b736809037dbbc53

Observation 3ae5b5bd-0fe7-4a8d-b6bc-83674a3b481c · outbound

This paper cites Attention is all you need,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Attention is all you need,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.780681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:48450e34c63e02b6cdc513909b2c0cc60acb5ac81183caa3539e95aaf8d9ac22

Observation d7593095-0001-459c-ab45-0e57c93f5cba · outbound

This paper cites Transfer learning for low- resource neural machine translation,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Transfer learning for low- resource neural machine translation,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.766224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:2af38353ca8719df82a41d2369254bbe5b64a5b66cc3a3eb869646a49ca63a61

Observation 6d6f100e-a24e-4b5b-a0a6-dc9ee4e73fa0 · outbound

This paper cites Rapid adaptation of neural machine translation to new languages,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Rapid adaptation of neural machine translation to new languages,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.768079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:f26aa77cc75c95d5c97ec3b2233db6f22c94a5ae38de1bbf587e4b7e0ce14f5d

Observation 630ac5f5-831e-4235-80ed-1b291be12edf · outbound

This paper cites Improving neural machine translation models with monolingual data,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Improving neural machine translation models with monolingual data,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.770100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:aaa6610876c0a62772efad6e5b02077a297aba0cf1bbdbe911f2adf2a25983e3

Observation 7904255b-418b-40a3-944d-c572cd8fe4ff · outbound

This paper cites Zero-shot cross-lingual transfer of neural machine translation with multilingual pretrained encoders,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Zero-shot cross-lingual transfer of neural machine translation with multilingual pretrained encoders,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.772219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:eb771b5bc255ddb68a0ad4fb174bef7b38ded70d1e85e9800d13076d15b2696e

Observation b3235f91-e97d-4073-94c1-d900656e2cad · outbound

This paper cites Towards robust in-context learning for machine translation with large language models,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Towards robust in-context learning for machine translation with large language models,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.762290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:6229fe7b042735ca119b81c3882159b8b0aa67a02922262968259db88882443f

Observation 3c35c5e7-2901-4c81-bc4a-6765cf4eec45 · outbound

This paper cites On the Multilingual Ability of Decoder-based Pre-trained Language Models: Finding and Controlling Language-Specific Neurons.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs On the Multilingual Ability of Decoder-based Pre-trained Language Models: Finding and Controlling Language-Specific Neurons

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T13:24:40.062828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:adb324c37a9385eee3c22b08e3b0166d0b574a013af7c5abcaed83ab3f79226f

Observation 3daf12aa-f28d-4ff7-8ae1-af5956447efb · outbound

This paper cites A paradigm shift: The future of machine translation lies with large language models,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs A paradigm shift: The future of machine translation lies with large language models,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.764367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:0c1d9bb506eb6518005266e58a25575e65f5de46c90cc623a228266b87c5fc0b

Observation 4fc343b9-c714-47cf-8f4d-59719ad2eecb · outbound

This paper cites Improving translation of out of vocabulary words using bilingual lexicon induction in low-resource machine translation,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Improving translation of out of vocabulary words using bilingual lexicon induction in low-resource machine translation,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.774256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:5f9af1ab4ed77f5b5a9d2ba16cfb3a6863e882427e4252e6df93997339a08987

Observation 4c7372dc-ca19-4be7-9b12-ea7ed3bfe8aa · outbound

This paper cites Exploiting domain-specific par- allel data on multilingual language models for low-resource language translation,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Exploiting domain-specific par- allel data on multilingual language models for low-resource language translation,

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-06-30T13:24:40.064123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:30237b00aadf0db45c80281fa218dce7254a4362f412b056db8d67850031d89f

Observation de0c6c74-c21a-4b4e-8ea1-a1699d7c7e57 · outbound

This paper cites Catastrophic interference in reinforcement learning: A solution based on context division and knowledge distillation,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Catastrophic interference in reinforcement learning: A solution based on context division and knowledge distillation,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.778655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:28ec4d4da48aab483f581cb1e19fa01541c4bf62629c25a46149015e95e98137

Observation f8d3c9d9-e3e5-4933-9356-f7a036571bb8 · outbound

This paper cites Understanding catas- trophic forgetting in language models via implicit inference,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Understanding catas- trophic forgetting in language models via implicit inference,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.784937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:f74c126084239a0d5709959f0b67a065b0fa87f60147874c2ce33b63ee359e81

Observation 9c50d113-166d-4950-a107-2d1b92021140 · outbound

This paper cites Efficient continual pre-training for building domain specific large language models,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Efficient continual pre-training for building domain specific large language models,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.753506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:6d0930afd9307dd5190047515a212e61bcfa59ee88c387d4be09ca67a1e88ae5

Observation 23baff92-09ec-44bc-b276-cfc749cfebd6 · outbound

This paper cites Overcoming catastrophic forgetting in graph neural networks,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Overcoming catastrophic forgetting in graph neural networks,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.755812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:d11d70a17bef0e9b0b52a8bf63d4427022f2ba4eb7f370cfada214df3664d5a1

Observation 7ff36dbe-4960-4c72-8e23-06371cd32365 · outbound

This paper cites A Rank Stabilization Scaling Factor for Fine-Tuning with LoRA.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs A Rank Stabilization Scaling Factor for Fine-Tuning with LoRA

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-06-30T13:24:40.044299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:c5a581229b2906dde87e3694e8114ffc37f3528d4d1d7e6e9749385a8a6f4e12

Observation 54776406-cfa2-4997-b3f1-36e6e72426e5 · outbound

This paper cites Efficient Large Scale Language Modeling with Mixtures of Experts.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Efficient Large Scale Language Modeling with Mixtures of Experts

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-06-30T13:24:40.034915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:783bd8ec00534eb11ebc2f518d461e030462852a7e48762377f02710ca18ab41

Observation b22a778a-affb-4b01-9be5-44d51091e2c0 · outbound

This paper cites Distributed learning of mixtures of experts,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Distributed learning of mixtures of experts,

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-06-30T13:24:40.053339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:122218b75df2f5aeee230148d46ff7695d4c3697c56ad49cb8133814cbb58f1b

Observation f7f23c40-a066-4dea-8b3e-e1434b54a81c · outbound

This paper cites Mmoe: Enhancing multimodal models with mixtures of multimodal interaction experts,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Mmoe: Enhancing multimodal models with mixtures of multimodal interaction experts,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.744323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:b65af0f782d343dec00144c91c19f15d03b440ad44aebc5e58d96778fd056411

Observation cc0a9091-386b-456b-97bd-f15c69fd2c6b · outbound

This paper cites From sparse to soft mixtures of experts,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs From sparse to soft mixtures of experts,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.739996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:fbb3cbf45d227f5d7dc938258b9fa68ccb39daced3d2adf906953bfea9b01588

Observation ebce3a09-c70a-49f7-b0dc-49d7aae20417 · outbound

This paper cites Efficient large scale language modeling with mixtures of experts,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Efficient large scale language modeling with mixtures of experts,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.742288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:db3628fa5aca326ed6dc26dbbd5b6a006fef2604dadbd014dd2bc55d0445c5b8

Observation b6661529-bc30-4564-9395-01fa3f54c404 · outbound

This paper cites A paradigm shift in machine translation: Boosting translation performance of large language models,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs A paradigm shift in machine translation: Boosting translation performance of large language models,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.747450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:2a3ff872bbf93db2a17567670de94169735ea38b398e7732b8e1241b84eae5ba

Observation 75733831-677b-4085-9f1c-567a41879d7d · outbound

This paper cites X-ALMA: plug & play modules and adaptive rejection for quality translation at scale,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs X-ALMA: plug & play modules and adaptive rejection for quality translation at scale,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.737130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:3bcf9bf76f1d742a84045528b27b7da912cd50d1dd97f79edb310374b7842335

Observation bac60625-ef22-4d28-aa72-3e384348031c · outbound

This paper cites Digital signal processing: signals systems and filters,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Digital signal processing: signals systems and filters,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.731430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:4d5695c516d54c60941efeed0c8726037ba77c389081ffabf77cf617d3c1e5b8

Observation 54fbd862-d15e-44c3-ae13-493d8aa42eff · outbound

This paper cites On the relation between linguistic typology and (limitations of) multilin- gual language modeling,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs On the relation between linguistic typology and (limitations of) multilin- gual language modeling,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.727347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:2d0cbeed7ba86573fda17715ebb20ee5bb37ac5fe505c2418ff9f1addab1e65b

Observation 5ed78e74-7e38-47d7-9807-ca2b36449346 · outbound

This paper cites Frequency of basic english grammatical structures: A corpus analysis,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Frequency of basic english grammatical structures: A corpus analysis,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.729448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:e2b23ce1767789a6e64139168f20572d03a8f26c6a57b8ee57ac4c117a013f7f

Observation 93450a82-3363-4d2f-9f85-8ccb0068db68 · outbound

This paper cites Deep Learning Based Natural Language Processing for End to End Speech Translation.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Deep Learning Based Natural Language Processing for End to End Speech Translation

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-06-30T13:24:40.061527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:8a394900b3a1598fa435b7e7565a12d55252603c81546cd81131ebf0e2da1d1d

Observation dd716170-496e-4968-afd0-1381307d20e6 · outbound

This paper cites Dnn-based cross-lingual voice conversion using bottleneck features,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Dnn-based cross-lingual voice conversion using bottleneck features,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.733641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:0d72033c36155744270c1db2ca4d26cc521736356bd54154f012e6b93e685b1a

Observation f828a622-6b8f-4025-afb5-705a06478711 · outbound

This paper cites Learn- ing a Fourier transform for linear relative positional encodings in transformers,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Learn- ing a Fourier transform for linear relative positional encodings in transformers,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.749593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:1eec9017e7b2ea2df95a68452c5714dbb9ccd639de16e4e0de325234e449c5b8

Observation 9990986c-9eeb-4881-85fb-e5ddc8365424 · outbound

This paper cites Learnable fourier features for multi-dimensional spatial positional encoding,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Learnable fourier features for multi-dimensional spatial positional encoding,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.757862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:be3d8902a53df70b1e9e8a0835ec07cc202e0900e44562a0679973ab861027d3

Observation bdca59fe-1cc1-4e18-bb3f-e8a47482b1de · outbound

This paper cites Fourier Position Embedding: Enhancing Attention's Periodic Extension for Length Generalization.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Fourier Position Embedding: Enhancing Attention's Periodic Extension for Length Generalization

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-06-30T13:24:40.046732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:2b0bd61df90830f8eaba4a6eead2d48e22dc33857b5b3c6347458e6313098da4

Observation 693ee2d9-dccb-4da6-b248-6149b7fb80bb · outbound

This paper cites Transformer feed-forward layers are key-value memories,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Transformer feed-forward layers are key-value memories,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.718880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:be45a675eb8ccf444bce944c72bb55362d90f7e36aeda8dc0083b0012d989b99

Observation 3e51d6fd-c253-49a3-b50b-5fa9a1d370a2 · outbound

This paper cites Knowledge neurons in pretrained transformers,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Knowledge neurons in pretrained transformers,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.709907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:408e63706922ba259acfe99b0eb7443eaf28bd787b1a7bfe1a6d7aec0ff14ddb

Observation f3456fb0-e2e1-4e01-9872-e75f87871e2e · outbound

This paper cites A mathematical framework for transformer circuits,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs A mathematical framework for transformer circuits,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.713977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:19f43b16a22c05306b30184c16a95741c477be93613e57b57c29ce2f7406b273

Observation a3c94869-ccd5-484f-8dab-7b8f4e98c635 · outbound

This paper cites Mixtral of Experts.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Mixtral of Experts

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-06-30T13:24:40.060223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:883e7607e07bbba6fe694dc14b84438727dfd5dec1356f37ba846d5a4f913cbd

Observation fc2fd4dc-d50b-433a-a1b7-e14bb062ee98 · outbound

This paper cites Switch transformers: Scaling to trillion parameter models with simple and efficient sparsity.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Switch transformers: Scaling to trillion parameter models with simple and efficient sparsity

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.704350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:97d176ff99ac79866e2d6076ff430104ad7a4d1cc1d8c7e162ea0ebf35ad2930

Observation 52883347-d3c3-494e-9d62-596a83e0b925 · outbound

This paper cites Mixtral of experts,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Mixtral of experts,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.706080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:3ed75c6a3a4e914e9e38d22b8b729b02fc6c7a0219e5b05e359fb94df8a09201

Observation b53b3918-19bd-483b-a301-bf0f97f2f32a · outbound

This paper cites LLaMA-MoE v2: Exploring Sparsity of LLaMA from Perspective of Mixture-of-Experts with Post-Training.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs LLaMA-MoE v2: Exploring Sparsity of LLaMA from Perspective of Mixture-of-Experts with Post-Training

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-06-30T13:24:40.055476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:db27c24dda41a206429abac403e585ed563e88b29d8f6cd31c552217f5dfedc4

Observation 837416b3-be07-438f-b32a-e80c8c2bebcd · outbound

This paper cites Llama-moe: Building mixture-of-experts from llama with continual pre- training,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Llama-moe: Building mixture-of-experts from llama with continual pre- training,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.712064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:6486350637481b51e4498341756124ef856826668a1bf9590a6e842222af60a6

Observation a8115eab-883b-4773-b1a9-dd2ad2b91beb · outbound

This paper cites OLMoE: Open Mixture-of-Experts Language Models.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs OLMoE: Open Mixture-of-Experts Language Models

Reference 52

Resolution
verified exact
local_arxiv, observed 2026-06-30T13:24:40.058817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:3d2b8843971028279f657d6b904c67e882bba119b61117be2cbe15d0b415f868

Observation d1d0f6f3-a27c-4917-bdff-06e6e73b8fb0 · outbound

This paper cites A call for clarity in reporting BLEU scores,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs A call for clarity in reporting BLEU scores,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.751537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:c24e2c338693330cd908abc6506bd5395e2a5eeea44c26415f30428ec76cb936

Observation d2d00e2b-e238-4c4a-8f6a-054af2b199df · outbound

This paper cites METEOR: an automatic metric for MT evalu- ation with improved correlation with human judgments,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs METEOR: an automatic metric for MT evalu- ation with improved correlation with human judgments,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.698687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:194f29611e5597a3e422f544498a2db0a6bb7b930352d76d0fe73abda6d78a3d

Observation b591dbee-3af1-4710-bffd-bf6e623fb035 · outbound

This paper cites COMET: A neural framework for MT evaluation,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs COMET: A neural framework for MT evaluation,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.700522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:29033ea994505ad832de8b8cb187097a02a15c32d3764fe9c080a4db21d1f365

Observation 9c98f2e7-23ba-4ca5-9eff-6653dc61fb32 · outbound

This paper cites Qa-lora: Quantization-aware low-rank adaptation of large language models,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Qa-lora: Quantization-aware low-rank adaptation of large language models,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.782722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:1e9d4e75614490567b4cb1ddbb5b6fbbaf838066b14df9874188142c2161e180

Observation de54399a-77fa-48b8-9e16-75b3a0a0cae0 · outbound

This paper cites Llama pro: Progressive llama with block expansion,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Llama pro: Progressive llama with block expansion,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.689527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:0b5bb51f098ae72c70b55143c444cc5a771cb8bea01a5240458871808c0c563b

Observation 6a2a1550-5278-4d9a-a1ec-b38b378cd7b4 · outbound

This paper cites Flashattention: Fast and memory-efficient exact attention with io-awareness,.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs Flashattention: Fast and memory-efficient exact attention with io-awareness,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:15:55.687728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:68ef65d8d88986ec0fefc4c8801c3692be1ea7304876489049590b6853a0a3c4

Observation 12e4b530-7b64-4062-8f87-51202af09a93 · outbound

This paper cites No Language Left Behind: Scaling Human-Centered Machine Translation.

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs No Language Left Behind: Scaling Human-Centered Machine Translation

Reference 59

Resolution
verified exact
local_arxiv, observed 2026-06-30T13:24:40.056217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:22:40.017922Z digest=sha256:9bb15334a01a49447e50cf10df15bcba74a1dfc62aaf5c7d6dcc1e28f6044428

Pith citing papers

No inbound Pith citation observations are available.