Pith. sign in

Paper Citation Record · LEDGER

Tokenization Matters: Improving Zero-Shot NER for Indic Languages

As of 21 August 2026, this Paper Citation Record lists 33 of 33 outbound references and 4 inbound Pith citation observations for arXiv:2504.16977.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.16977 v1

Coverage vector

measured 33 of 33 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:55:55.472875Z

measured 37 of 37 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T23:55:17.570331Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T11:35:19.928065Z

Reference resolution

33 of 33 outbound references displayed

  • verified exact2
  • verified fuzzy19
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c9e1467e-0ec8-42d1-a049-8daaac4a6dba · outbound

This paper cites Census of india 2011: Data on language and mother tongue,.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages Census of india 2011: Data on language and mother tongue,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:55:55.954887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:55:55.343099Z digest=sha256:5022b31e7947a02345a2e2c019591e817767018a687841e73e92885079208b4a

Observation a36e224b-5076-4af4-b8b1-dacb0b773152 · outbound

This paper cites Eighth schedule to the constitution of india,.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages Eighth schedule to the constitution of india,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:55:55.943095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:55:55.350381Z digest=sha256:422d76a3939d5ada34802e6bfb67d33adc15bc07314aad59d2fe3a61e4d13a8e

Observation ef213086-f0a6-4a3c-bd92-319a355af6da · outbound

This paper cites Review of reference generation methods in large language models,.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages Review of reference generation methods in large language models,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:55:55.930685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:55:55.354438Z digest=sha256:042df6a611e9d5fa3a5ea248fe83c7574073f493719149ec8d1fd53939b94e13

Observation a79e69e9-c53c-4e5f-906b-a65a4233a503 · outbound

This paper cites Retrofitting language models with dynamic tokenisation,.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages Retrofitting language models with dynamic tokenisation,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:55:55.919468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:55:55.358990Z digest=sha256:8112d4dbbf89ac977c4c98b860db9dce692bb3081d42036276956da8161ecd1f

Observation ff900633-a840-4e46-bc57-abbc82b16550 · outbound

This paper cites Clinical QA 2.0: Multi-Task Learning for Answer Extraction and Categorization.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages Clinical QA 2.0: Multi-Task Learning for Answer Extraction and Categorization

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T10:55:55.363340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:55:55.363340Z digest=sha256:d2e043835d9a455867bec34f4d0c76b327311800166394b9f4d341e1720084f0

Observation a879e7c6-2615-4359-bca8-9e8093af8534 · outbound

This paper cites Neural machine translation of rare words with subword units,.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages Neural machine translation of rare words with subword units,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:55:55.908224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:55:55.367615Z digest=sha256:2c85aefafeba724089275835cf64036a89b5b8431339af63adcec94b0db2b9c8

Observation c34d4be7-2e6b-47fa-af2a-5c9603745d22 · outbound

This paper cites Bert: Pre-training of deep bidirectional transformers for language understanding,.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages Bert: Pre-training of deep bidirectional transformers for language understanding,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:55:55.897329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:55:55.371879Z digest=sha256:3d0d445281b018bc15abd3a98f0dc57b8d6f668a6d2ad504c27c74acc9560b93

Observation 3eebeb5e-72dd-46cd-b3d8-985dd6919b9a · outbound

This paper cites Towards Leaving No Indic Language Behind: Building Monolingual Corpora, Benchmark and Models for Indic Languages.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages Towards Leaving No Indic Language Behind: Building Monolingual Corpora, Benchmark and Models for Indic Languages

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T10:55:55.376348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:55:55.376348Z digest=sha256:fec21fdcd139bc3cd7eeca9a77ea68c6d11325a39de72f7912ae6d2df98df93d

Observation 08629775-0382-4f57-8135-a0052f8be52e · outbound

This paper cites Un- supervised cross-lingual representation learning at scale,.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages Un- supervised cross-lingual representation learning at scale,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:55:55.885399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:55:55.380584Z digest=sha256:2cd0ba16a2af535651a10c66ad928f5d80c936fe7ab0bd98a215212ae847b985

Observation a087b0a7-3ec9-417c-b6b4-960f7bc59772 · outbound

This paper cites Enhancing Document AI Data Generation Through Graph-Based Synthetic Layouts.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages Enhancing Document AI Data Generation Through Graph-Based Synthetic Layouts

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T10:55:55.384566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:55:55.384566Z digest=sha256:65f5e37a0510b2776ef079f3884b8a0394cfb76ccf14bc042cde960db5286ab9

Observation 23928213-c50c-4df1-a4b3-8a55a04f7eda · outbound

This paper cites Fs-dag: Few shot domain adapting graph networks for visually rich document understanding,.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages Fs-dag: Few shot domain adapting graph networks for visually rich document understanding,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:55:55.873304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:55:55.388684Z digest=sha256:67682d9b7df03b66756eb34e3cf22f0a1f707221bd40836d3414747df5d7ca0c

Observation e21b69e9-faee-4321-b634-de9d7b7cf23b · outbound

This paper cites Continuous Spiking Graph Neural Networks.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages Continuous Spiking Graph Neural Networks

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T10:55:55.392269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:55:55.392269Z digest=sha256:9f861df5940208b867175e541d13108edb2e1f853e1fa2d3bfac1885f0c1277e

Observation 566d1c86-965b-4390-9f29-b8486e8d6035 · outbound

This paper cites Sentencepiece: A simple and language inde- pendent subword tokenizer and detokenizer for neural text processing,.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages Sentencepiece: A simple and language inde- pendent subword tokenizer and detokenizer for neural text processing,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:55:55.861382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:55:55.396358Z digest=sha256:cbc53a164e11ebcc98b5a6ddf782e9fc20ed007000b8ca3ac4e8bc5db37c8e8e

Observation b5e64fd4-877a-4005-8a00-f8bc67595333 · outbound

This paper cites A survey of cross-lingual word embedding models,.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages A survey of cross-lingual word embedding models,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:55:55.847073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:55:55.399981Z digest=sha256:c630bedeb362b4f9e26e149b9e8f43063d4d0c45d695b6e751976f9594a262b1

Observation ffe1b38d-25ce-4681-bfb4-c9cbfd8284c2 · outbound

This paper cites Survey of Large Multimodal Model Datasets, Application Categories and Taxonomy.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages Survey of Large Multimodal Model Datasets, Application Categories and Taxonomy

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T10:55:55.403585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:55:55.403585Z digest=sha256:9fd703bc847f6e1fdc1de07be0c851350e323d1892331ac9ab445338feb5ee7a

Observation 2892dfdc-75a2-482c-b14a-7c5c14a6c5cd · outbound

This paper cites Tokenizers for african languages,.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages Tokenizers for african languages,

Reference 16

Resolution
verified exact
raw_fallback, observed 2026-08-16T10:55:55.661203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:55:55.407397Z digest=sha256:c4917148c524db474d7a182c8226ec673aa7a9d03b21d1b9ed9b8e3fdb4fdfd5

Observation 87384943-aeb8-4ca6-ae35-b73445bf01cc · outbound

This paper cites MVTamperBench: Evaluating Robustness of Vision-Language Models.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages MVTamperBench: Evaluating Robustness of Vision-Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T10:55:55.410963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:55:55.410963Z digest=sha256:e830aedb18d91883521a8b979e8dade1f374d328762c77ad4f57743ca31fce41

Observation 0c446cb3-4fd4-4329-928c-831ebee2ac50 · outbound

This paper cites Hybrid machine learning and deep learning approaches for insult detection in roman urdu text,.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages Hybrid machine learning and deep learning approaches for insult detection in roman urdu text,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:55:55.834971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:55:55.414727Z digest=sha256:182ac4be022f54648b94e1b5bd6a45e89800e1527dc0feb61b3a50c8a0044ec0

Observation a092e4ed-44e1-46e1-b0ae-cb18ff71aaa6 · outbound

This paper cites Tokenization Standards for Linguistic Integrity: Turkish as a Benchmark.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages Tokenization Standards for Linguistic Integrity: Turkish as a Benchmark

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T10:55:55.418419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:55:55.418419Z digest=sha256:364edb4bae07308539b4c19e3b7aab34b1a9d7a20f359f336ac3025a59b4c6b4

Observation 349286a5-fcc5-46f2-96a1-d66c2289ec00 · outbound

This paper cites Pseudo-labelling based boot- strapping for semi supervised learning,.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages Pseudo-labelling based boot- strapping for semi supervised learning,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:55:55.821895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:55:55.422474Z digest=sha256:1073c7d33264eb3f44e5b6a7e75d9606ac278183beefe78284cd3486dd8d2ba1

Observation 53a8a361-9912-4d4c-957f-b283e3e6ed2a · outbound

This paper cites Leveraging cumeta for enhanced document classification in cursive languages with transformer stacking,.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages Leveraging cumeta for enhanced document classification in cursive languages with transformer stacking,

Reference 21

Resolution
verified exact
doi, observed 2026-08-16T10:55:55.509222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:55:55.426200Z digest=sha256:2a4611a0608aa216b1e2addd4780eda534b38b986b7b4b793a4d556f20781d9e

Observation c2bf618e-e22a-4081-b094-421200fc7a45 · outbound

This paper cites Augmented input representations in sequence generation models for decipherment and translation,.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages Augmented input representations in sequence generation models for decipherment and translation,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:55:55.809420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:55:55.429955Z digest=sha256:dee9e66db3d391b47bf10f43a55513341fafd1a544a35abe843eab972022d73d

Observation 412854b0-f140-45fb-8700-23ba039ec514 · outbound

This paper cites LLM for Barcodes: Generating Diverse Synthetic Data for Identity Documents.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages LLM for Barcodes: Generating Diverse Synthetic Data for Identity Documents

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T10:55:55.433489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:55:55.433489Z digest=sha256:07e0d0fb79f57676a840d7c507c333f32bcc45a5a52095ac49d0b1b8ba1b1656

Observation b76de85d-030c-4995-abe1-c97c8aca35d8 · outbound

This paper cites Indicnlp corpus: Monolingual corpora and word embeddings for indic languages,.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages Indicnlp corpus: Monolingual corpora and word embeddings for indic languages,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:55:55.797412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:55:55.437409Z digest=sha256:116e4c6b2a0017a1b0cb87cdd48cd8c54c0761f178e403602e9847b581b32cbe

Observation c9f53abe-0a6e-45c9-b8d1-326fd22353f0 · outbound

This paper cites A survey on recent approaches for natural language pro- cessing in low-resource scenarios,.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages A survey on recent approaches for natural language pro- cessing in low-resource scenarios,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:55:55.784377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:55:55.441153Z digest=sha256:90890c4de1a75ded01029a4a3e398d2b87c876eaab35e59946e94df019b4bfe2

Observation d960e47c-f474-44b9-8f75-0402fee4a801 · outbound

This paper cites Named entity recognition for indian languages,.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages Named entity recognition for indian languages,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:55:55.772389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:55:55.444925Z digest=sha256:155922eb025585df3e7d9faee69a1b24fa34152bb799bbf3d8268d111830cd60

Observation b3ec6b1c-fc64-4e46-979f-63e3fa5c12af · outbound

This paper cites When every token counts: Optimal segmentation for low-resource language models,.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages When every token counts: Optimal segmentation for low-resource language models,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:55:55.760413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:55:55.448994Z digest=sha256:5121b0bfea8290df851a00dbb17272accf5465fe468cc84fd101ee677db54a68

Observation 4a8c808c-8042-422a-837a-90d574b726e7 · outbound

This paper cites NER- RoBERTa: Fine-Tuning RoBERTa for Named Entity Recognition (NER) within low-resource languages.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages NER- RoBERTa: Fine-Tuning RoBERTa for Named Entity Recognition (NER) within low-resource languages

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T10:55:55.452661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:55:55.452661Z digest=sha256:c05c36637fb970dedd3d7c5d3f531b35531481ca42b2c24733293a3212ec59e5

Observation 4932a856-f10c-4cb9-af81-96cdc621db4d · outbound

This paper cites Adaptive subword tokenization for low- resource nlp: Balancing efficiency and generalization,.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages Adaptive subword tokenization for low- resource nlp: Balancing efficiency and generalization,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:55:55.747561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:55:55.457023Z digest=sha256:0ecc0ee55a82dff12cc19e60ae32ae319aca743fc8b103e83b308f57eff23e54

Observation 472eac41-6e95-4b58-9343-b1c85a10bd84 · outbound

This paper cites When Every Token Counts: Optimal Segmentation for Low-Resource Language Models.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages When Every Token Counts: Optimal Segmentation for Low-Resource Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-16T10:55:55.460756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:55:55.460756Z digest=sha256:506305c1c9aa675febf8b50a300a50e81ec5d2e255a88d8a10ea90bb43e96010

Observation c4439a7f-dcdb-4d4e-bf2a-d9ab3a38a0bc · outbound

This paper cites No Language Left Behind: Scaling Human-Centered Machine Translation.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages No Language Left Behind: Scaling Human-Centered Machine Translation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-16T10:55:55.464852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:55:55.464852Z digest=sha256:41784b9c0054bb58ba2c50814e5041ec8fce95e18452bfac91bd7a7243a0e3a9

Observation 272ee6f3-2696-477b-95b6-b35f9ba4683f · outbound

This paper cites Naamapadam: A Large-Scale Named Entity Annotated Data for Indic Languages.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages Naamapadam: A Large-Scale Named Entity Annotated Data for Indic Languages

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-16T10:55:55.469053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:55:55.469053Z digest=sha256:ae59a78cafb5b9e1ae7bda592e9843661b598ecd29063766b00393dbfafb1eef

Observation 9a96cd2e-d127-4149-bab2-041e609abca0 · outbound

This paper cites Multiclass text classi- fications of sindhi newspaper articles,.

Tokenization Matters: Improving Zero-Shot NER for Indic Languages Multiclass text classi- fications of sindhi newspaper articles,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:55:55.734321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:55:55.472875Z digest=sha256:85526b4802ea74c88d86fb2d29c2c3af804980746002ca7e4b87cc1e148ec5de

Pith citing papers

Observation b0f7d0ad-c704-4370-a2a7-e0e125a09789 · inbound

MVTamperBench: Evaluating Robustness of Vision-Language Models cites this paper.

MVTamperBench: Evaluating Robustness of Vision-Language Models Tokenization Matters: Improving Zero-Shot NER for Indic Languages

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T23:55:17.570331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:55:17.570331Z digest=sha256:f0addbd529c7f2efa655f1f391dad09eb67c782fa134516465afbe1a3e9bce5d

Observation 5b2fa618-4552-4e31-8336-e1f9e5468b8e · inbound

SweEval: Do LLMs Really Swear? A Safety Benchmark for Testing Limits for Enterprise Use cites this paper.

SweEval: Do LLMs Really Swear? A Safety Benchmark for Testing Limits for Enterprise Use Tokenization Matters: Improving Zero-Shot NER for Indic Languages

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:29.703994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:52:29.703994Z digest=sha256:ea4241bbfd14d0e4d5b7e9b49cb2f840a2006cf87229f21b59673735ee77ce91

Observation eb67cc13-93c6-4c67-9743-6163cda58732 · inbound

Hard Negative Mining for Domain-Specific Retrieval in Enterprise Systems cites this paper.

Hard Negative Mining for Domain-Specific Retrieval in Enterprise Systems Tokenization Matters: Improving Zero-Shot NER for Indic Languages

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T14:36:13.135559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:36:13.135559Z digest=sha256:5d7d76c3324ef6cba94bcb53a2a0caf8348ed3ab63863d9d073c95486c3d1775

Observation 21c10cc0-8486-461f-be9a-0afc31687822 · inbound

Hybrid AI for Responsive Multi-Turn Online Conversations with Novel Dynamic Routing and Feedback Adaptation cites this paper.

Hybrid AI for Responsive Multi-Turn Online Conversations with Novel Dynamic Routing and Feedback Adaptation Tokenization Matters: Improving Zero-Shot NER for Indic Languages

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:35:20.018282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T11:35:17.100580Z digest=sha256:2f4a8297ee285fe62a79c7a0fc2c2530ed9e85485fcc2cd2dd2666fc7835729a