Pith. sign in

Paper Citation Record · LEDGER

CodeMixBench: Evaluating Large Language Models on Code Generation with Code-Mixed Prompts

As of 18 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 0 inbound Pith citation observations for arXiv:2505.05063.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.05063 v1

Coverage vector

measured 25 of 25 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:18:35.704016Z

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

25 of 25 outbound references displayed

  • verified exact0
  • verified fuzzy16
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation deacab77-8d21-4d2d-8de8-97ac932730e8 · outbound

This paper cites Program Synthesis with Large Language Models.

CodeMixBench: Evaluating Large Language Models on Code Generation with Code-Mixed Prompts Program Synthesis with Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T23:18:35.249573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:18:35.249573Z digest=sha256:c53ab5cb2d09615700dcfc0c2a938c265141910efe7531cf665eb603c4555371

Observation 46580f22-0eac-43e2-8f2b-6f3677b6dc70 · outbound

This paper cites i am borrowing ya mixing?.

CodeMixBench: Evaluating Large Language Models on Code Generation with Code-Mixed Prompts i am borrowing ya mixing?

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:18:36.550112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T23:18:35.337434Z digest=sha256:718ecbfd8008e9c02175e97acc67ff3c949e4393be57d51a07e204444edea79b

Observation deddd272-a4f5-4c47-9b5c-c9ef2d0d3901 · outbound

This paper cites Phog: Probabilistic model for code generation.

CodeMixBench: Evaluating Large Language Models on Code Generation with Code-Mixed Prompts Phog: Probabilistic model for code generation

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:18:36.509484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T23:18:35.341603Z digest=sha256:eef7e798730aa7271fffb27c2232888c5ce15748ac20fd4281ad9bd3c3f60b9c

Observation 21852ec5-62ee-47d4-be96-db1984db72a0 · outbound

This paper cites Classification of code-mixed text using capsule networks.

CodeMixBench: Evaluating Large Language Models on Code Generation with Code-Mixed Prompts Classification of code-mixed text using capsule networks

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:18:36.405677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T23:18:35.345799Z digest=sha256:31c6b8bc2e0576b27f43257043ac4a9ad4c618f4ceaba6d8a8de66a01aad4fce

Observation 579c3317-9f18-4cb3-93c0-c6d139d82f01 · outbound

This paper cites Evaluating Large Language Models Trained on Code.

CodeMixBench: Evaluating Large Language Models on Code Generation with Code-Mixed Prompts Evaluating Large Language Models Trained on Code

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T23:18:35.350319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:18:35.350319Z digest=sha256:86b10825501b1f25864dcd985f1518183a05c9af979ad0fe5768e1569d720843

Observation f6b2fd87-52be-4d7b-b620-7955a6c31056 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

CodeMixBench: Evaluating Large Language Models on Code Generation with Code-Mixed Prompts DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T23:18:35.354863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:18:35.354863Z digest=sha256:5bf2ae8386a7bc163ff50e641fbda943e576ea55b34881f4e26c635b68244543

Observation eb97ed91-16b4-4328-917c-c0c5e2f4bf58 · outbound

This paper cites Mipe: A metric independent pipeline for effective code-mixed nlg evaluation.

CodeMixBench: Evaluating Large Language Models on Code Generation with Code-Mixed Prompts Mipe: A metric independent pipeline for effective code-mixed nlg evaluation

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:18:36.383219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T23:18:35.359965Z digest=sha256:467cdb4689ed6ce421cbc92c76252ea2b97bc11ab6765c2f594f09be16fee7c7

Observation bd49260b-b5db-40a1-8692-120a94d61b24 · outbound

This paper cites DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence.

CodeMixBench: Evaluating Large Language Models on Code Generation with Code-Mixed Prompts DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T23:18:35.455170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:18:35.455170Z digest=sha256:53c4000121f5c4fb6b900945d7170ebba55a7d09ba0737d6b06cdd67ea2d427c

Observation 7945a1a2-e56a-4b00-9335-94f4fe603128 · outbound

This paper cites Multilingual Controlled Generation And Gold-Standard-Agnostic Evaluation of Code-Mixed Sentences.

CodeMixBench: Evaluating Large Language Models on Code Generation with Code-Mixed Prompts Multilingual Controlled Generation And Gold-Standard-Agnostic Evaluation of Code-Mixed Sentences

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T23:18:35.516306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:18:35.516306Z digest=sha256:49e53deb77dd1df4e0025b77f1978d5f431eec518f54a750ae6dde92ada4d693

Observation 7a3f0f0e-6c1a-4809-a439-81d1af668cda · outbound

This paper cites Measuring coding competence with apps.

CodeMixBench: Evaluating Large Language Models on Code Generation with Code-Mixed Prompts Measuring coding competence with apps

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:18:36.371721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T23:18:35.520765Z digest=sha256:1d83788ade01604fa30123a13ac0d84888af551301796069a85e93c703e8a104

Observation 8b755cbe-4d94-49fe-9a83-c0bbfeee2452 · outbound

This paper cites OpenCoder: The Open Cookbook for Top-Tier Code Large Language Models.

CodeMixBench: Evaluating Large Language Models on Code Generation with Code-Mixed Prompts OpenCoder: The Open Cookbook for Top-Tier Code Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T23:18:35.525371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:18:35.525371Z digest=sha256:60122121a9fd06927413ab811e9231e246e533d9e15224b703376109022de712

Observation db7399f3-104d-4068-b8df-8fec8ea16e85 · outbound

This paper cites Qwen2.5-Coder Technical Report.

CodeMixBench: Evaluating Large Language Models on Code Generation with Code-Mixed Prompts Qwen2.5-Coder Technical Report

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T23:18:35.530570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:18:35.530570Z digest=sha256:f4cd72a06d5e3d6260694eb509e86e5f0d9124dcf8c886c2af0ba25855cf78a4

Observation cfc7534e-3e48-4727-a268-0b587693ed24 · outbound

This paper cites Xcodeeval: An execution-based large scale multilingual multitask benchmark for code understanding, generation, translation and retrieval.

CodeMixBench: Evaluating Large Language Models on Code Generation with Code-Mixed Prompts Xcodeeval: An execution-based large scale multilingual multitask benchmark for code understanding, generation, translation and retrieval

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:18:36.358665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T23:18:35.535050Z digest=sha256:2c09ee7a1f05225d0c21ddface6b795b0480d1d3ccbf74dbb26aac7f8ce83d86

Observation d1ba9b2e-80e2-46d4-b681-1eac09b371eb · outbound

This paper cites Starcoder: May the source be with you! Transactions on Machine Learning Research, 2023.

CodeMixBench: Evaluating Large Language Models on Code Generation with Code-Mixed Prompts Starcoder: May the source be with you! Transactions on Machine Learning Research, 2023

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:18:36.287324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T23:18:35.538539Z digest=sha256:d2c26b505daae1760fdde94b6a556512dca123eb779c210d03d726b328a24ade

Observation e0f9c429-54c4-43cc-ab5c-cb50e0ccb30e · outbound

This paper cites Competition-Level Code Generation with AlphaCode.

CodeMixBench: Evaluating Large Language Models on Code Generation with Code-Mixed Prompts Competition-Level Code Generation with AlphaCode

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T23:18:35.558304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:18:35.558304Z digest=sha256:5dc12cc6dd390efdf94940784429f974edbd65766f79af2207201bfc655c672e

Observation ee3af78b-dafd-4990-86a1-f1c754a9734c · outbound

This paper cites Social Motivations for Codeswitching: Evidence from Africa.

CodeMixBench: Evaluating Large Language Models on Code Generation with Code-Mixed Prompts Social Motivations for Codeswitching: Evidence from Africa

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:18:36.182343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T23:18:35.622218Z digest=sha256:b541babe8f8feff0ca8dbfaf7bbf1de65093de9674f122e28a1ef63a5f475c5d

Observation 719e9503-0e78-4627-a6de-d9aa743dbc84 · outbound

This paper cites L 3 C ube- H ing C orpus and H ing BERT : A code mixed H indi- E nglish dataset and BERT language models.

CodeMixBench: Evaluating Large Language Models on Code Generation with Code-Mixed Prompts L 3 C ube- H ing C orpus and H ing BERT : A code mixed H indi- E nglish dataset and BERT language models

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:18:36.169809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T23:18:35.627051Z digest=sha256:b3e45f8f7446eb6b956d13c52a809f1e5255ed3bc258c7dfa3e6bf3614a4cdeb

Observation 9c098994-e277-4348-8978-083009bc0613 · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

CodeMixBench: Evaluating Large Language Models on Code Generation with Code-Mixed Prompts Bleu: a method for automatic evaluation of machine translation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:18:36.156782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T23:18:35.630744Z digest=sha256:dc277eff21a6f446f720ea199d13adab5bbbf4a847182d56860633f7633c9331

Observation 1276fdf1-b787-4fa1-80ef-33f553aa67f1 · outbound

This paper cites Sentiment analysis of code-mixed text: A comprehensive review.

CodeMixBench: Evaluating Large Language Models on Code Generation with Code-Mixed Prompts Sentiment analysis of code-mixed text: A comprehensive review

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:18:36.098132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T23:18:35.634879Z digest=sha256:c0768fdfb6c82ec0a7b0654d507dca1957bacc5ffab955a2300a37ea7d08818c

Observation ce573e0f-2027-44c2-92f7-7d0548af6017 · outbound

This paper cites Code Llama: Open Foundation Models for Code.

CodeMixBench: Evaluating Large Language Models on Code Generation with Code-Mixed Prompts Code Llama: Open Foundation Models for Code

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T23:18:35.639926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:18:35.639926Z digest=sha256:9a8dec1209d18965cd4609481d87bc10c491752acf9aed15776cff2519fdfafc

Observation b766b0f7-d1bb-47a9-85f5-462ab6d0a138 · outbound

This paper cites Learning to predict code-switching points.

CodeMixBench: Evaluating Large Language Models on Code Generation with Code-Mixed Prompts Learning to predict code-switching points

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:18:36.028558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T23:18:35.687812Z digest=sha256:e2be9ef64cf7318aebd9b8ee12353433feb4a2747dbf74d7fa5a6ebbd6b2de8e

Observation 8aa3407d-a41b-4cf8-80a8-59a816ac0b4d · outbound

This paper cites Feature-rich part-of-speech tagging with a cyclic dependency network.

CodeMixBench: Evaluating Large Language Models on Code Generation with Code-Mixed Prompts Feature-rich part-of-speech tagging with a cyclic dependency network

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:18:36.015629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T23:18:35.692403Z digest=sha256:f1aeba58b2c070f2651d23e28e9f47872c14c1ec657a651fa31e50c4b674c663

Observation 5514d3b7-6fd8-47e3-b988-45232dec1f54 · outbound

This paper cites Adapting multilingual models for code-mixed translation.

CodeMixBench: Evaluating Large Language Models on Code Generation with Code-Mixed Prompts Adapting multilingual models for code-mixed translation

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:18:35.919580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T23:18:35.696359Z digest=sha256:7a77b3d713c8d5f73b27879860a54425952e8f2c35b390ee1b33a4e68e3ddcbe

Observation 76344f3a-9d8a-47b9-aa9e-9f7c320ea426 · outbound

This paper cites A syntactic neural model for general-purpose code generation.

CodeMixBench: Evaluating Large Language Models on Code Generation with Code-Mixed Prompts A syntactic neural model for general-purpose code generation

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:18:35.881327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T23:18:35.700248Z digest=sha256:8ce09ee30dc59a76e977767597ecb7a75fe71e29cddabbb1541640c02683cdd7

Observation 27270050-9858-4485-9aa9-24bd1e8d33fc · outbound

This paper cites Bigcodebench: Evaluating language models on tool-augmented code generation.

CodeMixBench: Evaluating Large Language Models on Code Generation with Code-Mixed Prompts Bigcodebench: Evaluating language models on tool-augmented code generation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:18:35.868455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T23:18:35.704016Z digest=sha256:c883392b3c0b171fcdfa605baae7f37aebc10de15ff3e2ac5dcb58dcfce89457

Pith citing papers

No inbound Pith citation observations are available.