Pith. sign in

Paper Citation Record · LEDGER

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests

As of 18 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2507.21447.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.21447 v2

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T12:50:13.860383Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

41 of 41 outbound references displayed

  • verified exact0
  • verified fuzzy34
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b1c3b57b-7765-4df4-a309-4fbd50fbeb16 · outbound

This paper cites Using a larg e language model as a building block to generate usablevalidation and v erification suite for openmp,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests Using a larg e language model as a building block to generate usablevalidation and v erification suite for openmp,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:15.201881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:09.384287Z digest=sha256:d059afcfd3e73609d7cdf532c199d9b5328435b0998b43ff41442b7873efadf9

Observation 16b7f98c-f4af-4b49-b55c-3123ca5bae56 · outbound

This paper cites Llm & hpc: Be nchmarking deepseek’s performance in high-performance computing tas ks,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests Llm & hpc: Be nchmarking deepseek’s performance in high-performance computing tas ks,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T12:50:09.481701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:50:09.481701Z digest=sha256:63d6eaad02a5747aa4f3ae3321cfee645bc7c335813c871b21eae3b98b3f414e

Observation eecdecaa-c935-4150-8c8b-75d755b83f13 · outbound

This paper cites Can large language models write parallel code?.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests Can large language models write parallel code?

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:15.179334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:09.549213Z digest=sha256:d33f08589237239427903057eaa3e263cfc4768d817f184850229999995701ef

Observation 3a70a8f6-8e42-4afa-94eb-93c4333d2207 · outbound

This paper cites The Landscape and Challenges of HPC Research and LLMs.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests The Landscape and Challenges of HPC Research and LLMs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T12:50:09.646346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:50:09.646346Z digest=sha256:0165c47968cf1f3815ed49672a6af90d80e176a900a2b5a292d82bf54af1c4b0

Observation 57a23e6c-d396-4fd9-a247-0b06549d7982 · outbound

This paper cites An assessment of large langu age models for openmp-based code parallelization: a user perspective ,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests An assessment of large langu age models for openmp-based code parallelization: a user perspective ,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:15.158181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:09.733462Z digest=sha256:cf59e4d96bfef3d20df06031f2c980efa9dba1522eaf731876c8c5955f059494

Observation 90a1c24f-3c9a-4709-a763-cf28f6d2737b · outbound

This paper cites Comparing Llama-2 and GPT-3 LLMs for HPC kernels generation.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests Comparing Llama-2 and GPT-3 LLMs for HPC kernels generation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T12:50:09.860741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:50:09.860741Z digest=sha256:b4f98576be72108f9cc24c0fa96ee245b75efdd0297e2d62d4b560ee771fca49

Observation e2f7b1c5-ebbc-4f7e-b0e6-27b700ea5b5a · outbound

This paper cites Large language model evaluation for high-perform ance com- puting software development,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests Large language model evaluation for high-perform ance com- puting software development,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:15.130899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:09.957015Z digest=sha256:2381010cb8b4f25b10793374136e062c407da893364a2b96d18665cb51a95ae4

Observation 8f341a23-e9b4-4d67-b0d7-61c7ef2d9a1c · outbound

This paper cites Boosting llm-based software generation by aligning code with requirements,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests Boosting llm-based software generation by aligning code with requirements,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:15.108393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:10.051512Z digest=sha256:aa126016cadfabb1d65c180a8c5c48462088ed9130b26a50142d825d21538360

Observation 718df2fd-d63b-41f8-a7c0-5c9be367efcf · outbound

This paper cites Generative Software Engineering.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests Generative Software Engineering

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T12:50:10.163166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:50:10.163166Z digest=sha256:fd2a4fd40073478c235da2d647c23ac489a2f10de54d09c1195f0f8a0f4182f3

Observation 46360c0a-4081-45d0-a45e-f14d4d665100 · outbound

This paper cites V altest: Automated vali dation of language model generated test cases,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests V altest: Automated vali dation of language model generated test cases,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T12:50:10.246055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:50:10.246055Z digest=sha256:5e5fec9c0287245e1a8b635ddfa191b77e417a03942a50ea292e2427ce7f33bb

Observation fb888320-a2f7-4fb3-a07f-3c4ef51ba2db · outbound

This paper cites Llm4vv : Developing llm-driven testsuite for compiler validation,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests Llm4vv : Developing llm-driven testsuite for compiler validation,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:15.086748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:10.347274Z digest=sha256:d41ea9af53eb3d29ffd24e6be681fa0aa3a543476873f84a2e8e045bfd4814ce

Observation c8605b88-4897-460f-828c-2d2e1e936568 · outbound

This paper cites Llm4vv: Exploring llm-as-a-judge for valida tion and verifi- cation testsuites,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests Llm4vv: Exploring llm-as-a-judge for valida tion and verifi- cation testsuites,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:15.062705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:10.422875Z digest=sha256:1c2e9dd75e7b8ec0c607ac442f8402d6f93798a47117334883d03c8086303c5a

Observation b61d5648-e940-480a-a810-c04f0890835c · outbound

This paper cites Hugging Face – The AI community building the future. — h ugging- face.co,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests Hugging Face – The AI community building the future. — h ugging- face.co,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:15.036019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:10.494290Z digest=sha256:6e3580f0e488cbb1006e19cede8de83b0427a3ee34a1bf767394d9d19961b1e2

Observation e1586907-3fee-4a25-9ac3-ee9629594389 · outbound

This paper cites library — ollama.com,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests library — ollama.com,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:15.010897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:10.646408Z digest=sha256:c99b746520d62185abd820dabf11050052244f549dba995fc2eac8f1c2ac9aef

Observation e677da6d-9c32-4cf3-94eb-98b5619c2300 · outbound

This paper cites LLM Leaderboard - Compare GPT-4o, Llama 3, Mistral, Ge mini & other models — Artificial Analysis — artificialanalysis.ai ,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests LLM Leaderboard - Compare GPT-4o, Llama 3, Mistral, Ge mini & other models — Artificial Analysis — artificialanalysis.ai ,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:14.985818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:10.761563Z digest=sha256:5247608ed6c4e48a39cbb233302337a73b76db3dbee78a98716c6113cb0106b7

Observation 68290ad6-fb57-428c-acc6-64253a76c5ef · outbound

This paper cites EvalPlus Leaderboard — evalplus.github.io,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests EvalPlus Leaderboard — evalplus.github.io,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:14.961494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:10.874956Z digest=sha256:d0cb14f463cf5fffe1d046116704addcb98cc506fb060e672c7ca394585972da

Observation 66082028-2d1a-41c2-a361-fadc7180d14f · outbound

This paper cites The Evolving Landscape of LLM Training Data — alibabacloud.com,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests The Evolving Landscape of LLM Training Data — alibabacloud.com,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:14.941545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:10.971989Z digest=sha256:8e94cfa581fe3c2c478861a3f87f8647d4ff8ac1450dbb1b968cf5780bd21b79

Observation 1e0c74e8-ce84-45d0-9edf-5a07670ef1ab · outbound

This paper cites EleutherAI/gpt-neo-2.7B · Hugging Face — huggingfac e.co,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests EleutherAI/gpt-neo-2.7B · Hugging Face — huggingfac e.co,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:14.920306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:11.095039Z digest=sha256:2794224c06ff7d9c6b8a7331373c3ef7f8544c3136d450f15d8891b0a454b6c7

Observation c1ff9d4c-fba2-48c2-83e9-0fbd95befe41 · outbound

This paper cites Mistral 7B — Mistral AI — mistral.ai,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests Mistral 7B — Mistral AI — mistral.ai,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:14.894136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:11.183607Z digest=sha256:cffb45954ab0654d08cdd77e79e354f4fd0cfa942bd35cb9cf2a3eaff03149cd

Observation 6decf3c5-3655-4524-8941-757fdf6f4483 · outbound

This paper cites nvidia/NVLM-D-72B · Hugging Face — huggingface.co,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests nvidia/NVLM-D-72B · Hugging Face — huggingface.co,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:14.872805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:11.300348Z digest=sha256:c81a209059e71a3b785345453418a7c1cb60165051926e7e0776f0e9c25708fb

Observation 91b85f4b-51eb-4794-a13b-38138cba9025 · outbound

This paper cites meta-llama/Llama-4-Maverick-17B-128E-Instruct · Hugging Face — huggingface.co,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests meta-llama/Llama-4-Maverick-17B-128E-Instruct · Hugging Face — huggingface.co,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:14.849405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:11.386946Z digest=sha256:f6dc69a33ee47153232c42f844b45f1afcb7ed6b5a9e38c265e70045c53d7e4d

Observation 687ec6db-1c66-4d68-9c1f-1375f57aa818 · outbound

This paper cites mistralai/Mistral-7B-Instruct-v0.3 · Hugging Face — huggingface.co,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests mistralai/Mistral-7B-Instruct-v0.3 · Hugging Face — huggingface.co,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:14.820333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:11.492658Z digest=sha256:c606243093ada4cd17f27df548d68d7f71858b5d99cc40d3a66e87b7b9bb5eed

Observation 7495472a-451e-488e-a857-7a3f3732330f · outbound

This paper cites The best small LLM-instruct models - a ehristoforu Col lec- tion — huggingface.co,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests The best small LLM-instruct models - a ehristoforu Col lec- tion — huggingface.co,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:14.795546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:11.676503Z digest=sha256:0848f88cb728f6337a5ecdb520b9007b7346660b74277ec623221d0844b133dc

Observation 2330b7a8-db15-4d28-81f2-74091a6dcd71 · outbound

This paper cites EvalPerf: Evaluating Language Models for Efficient Co de Generation — evalplus.github.io,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests EvalPerf: Evaluating Language Models for Efficient Co de Generation — evalplus.github.io,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:14.772174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:11.835725Z digest=sha256:2792eb568f8ddd8ee05551c9ddeff1c107ad1efc3ae5ca15c4e0fa59e4299e6a

Observation 3d60d9cb-d8ef-4d3a-ad78-4daa14dddb73 · outbound

This paper cites Qwen/Qwen2.5-7B-Instruct · Hugging Face — huggingfa ce.co,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests Qwen/Qwen2.5-7B-Instruct · Hugging Face — huggingfa ce.co,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:14.750648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:11.981521Z digest=sha256:b27beadb99998a2a6e01be2ed76ad116267025c96d97723431e7f493b4b4be24

Observation 744604f0-f53b-419e-9e85-7c1fb7ba7ec8 · outbound

This paper cites Qwen/Qwen2.5-Coder-32B-Instruct · Hugging Face — hu gging- face.co,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests Qwen/Qwen2.5-Coder-32B-Instruct · Hugging Face — hu gging- face.co,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:14.724289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:12.099770Z digest=sha256:9b460914db393b953b30e49f5a9ac50c09c4739b9323110a502cfa7e4189316b

Observation f9190357-cef7-43f8-ab3a-f8726caacaeb · outbound

This paper cites deepseek-ai/deepseek-coder-33b-instruct · Huggin g Face — huggingface.co,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests deepseek-ai/deepseek-coder-33b-instruct · Huggin g Face — huggingface.co,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:14.701451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:12.267027Z digest=sha256:a74f23a46dbcbe0e308209e43348bfaaa6091b0c469256d97dd6fafe1e566d2d

Observation 70b55802-5421-4d2f-9456-82be4e8e8665 · outbound

This paper cites Phind/Phind-CodeLlama-34B-v2 · Hugging Face — huggi ngface.co,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests Phind/Phind-CodeLlama-34B-v2 · Hugging Face — huggi ngface.co,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:14.676438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:12.425337Z digest=sha256:3eecac60a3dade905ca4bae770eaabe7f8b76a222a4fc61948cdca2efa7c49c8

Observation 91bbf7d9-0e02-45df-bf56-0a52ade23415 · outbound

This paper cites OpenAI Platform — platform.openai.com,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests OpenAI Platform — platform.openai.com,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:14.649577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:12.583469Z digest=sha256:7dd6de7b61d9af8dc5e76d4d371dae6bd6afcdfbbc604bb1b41285a4ae54dcd2

Observation 27d6fbd8-db8d-4c5c-8be9-a32d55743fe2 · outbound

This paper cites meta-llama/Llama-3.3-70B-Instruct · Hugging Face — hugging- face.co,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests meta-llama/Llama-3.3-70B-Instruct · Hugging Face — hugging- face.co,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:14.626814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:12.656027Z digest=sha256:d23371180ce70a23421a15a2d9ae96e32e5bee76e9907702162e60b6d4e856c6

Observation 3a3f97ce-6060-4d91-9ba5-346290e41778 · outbound

This paper cites nvidia/Llama-3.1-Nemotron-70B-Instruct-HF · Hugg ing Face — huggingface.co,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests nvidia/Llama-3.1-Nemotron-70B-Instruct-HF · Hugg ing Face — huggingface.co,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:14.600219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:12.793677Z digest=sha256:274e4809de8aee5fdac8cdf50a4dd115a0b9c7e7ffe6aedd653a7fa4d5d88f1b

Observation 0ae0fe1c-e412-46c0-9edd-712890205f04 · outbound

This paper cites meta-llama/Llama-3.1-405B-Instruct · Hugging Face — hugging- face.co,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests meta-llama/Llama-3.1-405B-Instruct · Hugging Face — hugging- face.co,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:14.559838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:12.933053Z digest=sha256:14554296b02cb5637d31935386bfe2ac0e724ca70232666a2d726720d30a4544

Observation 4a5942b4-8870-4441-87f7-72a1c655aad7 · outbound

This paper cites Introducing Claude 3.5 Sonnet — anthropic.com,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests Introducing Claude 3.5 Sonnet — anthropic.com,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:14.516971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:13.085098Z digest=sha256:8a8c9a142be964660af4df84cfa12f115c21f3b95a9bd38d0443f7ab057bef8e

Observation 12e9537c-d416-4c0c-93d4-c62a7a88a7cd · outbound

This paper cites GitHub - lmarena/arena-hard-auto: Arena-Hard-Auto : An auto- matic LLM benchmark. — github.com,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests GitHub - lmarena/arena-hard-auto: Arena-Hard-Auto : An auto- matic LLM benchmark. — github.com,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:14.489428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:13.237508Z digest=sha256:85b296dfe29b5b042faf7800522f7d74c81e1e2eb92ffbec3da65a764d5cdcc3

Observation 88fef424-19ec-4f88-b427-a99c79f165ce · outbound

This paper cites AlpacaEval Leaderboard — tatsu-lab.github.io,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests AlpacaEval Leaderboard — tatsu-lab.github.io,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:14.459155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:13.458467Z digest=sha256:57fe987b4c692c7d86737366032207508138010985a76472038b2030e7bde14f

Observation ca814d7c-8562-417c-a3f3-9dd13defb8eb · outbound

This paper cites MT Bench - a Hugging Face Space by lmsys — huggingface.c o,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests MT Bench - a Hugging Face Space by lmsys — huggingface.c o,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:14.419272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:13.539886Z digest=sha256:24145e9996d36fae5d00cc7cd3693d0e7caee82bf4467de16ab13487df28e55d

Observation 60bfe24d-843a-4810-ad62-10a80d216f2f · outbound

This paper cites an unresolved cited work.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:50:14.397870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:13.611912Z digest=sha256:8e96d3836fe9985175354fc5ee984ee7016fab745684027aba7faf70a55304cd

Observation 4b99a597-de35-4ef9-8556-60055d68f0d5 · outbound

This paper cites GitHub - OpenACCUserGroup/OpenACCV-V: OpenACC Vali dation and Verification Testsuite repository — github.com,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests GitHub - OpenACCUserGroup/OpenACCV-V: OpenACC Vali dation and Verification Testsuite repository — github.com,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:14.370788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:13.728694Z digest=sha256:4b9258b40c8b8695c7647a946e47728ed1821e934daae19cb18f405348845cad

Observation 769588b4-a41d-471e-9f19-5689c0b4e4a4 · outbound

This paper cites C-pack: P ackaged resources to advance general chinese embedding,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests C-pack: P ackaged resources to advance general chinese embedding,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:14.342774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:13.792328Z digest=sha256:3df2938098645a88d667dec2de11aa8f398914e2dd7c787046cf84fe01510b44

Observation 9d27ba48-7282-469f-a545-9c5fdbaf92c2 · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T12:50:13.845788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:50:13.845788Z digest=sha256:bc8a0e1a7f425aeb5064f283d64aa4c530b482099cb9de92fda98a8f0f4b0083

Observation 244755d9-17a2-4c5d-ba5d-17c2087382f2 · outbound

This paper cites Architecture - NERSC Documentation — docs.ner sc.gov,.

LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests Architecture - NERSC Documentation — docs.ner sc.gov,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:50:14.311046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T12:50:13.860383Z digest=sha256:9b4fc536cd4e7d7aac09bd8715d82702ae1e2a8ea9082340edf2a34b1f4d0669

Pith citing papers

No inbound Pith citation observations are available.