Pith. sign in

Paper Citation Record · LEDGER

Assessing Reliability of BERT-Based Models on Question Answering Tasks

As of 20 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 0 inbound Pith citation observations for arXiv:2608.10806.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.10806 v1

Coverage vector

measured 32 of 32 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T17:07:13.928037Z

measured 32 of 32 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

32 of 32 outbound references displayed

  • verified exact1
  • verified fuzzy22
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1a1be73d-317c-4a57-98c0-fb99681b7f07 · outbound

This paper cites A deep network model for paraphrase detection in short text 29 messages.Information Processing & Management, 54(6):922–937, 2018.

Assessing Reliability of BERT-Based Models on Question Answering Tasks A deep network model for paraphrase detection in short text 29 messages.Information Processing & Management, 54(6):922–937, 2018

Reference 1

Resolution
verified exact
doi, observed 2026-08-12T17:07:13.981666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T17:07:13.796233Z digest=sha256:a55d074061e7acd9852c823f2546050cbfb28561d49df3177e0663b128840aa8

Observation c0d6c6af-989f-4cad-9eb2-fd5b364f1c3a · outbound

This paper cites Evaluating ChatGPT as a Question Answering System: A Comprehensive Analysis and Comparison with Existing Models.

Assessing Reliability of BERT-Based Models on Question Answering Tasks Evaluating ChatGPT as a Question Answering System: A Comprehensive Analysis and Comparison with Existing Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T17:07:13.802210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:07:13.802210Z digest=sha256:a31a7deba0350d6deb3af47731187c67747f926cac33915f55d4975ff2509999

Observation 809d3d01-6bef-4d9c-bd9e-9eae6d8a8114 · outbound

This paper cites Understanding dropout.Advances in neural information processing systems, 26, 2013.

Assessing Reliability of BERT-Based Models on Question Answering Tasks Understanding dropout.Advances in neural information processing systems, 26, 2013

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:07:14.387198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T17:07:13.807297Z digest=sha256:78a35c97327ae5d28c79df8ce002b19f48188c800c5cfc2279cd61ff3cebe5da

Observation fef1ec1e-a644-4c25-954c-594549ec1180 · outbound

This paper cites Language models are few-shot learners.Advances in neural information processing systems, 33:1877–1901, 2020.

Assessing Reliability of BERT-Based Models on Question Answering Tasks Language models are few-shot learners.Advances in neural information processing systems, 33:1877–1901, 2020

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T17:07:13.811516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:07:13.811516Z digest=sha256:d4cc2524013cfd3625836c3eb60b784149b4d804edccebb0a61afdf56c3c994a

Observation 61c26466-de22-4f5c-b7be-709bbd1b0260 · outbound

This paper cites QuAC: Question answering in context.

Assessing Reliability of BERT-Based Models on Question Answering Tasks QuAC: Question answering in context

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:07:14.367584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T17:07:13.815803Z digest=sha256:5a6a86433dbc63f07aab2873ec2db3ee9db170f3c7fbaec48fb6bdc3e500b1b3

Observation ace65953-b12c-4f6a-91f8-cd1ec0e1ac54 · outbound

This paper cites BERT: Pre-training of deep bidirectional transformers for language un- derstanding.

Assessing Reliability of BERT-Based Models on Question Answering Tasks BERT: Pre-training of deep bidirectional transformers for language un- derstanding

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:07:14.353272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T17:07:13.820693Z digest=sha256:3f5f1761dddc9916c97801cfa0a0250c1689f8d6ae76b401346416c3c2075d33

Observation db6180b0-1349-45fd-8a43-e49130a750d6 · outbound

This paper cites A review on different methods of paraphras- ing.

Assessing Reliability of BERT-Based Models on Question Answering Tasks A review on different methods of paraphras- ing

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:07:14.340648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T17:07:13.825247Z digest=sha256:1b0c522504e69797c70cf37ad2a0f10d5564254eba1e2d2c0ea02261be4593c5

Observation f1a20b2e-63dd-4963-b858-61a04f68f717 · outbound

This paper cites Natural language based analysis of SQuAD: An analytical approach for BERT.Expert Systems with Applications, 195:116592, 2022.

Assessing Reliability of BERT-Based Models on Question Answering Tasks Natural language based analysis of SQuAD: An analytical approach for BERT.Expert Systems with Applications, 195:116592, 2022

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:07:14.329348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T17:07:13.829444Z digest=sha256:9a2465806155e0a3f7a62e4fde59e6ccc7a8cbd7205f091dd7af151c0742c089

Observation 1c6a893d-71ea-4184-b24e-4d76ec5f1ca1 · outbound

This paper cites A survey on hallucination in large language models: Principles, tax- onomy, challenges, and open questions.ACM Transactions on Information Systems, 43(2):1–55, 2025.

Assessing Reliability of BERT-Based Models on Question Answering Tasks A survey on hallucination in large language models: Principles, tax- onomy, challenges, and open questions.ACM Transactions on Information Systems, 43(2):1–55, 2025

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:07:14.317049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T17:07:13.833219Z digest=sha256:673f7b05eef5038d637d59218c47a21610edbb7150948588f92f8a8efca29aa0

Observation bbbac51a-40b2-4f7e-a93b-09afb0ad4b3a · outbound

This paper cites AlBERT: A lite BERT for self-supervised learning of language representations.International Conference on Learning Representations., 2020.

Assessing Reliability of BERT-Based Models on Question Answering Tasks AlBERT: A lite BERT for self-supervised learning of language representations.International Conference on Learning Representations., 2020

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:07:14.303939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T17:07:13.837113Z digest=sha256:fa01ebabad5e71d3a38b89fd1bcb5ef16dee3c27c1642f1bf28d07ac1c887626

Observation 6f110daa-c66d-4820-813e-d57af02c7443 · outbound

This paper cites The measurement of observer agree- ment for categorical data.biometrics, pages 159–174, 1977.

Assessing Reliability of BERT-Based Models on Question Answering Tasks The measurement of observer agree- ment for categorical data.biometrics, pages 159–174, 1977

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:07:14.291640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T17:07:13.840547Z digest=sha256:cacc812a0b872b2a3ac57af8ba928beb2fcc42f290f89bb072da835e7b4d701d

Observation 44efd00f-0caf-4c1e-bc4d-24bbcd777ece · outbound

This paper cites Ensemble ALBERT on SQuAD 2.0.

Assessing Reliability of BERT-Based Models on Question Answering Tasks Ensemble ALBERT on SQuAD 2.0

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T17:07:13.844836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:07:13.844836Z digest=sha256:dab5a7568c8188e20d280009648d10a0547cfa906a533cf7c41179b67f848142

Observation 6431219a-9d4f-499a-bc88-067e32f36e0e · outbound

This paper cites How can rec- ommender systems benefit from large language models: A survey.ACM Transactions on Information Systems, 43(2):1–47, 2025.

Assessing Reliability of BERT-Based Models on Question Answering Tasks How can rec- ommender systems benefit from large language models: A survey.ACM Transactions on Information Systems, 43(2):1–47, 2025

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:07:14.280073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T17:07:13.849240Z digest=sha256:9ea542f1e42eac6c53ebc2c9089dfa311a39b938f11cd66156cfbe9d1d5c7120

Observation e7313363-cee6-4234-85c5-b501c2b4225c · outbound

This paper cites RoBERTa: A Robustly Optimized BERT Pretraining Approach.

Assessing Reliability of BERT-Based Models on Question Answering Tasks RoBERTa: A Robustly Optimized BERT Pretraining Approach

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T17:07:13.853253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:07:13.853253Z digest=sha256:17a731d0d7cf3ddb59e15515fc978274501e63edb465f34032f6b92c6da58dbf

Observation 256fdb18-4bbf-4da5-afd8-a82554963bba · outbound

This paper cites Le- dlcm: Decoupled learner and course modeling with large language mod- els for enhanced course recommendation.Knowledge-Based Systems, page 115135, 2025.

Assessing Reliability of BERT-Based Models on Question Answering Tasks Le- dlcm: Decoupled learner and course modeling with large language mod- els for enhanced course recommendation.Knowledge-Based Systems, page 115135, 2025

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:07:14.267735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T17:07:13.857645Z digest=sha256:9cb2d3877b1634d12063c6e794781e3d3fc18d08f423d00d80e9d5a21d78bf0c

Observation cb0c8fec-297e-4c44-9b74-4d9a180f41b1 · outbound

This paper cites Prediction uncertainty estimation for hate speech classi- fication.

Assessing Reliability of BERT-Based Models on Question Answering Tasks Prediction uncertainty estimation for hate speech classi- fication

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:07:14.254860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T17:07:13.861205Z digest=sha256:86900847f1c9f1953171858ce193dc9014a99383c1271ee9c80a4d13808233ce

Observation bfd236f4-cd8f-4427-9e37-2d7407aea040 · outbound

This paper cites To BAN or not to BAN: Bayesian attention networks for reliable hate speech detection.Cognitive Computation, 14(1):353–371, 2022.

Assessing Reliability of BERT-Based Models on Question Answering Tasks To BAN or not to BAN: Bayesian attention networks for reliable hate speech detection.Cognitive Computation, 14(1):353–371, 2022

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:07:14.239856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T17:07:13.864818Z digest=sha256:07c5fcb86e90dabdbc8cf274e774da960505e630c7acb1f3aff742d635bf9b88

Observation 4f963bdd-a472-4ff5-8847-e890ca0386b9 · outbound

This paper cites Extractive text summarization.International Journal of Current Engineering and Technology, 4(2), 2014.

Assessing Reliability of BERT-Based Models on Question Answering Tasks Extractive text summarization.International Journal of Current Engineering and Technology, 4(2), 2014

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:07:14.222950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T17:07:13.868822Z digest=sha256:e927374ea617c4c26178a20c38f3cbb020cd2cf32e46cb8b9baa912f0cb04793

Observation 3d77e757-ddd7-4a69-87f7-00f970367124 · outbound

This paper cites Transformer models used for text- based question answering systems.Applied Intelligence, 53(9):10602–10635, 2023.

Assessing Reliability of BERT-Based Models on Question Answering Tasks Transformer models used for text- based question answering systems.Applied Intelligence, 53(9):10602–10635, 2023

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:07:14.207782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T17:07:13.872173Z digest=sha256:c1457ffe91ee685ed701a0449ef626806a1b320310ff98e2f1ee690f29f85fa2

Observation 70b058c0-b6d8-4991-b803-5ca724cee30b · outbound

This paper cites Comparative analysis of state-of-the-art Q&A models: BERT, RoBERTa, DistilBERT, and ALBERT on SQuAD v2 dataset.

Assessing Reliability of BERT-Based Models on Question Answering Tasks Comparative analysis of state-of-the-art Q&A models: BERT, RoBERTa, DistilBERT, and ALBERT on SQuAD v2 dataset

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:07:14.191199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T17:07:13.876174Z digest=sha256:c76c98369a0c83c63412ea1b8d26866ddc6b953a78de90592853b7aeb949d72b

Observation df9fb125-105e-4acb-912d-60417516aef7 · outbound

This paper cites A Comparative Study of Transformer-Based Language Models on Extractive Question Answering.

Assessing Reliability of BERT-Based Models on Question Answering Tasks A Comparative Study of Transformer-Based Language Models on Extractive Question Answering

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T17:07:13.879930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:07:13.879930Z digest=sha256:95a68099a8b4ff7fbcb5361379ed4f1183160c152aad46e520c467a1e4d305b5

Observation 1464a2e5-5be2-4b48-8107-a2b35a8a81b6 · outbound

This paper cites Measuring Reliability of Large Language Models through Semantic Consistency.

Assessing Reliability of BERT-Based Models on Question Answering Tasks Measuring Reliability of Large Language Models through Semantic Consistency

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T17:07:13.884589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:07:13.884589Z digest=sha256:ec0c2b9c63ec75a376ecc05f86f67a28d4b77580a36eb39122e11605c2bfb3f1

Observation 19cf3f82-9002-42f9-acce-68aa7cca06f3 · outbound

This paper cites Semantic Consistency for Assuring Reliability of Large Language Models.

Assessing Reliability of BERT-Based Models on Question Answering Tasks Semantic Consistency for Assuring Reliability of Large Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T17:07:13.888804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:07:13.888804Z digest=sha256:58f02f94fa38588f0293b22c9daaa075a21061e51313e0c8c7f43b5ed4f64b6b

Observation 031251a5-ce00-49fd-942d-adc4b08a7d4a · outbound

This paper cites SQuAD: 100,000+ questions for machine comprehension of text.

Assessing Reliability of BERT-Based Models on Question Answering Tasks SQuAD: 100,000+ questions for machine comprehension of text

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T17:07:13.893018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:07:13.893018Z digest=sha256:985ea4a901c705e6cef24781cceadc75cbd82e8f3234e26b908d3a8a5082478b

Observation 38b4b519-2685-4ac9-8a9c-b66c043b5284 · outbound

This paper cites Comparative analysis of transformer based models for question answering.

Assessing Reliability of BERT-Based Models on Question Answering Tasks Comparative analysis of transformer based models for question answering

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:07:14.171760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T17:07:13.896988Z digest=sha256:87fd8295100f72f0c330babe83888e8cea46fc922bb70f4b23b7ea1424fd85df

Observation 6d96e53a-adc5-4eea-9b88-941fcc0de66f · outbound

This paper cites Llm4rec: a comprehensive sur- vey on the integration of large language models in recommender sys- tems—approaches, applications and challenges.Future Internet, 17(6):252, 2025.

Assessing Reliability of BERT-Based Models on Question Answering Tasks Llm4rec: a comprehensive sur- vey on the integration of large language models in recommender sys- tems—approaches, applications and challenges.Future Internet, 17(6):252, 2025

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:07:14.157317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T17:07:13.900875Z digest=sha256:4a9b4aadbe94124bba5b74a322db2db55a091419e3ce165056220d30007702b3

Observation 8e758bb8-2414-493b-9e79-6fb320c7ac7c · outbound

This paper cites How does BERT answer questions? A layer-wise analysis of transformer representations.

Assessing Reliability of BERT-Based Models on Question Answering Tasks How does BERT answer questions? A layer-wise analysis of transformer representations

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:07:14.141090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T17:07:13.905068Z digest=sha256:bcd81db1b43f91ef681479b853170779a25ca3bdef6f0e76b9e1ad274d61ea36

Observation c8df97b6-44f9-4f9f-a6e3-fc43e0bd5e64 · outbound

This paper cites Attention is all you need.Advances in Neural Information Processing Systems, 2017.

Assessing Reliability of BERT-Based Models on Question Answering Tasks Attention is all you need.Advances in Neural Information Processing Systems, 2017

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:07:14.127832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T17:07:13.909022Z digest=sha256:b3a20987d71a26940db6baccd67d68e74d2669653dc6006ab9fd4a47e5dd403c

Observation 6347be6f-55f6-4047-b467-32c692df14b7 · outbound

This paper cites Assessing factual reliability of large language model knowledge.

Assessing Reliability of BERT-Based Models on Question Answering Tasks Assessing factual reliability of large language model knowledge

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:07:14.115356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T17:07:13.913391Z digest=sha256:10d70b83779b8e29cb4ab0b1816c5f94b4639748ab692c7cb59a3b5caca3cada

Observation d5a1f25d-9ccd-4c88-a047-9ac803115b01 · outbound

This paper cites A survey on large language models for recommendation.World Wide Web, 27(5):60, 2024.

Assessing Reliability of BERT-Based Models on Question Answering Tasks A survey on large language models for recommendation.World Wide Web, 27(5):60, 2024

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T17:07:13.916977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:07:13.916977Z digest=sha256:d54b8d49103d1693595c559463893e25f385b8bd0f6113dcfe77528f587130ab

Observation 45d470b3-ae63-4002-befa-9877fe02bbe9 · outbound

This paper cites Paraphrasing for style.

Assessing Reliability of BERT-Based Models on Question Answering Tasks Paraphrasing for style

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:07:14.096741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T17:07:13.920638Z digest=sha256:6b5d394cbd46a7e3068bb7e872cd3a3d64ffb1c31c7d3c24e40b74ab00357719

Observation f08d81ca-eb04-4ba1-8a8f-e35d39b4b52d · outbound

This paper cites Pegasus: Pre- training with extracted gap-sentences for abstractive summarization.

Assessing Reliability of BERT-Based Models on Question Answering Tasks Pegasus: Pre- training with extracted gap-sentences for abstractive summarization

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:07:14.085150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T17:07:13.928037Z digest=sha256:47a2942b3b266edde6bc7000e4759fad3384b7f96ea19e895893ed87c9382918

Pith citing papers

No inbound Pith citation observations are available.