Pith. sign in

Paper Citation Record · LEDGER

Parameter-Efficient Transformer Embeddings

As of 23 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 0 inbound Pith citation observations for arXiv:2505.02266.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.02266 v1

Coverage vector

measured 21 of 21 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T01:01:52.792606Z

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

21 of 21 outbound references displayed

  • verified exact1
  • verified fuzzy12
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7268a1ca-4253-40fc-bf29-18caabc840d6 · outbound

This paper cites Quantifying unique information.

Parameter-Efficient Transformer Embeddings Quantifying unique information

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T01:01:53.090177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T01:01:52.712574Z digest=sha256:d6e20ef9e15fe929448f9b1d2281c59b53044ac54d4d1a5bf7f139376affa933

Observation 6b468769-f974-4d6e-ab62-d843fd89866e · outbound

This paper cites A large an- notated corpus for learning natural language inference.

Parameter-Efficient Transformer Embeddings A large an- notated corpus for learning natural language inference

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T01:01:53.078977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T01:01:52.717018Z digest=sha256:87e670e84f585ca0b9ada8891d760392832286187c1c0caaf0dd9e38c8dc762b

Observation f927bbec-2066-416a-891e-e5843119dcd4 · outbound

This paper cites From Wide to Deep: Dimension Lifting Network for Parameter-efficient Knowledge Graph Embedding.

Parameter-Efficient Transformer Embeddings From Wide to Deep: Dimension Lifting Network for Parameter-efficient Knowledge Graph Embedding

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-08-16T01:01:52.934507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T01:01:52.721592Z digest=sha256:cbcb97443393cfd264b69b18e33c505a8ff3569228d9d1a75b0b0767e95bea99

Observation eff6d408-f644-4f66-914c-f1cf866fe0cb · outbound

This paper cites Bert: Pre-training of deep bidirectional transformers for language understan ding.

Parameter-Efficient Transformer Embeddings Bert: Pre-training of deep bidirectional transformers for language understan ding

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T01:01:53.066256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T01:01:52.726094Z digest=sha256:dba84fd61997b1567b0841faaea3229d32970f708602feda38418870df29b45b

Observation 3029fb09-463b-457a-bfa9-fb02e1c07f6c · outbound

This paper cites ALBERT: A Lite BERT for Self-supervised Learning of Language Representations.

Parameter-Efficient Transformer Embeddings ALBERT: A Lite BERT for Self-supervised Learning of Language Representations

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T01:01:52.730325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T01:01:52.730325Z digest=sha256:c40cac53e788c481b88dcc33ddd1ebbae2869d19548c36452dace933706340b2

Observation 5bce0399-1d85-444c-8fd7-3d5e998e7d3b · outbound

This paper cites Root Mean Square Layer Normalization.

Parameter-Efficient Transformer Embeddings Root Mean Square Layer Normalization

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T01:01:52.734574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T01:01:52.734574Z digest=sha256:92cf30f58457ba4759d9e50a79302101f52bf242eeb059f5325d75eae6e662a2

Observation a4bdd1ff-6e5c-4b2c-808a-710f90715f7a · outbound

This paper cites Decoupled weight deca y regularization.

Parameter-Efficient Transformer Embeddings Decoupled weight deca y regularization

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T01:01:53.054043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T01:01:52.739362Z digest=sha256:37b98233a998ee6ca1631aaa31f315f0394f1970dfb171e011f028a27ba0f6bf

Observation 8f503cdf-3b12-47c3-97a8-bf4648ce8a23 · outbound

This paper cites Learning Transferable Visual Models From Natural Language Supervision.

Parameter-Efficient Transformer Embeddings Learning Transferable Visual Models From Natural Language Supervision

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T01:01:52.743107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T01:01:52.743107Z digest=sha256:03b8e25b0593e2a9ba5e753da202e96b1945f69ad3b1fbfd96f7731acc7f0e63

Observation 8181d766-724f-44e4-8c91-916fe72b84bb · outbound

This paper cites ZeRO: Memory Optimizations Toward Training Trillion Parameter Models.

Parameter-Efficient Transformer Embeddings ZeRO: Memory Optimizations Toward Training Trillion Parameter Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T01:01:52.747306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T01:01:52.747306Z digest=sha256:4b272b208cd5086f5c36b9d3d086b2909f1959063e146362ba4247acb3f2f179

Observation 9519b7a1-914e-41b2-aec8-e7f355a963b0 · outbound

This paper cites Neur al machine translation of rare words with subword units.

Parameter-Efficient Transformer Embeddings Neur al machine translation of rare words with subword units

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T01:01:53.041960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T01:01:52.751113Z digest=sha256:dc710a565ca5017a5540025202bdeaaa296135d1ff34496326b79fb1a0c797d1

Observation a3f71735-eac3-4f28-b396-fda14e2db219 · outbound

This paper cites Glu variants improve transformer.

Parameter-Efficient Transformer Embeddings Glu variants improve transformer

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T01:01:53.028979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T01:01:52.754626Z digest=sha256:c2253d99168674b52ca2ecc5f6c8158e283a55c2ec716a586deb0ac49750108c

Observation 9c7d9da3-2e58-4797-a5e0-4305e2ac90ed · outbound

This paper cites Q-bert: Hessian-based ul tra low precision quantization of bert.

Parameter-Efficient Transformer Embeddings Q-bert: Hessian-based ul tra low precision quantization of bert

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T01:01:53.017014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T01:01:52.758537Z digest=sha256:96fc18cd709e5cdc653bdec2dde540c5cb046c464c7ebcf4ab961618c7caaa85

Observation bf7b2bdb-be15-4a81-bfd0-f4ff4b7c9e40 · outbound

This paper cites The Gram-Schmidt Walk: A Cure for the Banaszczyk Blues.

Parameter-Efficient Transformer Embeddings The Gram-Schmidt Walk: A Cure for the Banaszczyk Blues

Reference 13

Resolution
metadata mismatch
local_arxiv, observed 2026-08-16T01:01:52.865197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T01:01:52.762021Z digest=sha256:01f7df06481b2f3552072ace3743d12e56392ddb06abb0d7328344e15b222623

Observation 85ed60b4-1055-4f70-8cb8-5f08c554a975 · outbound

This paper cites RoFormer: Enhanced Transformer with Rotary Position Embedding.

Parameter-Efficient Transformer Embeddings RoFormer: Enhanced Transformer with Rotary Position Embedding

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T01:01:52.765804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T01:01:52.765804Z digest=sha256:04fcb4f3d9fdd79d51a80241a922f46e1ce5d2f9bdb7117cd7cc7200f1902e75

Observation 3402f499-17f1-4bac-9b74-3e7a488178ad · outbound

This paper cites Hash embeddings for efficient word representations.

Parameter-Efficient Transformer Embeddings Hash embeddings for efficient word representations

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T01:01:53.003744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T01:01:52.769155Z digest=sha256:beff75c90723d8be8a084ec932b5f459a231d02acb7ba5f6026b127abd47af31

Observation 53b2c9fa-9ebf-434b-a002-6125e3b27baf · outbound

This paper cites Represe ntation learning with contrastive predictive coding.

Parameter-Efficient Transformer Embeddings Represe ntation learning with contrastive predictive coding

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T01:01:52.989063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T01:01:52.772557Z digest=sha256:64105d907876820dd9094101ed0a5a695cac2a069698456da95687f3baa7500c

Observation d6902805-443d-4e1d-b822-6fd4001366de · outbound

This paper cites Structured embedding compression, 20 20.

Parameter-Efficient Transformer Embeddings Structured embedding compression, 20 20

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T01:01:52.975963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T01:01:52.776509Z digest=sha256:24da9d1f0e2aee7acdbc7d79ffa0230001bc52d8c7bc599e7fcc7126914a3eab

Observation 7032d44b-8eef-4996-88b7-f7d9691f1c74 · outbound

This paper cites A br oad-coverage challenge cor- pus for sentence understanding through inference.

Parameter-Efficient Transformer Embeddings A br oad-coverage challenge cor- pus for sentence understanding through inference

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T01:01:52.960466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T01:01:52.780103Z digest=sha256:8b777c468fb5a497d4f4a51d36391b4aada16ab04ca73140177d7e3b629de2b5

Observation 3254060e-3a39-454a-98a8-4d452d197d1e · outbound

This paper cites Nonnegative Decomposition of Multivariate Information.

Parameter-Efficient Transformer Embeddings Nonnegative Decomposition of Multivariate Information

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T01:01:52.783725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T01:01:52.783725Z digest=sha256:c20ecfec4c87e1e87ac5230c9ed973a86534a22047643d1d2425e4353e0cac81

Observation 3dcfb862-b619-4c9d-8ce5-4afebc9c34d2 · outbound

This paper cites TensorGPT: Efficient Compression of Large Language Models based on Tensor-Train Decomposition.

Parameter-Efficient Transformer Embeddings TensorGPT: Efficient Compression of Large Language Models based on Tensor-Train Decomposition

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T01:01:52.788232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T01:01:52.788232Z digest=sha256:5389154f254d5eff08b7563a18ee7d5c00297dd382d444c22f58a830c5c9b016

Observation e625b904-9e1a-4b55-9b8c-b97057112d23 · outbound

This paper cites Adaptively-masked twins-based layer f or efficient embeddings, 2021.

Parameter-Efficient Transformer Embeddings Adaptively-masked twins-based layer f or efficient embeddings, 2021

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T01:01:52.947546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T01:01:52.792606Z digest=sha256:7b9b03329dd70ce8d9bc157c4ffb0722994e955d134fb7522186f241b295d2c9

Pith citing papers

No inbound Pith citation observations are available.