Pith. sign in

Paper Citation Record · LEDGER

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones

As of 8 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 0 inbound Pith citation observations for arXiv:2506.05641.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.05641 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:20:23.528274Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

44 of 44 outbound references displayed

  • verified exact4
  • verified fuzzy12
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6bee9b31-2d7a-411d-aea9-c1326c1eb397 · outbound

This paper cites write newline.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:23.314757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:23.314757Z digest=sha256:2f6249dddc7ac74d32e8e4f7a80f605ed7e05f949b2367fe71527c6b30927734

Observation 0d953fd4-9338-4c1a-abfe-bfa82963d17e · outbound

This paper cites Hyperstyle: Stylegan inversion with hypernetworks for real image editing.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Hyperstyle: Stylegan inversion with hypernetworks for real image editing

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:20:24.212617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:20:23.320469Z digest=sha256:32be64cf09f624201581426cfb9f057071e700fbe35a23221c6d964b36b0c405

Observation 916cede7-1b1b-463c-b869-797984d7d730 · outbound

This paper cites Generative pretraining from pixels.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Generative pretraining from pixels

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:20:24.196950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:20:23.325266Z digest=sha256:68ec81bc6703dfde6d22e26702659368b60a79db4e32a732d2817c0cadac28d8

Observation 715d2f6c-a967-4153-ac26-594df71f9257 · outbound

This paper cites an unresolved cited work.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:20:24.181790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:20:23.330119Z digest=sha256:80ebffeed1c260c6596d3730d1ff2b3f0123e392672666e190030fd45e33f38a

Observation 4cc4256c-a7b8-4d46-9afe-4135f99a4ca6 · outbound

This paper cites Streamlining Redundant Layers to Compress Large Language Models.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Streamlining Redundant Layers to Compress Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:23.334815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:23.334815Z digest=sha256:cbe77111238daa6cbe163f7791ecbafeaa3fc7d37644252897bcffc9b9c9c17a

Observation 275a067c-b886-4fd5-ae53-c8023bd7de5b · outbound

This paper cites Boosting Natural Language Generation from Instructions with Meta-Learning.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Boosting Natural Language Generation from Instructions with Meta-Learning

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-07T10:20:23.934119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:20:23.340824Z digest=sha256:1848361abdbdb13e6468685c7a3dff518e7876ffb2342ddaf3b1a186e91ab8e3

Observation d699fe1f-2075-47aa-898e-b4be65cea802 · outbound

This paper cites Imagenet: A large-scale hierarchical image database.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Imagenet: A large-scale hierarchical image database

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:23.346680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:23.346680Z digest=sha256:237ea04d1e577af1b47b0a019242306f4ba5b9a8cd95770bcd4950bc12e4f389

Observation b893d2be-ada6-4e84-9d62-fa5cbdf9b7b3 · outbound

This paper cites M., Tran, A.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones M., Tran, A

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:20:24.166641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:20:23.352660Z digest=sha256:3e4c56d2085832629c3b508961f4a112745ce065d17aae8f9413b29a16e49073

Observation 7a8ead0f-a2de-40b5-9a6a-d8562cdae8d8 · outbound

This paper cites HyperNetworks.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones HyperNetworks

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:23.357151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:23.357151Z digest=sha256:828114fc2b37922d9854ee466654f7fdb5776a27053d7be74bb9a384ae394264

Observation 5d7068fc-e95b-4a07-a3c5-91f7637e81c9 · outbound

This paper cites Hyperprompt: Prompt-based task-conditioning of transformers.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Hyperprompt: Prompt-based task-conditioning of transformers

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:23.362579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:23.362579Z digest=sha256:aa260ed16e8ae2dca529d2d2a65406cef32b52164be330b2194f25cd22ce9ddd

Observation cf7bb954-6bd5-489c-9a54-29b05de8df58 · outbound

This paper cites Hyperdecoders: Instance-specific decoders for multi-task NLP.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Hyperdecoders: Instance-specific decoders for multi-task NLP

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-07T10:20:23.823298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:20:23.367230Z digest=sha256:61e93adaf0dfb33d713836d650b04aa15ddab4ceff05c6b06c1336a0aad7d724

Observation 43eac6e2-e438-441b-a128-46d665c518d1 · outbound

This paper cites HINT: Hypernetwork Instruction Tuning for Efficient Zero- & Few-Shot Generalisation.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones HINT: Hypernetwork Instruction Tuning for Efficient Zero- & Few-Shot Generalisation

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-08-07T10:20:23.802443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:20:23.372648Z digest=sha256:14405e566027298087e777d2501f57e93bb5ae8d7a378c457b53039fbc83004a

Observation 86b4b6db-6770-48ab-86ef-a2137cca3ae3 · outbound

This paper cites Scaling up gans for text-to-image synthesis.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Scaling up gans for text-to-image synthesis

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:23.377682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:23.377682Z digest=sha256:24d37cbf83de9de7ae716c69e2399b975ba80271e26a96e9b72f5019aa2d1cf4

Observation de91b97f-d021-438a-984d-7936166ea8a4 · outbound

This paper cites W., and Romero Soriano, A.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones W., and Romero Soriano, A

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:20:24.133499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:20:23.382346Z digest=sha256:d7f290609bf3b34c1195f49d4b34a0d6e9d1baf1094f7d22dcffd80c76d03e01

Observation 21e1d121-5618-406f-ad35-ba98fc176130 · outbound

This paper cites MEND: Meta dEmonstratioN Distillation for Efficient and Effective In-Context Learning.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones MEND: Meta dEmonstratioN Distillation for Efficient and Effective In-Context Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:23.386981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:23.386981Z digest=sha256:2703aa804c888945d64f6bd0c560bba3372a9cd291a0ecf641e6f26ee5effd8f

Observation 2a602a46-a976-471e-9d45-cffde6b3741a · outbound

This paper cites Hart: Efficient adaptation via regularized autoregressive parameter generation.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Hart: Efficient adaptation via regularized autoregressive parameter generation

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:20:24.118920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:20:23.392191Z digest=sha256:5f3a6abd2cfb2de3ae4598ce15573e40d689fd6b67377518ca93c08742c40fd2

Observation 7543a2fd-2b45-4c67-a734-077fe83da26d · outbound

This paper cites Weight Distillation: Transferring the Knowledge in Neural Network Parameters.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Weight Distillation: Transferring the Knowledge in Neural Network Parameters

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:23.397051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:23.397051Z digest=sha256:fa6c112e1b196bdf1017787b2b0fa9bf77a229b34df68d9fa850e9f958eb2ccd

Observation 1995b6ab-1e12-460f-bfa4-9baa28950a70 · outbound

This paper cites and Wolf, L.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones and Wolf, L

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:20:24.104202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:20:23.401458Z digest=sha256:82db5622a8edf6e0392eb74b4759e52539ac1c2c0b791e3362d083a588681901

Observation 7be6e0e3-43ae-4cf7-91b1-9025df43c6bd · outbound

This paper cites LLM-Pruner: On the Structural Pruning of Large Language Models.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones LLM-Pruner: On the Structural Pruning of Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:23.406830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:23.406830Z digest=sha256:154046e5adc6b6de27d8960d7d842fe7fecd6e4cc6eaa4ae7145ab932bb7b1b4

Observation 4e24ea5f-34e9-46ab-b714-0d79f5411bd6 · outbound

This paper cites Parameter-efficient Multi-task Fine-tuning for Transformers via Shared Hypernetworks.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Parameter-efficient Multi-task Fine-tuning for Transformers via Shared Hypernetworks

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:23.411520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:23.411520Z digest=sha256:a28f0d7bfb0a6e30587fd2d30bdd5f1e5807aa67e15d067df334f44bb55d9e9c

Observation 5a6f7cea-f05d-48be-bcb6-273463a35af1 · outbound

This paper cites Learning to compress prompts with gist tokens.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Learning to compress prompts with gist tokens

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:23.416110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:23.416110Z digest=sha256:db75a5eeaab6f047e4a68bff2ebd3dfe59824795a597203f741c03e7dd03fb9c

Observation 6e90fde6-bde9-4a93-af25-d3cb7c3e90e9 · outbound

This paper cites Hyperseg: Patch-wise hypernetwork for real-time semantic segmentation.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Hyperseg: Patch-wise hypernetwork for real-time semantic segmentation

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:20:24.079707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:20:23.420255Z digest=sha256:24547148f90e50694339a8a8a810fcc510af7ebf9ace00d30f13387c177c27d4

Observation 18a11af9-940f-4006-aff5-4dc8e24ac5b9 · outbound

This paper cites Investigating the Effectiveness of HyperTuning via Gisting.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Investigating the Effectiveness of HyperTuning via Gisting

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-07T10:20:23.740739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:20:23.425212Z digest=sha256:a20d85a53d9f3e4ff18e6b16f5ee1af39b9a87db4cc8eba0e7491f769188fdc3

Observation fda714b2-9a56-44bc-8f33-d90ac59faf0c · outbound

This paper cites Hypertuning: Toward adapting large language models without back-propagation.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Hypertuning: Toward adapting large language models without back-propagation

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:20:24.066240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:20:23.429853Z digest=sha256:57acb62caca65d8c6270eede78c66338ac275d02487f5d89ee76fcfd006bd92b

Observation 5d23275c-0f2c-41db-aa97-77714cb430f8 · outbound

This paper cites Language models are unsupervised multitask learners.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Language models are unsupervised multitask learners

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:23.434658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:23.434658Z digest=sha256:47d5c2e7c2fd9e537fd8ba3836dc5ea7b35a8cdeb9bb1eb17fe6a638dafe9e2e

Observation e0fbae4a-22a1-4de7-aef7-a07c02a75157 · outbound

This paper cites an unresolved cited work.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:23.438991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:23.438991Z digest=sha256:43636090aa17fde843d675675f778350d4feb94485a223aecf64d622222e6535

Observation 11cb2282-9177-43e1-985b-80f56214a114 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones High-resolution image synthesis with latent diffusion models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:23.443960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:23.443960Z digest=sha256:dbaab669eadb952eb99273ca9f3b146ba1b70654716b1705d03645e53ebb2ea3

Observation e7114082-f259-48a8-87f8-7a48ba4fbcec · outbound

This paper cites Weight subcloning: direct initialization of transformers using larger pretrained ones.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Weight subcloning: direct initialization of transformers using larger pretrained ones

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:23.449360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:23.449360Z digest=sha256:66984ead74c1227b5790396760fed6b4a2621c2628fd4e4265e8e658719cf07d

Observation 8fedaeba-9c02-41dc-bbb2-d6558d894bd0 · outbound

This paper cites Implicit neural representations with periodic activation functions.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Implicit neural representations with periodic activation functions

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:20:24.024695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:20:23.454384Z digest=sha256:76c6f31ccfdb2a80967f81b63f9da05082cfcc9191ba75cff774eb4c7bbd8da9

Observation ee564d52-2cc6-430b-982e-d0a93c16ee8e · outbound

This paper cites K., Tabor, J., Trzci \'n ski, T., et al.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones K., Tabor, J., Trzci \'n ski, T., et al

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:20:24.010760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:20:23.459631Z digest=sha256:334565907e3bc4290c9288407810bea82009f4c2b1b4aba1b796fc29e74ef0e3

Observation 9a85040c-d887-405b-933e-432fe6ffff29 · outbound

This paper cites Online Adaptation of Language Models with a Memory of Amortized Contexts.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Online Adaptation of Language Models with a Memory of Amortized Contexts

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:23.464106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:23.464106Z digest=sha256:0fae3a0a3de37fe0b69445d5935d46d231380f5684697f480e30d27fb4a7779c

Observation dcd730a1-9d8a-4eb8-909d-20fbf5061899 · outbound

This paper cites A Survey on Transformer Compression.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones A Survey on Transformer Compression

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:23.468653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:23.468653Z digest=sha256:0450ba038c3828d97e420e20574c0ea058ed032932593adae3bcee54bdb49849

Observation c39bb78e-fb8c-454e-835d-73e4588b7c42 · outbound

This paper cites Hypergrid transformers: Towards a single model for multiple tasks.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Hypergrid transformers: Towards a single model for multiple tasks

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:20:23.996889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:20:23.474236Z digest=sha256:d0c662bf1f80d1263ab8020e8dd4820b1303fbe59ce59b2c2a7c632fa96ca5f7

Observation 32d1dfa4-b204-442a-bc48-ab1ce2b3b453 · outbound

This paper cites Neural discrete representation learning.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Neural discrete representation learning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:23.478703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:23.478703Z digest=sha256:b93f5fae0a6ef53d20dfbb28d10e8a867194e13e199382657a972593a7489f7a

Observation cbfe6837-2078-4eca-9805-9b989624ca24 · outbound

This paper cites Example-based hypernetworks for multi-source adaptation to unseen domains.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Example-based hypernetworks for multi-source adaptation to unseen domains

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:20:23.972941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:20:23.483346Z digest=sha256:6b793f66244dd3af8c27fb064eed56e9a0db528a464c4a31d12e5337949feb9a

Observation 683114a4-2df1-4d58-8ce5-d93720fcb929 · outbound

This paper cites Continual learning with hypernetworks.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Continual learning with hypernetworks

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:23.488791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:23.488791Z digest=sha256:ca98bd4fb7c56cccd29eb7a9a826c3b33069e68d2ce51d24722bc07cdc49eb09

Observation 9da367a4-6e4d-4f94-906e-d6db61245e1d · outbound

This paper cites Learngene: Inheriting Condensed Knowledge from the Ancestry Model to Descendant Models.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Learngene: Inheriting Condensed Knowledge from the Ancestry Model to Descendant Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:23.493636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:23.493636Z digest=sha256:36f4161b73b5591211bd4aa691952f839bb2caf8728d66c745f4b3c1eff9b9a5

Observation 672c3344-f7a9-49ba-95cb-9c901ddbcfe5 · outbound

This paper cites Model Compression and Efficient Inference for Large Language Models: A Survey.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Model Compression and Efficient Inference for Large Language Models: A Survey

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:23.498590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:23.498590Z digest=sha256:5b480875dc0fa0c5fa9481c79c7c5e4c6e1407d85b2eeaf9ebbea4bf80e4ef08

Observation ae4335e1-87c5-4a7d-b29c-32a55d66cbc2 · outbound

This paper cites Sheared LLaMA: Accelerating Language Model Pre-training via Structured Pruning.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Sheared LLaMA: Accelerating Language Model Pre-training via Structured Pruning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:23.503135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:23.503135Z digest=sha256:e572506b5a9ca41ab057656683d3dd230f6f75f6c020ce4b5c852af8212a9e37

Observation 0d68f145-6488-44c9-aefe-85cb92153452 · outbound

This paper cites Initializing Models with Larger Ones.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Initializing Models with Larger Ones

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:23.507736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:23.507736Z digest=sha256:fc329333da8f635d6a9951ddc768f25baa63a9f54e7613782fe3ffaab2596ed4

Observation 5c73e39c-e67a-4057-8758-4ba43943c7b7 · outbound

This paper cites Learning to Generate Task-Specific Adapters from Task Description.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Learning to Generate Task-Specific Adapters from Task Description

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:23.512717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:23.512717Z digest=sha256:29caaeee26b788fa47112ef9c27e99c072f119acb1863f7d43c1907bd67bb665

Observation 84fcaa74-7a48-4441-b0e1-242d3319bb33 · outbound

This paper cites Graph HyperNetworks for Neural Architecture Search.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Graph HyperNetworks for Neural Architecture Search

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:23.517790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:23.517790Z digest=sha256:c0478673c283b4dc9ad0276ecb8284afd934cb495d2f92429bbdd925b7128851

Observation da0e948b-b922-44f4-97a2-54290bd35183 · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones Adding conditional control to text-to-image diffusion models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:23.522276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:23.522276Z digest=sha256:15ce75e567ecf94ff2b1a17726d562325d1af29e3cdc1802782890235b0ce5be

Observation 0ebd817a-a2be-47aa-8a36-9740174c8161 · outbound

This paper cites HyperMoE: Towards Better Mixture of Experts via Transferring Among Experts.

Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones HyperMoE: Towards Better Mixture of Experts via Transferring Among Experts

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:23.528274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:23.528274Z digest=sha256:34e17348c847563e95fc8bf7d91cd96b59ad2e8c74d00d69962c86855681dc0a

Pith citing papers

No inbound Pith citation observations are available.