Pith. sign in

REVIEW 12 cited by

Relative representations enable zero-shot latent space communication

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2209.15430 v2 pith:IVKMT2P6 submitted 2022-09-30 cs.LG cs.AI

classification cs.LGcs.AI
keywords latentdataspacerepresentationstrainingarchitecturescommunicationneural
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Neural networks embed the geometric structure of a data manifold lying in a high-dimensional space into latent representations. Ideally, the distribution of the data points in the latent space should depend only on the task, the data, the loss, and other architecture-specific constraints. However, factors such as the random weights initialization, training hyperparameters, or other sources of randomness in the training phase may induce incoherent latent spaces that hinder any form of reuse. Nevertheless, we empirically observe that, under the same data and modeling choices, the angles between the encodings within distinct latent spaces do not change. In this work, we propose the latent similarity between each sample and a fixed set of anchors as an alternative data representation, demonstrating that it can enforce the desired invariances without any additional training. We show how neural architectures can leverage these relative representations to guarantee, in practice, invariance to latent isometries and rescalings, effectively enabling latent space communication: from zero-shot model stitching to latent space comparison between diverse settings. We extensively validate the generalization capability of our approach on different datasets, spanning various modalities (images, text, graphs), tasks (e.g., classification, reconstruction) and architectures (e.g., CNNs, GCNs, transformers).

Discussion (0). Sign in to comment.

Forward citations

Cited by 12 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Promptbreeder: Self-Referential Self-Improvement Via Prompt Evolution

    cs.CL 2023-09 unverdicted novelty 8.0 of 10

    Promptbreeder evolves both task prompts and the mutation prompts that improve them using LLMs, outperforming Chain-of-Thought and Plan-and-Solve on arithmetic and commonsense reasoning benchmarks.

  2. OPRD: On-Policy Representation Distillation

    cs.LG 2026-06 unverdicted novelty 7.0 of 10

    OPRD performs distillation in hidden-state space on on-policy data for deterministic gradients and better math benchmark performance, plus OPRD-Bridge for cross-architecture transfer via low-rank projectors.

  3. A Stitch in Time Saves Nine: Preserving Policy Compatibility Under Perception Updates in End-to-End Autonomous Driving

    cs.RO 2026-06 unverdicted novelty 6.0 of 10

    Lightweight model stitching preserves over 91% of driving performance in cross-domain perception updates for end-to-end autonomous driving, cutting adaptation time from 22 hours to under 1 hour.

  4. You Had One Job: Per-Task Quantization Using LLMs' Hidden Representations

    cs.CL 2025-11 conditional novelty 6.0 of 10

    TAQ estimates per-layer importance from hidden representations and output sensitivity on task calibration data to allocate mixed precision in a training-free PTQ setting, outperforming task-agnostic baselines on accur...

  5. Game-Theoretic Latent Space Alignment for Multi-user Semantic MIMO Communications

    cs.GT 2026-06 unverdicted novelty 5.0 of 10

    Formulates latent-space alignment in semantic MIMO interference networks as a non-cooperative game, derives closed-form linear transceiver solution, reduces to power-allocation game, and proves convergence to Nash equ...

  6. Improving Relative Representations with Learned Anchors and Whitened Inner Products

    cs.LG 2026-05 unverdicted novelty 5.0 of 10

    Learned anchors as semantic prototypes combined with whitened inner products improve relative representations, enabling nearly lossless zero-shot communication between heterogeneous neural models on vision and language tasks.

  7. You Had One Job: Per-Task Quantization Using LLMs' Hidden Representations

    cs.CL 2025-11 reject novelty 5.0 of 10

    A per-layer, per-task bit-allocation rule based on hidden-activation entropy and variance reportedly matches full-precision TriviaQA accuracy within about one point — but the paper's tables and abstract are internally...

  8. What are you sinking? A geometric approach on attention sink

    cs.LG 2025-08 reject novelty 5.0 of 10

    Attention sinks in transformers are reinterpreted as geometric reference frames, with three architecture-dependent types: centralized, distributed, and bidirectional.

  9. The Platonic Representation Hypothesis

    cs.LG 2024-05 unverdicted novelty 5.0 of 10

    Representations learned by large AI models are converging toward a shared statistical model of reality.

  10. A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions

    cs.CL 2023-11 unverdicted novelty 5.0 of 10

    The paper surveys hallucination in LLMs with an innovative taxonomy, factors, detection methods, benchmarks, mitigation strategies, and open research directions.

  11. There Will Be a Scientific Theory of Deep Learning

    stat.ML 2026-04 unverdicted novelty 2.0 of 10

    A mechanics of the learning process is emerging in deep learning theory, characterized by dynamics, coarse statistics, and falsifiable predictions across idealized settings, limits, laws, hyperparameters, and universa...

  12. Harnessing Multiple Large Language Models: A Survey on LLM Ensemble

    cs.CL 2025-02 unverdicted novelty 2.0 of 10

    A systematic survey of LLM ensemble methods organized into a taxonomy of ensemble-before-inference, ensemble-during-inference, and ensemble-after-inference stages, with review of benchmarks, applications, and future d...

Pith tools