Pith. sign in

REVIEW 10 cited by

Getting aligned on representational alignment

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2310.13018 v3 pith:MVIS7GEQ submitted 2023-10-18 q-bio.NC cs.AIcs.LGcs.NE

classification q-bio.NCcs.AIcs.LGcs.NE
keywords alignmentrepresentationalfieldsrepresentationsresearchsystemsanothercognitive
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Biological and artificial information processing systems form representations of the world that they can use to categorize, reason, plan, navigate, and make decisions. How can we measure the similarity between the representations formed by these diverse systems? Do similarities in representations then translate into similar behavior? If so, then how can a system's representations be modified to better match those of another system? These questions pertaining to the study of representational alignment are at the heart of some of the most promising research areas in contemporary cognitive science, neuroscience, and machine learning. In this Perspective, we survey the exciting recent developments in representational alignment research in the fields of cognitive science, neuroscience, and machine learning. Despite their overlapping interests, there is limited knowledge transfer between these fields, so work in one field ends up duplicated in another, and useful innovations are not shared effectively. To improve communication, we propose a unifying framework that can serve as a common language for research on representational alignment, and map several streams of existing work across fields within our framework. We also lay out open problems in representational alignment where progress can benefit all three of these fields. We hope that this paper will catalyze cross-disciplinary collaboration and accelerate progress for all communities studying and developing information processing systems.

Discussion (0). Sign in to comment.

Forward citations

Cited by 10 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Beyond Color Geometry: Evaluating Human-Like Color Representations in Vision Models

    cs.CV 2026-07 conditional novelty 6.0 of 10

    MAE encoders show significantly stronger alignment with human fuzzy color categories than other ViTs, beyond what perceptual color geometry explains.

  2. Accuracy Does Not Guarantee Human-Likeness: Cross-Domain Human-Centered Benchmark in Monocular Depth Estimation

    cs.CV 2025-12 conditional novelty 6.0 of 10

    Across 69 monocular depth estimators, human-likeness of error patterns peaks near human-level accuracy and declines for the most accurate models: accuracy does not guarantee human-like depth perception.

  3. The Geometry of Grokking: Norm Minimization on the Zero-Loss Manifold

    cs.LG 2025-11 conditional novelty 6.0 of 10

    Post-memorization learning in grokking is equivalent to minimizing the weight norm on the zero-loss manifold, with a closed-form approximation for two-layer networks.

  4. Can Biologically Plausible Temporal Credit Assignment Rules Match BPTT for Neural Similarity? E-prop as an Example

    cs.NE 2025-06 conditional novelty 6.0 of 10

    At matched task accuracy, e-prop trained RNNs reach neural data similarity comparable to BPTT trained RNNs on Mante 2013 and Sussillo 2015 datasets, with initialization and architecture influencing similarity more tha...

  5. The Principle of Isomorphism: A Theory of Population Activity in Grid Cells and Beyond

    q-bio.NC 2025-10 conditional novelty 5.0 of 10

    Grid-cell population activity is toroidal because path integration and the neural metric both require a compact flat/commutative structure, and hexagonal single-cell fields emerge only in a narrow range of torus sizes.

  6. Linear Spatial World Models Emerge in Large Language Models

    cs.AI 2025-06 reject novelty 5.0 of 10

    Spatial relation words in LLaMA and Qwen models form antipodal, roughly orthogonal directions in a low-dimensional subspace, and steering along these directions changes the model's output.

  7. Using LLMs to Advance the Cognitive Science of Collectives

    q-bio.NC 2025-05 conditional novelty 5.0 of 10

    A position paper arguing that LLMs can help cognitive scientists study collective behavior along structural, interactional, and individual complexity axes, with cautions about bias and reproducibility.

  8. Evaluating Steering Techniques using Human Similarity Judgments

    cs.AI 2025-05 conditional novelty 5.0 of 10

    Prompt-based steering outperformed activation-based steering on accuracy, but no method produced representations well aligned with human judgments, especially for size.

  9. The Representational Alignment between Humans and Language Models is implicitly driven by a Concreteness Effect

    cs.CL 2025-05 conditional novelty 5.0 of 10

    For 40 German nouns, human similarity judgments and language model embeddings align mainly because both are organized along the concreteness dimension, not along frequency, length, or orthographic similarity.

  10. Representation biases: will we achieve complete understanding by analyzing representations?

    q-bio.NC 2025-07 accept novelty 4.0 of 10

    Feature representation biases in trained models can distort PCA, regression, RSA, and model-brain comparisons, so representational analyses may not reveal all of a system's computations.

Pith tools