multi-model-geometric-convergence

IN derived (depth 1)

Created 2026-08-25T03:00:40+00:00 · Reviewed 2026-08-25T04:28:09+00:00

Both the polytope/orthogonality geometry (Park, validated on Gemma-2B and LLaMA-3-8B) and sparse feature structure (SAE, universal across architectures) converge on the finding that transformer representation spaces carry model-independent geometric invariants.

Summary

Two independent methods for peering into how transformers organize concepts in their internal space agree on the same conclusion: the geometric layout of meaning is shaped by the structure of language itself, not by which particular model architecture you plug in. This means findings about how concepts relate to one another in a model's head can be transferred across architectures, and it points toward a shared conceptual geometry that different models are all independently converging on.

Justifications

This belief has 2 justifications — it is IN if any one holds.

SL — Park's Theorem 8 orthogonality holds across Gemma-2B (2B params) and LLaMA-3-8B; SAEs trained on GPT-2, Mistral, and Claude produce mutually more similar features than their random baselines. Two independent geometric analyses confirm architecture-invariant structure.

Antecedents (all must be IN):

  • IN park-2025-validation-models-wordnet — The results were empirically validated on Gemma-2B and LLaMA-3-8B using 593 noun and 364 verb WordNet synsets (retained if containing ≥50 words in the model vocabulary).
SL — Park's Theorem 8 orthogonality holds across Gemma-2B (2B params) and LLaMA-3-8B; SAEs trained on GPT-2, Mistral, and Claude produce mutually more similar features than their random baselines. Two independent geometric analyses confirm architecture-invariant structure.

Antecedents (all must be IN):

  • IN sae-universality-across-models — SAEs applied to different transformer models produce mostly similar features—more similar to each other than to their own model's neurons—suggesting features reflect data structure rather than architecture

Dependents

These beliefs depend on this one: