DSReg: Provably Recovering Individual World Latents without Reconstruction
Abstract
Methods that recover individual latent variables of the world, from nonlinear ICA to dictionary learning and causal representation learning, anchor the latents to observations through reconstruction, auxiliary supervision, or distributional asymmetries such as non-Gaussianity. Methods without these anchors, including joint-embedding predictive architectures (JEPAs), identify the latent state only up to a linear transformation, so individual latents remain mixed. We close this gap: individual world latents can be provably recovered with no reconstruction, no decoder, and no labels. The key condition is Structural Diversity: different latents leave distinct dependency footprints on observations, just as no two snowflakes are alike. Building on the linear identifiability that LeJEPA provides, we prove that under Structural Diversity, DSReg (Dependency-Sparsity Regularization) recovers individual world latents up to signed permutation, without reconstruction or a decoder. It applies post hoc to any linearly identified representation, reusing trained checkpoints at no loss over joint training, and establishes the first fully identifiable JEPA that recovers every world latent. Moreover, as a condition on dependency footprints, Structural Diversity is strictly weaker than all structural conditions of prior identifiable latent variable models. Across synthetic regimes, world model probes, learned visual encoders, and external renderers, DSReg preserves dense prediction while improving individual-latent recovery and downstream use with scales.
Community
Is it possible to provably recover individual latent variables of the true world, even without reconstruction (e.g., JEPA)?
Yes, with DSReg, a simple regularization that can be applied post hoc to your pretrained model!
Project: https://dsreg.github.io
This is an automated message from the Librarian Bot. I found the following papers similar to this paper.
The following papers were recommended by the Semantic Scholar API
- Frozen in a Frame: The Velocity Blind Spot in JEPA World Models (2026)
- Low-Interaction-Rank Learning: Unifying Multiplicative Dual-Encoder Heads (2026)
- Beyond Gaussian Worlds: Latent Geometry Matters for JEPAs (2026)
- Identifiable World Models from Pretrained Diffusion Representations (2026)
- Algebraic Consistency Alone Does Not Certify Temporal Structure in Latent Action Models (2026)
- Jigsaw-CRL: Recovering Global Latent Causal Order from Fragmented Multi-Client Interventions (2026)
- Causal Representation Learning with Instantaneous and Lagged Relations via Nonstationarity (2026)
Please give a thumbs up to this comment if you found it helpful!
If you want recommendations for any Paper on Hugging Face checkout this Space
You can directly ask Librarian Bot for paper recommendations by tagging it in a comment: @librarian-bot recommend
Get this paper in your agent:
hf papers read 2610.09457 Don't have the latest CLI?
curl -LsSf https://hf.co/cli/install.sh | bash Models citing this paper 0
No model linking this paper
Datasets citing this paper 0
No dataset linking this paper
Spaces citing this paper 0
No Space linking this paper
Collections including this paper 0
No Collection including this paper