Unveiling Transformer Perception by Exploring Input Manifolds (2410.06019v1)

Published 8 Oct 2024 in cs.LG, cs.AI, and cs.CL

Abstract: This paper introduces a general method for the exploration of equivalence classes in the input space of Transformer models. The proposed approach is based on sound mathematical theory which describes the internal layers of a Transformer architecture as sequential deformations of the input manifold. Using eigendecomposition of the pullback of the distance metric defined on the output space through the Jacobian of the model, we are able to reconstruct equivalence classes in the input space and navigate across them. We illustrate how this method can be used as a powerful tool for investigating how a Transformer sees the input space, facilitating local and task-agnostic explainability in Computer Vision and Natural Language Processing tasks.

Collections

Summary

We haven't generated a summary for this paper yet.

Summarize Now

Paper Prompts

Explore 10 Community Prompts

Follow-up Questions

We haven't generated follow-up questions for this paper yet.

Generate Now

Unveiling Transformer Perception by Exploring Input Manifolds (2410.06019v1)

Collections

Summary

Paper Prompts

Follow-up Questions

Related Papers

Authors (5)