Transformer Visualization #11

helblazer811 · 2023-01-12T12:14:06Z

I want to make visualization systems for visualizing transformers, specifically self-attention. It would be nice if it worked for Vision Transformers as well as Language Models.

helblazer811 · 2023-01-12T12:36:20Z

Vision Transformer

I think it would be interesting to visualize how the vision transformer works by splitting an image into a bunch of patches.

helblazer811 · 2023-01-12T13:11:52Z

Self-Attention

This is the most important component (imo) to visualize property in the Transformer Architecture. I can think of two levels of visualization for this.

In-depth visualization

This visualization will show (1) the Key, Value, and Query feed-forward layers, (2) the matrices returned by these layers that are then multiplied, (3) the softmax operation combining the Key and Query into a score (4) the linear combination of the values into final values.

A high-level conceptual visualization

This is the layer that I think should be a NeuralNetworkLayer.

It should take in either text (broken down into tokens), an image (broken into patches), or vectors (output of a feed forward layer). These should then be passed into a self-attention layer. This layer should put the tokens (whatever type) onto the left and top side of a matrix visualization. The matrix visualization should be a 2D heatmap of the softmaxed (normalized) attention scores. Finally, the scores should be combined with the values to form the output of the self-attention module.

ImageToPatches

I will need to make a layer for splitting up an image into patches. The patches are necessary to represent the image as a sequence.

helblazer811 added the enhancement New feature or request label Jan 12, 2023

helblazer811 mentioned this issue Jan 15, 2023

Nested Neural Networks #20

Open

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Transformer Visualization #11

Transformer Visualization #11

helblazer811 commented Jan 12, 2023

helblazer811 commented Jan 12, 2023

helblazer811 commented Jan 12, 2023 •

edited

Loading

Transformer Visualization #11

Transformer Visualization #11

Comments

helblazer811 commented Jan 12, 2023

helblazer811 commented Jan 12, 2023

Vision Transformer

helblazer811 commented Jan 12, 2023 • edited Loading

Self-Attention

In-depth visualization

A high-level conceptual visualization

ImageToPatches

helblazer811 commented Jan 12, 2023 •

edited

Loading