✳flâneur — a map of the web's best reading
Set Transformer: A Framework for Attention-basedPermutation-Invariant Neural Networks
arxiv.org · saved by 1 readers
N/A
Explore this link on the map →related reading
- A Mathematical Framework for Transformer Circuitstransformer-circuits.pub
- Transformer Explainer: LLM Transformer Model Visually Explainedpoloclub.github.io
- The Annotated Transformernlp.seas.harvard.edu
- Transformers are RNNs: Fast Autoregressive Transformers with Linear Attentionarxiv.org
- Sparser Block-Sparse Attention via Token Permutationarxiv.org
- Jacobian Lens – Qwen3.6-27B | Neuronpedianeuronpedia.org
- Efficiently Scaling Transformer Inferencearxiv.org
- Pretraining Data Mixtures Enable Narrow Model Selection Capabilities in Transformer Modelsarxiv.org
- Annotated Research Paper Implementations: Transformers, StyleGAN, Stable Diffusion, DDPM/DDIM, LayerNorm, Nucleus Sampling and morenn.labml.ai
- The Annotated Transformernlp.seas.harvard.edu
- Neuronpedianeuronpedia.org
- Full Stack Optimization of Transformer Inference: a Surveyarxiv.org