Publications
-
1
Testing $k$-submodularity
-
2
Efficient Algorithms for Influence Maximization in General Models and Observed Cascades
-
3
Layerwise Dynamics for In-Context Classification in TransformersWe show how Transformers geometrically separate classes layer-by-layer during in-context learning.
-
4
Noise Stability of Transformer ModelsWe introduce noise stability to measure model simplicity and accelerate Transformer training and grokking.
-
5
Fast-MWEM: Private Data Release in Sublinear Time
-
6
Efficient Algorithms for Adversarially Robust Approximate Nearest Neighbor Search
- NeurIPS 2025 Workshop: Reliable ML from Unreliable Data
- WoLA 2026 Poster
-
7
Estimating Hitting Times Locally At Scale
-
8
Compression Barriers for Autoregressive TransformersWe prove fundamental limits on KV cache compression, showing when sublinear memory is impossible.
-
9
$k$NN Attention Demystified: A Theoretical Exploration for Scalable TransformersWe provide the first theoretical guarantees and fast sub-quadratic algorithms for $k$NN attention.
-
10
Counting Simplices in Hypergraph Streams
-
11
Teaching American Sign Language in Mixed Reality