Picture for Chris Maddison

Chris Maddison

The Shaped Transformer: Attention Models in the Infinite Depth-and-Width Limit

Add code
Jun 30, 2023
Viaarxiv icon