Papers
Communities
Events
Blog
Pricing
Search
Open menu
Home
Papers
2406.02080
Cited By
LongSSM: On the Length Extension of State-space Models in Language Modelling
4 June 2024
Shida Wang
Re-assign community
ArXiv
PDF
HTML
Papers citing
"LongSSM: On the Length Extension of State-space Models in Language Modelling"
3 / 3 papers shown
Title
Griffin: Mixing Gated Linear Recurrences with Local Attention for Efficient Language Models
Soham De
Samuel L. Smith
Anushan Fernando
Aleksandar Botev
George-Christian Muraru
...
David Budden
Yee Whye Teh
Razvan Pascanu
Nando de Freitas
Çağlar Gülçehre
Mamba
53
116
0
29 Feb 2024
Train Short, Test Long: Attention with Linear Biases Enables Input Length Extrapolation
Ofir Press
Noah A. Smith
M. Lewis
242
690
0
27 Aug 2021
The Pile: An 800GB Dataset of Diverse Text for Language Modeling
Leo Gao
Stella Biderman
Sid Black
Laurence Golding
Travis Hoppe
...
Horace He
Anish Thite
Noa Nabeshima
Shawn Presser
Connor Leahy
AIMat
245
1,977
0
31 Dec 2020
1