File size: 194 Bytes
0ca31bd
 
 
 
 
 
1
2
3
4
5
6
---
license: mit
---
This is the fastText pretraining data filter targeting
the LAMBADA ES task, discussed in the main text of the Perplexity
Correlations paper: https://arxiv.org/abs/2409.05816