Filtered Corpus Training (FiCT) Shows that Language Models can Generalize from Indirect Evidence
Comments Forthcoming in Transactions of the Association for Computational Linguistics (TACL). This is a pre-MIT Press publication version. For code and trained models, see http://github.com/CLMBRs/corpus-filtering