Evaluation and Generalization
We develop rigorous ways to evaluate language models and understand when, how, and why they generalize.
Related publications
2026
-
-
In Findings of the Association for Computational Linguistics: ACL 2026, Dec 2026ACL 2026 Findings -
2025
-
In Proceedings of the 2025 Conference of the Nations of the Americas Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 1: Long Papers), Dec 2025Best Paper Award -
In Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), Jul 2025
2024
-
In Advances in Neural Information Processing Systems, Jul 2024 -
In Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing, Nov 2024
2023
-
-
In Thirty-seventh Conference on Neural Information Processing Systems, Nov 2023NeurIPS 2023 Spotlight
2022
-
In AAAI Conference on Artificial Intelligence, Nov 2022
2021
-
In Advances in Neural Information Processing Systems, Nov 2021 -
In Advances in Neural Information Processing Systems, Nov 2021NeurIPS 2021 Outstanding Paper Award