LongLeader: A Comprehensive Leaderboard for Large Language Models in Long-context Scenarios

Publication
Proceedings of the 2025 Conference of the North American Chapter of the Association for Computational Linguistics
Pei (Patrick) Chen
Pei (Patrick) Chen
Applied Scientist at Amazon

Applied Scientist at Amazon working across the LLM training and evaluation stack — mid-training, post-training, RL, and rubric-based evaluation — on 400B+ MoE backbones for customer-facing agentic systems.