Two papers accepted by EMNLP 2026!
Two papers from our team were accepted by EMNLP 2026, including one main-conference paper and one Findings paper. Both works study computer-use agents in long-horizon real-world or clinical environments.
Reference
- WeaveBench: A Long-Horizon, Real-World Benchmark for Computer-Use Agents with Hybrid Interfaces
Wanli Li, Bowen Zhou, Yunyao Yu, Zhou Xu, Yifan Yang, Dongsheng Li, Caihua Shan.
The 2026 Conference on Empirical Methods in Natural Language Processing (EMNLP). 2026 - MedCUA-Bench: A Screenshot-Only Benchmark for Clinical Computer-Use Agents
Jia Yu, Zilong Wang, Xinyang Jiang, Dongsheng Li, Shuo Wang.
Findings of the Association for Computational Linguistics: EMNLP (EMNLP Findings). 2026