![]() | HU, Q.* , WEN, J.* , Zhang, Y., CORDY, M., & Lyu, Y. (2026). On the Evaluation of Capability Estimation Methods for Large Language Models. In S. Koenig, C. Jenkins, ... M. E. Taylor (Eds.), Proceedings of the AAAI Conference on Artificial Intelligence. Association for the Advancement of Artificial Intelligence. doi:10.1609/aaai.v40i37.40368 Peer reviewed* These authors have contributed equally to this work. |
![]() | Quan, L., WEN, J., Hu, Q., CORDY, M., Huang, Y., Ma, L., & Li, X. (2025). Evaluation and Improvement of Test Selection for Large Language Models. Journal of Software: Evolution and Process. doi:10.22541/au.175033647.70907223/v1 Peer Reviewed verified by ORBi |
![]() | WEN, J., HU, Q., GUO, Y., CORDY, M., & Le Traon, Y. (2025). Variable Renaming-Based Adversarial Test Generation for Code Model: Benchmark and Enhancement. ACM Transactions on Software Engineering and Methodology. doi:10.1145/3723353 Peer Reviewed verified by ORBi |