Publications
2026
-
EMNLPHarnessing LLM Agents with Skill ProgramsIn Conference on Empirical Methods in Natural Language Processing (EMNLP) 2026
-
ACLHard2Verify: A Step-Level Verification Benchmark for Open-Ended Frontier MathIn Annual Conference of the Association for Computational Linguistics (ACL) 2026
-
ACLDiscovering the Gems in Early Layers: Accelerating Long-Context LLMs with 1000x Input Token ReductionIn Annual Conference of the Association for Computational Linguistics (ACL) 2026
2025
-
TMLRA Survey of Frontiers in LLM Reasoning: Inference Scaling, Learning to Reason, and Agentic SystemsIn Transactions on Machine Learning Research (TMLR) 2025
-
TMLRGeneralized Out-of-Distribution Detection and Beyond in Vision Language Model Era: A SurveyIn Transactions on Machine Learning Research (TMLR) 2025
2024
-
NeurIPSIs A Picture Worth A Thousand Words? Delving Into Spatial Reasoning for Vision Language ModelsIn Neural Information Processing Systems (NeurIPS) 2024
-
ICMLUnderstanding Retrieval-Augmented Task Adaptation for Vision-Language ModelsIn International Conference on Machine Learning (ICML) 2024
-
ICLRProvable Out-of-Distribution Generalization in HypersphereIn International Conference on Learning Representations (ICLR) 2024
2023
-
CPAL
Oral Domain Generalization via Nuclear Norm RegularizationIn Conference on Parsimony and Learning (CPAL) 2023 -
EMNLPA Critical Analysis of Document Out-of-Distribution DetectionIn Empirical Methods in Natural Language Processing (EMNLP Findings) 2023
-
IJCVHow Does Fine-Tuning Impact Out-of-Distribution Detection for Vision-Language Models?In International Journal of Computer Vision (IJCV) 2023
2022
-
NeurIPSDomain Generalization with Nuclear Norm RegularizationIn Neural Information Processing Systems (NeurIPS’W) DistShift Workshop 2022
-
ICMLAre Vision Transformers Robust to Spurious Correlations?In International Conference on Machine Learning (ICML’W), SCIS Workshop 2022