Publications
My advisees*. Equal contribution^.
2026
- TOSEM’26Guiding AI to Fix Its Own Flaws: An Empirical Study on LLM-Driven Secure Code GenerationACM Trans. Softw. Eng. Methodol., Sep 2026
- TMLR’26Lens: A Knowledge-Guided Foundation Model for Network TrafficTransactions on Machine Learning Research, Sep 2026
- CHI’26Designing AI Peers for Collaborative Mathematical Problem Solving with Middle School Students: A Participatory Design StudyarXiv preprint arXiv:2601.17962 (To appear at ACM CHI 2026), Sep 2026
- TOSEM’26Reassessing Code Authorship Attribution in the Era of Language ModelsAccepted to ACM TOSEM, Sep 2026
- ACL-BEA’26PeerMathDial: A Middle School Dialogue Dataset for Student Collaborative Math Problem SolvingAccepted to the 21st Workshop on Innovative Use of NLP for Building Educational Applications (BEA) at ACL, Sep 2026
2025
- PreprintRevisiting Prompt Optimization with Large Reasoning Models-A Case Study on Event ExtractionarXiv preprint arXiv:2504.07357, Sep 2025
- PreprintA Survey on Large Language Models for Automated PlanningarXiv preprint arXiv:2502.12435, Sep 2025
- ICDM’25WEvaluating the Effectiveness of Persona Simulation in Opinion Prediction with GPT-4.1In 2025 IEEE International Conference on Data Mining Workshops (ICDMW - Undergraduate and Graduate Honor Symposium), Sep 2025
- EMNLP’25Feature Extraction and Steering for Enhanced Chain-of-Thought Reasoning in Language ModelsarXiv preprint arXiv:2505.15634 (to appear at EMNLP 2025 Main), Sep 2025
- EMNLP’25All for One: LLMs Solve Mental Math at the Last Token With Information Transferred From Other Tokens(Covered by Rohan Paul (AI Influencer at X) and “AI: post transformers” Podcast)To appear at EMNLP 2025 Main, Sep 2025
- EMNLP’25 FindingsA survey on sparse autoencoders: Interpreting the internal mechanisms of large language modelsarXiv preprint arXiv:2503.05613 (to appear at EMNLP 2025 Findings), Sep 2025
- EMNLP’25 FindingsBeneath the Surface: How Large Language Models Reflect Hidden BiasarXiv preprint arXiv:2502.19749 (to appear at EMNLP 2025 Findings), Sep 2025
- COLM’25WCan LLMs Simulate Personas with Reversed Performance? A Benchmark for Counterfactual Instruction FollowingCOLM Workshop on Social Simulation with LLMs, Sep 2025
- IROS’25Autospatial: Visual-language reasoning for social robot navigation through efficient spatial reasoning learningarXiv preprint arXiv:2503.07557 (to appear at IROS 2025), Sep 2025
- AAAI’25WMechanistic Understanding of Language Models in Syntactic Code CompletionAAAI Workshop on Towards Knowledgeable Foundation Models, Sep 2025
- ICLR’25DOTS: Learning to Reason Dynamically in LLMs via Optimal Reasoning Trajectories Search(Covered by MIT Tech Review China)The Thirteenth International Conference on Learning Representations, Sep 2025
- AAAI’25WMathVC: An LLM-Simulated Multi-Character Virtual Classroom for Mathematics Education(Invited Presentation at Wolfram Research LLM Agent Colloquium; media coverage by GovTech and The George)AAAI AI4Edu Workshop, Sep 2025
2024
- PreprintUnderstanding the Effect of Algorithm Transparency of Model Explanations in Text-to-SQL Semantic ParsingarXiv preprint arXiv:2410.16283, Sep 2024
- EMNLP’24
- ACL’24An Investigation of Neuron Activation as a Unified Lens to Explain Chain-of-Thought Eliciting Arithmetic Reasoning of LLMs(Covered by MIT Technology Review China [English Translate])ACL, Sep 2024
- Preprint
- ICLR’24Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning(Featured in Hugging Face Daily Papers and AWS Builder Blog)The Twelfth International Conference on Learning Representations (also at ICLR Workshop on Reliable and Responsible Foundation Models), Sep 2024
2023
- ACL’23Improving Generalization in Language Model-based Text-to-SQL Semantic Parsing: Two Simple Semantic Boundary-based Techniques(Rank#8 on Spider leaderboard as of Aug 2023)In Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers), Sep 2023
- AAAI’23 SAExplaining Large Language Model-Based Neural Semantic Parsers (Student Abstract)AAAI Student Abstract, Sep 2023
- JBP’23A paradigm shift from “human writing” to “machine generation” in personality test development: An application of state-of-the-art natural language processing(Editor Commendation, one of 13 out of 1,000+ submissions in 2022)Journal of Business and Psychology, Sep 2023
2022
- ICLR’22 DL4CodeCode Editing from Few Exemplars by Adaptive Multi-Extent CompositionIn Deep Learning for Code Workshop at International Conference on Learning Representations, Sep 2022
2021
- DissertationOn Advancing Natural Language Interfaces: Data Collection, Model Development, and User InteractionSep 2021
2020
2019
- EMNLP-IJCNLP’19
2018
- KDD’18 DL DayA comprehensive study of staqc for deep code summarizationIn Deep Learning Day at KDD, Sep 2018
2016
- AAAI’16Semi-supervised multinomial naive bayes for text classification by leveraging word-level statistical constraintIn Proceedings of the AAAI Conference on Artificial Intelligence, Sep 2016