Direct Multi-Token DecodingXuan Luo, Weizhi Wang, Xifeng Yan
In Proceedings of the Third Conference on Language Modeling (COLM 2026).
[Paper]
PhyWorldBench: A Comprehensive Evaluation of Physical Realism in Text-to-Video ModelsJing Gu, Xian Liu, Yu Zeng, Ashwin Nagarajan, Fangrui Zhu, Daniel Hong, Yue Fan, Qianqi Yan, Kaiwen Zhou, Ming-Yu Liu, Xin Eric WangICLR 2026 (Oral).
Presenting a Paper is an Art: Self-Improvement Aesthetic Agents for Academic PresentationsChengzhi Liu, Yuzhe Yang, Kaiwen Zhou, Zhen Zhang, Yue Fan, Yanan Xie, Peng Qi, Xin Eric WangICLR 2026.
SAFER: Risk-Constrained Sample-then-Filter in Large Language ModelsQingni Wang, Yue Fan, Xin Eric WangICLR 2026.
Interleaved Vision-and-Language Generation via Generative VokensKaizhi Zheng*, Xuehai He*, Xin Eric WangWACV 2026.
OpenSage: Self-programming Agent Generation EngineHongwei Li, Zhun Wang, Qinrun Dai, Yuzhou Nie, Jinjun Peng, Ruitong Liu, Jingyang Zhang, Kaijie Zhu, Jingxuan He, Lun Wang, Yangruibo Ding, Yueqi Chen, Wenbo Guo, Dawn Song
In International Conference on Machine Learning (ICML 2026).
[Paper]
BlueCodeAgent: A Blue Teaming Agent Powered by Automated Red Teaming for CodeGen AIChengquan Guo, Yuzhou Nie, Chulin Xie, Zinan Lin, Wenbo Guo, Bo Li
In International Conference on Machine Learning (ICML 2026).
[Paper]
rePIRL: Learn PRM with Inverse RL for LLM ReasoningXian Wu, Kaijie Zhu, Ying Zhang, Lun Wang, Wenbo Guo
In International Conference on Machine Learning (ICML 2026).
[Paper]
CyberCycle: Scalable Real-World Benchmark for AI Agents' End-to-End Cybersecurity CapabilitiesTianneng Shi, Robin Rheem, Dongwei Jiang, Francisco De La Riega, Mona Wang, Zhun Wang, Jingzhi Jiang, Alexander Cheung, Sean Tai, Jonah Cha, Jianhong Tu, Gabriel Han, Chenguang Wang, Wenbo Guo, Jingxuan He, Dawn Song
In International Conference on Machine Learning (ICML 2026).
[Paper]
Position: To Defend Against Cyber Attacks, We Must Teach AI Agents to HackTerry Yue Zhuo, Yangruibo Ding, Wenbo Guo, Ruijie Meng
In International Conference on Machine Learning (ICML 2026).
[Paper]
AgentBeats: Agentifying Agent Assessment for Openness, Standardization, and ReproducibilityXiaoyuan Liu, Jianhong Tu, Yuqi Chen, Siyuan Xie, Sihan Ren, Tianneng Shi, Gal Gantar, Evan Sandoval, Donghyun Lee, Daniel Miao, Peter J. Gilbert, Nick Hynes, Mauro Staver, Warren He, David Marn, Andrew Low, Xi Zhang, Elron Bandel, Michal Shmueli-Scheuer, Siva Reddy, Alexandre Drouin, Alexandre Lacoste, Ramayya Krishnan, Elham Tabassi, Yu Su, Victor Barres, Chenguang Wang, Wenbo Guo, Dawn Song
In International Conference on Machine Learning (ICML 2026).
[Paper]
PAGENT: Program Analysis Guided LLM Agent for Proof-of-Concept GenerationAchintya Desai, Md Shafiuzzaman, Wenbo Guo, Tevfik Bultan
In ACM SIGSOFT International Symposium on Software Testing and Analysis (ISSTA 2026).
[Paper]
PDFuzzer: Fuzzing PDF Readers with LLM-Generated API Call SequencesSuyue Guo, Stijn Pletinckx, Tianle Yu, Yigitcan Kaya, Saad Ullah, Wenbo Guo, Christopher Kruegel, Giovanni Vigna
In ACM Conference on Computer and Communications Security (CCS 2026).
CVE-Genie: Automated Reproduction of Real-World Vulnerabilities with LLM-Based Multi-Agent FrameworkSaad Ullah, Praneeth Balasubramanian, Wenbo Guo, Amanda Burnett, Hammond Pearce, Christopher Kruegel, Giovanni Vigna, Gianluca Stringhini
In ACM Conference on Computer and Communications Security (CCS 2026).
[Paper]
DevOps-Gym: Benchmarking AI Agents in Software DevOps CycleYuheng Tang**, Kaijie Zhu**, Bonan Ruan, Chuqi Zhang, Michael Yang, Hongwei Li, Suyue Guo, Tianneng Shi, Zekun Li, Christopher Kruegel, Giovanni Vigna, Dawn Song, William Yang Wang, Lun Wang, Yangruibo Ding, Zhenkai Liang, Wenbo Guo
In International Conference on Learning Representations (ICLR 2026).
[Paper]
SoK: Attack and Defense Landscape of Agentic AI SystemsJuhee Kim, Xiaoyuan Liu, Wenbo Guo, Dawn Song
In USENIX Security Symposium (USENIX Security 2026).
[Paper]
Graph Neural Networks for Residential Location Choice: Connection to Classical Logit ModelsZ. Cheng, L. Hu, Y. Bu, Y. Zhou, S. Wang
Transportation Research Part B: Methodological, vol. 209, Art. 103464, Jul. 2026.
Position: AI-Agent Pricing Should Become More Outcome-Dependent: An Economic PerspectiveY. Bu, Y. Ma
Preprint, 2026.
Information-Theoretic Classifier-Free Guidance with Adaptive Schedule OptimizationH. Chen, X. Xu, Y. Bu
arXiv preprint, 2026.
[Paper]
Detector-Evasive LLM Paraphrasing via Constrained Policy OptimizationM. Wang, Z. Shen, Y. Bu, S. Zou
arXiv preprint, 2026.
[Paper]
WorldMemArena: Evaluating Multimodal Agent Memory Through Action-World InteractionC. Liu, Y. Yang, S. X. Pu, Y. Liu, L. Long, Y. Guo, N. Chen, Z. Weng, E. Kochkina, S. Kaur, C. Smiley, X. Liu, J. Zou, S. Liu, Y. Bu, S. Peng, X. E. Wang
arXiv preprint, 2026.
[Paper]
Auditing Agent Harness SafetyC. Liu, Y. Guo, Y. Liu, Y. Yang, Q. Yan, X. Zhao, W. Hua, S. Liu, S. Li, Y. Bu, X. E. Wang
arXiv preprint, 2026.
[Paper]
Fundamental Trade-Offs in Multi-Bit Watermarking of Stochastic ProcessesH. He, Y. Liu, Z. Shen, Z. Wang, Y. Mao, Y. Bu
arXiv preprint, 2026.
[Paper]
Length Value Model: Scalable Value Pretraining for Token-Level Length ModelingZ. Zhang, C. Yang, Z. Xia, Z. Yang, C. Liu, Z. Weng, Y. Liu, H. Chen, J. Pan, C. Zhao, Y. Bu, A. Patel, Z. Gan, X. E. Wang
arXiv preprint, 2026.
[Paper]
ConvexBench: Can LLMs Recognize Convex Functions?Y. Liu, Y. Huang, Y.-X. Wang, Y. Liang, Y. Bu
International Conference on Machine Learning (ICML 2026).
[Paper]
Position: LLM Watermarking Should Align Stakeholders’ Incentives for Practical AdoptionY. Liu, X. Zhao, D. Song, G. W. Wornell, Y. Bu
Findings of the Association for Computational Linguistics (ACL 2026).
[Paper]
In-Context Watermarks for Large Language ModelsY. Liu, X. Zhao, C. Kruegel, D. Song, Y. Bu
International Conference on Learning Representations (ICLR 2026).
[Paper]
A Reinforcement Learning Framework for Robust and Secure LLM WatermarkingL. An, Y. Liu, Y. Liu, Y. Bu, Y. Zhang, S. Chang
European Chapter of the Association for Computational Linguistics (EACL 2026).
[Paper]
TrustEnergy: A Unified Framework for Accurate and Reliable User-level Energy Usage PredictionD. Yu, R. Xu, D. Zhuang, Y. Bu, S. Wang, G. Wang
AAAI Conference on Artificial Intelligence (AAAI 2026).
2025
ThoughtTerminator: Benchmarking, Calibrating, and Mitigating Overthinking in Reasoning ModelsXiao Pu, Michael Saxon, Wenyue Hua, William Yang Wang
In Proceedings of the Second Conference on Language Modeling (COLM 2025).
[Paper]
T2V-Turbo-v2: Enhancing Video Model Post-Training through Data, Reward, and Conditional Guidance DesignJiachen Li, Qian Long, Jian Zheng, Xiaofeng Gao, Robinson Piramuthu, Wenhu Chen, William Yang Wang
In Proceedings of The Thirteenth International Conference on Learning Representations (ICLR 2025).
Generalization v.s. Memorization: Tracing Language Models' Capabilities Back to Pretraining DataXinyi Wang, Antonis Antoniades, Yanai Elazar, Alfonso Amayuelas, Alon Albalak, Kexun Zhang, William Yang Wang
In Proceedings of The Thirteenth International Conference on Learning Representations (ICLR 2025).
Enhancing Software Agents with Monte Carlo Tree Search and Hindsight FeedbackAntonis Antoniades, Albert Örwall, Kexun Zhang, Yuxi Xie, Anirudh Goyal, William Yang Wang
In Proceedings of The Thirteenth International Conference on Learning Representations (ICLR 2025).
MMWorld: Towards Multi-discipline Multi-faceted World Model Evaluation in VideosXuehai He, Weixi Feng, Kaizhi Zheng, Yujie Lu, Wanrong Zhu, Jiachen Li, Yue Fan, Jianfeng Wang, Linjie Li, Zhengyuan Yang, Kevin Lin, William Yang Wang, Lijuan Wang, Xin Eric Wang
In Proceedings of The Thirteenth International Conference on Learning Representations (ICLR 2025).
Speculative Knowledge Distillation: Bridging the Teacher-Student Gap Through Interleaved SamplingWenda Xu, Rujun Han, Zifeng Wang, Long Le, Dhruv Madeka, Lei Li, William Yang Wang, Rishabh Agarwal, Chen-Yu Lee, Tomas Pfister
In Proceedings of The Thirteenth International Conference on Learning Representations (ICLR 2025).
CBT-Bench: Evaluating Large Language Models on Assisting Cognitive Behavior TherapyMian Zhang, Xianjun Yang, Xinlu Zhang, Travis Labrum, Jamie C. Chiu, Shaun M. Eack, Fei Fang, William Yang Wang, Zhiyu Chen
In Proceedings of the 2025 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL 2025).
Investigating the Transferability of Code Repair for Low-Resource Programming LanguagesKyle Wong, Alfonso Amayuelas, Liangming Pan, William Yang Wang
In Findings of the 2025 Annual Conference of the Nations of the Americas Chapter of the ACL (Findings of NAACL 2025).
Scaling LLM Inference Efficiently with Optimized Sample Compute AllocationKexun Zhang, Shang Zhou, Danqing Wang, William Yang Wang, Lei Li
In Findings of the 2025 Annual Conference of the Nations of the Americas Chapter of the ACL (Findings of NAACL 2025).
Unveiling the Impact of Coding Data Instruction Fine-Tuning on Large Language Models ReasoningXinlu Zhang, Zhiyu Chen, Xi Ye, Xianjun Yang, Lichang Chen, William Yang Wang, Linda Ruth Petzold
In Proceedings of The 39th Annual AAAI Conference on Artificial Intelligence (AAAI 2025).
Combating Multimodal LLM Hallucination via Bottom-up Holistic ReasoningShengqiong Wu, Hao Fei, Liangming Pan, William Yang Wang, Shuicheng YAN, Tat-Seng Chua
In Proceedings of The 39th Annual AAAI Conference on Artificial Intelligence (AAAI 2025).
Soft Thinking: Unlocking the Reasoning Potential of LLMs in Continuous Concept SpaceZhen Zhang, Xuehai He, Weixiang Yan, Ao Shen, Chenyang Zhao, Shuohang Wang, Yelong Shen, Xin Eric WangNeurIPS 2025.
GRIT: Teaching MLLMs to Think with ImagesYue Fan, Xuehai He, Diji Yang, Kaizhi Zheng, Ching-Chen Kuo, Yuting Zheng, Sravana Jyothi Narayanaraju, Xinze Guan, Xin Eric WangNeurIPS 2025.
More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning ModelsChengzhi Liu, Zhongxing Xu, Qingyue Wei, Juncheng Wu, James Zou, Xin Eric Wang, Yuyin Zhou, Sheng LiuNeurIPS 2025.
SafeKey: Amplifying Aha-Moment Insights for Safety ReasoningKaiwen Zhou, Xuandong Zhao, Gaowen Liu, Jayanth Srinivasa, Aosong Feng, Dawn Song, Xin Eric WangEMNLP 2025.
Hidden in Plain Sight: Probing Implicit Reasoning in Multimodal Language ModelsQianqi Yan, Hongquan Li, Shan Jiang, Yang Zhao, Xinze Guan, Ching-Chen Kuo, Xin Eric WangEMNLP 2025.
GUI-Bee: Align GUI Action Grounding to Novel Environments via Autonomous ExplorationYue Fan, Handong Zhao, Ruiyi Zhang, Yu Shen, Xin Eric Wang, Gang WuEMNLP 2025.
Dynamic Evaluation for Oversensitivity in LLMsSophia Xiao Pu, Sitao Cheng, Xin Eric Wang, William Yang WangFindings of EMNLP 2025.
Agent S2: A Compositional Generalist-Specialist Framework for Computer Use AgentsSaaket Agashe*, Kyle Wong*, Vincent Tu*, Jiachen Yang, Ang Li, Xin Eric WangCOLM 2025.
VLM4D: Towards Spatiotemporal Awareness in Vision Language ModelsShijie Zhou*, Alexander Vilesov*, Xuehai He*, Ziyu Wan, Shuwang Zhang, Aditya Nagachandra, Di Chang, Dongdong Chen, Xin Eric Wang, Achuta KadambiICCV 2025.
Multimodal Inconsistency Reasoning (MMIR): A New Benchmark for Multimodal Reasoning ModelsQianqi Yan, Yue Fan, Hongquan Li, Shan Jiang, Yang Zhao, Xinze Guan, Ching-Chen Kuo, Xin Eric WangFindings of ACL 2025.
Worse than Random? An Embarrassingly Simple Probing Evaluation of Large Multimodal Models in Medical VQAQianqi Yan, Xuehai He, Xiang Yue, Xin Eric WangFindings of ACL 2025 NeurIPS 2024 Workshop on GenAI for Health.
Agent S: An Open Agentic Framework that Uses Computers Like a HumanSaaket Agashe*, Jiuzhou Han*, Shuyu Gan, Jiachen Yang, Ang Li, Xin Eric WangICLR 2025 Best Paper Award (ICLR 2025 Agentic AI for Science Workshop).
Multimodal Situational SafetyKaiwen Zhou*, Chengzhi Liu*, Xuandong Zhao, Anderson Compalas, Dawn Song, Xin Eric WangICLR 2025.
EditRoom: LLM-parameterized Graph Diffusion for Composable 3D Room Layout EditingKaizhi Zheng, Xiaotong Chen, Xuehai He, Jing Gu, Linjie Li, Zhengyuan Yang, Kevin Lin, Jianfeng Wang, Lijuan Wang, Xin Eric WangICLR 2025.
LLM-Coordination: Evaluating and Analyzing Multi-Agent Coordination Abilities in Large Language ModelsSaaket Agashe, Yue Fan, Anthony Reyna, Xin Eric WangFindings of NAACL 2025.
KVLink: Accelerating Large Language Models via Efficient KV Cache ReuseJingbo Yang, Bairu Hou, Wei Wei, Yujia Bao, Shiyu Chang
Advances in Neural Information Processing Systems (NeurIPS 2025).
[Paper][Code]
Defending LLM Watermarking Against Spoofing Attacks with Contrastive Representation LearningLi An, Yujian Liu, Yepeng Liu, Yang Zhang, Yuheng Bu, Shiyu Chang
Conference on Language Modelling (COLM 2025).
[Paper][Code]
VSP: Assessing the Dual Challenges of Perception and Reasoning in Spatial Planning Tasks for VLMsQiucheng Wu, Handong Zhao, Michael Saxon, Trung Bui, William Yang Wang, Yang Zhang, Shiyu Chang
International Conference on Computer Vision (ICCV 2025).
[Paper][Code]
Instruction-Following Pruning for Large Language ModelsBairu Hou, Qibin Chen, Jianyu Wang, Guoli Yin, Chong Wang, Nan Du, Ruoming Pang, Shiyu Chang, Tao Lei
International Conference on Machine Learning (ICML 2025).
[Paper]
Fictitious Synthetic Data Can Improve LLM Factuality via Prerequisite LearningYujian Liu, Shiyu Chang, Tommi S. Jaakkola, Yang Zhang
International Conference on Learning Representations (ICLR 2025).
[Paper][Code]
A Probabilistic Framework for LLM Hallucination Detection via Belief Tree PropagationBairu Hou, Yang Zhang, Jacob Andreas, Shiyu Chang
Annual Conference of the North American Chapter of the Association for Computational Linguistics (NAACL 2025).
[Code]
Defending Large Language Models against Jailbreak Attacks via Semantic SmoothingJiabao Ji, Bairu Hou, Alexander Robey, George J. Pappas, Hamed Hassani, Yang Zhang, Eric Wong, Shiyu Chang
International Joint Conference on Natural Language Processing & Asia-Pacific Chapter of the Association for Computational Linguistics (IJCNLP-AACL 2025).
[Paper][Code]
Augment before You Try: Knowledge-Enhanced Table Question Answering via Table ExpansionYujian Liu, Jiabao Ji, Tong Yu, Ryan A. Rossi, Sungchul Kim, Handong Zhao, Ritwik Sinha, Yang Zhang, Shiyu Chang
Conference on Empirical Methods in Natural Language Processing (EMNLP-Findings 2025).
[Paper][Code]
Train a Unified Multimodal Data Quality Classifier with Synthetic DataWeizhi Wang, Rongmei Lin, Shiyang Li, Colin Lockard, Ritesh Sarkhel, Sanket Lokegaonkar, Jingbo Shang, Xifeng Yan, Nasser Zalmout, Xian Li
Findings of the Association for Computational Linguistics (Findings of EMNLP 2025).
[Paper]
Context-Aware Language Models for Forecasting Market Impact from Sequences of Financial NewsRoss Koval, Nicholas Andrews, Xifeng Yan
arXiv:2504.14787 preprint, 2025.
[Paper]
Multimodal Language Models with Modality-Specific Experts for Financial Forecasting from Interleaved Sequences of Text and Time SeriesRoss Koval, Nicholas Andrews, Xifeng Yan
Proceedings of International Joint Conference on Natural Language Processing (IJCNLP-AACL 2025).
[Paper]
Demystifying Network Foundation ModelsRoman Beltiukov, Satyandra Guthula, Wenbo Guo, Walter Willinger, Arpit Gupta
In Annual Conference on Neural Information Processing Systems (NeurIPS 2025).
SECODEPLT: A Unified Platform for Evaluating the Security of Code GenAIYuzhou Nie, Zhun Wang, Yu Yang, Ruizhe Jiang, Yuheng Tang, Xander Davies, Yarin Gal, Bo Li, Wenbo Guo, Dawn Song
In Annual Conference on Neural Information Processing Systems (NeurIPS 2025).
[Paper][Code]
Co-PatcheR: Collaborative Software Patching with Component-specific Small Reasoning ModelsYuheng Tang, Hongwei Li, Kaijie Zhu, Michael Yang, Yangruibo Ding, Wenbo Guo
In Annual Conference on Neural Information Processing Systems (NeurIPS 2025).
[Paper]
BlockFound: Customized blockchain foundation model for anomaly detectionJiahao Yu, Xian Wu, Hao Liu, Wenbo Guo, Xinyu Xing
In Annual Conference on Neural Information Processing Systems (NeurIPS 2025).
[Paper]
Temporal Logic-Based Multi-Vehicle Backdoor Attacks against Offline RL Agents in End-to-end Autonomous DrivingXuan Chen, Shiwei Feng, Zikang Xiong, Shengwei An, Yunshu Mao, Lu Yan, Guanhong Tao, Wenbo Guo, Xiangyu Zhang
In Annual Conference on Neural Information Processing Systems (NeurIPS 2025).
AGENTVIGIL: Generic Black-Box Red-teaming for Indirect Prompt Injection against LLM AgentsZhun Wang, Vincent Siu, Zhe Ye, Tianneng Shi, Yuzhou Nie, Xuandong Zhao, Chenguang Wang, Wenbo Guo, Dawn Song
In Empirical Methods in Natural Language Processing (EMNLP 2025).
[Paper]
LeakAgent: RL-based Red-teaming Agent for LLM Privacy LeakageYuzhou Nie, Zhun Wang, Ye Yu, Xian Wu, Xuandong Zhao, Nathaniel D. Bastian, Wenbo Guo, Dawn Song
In Conference on Language Modeling (COLM 2025).
[Paper]
PatchPilot: A Cost-Efficient Software Engineering Agent with Early Attempts on Formal VerificationHongwei Li, Yuheng Tang, Shiqi Wang, Wenbo Guo
In International Conference on Machine Learning (ICML 2025).
[Paper]
MELON: Provable Defense Against Indirect Prompt Injection Attacks in AI AgentsKaijie Zhu, Xianjun Yang, Jindong Wang, Wenbo Guo, William Yang Wang
In International Conference on Machine Learning (ICML 2025).
[Paper]
Mind the Inconspicuous: Revealing the Hidden Weakness in Aligned LLMs’ Ethical BoundariesJiahao Yu, Haozheng Luo, Yao-Chieh Hu, Yan Chen, Wenbo Guo, Han Liu, Xinyu Xing
In USENIX Security Symposium (USENIX Security 2025).
[Paper]
F-Fidelity: A Robust Framework for Faithfulness Evaluation of Explainable AIXu Zheng, Farhad Shirani, Zhuomin Chen, Chaohao Lin, Wei Cheng, Wenbo Guo, Dongsheng Luo
In International Conference on Learning Representations (ICLR 2025).
[Paper]
Class-wise Generalization Error: An Information-Theoretic AnalysisF. Laakom, M. Gabbouj, J. Schmidhuber, Y. Bu
Transactions on Machine Learning Research (TMLR), 2025.
[Paper]
NestGNN: A Graph Neural Network Framework Generalizing the Nested Logit Model for Travel Mode ChoiceY. Zhou, Z. Cheng, L. Hu, Y. Bu, S. Wang
arXiv preprint, 2025.
[Paper]
SAM2-SGP: Enhancing SAM2 for Medical Image Segmentation via Support-Set Guided PromptingY. Xing, J. Wu, Y. Bu, K. Gong
arXiv preprint, 2025.
[Paper]
Dataset Protection via Watermarked Canaries in Retrieval-Augmented LLMsY. Liu, X. Zhao, D. Song, Y. Bu
arXiv preprint, 2025.
[Paper]
Theoretically Grounded Framework for LLM Watermarking: A Distribution-Adaptive ApproachH. He, Y. Liu, Z. Wang, Y. Mao, Y. Bu
Conference on Neural Information Processing Systems (NeurIPS 2025).
[Paper]
UQGNN: Uncertainty Quantification of Graph Neural Networks for Multivariate Spatiotemporal PredictionD. Yu, D. Zhuang, L. Jiang, R. Xu, X. Ye, Y. Bu, S. Wang, G. Wang
International Conference on Advances in Geographic Information Systems (ACM SIGSPATIAL 2025).
Distributional Information Embedding: A Framework for Multi-bit WatermarkingH. He, Y. Liu, Z. Wang, Y. Mao, Y. Bu
Asia Pacific Workshop on Data Science and Information Theory (APWDSIT 2025).
Fairness Overfitting in Machine Learning: An Information-Theoretic PerspectiveF. Laakom, H. Chen, J. Schmidhuber, Y. Bu
International Conference on Machine Learning (ICML 2025).
Image Watermarks are Removable using Controllable Regeneration from Clean NoiseY. Liu, Y. Song, H. Ci, Y. Zhang, H. Wang, Z. Shou, Y. Bu
International Conference on Learning Representations (ICLR 2025).
[Paper][Code]
2024
Who Evaluates the Evaluations? Objectively Scoring Text-to-Image Prompt Coherence Metrics with T2IScoreScore (TS2)Michael Saxon, Fatima Jahara, Mahsa Khoshnoodi, Yujie Lu, Aditya Sharma, William Yang Wang
In Proceedings of The Thirty-eighth Annual Conference on Neural Information Processing Systems (NeurIPS 2024). Spotlight.
T2V-Turbo: Breaking the Quality Bottleneck of Video Consistency Model with Mixed Reward FeedbackJiachen Li, Weixi Feng, Tsu-Jui Fu, Xinyi Wang, S Basu, Wenhu Chen, William Yang Wang
In Proceedings of The Thirty-eighth Annual Conference on Neural Information Processing Systems (NeurIPS 2024).
FASTopic: Pretrained Transformer is a Fast, Adaptive, Stable, and Transferable Topic ModelXiaobao Wu, Thong Thanh Nguyen, Delvin Ce Zhang, William Yang Wang, Anh Tuan Luu
In Proceedings of The Thirty-eighth Annual Conference on Neural Information Processing Systems (NeurIPS 2024).
WildVision: Evaluating Vision-Language Models in the Wild with Human PreferencesYujie Lu, Dongfu Jiang, Wenhu Chen, William Yang Wang, Yejin Choi, Bill Yuchen Lin
In Proceedings of The Thirty-eight Conference on Neural Information Processing Systems Datasets and Benchmarks Track (NeurIPS 2024 Datasets and Benchmarks Track).
Multimodal Procedural Planning via Dual Text-Image PromptingYujie Lu, Pan Lu, Zhiyu Chen, Wanrong Zhu, Xin Eric Wang, William Yang Wang
In Findings of The 2024 Conference on Empirical Methods in Natural Language Processing (Findings of EMNLP 2024).
A Survey on Detection of LLMs-Generated ContentXianjun Yang, Liangming Pan, Xuandong Zhao, Haifeng Chen, Linda Ruth Petzold, William Yang Wang, Wei Cheng
In Findings of The 2024 Conference on Empirical Methods in Natural Language Processing (Findings of EMNLP 2024).
AKEW: Assessing Knowledge Editing in the WildXiaobao Wu, Liangming Pan, William Yang Wang, Anh Tuan Luu
In Proceedings of The 2024 Conference on Empirical Methods in Natural Language Processing (EMNLP 2024).
MultiAgent Collaboration Attack: Investigating Adversarial Attacks in Large Language Model Collaborations via DebateAlfonso Amayuelas, Xianjun Yang, Antonis Antoniades, Wenyue Hua, Liangming Pan, William Yang Wang
In Findings of The 2024 Conference on Empirical Methods in Natural Language Processing (Findings of EMNLP 2024).
BPO: Staying Close to the Behavior LLM Creates Better Online LLM AlignmentWenda Xu, Jiachen Li, William Yang Wang, Lei Li
In Proceedings of The 2024 Conference on Empirical Methods in Natural Language Processing (EMNLP 2024).
Losing Visual Needles in Image Haystacks: Vision Language Models are Easily Distracted in Short and Long Contexts", to appear in Findings of The 2024 Conference on Empirical Methods in Natural Language Processing (Findings of EMNLP 2024), Miami, Florida, ACL. 228. Rujun Han, Yuhao Zhang, Peng Qi, Yumo Xu, Jenyuan Wang, Lan Liu, William Yang Wang, Bonan Min, Vittorio Castelli, "RAG-QA Arena: Evaluating Domain Robustness for Long-form Retrieval Augmented Question AnsweringAditya Sharma, Michael Saxon, William Yang Wang
In Proceedings of The 2024 Conference on Empirical Methods in Natural Language Processing (EMNLP 2024).
Reward Guided Latent Consistency DistillationJiachen Li, Weixi Feng, Wenhu Chen, William Yang Wang
In Transactions on Machine Learning Research (TMLR), 2024.
A Survey on Large Language Models for Critical Societal Domains: Finance, Healthcare, and LawZhiyu Chen, Jing Ma, Xinlu Zhang, Nan Hao, An Yan, Armineh Nourbakhsh, Xianjun Yang, Julian McAuley, Linda Ruth Petzold, William Yang Wang
In Transactions on Machine Learning Research (TMLR), 2024. Survey Certification.
A Survey on Data Selection for Language ModelsAlon Albalak, Yanai Elazar, Sang Michael Xie, Shayne Longpre, Nathan Lambert, Xinyi Wang, Niklas Muennighoff, Bairu Hou, Liangming Pan, Haewon Jeong, Colin Raffel, Shiyu Chang, Tatsunori Hashimoto, William Yang Wang
In Transactions on Machine Learning Research (TMLR), 2024.
Discffusion: Discriminative Diffusion Models as Few-shot Vision and Language LearnersXuehai He, Weixi Feng, Tsu-Jui Fu, Varun Jampani, Arjun Reddy Akula, Pradyumna Narayana, S Basu, William Yang Wang, Xin Eric Wang
In Transactions on Machine Learning Research (TMLR), 2024.
Benchmarks as Microscopes: A Call for Model MetrologyMichael Saxon, Ari Holtzman, Peter West, William Yang Wang, Naomi Saphra
In Proceedings of the First Conference on Language Modeling (COLM 2024).
Guiding Language Model Reasoning with Planning TokensXinyi Wang, Lucas Caccia, Oleksiy Ostapenko, Xingdi Yuan, William Yang Wang, Alessandro Sordoni
In Proceedings of the First Conference on Language Modeling (COLM 2024).
Commonsense-T2I Challenge: Can Text-to-Image Generation Models Understand Commonsense?Xingyu Fu, Muyu He, Yujie Lu, William Yang Wang, Dan Roth
In Proceedings of the First Conference on Language Modeling (COLM 2024).
Ferret-v2: An Improved Baseline for Referring and Grounding with Large Language ModelsHaotian Zhang, Haoxuan You, Philipp Dufter, Bowen Zhang, Chen Chen, Hong-You Chen, Tsu-Jui Fu, William Yang Wang, Shih-Fu Chang, Zhe Gan, Yinfei Yang
In Proceedings of the First Conference on Language Modeling (COLM 2024).
Perils of Self-Feedback: Self-Bias Amplifies in Large Language ModelsWenda Xu, Guanglei Zhu, Xuandong Zhao, Liangming Pan, Lei Li, William Yang Wang
In Proceedings of The 62nd Annual Meeting of the Association for Computational Linguistics (ACL 2024).
The Knowledge Alignment Problem: Bridging Human and External Knowledge for Large Language ModelsShuo Zhang, Liangming Pan, Junzhou Zhao, William Yang Wang
In Findings of The 62nd Annual Meeting of the Association for Computational Linguistics (Findings of ACL 2024).
Hire a Linguist!: Learning Endangered Languages with In-Context Linguistic DescriptionsKexun Zhang, Yee Man Choi, Zhenqiao Song, Taiqi He, William Yang Wang, Lei Li
In Findings of The 62nd Annual Meeting of the Association for Computational Linguistics (Findings of ACL 2024).
Knowledge of Knowledge: Exploring Known-Unknowns Uncertainty with Large Language ModelsAlfonso Amayuelas, Kyle Wong, Liangming Pan, Wenhu Chen, William Yang Wang
In Findings of The 62nd Annual Meeting of the Association for Computational Linguistics (Findings of ACL 2024).
Global Human-guided Counterfactual Explanations for Molecular Properties via Reinforcement LearningDanqing Wang, Antonis Antoniades, Kha-Dinh Luong, Edwin Zhang, Mert Kosan, Jiachen Li, Ambuj Singh, William Yang Wang, Lei Li
In Proceedings of 30th SIGKDD Conference on Knowledge Discovery and Data Mining (KDD 2024).
Understanding Reasoning Ability of Language Models From the Perspective of Reasoning Paths AggregationXinyi Wang, Alfonso Amayuelas, Kexun Zhang, Liangming Pan, Wenhu Chen, William Yang Wang
In Proceedings of 41st International Conference on Machine Learning (ICML 2024).
Mastering Robot Manipulation with Multimodal Prompts through Pretraining and Multi-task Fine-tuningJiachen Li, Qiaozi Gao, Michael Johnston, Xiaofeng Gao, Xuehai He, Suhaila Shakiah, Hangjie Shi, Reza Ghanadan, William Yang Wang
In Proceedings of 41st International Conference on Machine Learning (ICML 2024).
Position Paper: Understanding the Role of Social Media Influencers in AI Research VisibilityIain Xie Weissburg, Mehir Arora, Xinyi Wang, Liangming Pan, William Yang Wang
In Proceedings of 41st International Conference on Machine Learning (ICML 2024).
Position Paper: TrustLLM: Trustworthiness in Large Language ModelsHuang et al
In Proceedings of 41st International Conference on Machine Learning (ICML 2024).
Lost in Translation? Translation Errors and Challenges for Fair Assessment of Text-to-Image Models on Multilingual ConceptsMichael Saxon, Yiran Lawrence Luo, Sharon Levy, Chitta Baral, Yezhou Yang, William Yang Wang
In Proceedings of 2024 Annual Conference of the North American Chapter of the Association for Computational Linguistics (NAACL 2024).
LLMMend: Pinpointing and Refining Large Language Models via Fine-Grained Actionable FeedbackWenda Xu, Daniel Deutsch, Mara Finkelstein, Juraj Juraska, Biao Zhang, Zhongtao Liu, William Yang Wang, Lei Li, Markus Freitag
In Findings of 2024 Annual Conference of the North American Chapter of the Association for Computational Linguistics (Findings of NAACL 2024).
Automatically Correcting Large Language Models: Surveying the landscape of diverse self-correction strategiesLiangming Pan, Michael Saxon, Wenda Xu, Deepak Nathani, Xinyi Wang, William Yang Wang
In Transactions of the Association for Computational Linguistics (TACL 2024).
Neuroformer: Multimodal and Multitask Generative Pretraining for Brain DataAntonis Antoniades, Yiyi Yu, Joe S Canzano, William Yang Wang, Spencer Smith
In Proceedings of the Twelfth International Conference on Learning Representations (ICLR 2024).
DNA-GPT: Divergent N-Gram Analysis for Training-Free Detection of GPT-Generated TextXianjun Yang, Wei Cheng, Yue Wu, Linda Ruth Petzold, William Yang Wang, Haifeng Chen
In Proceedings of the Twelfth International Conference on Learning Representations (ICLR 2024).
Language Control Diffusion: Efficiently Scaling through Space, Time, and TasksEdwin Zhang, Yujie Lu, Shinda Huang, William Yang Wang, Amy Zhang
In Proceedings of the Twelfth International Conference on Learning Representations (ICLR 2024).
Guiding Instruction-based Image Editing via Multimodal Large Language ModelsTsu-Jui Fu, Wenze Hu, Xianzhi Du, William Yang Wang, Yinfei Yang, Zhe Gan
In Proceedings of the Twelfth International Conference on Learning Representations (ICLR 2024).
VELMA: Verbalization Embodiment of LLM Agents for Vision and Language Navigation in Street ViewRaphael Schumann, Wanrong Zhu, Weixi Feng, Tsu-Jui Fu, Stefan Riezler, William Yang Wang
In Proceedings of the 38th Annual AAAI Conference on Artificial Intelligence (AAAI 2024).
Emergent human-like covert attention in feedforward convolutional neural networksSudhanshu Srivastava, William Yang Wang, Miguel Eckstein
In Current Biology, Volume 34, Issue 3, 5 February 2024, Pages 579-593.e12.
Read Anywhere Pointed: Layout-aware GUI Screen Reading with Tree-of-Lens GroundingYue Fan, Lei Ding, Ching-Chen Kuo, Shan Jiang, Yang Zhao, Xinze Guan, Jie Yang, Yi Zhang, Xin Eric WangEMNLP 2024.
Active Listening: Personalized Question Generation in Open-Domain Social Conversation with User Model Based PromptingKevin Bowden, Yue Fan, Winsom Chen, Wen Cui, Davan Harrison, Xin Eric Wang, Marilyn WalkerFindings of EMNLP 2024.
FlexEControl: Flexible and Efficient Multimodal Control for Text-to-Image GenerationXuehai He, Jian Zheng, Jacob Zhiyuan Fang, Robinson Piramuthu, Mohit Bansal, Vicente Ordonez, Gunnar A Sigurdsson, Nanyun Peng, Xin Eric WangTransactions on Machine Learning Research (TMLR) 2024.
SwapAnything: Enabling Arbitrary Object Swapping in Personalized Visual EditingJing Gu, Yilin Wang, Nanxuan Zhao, Wei Xiong, Qing Liu, Zhifei Zhang, He Zhang, Jianming Zhang, HyunJoon Jung, Xin Eric WangECCV 2024.
NavGPT-2: Unleashing Navigational Reasoning Capability for Large Vision-Language ModelsGengze Zhou, Yicong Hong, Zun Wang, Xin Eric Wang, Qi WuECCV 2024.
Muffin or Chihuahua? Challenging Large Vision-Language Models with Multipanel VQAYue Fan, Jing Gu, Kaiwen Zhou, Qianqi Yan, Shan Jiang, Ching-Chen Kuo, Xinze Guan, Xin Eric WangACL 2024.
ViCor: Bridging Visual Understanding and Commonsense Reasoning with Large Language ModelsKaiwen Zhou, Kwonjoon Lee, Teruhisa Mitsu, Xin Eric WangFindings of ACL 2024.
Navigation as Attackers Wish? Towards Building Byzantine-Robust Embodied Agents under Federated LearningYunchao Zhang, Zonglin Di, Kaiwen Zhou, Cihang Xie, Xin Eric WangNAACL 2024.
ComCLIP: Training-Free Compositional Image and Text MatchingKenan Jiang*, Xuehai He*, Ruize Xu, Xin Eric WangNAACL 2024.
Reversing the Forget-Retain Objectives: An Efficient LLM Unlearning Framework from Logit DifferenceJiabao Ji, Yujian Liu, Yang Zhang, Gaowen Liu, Ramana Rao Kompella, Sijia Liu, Shiyu Chang
Advances in Neural Information Processing Systems (NeurIPS 2024).
[Paper][Code]
Revisiting Who's Harry Potter: Towards Targeted Unlearning from a Causal Intervention PerspectiveYujian Liu, Yang Zhang, Tommi S. Jaakkola, Shiyu Chang
Conference on Empirical Methods in Natural Language Processing (EMNLP 2024).
[Paper][Code]
Decomposing Uncertainty for Large Language Models through Input Clarification EnsemblingBairu Hou, Yujian Liu, Kaizhi Qian, Jacob Andreas, Shiyu Chang, Yang Zhang
International Conference on Machine Learning (ICML 2024).
[Paper][Code]
Advancing the Robustness of Large Language Models through Self-Denoised SmoothingJiabao Ji, Bairu Hou, Zhen Zhang, Guanhua Zhang, Wenqi Fan, Qing Li, Yang Zhang, Gaowen Liu, Sijia Liu, Shiyu Chang
Annual Conference of the North American Chapter of the Association for Computational Linguistics (NAACL 2024).
[Paper][Code]
Correcting Diffusion Generation through ResamplingYujian Liu, Yang Zhang, Tommi S. Jaakkola, Shiyu Chang
IEEE Computer Vision and Pattern Recognition (CVPR 2024).
[Paper][Code]
Improving Diffusion Models for Scene Text Editing with Dual EncodersJiabao Ji, Guanhua Zhang, Zhaowen Wang, Bairu Hou, Zhifei Zhang, Brian Price, Shiyu Chang
Transactions on Machine Learning Research (TMLR 2024).
[Paper][Code]
Financial Forecasting from Textual and Tabular Time SeriesRoss Koval, Nicholas Andrews, Xifeng Yan
Findings of the Association for Computational Linguistics (Findings of EMNLP 2024).
[Paper]
Evaluating the Instruction-Following Robustness of Large Language Models to Prompt InjectionZekun Li, Baolin Peng, Pengcheng He, Xifeng Yan
Findings of the Association for Computational Linguistics (Findings of EMNLP 2024).
[Paper]
Creative and Context-Aware Translation of East Asian Idioms with GPT-4K. Tang, P. Song, Y. Qin, X. Yan
Findings of the Association for Computational Linguistics (Findings of EMNLP 2024).
Finetuned Multimodal Language Models Are High-Quality Image-Text Data FiltersW. Wang, K. Mrini, L. Yang, S. Kumar, Y. Tian, X. Yan, H. Wang
arXiv:2403.02677 preprint, 2024.
[Paper]
Can Editing LLMs Inject Harm?C. Chen, B. Huang, Z. Li, Z. Chen, S. Lai, X. Xu, J.-C. Gu, J. Gu, H. Yao, C. Xiao, X. Yan, W. Wang, P. Torr, D. Song, K. Shu
arXiv:2407.20224 preprint, 2024.
[Paper]
MMSci: A Multimodal Multi-Discipline Dataset for PhD-Level Scientific ComprehensionZ. Li, X. Yang, K. Choi, W. Zhu, R. Hsieh, HJ Kim, JH Lim, S. Ji, B. Lee, X. Yan, L. Petzold, S. Wilson, W. Lim, W. Wang
AI4MAT'24 (AI for Accelerated Materials Design) 2024 Spotlights.
[Paper]
Bot or Human? Detecting ChatGPT Imposters with A Single QuestionH. Wang, X. Luo, W. Wang, M. Yu, X. Yan
Conference on Language Modeling (COLM 2024).
[Paper]
Large Language Models as Zero-shot Dialogue State Tracker through Function CallingZ. Li, Z. Chen, M. Ross, P. Huber, S. Moon, Z. Lin, X. Dong, A. Sagar, X. Yan, P. Crook [arxiv]
Proc. of the Annual Meeting of the Association for Computational Linguistics (ACL 2024).
[Paper]
Learning to Compare Financial Reports for Financial ForecastingR. Koval, N. Andrews and X. Yan
Findings of the Association for Computational Linguistics (Findings of EACL 2024).
[Paper]
When LLM Meets DRL: Advancing Jailbreaking Efficiency via DRL-guided SearchXuan Chen, Yuzhou Nie, Wenbo Guo, Xiangyu Zhang
In Annual Conference on Neural Information Processing Systems (NeurIPS 2024).
[Paper]
DFBA: Data Free Backdoor AttacksBochuan Cao, Jinyuan Jia, Chuxuan Hu, Wenbo Guo, Zhen Xiang, Jinghui Chen, Bo Li, Dawn Song
In Annual Conference on Neural Information Processing Systems (NeurIPS 2024).
BandFuzz: A Practical Framework for Collaborative Fuzzing with Reinforcement LearningWenxuan Shi, Hongwei Li, Jiahao Yu, Wenbo Guo, Xinyu Xing
In International Workshop on Search-Based and Fuzz Testing (SBFT 2024).
[Paper]
FORAY: Towards Effective Attack Synthesis against Deep Logical Vulnerabilities in DeFi ProtocolsHongbo Wen, Hanzhi Liu, Jiaxin Song, Yanju Chen, Wenbo Guo, Yu Feng
In ACM Conference on Computer and Communications Security (CCS 2024).
GuideEnricher: Protecting the Anonymity of Ethereum Mixing Service Users with Deep Reinforcement LearningRavindu De Silva, Wenbo Guo, Nicola Ruaro, Ilya Grishchenko, Christopher Kruegel, Giovanni Vigna
In USENIX Security Symposium (USENIX Security 2024).
[Code]
SHINE: Shielding Backdoors in Deep Reinforcement LearningZhuowen Yuan, Wenbo Guo, Jinyuan Jia, Bo Li, Dawn Song
In International Conference on Machine Learning (ICML 2024).
[Paper][Code]
BOXRR-23: 4.7 Million Motion Capture Recordings from 105,000 VR UsersVivek Nair, Wenbo Guo, Rui Wang, James F. O'Brien, Louis Rosenberg, Dawn Song
In IEEE Conference on Virtual Reality and 3D User Interfaces (VR 2024).
[Code]
TextGuard: Provable Defense against Backdoor Attacks on Text ClassificationHengzhi Pei, Jinyuan Jia, Wenbo Guo, Bo Li, Dawn Song
In The Conference on Network and Distributed System Security Symposium (NDSS 2024).
[Paper][Code]
Information-Theoretic Characterizations of Generalization Error for the Gibbs AlgorithmG. Aminian, Y. Bu, L. Toni, M. R. Rodrigues, G. W. Wornell ( equal contribution)
IEEE Transactions on Information Theory, vol. 70, no. 1, pp. 632–655, Jan. 2024.
Are Uncertainty Quantification Capabilities of Evidential Deep Learning a Mirage?J. J. Ryu, M. Shen, S. Ghosh, Y. Bu, P. Sattigeri, S. Das, G. W. Wornell
Conference on Neural Information Processing Systems (NeurIPS 2024).
[Paper]
Information-theoretic Analysis of the Gibbs Algorithm: An Individual Sample ApproachY. Zhu, Y. Bu
IEEE Information Theory Workshop (ITW 2024).
SAUC: Sparsity-Aware Uncertainty Calibration for Spatiotemporal Prediction with Graph Neural NetworksD. Zhuang, Y. Bu, G. Wang, S. Wang, J. Zhao
International Conference on Advances in Geographic Information Systems (ACM SIGSPATIAL 2024).
Learning Orthonormal Features in Self-Supervised Learning using Functional Maximal CorrelationB. Hu, Y. Bu, J. C. Príncipe
IEEE International Conference on Image Processing (ICIP 2024).
[Paper]
An Algorithm for Computing the Capacity of Symmetrized KL Information for Discrete ChannelsH. Chen, G. Aminian, Y. Bu
Allerton Conference on Communication, Control, and Computing (Allerton 2024).
Information-Theoretic Opacity-Enforcement in Markov Decision ProcessesC. Shi, Y. Bu, J. Fu
International Joint Conference on Artificial Intelligence (IJCAI 2024).
Adaptive Text Watermark for Large Language ModelsY. Liu, Y. Bu
International Conference on Machine Learning (ICML 2024).
[Code]
Operator SVD with Neural Networks via Nested Low-Rank ApproximationJ. J. Ryu, X. Xu, H. SM Erol, Y. Bu, L. Zheng, G. W. Wornell
International Conference on Machine Learning (ICML 2024).
[Paper][Code]
Towards Optimal Inverse Temperature in the Gibbs AlgorithmY. Bu
IEEE International Symposium on Information Theory (ISIT 2024).
Group Fairness with Uncertain Sensitive AttributesA. Shah, M. Shen, J. J. Ryu, S. Das, P. Sattigeri, Y. Bu, G. W. Wornell
IEEE International Symposium on Information Theory (ISIT 2024).
[Paper]
Gibbs-Based Information Criteria and the Over-Parameterized RegimeH. Chen, G. W. Wornell, Y. Bu
International Conference on Artificial Intelligence and Statistics (AISTATS 2024).
[Paper]
2023
INSTRUCTSCORE: Towards Explainable Text Generation Evaluation with Automatic FeedbackWenda Xu, Danqing Wang, Liangming Pan, Zhenqiao Song, Markus Freitag, William Yang Wang, Lei Li
In Proceedings of The 2023 Conference on Empirical Methods in Natural Language Processing (EMNLP 2023).
MAF: Multi-Aspect Feedback for Improving Reasoning in Large Language ModelsDeepak Nathani, David Wang, Liangming Pan, William Yang Wang
In Proceedings of The 2023 Conference on Empirical Methods in Natural Language Processing (EMNLP 2023).
EDIS: Entity-Driven Image Search over Multimodal Web ContentSiqi Liu, Weixi Feng, Tsu-Jui Fu, Wenhu Chen, William Yang Wang
In Proceedings of The 2023 Conference on Empirical Methods in Natural Language Processing (EMNLP 2023).
Let's Think Frame by Frame with VIP: A Video Infilling and Prediction Dataset for Evaluating Video Chain-of-ThoughtVaishnavi Himakunthala, Andy Ouyang, Daniel Philip Rose, Ryan He, Alex Mei, Yujie Lu, Chinmay Sonar, Michael Saxon, William Yang Wang
In Proceedings of The 2023 Conference on Empirical Methods in Natural Language Processing (EMNLP 2023).
Collaborative Generative AI: Integrating GPT-k for Efficient Editing in Text-to-Image GenerationWanrong Zhu, Xinyi Wang, Yujie Lu, Tsu-Jui Fu, Xin Eric Wang, Miguel Eckstein, William Yang Wang
In Proceedings of The 2023 Conference on Empirical Methods in Natural Language Processing (EMNLP 2023), Short Paper.
Text-guided 3D Human Generation from 2D CollectionsTsu-Jui Fu, Wenhan Xiong, Yixin Nie, Jingyu Liu, Barlas Oguz, William Yang Wang
In Findings of The 2023 Conference on Empirical Methods in Natural Language Processing (EMNLP 2023 Findings).
Logic-LM: Empowering Large Language Models with Symbolic Solvers for Faithful Logical ReasoningLiangming Pan, Alon Albalak, Xinyi Wang, William Yang Wang
In Findings of The 2023 Conference on Empirical Methods in Natural Language Processing (EMNLP 2023 Findings).
Knowledge-Selective Pretraining for Attribute Value ExtractionHui Liu, Qingyu Yin, Zhengyang Wang, Chenwei Zhang, Haoming Jiang, Yifan Gao, Zheng Li, Xian Li, Chao Zhang, Bing Yin, William Yang Wang, Xiaodan Zhu
In Findings of The 2023 Conference on Empirical Methods in Natural Language Processing (EMNLP 2023 Findings).
ASSERT: Automated Safety Scenario Red Teaming for Evaluating the Robustness of Large Language ModelsAlex Mei, Sharon Levy, William Yang Wang
In Findings of The 2023 Conference on Empirical Methods in Natural Language Processing (EMNLP 2023 Findings).
On the Risk of Misinformation Pollution with Large Language ModelsYikang Pan, Liangming Pan, Wenhu Chen, Preslav Nakov, Min-Yen Kan, William Yang Wang
In Findings of The 2023 Conference on Empirical Methods in Natural Language Processing (EMNLP 2023 Findings).
Empowering Psychotherapy with Large Language Model: Cognitive Distortion Detection through Diagnosis of Thought PromptingZhiyu Chen, Yujie Lu, William Yang Wang
In Findings of The 2023 Conference on Empirical Methods in Natural Language Processing (EMNLP 2023 Findings).
LayoutGPT: Compositional Visual Planning and Generation with Large Language ModelsWeixi Feng*, Wanrong Zhu*, Tsu-Jui Fu, Varun Jampani, Arjun Reddy Akula, Xuehai He, S Basu, Xin Eric Wang, William Yang Wang
In Proceedings of Thirty-seventh Conference on Neural Information Processing Systems (NeurIPS 2023).
Improving Few-Shot Generalization by Exploring and Exploiting Auxiliary DataAlon Albalak, Colin Raffel, William Yang Wang
In Proceedings of Thirty-seventh Conference on Neural Information Processing Systems (NeurIPS 2023).
Large Language Models Are Implicitly Topic Models: Explaining and Finding Good Demonstrations for In-Context LearningXinyi Wang, Wanrong Zhu, Michael Saxon, Mark Steyvers, William Yang Wang
In Proceedings of Thirty-seventh Conference on Neural Information Processing Systems (NeurIPS 2023).
LLMScore: Unveiling the Power of Large Language Models in Text-to-Image Synthesis EvaluationYujie Lu, Xianjun Yang, Xiujun Li, Xin Eric Wang, William Yang Wang
In Proceedings of Thirty-seventh Conference on Neural Information Processing Systems (NeurIPS 2023).
ALGO: Synthesizing Algorithmic Programs with Generated Oracle VerifiersKexun Zhang, Danqing Wang, Jingtao Xia, William Yang Wang, Lei Li
In Proceedings of Thirty-seventh Conference on Neural Information Processing Systems (NeurIPS 2023).
Flexible Attention-Based Multi-Policy Fusion for Efficient Deep Reinforcement LearningZih-Yun Chiu, Yi-Lin Tuan, William Yang Wang, Michael C. Yip
In Proceedings of Thirty-seventh Conference on Neural Information Processing Systems (NeurIPS 2023).
Multimodal C4: An Open, Billion-scale Corpus of Images Interleaved with TextWanrong Zhu*, Jack Hessel*, Anas Awadalla, Samir Yitzhak Gadre, Jesse Dodge, Alex Fang, Youngjae Yu, Ludwig Schmidt, William Yang Wang, Yejin Choi
In Proceedings of Thirty-seventh Conference on Neural Information Processing Systems Datasets and Benchmarks Track (NeurIPS D&B 2023).
Attacking Open-domain Question Answering by Injecting MisinformationLiangming Pan, Wenhu Chen, Min-Yen Kan and William Yang Wang
In Proceedings of the 13th International Joint Conference on Natural Language Processing and the 3rd Conference of the Asia-Pacific Chapter of the Association for Computational Linguistics (IJCNLP-AACL 2023).
Pre-trained Language Models can be Fully Zero-Shot LearnersXuandong Zhao, Siqi Ouyang, Zhiguo Yu, Ming Wu and Lei Li
In Proceedings of the 61th Annual Meeting of the Association for Computational Linguistics (ACL 2023), Long Paper.
Say What You Mean! Large Language Models Speak Too Positively about Negative Commonsense KnowledgeJiangjie Chen, Wei Shi, Ziquan Fu, Sijie Cheng, Lei Li and Yanghua Xiao
In Proceedings of the 61th Annual Meeting of the Association for Computational Linguistics (ACL 2023), Long Paper.
WACO: Word-Aligned Contrastive Learning for Speech TranslationSiqi Ouyang, Rong Ye and Lei Li
In Proceedings of the 61th Annual Meeting of the Association for Computational Linguistics (ACL 2023), Long Paper.
Multilingual Conceptual Coverage in Text-to-Image ModelsMichael Saxon and William Yang Wang
In Proceedings of the 61th Annual Meeting of the Association for Computational Linguistics (ACL 2023), Long Paper.
Retrieval Augmented Pretraining for Text Generation EvaluationWenda Xu, Xian Qian, Mingxuan Wang, Lei Li, and William Yang Wang
In Proceedings of the 61th Annual Meeting of the Association for Computational Linguistics (ACL 2023), Long Paper.
Fact-Checking Complex Claims with Program-Guided ReasoningLiangming Pan, Xiaobao Wu, Xinyuan Lu, Anh Tuan Luu, William Yang Wang, Min-Yen Kan, and Preslav Nakov
In Proceedings of the 61th Annual Meeting of the Association for Computational Linguistics (ACL 2023), Long Paper.
CausalDialogue: Modeling Utterance-level Causality in ConversationsYi-Lin Tuan, Alon Albalak, Wenda Xu, Michael Saxon, Connor F. Pryor, Lise Getoor, and William Yang Wang
In Findings of 61th Annual Meeting of the Association for Computational Linguistics (Findings of ACL 2023), Long Paper.
Lego-MT: Learning Detachable Models for Massively Multilingual Machine TranslationFei Yuan, Yinquan Lu, Wenhao Zhu, Lingpeng Kong, Lei Li, Yu Qiao and Jingjing Xu
In Findings of 61th Annual Meeting of the Association for Computational Linguistics (Findings of ACL 2023), Long Paper.
Foveate, Attribute, and Rationalize: Towards Safe and Trustworthy AIAlex Mei, Sharon Levy, and William Yang Wang
In Findings of 61th Annual Meeting of the Association for Computational Linguistics (Findings of ACL 2023), Long Paper.
Towards Coherent Image Inpainting Using Denoising Diffusion Implicit ModelsGuanhua Zhang, Jiabao Ji, Yang Zhang, Mo Yu, Tommi Jaakkola, Shiyu Chang
In Proceedings of Fortieth International Conference on Machine Learning (ICML 2023).
PromptBoosting: Black-Box Text Classification with Ten Forward PassesBairu Hou, Joe O'Connor, Jacob Andreas, Shiyu Chang, Yang Zhang
In Proceedings of Fortieth International Conference on Machine Learning (ICML 2023).
Importance Weighted Variational Bayes for Protein Sequence DesignZhenqiao Song, Lei Li
In Proceedings of Fortieth International Conference on Machine Learning (ICML 2023).
Protecting Language Generation Models via Invisible WatermarkingXuandong Zhao, Yu-Xiang Wang, Lei Li
In Proceedings of Fortieth International Conference on Machine Learning (ICML 2023).
ReDi: Efficient Learning-Free Diffusion Inference via Trajectory RetrievalKexun Zhang, Xianjun Yang, William Yang Wang, and Lei Li
In Proceedings of Fortieth International Conference on Machine Learning (ICML 2023).
Offline Reinforcement Learning with Closed-Form Policy Improvement OperatorsJiachen Li, Eddie Zhang, Ming Yin, Qinxun Bai, Yu-Xiang Wang, and William Yang Wang
In Proceedings of Fortieth International Conference on Machine Learning (ICML 2023).
WikiWhy: Answering and Explaining Cause-and-Effect QuestionsMatthew Ho, Aditya Sharma, Justin Chang, Michael Saxon, Sharon Levy, Yujie Lu, William Yang Wang
In Proceedings of Eleventh International Conference on Learning Representations (ICLR 2023). Oral Paper: Top 5% out of all 4019 submissions
Neuro-Symbolic Procedural Planning with Commonsense PromptingYujie Lu, Weixi Feng, Wanrong Zhu, Wenda Xu, Xin Eric Wang, Miguel Eckstein, William Yang Wang
In Proceedings of Eleventh International Conference on Learning Representations (ICLR 2023), Spotlight Paper.
Training-Free Structured Diffusion Guidance for Compositional Text-to-Image SynthesisWeixi Feng, Xuehai He, Tsu-Jui Fu, Varun Jampani, Arjun Reddy Akula, Pradyumna Narayana, Sugato Basu, Xin Eric Wang, William Yang Wang
In Proceedings of Eleventh International Conference on Learning Representations (ICLR 2023).
Causal Balancing for Domain GeneralizationXinyi Wang, Michael Saxon, Jiachen Li, Hongyang Zhang, Kun Zhang, William Yang Wang
In Proceedings of Eleventh International Conference on Learning Representations (ICLR 2023).
TextGrad: Advancing Robustness Evaluation in NLP by Gradient-Driven OptimizationBairu Hou, Jinghan Jia, Yihua Zhang, Guanhua Zhang, Yang Zhang, Sijia Liu, and Shiyu Chang
In Proceedings of Eleventh International Conference on Learning Representations (ICLR 2023).
PECO: Examining Single Sentence Label Leakage in Natural Language Inference Datasets through Progressive Evaluation of Cluster OutliersMichael S. Saxon, Xinyi Wang, Wenda Xu and William Yang Wang
In Proceedings of the 17th Conference of the European Chapter of the Association for Computational Linguistics (EACL 2023), Long Paper.
Visualize Before You Write: Imagination-Guided Open-Ended Text GenerationWanrong Zhu, An Yan, Yujie Lu, Wenda Xu, Xin Eric Wang, Miguel Eckstein and William Yang Wang
In Findings of the 17th Conference of the European Chapter of the Association for Computational Linguistics (Findings of EACL 2023), Long Paper.
ImaginE: An Imagination-Based Automatic Evaluation Metric for Natural Language GenerationWanrong Zhu, Xin Eric Wang, An Yan, Miguel Eckstein and William Yang Wang
In Findings of the 17th Conference of the European Chapter of the Association for Computational Linguistics (Findings of EACL 2023), Long Paper.
Uncovering the Disentanglement Capability in Text-to-Image Diffusion ModelsQiucheng Wu, Yujian Liu, Handong Zhao, Ajinkya Kale, Trung Bui, Tong Yu, Zhe Lin, Yang Zhang, Shiyu Chang
In Proceedings of The Thirty-Fourth IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR 2023).
Tell Me What Happened: Unifying Text-guided Video Completion via Multimodal Masked Video GenerationTsu-Jui Fu, Licheng Yu, Ning Zhang, Cheng-Yang Fu, Jong-Chyi Su, William Yang Wang, Sean Bell
In Proceedings of The Thirty-Fourth IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR 2023).
An Empirical Study of End-to-End Video-Language Transformers with Masked Visual ModelingTsu-Jui Fu, Linjie Li, Zhe Gan, Kevin Lin, William Yang Wang, Lijuan Wang, Zicheng Liu
In Proceedings of The Thirty-Fourth IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR 2023).
Converge to the Truth: Factual Error Correction via Iterative Constrained EditingJiangjie Chen, Rui Xu, Wenxuan Zeng, Changzhi Sun, Lei Li, and Yanghua Xiao.
In Proceedings of the AAAI Conference on Artificial Intelligence (AAAI 2023).
2022
FETA: A Benchmark for Few-Sample Task Transfer in Open-Domain DialogueAlon Albalak, Yi-Lin Tuan, Pegah Jandaghi, Connor Pryor, Luke Yoffe, Deepak Ramachandran, Lise Getoor, Jay Pujara and William Yang Wang
In Proceedings of Empirical Methods in Natural Language Processing, (EMNLP 2022).
SafeText: A Benchmark for Exploring Physical Safety in Language ModelsSharon Levy, Emily Allaway, Melanie Subbiah, Lydia Chilton, Desmond Patton, Kathleen McKeown and William Yang Wang
In Proceedings of Empirical Methods in Natural Language Processing, (EMNLP 2022).
ULN: Towards Underspecified Vision-and-Language NavigationWeixi Feng, Tsu-Jui Fu, Yujie Lu and William Yang Wang
In Proceedings of Empirical Methods in Natural Language Processing, (EMNLP 2022).
ConvFinQA: Exploring the Chain of Numerical Reasoning in Conversational Finance Question AnsweringZhiyu Chen, Shiyang Li, Charese Smiley, Zhiqiang Ma, Sameena Shah and William Yang Wang
In Proceedings of Empirical Methods in Natural Language Processing, (EMNLP 2022).
CPL: Counterfactual Prompt Learning for Vision and Language ModelsXuehai He, Diji Yang, Weixi Feng, Tsu-Jui Fu, Arjun Akula, Varun Jampani, Pradyumna Narayana, Sugato Basu, William Yang Wang and Xin Eric Wang
In Proceedings of Empirical Methods in Natural Language Processing, (EMNLP 2022).
Not All Errors are Equal: Learning Text Generation Metrics using Stratified Error SynthesisWenda Xu, Yi-Lin Tuan, Yujie Lu, Michael S. Saxon, Lei Li and William Yang Wang
In Findings of Empirical Methods in Natural Language Processing, (Findings of EMNLP 2022).
Bridging the Training-Inference Gap for Dense Phrase RetrievalGyuwan Kim, Jinhyuk Lee, Barlas Oguz, Wenhan Xiong, Yizhe Zhang, Yashar Mehdad and William Yang Wang
In Findings of Empirical Methods in Natural Language Processing, (Findings of EMNLP 2022).
Mitigating Covertly Unsafe Text within Natural Language SystemsAlex Mei, Anisha Kabir, Sharon Levy, Melanie Subbiah, Emily Allaway, John N. Judge, Desmond Patton, Bruce Bimber, Kathleen McKeown and William Yang Wang
In Findings of Empirical Methods in Natural Language Processing, (Findings of EMNLP 2022).
Calibrating Factual Knowledge in Pretrained Language ModelsQingxiu Dong, Damai Dai, Yifan Song, Jingjing Xu, Zhifang Sui, and Lei Li
In Findings of Empirical Methods in Natural Language Processing, (Findings of EMNLP 2022).
Distillation-Resistant Watermarking for Model Protection in NLPXuandong Zhao, Lei Li, and Yu-Xiang Wang
In Findings of Empirical Methods in Natural Language Processing, (Findings of EMNLP 2022).
LightSeq2: Accelerated Training for Transformer-based Models on GPUsXiaohui Wang, Yang Wei, Ying Xiong, Guyue Huang, Xian Qian, Yufei Ding, Mingxuan Wang, and Lei Li
In Proceedings of The International Conference for High Performance Computing, Networking, Storage and Analysis (SC'22).
Uncovering the Heterogeneous Effects of Preference Diversity on User Activeness: A Dynamic Mixture ModelYunfei Lu, Peng Cui, Linyun Yu, Lei Li, and Wenwu Zhu
In the 28th SIGKDD Conference on Knowledge Discovery and Data Mining (KDD 2022).
Language-Driven Artistic Style TransferTsu-Jui Fu, Xin Eric Wang, William Yang Wang
In Proceedings of European Conference on Computer Vision 2022 (ECCV 2022).
Revisiting and Advancing Fast Adversarial Training Through The Lens of Bi-Level OptimizationYihua Zhang, Guanhua Zhang, Prashant Khanduri, Mingyi Hong, Shiyu Chang, Sijia Liu
In Proceedings of the 39th International Conference on Machine Learning (ICML 2022).
Learning Stable Classifiers by Transferring Unstable FeaturesYujia Bao, Shiyu Chang, Regina Barzilay
In Proceedings of the 39th International Conference on Machine Learning (ICML 2022).
Improving Self-Supervised Speech Representations by Disentangling SpeakersKaizhi Qian‡, Yang Zhang‡, Heting Gao, Junru Ni, Cheng-I Lai, David Cox, Mark A. Hasegawa-Johnson, Shiyu Chang
In Proceedings of the 39th International Conference on Machine Learning (ICML 2022).
Data-Efficient Double-Win Lottery Tickets from Robust Pre-trainingTianlong Chen, Zhenyu Zhang, Sijia Liu, Yang Zhang, Shiyu Chang, Zhangyang Wang
In Proceedings of the 39th International Conference on Machine Learning (ICML 2022).
Linearity Grafting: How Neuron Pruning Helps Certifiable RobustnessTianlong Chen, Huan Zhang, Zhenyu Zhang, Shiyu Chang, Sijia Liu, Pin-Yu Chen, Zhangyang Wang
In Proceedings of the 39th International Conference on Machine Learning (ICML 2022).
On the Learning of Non-autoregressive TransformersFei Huang, Tianhua Tao, Hao Zhou, Lei Li, and Minlie Huang
In Proceedings of the 39th International Conference on Machine Learning (ICML 2022).
Imagination-Augmented Natural Language UnderstandingYujie Lu, Wanrong Zhu, Xin Eric Wang, Miguel Eckstein, William Yang Wang
In Proceedings of 2022 Annual Conference of the North American Chapter of the Association for Computational Linguistics (NAACL 2022), Long Paper.
Diagnosing Vision-and-Language Navigation: What Really MattersWanrong Zhu, Yuankai Qi, Pradyumna Narayana, Kazoo Sone, Sugato Basu, Xin Eric Wang, Qi Wu, Miguel P. Eckstein, William Yang Wang
In Proceedings of 2022 Annual Conference of the North American Chapter of the Association for Computational Linguistics (NAACL 2022), Long Paper.
DiffCSE: Difference-based Contrastive Learning for Sentence EmbeddingsYung-Sung Chuang, Rumen Dangovski, Hongyin Luo, Yang Zhang, Shiyu Chang, Marin Soljačić, Shang-Wen Li, Wen-tau Yih, Yoon Kim, James Glass
In Proceedings of 2022 Annual Conference of the North American Chapter of the Association for Computational Linguistics (NAACL 2022), Long Paper.
Cross-modal Contrastive Learning for Speech TranslationRong Ye, Mingxuan Wang, and Lei Li
In Proceedings of 2022 Annual Conference of the North American Chapter of the Association for Computational Linguistics (NAACL 2022), Long Paper.
Provably Confidential Language ModellingXuandong Zhao, Lei Li, and Yu-Xiang Wang
In Proceedings of 2022 Annual Conference of the North American Chapter of the Association for Computational Linguistics (NAACL 2022), Long Paper.
MTG: A Benchmark Suite for Multilingual Text GenerationYiran Chen, Zhenqiao Song, Xianze Wu, Danqing Wang, Jingjing Xu, Jiaze Chen, Hao Zhou, and Lei Li
In Findings of 2022 Annual Conference of the North American Chapter of the Association for Computational Linguistics (Findings of NAACL 2022), Long Paper.
KETOD: Knowledge-Enriched Task-Oriented DialogueZhiyu Chen, Bing Liu, Seungwhan Moon, Chinnadhurai Sankar, Paul A. Crook, William Yang Wang
In Findings of 2022 Annual Conference of the North American Chapter of the Association for Computational Linguistics (Findings of NAACL 2022), Long Paper.
M3L: Language-based Video Editing via Multi-Modal Multi-Level TransformersTsu-Jui Fu, Xin Eric Wang, Scott Grafton, Miguel Eckstein, William Yang Wang
In Proceedings of Thirty-Third IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR 2022), Long Paper.
Quarantine: Sparsity Can Uncover the Trojan Attack Trigger for FreeTianlong Chen‡, Zhenyu Zhang‡, Yihua Zhang‡, Shiyu Chang, Sijia Liu, and Zhangyang Wang
In Proceedings of Thirty-Third IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR 2022), Long Paper.
Learning Design and Construction with Varying-Sized Materials via Prioritized Memory ResetsYunfei Li, Tao Kong, Lei Li, and Yi Wu
In IEEE International Conference on Robotics and Automation (ICRA 2022).
On the Impact of Noises in Crowd-Sourced Data for Speech TranslationSiqi Ouyang, Rong Ye, and Lei Li
In Proceedings of the 19th International Conference on Spoken Language Translation (IWSLT 2022).
latent-GLAT: Glancing at Latent Variables for Parallel Text GenerationYu Bao, Hao Zhou, Shujian Huang, Dongqi Wang, Lihua Qian, Xinyu Dai, Jiajun Chen, and Lei Li
In Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (ACL 2022), Long Paper.
Learning When to Translate for Streaming SpeechQianqian Dong, Yaoming Zhu, Mingxuan Wang, and Lei Li
In Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (ACL 2022), Long Paper.
STEMM: Self-learning with Speech-text Manifold Mixup for Speech TranslationQingkai Fang, Rong Ye, Lei Li, Yang Feng, and Mingxuan Wang
In Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (ACL 2022), Long Paper.
Contextual Representation Learning beyond Masked Language ModelingZhiyi Fu, Wangchunshu Zhou, Jingjing Xu, Hao Zhou, and Lei Li
In Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (ACL 2022), Long Paper.
HybriDialogue: An Information-Seeking Dialogue Dataset Grounded on Tabular and Textual DataKai Nakamura, Sharon Levy, Yi-Lin Tuan, Wenhu Chen, William Yang Wang
In Findings of 60th Annual Meeting of the Association for Computational Linguistics (Findings of ACL 2022), Long Paper.
Towards Large-Scale Interpretable Knowledge Graph Reasoning for Dialogue SystemsYi-Lin Tuan, Sajjad Beygi, Maryam Fazel-Zarandi, Qiaozi Gao, Alessandra Cervone, William Yang Wang
In Findings of 60th Annual Meeting of the Association for Computational Linguistics (Findings of ACL 2022), Long Paper.
Query and Extract: Refining Event Extraction as Type-oriented Binary DecodingSijia Wang, Mo Yu, Shiyu Chang, Lichao Sun, Lifu Huang
In Findings of 60th Annual Meeting of the Association for Computational Linguistics (Findings of ACL 2022), Long Paper.
E-KAR: A Benchmark for Rationalizing Natural Language Analogical ReasoningJiangjie Chen, Rui Xu, Ziquan Fu, Wei Shi, Zhongqiao Li, Xinbo Zhang, Changzhi Sun, Lei Li, Yanghua Xiao, and Hao Zhou
In Findings of 60th Annual Meeting of the Association for Computational Linguistics (Findings of ACL 2022), Long Paper.
Rethinking Document-level Neural Machine TranslationZewei Sun, Mingxuan Wang, Hao Zhou, Chengqi Zhao, Shujian Huang, Jiajun Chen, and Lei Li
In Findings of 60th Annual Meeting of the Association for Computational Linguistics (Findings of ACL 2022), Long Paper.
Compressing Sentence Representation via Homomorphic Projective DistillationXuandong Zhao, Zhiguo Yu, Ming Wu, and Lei Li
In Findings of 60th Annual Meeting of the Association for Computational Linguistics (Findings of ACL 2022), Short Paper.
How to Robustify Black-Box ML Models? A Zeroth-Order Optimization PerspectiveYimeng Zhang, Yuguang Yao, Jinghan Jia, Jinfeng Yi, Mingyi Hong, Shiyu Chang, Sijia Liu
In Proceedings of the International Conference on Learning Representations (ICLR 2022).
Adversarial Support AlignmentShangyuan Tong, Timur Garipov‡, Yang Zhang, Shiyu Chang, Tommi S. Jaakkola
In Proceedings of the International Conference on Learning Representations (ICLR 2022).
Optimizer AmalgamationTianshu Huang, Tianlong Chen, Sijia Liu, Shiyu Chang, Lisa Amini, Zhangyang Wang
In Proceedings of the International Conference on Learning Representations (ICLR 2022).
switch-GLAT: Multilingual Parallel Machine Translation via Code-switch DecoderZhenqiao Song, Hao Zhou, Lihua Qian, Jingjing Xu, Shanbo Cheng, Mingxuan Wang, and Lei Li
In Proceedings of the International Conference on Learning Representations (ICLR 2022).
Enhancing Cross-lingual Transfer by Manifold MixupHuiyun Yang, Huadong Chen, Hao Zhou, and Lei Li
In Proceedings of the International Conference on Learning Representations (ICLR 2022).
Self-Supervised Knowledge Assimilation for Expert-Layman Text Style TransferWenda Xu, Michael Saxon, Misha Sra, William Yang Wang
In Proceedings of the Thirty-Sixth AAAI Conference on Artificial Intelligence (AAAI 2022).
DOC2PPT: Automatic Presentation Slides Generation from Scientific DocumentsTsu-Jui Fu, William Yang Wang, Daniel McDuff, Yale Song
In Proceedings of the Thirty-Sixth AAAI Conference on Artificial Intelligence (AAAI 2022).
LOREN: Logic-Regularized Reasoning for Interpretable Fact VerificationJiangjie Chen, Qiaoben Bao, Changzhi Sun, Xinbo Zhang, Jiaze Chen, Hao Zhou, Yanghua Xiao, and Lei Li
In Proceedings of the Thirty-Sixth AAAI Conference on Artificial Intelligence (AAAI 2022).
Unsupervised Editing for Counterfactual StoriesJiangjie Chen, Chun Gan, Sijie Cheng, Hao Zhou, Yanghua Xiao, and Lei Li
In Proceedings of the Thirty-Sixth AAAI Conference on Artificial Intelligence (AAAI 2022).
Non-Autoregressive Translation with Layer-Wise Prediction and Deep SupervisionChenyang Huang, Hao Zhou, Osmar Zaiane, Lili Mou, and Lei Li
In Proceedings of the Thirty-Sixth AAAI Conference on Artificial Intelligence (AAAI 2022).
Towards Understanding Gender-Seniority Compound Bias in Natural Language GenerationSamhita Honnavalli, Aesha Parekh, Lily Ou, Sophie Groenwold, Sharon Levy, Vicente Ordonez and William Yang Wang
In Proceedings of The 2022 International Conference on Language Resources and Evaluation (LREC 2022), Oral Paper.
Learning to Prioritize: Precision-Driven Sentence Filtering for Long Text SummarizationAlex Mei, Anisha Kabir, Rukmini Bapat, John N. Judge, Tony Sun and William Yang Wang
In Proceedings of The 2022 International Conference on Language Resources and Evaluation (LREC 2022), Oral Paper.
ICM-3D: Instantiated Category Modeling for 3D Instance SegmentationRuihang Chu, Yukang Chen, Tao Kong, Lu Qi, and Lei Li.
In IEEE Robotics and Automation Letters (RA-L 2022).
2021
Duplex Sequence-to-Sequence Learning for Reversible Machine TranslationZaixiang Zheng, Hao Zhou, Shujian Huang, Jiajun Chen, Jingjing Xu, and Lei Li
In Proceedings of Thirty-fifth Conference on Neural Information Processing Systems (NeurIPS 2021), Long Paper.
Understanding Interlocking Dynamics of Cooperative RationalizationMo Yu, Yang Zhang, Shiyu Chang, Tommi S. Jaakkola
In Proceedings of Thirty-fifth Conference on Neural Information Processing Systems (NeurIPS 2021), Long Paper.
TransGAN: Two Pure Transformers Can Make One Strong GAN, and That Can Scale UpYifan Jiang, Shiyu Chang, Zhangyang Wang
In Proceedings of Thirty-fifth Conference on Neural Information Processing Systems (NeurIPS 2021), Long Paper.
Counterfactual Maximum Likelihood Estimation for Training Deep NetworksXinyi Wang, Wenhu Chen, Michael Saxon, William Yang Wang
In Proceedings of Thirty-fifth Conference on Neural Information Processing Systems (NeurIPS 2021), Long Paper.
Local Explanation of Dialogue Response GenerationYi-Lin Tuan, Connor Pryor, Wenhu Chen, Lise Getoor, William Yang Wang
In Proceedings of Thirty-fifth Conference on Neural Information Processing Systems (NeurIPS 2021), Long Paper.
A Dataset for Answering Time-Sensitive QuestionsWenhu Chen, Xinyi Wang, William Yang Wang
In Proceedings of Thirty-fifth Conference on Neural Information Processing Systems Datasets and Benchmarks Track (NeurIPS 2021), Long Paper.
VALUE: A Multi-Task Benchmark for Video-and-Language Understanding EvaluationLinjie Li, Jie Lei, Zhe Gan, Licheng Yu, Yen-Chun Chen, Rohit Pillai, Yu Cheng, Luowei Zhou, Xin Eric Wang, William Yang Wang, Tamara Lee Berg, Mohit Bansal, Jingjing Liu, Lijuan Wang, Zicheng Liu
In Proceedings of Thirty-fifth Conference on Neural Information Processing Systems Datasets and Benchmarks Track (NeurIPS 2021), Long Paper.
FinQA: A Dataset of Numerical Reasoning over Financial DataZhiyu Chen, Wenhu Chen, Charese Smiley, Sameena Shah, Iana Borova, Dylan Langdon, Reema Moussa, Matt Beane, Ting-Hao Huang, Bryan R. Routledge, William Yang Wang
In Proceedings of The 2021 Conference on Empirical Methods in Natural Language Processing (EMNLP 2021), Long Paper, ACL.
Modeling Disclosive Transparency in NLP Application DescriptionsMichael Saxon, Sharon Levy, Xinyi Wang, Alon Albalak and William Yang Wang
In Proceedings of The 2021 Conference on Empirical Methods in Natural Language Processing (EMNLP 2021), Long Paper, ACL.
A Massively Multilingual Analysis of Cross-linguality in Shared Embedding SpaceAlexander Jones, William Yang Wang and Kyle Mahowald
In Proceedings of The 2021 Conference on Empirical Methods in Natural Language Processing (EMNLP 2021), Long Paper, ACL.
Open-Domain Question-Answering for COVID-19 and Other Emergent DomainsSharon Levy, Kevin Mo, Wenhan Xiong and William Yang Wang
In Proceedings of The 2021 Conference on Empirical Methods in Natural Language Processing (EMNLP 2021), Demos Track, ACL.
Zero-Shot Fact Verification by Claim GenerationLiangming Pan, Wenhu Chen, Wenhan Xiong, Min-Yen Kan and William Yang Wang
In Findings of The Joint Conference of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Findings of ACL-IJCNLP 2021), Short Paper, ACL.
Neural Stylistic Response Generation with Disentangled Latent VariablesQingfu Zhu, Wei-Nan Zhang, Ting Liu and William Yang Wang
In Findings of The Joint Conference of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Findings of ACL-IJCNLP 2021), Long Paper, ACL.
Investigating Memorization of Conspiracy Theories in Text GenerationSharon Levy, Michael Saxon and William Yang Wang
In Findings of The Joint Conference of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Findings of ACL-IJCNLP 2021), Long Paper, ACL.
Unsupervised Multi-hop Question Answering by Question GenerationLiangming Pan, Wenhu Chen, Wenhan Xiong, Min-Yen Kan and William Yang Wang
In Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics - Human Language Technologies (NAACL-HLT 2021), Long Paper, ACL.
Semi-Supervised Policy Initialization for Playing Games with Language HintsTsu-Jui Fu and William Yang Wang
In Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics - Human Language Technologies (NAACL-HLT 2021), Short Paper, ACL.
Open Question Answering over Tables and TextWenhu Chen, Ming-Wei Chang, Eva Schlinger, William Yang Wang, William Cohen
In Proceedings of the International Conference on Learning Representations (ICLR 2021), Full Paper.
Answering Complex Open-Domain Questions with Multi-Hop Dense RetrievalWenhan Xiong, Xiang Lorraine Li, Srini Iyer, Jingfei Du, Patrick Lewis, William Yang Wang, Yashar Mehdad, Wen-tau Yih, Sebastian Riedel, Douwe Kiela, Barlas Oguz
In Proceedings of the International Conference on Learning Representations (ICLR 2021), Full Paper.
Progressively Pretrained Dense Corpus Index for Open-Domain Question AnsweringWenhan Xiong, Hong Wang and William Yang Wang
In Proceedings of the 2021 Conference of the European Chapter of the Association for Computational Linguistics (EACL 2021), Long Paper, ACL.
On Hallucination and Predictive Uncertainty in Conditional Language GenerationYijun Xiao and William Yang Wang
In Proceedings of the 2021 Conference of the European Chapter of the Association for Computational Linguistics (EACL 2021), Long Paper, ACL.
L2C: Describing Visual Differences Needs Semantic Understanding of IndividualsAn Yan, Xin Wang, Tsu-Jui Fu and William Yang Wang
In Proceedings of the 2021 Conference of the European Chapter of the Association for Computational Linguistics (EACL 2021), Short Paper, ACL.
Multimodal Text Style Transfer for Outdoor Vision-and-Language NavigationWanrong Zhu, Xin Wang, Tsu-Jui Fu, An Yan, Pradyumna Narayana, Kazoo Sone, Sugato Basu and William Yang Wang
In Proceedings of the 2021 Conference of the European Chapter of the Association for Computational Linguistics (EACL 2021), Long Paper, ACL.
HULK: An Energy Efficiency Benchmark Platform for Responsible Natural Language ProcessingXiyou Zhou, Zhiyu Chen, Xiaoyong Jin and William Yang Wang
In Proceedings of the 2021 Conference of the European Chapter of the Association for Computational Linguistics (EACL 2021), Demo Paper, ACL.
Meta Module Network for Compositional Visual ReasoningWenhu Chen, Zhe Gan, Linjie Li, Yu Cheng, William Yang Wang, Jingjing Liu
In Proceedings of Winter Conference on Applications of Computer Vision (WACV 2021), Oral Paper, IEEE/CVF.
Best Student Paper Award Honorable Mention.
2020
Investigating African-American Vernacular English in Transformer-Based Text GenerationSophie Groenwold, Lily Ou, Aesha Parekh, Samhita Honnavalli, Sharon Levy, Diba Mirza and William Yang Wang
In Proceedings of Conference on Empirical Methods in Natural Language Processing (EMNLP 2020), Short Paper, ACL.
KGLM: Pretrained Knowledge-Grounded Language Model for Data-to-Text GenerationWenhu Chen, Yu Su, Xifeng Yan and William Yang Wang
In Proceedings of Conference on Empirical Methods in Natural Language Processing (EMNLP 2020), Long Paper, ACL.
Towards Understanding Sample Variance in Visually Grounded Language Generation: Evaluations and ObservationsWanrong Zhu, Xin Wang, Pradyumna Narayana, Kazoo Sone, Sugato Basu, William Yang Wang
In Proceedings of Conference on Empirical Methods in Natural Language Processing (EMNLP 2020), Short Paper, ACL.
Iterative Language-Based Image Editing via Self-Supervised Counterfactual ReasoningTsu-Jui Fu, Xin Wang, Scott Grafton, Miguel Eckstein, William Yang Wang
In Proceedings of Conference on Empirical Methods in Natural Language Processing (EMNLP 2020), Long Paper, ACL.
Counterfactual Off-Policy Training for Neural Dialogue GenerationQingfu Zhu, Wei-Nan Zhang, Ting Liu and William Yang Wang
In Proceedings of Conference on Empirical Methods in Natural Language Processing (EMNLP 2020), Long Paper, ACL.
Logic2Text: High-Fidelity Natural Language Generation from Logical FormsZhiyu Chen, Wenhu Chen, Hanwen Zha, Xiyou Zhou, Yunkai Zhang, Sairam Sundaresan, William Yang Wang
In Findings of Conference on Empirical Methods in Natural Language Processing (EMNLP 2020), Findings of EMNLP, ACL.
HybridQA: A Dataset of Multi-Hop Question Answering over Tabular and Textual DataWenhu Chen, Hanwen Zha, Zhiyu Chen, Wenhan Xiong, Hong Wang, William Wang
In Findings of Conference on Empirical Methods in Natural Language Processing (EMNLP 2020), Findings of EMNLP, ACL.
Learning to Stop: A Simple yet Effective Approach to Urban Vision-Language NavigationJiannan Xiang, Xin Wang and William Yang Wang
In Findings of Conference on Empirical Methods in Natural Language Processing (EMNLP 2020), Findings of EMNLP, ACL.
Counterfactual Vision-and-Language Navigation via Adversarial Path SamplerTsu-Jui Fu, Xin Wang, Matthew F Peterson, Scott T. Grafton, Miguel Eckstein, William Yang Wang
In Proceedings of the 16th European Conference on Computer Vision (ECCV 2020), Spotlight Paper: Top 5% out of 5025 submissions , Springer.
Environment-agnostic Multitask Learning for Natural Language Grounded NavigationXin Wang, Vihan Jain, Eugene Ie, William Yang Wang, Zornitsa Kozareva, Sujith Ravi
In Proceedings of the 16th European Conference on Computer Vision (ECCV 2020), full paper, Springer
SafeRoute: Learning to Navigate Streets Safely in an Urban EnvironmentSharon Levy, Wenhan Xiong, Elizabeth Belding, and William Yang Wang
in ACM Transactions on Intelligent Systems and Technology (ACM TIST), journal paper, ACM, 2020
Logical Natural Language Generation from Open-Domain TablesWenhu Chen, Jianshu Chen, Yu Su, Zhiyu Chen and William Yang Wang
In Proceedings of The 58th Annual Meeting of the Association for Computational Linguistics (ACL 2020)
Couple-VAE: Mitigating the Encoder-Decoder Incompatibility in Variational Text Modeling with Coupled Deterministic NetworksChen Wu, Prince Zizhuang Wang, William Yang Wang
In Proceedings of The 58th Annual Meeting of the Association for Computational Linguistics (ACL 2020)
Towards Understanding Gender Bias in Relation ExtractionAndrew Gaut, Tony Sun, Shirlyn Tang, Yuxin Huang, Jing Qian, Mai ElSherief, Jieyu Zhao, Diba Mirza, Elizabeth Belding, Kai-Wei Chang, William Yang Wang
In Proceedings of The 58th Annual Meeting of the Association for Computational Linguistics (ACL 2020)
Few-Shot NLG with Pre-Trained Language ModelZhiyu Chen, Harini Eavani, Wenhu Chen, Yinyin Liu, William Yang Wang
In Proceedings of The 58th Annual Meeting of the Association for Computational Linguistics (ACL 2020)
Vision-Language Navigation Policy Learning and AdaptationXin Wang, Qiuyuan Huang, Asli Celikyilmaz, Jianfeng Gao, Dinghan Shen, Yuan-Fang Wang, William Yang Wang, Lei Zhang
IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI)
REVERIE: Remote Embodied Visual Referring Expression in Real Indoor EnvironmentsYuankai Qi, Qi Wu, Peter Anderson, Xin Wang, William Yang Wang, Chunhua Shen, Anton van den Hengel
In Proceedings of the 2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR 2020)[PDF]
Unsupervised Reinforcement Learning of Transferable Meta-Skills for Embodied NavigationJuncheng Li, Xin Wang, Siliang Tang, Haizhou Shi, Fei Wu, Yueting Zhuang, William Yang Wang
In Proceedings of the 2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR 2020)[PDF]
Fakeddit: A New Multimodal Benchmark Dataset for Fine-grained Fake News DetectionKai Nakamura, Sharon Levy, William Yang Wang
In roceedings of 12th International Conference on Language Resources and Evaluation (LREC 2020)[PDF]
A Survey on Natural Language Processing for Fake News DetectionRay Oshikawa, Jing Qian, William Yang Wang
In Proceedings of 12th International Conference on Language Resources and Evaluation (LREC 2020)[PDF]
Pretrained Encyclopedia: Weakly Supervised Knowledge-Pretrained Language ModelWenhan Xiong, Jingfei Du, William Yang Wang, Veselin Stoyanov
In Proceedings of the International Conference on Learning Representations (ICLR 2020)[PDF]
TabFact: A Large-scale Dataset for Table-based Fact VerificationWenhu Chen, Hongmin Wang, Jianshu Chen, Yunkai Zhang, Hong Wang, Shiyang Li, Xiyou Zhou, William Yang Wang
In Proceedings of the International Conference on Learning Representations (ICLR 2020)[PDF]
Generative Adversarial Zero-Shot Relational Learning for Knowledge GraphsPengda Qin, Xin Wang, Wenhu Chen, Chunyun Zhang, Weiran Xu, William Yang Wang
In Proceedings of the Thirty-Fourth AAAI Conference on Artificial Intelligence (AAAI 2020)
Multi-Task Self-Supervised Learning for Disfluency DetectionShaolei Wang, Wanxiang Che, Qi Liu, Pengda Qin, Ting Liu, William Yang Wang
In Proceedings of the Thirty-Fourth AAAI Conference on Artificial Intelligence (AAAI 2020)
2019
Neural Gaussian Copula for Variational AutoencoderPrince Zizhuang Wang, William Yang Wang
In Proceedings of 2019 Conference on Empirical Methods in Natural Language Processing and 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP 2019)[PDF][BIB]
Deep Reinforcement Learning with Distributional Semantic Rewards for Abstractive SummarizationSiyao Li, Deren Lei, Pengda Qin, William Yang Wang
In Proceedings of 2019 Conference on Empirical Methods in Natural Language Processing and 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP 2019)[PDF][BIB]
A Benchmark Dataset for Learning to Intervene in Online Hate SpeechJing Qian, Anna Bethke, Yinyin Liu, Elizabeth Belding, William Yang Wang
In Proceedings of 2019 Conference on Empirical Methods in Natural Language Processing and 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP 2019)[PDF][BIB]
VATEX: A Large-Scale, High-Quality Multilingual Dataset for Video-and-Language ResearchXin Wang, Jiawei Wu, Junkun Chen, Lei Li, Yuan-Fang Wang, William Yang Wang
In Proceedings of the 17th CVF/IEEE International Conference on Computer Vision (ICCV 2019)[PDF][BIB][WEBSITE]
Semantically Conditioned Dialog Response Generation via Hierarchical Disentangled Self-AttentionWenhu Chen, Jianshu Chen, Pengda Qin, Xifeng Yan, William Yang Wang
In Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics (ACL 2019).
[Paper]
[Bib]
[Code]
Improving Question Answering over Incomplete KBs with Knowledge-Aware ReaderWenhan Xiong, Mo Yu, Shiyu Chang, Xiaoxiao Guo, William Yang Wang
In Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics (ACL 2019).
[Paper]
[Bib]
[Code]
TWEETQA: A Social Media Focused Question Answering DatasetWenhan Xiong, Jiawei Wu, Hong Wang, Vivek Kulkarni, Mo Yu, Xiaoxiao Guo, Shiyu Chang, William Yang Wang
In Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics (ACL 2019).
[Preprint coming soon]
[Bib]
[Code]
What Should I Ask? Using Conversationally Informative Rewards for Goal-Oriented Visual DialogPushkar shukla, Carlos Elmadjian, Richika Sharan, Vivek Kulkarni, Matthew Turk, William Yang Wang
In Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics (ACL 2019).
[Preprint coming soon]
[Bib]
[Code]
Self-Supervised Learning for Contextualized Extractive SummarizationHong Wang, Xin Wang, Wenhan Xiong, Mo Yu, Xiaoxiao Guo, Shiyu Chang, William Yang Wang
In Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics (ACL 2019).
[Paper]
[Bib]
[Code]
Towards Explainable NLP: A Generative Explanation Framework for Text ClassificationHui Liu, Qingyu Yin, William Yang Wang
In Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics (ACL 2019).
[Paper]
[Bib]
[Code]
Mitigating Gender Bias in Natural Language Processing: Literature ReviewTony Sun, Andrew Gaut, Shirlyn Tang, Yuxin Huang, Mai ElSherief, Jieyu Zhao, Diba Mirza, Elizabeth Belding, Kai-Wei Chang, William Yang Wang
In Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics (ACL 2019).
[Paper]
[Bib]
[Code]
Reinforced Cross-Modal Matching and Self-Supervised Imitation Learning for Vision-Language NavigationXin Wang, Qiuyuan Huang, Asli Celikyilmaz, Jianfeng Gao, Dinghan Shen, Yuan-Fang Wang, William Yang Wang, Lei Zhang
In Proceedings of the 31st IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR 2019).
[Paper][Bib]
[Code]
*Best Student Paper Award
Sentence Embedding Alignment for Lifelong Relation ExtractionHong Wang, Wenhan Xiong, Mo Yu, Xiaoxiao Guo, Shiyu Chang, William Yang Wang
In Proceedings of the 2019 Annual Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL-HLT 2019).
[Paper]
[Bib]
[Code]
Riemannian Normalizing Flow on Variational Wasserstein Autoencoder for Text ModelingPrince Zizhuang Wang, William Yang Wang
In Proceedings of the 2019 Annual Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL-HLT 2019).
[Paper]
[Bib]
[Code]
Imposing Label-Relational Inductive Bias for Extremely Fine-Grained Entity TypingWenhan Xiong, Jiawei Wu, Deren Lei, Mo Yu, Shiyu Chang, Xiaoxiao Guo, William Yang Wang
In Proceedings of the 2019 Annual Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL-HLT 2019).
[Paper]
[Bib]
[Code]
Learning to Decipher Hate SymbolsJing Qian, Mai ElSherief, Elizabeth Belding, William Yang Wang
In Proceedings of the 2019 Annual Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL-HLT 2019).
[Paper]
[Bib]
[Code]
How Large A Vocabulary Does Text Classification Need? A Variational Approach on Vocabulary SelectionWenhu Chen, Yu Su, Yilin Shen, Zhiyu Chen, Xifeng Yan, William Yang Wang
In Proceedings of the 2019 Annual Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL-HLT 2019).
[Paper]
[Bib]
[Code]
Learning to Compose Topic-Aware Mixture of Experts for Zero-Shot Video CaptioningXin Wang, Jiawei Wu, Da Zhang, Yu Su, William Yang Wang
In Proceedings of the Thirty-Third AAAI Conference on Artificial Intelligence (AAAI 2019).
[Paper][Bib][Code]
Quantifying Uncertainties in Natural Language Processing TasksYijun Xiao, William Yang Wang
In Proceedings of the Thirty-Third AAAI Conference on Artificial Intelligence (AAAI 2019).
[Paper][Bib]
[Code]
2018
One-Shot Relational Learning for Knowledge GraphsWenhan Xiong, Mo Yu, Shiyu Chang, Xiaoxiao Guo, William Yang Wang
In Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing (EMNLP 2018).
[Paper][Bib][Code]
Hierarchical CVAE for Fine-Grained Hate Speech ClassificationJing Qian, Mai ElSherief, Elizabeth Belding, William Yang Wang
In Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing (EMNLP 2018).
[Paper][Bib]
[Code]
Multi-view Models for Political Ideology Detection of News ArticlesVivek Kulkarni, Junting Ye, Steve Skienam, William Yang Wang
In Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing (EMNLP 2018).
[Paper][Bib]
[Code]
XL-NBT: A Cross-lingual Neural Belief Tracking FrameworkWenhu Chen, Jianshu Chen, Yu Su, Xin Wang, Dong Yu, Xifeng Yan, William Yang Wang
In Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing (EMNLP 2018).
[Paper][Bib][Code]
Look Before You Leap: Bridging Model-Free and Model-Based Reinforcement Learning for Planned-Ahead Vision-and-Language NavigationXin Wang*, Wenhan Xiong*, Hongmin Wang, William Yang Wang (* Equal contribution)
In Proceedings of the 15th European Conference on Computer Vision (ECCV 2018).
[Paper][Bib]
[Code]
No Metrics Are Perfect: Adversarial Reward Learning for Visual StorytellingXin Wang*, Wenhu Chen*, Yuan-Fang Wang, William Yang Wang (* Equal contribution)
In Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (ACL 2018).
[Paper][Bib][Code]
MOJITALK: Generating Emotional Responses at ScaleXianda Zhou, William Yang Wang
In Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (ACL 2018).
[Paper][Bib][Code]
DSGAN: Generative Adversarial Training for Robust Distant Supervision Relation ExtractionPengda Qin, Weiran Xu, William Yang Wang
In Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (ACL 2018).
[Paper][Bib]
[Code]
Deep Reinforcement Learning for Chinese Zero Pronoun ResolutionQingyu Yin, Yu Zhang, Wei-Nan Zhang, Ting Liu, William Yang Wang
In Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (ACL 2018).
[Paper]
[Bib]
[Code]
Robust Distant Supervision via Deep Reinforcement LearningPengda Qin, Weiran Xu, William Yang Wang
In Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (ACL 2018).
[Paper][Bib]
[Code]
Scheduled Policy Optimization for Natural Language Communication with Intelligent AgentsWenhan Xiong, Xiaoxiao Guo, Mo Yu, Shiyu Chang, Bowen Zhou, William Yang Wang
In Proceedings of the 27th International Joint Conference on Artificial Intelligence and the 23rd European Conference on Artificial Intelligence (IJCAI-ECAI 2018).
[Paper][Bib][Code]
Hate Lingo: A Target-based Linguistic Analysis of Hate Speech in Social MediaMai Elsherief, Vivek Kulkarni, Dana Nguyen, William Yang Wang, Elizabeth Belding
In Proceedings of the 12th International AAAI Conference on Web and Social Media (ICWSM 2018).
[Paper][Bib]
[Code]
Video Captioning via Hierarchical Reinforcement LearningXin Wang, Wenhu Chen, Jiawei Wu, Yuan-Fang Wang, William Yang Wang
In Proceedings of the Thirtieth IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR 2018).
[Paper][Bib][Data]
[Code]
Leveraging Intra-User and Inter-User Representation Learning for Automated Hate Speech DetectionJing Qian, Mai ElSherief, Elizabeth Belding, William Yang Wang
In Proceedings of the 16th Annual Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL-HLT 2018).
[Paper][Bib]
[Code]
Watch, Listen, and Describe: Globally and Locally Aligned Cross-Modal Attentions for Video CaptioningXin Wang, Yuan-Fang Wang, William Yang Wang
In Proceedings of the 16th Annual Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL-HLT 2018).
[Paper][Bib]
[Code]
Variational Knowledge Graph ReasoningWenhu Chen, Wenhan Xiong, Xifeng Yan, William Yang Wang
In Proceedings of the 16th Annual Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL-HLT 2018).
[Paper][Bib]
[Code]
Simple Models for Word Formation in SlangVivek Kulkarni, William Yang Wang
In Proceedings of the 16th Annual Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL-HLT 2018).
[Paper][Bib]
[Code]
KBGAN: Adversarial Learning for Knowledge Graph EmbeddingsLiwei Cai, William Yang Wang
In Proceedings of the 16th Annual Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL-HLT 2018).
[Paper][Bib][Code]
2017
Learning to Explain Non-Standard English Words and PhrasesKe Ni, William Yang Wang
In Proceedings of the 8th International Joint Conference on Natural Language Processing (IJCNLP 2017).
[Paper][Bib][Data]
DeepPath: A Reinforcement Learning Method for Knowledge Graph ReasoningWenhan Xiong, Thien Hoang, William Yang Wang
In Proceedings of the 2017 Conference on Empirical Methods in Natural Language Processing (EMNLP 2017).
[Paper][Bib][Code][NELL-995 Dataset]
Deep Residual Learning for Weakly-Supervised Relation ExtractionYi Yao Huang, William Yang Wang
In Proceedings of the 2017 Conference on Empirical Methods in Natural Language Processing (EMNLP 2017).
[Paper][Bib][Code]
"Liar, Liar Pants on Fire": A New Benchmark Dataset for Fake News DetectionWilliam Yang Wang
In Proceedings of the 55th Annual Meeting of the Association for Computational Linguistics (ACL 2017).
[Paper][Bib][Data]
2016 and before
Learning First-Order Logic Embeddings via Matrix FactorizationWilliam Yang Wang, William W. Cohen
In Proceedings of the 25th International Joint Conference on Artificial Intelligence (IJCAI 2016).
[Paper][Bib]
A Low-Rank Approximation Approach to Learning Joint Embeddings of News Stories and Images for Timeline SummarizationWilliam Yang Wang, Yashar Mehdad, Drago Radev, Amanda Stent
In Proceedings of the 15th Annual Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL HLT 2016).
[Paper][Bib]
That's So Annoying!!!: A Lexical and Frame-Semantic Embedding Based Data Augmentation Approach to Automatic Categorization of Annoying Behaviors using #petpeeve TweetsWilliam Yang Wang, Diyi Yang
In Proceedings of the 2015 Conference on Empirical Methods in Natural Language Processing (EMNLP 2015).
[Paper][Bib][Data]
[EMNLP 2015 Notable Data Set Award: 2/1300, 0.2%]
A Soft Version of Predicate Invention Based on Structured SparsityWilliam Yang Wang, Kathryn Mazaitis, William W. Cohen
In Proceedings of the 24th International Joint Conference on Artificial Intelligence (IJCAI 2015).
[Paper][Bib][Data]
Joint Information Extraction and Reasoning: A Scalable Statistical Relational Learning ApproachWilliam Yang Wang, William W. Cohen
In Proceedings of the 53rd Annual Meeting of the Association for Computational Linguistics and the 7th International Joint Conference of the Asian Federation of Natural Language Processing (ACL-IJCNLP 2015).
[Paper][Bib][Data]
Matrix Factorization with Knowledge Graph Propagation for Unsupervised Spoken Language UnderstandingYun-Nung Chen, William Yang Wang, Anatole Gershman, Alexander Rudnicky
In Proceedings of the 53rd Annual Meeting of the Association for Computational Linguistics and The 7th International Joint Conference of the Asian Federation of Natural Language Processing (ACL-IJCNLP 2015).
[Paper][Bib]
Efficient Inference and Learning in a Large Knowledge Base: Reasoning with Extracted Information using a Locally Groundable First-Order Probabilistic LogicWilliam Yang Wang, Kathryn Mazaitis, Ni Lao, William W. Cohen
In Machine Learning Journal (MLJ 2015).
[Paper][Bib][Code]
I Can Has Cheezburger? A Nonparanormal Approach to Combining Textual and Visual Information for Predicting and Generating Popular Meme DescriptionsWilliam Yang Wang, Miaomiao Wen
In Proceedings of the 2015 Conference of the North American Chapter of the Association for Computational Linguistics – Human Language Technologies (NAACL HLT 2015).
[Paper][Bib][Data][Covered by Mental Floss]
Jointly Modeling Inter-Slot Relations by Random Walk on Knowledge Graphs for Unsupervised Spoken Language UnderstandingYun-Nung Chen, William Yang Wang, Alex Rudnicky
In Proceedings of the 2015 Conference of the North American Chapter of the Association for Computational Linguistics – Human Language Technologies (NAACL HLT 2015).
[Paper][Bib]
Learning Semantic Hierarchy with Distributed Representations for Unsupervised Spoken Language UnderstandingYun-Nung Chen, William Yang Wang, Alex Rudnicky
In Proceedings of the 16th Annual Conference of the International Speech Communication Association (INTERSPEECH 2015).
[Paper][Bib]
Structure Learning via Parameter LearningWilliam Yang Wang, Kathryn Mazaitis, William W. Cohen
In Proceedings of the 23rd ACM International Conference on Information and Knowledge Management (CIKM 2014).
[Paper][Bib]
Dependency Parsing for Weibo: An Efficient Probabilistic Logic Programming ApproachWilliam Yang Wang, Lingpeng Kong, Kathryn Mazaitis, William W. Cohen
In Proceedings of the Conference on Empirical Methods in Natural Language Processing (EMNLP 2014).
[Paper][Bib][Data]
A Semiparametric Gaussian Copula Regression Model for Predicting Financial Risks from Earnings CallsWilliam Yang Wang, Zhenhao Hua
In Proceedings of the 52nd Annual Meeting of the Association for Computational Linguistics (ACL 2014).
[Paper][Bib][Data]
ProPPR: Efficient First-Order Probabilistic Logic Programming for Structure Discovery, Parameter Learning, and Scalable InferenceWilliam Yang Wang, Kathryn Mazaitis, William W. Cohen
In Proceedings of the AAAI 2014 Workshop on Statistical Relational AI (StarAI 2014).
[Paper][Bib][Code]
Leveraging Frame Semantics and Distributional Semantics for Unsupervised Semantic Slot Induction for Spoken Dialogue SystemsYun-Nung Chen, William Yang Wang, Alex Rudnicky
In Proceedings of the 2014 IEEE Workshop Spoken Language Technology (SLT 2014).
An abstract was presented at ACL 2014 Workshop on Semantic Parsing (SP 2014).
[Paper][Bib]
Programming with Personalized PageRank: A Locally Groundable First-Order Probabilistic LogicWilliam Yang Wang, Kathryn Mazaitis, William W. Cohen
In Proceedings of the 22nd ACM International Conference on Information and Knowledge Management (CIKM 2013).
An extended abstract was presented at ICML 2013 workshop on Inferning: Interactions between Inference and Learning.
[Paper][Bib][Code][Video-ICML][Slides-CIKM]
[CIKM 2013 Best Paper Honorable Mention Award: 3/848, 0.3%]
This Text has the Scent of Starbucks: A Laplacian Structured Sparsity Model for Computational Branding AnalyticsWilliam Yang Wang, Ed Lin, John Kominek
In Proceedings of the 2013 Conference on Empirical Methods in Natural Language Processing (EMNLP 2013).
[Paper][Bib][Data]
Automatic Domain Partitioning for Multi-Domain LearningDi Wang, Chenyan Xiong, William Yang Wang
In Proceedings of the 2013 Conference on Empirical Methods in Natural Language Processing (EMNLP 2013).
[Paper][Bib]
Unsupervised Induction and Filling of Semantic Slots for Spoken Dialogue Systems using Frame-Semantic ParsingYun-Nung Chen, William Yang Wang, Alex Rudnicky
In Proceedings of the IEEE Automatic Speech Recognition and Understanding Workshop (ASRU 2013).
[Paper][Bib]
[ASRU 2013 Best Student Paper Award: 1/~170, 0.6%]
An Empirical Investigation of Sparse Log-Linear Models for Improved Dialogue Act ClassificationYun-Nung Chen, William Yang Wang, Alex Rudnicky
In Proceedings of 38th International Conference on Acoustics, Speech, and Signal Processing (ICASSP 2013).
[Paper][Bib]
Crowdsourcing the Acquisition of Natural Language Corpora: Methods and ObservationsWilliam Yang Wang, Dan Bohus, Ece Kamar, Eric Horvitz
In Proceedings of the 2012 IEEE Workshop Spoken Language Technology (SLT 2012).
[Paper][Bib]
Historical Analysis of Legal Opinions with a Sparse Mixed-Effects Latent Variable ModelWilliam Yang Wang, Elijah Mayfield, Suresh Naidu, Jeremiah Dittmar
In Proceedings of the 50th Annual Meeting of the Association for Computational Linguistics (ACL 2012).
[Paper][Bib]
"Love ya, jerkface": using Sparse Log-Linear Models to Build Positive (and Impolite) Relationships with TeensWilliam Yang Wang, Samantha Finkelstein, Amy Ogan, Alan W Black, Justine Cassell
In Proceedings of the 13th Annual SIGdial Meeting on Discourse and Dialogue (SIGDIAL 2012).
[Paper][Bib][Video]
Automatic Detection of Speaker State: Lexical, Prosodic, and Phonetic Approaches to Level-of-Interest and Intoxication ClassificationWilliam Yang Wang, Fadi Biadsy, Andrew Rosenberg, Julia Hirschberg
In the Computer Speech and Language Journal, 2012.
[Paper][Bib]
Automatic Detection of Unnatural Word-Level Segments in Unit-Selection Speech SynthesisWilliam Yang Wang, Kallirroi Georgila
In Proceedings of the IEEE Automatic Speech Recognition and Understanding Workshop (ASRU 2011).
[Paper][Bib]
Identifying Event Descriptions using Co-training with Online News SummariesWilliam Yang Wang, Kapil Thadani, Kathleen R. McKeown
In Proceedings of the 5th International Joint Conference on Natural Language Processing (IJCNLP 2011).
[Paper][Bib]
Improving Spoken Dialogue Understanding Using Phonetic Mixture ModelsWilliam Yang Wang, Ron Artstin, Anton Leuski, David Traum
In Cross-Disciplinary Advances in Applied Natural Language Processing: Issues and Approaches, IGI Global.
[An invited book chapter based on the FLAIRS 2011 paper]
Intoxication Detection using Phonetic, Phonotactic and Prosodic CuesFadi Biadsy, William Yang Wang, Andrew Rosenberg, Julia Hirschberg
In Proceedings of the 12th Annual Conference of the International Speech Communication Association (INTERSPEECH 2011).
[Paper][Bib]
Detecting Levels of Interest from Spoken Dialog with Multistream Prediction Feedback and Similarity Based Hierarchical Fusion LearningWilliam Yang Wang, Julia Hirschberg
In Proceedings of the 12th Annual SIGdial Meeting on Discourse and Dialogue (SIGDIAL 2011).
[Paper][Bib]
Improving Spoken Dialogue Understanding Using Phonetic Mixture ModelsWilliam Yang Wang, Ron Artstein, Anton Leuski, David Traum
In Proceedings of the 24th International Florida Artificial Intelligence Research Society Conference (FLAIRS-24).
[Paper][Bib]
[FLAIRS 2011 Best Paper Finalist: 6/179, 3%]
Got You!: Automatic Vandalism Detection in Wikipedia with Web-based Syntactic-Semantic ModelingWilliam Yang Wang, Kathleen McKeown
In Proceedings of the 23rd International Conference on Computational Linguistics (COLING 2010).
[Paper][Bib]