Bowen Wei

prof_pic.jpg

Hello! My name is Bowen Wei, and I am a third-year Ph.D. student in Computer Science at George Mason University. I am fortunate to be advised by Professor Ziwei Zhu.

My research spans trustworthy and interpretable AI and agentic reinforcement learning (RL) for large language models. I develop prototypebased, symbolic, and explanation-driven methods to make model behavior transparent, and robust, enabling users to understand and trust AI decisions in high-stakes settings. In parallel, I study RL and post-training techniques that distill multi-agent reasoning into single, verifiable agentsβ€”improving reasoning quality, evidence attribution, and causal grounding. Together, these directions aim to advance AI systems that are both interpretable and competent in reasoning.

News

Aug 24, 2026 πŸŽ‰ Our paper β€œCOSE: Confidence-Orchestrated Self-Evolution for Effective LLM Reasoning” has been accepted to the Main Conference at EMNLP 2026!
Apr 14, 2026 πŸŽ‰ Two papers accepted to ACL 2026! β€œVIGNETTE: Socially Grounded Bias Evaluation for Vision-Language Models” as an oral in the Main Conference, and β€œContext-Aware Decoding for Faithful Vision-Language Generation” in Findings.
Nov 08, 2025 πŸŽ‰ Our paper β€œMaking Sense of LLM Decisions: A Prototype-based Framework for Explainable Classification” has been accepted for an oral presentation at AAAI 2026!
May 15, 2025 πŸŽ‰ Our paper β€œProtoLens: Advancing Prototype Learning for Fine-Grained Interpretability in Text Classification” has been accepted to the main conference at ACL 2025!

Selected Publications

  1. EMNLP 2026
    Main
    COSE: Confidence-Orchestrated Self-Evolution for Effective LLM Reasoning
    Bowen Wei, Nan Wang, Yuqing Zhou, and 2 more authors
    In Proceedings of the 2026 Conference on Empirical Methods in Natural Language Processing, 2026
    Main Conference acceptance rate: 15.4% (2,719 / 17,669).
  2. AAAI 2026
    Oral
    Making Sense of LLM Decisions: A Prototype-based Framework for Explainable Classification
    Bowen Wei, Mehrdad Fazli, and Ziwei Zhu
    In Proceedings of the AAAI Conference on Artificial Intelligence, 2026
    Overall acceptance rate: 17.6% (4,167 / 23,680); oral presentation rate: 4.5% of submissions (1,058 / 23,680).
  3. NeurIPS LAW 2025
    CORTEX: Collaborative LLM Agents for High-Stakes Alert Triage
    Bowen Wei, Yuan Shen Tay, Howard Liu, and 4 more authors
    2025
  4. ACL 2025
    Main
    ProtoLens: Advancing Prototype Learning for Fine-Grained Interpretability in Text Classification
    Bowen Wei and Ziwei Zhu
    In Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), Jul 2025
    Main Conference acceptance rate: 20.3% (1,699 / 8,360).
  5. arXiv
    Preprint
    A Logical-Rule Autoencoder for Interpretable Recommendations
    Bowen Wei*, Jinhao Pan*, Yuqing Zhou, and 1 more author
    Jul 2026
    * Equal contribution.
  6. arXiv
    Preprint
    ClawSafety: "Safe" LLMs, Unsafe Agents
    Bowen Wei, Yunbei Zhang, Jinhao Pan, and 5 more authors
    Jul 2026
  7. arXiv
    Preprint
    Neural Symbolic Logical Rule Learner for Interpretable Learning
    Bowen Wei and Ziwei Zhu
    Jul 2024
  8. ICML 2026
    Main
    Knowing Bias, Doing Better: Mitigating Social Bias in LLMs via Know-Bias Neuron Enhancement
    Jinhao Pan, Chahat Raj, Anjishnu Mukherjee, and 4 more authors
    Jul 2026
  9. ACL 2026
    Oral
    VIGNETTE: Socially Grounded Bias Evaluation for Vision-Language Models
    Chahat Raj, Bowen Wei, Aylin Caliskan, and 2 more authors
    In Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics, Jul 2026
    Selected as an SAC Highlight at ACL 2026.
  10. ACL 2026
    Findings
    Context-Aware Decoding for Faithful Vision-Language Generation
    Mehrdad Fazli, Bowen Wei, and Ziwei Zhu
    In Findings of the Association for Computational Linguistics: ACL 2026, Jul 2026
  11. WACV 2026
    Main
    Mitigating Hallucination in Large Vision-Language Models via Adaptive Attention Calibration
    Mehrdad Fazli, Bowen Wei, Ahmet Sari, and 1 more author
    Jul 2025