Majority of work done or significant contribution by MAIL/SAIL members!
Submission History shows the venues where the work has been submitted (🙃 including rejections 🙃). I hope some of my poor rejection/failure histories (record now is 10 rejections 😅) give you some encouragement to try again when things don't work out (don't give up -- good work doesn't need to be rushed)!
publications by categories in reversed chronological order. An up-to-date list is available on Google Scholar.

  1. PREPRINT PRISM: A Multi-Dimensional Benchmark for Evaluating LLM Peer Reviewers
    Ngoc PP Loc, Toan LV Huynh, Thanh Tran Khanh, Duy A Nguyen, Tuan Anh Nguyen Pham, Thanh Nguyen, Nitesh V Chawla, Wray Buntine, Kok-Seng Wong, Khoa D Doan, and Binh T Nguyen
    arxiv preprint 2026
  2. Interspeech ViP-VL: Vietnamese Self-supervised Speech Pretraining Model with Vector-Quantization Learning
    Duy K Le, Kiet Anh Hoang, Bao Nguyen, Duy Vo, Dung Vo, Thai Tran, Linh Pham, and Khoa D Doan
    In Interspeech 2026
  3. PREPRINT Pop-Up Distractions Reveal Bag-of-Events Behavior in Video Large Language Models
    Oscar Chew, Serhii Honcharenko, Qian-Hui Chen, Patricia Lu, Dishant Zaveri, Khoa D Doan, and Kuan-Hao Huang
    arxiv preprint 2026
  4. PREPRINT SparseSAM: Structured Sparsification of Activations in Segment Anything Models
    Chau H Tran, Chi H Nguyen, Duy MH Nguyen, Mathias Niepert, Fan Lai, and Khoa D Doan
    arxiv preprint 2026
  5. PREPRINT Decoding the Critique Mechanism in Large Reasoning Models
    Phan Hoang, Quang H Nguyen, Hung T. Q. Le, Xiusi Chen, Heng Ji, and Khoa D Doan
    arxiv preprint 2026
  6. PREPRINT The Reasoning Boundary Paradox: How Reinforcement Learning Constrains Language Models
    Phuc M Nguyen, Chinh D La, Duy MH Nguyen, Nitesh V Chawla, Binh T Nguyen, and Khoa D Doan
    arxiv preprint 2026
  7. COLM Why Do Reasoning Models Lose Coverage? The Role of Data and Forks in the Road
    Ngoc-Hieu Nguyen, Parshin Shojaee, Phuc M Nguyen, Nan Zhang, Chandan K Reddy, Khoa D Doan, and Rui Zhang
    In Conference on Language Modeling 2026
  8. TMLR FEATURED Retrospective Feature Estimation for Continual Learning
    Nghia D Nguyen, Hieu T Nguyen, Ang Li, Hoang V Pham, Viet Anh Nguyen, and Khoa D Doan
    Transactions on Machine Learning Research 2026
  9. EACL SpARK: An Embarrassingly Simple Sparse Watermarking in LLMs with Enhanced Text Quality
    Cao-Duy Hoang, Hung T. Q. Le, Rui Chu, Ping Li, Weijie Zhao, Yingjie Lao, and Khoa D Doan
    In Findings of the European Chapter of the Association for Computational Linguistics 2026
  10. AAAI Clean-Label Physical Backdoor Attacks with Data Distillation
    Thinh Dao, Khoa D Doan, and Kok-Seng Wong
    In AAAI Conference on Artificial Intelligence 2026
  11. PREPRINT Rethinking Molecular Graph Backdoors under Chemistry-aware Admission
    Thinh T. H. Nguyen, Sze Jue Yang, Khoa D Doan, Chee Seng Chan, and Kok-Seng Wong
    arxiv preprint 2026
  12. PREPRINT Look Before You Zoom: Adaptive Routing for the Resolution-Context Trade-off in Visual RAG
    Oanh Tran, Hung T. Q. Le, Oscar Chew, Kuan-Hao Huang, and Khoa D Doan
    arxiv preprint 2026
  13. ICML FOCA: Future-Oriented Conditioning for Data-Efficient Vision-Language-Action Adaptation
    Duc Minh Nguyen, Nghiem Tuong Diep, Binh Gia Nguyen, Trong-Bao Ho, Doanh Le, Tan Q. Nguyen, Thien-Loc Ha, Nhiem Tran, Bao Thach, Nhat X. Tran, Tuan A. Tran, Artur Habuda, Philip Lund Møller, Tran Nguyen Le, Daniel Sonntag, Matthias Niepert, Khoa D Doan, Vu Duong, Hung Ngo, Minh N. Vu, Duy MH Nguyen, An Thai Le, and Ngo Anh Vien
    In International Conference on Machine Learning 2026
  14. PREPRINT When Generator Replay Degrades: Projected Rehearsal Orchestration for Heterogeneous Federated Class-Incremental Learning
    Thinh T. H. Nguyen, Khoa D Doan, Binh T Nguyen, Danh Le-Phuoc, and Kok-Seng Wong
    arxiv preprint 2026
  15. PREPRINT Self-Improving VLA Policies: Selected Diffusion Noise for Spurious-Robust Action Smoothing
    Duc Minh Nguyen, Bao-Ngoc Dao, Tung M. Luu, Binh Gia Nguyen, Vinh Tong, Anji Liu, Vu N. Duong, Dung D. Le, Daniel Sonntag, Trung Le, Ngan Le, Jan Peter, An Thai Le, Minh Nhat Vu, Mathias Niepert, Khoa D Doan, Duy MH Nguyen, and Vien Anh Ngo
    arxiv preprint 2026
  16. PREPRINT DebiasRAG: A Tuning-Free Path to Fair Generation in Large Language Models through Retrieval-Augmented Generation
    Rui Chu, Bingyin Zhao, Hung T. Q. Le, Cao-Duy Hoang, Huawei Lin, Ping Li, Weijie Zhao, Khoa D Doan, and Yingjie Lao
    arxiv preprint 2026
  17. COLM Is CLIP Cross-Eyed? Revealing and Mitigating Center Bias in the CLIP Family
    Oscar Chew, Hsiao-Ying Huang, Kunal Jain, Tai-I Chen, Khoa D Doan, and Kuan-Hao Huang
    In Conference on Language Modeling 2026
  1. NeurIPS How Many Tokens Do 3D Point Cloud Transformer Architectures Really Need?
    Tuan A Tran, Duy MH Nguyen, Chau H Tran, and others
    In Advances in Neural Information Processing Systems 2025
  2. NeurIPS Mitigating Reward Over-optimization in Direct Alignment Algorithms with Adaptive Importance Sampling
    Phuc M Nguyen, Ngoc-Hieu Nguyen, Binh T Nguyen, and Khoa D Doan
    In Advances in Neural Information Processing Systems 2025
  3. NeurIPS Unveiling Concept Attribution in Diffusion Models
    Quang H Nguyen, Phan Hoang, and Khoa D Doan
    In Advances in Neural Information Processing Systems 2025
  4. ICML ORAL LLM-SRBench: A New Benchmark for Scientific Equation Discovery with Large Language Models
    Parshin Shojaee, Ngoc-Hieu Nguyen, Kazem Meidani, Amir Barati Farimani, Khoa D Doan, and Chandan K Reddy
    In International Conference on Machine Learning 2025
  5. ICLR Wicked Oddities: Selectively Poisoning for Effective Clean-Label Backdoor Attacks
    Quang H Nguyen, Ngoc-Hieu Nguyen, The-Anh Ta, Thanh Nguyen-Tang, Kok-Seng Wong, Hoang Thanh-Tung, and Khoa D Doan
    In The Twelfth International Conference on Learning Representations 2025
  6. PREPRINT BackFed: An Efficient & Standardized Benchmark Suite for Backdoor Attacks in Federated Learning
    Thinh Dao, Dung Thuy Nguyen, Khoa D Doan, and Kok-Seng Wong
    arxiv preprint 2025
  1. CODS-COMAD Class-Aware Contrastive Optimization for Imbalanced Text Classification
    Grigorii Khvatskii, Nuno Moniz, Khoa D Doan, and Nitesh V Chawla
    In 8th International Conference on Data Science and Management of Data 2024
  2. ECCV ORAL Flatness-aware Sequential Learning Generates Resilient Backdoors
    Hoang V Pham, The-Anh Ta, Anh Tran, and Khoa D Doan
    In European Conference on Computer Vision 2024
  3. ECCV Data Poisoning Quantization Backdoor Attack
    Tran Huynh, Anh Tran, Khoa D Doan, and Tung Pham
    In European Conference on Computer Vision 2024
  4. ICPR Composite Concept Extraction through Backdooring
    Banibrata Ghosh, Haripriya Harikumar, Khoa D Doan, Svetha Venkatesh, and Santu Rana
    In 27th International Conference on Pattern Recognition 2024
  5. PREPRINT Non-Cooperative Backdoor Attacks in Federated Learning: A New Threat Landscape
    Tuan M Nguyen, Dung T Nguyen, Khoa D Doan, and Kok-Seng Wong
    2024
  6. PREPRINT MetaLLM: A High-performant and Cost-efficient Dynamic Framework for Wrapping LLMs
    Quang H Nguyen, Cao-Duy Hoang, Juliette Decugis, Saurav Manchanda, Nitesh V Chawla, and Khoa D Doan
    2024
  7. PREPRINT Overcoming Catastrophic Forgetting in Federated Class-Incremental Learning via Federated Global Twin Generator
    Thinh Nguyen, Khoa D Doan, Binh T Nguyen, Danh Le-Phuoc, and Kok-Seng Wong
    arxiv preprint 2024
  8. PREPRINT Venomancer: Towards Imperceptible and Target-on-Demand Backdoor Attacks in Federated Learning
    Son Nguyen, Thinh Nguyen, Khoa D Doan, and Kok-Seng Wong
    arxiv preprint 2024
  9. PREPRINT Synthesizing Physical Backdoor Datasets: An Automated Framework Leveraging Deep Generative Models
    Sze Jue Yang, Chinh D La, Quang H Nguyen, Eugene Bagdasaryan, Kok-Seng Wong, Anh T Tran, Chee Seng Chan, and Khoa D Doan
    2024
  10. PREPRINT Forget-Me-Not: Making Backdoor Hard to be Forgotten in Fine-tuning
    Tran Ngoc Huynh, Anh T Tran, Khoa D Doan, and Tung Pham
    2024
  11. ACL-Findings Fooling the Textual Fooler via Randomizing Latent Representations
    Cao-Duy Hoang, Quang H Nguyen, Saurav Manchanda, Minlong Peng, Kok-Seng Wong, and Khoa D Doan
    In Findings of the Association for Computational Linguistics 2024
  12. PREPRINT Everyone Can Attack: Repurpose Lossy Compression as a Natural Backdoor Attack
    Sze Jue Yang, Quang H Nguyen, Chee Seng Chan, and Khoa D Doan
    arXiv preprint arXiv:2308.16684 2024
  13. PREPRINT CoopHash: Cooperative Learning of Multipurpose Descriptor and Contrastive Pair Generator via Variational MCMC Teaching for Supervised Image Hashing
    Khoa D Doan, Jianwen Xie, Yaxuan Zhu, Yang Zhao, and Ping Li
    2024
  14. ICLR Understanding the Robustness of Randomized Feature Defense Against Query-Based Adversarial Attacks
    Quang H Nguyen, Yingjie Lao, Tung Pham, Kok-Seng Wong, and Khoa D Doan
    In The Twelfth International Conference on Learning Representations 2024
  15. NeurIPS Iba: Towards irreversible backdoor attacks in federated learning
    Thuy Dung Nguyen, Tuan M Nguyen, Anh T Tran, Khoa D Doan, and Kok-Seng Wong
    Advances in Neural Information Processing Systems 2024
  16. EAAI Backdoor attacks and defenses in federated learning: Survey, challenges and future research directions
    Thuy Dung Nguyen, Tuan M Nguyen, Phi Le Nguyen, Hieu H Pham, Khoa D Doan, and Kok-Seng Wong
    Engineering Applications of Artificial Intelligence 2024
  1. UAI Cold-start Recommendation by Personalized Embedding Region Elicitation
    Hieu Trung Nguyen, Duy Nguyen, Khoa D Doan, and Viet Anh Nguyen
    In The Conference on Uncertainty in Artificial Intelligence 2023
  2. NeurIPS-W Clean-label Backdoor Attacks by Selectively Poisoning with Limited Information from Target Class
    Quang H Nguyen, Ngoc-Hieu Nguyen, The-Anh Ta, Thanh T Nguyen, Thanh-Tung Hoang, and Khoa D Doan
    In NeurIPS 2023 Workshop on Backdoors in Deep Learning-The Good, the Bad, and the Ugly 2023
  3. ACML Empirical Study of Federated Unlearning: Efficiency and Effectiveness
    Thai-Hung Nguyen, Hong-Phuc Vu, Dung Thuy Nguyen, Tuan Minh Nguyen, Khoa D Doan, and Kok-Seng Wong
    In Asian Conference on Machine Learning 2023
  4. SIGIR Asymmetric Hashing for Fast Ranking via Neural Network Measures
    Khoa D Doan, Shulong Tan, Weijie Zhao, and Ping Li
    In 46th International ACM SIGIR Conference on Research and Development in Information Retrieval 2023
  5. ICML-W A Cosine Similarity-based Method for Out-of-Distribution Detection
    Ngoc-Hieu Nguyen, Quang H Nguyen, Thanh T Nguyen, Khoa D Doan, and Thanh-Tung Hoang
    In ICML 2023 Workshop on Spurious Correlations, Invariance, and Stability 2023
  6. AAAI Defending backdoor attacks on vision transformer via patch processing
    Khoa D Doan, Yingjie Lao, and Ping Li
    In AAAI Conference on Artificial Intelligence 2023
  1. ACCV Unified Learning of Multipurpose Energy Based Generative Hashing Network
    Khoa D Doan, and Chandan K Reddy
    In Sixteenth Asian Conference on Computer Vision 2022
  2. NeurIPS Marksman Backdoor: Backdoor Attacks with Arbitrary Target Class
    Khoa D Doan, Yingjie Lao, and Ping Li
    In Thirty-Sixth Conference on Neural Information Processing Systems 2022
  3. CVPR One Loss for Quantization: Deep Hashing with Discrete Wasserstein Distributional Matching
    Khoa D Doan, Peng Yang, and Ping Li
    In Conference on Computer Vision and Pattern Recognition 2022
  1. NeurIPS Backdoor Attack with Imperceptible Input and Latent Modification
    Khoa D Doan, Yingjie Lao, and Ping Li
    In Thirty-Fifth Conference on Neural Information Processing Systems 2021
  2. ICCV LIRA: Learnable, Imperceptible and Robust Backdoor Attacks
    Khoa D Doan, Yingjie Lao, Weijie Zhao, and Ping Li
    In International Conference on Computer Vision 2021
  3. SIGIR Interpretable Graph Similarity Computation via Differentiable Optimal Alignment of Node Embeddings
    Khoa D Doan, Saurav Manchanda, Suchismit Mahapatra, and Chandan K Reddy
    In 44th International ACM SIGIR Conference on Research and Development in Information Retrieval 2021
  1. WWW Efficient Implicit Unsupervised Text Hashing Using Adversarial Autoencoder
    Khoa D Doan, and Chandan K Reddy
    In Proceedings of The Web Conference 2020
  2. arXiv Image Generation Via Minimizing Fréchet Distance in Discriminator Feature Space
    arXiv preprint arXiv:2003.11774 2020
  3. arXiv Image Hashing by Minimizing Discrete Component-wise Wasserstein Distance
    arXiv preprint arXiv:2003.00134 2020
  4. arXiv Regression via implicit models and optimal transport cost minimization
    Saurav Manchanda, Khoa D Doan, Pranjul Yadav, and Sathiya K Selvaraj
    arXiv preprint arXiv:2003.01296 2020
  5. arXiv Gradient boosting neural networks: Grownet
    arXiv preprint arXiv:2002.07971 2020
  1. BigData Targeted display advertising: the case of preferential attachment
    Saurav Manchanda, Pranjul Yadav, Khoa D Doan, and Sathiya K Selvaraj
    In Proceedings of the 2019 IEEE International Conference on Big Data 2019
  2. CIKM Adversarial Factorization Autoencoder for Look-Alike Modeling
    Khoa D Doan, Pranjul Yadav, and Chandan K Reddy
    In Proceedings of the 28th ACM International Conference on Information and Knowledge Management 2019
  3. PAKDD An Attentive Spatio-Temporal Neural Model for Successive Point of Interest Recommendation.
    Khoa D Doan, Guolei Yang, and Chandan K Reddy
    In Proceedings of the 2019 Pacific-Asia Conference on Knowledge Discovery and Data Mining (PAKDD) 2019
  1. BigData Quest for Value in Big Earth Data
    Kwo-Sen Kuo, Amidu O Oloso, Mike L Rilee, Khoa D Doan, Thomas L Clune, and Hongfeng Yu
    In EGU General Assembly Conference Abstracts 2017
  1. BigData Evaluating the impact of data placement to spark and SciDB with an Earth Science use case
    Khoa D Doan, Amidu O Oloso, Kwo-Sen Kuo, Thomas L Clune, Hongfeng Yu, Brian Nelson, and Jian Zhang
    In Proceedings of the 2016 IEEE International Conference on Big Data 2016
  2. IGARSS Implications of data placement strategy to Big Data technologies based on shared-nothing architecture for geosciences
    Kwo-Sen Kuo, Amidu Oloso, Khoa D Doan, Thomas L Clune, and Hongfeng Yu
    In 2016 IEEE International Geoscience and Remote Sensing Symposium (IGARSS) 2016
  1. AGU SciDB versus Spark: A preliminary comparison based on an Earth science use case
    Thomas Clune, Kwo-Sen Kuo, Khoa D Doan, and Amidu Oloso
    In AGU Fall Meeting Abstracts 2015
  1. BigData Performance comparison of big-data technologies in locating intersections in satellite ground tracks
    Khoa D Doan, Amidu Oloso, Kwo-Sen Kuo, Thomas L Clune, and LLC Bayesics
    In Proceedings of the 2014 ASE BigData/SocialInformatics/PASSAT/BioMedCom Conference 2014
  1. AGU A Demonstration of Big Data Technology for Data Intensive Earth Science
    K Kuo, T Clune, R Ramachandran, J Rushing, G Fekete, A Lin, KD Doan, AO Oloso, and D Duffy
    In AGU Fall Meeting Abstracts 2013
The brick walls are there for a reason. The brick walls are not there to keep us out. The brick walls are there to give us a chance to show how badly we want something -- Randy Pausch