Scheduled Maintenance Notice
Please note that Researcher Profiles will be undergoing scheduled maintenance on Wednesday 7th Oct, from 8:00am to 9:00am. During this time, the Researcher Profiles system will be unavailable. We apologise for any inconvenience and appreciate your understanding.
Select Publications
Conference Papers
, 2027, 'Uncovering Logit Suppression Vulnerabilities in LLM Safety Alignment', in Lecture Notes in Computer Science, pp. 61 - 76, http://dx.doi.org/10.1007/978-3-032-31583-0_5
, 2026, 'NgCaptcha: A CAPTCHA Bridging the Past and the Future', in Www Companion 2026 Companion Proceedings of the ACM Web Conference 2026, pp. 104 - 107, http://dx.doi.org/10.1145/3774905.3793112
, 2026, 'LifeFuzz: Lifecycle-Guided Fuzzing for Windows Driver Cross-Handler Vulnerabilities', in Eurosys 2026 Proceedings of the 2026 European Conference on Computer Systems, pp. 2022 - 2036, http://dx.doi.org/10.1145/3767295.3803624
, 2026, 'SpecGuru: Hierarchical LLM-Driven API Points-to Specification Generation with Self-Validation', Association for Computing Machinery (ACM), pp. 1200 - 1212, presented at Proceedings of the 2026 IEEE/ACM 48th International Conference on Software Engineering, http://dx.doi.org/10.1145/3744916.3773209
, 2026, 'Diverse Claire: An AI-Powered UDL Approach Bridging Educational Experience Gaps in CS1', Association for Computing Machinery (ACM), pp. 1714 - 1714, presented at Proceedings of the 57th ACM Technical Symposium on Computer Science Education V.2, http://dx.doi.org/10.1145/3770761.3777125
, 2026, 'CSTutorBench: Benchmarking Large Language Models for Realistic Computer Science Tutoring', in SIGCSE TS 2026 Proceedings of the 57th ACM Technical Symposium on Computer Science Education V 2, pp. 1263 - 1264, http://dx.doi.org/10.1145/3770761.3777333
, 2026, 'DiverseClaire: Simulating Students to Improve Introductory Programming Course Materials for All CS1 Learners', in SIGCSE TS 2026 Proceedings of the 57th ACM Technical Symposium on Computer Science Education V 2, pp. 1585 - 1586, http://dx.doi.org/10.1145/3770761.3777340
, 2026, 'Exploring Trust in Human-LLM Feedback Systems: Observation of Student Behaviour in Software Engineering Education', in SIGCSE TS 2026 Proceedings of the 57th ACM Technical Symposium on Computer Science Education V 2, pp. 1419 - 1420, http://dx.doi.org/10.1145/3770761.3777342
, 2026, 'A Multi-Agent System for Inclusive Automatic Speech Recognition for People Who Stutter', in Proceedings of the Aaai Conference on Artificial Intelligence, pp. 39495 - 39503, http://dx.doi.org/10.1609/aaai.v40i46.41300
, 2026, 'An Ontology-Driven Service-Oriented System for ESG Metric Computation and Reporting', in Proceedings of the IEEE International Conference on Web Services Icws, pp. 347 - 353, http://dx.doi.org/10.1109/ICWS72778.2026.00052
, 2026, 'PufferDoS: Efficient and Effective Attack String Generation for Regular Expression Denial of Service Vulnerabilities', in Proceedings IEEE Symposium on Security and Privacy, pp. 3984 - 4002, http://dx.doi.org/10.1109/SP63933.2026.00169
, 2026, 'SceneJailEval: A Scenario-Adaptive Multi-Dimensional Framework for Jailbreak Evaluation', in Proceedings of the Aaai Conference on Artificial Intelligence, pp. 35553 - 35561, http://dx.doi.org/10.1609/aaai.v40i42.40866
, 2026, 'Socrates or Smartypants: Testing Logic Reasoning Capabilities of Large Language Models with Logic Programming-Based Test Oracles', in Proceedings of the Aaai Conference on Artificial Intelligence, pp. 19433 - 19440, http://dx.doi.org/10.1609/aaai.v40i23.39021
, 2025, 'IllusionCAPTCHA: A CAPTCHA based on Visual Illusion', in Www 2025 Proceedings of the ACM Web Conference, pp. 3683 - 3691, http://dx.doi.org/10.1145/3696410.3714726
, 2025, 'A Large Scale Study of AI-based Binary Function Similarity Detection Techniques for Security Researchers and Practitioners', in Proceedings 2025 40th IEEE ACM International Conference on Automated Software Engineering Ase 2025, pp. 1070 - 1082, http://dx.doi.org/10.1109/ASE63991.2025.00093
, 2025, 'A Methodology for Replicating Historical Exploits on EVM-Compatible Blockchains', in Proceedings 2025 IEEE ACM 7th International Workshop on Emerging Trends in Software Engineering for Blockchain Wetseb 2025, pp. 57 - 60, http://dx.doi.org/10.1109/WETSEB66605.2025.00014
, 2025, 'A Rusty Link in the AI Supply Chain: Detecting Evil Configurations in Model Repositories', in Proceedings 46th IEEE Symposium on Security and Privacy Workshops Spw 2025, pp. 260 - 264, http://dx.doi.org/10.1109/SPW67851.2025.00036
, 2025, 'Continuous Embedding Attacks via Clipped Inputs in Jailbreaking Large Language Models', in Proceedings 46th IEEE Symposium on Security and Privacy Workshops Spw 2025, pp. 270 - 277, http://dx.doi.org/10.1109/SPW67851.2025.00038
, 2025, 'From Constraints to Cracks: Constraint Semantic Inconsistencies as Vulnerability Beacons for Embedded Systems', in Proceedings of the 34th Usenix Security Symposium, pp. 685 - 704
, 2025, 'Good News for Script Kiddies? Evaluating Large Language Models for Automated Exploit Generation', in Proceedings 46th IEEE Symposium on Security and Privacy Workshops Spw 2025, pp. 278 - 282, http://dx.doi.org/10.1109/SPW67851.2025.00039
, 2025, 'It Only Gets Worse: Revisiting DL-Based Vulnerability Detectors from a Practical Perspective', in Proceedings Asia Pacific Software Engineering Conference APSEC, pp. 947 - 956, http://dx.doi.org/10.1109/APSEC66846.2025.00109
, 2025, 'KRAKEN-FUZZ: Minimizing Corpus During Ensemble Fuzzing', in Proceedings 2025 IEEE ACM International Workshop on Search Based and Fuzz Testing Sbft 2025, pp. 47 - 48, http://dx.doi.org/10.1109/SBFT66712.2025.00017
, 2025, 'MMLU-ProX: A Multilingual Benchmark for Advanced Large Language Model Evaluation', in Emnlp 2025 2025 Conference on Empirical Methods in Natural Language Processing Proceedings of the Conference, pp. 1513 - 1532, http://dx.doi.org/10.18653/v1/2025.emnlp-main.79
, 2025, 'Source Code Summarization in the Era of Large Language Models', in Proceedings International Conference on Software Engineering, pp. 1882 - 1894, http://dx.doi.org/10.1109/ICSE55347.2025.00034
, 2025, 'TOMBRAIDER: Entering the Vault of History to Jailbreak Large Language Models', in Emnlp 2025 2025 Conference on Empirical Methods in Natural Language Processing Proceedings of the Conference, pp. 5478 - 5493, http://dx.doi.org/10.18653/v1/2025.emnlp-main.279
, 2025, 'TransferFuzz: Fuzzing with Historical Trace for Verifying Propagated Vulnerability Code', in Proceedings International Conference on Software Engineering, pp. 268 - 280, http://dx.doi.org/10.1109/ICSE55347.2025.00061
, 2025, 'Truman: A Large Language Model-based Multi-agent Simulator for Synthetic Money Laundering Data Generation', in Proceedings of the International Joint Conference on Autonomous Agents and Multiagent Systems Aamas, pp. 2594 - 2596
, 2024, 'Demystifying RCE Vulnerabilities in LLM-Integrated Apps', in Ccs 2024 Proceedings of the 2024 ACM Sigsac Conference on Computer and Communications Security, pp. 1716 - 1730, http://dx.doi.org/10.1145/3658644.3690338
, 2024, 'Rust-twins: Automatic Rust Compiler Testing through Program Mutation and Dual Macros Generation', in Proceedings 2024 39th ACM IEEE International Conference on Automated Software Engineering Ase 2024, pp. 631 - 642, http://dx.doi.org/10.1145/3691620.3695059
, 2024, 'Bugs in Pods: Understanding Bugs in Container Runtime Systems', in Issta 2024 Proceedings of the 33rd ACM SIGSOFT International Symposium on Software Testing and Analysis, pp. 1364 - 1376, http://dx.doi.org/10.1145/3650212.3680366
, 2024, 'How Effective Are They? Exploring Large Language Model Based Fuzz Driver Generation', in Issta 2024 Proceedings of the 33rd ACM SIGSOFT International Symposium on Software Testing and Analysis, pp. 1223 - 1235, http://dx.doi.org/10.1145/3650212.3680355
, 2024, 'A Comprehensive Study of Jailbreak Attack versus Defense for Large Language Models', in Martins A; Srikumar V; Ku LW (eds.), FINDINGS OF THE ASSOCIATION FOR COMPUTATIONAL LINGUISTICS: ACL 2024, ASSOC COMPUTATIONAL LINGUISTICS-ACL, THAILAND, Bangkok, pp. 7432 - 7449, presented at 62nd Annual Meeting of the Association-for-Computational-Linguistics (ACL) / Student Research Workshop (SRW), THAILAND, Bangkok, 11 August 2024 - 16 August 2024
, 2024, 'A Hitchhiker’s Guide to Jailbreaking ChatGPT via Prompt Engineering', in Sea4dq 2024 Proceedings of the 4th International Workshop on Software Engineering and AI for Data Quality in Cyber Physical Systems Internet of Things Co Located with Esec Fse 2024, pp. 12 - 21, http://dx.doi.org/10.1145/3663530.3665021
, 2024, 'Medusa: Unveil Memory Exhaustion DoS Vulnerabilities in Protocol Implementations', in Www 2024 Proceedings of the ACM Web Conference, pp. 1668 - 1679, http://dx.doi.org/10.1145/3589334.3645476
, 2024, 'MeTMaP: Metamorphic Testing for Detecting False Vector Matching Problems in LLM Augmented Generation', in Proceedings 2024 IEEE ACM 1st International Conference on AI Foundation Models and Software Engineering Forge 2024, pp. 12 - 23, http://dx.doi.org/10.1145/3650105.3652297
, 2024, 'Leveraging Semantic Relations in Code and Data to Enhance Taint Analysis of Embedded Systems', in Proceedings of the 33rd Usenix Security Symposium, pp. 7067 - 7084
, 2024, 'MASTERKEY: Automated Jailbreaking of Large Language Model Chatbots', Internet Society, presented at Proceedings 2024 Network and Distributed System Security Symposium, http://dx.doi.org/10.14722/ndss.2024.24188
, 2024, 'PENTESTGPT: Evaluating and Harnessing Large Language Models for Automated Penetration Testing', in Proceedings of the 33rd Usenix Security Symposium, pp. 847 - 864
, 2023, 'Monitoring Automotive Software Security Health through Trustworthiness Score', in Proceedings Cscs 2023 7th ACM Computer Science in Cars Symposium, http://dx.doi.org/10.1145/3631204.3631859
, 2023, 'ACETest: Automated Constraint Extraction for Testing Deep Learning Operators', in Issta 2023 Proceedings of the 32nd ACM SIGSOFT International Symposium on Software Testing and Analysis, pp. 690 - 702, http://dx.doi.org/10.1145/3597926.3598088
, 2023, 'ASTER: Automatic Speech Recognition System Accessibility Testing for Stutterers', in Proceedings 2023 38th IEEE ACM International Conference on Automated Software Engineering Ase 2023, pp. 510 - 521, http://dx.doi.org/10.1109/ASE56229.2023.00107
, 2023, 'HasteFuzz: Full-Speed Fuzzing', in Proceedings 2023 IEEE ACM International Workshop on Search Based and Fuzz Testing Sbft 2023, pp. 73 - 75, http://dx.doi.org/10.1109/SBFT59156.2023.00022
, 2023, 'NAUTILUS: Automated RESTful API Vulnerability Detection', in 32nd Usenix Security Symposium Usenix Security 2023, pp. 5593 - 5610
, 2023, 'PumpChannel: An Efficient and Secure Communication Channel for Trusted Execution Environment on ARM-FPGA Embedded SoC', in Proceedings Design Automation and Test in Europe Date, http://dx.doi.org/10.23919/DATE56975.2023.10137170
, 2023, 'RSFuzzer: Discovering Deep SMI Handler Vulnerabilities in UEFI Firmware with Hybrid Fuzzing', in Proceedings IEEE Symposium on Security and Privacy, pp. 2155 - 2169, http://dx.doi.org/10.1109/SP46215.2023.10179421
, 2022, 'More Secure Collaborative APIs resistant to Flush-Based Cache Attacks on Cortex-A9 Based Automotive System', in Proceedings Cscs 2022 6th ACM Computer Science in Cars Symposium, http://dx.doi.org/10.1145/3568160.3570227
, 2022, 'Morest: Industry Practice of Automatic RESTful API Testing', in ACM International Conference Proceeding Series, http://dx.doi.org/10.1145/3551349.3559498
, 2022, 'Efficient greybox fuzzing of applications in Linux-based IoT devices via enhanced user-mode emulation', in Issta 2022 Proceedings of the 31st ACM SIGSOFT International Symposium on Software Testing and Analysis, pp. 417 - 428, http://dx.doi.org/10.1145/3533767.3534414
, 2022, 'Morest: Model-based RESTful API Testing with Execution Feedback', in Proceedings International Conference on Software Engineering, pp. 1406 - 1417, http://dx.doi.org/10.1145/3510003.3510133
, 2022, 'Windranger: A Directed Greybox Fuzzer driven by Deviation Basic Blocks', in Proceedings International Conference on Software Engineering, pp. 2440 - 2451, http://dx.doi.org/10.1145/3510003.3510197