Select Publications

Preprints

Xie X; Fan G; Lin X; Zhou A; Li S; Zheng X; Liang Y; Zhang Y; Yu N; Li H; Chen X; Chen Y; Zhen Y; Dong D; Fu X; Su J; Pan F; Luo P; Feng Y; Hu R; Guo H; Fan J; Xiao X; Di P, 2026, Principles and Practices of Large-Scale Code Analysis at Ant Group: A Data- and Logic-Oriented Approach, http://dx.doi.org/10.48550/arxiv.2401.01571

Guo Y; Liu P; Ma W; Deng Z; Zhu X; Di P; Xiao X; Wen S, 2026, MCPXKIT: The Unified Toolkit for Analyzing Model Context Protocol Security, http://dx.doi.org/10.48550/arxiv.2508.12538

Li T; Yan Z; Liu J; Di P; Zhang X, 2026, Guiding LLM-based Loop Invariant Synthesis via Feedback on Local Reasoning Errors, http://dx.doi.org/10.48550/arxiv.2605.17914

Zhou C; Chen D; Shen Z; jiang W; Li Y; Di P, 2026, Explaining the "Why": A Unified Framework for the Additive Attribution of Changes in Arbitrary Measures, https://arxiv.org/abs/2604.26266v1

Chen X; Xiao S; Zhu X; Xie J; Liang M; Chen D; Jiang W; Li Y; Di P, 2026, NES: An Instruction-Free, Low-Latency Next Edit Suggestion Framework Powered by Learned Historical Editing Trajectories, http://dx.doi.org/10.48550/arxiv.2508.02473

Di P; Chen F; Bai X; Yang H; Li Q; Wei G; Mou J; Shi F; Chen K; Tang P; Shen Z; Li Z; Shi W; Guo J; Yu H, 2025, OpenDerisk: An Industrial Framework for AI-Driven SRE, with Design, Implementation, and Case Studies, https://doi.org/10.1145/3786583.3786864

Guo H; Xie X; Dai H-N; Di P; Zhang Y; Tao B; Zheng Z, 2025, Accelerating Automatic Program Repair with Dual Retrieval-Augmented Fine-Tuning and Patch Generation on Large Language Models, http://dx.doi.org/10.48550/arxiv.2507.10103

Codefuse ; Team L; : ; Cai W; Cao Y; Chen C; Chen C; Chen S; Cui Q; Di P; Fang J; Gong Z; Guo T; He Z; Huang Y; Li C; Li J; Li Z; Lian S; Liu B; Luo S; Mao S; Shen M; Wu J; Yang J; Yang W; Ye T; Yu H; Zhang W; Zhang Z; Zhao H; Zheng X; Zhou J, 2025, Every Sample Matters: Leveraging Mixture-of-Experts and High-Quality Data for Efficient and Accurate Code LLM, http://dx.doi.org/10.48550/arxiv.2503.17793

Liu P; Sun C; Zheng Y; Feng X; Qin C; Wang Y; Xu Z; Li Z; Di P; Jiang Y; Sun L, 2025, Harnessing the Power of LLM to Support Binary Taint Analysis, http://dx.doi.org/10.48550/arxiv.2310.08275

Wu M; Xiang J; Chen K; DI P; Tan SH; Cui H; Zhang Y, 2024, Tumbling Down the Rabbit Hole: How do Assisting Exploration Strategies Facilitate Grey-box Fuzzing?, http://dx.doi.org/10.48550/arxiv.2409.14541

Di P; Li J; Yu H; Jiang W; Cai W; Cao Y; Chen C; Chen D; Chen H; Chen L; Fan G; Gong J; Gong Z; Hu W; Guo T; Lei Z; Li T; Li Z; Liang M; Liao C; Liu B; Liu J; Liu Z; Lu S; Shen M; Wang G; Wang H; Wang Z; Xu Z; Yang J; Ye Q; Zhang G; Zhang Y; Zhao Z; Zheng X; Zhou H; Zhu L; Zhu X, 2024, CodeFuse-13B: A Pretrained Multi-lingual Code Large Language Model, http://dx.doi.org/10.48550/arxiv.2310.06266

Di P; Liu B; Gao Y, 2024, MicroFuzz: An Efficient Fuzzing Framework for Microservices, http://dx.doi.org/10.48550/arxiv.2401.05529

Liu X; Wang J; Sun J; Yuan X; Dong G; Di P; Wang W; Wang D, 2023, Prompting Frameworks for Large Language Models: A Survey, http://dx.doi.org/10.48550/arxiv.2311.12785

Fan G; Xie X; Zheng X; Liang Y; Di P, 2023, Static Code Analysis in the AI Era: An In-depth Exploration of the Concept, Function, and Potential of Intelligent Code Analysis Agents, http://dx.doi.org/10.48550/arxiv.2310.08837

Liu J; Liu J; Di P; Wu D; Zheng H; Liu A; Xue J, 2022, Hybrid Inlining: A Compositional and Context Sensitive Static Analysis Framework, http://dx.doi.org/10.48550/arxiv.2210.14436


Back to profile page