Select Publications
Journal articles
, 2026, 'Exploiting Class-agnostic Visual Prior for Few-shot Keypoint Detection', International Journal of Computer Vision, 134, http://dx.doi.org/10.1007/s11263-025-02671-5
, 2026, 'An Efficient End-to-End Framework for Localized Second-Order Pooling', Transactions on Machine Learning Research, 2026-June
, 2026, 'Evolving Multimodal Models for Physical Dynamics: A Multi-objective Neuroevolution Approach', IEEE Transactions on Evolutionary Computation, http://dx.doi.org/10.1109/TEVC.2026.3698641
, 2025, 'Feature Hallucination for Self-supervised Action Recognition', International Journal of Computer Vision, 133, pp. 7612 - 7646, http://dx.doi.org/10.1007/s11263-025-02513-4
, 2025, 'Beyond decoders: Learning prompt-aware features for few-shot object counting', Neurocomputing, 651, http://dx.doi.org/10.1016/j.neucom.2025.130997
, 2025, 'WWW 2025 Workshop Chairs' Welcome', Www Companion 2025 Companion Proceedings of the ACM Web Conference 2025, pp. 1563 - 1564, http://dx.doi.org/10.1145/3701716.3725685
, 2025, 'Point-PNG: Conditional Pseudo-Negatives Generation for Point Cloud Pre-Training', IEEE Access, 13, pp. 208612 - 208625, http://dx.doi.org/10.1109/ACCESS.2025.3641626
, 2024, 'Multivariate prototype representation for domain-generalized incremental learning', Computer Vision and Image Understanding, 249, http://dx.doi.org/10.1016/j.cviu.2024.104215
, 2024, 'Saliency-guided meta-hallucinator for few-shot learning', Science China Information Sciences, 67, http://dx.doi.org/10.1007/s11432-023-4113-1
, 2024, 'Meet JEANIE: A Similarity Measure for 3D Skeleton Sequences via Temporal-Viewpoint Alignment', International Journal of Computer Vision, 132, pp. 4091 - 4122, http://dx.doi.org/10.1007/s11263-024-02070-2
, 2024, 'Synergizing triple attention with depth quality for RGB-D salient object detection', Neurocomputing, 589, http://dx.doi.org/10.1016/j.neucom.2024.127672
, 2024, 'Graph neural networks-enhanced relation prediction for ecotoxicology (GRAPE)', Journal of Hazardous Materials, 472, http://dx.doi.org/10.1016/j.jhazmat.2024.134456
, 2024, 'Traffic forecasting on new roads using spatial contrastive pre-training (SCPT)', Data Mining and Knowledge Discovery, 38, pp. 913 - 937, http://dx.doi.org/10.1007/s10618-023-00982-0
, 2023, 'Exploiting Field Dependencies for Learning on Categorical Data', IEEE Transactions on Pattern Analysis and Machine Intelligence, 45, pp. 13509 - 13522, http://dx.doi.org/10.1109/TPAMI.2023.3298028
, 2023, 'Accurate 3-DoF Camera Geo-Localization via Ground-to-Satellite Image Matching', IEEE Transactions on Pattern Analysis and Machine Intelligence, 45, pp. 2682 - 2697, http://dx.doi.org/10.1109/TPAMI.2022.3189702
, 2023, 'Event-guided Multi-patch Network with Self-supervision for Non-uniform Motion Deblurring', International Journal of Computer Vision, 131, pp. 453 - 470, http://dx.doi.org/10.1007/s11263-022-01708-3
, 2023, 'Multi-Level Second-Order Few-Shot Learning', IEEE Transactions on Multimedia, 25, pp. 2111 - 2126, http://dx.doi.org/10.1109/TMM.2022.3142955
, 2023, 'Preface', Lecture Notes in Computer Science, 13848 LNCS, pp. v
, 2022, 'Power Normalizations in Fine-Grained Image, Few-Shot Image and Graph Classification', IEEE Transactions on Pattern Analysis and Machine Intelligence, 44, pp. 591 - 609, http://dx.doi.org/10.1109/TPAMI.2021.3107164
, 2022, 'Predicting flight delay with spatio-temporal trajectory convolutional network and airport situational awareness map', Neurocomputing, 472, pp. 280 - 293, http://dx.doi.org/10.1016/j.neucom.2021.04.136
, 2022, 'Tensor Representations for Action Recognition', IEEE Transactions on Pattern Analysis and Machine Intelligence, 44, pp. 648 - 665, http://dx.doi.org/10.1109/TPAMI.2021.3107160
, 2020, 'A Comparative Review of Recent Kinect-Based Action Recognition Algorithms', IEEE Transactions on Image Processing, 29, pp. 15 - 28, http://dx.doi.org/10.1109/TIP.2019.2925285
, 2019, 'Identity-Preserving Face Recovery from Stylized Portraits', International Journal of Computer Vision, 127, pp. 863 - 883, http://dx.doi.org/10.1007/s11263-019-01169-1
, 2017, 'Higher-Order Occurrence Pooling for Bags-of-Words: Visual Concept Detection', IEEE Transactions on Pattern Analysis and Machine Intelligence, 39, pp. 313 - 326, http://dx.doi.org/10.1109/TPAMI.2016.2545667
, 2014, 'Robust multi-speaker tracking via dictionary learning and identity modeling', IEEE Transactions on Multimedia, 16, pp. 864 - 880, http://dx.doi.org/10.1109/TMM.2014.2301977
, 2013, 'A robust and scalable visual category and action recognition system using kernel discriminant analysis with spectral regression', IEEE Transactions on Multimedia, 15, pp. 1653 - 1664, http://dx.doi.org/10.1109/TMM.2013.2264927
, 2013, 'Comparison of mid-level feature coding approaches and pooling strategies in visual concept detection', Computer Vision and Image Understanding, 117, pp. 479 - 492, http://dx.doi.org/10.1016/j.cviu.2012.10.010
Conference Papers
, 2026, 'Learning Time in Static Classifiers', in Proceedings of the Aaai Conference on Artificial Intelligence, pp. 20816 - 20825, http://dx.doi.org/10.1609/aaai.v40i25.39221
, 2025, 'Stabilizing Modality Gap & Lowering Gradient Norms Improve Zero-Shot Adversarial Robustness of VLMs', in Proceedings of the ACM SIGKDD International Conference on Knowledge Discovery and Data Mining, pp. 236 - 247, http://dx.doi.org/10.1145/3690624.3709296
, 2025, 'Understanding and Mitigating Hyperbolic Dimensional Collapse in Graph Contrastive Learning', in Proceedings of the ACM SIGKDD International Conference on Knowledge Discovery and Data Mining, pp. 1984 - 1995, http://dx.doi.org/10.1145/3690624.3709249
, 2025, 'Graph Self-Supervised Learning with Learnable Structural and Positional Encodings', in Www 2025 Proceedings of the ACM Web Conference, pp. 4053 - 4067, http://dx.doi.org/10.1145/3696410.3714745
, 2025, 'Inductive Graph Few-shot Class Incremental Learning', in Wsdm 2025 Proceedings of the 18th ACM International Conference on Web Search and Data Mining, pp. 466 - 474, http://dx.doi.org/10.1145/3701551.3703578
, 2025, 'Adaptive Multi-head Contrastive Learning', in Lecture Notes in Computer Science Including Subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics, pp. 404 - 421, http://dx.doi.org/10.1007/978-3-031-72890-7_25
, 2025, 'Adversarially Robust Distillation by Reducing the Student-Teacher Variance Gap', in Lecture Notes in Computer Science Including Subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics, pp. 92 - 111, http://dx.doi.org/10.1007/978-3-031-73235-5_6
, 2025, 'Amortized Active Generation of Pareto Sets', in Advances in Neural Information Processing Systems, pp. 95495 - 95520
, 2025, 'BiLoRA: Almost-orthogonal Parameter Spaces for Continual Learning', in Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition, pp. 25613 - 25622, http://dx.doi.org/10.1109/CVPR52734.2025.02385
, 2025, 'CrossSpectra: Exploiting Cross-Layer Smoothness for Parameter-Efficient Fine-Tuning', in Advances in Neural Information Processing Systems, pp. 156796 - 156818
, 2025, 'Graph Your Own Prompt', in Advances in Neural Information Processing Systems, pp. 42542 - 42587
, 2025, 'Improving Zero-Shot Adversarial Robustness in Vision-Language Models by Closed-form Alignment of Adversarial Path Simplices', in Proceedings of Machine Learning Research, pp. 14061 - 14078
, 2025, 'LEARNABLE EXPANSION OF GRAPH OPERATORS FOR MULTI-MODAL FEATURE FUSION', in 13th International Conference on Learning Representations Iclr 2025, pp. 78379 - 78401
, 2025, 'Machine Unlearning viaTask Simplex Arithmetic', in Advances in Neural Information Processing Systems, pp. 184768 - 184799
, 2025, 'Noise Consistency Regularization for Improved Subject-Driven Image Synthesis', in IEEE Computer Society Conference on Computer Vision and Pattern Recognition Workshops, pp. 3107 - 3117, http://dx.doi.org/10.1109/CVPRW67362.2025.00294
, 2025, 'Open-World Objectness Modeling Unifies Novel Object Detection', in Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition, pp. 30332 - 30342, http://dx.doi.org/10.1109/CVPR52734.2025.02824
, 2025, 'OpenKD: Opening Prompt Diversity for Zero- and Few-Shot Keypoint Detection', in Lecture Notes in Computer Science Including Subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics, pp. 148 - 165, http://dx.doi.org/10.1007/978-3-031-72655-2_9
, 2025, 'Primitive Vision: Improving Diagram Understanding in MLLMs', in Proceedings of Machine Learning Research, pp. 74732 - 74755
, 2025, 'Protein Fitness Landscape: Spectral Graph Theory Perspective', in Proceedings of Machine Learning Research, pp. 2827 - 2835
, 2025, 'Robust SuperAlignment: Weak-to-Strong Robustness Generalization for Vision-Language Models', in Advances in Neural Information Processing Systems, pp. 20822 - 20854
, 2025, 'Robustifying Zero-Shot Vision Language Models by Subspaces Alignment', in Proceedings of the IEEE International Conference on Computer Vision, pp. 21037 - 21047, http://dx.doi.org/10.1109/ICCV51701.2025.01955
, 2024, 'Geometric View of Soft Decorrelation in Self-Supervised Learning', in Proceedings of the ACM SIGKDD International Conference on Knowledge Discovery and Data Mining, pp. 4338 - 4349, http://dx.doi.org/10.1145/3637528.3671914
, 2024, 'Detect Any Keypoints: An Efficient Light-Weight Few-Shot Keypoint Detector', in Proceedings of the Aaai Conference on Artificial Intelligence, pp. 3882 - 3890, http://dx.doi.org/10.1609/aaai.v38i4.28180