教育经历#

深圳大学

计算机科学与技术 · 硕士

2026.09 – 2029.06 广东深圳

桂林电子科技大学

人工智能 · 本科

2022.09 – 2026.06 广西桂林

研究兴趣#

Diffusion Models Flow Matching LLM VLM MLLM Medical Image Segmentation Agent

动态#

论文发表#

MSConvFormer: A Multi-Scale Depthwise Convolution Transformer for Medical Time Series Classification

作者列表待补充。

Digital Signal Processing(JCR-Q2),审稿中。

详情

Abstract

摘要待补充。

Hierarchical Progressive Cross-modal Information Interaction for Incomplete Multimodal Brain Tumor Segmentation

作者列表待补充。

MICCAI 2026(CCF-B),审稿中。

详情

Abstract

摘要待补充。

DMKformer: A Dual-Path Routing Transformer with Hybrid Mamba and KAN for Medical Time Series

Zhizun Zeng, Qiuyuan Gan, Wei Zhang, Yujie Li

Signal, Image and Video Processing,2026。(JCR-Q3)

详情

Abstract

Transformer has shown great potential in modeling global dependencies in sequential data, research on Transformer-based variant architectures for medical time series analysis has attracted growing attention. However, challenges remain in capturing complex local patterns and long-term dependencies in medical time series due to their high-dimensional nonlinearities and temporal correlations. To address this, we propose a Transformer model DMKformer which incorporates a new feature fusion strategy named 2R-Attention. Specifically, on the basis of the original Transformer, we adopt the STAR and the self-attention structure to enhance the ability of the attention module to capture and utilize the information of long-term dependencies in the global signal. Meanwhile, a one-dimensional DWC module is introduced to improve local information capture and reduce overfitting. Additionally, we employ Mamba to optimize FFN for long sequences and utilize a KAN network to strengthen classification capabilities. We conduct extensive experiments on three publicly available datasets under both subject-independent and subject-dependent settings. The experimental results demonstrate that our method outperforms 11 benchmark models, achieving an average F1 score improvement of 2%, validating its effectiveness.

Gaze Estimation based on Visual State Space Model with Hybrid Features

Yujie Li, Rongjie Liu, Zhizun Zeng, Ziwen Wang, Yuhang Hong, Benying Tan

ACM Transactions on Sensor Networks,2025。(JCR-Q1/CCF-B)

详情

Abstract

Visual State Space Model (denoted as VMamba), a vision-based model proposed to introduce Mamba into computer vision, has shown strong performance in recent work on computer vision tasks. However, the performance of VMamba in gaze estimation is still to be explored. In this paper, we propose two VMamba-based gaze estimation approaches: GazeVM-Pure based on pure VMamba and GazeVM-Hybrid based on hybrid VMamba. GazeVM-Pure is used to estimate gaze direction according to the original VMamba structure. GazeVM-Hybrid combines Convolutional Neural Network (CNN) and VMamba, where the Visual State Space (VSS) Block (the core module of VMamba) is used as a complementary component of the CNN. In GazeVM-Hybrid, the convolutional layers of ResNet-34 are used to learn local feature maps from face images, and VSS Block is used to capture global relations from feature maps. The experimental results show that GazeVM-Hybrid exhibits superior performance compared with existing state-of-the-art techniques with a nearly 0.11 decrease in angle error compared with Static Transformer Temporal Differential Network (STTDN) on the EyeDiap dataset.

GA-YOLOv8: A Road Defect Detection Model Based on Improved YOLOv8

ZhiZun Zeng, RongJie Liu, YongHang Huang, YuJie Li, BenYing Tan

VSIP 2025。(EI-Core/ACM)

详情

Abstract

Timely and accurate detection of pavement defects is essential for maintaining road safety. Traditional manual detection methods are labor-intensive, costly, and inefficient. To address the challenges posed by various defect scales, complex backgrounds, and subtle visibility of defects, this paper proposes an enhanced pavement defect detection model, GA-YOLOv8s, based on YOLOv8s. Our approach contains several key improvements. First, we add an additional P6 layer to YOLOv8s, thus introducing a multi-scale detection strategy that enables the model to capture deeper semantic features and improve the detection of small targets such as cracks. Second, we propose the C2f-GDC module, which maintains the model’s ability to extract fine-grained features, especially for small defects such as cracks, while reducing GFLOPs. Finally, we integrate the Adaptive Graphic Channel Attention (AGCA) mechanism to improve cross-layer feature fusion to effectively balance deep semantic information and shallow geometric details. Experimental results show that the mAP value is improved by 3.5% compared to the YOLOv8s model. Compared with other algorithms, the method proposed in this paper performs better in small and mild damage detection.

MamKanformer: A hybrid Mamba and KAN model based on Medformer for Medical Time-Series Classification

Zhizun Zeng, Qiuyuan Gan, Zhoukun Yan, Wei Zhang

CAIT 2024。(EI-Core/IEEE)

详情

Abstract

The rapid advancements in sensing technology and artificial intelligence have spurred a substantial number of novel demands and challenges within the realm of medical time series analysis. Notably, the effective extraction and utilization of the long-term and short-term dependency relationships, along with the features at various granularities encapsulated within high-dimensional and multivariate time series, have emerged as one of the pressing issues that demand immediate resolution. To address these, we propose a deep learning model MamKanformer for multichannel EEG signal processing. The model firstly solves the problem of long-short term dependencies which are difficult to be modeled effectively by Feed Forward in Transformer by using Mamba architecture. Subsequently, Kan is utilized to capture the critical feature information of different granularities contained in the time series data to assist in enhancing the output of the model. Finally, extensive experiments conducted on three public datasets have effectively verified the efficacy of the method proposed in this paper. Specifically, the MamKanformer model has demonstrated a superior performance over the other 11 baselines. It has achieved an average top ranking across all six metrics within the three datasets, which unequivocally attests to the outstanding capabilities of our proposed method.

奖项与荣誉#

学业表现

  • GPA 3.84/5

    专业前 2%

  • 校级一等奖学金

    年度奖项

学科竞赛

  • 华为 ICT 大赛 - 昇腾 AI 赛道

    Global · 一等奖

  • 美国大学生数学建模竞赛

    Honorable Mention

  • 挑战杯"揭榜挂帅"虚实联动专项赛

    National · 二等奖

  • "蓝桥杯"全国软件和信息技术专业人才大赛(B 组)

    National · 三等奖

  • 全国大学生数学建模竞赛

    National · 二等奖

  • 全国大学生数学竞赛

    Provincial · 一等奖