- π Student at Zhejiang University
- π€ Passionate about AI Infra, AI Agents, Multimodal Models
- π§ Contributing to open-source projects in the AI ecosystem
- π Multiple academic competition awards
LiquidIdentification β Bottle Liquid-Level Detection
Built a pipeline for transparent bottle residual liquid identification using YOLO OBB and traditional ML.
- Custom BottleDataset (218 images) with LCDTC integration
- YOLO11m-OBB rotation detection + 247-dimensional interpretable features
- Compared SVM, Random Forest, XGBoost; achieved mAP@0.5 = 0.811
minimind-llava-v β Lightweight Multimodal Model
Extended MiniMind with vision-language capabilities using a LLaVA-style architecture.
- Adapted vision encoder + language model with projection / cross-attention interfaces
- Training script with mixed precision, gradient accumulation, checkpoint recovery
ATTI β Auto AI Questionnaire Evaluator
Edge browser extension for automated AI questionnaire evaluation.
- Supports MBTI / Truity information extraction
- Session persistence, auto-fill, and memory-bank rule maintenance
- π¬ Multimodal models and vision-language architectures
- ποΈ AI infrastructure and efficient training/inference
- π€ Agent systems, MCP, and tool-use frameworks