Skip to content
View ZXT-zjbiliy's full-sized avatar
  • Zhejiang University
  • Zhejiang University

Highlights

  • Pro

Block or report ZXT-zjbiliy

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
ZXT-zjbiliy/README.md

ZJbiliy

Student @ Zhejiang University | AI & Open Source Enthusiast

GitHub followers GitHub stars


About Me

  • πŸŽ“ Student at Zhejiang University
  • πŸ€– Passionate about AI Infra, AI Agents, Multimodal Models
  • πŸ”§ Contributing to open-source projects in the AI ecosystem
  • πŸ† Multiple academic competition awards

Tech Stack

Python C++ C PyTorch TypeScript Docker Git Linux CUDA


Featured Projects

LiquidIdentification β€” Bottle Liquid-Level Detection

Built a pipeline for transparent bottle residual liquid identification using YOLO OBB and traditional ML.

  • Custom BottleDataset (218 images) with LCDTC integration
  • YOLO11m-OBB rotation detection + 247-dimensional interpretable features
  • Compared SVM, Random Forest, XGBoost; achieved mAP@0.5 = 0.811

minimind-llava-v β€” Lightweight Multimodal Model

Extended MiniMind with vision-language capabilities using a LLaVA-style architecture.

  • Adapted vision encoder + language model with projection / cross-attention interfaces
  • Training script with mixed precision, gradient accumulation, checkpoint recovery

ATTI β€” Auto AI Questionnaire Evaluator

Edge browser extension for automated AI questionnaire evaluation.

  • Supports MBTI / Truity information extraction
  • Session persistence, auto-fill, and memory-bank rule maintenance

Research Interests

  • πŸ”¬ Multimodal models and vision-language architectures
  • πŸ—οΈ AI infrastructure and efficient training/inference
  • πŸ€– Agent systems, MCP, and tool-use frameworks

Contribution Snake

github contribution grid snake animation

"The best way to predict the future is to invent it." β€” Alan Kay

Visitors

Pinned Loading

  1. minimind-llava-v minimind-llava-v Public

    MiniMind-LLaVA-V is a lightweight multimodal project that extends MiniMind with vision-language capabilities. It combines a MiniMind language model, a vision tower, and a LLaVA-style projector to s…

    Python 1

  2. ATTI ATTI Public

    Auto-Type-Indicator ATTI

    TypeScript 3