Skip to content
View Dovis01's full-sized avatar
🎯
Focusing
🎯
Focusing

Highlights

  • Pro

Block or report Dovis01

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Dovis01/README.md

Welcome to my profile

Technology banner

Shijin Zhang, also known as Doven — Open-Source LLM Inference and High-Performance AI Systems

Profile status  Python 3.12  visitors

Hi, I'm Shijin Zhang — you can call me Doven.
I'm passionate about open-source LLM inference, high-performance computing, and efficient AI systems.
Feel free to reach me at dovis.zhang02@gmail.com. Thanks for visiting!

🛠 Tech Stack

Property Data
Language / IDE Python  Java  PyCharm  Django  PyQt  C  C++  Bash 
Domain Knowledge Machine Learning  Computer Science  Software Development  High Performance Computing 
LLM Inference / Acceleration Transformers  vLLM  vLLM-Omni  SGLang  CUDA  DeepGEMM  FlashMLA  FlashInfer  FlashAttention 
Web Development TypeScript  Next.js 
CI / CD & Tools Git  GitHub  GitLab  Docker  VS Code 
Databases MySQL  SQLite  PostgreSQL  MongoDB 
ML / Deep Learning Frameworks Jupyter Notebook  Scikit-learn  PyTorch  TensorFlow  ChatGPT  OpenCV 

📈 GitHub Contribution Activity

Dovis01 GitHub contribution snake animation

Pinned Loading

  1. FlashMLA FlashMLA Public

    Forked from sgl-project/FlashMLA

    FlashMLA: Efficient Multi-head Latent Attention Kernels

    C++

  2. flashinfer flashinfer Public

    Forked from flashinfer-ai/flashinfer

    FlashInfer: Kernel Library for LLM Serving

    Python

  3. sglang sglang Public

    Forked from sgl-project/sglang

    SGLang is a high-performance serving framework for large language models and multimodal models.

    Python

  4. transformers transformers Public

    Forked from huggingface/transformers

    🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.

    Python

  5. vllm vllm Public

    Forked from vllm-project/vllm

    A high-throughput and memory-efficient inference and serving engine for LLMs

    Python

  6. vllm-omni vllm-omni Public

    Forked from vllm-project/vllm-omni

    A framework for efficient model inference with omni-modality models

    Python