XIAO YOUWEI avatar

XIAO YOUWEI

Community

@uv-xiao

87Followers
|
67Public Repos
|
54Published Skills

Fourth-year Ph.D. student at Peking University, researching language and compiler techniques for MLSys/Arch/Hardware design.

Agent Skills by XIAO YOUWEI

Showing 54 vetted skills indexed across 1 GitHub repositories.

uv-xiaouv-xiao
1

uv-example-skill

Validate skill mirror and manifest generation in automated tests.

Community
Basic
uv-xiaouv-xiao
1

uv-find-skills

Search the open agent skills ecosystem and guide installation via the Skills CLI.

Community
Basic
uv-xiaouv-xiao
1

uv-bootstrap-skill-linking

Define prerequisite, companion, follow-up, and escalation relationships between skills.

Community
Intermediate
uv-xiaouv-xiao
1

uv-bootstrap-ml-knowledge-authoring

Scaffold new ML knowledge skills with predefined repository structure and naming conventions.

Community
Intermediate
uv-xiaouv-xiao
1

uv-skill-evolution-manager

Persist session feedback into evolution.json and stitch learnings into SKILL.md.

Community
Intermediate
uv-xiaouv-xiao
1

uv-bootstrap-skill-maintenance

Manage pkbllm skills repository additions, imports, and refactors.

Community
Intermediate
uv-xiaouv-xiao
1

uv-using-pkb

Guide users through discovering and invoking uv-* skills in pkbllm.

Community
Basic
uv-xiaouv-xiao
1

uv-start-task

Assemble relevant skill notes into a project's AGENTS.md file.

Community
Basic
uv-xiaouv-xiao
1

uv-ml-paper-writing

Creates ML/CS conference submissions with LaTeX templates and citation verification workflows.

Community
Advanced
uv-xiaouv-xiao
1

uv-serving-llms-vllm

Serve LLMs with vLLM using OpenAI-compatible endpoints and quantization.

Community
Advanced
uv-xiaouv-xiao
1

uv-tensorrt-llm

Optimize LLM inference with NVIDIA TensorRT-LLM on NVIDIA GPUs.

Community
Advanced
uv-xiaouv-xiao
1

uv-sglang

Serve LLMs with RadixAttention prefix caching and structured generation.

Community
Advanced
uv-xiaouv-xiao
1

uv-llama-cpp

Run LLM inference on CPUs, Apple Silicon, and non-NVIDIA GPUs with GGUF quantization.

Community
Intermediate
uv-xiaouv-xiao
1

uv-deepspeed

Guide distributed training with DeepSpeed ZeRO, pipeline parallelism, and mixed precision.

Community
Advanced
uv-xiaouv-xiao
1

uv-pytorch-fsdp2

Integrate PyTorch FSDP2 fully_shard into training scripts with DCP checkpointing.

Community
Advanced
uv-xiaouv-xiao
1

uv-ray-train

Orchestrates distributed ML training across multi-node clusters using Ray Train.

Community
Advanced
uv-xiaouv-xiao
1

uv-slime-rl-training

Run GRPO reinforcement learning training for LLMs with Megatron-LM and SGLang.

Community
Advanced
uv-xiaouv-xiao
1

uv-miles-rl-training

Train large-scale Mixture-of-Experts models with FP8/INT4 quantization-aware training and speculative RL.

Community
Advanced
uv-xiaouv-xiao
1

uv-verl-rl-training

Run large-scale RL training for LLMs with verl and GRPO.

Community
Advanced
uv-xiaouv-xiao
1

uv-moe-training

Train Mixture of Experts models with DeepSpeed and HuggingFace Transformers.

Community
Advanced
uv-xiaouv-xiao
1

uv-implementing-llms-litgpt

Implement and fine-tune LLM architectures with LitGPT.

Community
Advanced
uv-xiaouv-xiao
1

uv-mamba-architecture

Implement the Mamba state-space model architecture with O(n) complexity.

Community
Intermediate
uv-xiaouv-xiao
1

uv-rwkv-architecture

Explain the RWKV hybrid RNN-Transformer architecture with linear-time inference.

Community
Advanced
uv-xiaouv-xiao
1

uv-speculative-decoding

Accelerate LLM inference with speculative decoding and Medusa multiple heads.

Community
Advanced