庄表伟
Community@zhuangbiaowei · Shanghai, China
Agent Skills by 庄表伟
Showing 99 vetted skills indexed across 2 GitHub repositories.
master-control-expert
Classify requests into governance-backed workflows with control plans and decision logs.
dspy
Create declarative language model programs with automatic prompt optimization in DSPy.
openrlhf-training
Train large language models with RLHF algorithms using Ray and vLLM.
nemo-guardrails
Implement runtime safety guardrails for LLM applications using NeMo Guardrails and Colang 2.0 DSL.
llamaindex
Build LLM applications with RAG using data connectors and vector indices.
deepspeed
Optimize distributed training and inference for large-scale AI models.
crewai-multi-agent
Orchestrate role-based multi-agent systems for sequential and hierarchical task execution.
moe-training
Train Mixture of Experts models using DeepSpeed and HuggingFace Transformers.
test_script_type
Run Python scripts and capture their output and arguments.
nnsight-remote-interpretability
Inspect and modify PyTorch model activations with optional NDIF remote execution.
langsmith-observability
Trace, evaluate, and monitor LLM applications with LangSmith.
audiocraft-audio-generation
Generate music and sound effects from text using AudioCraft.
nemo-curator
Curate multimodal LLM training datasets with GPU-accelerated deduplication and filtering.
huggingface-accelerate
Simplify distributed PyTorch training with a unified API for DDP, DeepSpeed, and FSDP.
ml-paper-writing
Automate writing and formatting ML papers for NeurIPS, ICML, and ICLR.
chroma
Manage an open-source embedding database for semantic search and retrieval.
prompt-guard
Detect prompt injections and jailbreak attempts in LLM inputs.
evaluating-llms-harness
Evaluate LLMs across academic benchmarks like MMLU and GSM8K.
llama-factory
Fine-tune LLMs with LLaMA-Factory WebUI and QLoRA quantization.
youtube-downloader
Download YouTube videos and audio with yt-dlp in configurable formats.
knowledge-distillation
Compress large language models into smaller student models using knowledge distillation.
simpo-training
Optimize LLM alignment with reference-free SimPO preference optimization.
slime-rl-training
Trains large language models with GRPO reinforcement learning using Megatron-LM and SGLang.
blip-2-vision-language
Bridges frozen image encoders with LLMs for vision-language tasks like image captioning, VQA, and image-text retrieval.