Phạm Phú Ngọc Trai
Community@jayll1303 · HCM Viet Nam
FPTU K16 @JayLL-vn
Agent Skills by Phạm Phú Ngọc Trai
Showing 31 vetted skills indexed across 1 GitHub repositories.
semantic-router
Routes user queries to handlers using semantic vector matching and configurable encoders/indexes.
notebook-workflows
Automate Jupyter and Colab notebook creation and editing via nbformat and nbconvert.
ml-brainstorm
Brainstorm ML/AI decisions and recommend paths from repository context.
openai-audio-api
Build OpenAI-compatible TTS endpoints with streaming and batch inference.
hf-transformers-trainer
Fine-tune and align local LLMs with HuggingFace Trainer, PEFT, and TRL.
sherpa-onnx
Perform offline speech processing with ONNX models via sherpa-onnx.
opentelemetry
Instrument Python and Node.js apps with OpenTelemetry to collect traces, metrics, and logs.
hf-speech-to-speech-pipeline
Orchestrate real-time speech-to-speech pipelines connecting VAD, STT, LLM, and TTS components.
python-project-setup
Bootstrap Python projects with uv and a single pyproject.toml.
paddleocr
Train, fine-tune, and export PaddleOCR models for text detection and recognition.
ollama-local-llm
Manage local LLMs via Ollama CLI and API with Modelfile creation.
python-quality-testing
Add type annotations, contract docstrings, Hypothesis tests, and mutmut mutation testing to Python ML code.
hf-hub-datasets
Download and upload HuggingFace Hub models and datasets with authentication.
tensorrt-llm
Convert HuggingFace checkpoints into optimized TensorRT-LLM engines with quantization.
text-embeddings-rag
Generate local embeddings and build retrieval-augmented generation pipelines with sentence-transformers and FAISS.
arxiv-reader
Fetch and analyze arXiv papers via HTML-rendered pages.
aie-skills-installer
Detect technology signals in Python and ML repositories to recommend and install matching AIE-Skills.
experiment-tracking
Deploy self-hosted MLflow and Weights & Biases trackers for ML experiments.
llama-cpp-inference
Run local GGUF model inference and OpenAI-compatible serving with llama.cpp.
sglang-serving
Launch and tune SGLang servers with constrained decoding and RadixAttention.
vllm-tgi-inference
Deploy local vLLM or HuggingFace TGI inference servers with OpenAI-compatible APIs.
docker-gpu-setup
Build GPU-enabled Docker containers with NGC base images and uv.
ultralytics-yolo
Train and export Ultralytics YOLO models for detection and tracking.
k2-training-pipeline
Train ASR and TTS models with k2, icefall, and lhotse pipelines.