唯品会 avatar

唯品会

Official

@vipshop · Guangzhou China

0Followers
|
36Public Repos
|
7Published Skills

全球精选,正品特卖 (NASDAQ: VIPS)

Skills Distribution
DomainAI Models & ...GPU Kernel Optimiz.. (40%)Operator Migration (30%)High-Performance C.. (30%)

Agent Skills by 唯品会

Showing 7 vetted skills indexed across 1 GitHub repositories.

Frequently Asked Questions About 唯品会

FAQPage Schema
What specific tasks are enabled by these kernel optimization skills?

These skills enable the implementation and optimization of high-performance GPU kernels using CuTe DSL and CUTLASS. They facilitate the migration of native CUDA or Triton operators into the cache-dit environment, ensuring rigorous validation and performance parity across NVIDIA hardware architectures.

Which engineering personas benefit from these technical capabilities?

These capabilities are designed for GPU performance engineers, machine learning infrastructure developers, and systems architects focused on low-level hardware acceleration. They are specifically intended for engineers tasked with optimizing tensor computation layers and managing operator portability within high-performance computing environments.

What are the primary prerequisites for implementing these kernels?

Implementation requires a deep understanding of NVIDIA GPU architectures, proficiency in C++ for high-performance computing, and familiarity with the CUTLASS and CuTe programming models. Users must also have an existing cache-dit environment to integrate and validate the migrated operators.