---
name: "ascend-triton-operator-development"
description: "Develop a first correct Ascend Triton operator from a PyTorch reference or migrate an existing GPU Triton kernel to Ascend, including semantic audit, explicit task contracts, hardware-aware grid and tiling design, implementation, and handoff to correctness validation. Use for new kernel implementation, CUDA/GPU Triton migration, or repairing a candidate that has not yet passed correctness. Do not use for a kernel that already passes all planned cases and only needs performance tuning, for isolated torch_npu or ACLNN debugging, or for model-level graph failures."
---

<!-- Generated from .agents/skills/ascend-triton-operator-development/SKILL.md. Do not edit. -->

# ascend-triton-operator-development

Read `.agents/skills/ascend-triton-operator-development/SKILL.md` and only the references needed
for the current task. The canonical skill owns the workflow.
