★ 1 · Updated 2026-09-15
Writes numerical equivalence tests comparing a tensor-descriptor kernel against its original implementation, saving to tests/triton/test_<name>_td.py.
Browse skills that share this tag.
★ 1 · Updated 2026-09-15
Writes numerical equivalence tests comparing a tensor-descriptor kernel against its original implementation, saving to tests/triton/test_<name>_td.py.
★ 602 · Updated 2026-06-30
Write docstrings for PyTorch functions and methods following PyTorch Sphinx/reStructuredText conventions.
★ 18 · Updated 2026-06-30
Optimizes transformer attention using Flash Attention via PyTorch SDPA, flash-attn library, H100 FP8, and sliding window attention.
★ 602 · Updated 2026-06-30
PyTorch-based RL algorithms library for training agents and creating custom environments
★ 23 · Updated 2026-06-30
PyTorch-based RL algorithms for training agents, custom environments, callbacks, and workflow optimization
★ 602 · Updated 2026-06-30
Implement Metal/MPS kernels for PyTorch operators on Apple Silicon
★ 0 · Updated 2026-06-30
Implement Metal/MPS kernels for PyTorch operators using native c10/metal infrastructure
★ 4,344 · Updated 2026-06-15
Catalog common PyTorch mistakes and provide battle-tested training patterns and optimization techniques
★ 18 · Updated 2026-05-11
Routes GitHub issues to teams and applies labels
★ 602 · Updated 2026-05-11
Triages GitHub issues by routing to oncall teams, applying labels, and closing questions
★ 1 · Updated 2026-03-25
Write PyTorch and MLX code for Apple Silicon M-series hardware without CUDA
★ 18 · Updated 2026-03-20
Complete machine learning environment for model training, GPU acceleration, and data science workflows
★ 581 · Updated 2026-03-10
Integrate Counter_Guide module with Adaptive_Weight into dual-branch ViT for RGB/Event fusion using Multi_Context architecture
★ 0 · Updated 2026-03-10
Design and implement PyTorch modules with best practices for memory, DDP, and performance.
★ 2 · Updated 2026-02-13
Train neural networks for dynamical systems and control applications
★ 18 · Updated 2026-02-11
This skill helps access high-performance GPUs through Basilica's CLI for machine learning training, distributed jobs, and compute resource management
★ 602 · Updated 2026-02-11
Manage GPU rentals, ML training jobs, and compute resources on Basilica's decentralized GPU marketplace