ighoshsubho GitHub avatar

awesome-kernel-skills

ighoshsubho

A public repository showcasing kernel writing skills using CuTeDSL, Triton, Tilelang, and CUDA.

Stars

24

7-day growth

No data

Forks

2

Open issues

0

License

No data

Last updated

2026-07-07

AI repository intelligence
FR-AI / ANALYSIS

Why it is worth attention

It provides a rare comparison of kernel development across multiple modern GPU DSLs and CUDA in a single repository.

Who it is for

  • GPU kernel developers seeking examples across DSLs
  • Researchers comparing programming models for GPU computation
  • Students learning advanced GPU programming techniques
  • Performance engineers evaluating different kernel implementations

Use cases

  • Learning how to write efficient GPU kernels in CuTeDSL, Triton, Tilelang, and CUDA
  • Comparing syntax and performance characteristics of different kernel programming approaches
  • Using as a reference for porting kernels between DSLs
  • Training material for GPU computing workshops or courses

Strengths

  • Covers four distinct kernel programming frameworks/DSLs in one place
  • Directly relevant to cutting-edge GPU computing research and practice
  • Likely contains practical code examples (based on repository purpose)
  • Open-source and publicly accessible for learning and collaboration

Considerations

  • Extremely minimal README with no documentation, examples, or structure described
  • Unclear if repository contains substantial code or just placeholders
  • No information about maintenance status or completeness of coverage for each DSL

Description

Public repo for kernel writing skills in CuTeDSL, Triton, Tilelang and CUDA

Related repositories

Similar projects matched by category, topics, and programming language.

lopopolo
Featured
lopopolo GitHub avatar

harness-engineering

Harness Engineering is a methodology for improving coding agent outputs by carefully crafting the environment around them—providing curated context, tools, and executable constraints that encode an organization’s nonfunctional requirements and cumulative lessons.

AI & Machine LearningAI Agents
2,390
slvDev
Featured
slvDev GitHub avatar

esp32-ai

A 28.9 million parameter language model runs on an $8 ESP32-S3 microcontroller entirely on-device, generating simple stories at about 9.5 tokens per second.

AI & Machine LearningLarge Language Models
1,960
littledivy
Featured
littledivy GitHub avatar

mimic

mimic captures traffic from any iOS or web app and automatically generates a Python client library that lets you call the app's API like a regular library.

AI & Machine Learning
1,482