🚀🚀 Efficient implementations of Native Sparse Attention
-
Updated
Sep 29, 2025 - Python
🚀🚀 Efficient implementations of Native Sparse Attention
Official pytorch implementation of the paper "Bayesian Meta-Learning for the Few-Shot Setting via Deep Kernels" (NeurIPS 2020)
This is the full file system fuzzing framework that I presented at the Hack in the Box 2020 Lockdown Edition conference in April.
Learn how to develop kernels
FireRed-Image-Edit-1.0-Fast is a high-performance, AI-driven image editing application that utilizes advanced diffusers and the QIE+ Pipeline for precise, prompt-based image modifications.
Supplementary code for the AISTATS 2021 paper "Matern Gaussian Processes on Graphs".
Supplementary code for the paper "Stationary Kernels and Gaussian Processes on Lie Groups and their Homogeneous Spaces"
Foundational library for Kernel methods in pattern analysis and machine learning
[MLSys 26] 🥇 Solution for Gated Delta Net Track of MLSys 26 Flash infer competition
High-performance GPU kernels for LLM inference in OpenAI Triton. Fused RMSNorm, SwiGLU, INT8 GEMM with benchmarks and roofline analysis.
An open attempt at reproducing a simplified version of Google's Genie World Model API on a single H100.
Operating System-based projects explored and implemented chapter-wise. The programs are inspired from end-of-chapter projects in Silberchatz's "Operating System Concepts".
Fast embedding-based graph classification with connections to kernels
QIE-Object-Remover-Bbox is an advanced, AI-powered image editing application specifically designed to perform precise object removal and background inpainting based on user-defined bounding box coordinates.
Code for the NeurIPS 2021 paper "Higher Order Kernel Mean Embeddings to Capture Filtrations of Stochastic Processes".
Add a description, image, and links to the kernels topic page so that developers can more easily learn about it.
To associate your repository with the kernels topic, visit your repo's landing page and select "manage topics."