Skip to content
#

kernel-fusion

Here are 17 public repositories matching this topic...

Fused Triton kernels for TurboQuant KV cache compression — 2-4 bit quantization with RHT rotation. Drop-in HuggingFace & vLLM integration. Up to 4.9x KV cache compression for Llama, Qwen, Mistral, and more.

  • Updated Jun 21, 2026
  • Python

Multi-Engine (PyTorch & JAX/XLA) Zero-Branching Geometric Acceleration Core. Enforces 0% Graph Breaks & Real-time Fault-Isolation via hardware-native bitwise MUX operations (torch.where / jax.lax.select) to permanently eliminate 'jmp' instructions and host-device synchronization fences.

  • Updated Jul 4, 2026
  • Python

Improve this page

Add a description, image, and links to the kernel-fusion topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the kernel-fusion topic, visit your repo's landing page and select "manage topics."

Learn more