Instructions to use replicate/flashinfer-draft with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Kernels
How to use replicate/flashinfer-draft with Kernels:
# !pip install kernels from kernels import get_kernel kernel = get_kernel("replicate/flashinfer-draft") - Notebooks
- Google Colab
- Kaggle
Add model repository removal notice
Browse files
README.md
CHANGED
|
@@ -4,6 +4,8 @@ tags:
|
|
| 4 |
- kernels
|
| 5 |
---
|
| 6 |
|
|
|
|
|
|
|
| 7 |
|
| 8 |
This kernel is a work in progress and requires more work to correctly add all of the FlashInfer kernels.
|
| 9 |
|
|
|
|
| 4 |
- kernels
|
| 5 |
---
|
| 6 |
|
| 7 |
+
> [!CAUTION]
|
| 8 |
+
> Starting from September 13, 2026, we will be removing the "model" type repositories of kernels (e.g., kernels-community/flash-attn3). Make sure you're using a latest version of kernels. If you face any disruption, please report them here: https://github.com/huggingface/kernels/issues/new.
|
| 9 |
|
| 10 |
This kernel is a work in progress and requires more work to correctly add all of the FlashInfer kernels.
|
| 11 |
|