Craig Jones
grandcodepope
AI & ML interests
None yet
Recent Activity
updated a model about 7 hours ago
grandcodepope/buyasoul-gsk-complete new activity about 14 hours ago
kernels-community/triton_kernels:Update build/torch-universal/triton_kernels/target_info.py new activity about 14 hours ago
kernels-community/megablocks:Update build.tomlOrganizations
Update build/torch-universal/triton_kernels/target_info.py
6
#7 opened 6 months ago
by
KernelMC
Update build.toml
3
#4 opened 9 months ago
by
ahadnagy
Add flash_attn_with_kvcache paged-decode kernel (port from huggingface/transformers#45977)
1
#4 opened 4 months ago
by
ArthurZ
Add Metal (Apple Silicon) build variants
2
#6 opened 6 months ago
by
robtaylor-chipflow
[WIP] Add sliding-window attention support to the varlen kernel
5
#5 opened 3 months ago
by
ArthurZ
Update paged-attention-metal/attention/paged_attention.metal
3
#2 opened about 1 year ago
by
ArthurZ
Update autotune configuration to avoid crash on AMD devices
2
#2 opened over 1 year ago
by
ror
Available methods in a kernel
2
#5 opened 9 months ago
by
christopher5106
flash-attn 3 support for rtx 5090?
2
#2 opened 8 months ago
by
hydraofm0
TemporalMesh Transformer: 29.4 PPL at 48% compute — beats Mamba, new open-source architecture
2
#47 opened 3 months ago
by
vigneshwar234
Bug report + proposed fix for chat_template.jinja
👍🔥 2
1
#142 opened about 1 month ago
by
Khalidnass
Add exported onnx model 'model.onnx'
3
#63 opened 11 days ago
by
I-Vi
Request: DOI
4
#148 opened 20 days ago
by
Geminiredirect
Internal model structure preview
2
#17 opened 17 days ago
by
svetoviz
Internal model structure preview
3
#40 opened 17 days ago
by
svetoviz
Our most-deployed mobile model — 16.9 t/s on Snapdragon
2
#22 opened 2 months ago
by
3morixd
Request: DOI
3
#82 opened 3 months ago
by
kanha077