MLX: quantize embeddings (4w/64) and emit get_model_schema f82677f verified msluszniak commited on 10 days ago
Add LFM2.5-ColBERT-350M (XNNPACK 8da4w + MLX int4) for react-native-executorch d5fa46d verified nklockiewicz commited on Jun 22