-
Na0s/Llama-3.2-3B-Instruct-Medical-Chatbot-LoRA-FT
Text Generation • 3B • Updated • 20 • 1 -
Na0s/Llama-3.2-3B-Medical-Chatbot-LoRA-FT
Text Generation • 3B • Updated • 24 • 1 -
meta-llama/Llama-3.2-3B-Instruct
Text Generation • 3B • Updated • 1.61M • • 2.72k -
meta-llama/Llama-3.2-3B
Text Generation • 3B • Updated • 335k • • 994
Ali Janati
Na0s
AI & ML interests
NLP, Speech Recognition, Computer Vision, Time Series Forecasting.
Recent Activity
published a model 2 days ago
frisson-labs/Faynt-75M-Arena published a model 2 days ago
frisson-labs/Faynt-10M-Arena published a model 2 days ago
frisson-labs/Faynt-75M-ExpertOrganizations
Depth pruned and fine tuned Llama-3.1-8B
-
Na0s/Llama-3.1-8B-Pruned-4-Layers_LoRA-PEFT-3.0
Text Generation • 7B • Updated • 38 • 4 -
Na0s/Llama-3.1-8B-Pruned-4-Layers_LoRA-PEFT-2.0
Text Generation • 7B • Updated • 20 • 1 -
Na0s/Llama-3.1-8B-Pruned-4-Layers_LoRA-PEFT-1.0
Text Generation • 7B • Updated • 12 -
Na0s/Llama-3.1-8B-Pruned-4-Layers_LoRA-PEFT
Text Generation • 7B • Updated • 11
Pruned MoEs (Mixtral-8x7B-Instruct-v0.1)
Mixtral expert-pruning checkpoints. NeurIPS 2026 AXIOM Workshop: https://arxiv.org/abs/2608.07890. Builds on Chowdhury et al. (ICML 2024).
-
mistralai/Mixtral-8x7B-Instruct-v0.1
47B • Updated • 215k • 4.76k -
Na0s/Mixtral-8x7B-Instruct-v0.1-LoRA-on-Gates
Text Generation • 47B • Updated • 25 • 1 -
Na0s/Mixtral-8x7B-Instruct-v0.1-exhaustive-LoRA
Text Generation • 47B • Updated • 14 -
Na0s/Mixtral-8x7B-v0.1-instruct-pruned-random-1-experts
Text Generation • 41B • Updated • 20
Medical Chatbot
-
Na0s/Llama-3.2-3B-Instruct-Medical-Chatbot-LoRA-FT
Text Generation • 3B • Updated • 20 • 1 -
Na0s/Llama-3.2-3B-Medical-Chatbot-LoRA-FT
Text Generation • 3B • Updated • 24 • 1 -
meta-llama/Llama-3.2-3B-Instruct
Text Generation • 3B • Updated • 1.61M • • 2.72k -
meta-llama/Llama-3.2-3B
Text Generation • 3B • Updated • 335k • • 994
Differential transformers
Fine-tuning foundation Llama-3.2-3B-Instruct on medical Q&A using differential attention (In progress). Paper: https://arxiv.org/pdf/2410.05258
Depth pruned and fine tuned Llama-3.1-8B
-
Na0s/Llama-3.1-8B-Pruned-4-Layers_LoRA-PEFT-3.0
Text Generation • 7B • Updated • 38 • 4 -
Na0s/Llama-3.1-8B-Pruned-4-Layers_LoRA-PEFT-2.0
Text Generation • 7B • Updated • 20 • 1 -
Na0s/Llama-3.1-8B-Pruned-4-Layers_LoRA-PEFT-1.0
Text Generation • 7B • Updated • 12 -
Na0s/Llama-3.1-8B-Pruned-4-Layers_LoRA-PEFT
Text Generation • 7B • Updated • 11
Medical Whisper
Fine-tuned Whisper Large v3 on Doctor/ Patient consultations.
Pruned MoEs (Mixtral-8x7B-Instruct-v0.1)
Mixtral expert-pruning checkpoints. NeurIPS 2026 AXIOM Workshop: https://arxiv.org/abs/2608.07890. Builds on Chowdhury et al. (ICML 2024).
-
mistralai/Mixtral-8x7B-Instruct-v0.1
47B • Updated • 215k • 4.76k -
Na0s/Mixtral-8x7B-Instruct-v0.1-LoRA-on-Gates
Text Generation • 47B • Updated • 25 • 1 -
Na0s/Mixtral-8x7B-Instruct-v0.1-exhaustive-LoRA
Text Generation • 47B • Updated • 14 -
Na0s/Mixtral-8x7B-v0.1-instruct-pruned-random-1-experts
Text Generation • 41B • Updated • 20