-
Na0s/Llama-3.2-3B-Instruct-Medical-Chatbot-LoRA-FT
Text Generation • 3B • Updated • 25 • 1 -
Na0s/Llama-3.2-3B-Medical-Chatbot-LoRA-FT
Text Generation • 3B • Updated • 24 • 1 -
meta-llama/Llama-3.2-3B-Instruct
Text Generation • 3B • Updated • 1.49M • • 2.76k -
meta-llama/Llama-3.2-3B
Text Generation • 3B • Updated • 281k • • 1.01k
Ali Janati
Na0s
AI & ML interests
NLP, Speech Recognition, Computer Vision, Time Series Forecasting.
Recent Activity
published a model 10 days ago
frisson-labs/Faynt-75M-Arena published a model 10 days ago
frisson-labs/Faynt-10M-Arena published a model 10 days ago
frisson-labs/Faynt-75M-ExpertOrganizations
Depth pruned and fine tuned Llama-3.1-8B
-
Na0s/Llama-3.1-8B-Pruned-4-Layers_LoRA-PEFT-3.0
Text Generation • 7B • Updated • 45 • 4 -
Na0s/Llama-3.1-8B-Pruned-4-Layers_LoRA-PEFT-2.0
Text Generation • 7B • Updated • 28 • 1 -
Na0s/Llama-3.1-8B-Pruned-4-Layers_LoRA-PEFT-1.0
Text Generation • 7B • Updated • 19 -
Na0s/Llama-3.1-8B-Pruned-4-Layers_LoRA-PEFT
Text Generation • 7B • Updated • 18
Pruned MoEs (Mixtral-8x7B-Instruct-v0.1)
Mixtral expert-pruning checkpoints. NeurIPS 2026 AXIOM Workshop: https://arxiv.org/abs/2608.07890. Builds on Chowdhury et al. (ICML 2024).
-
mistralai/Mixtral-8x7B-Instruct-v0.1
47B • Updated • 216k • 4.76k -
Na0s/Mixtral-8x7B-Instruct-v0.1-LoRA-on-Gates
Text Generation • 47B • Updated • 35 • 1 -
Na0s/Mixtral-8x7B-Instruct-v0.1-exhaustive-LoRA
Text Generation • 47B • Updated • 22 -
Na0s/Mixtral-8x7B-v0.1-instruct-pruned-random-1-experts
Text Generation • 41B • Updated • 23
Medical Chatbot
-
Na0s/Llama-3.2-3B-Instruct-Medical-Chatbot-LoRA-FT
Text Generation • 3B • Updated • 25 • 1 -
Na0s/Llama-3.2-3B-Medical-Chatbot-LoRA-FT
Text Generation • 3B • Updated • 24 • 1 -
meta-llama/Llama-3.2-3B-Instruct
Text Generation • 3B • Updated • 1.49M • • 2.76k -
meta-llama/Llama-3.2-3B
Text Generation • 3B • Updated • 281k • • 1.01k
Differential transformers
Fine-tuning foundation Llama-3.2-3B-Instruct on medical Q&A using differential attention (In progress). Paper: https://arxiv.org/pdf/2410.05258
Depth pruned and fine tuned Llama-3.1-8B
-
Na0s/Llama-3.1-8B-Pruned-4-Layers_LoRA-PEFT-3.0
Text Generation • 7B • Updated • 45 • 4 -
Na0s/Llama-3.1-8B-Pruned-4-Layers_LoRA-PEFT-2.0
Text Generation • 7B • Updated • 28 • 1 -
Na0s/Llama-3.1-8B-Pruned-4-Layers_LoRA-PEFT-1.0
Text Generation • 7B • Updated • 19 -
Na0s/Llama-3.1-8B-Pruned-4-Layers_LoRA-PEFT
Text Generation • 7B • Updated • 18
Medical Whisper
Fine-tuned Whisper Large v3 on Doctor/ Patient consultations.
Pruned MoEs (Mixtral-8x7B-Instruct-v0.1)
Mixtral expert-pruning checkpoints. NeurIPS 2026 AXIOM Workshop: https://arxiv.org/abs/2608.07890. Builds on Chowdhury et al. (ICML 2024).
-
mistralai/Mixtral-8x7B-Instruct-v0.1
47B • Updated • 216k • 4.76k -
Na0s/Mixtral-8x7B-Instruct-v0.1-LoRA-on-Gates
Text Generation • 47B • Updated • 35 • 1 -
Na0s/Mixtral-8x7B-Instruct-v0.1-exhaustive-LoRA
Text Generation • 47B • Updated • 22 -
Na0s/Mixtral-8x7B-v0.1-instruct-pruned-random-1-experts
Text Generation • 41B • Updated • 23