T1-Mini-Preview / README.md
Bc-AI's picture
Update README.md
26e9f87 verified
|
Raw History Blame Contribute Delete
2.01 kB
---
language:
- en
license: apache-2.0
base_model: Qwen/Qwen3.5-4B
datasets:
- Bc-AI/SFT-Ultra
tags:
- full-fine-tune
- sft
- dpo
- qwen3.5
- smilyai
library_name: transformers
pipeline_tag: image-text-to-text
---
# SmilyAI Labs T1-Mini-Preview
**T1-Mini-Preview** is an early preview of the upcoming **T1-Mini** model from **SmilyAI Labs**.
T1-Mini-Preview is based on **Qwen/Qwen3.5-4B** and was fully fine-tuned on **Bc-AI/SFT-Ultra**, a curated instruction-tuning dataset prepared for SmilyAI's small-model research.
## Training
The model was trained in two main stages:
1. **Supervised Fine-Tuning (SFT)** β€” trained the base model on our curated instruction and reasoning data.
2. **Direct Preference Optimization (DPO)** β€” further refined the model's responses using preference-based training.
This two-stage pipeline was designed to improve instruction following, response quality, and overall conversational behavior while keeping the model relatively small and efficient.
## Model Status
**T1-Mini-Preview is a preview release**, not the final T1-Mini model. Training, evaluation, and further refinement are still ongoing.
We are releasing this version so the community can experiment with it and provide feedback while development continues.
## Base Model
- **Base:** Qwen/Qwen3.5-4B
- **Training:** Full fine-tuning
- **Primary language:** English
- **Fine-tuning dataset:** Bc-AI/SFT-Ultra
- **Training stages:** SFT β†’ DPO
- **License:** Apache 2.0
## About SmilyAI Labs
**SmilyAI Labs** is a small open-source AI project focused on building capable, efficient, and accessible AI models.
We're experimenting with smaller models that can deliver strong performance without requiring enormous amounts of compute.
πŸš€ **T1-Mini-Preview is one step toward that goal.**
## Disclaimer
This is an experimental preview model. Its behavior and capabilities may differ from the final T1-Mini release, and it may occasionally produce incorrect, inconsistent, or undesirable outputs.