SDPO under Continual Learning
Meng Wang
Moenupa
AI & ML interests
MLLM Post-Training & Alignment
Recent Activity
updated a model 5 minutes ago
MLLM-CL/Qwen3-4B-Instruct-2507-SFT-MathChemTool updated a model 9 minutes ago
MLLM-CL/Qwen3-4B-Instruct-2507-SFT-MathChem updated a model 12 minutes ago
MLLM-CL/Qwen3-4B-Instruct-2507-SFT-Math