modrill commited on
Commit
e4c12ff
·
verified ·
1 Parent(s): d5b575c

Date unify to 20260908; formerly code-think-o7b-20260909

Browse files
Files changed (1) hide show
  1. README.md +3 -3
README.md CHANGED
@@ -10,7 +10,7 @@ pipeline_tag: text-generation
10
  library_name: transformers
11
  ---
12
 
13
- # code-think-o7b-20260909
14
 
15
  Research checkpoint: **allenai/Olmo-3-1025-7B** (`a81bae42db3975be1671e27b9c9a56da1a9f980f`) after V4 LoRA SFT on Qwen3-30B-A3B-Thinking-2507 traces (OLMo-tokenized V4 payload), then merged to full weights.
16
 
@@ -50,7 +50,7 @@ Merged weights are the 2-epoch endpoint (`step-000904`, 63,378,772 assistant tok
50
 
51
  | Model | pass@1 | Cap | Notes |
52
  |---|---:|---:|---|
53
- | **code-think-o7b-20260909** | **56/256** | **149** | this repo; seed 3407 |
54
  | Olmo-3-1025-7B (same contract, think) | 15/256 | 104 | bare base, seed 3407 |
55
 
56
  Single seed. These are research checkpoints, not product scores.
@@ -62,7 +62,7 @@ Merged full weights; no PEFT required at inference. The pinned OLMo template sup
62
  ```python
63
  from transformers import AutoModelForCausalLM, AutoTokenizer
64
 
65
- repo = "modrill/code-think-o7b-20260909"
66
  tokenizer = AutoTokenizer.from_pretrained(repo)
67
  model = AutoModelForCausalLM.from_pretrained(
68
  repo, torch_dtype="bfloat16", device_map="auto"
 
10
  library_name: transformers
11
  ---
12
 
13
+ # code-think-o7b-20260908
14
 
15
  Research checkpoint: **allenai/Olmo-3-1025-7B** (`a81bae42db3975be1671e27b9c9a56da1a9f980f`) after V4 LoRA SFT on Qwen3-30B-A3B-Thinking-2507 traces (OLMo-tokenized V4 payload), then merged to full weights.
16
 
 
50
 
51
  | Model | pass@1 | Cap | Notes |
52
  |---|---:|---:|---|
53
+ | **code-think-o7b-20260908** | **56/256** | **149** | this repo; seed 3407 |
54
  | Olmo-3-1025-7B (same contract, think) | 15/256 | 104 | bare base, seed 3407 |
55
 
56
  Single seed. These are research checkpoints, not product scores.
 
62
  ```python
63
  from transformers import AutoModelForCausalLM, AutoTokenizer
64
 
65
+ repo = "modrill/code-think-o7b-20260908"
66
  tokenizer = AutoTokenizer.from_pretrained(repo)
67
  model = AutoModelForCausalLM.from_pretrained(
68
  repo, torch_dtype="bfloat16", device_map="auto"