This collection is provided for reproducibility of the paper's main claim
Andrey PRO
Bochkov
AI & ML interests
None yet
Recent Activity
published an article about 9 hours ago
Token Identity Is Not Meaning: What Fixed-Input Language Models Can Teach Us updated a Space about 12 hours ago
BEMSH-BVV/README published a Space about 12 hours ago
BEMSH-BVV/READMEOrganizations
Language Models Without a Trainable Input Embedding Table
This collection is provided for reproducibility of the paper's main claim
-
Bochkov/llm-fix-min-baseline-learned-input-table-model-classic
Text Generation • 0.5B • Updated • 316 -
Bochkov/llm-fix-min-fixed-minimal-binary-code
Text Generation • 0.5B • Updated • 321 -
Bochkov/llm-fix-min-affine-recoded-minimal-code-table-free
Text Generation • 0.5B • Updated • 307 -
Language Models Without a Trainable Input Embedding Table: Learning from Fixed Minimal Binary Token Codes
Paper • 2605.09751 • Published
Growing Transformers:Layer-wise Expansion Comparative Study
Paper: 2507.07129 'Growing Transformers: Modular Composition and Layer-wise Expansion on a Frozen Substrate' (4.2.2, 5.2. Results)
-
Bochkov/growing-transformers-model-16-bit-1-9-181m
Text Generation • 0.2B • Updated • 158 -
Bochkov/growing-transformers-model-unicode-1-9-247m
Text Generation • 0.2B • Updated • 29 -
Bochkov/growing-transformers-model-unfrozen-1-9-247m
Text Generation • 0.2B • Updated • 145 -
Bochkov/growing-transformers-model-frozen-16-bit-baseline-monolyth-181m
Text Generation • 0.2B • Updated • 25
Do Language Models Need a Trainable Input Embedding Table?
This collection is provided for reproducibility of the paper's main claim
-
Bochkov/ab_ext_learned
Text Generation • 2B • Updated • 448 -
Bochkov/ab_ext_binary16
Text Generation • 2B • Updated • 416 -
Bochkov/ab_ext_gf2
Text Generation • 2B • Updated • 375 -
Do Language Models Need a Trainable Input Embedding Table? Fixed Minimal Token Codes at 1.7B-Class Scale
Paper • 2610.04002 • Published • 3
Emergent Semantics Beyond Token Embeddings
Paper: 2507.04886 (TMLR, Oct 2025). 'Emergent Semantics Beyond Token Embeddings: Transformer LMs with Frozen Visual Unicode Representations'
-
Bochkov/emergent-semantics-model-uni-glyph-335m
Text Generation • 0.3B • Updated • 20 -
Bochkov/emergent-semantics-model-unfrozen-335m
Text Generation • 0.3B • Updated • 18 -
Bochkov/emergent-semantics-model-16-bit-269m
Text Generation • 0.3B • Updated • 21 • 1 -
Bochkov/emergent-semantics-model-64-bit-272m
Text Generation • 0.3B • Updated • 130
Tokenizers
This collection features frozen, precomputed token embedding tensors designed for experimentation with semantic emergence in language models.
Beyond the Parameter Monolith: Modular Language Modeling
This collection is provided for reproducibility of the paper's main claim
Do Language Models Need a Trainable Input Embedding Table?
This collection is provided for reproducibility of the paper's main claim
-
Bochkov/ab_ext_learned
Text Generation • 2B • Updated • 448 -
Bochkov/ab_ext_binary16
Text Generation • 2B • Updated • 416 -
Bochkov/ab_ext_gf2
Text Generation • 2B • Updated • 375 -
Do Language Models Need a Trainable Input Embedding Table? Fixed Minimal Token Codes at 1.7B-Class Scale
Paper • 2610.04002 • Published • 3
Language Models Without a Trainable Input Embedding Table
This collection is provided for reproducibility of the paper's main claim
-
Bochkov/llm-fix-min-baseline-learned-input-table-model-classic
Text Generation • 0.5B • Updated • 316 -
Bochkov/llm-fix-min-fixed-minimal-binary-code
Text Generation • 0.5B • Updated • 321 -
Bochkov/llm-fix-min-affine-recoded-minimal-code-table-free
Text Generation • 0.5B • Updated • 307 -
Language Models Without a Trainable Input Embedding Table: Learning from Fixed Minimal Binary Token Codes
Paper • 2605.09751 • Published
Emergent Semantics Beyond Token Embeddings
Paper: 2507.04886 (TMLR, Oct 2025). 'Emergent Semantics Beyond Token Embeddings: Transformer LMs with Frozen Visual Unicode Representations'
-
Bochkov/emergent-semantics-model-uni-glyph-335m
Text Generation • 0.3B • Updated • 20 -
Bochkov/emergent-semantics-model-unfrozen-335m
Text Generation • 0.3B • Updated • 18 -
Bochkov/emergent-semantics-model-16-bit-269m
Text Generation • 0.3B • Updated • 21 • 1 -
Bochkov/emergent-semantics-model-64-bit-272m
Text Generation • 0.3B • Updated • 130
Growing Transformers:Layer-wise Expansion Comparative Study
Paper: 2507.07129 'Growing Transformers: Modular Composition and Layer-wise Expansion on a Frozen Substrate' (4.2.2, 5.2. Results)
-
Bochkov/growing-transformers-model-16-bit-1-9-181m
Text Generation • 0.2B • Updated • 158 -
Bochkov/growing-transformers-model-unicode-1-9-247m
Text Generation • 0.2B • Updated • 29 -
Bochkov/growing-transformers-model-unfrozen-1-9-247m
Text Generation • 0.2B • Updated • 145 -
Bochkov/growing-transformers-model-frozen-16-bit-baseline-monolyth-181m
Text Generation • 0.2B • Updated • 25
Tokenizers
This collection features frozen, precomputed token embedding tensors designed for experimentation with semantic emergence in language models.