Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
🦀
dancing crab
26.6
TFLOPS
Damien
el4
11
7
22
Follow
indorilml's profile picture
BermiTechs's profile picture
SeaWolf-AI's profile picture
5 followers
·
20 following
https://codeberg.org/el_4
AI & ML interests
I do anything really | #quantwizard | ETH: 0xDEE7fa8C421BD038D32e4441ea1aDe72fE973706
Recent Activity
liked
a Space
1 day ago
FINAL-Bench/model-galaxy
reacted
to
SeaWolf-AI
's
post
with 🔥
2 days ago
A small gift for anyone building or studying foundation models. Most "open" models hand you the weights and stop there. With Aether-7B-5Attn we wanted to hand over the whole thing — so you can actually learn from it, reproduce it, and build on it: the data recipe, the training code, every hyperparameter, the complete logs, and the intermediate checkpoints. All Apache-2.0, reproducible byte-for-byte. What you can do with it: 🔁 Rebuild it from scratch, or fork the recipe for your own model 🔬 Study a real heterogeneous-attention MoE — 49 layers place 5 attention mechanisms on a 7×7 Latin square, arranged as a clean, attributable ablation 📈 Trace training dynamics across the released checkpoints (110k / 115k / 162k) It's a modest 6.59B model, and an honest one — the limitations (no KV-cache in this build, small scale) are written right in the card. We're not claiming it's special. If any piece of it saves you time or teaches you something, that's exactly what we hoped for. 🤗 📖 Full write-up → [blog] · https://huggingface.co/blog/FINAL-Bench/opensource-llm 📦 5 Attention Base · https://huggingface.co/FINAL-Bench/Aether-7B-5Attn 🎯 5 Attention Instruct · https://huggingface.co/FINAL-Bench/Aether-7B-5Attn-it 🚀 5 Attention Live demo · https://huggingface.co/spaces/FINAL-Bench/Aether-Sovereign-AI 📦 7 Attention Base · https://huggingface.co/FINAL-Bench/Aether-7B-7Attn-base 📦 11 Attention Base · https://huggingface.co/FINAL-Bench/Aether-6B-11Attn-base 🧬 Collection · https://huggingface.co/collections/FINAL-Bench/aether-foundation-model #opensource #LLM #MoE #reproducibility #Apache2
upvoted
an
article
3 days ago
Aether-7B-5Attn: A 100% Open-Source Sovereign Foundation Model — and a Controlled Experiment in Heterogeneous Attention
View all activity
Organizations
None yet
el4
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
liked
a Space
1 day ago
Running
Agents
19
Model Galaxy
🌌
19
Darwin family + 2026 trending models on the HF galaxy
liked
a dataset
4 days ago
AlexWortega/karp-autoresearch-distill
Updated
Jun 14
•
17
•
2
liked
a model
4 days ago
el4/Xenon-26B-A4B-OPAL-GGUF
Text Generation
•
25B
•
Updated
4 days ago
•
64
•
2
liked
2 datasets
4 days ago
greghavens/gpt-5.6-sol-coding-and-debugging-traces
Preview
•
Updated
4 days ago
•
1.45k
•
12
greghavens/glm-5.2-coding-and-debugging-traces
Viewer
•
Updated
5 days ago
•
1.63k
•
470
•
6
liked
a dataset
5 days ago
greghavens/kimi-k3-coding-and-debugging-traces
Viewer
•
Updated
about 3 hours ago
•
2.3k
•
82
•
23
liked
2 models
5 days ago
mradermacher/model_requests
Updated
Jun 6
•
187
el4/Xenon-26B-A4B
Image-Text-to-Text
•
Updated
5 days ago
•
67
•
4
liked
2 datasets
5 days ago
greghavens/fable-5-coding-and-debugging-traces
Updated
29 minutes ago
•
1.35k
•
16
el4/xenon-mixed-agentic-datasetv1
Viewer
•
Updated
5 days ago
•
4.33k
•
48
•
1
liked
a model
7 days ago
unsloth/inkling-GGUF
Image-Text-to-Text
•
947B
•
Updated
7 days ago
•
7.38k
•
120
liked
a model
8 days ago
prism-ml/Ternary-Bonsai-27B-gguf
Text Generation
•
4B
•
Updated
5 days ago
•
432k
•
•
952
liked
a model
11 days ago
InternScience/Agents-A1
Text Generation
•
35B
•
Updated
8 days ago
•
36.6k
•
601
liked
a model
16 days ago
el4/Qwopus3.6-35B-A3B-Coder-OPAL-GGUF
Text Generation
•
36B
•
Updated
14 days ago
•
15.3k
•
4
liked
2 models
17 days ago
el4/Darwin-36B-Opus-OPAL-GGUF
Text Generation
•
35B
•
Updated
14 days ago
•
11.9k
•
11
FINAL-Bench/Darwin-36B-Opus
Text Generation
•
35B
•
Updated
27 minutes ago
•
400
•
87
liked
a Space
17 days ago
Running
2
OPAL
💻
2
Optimized Precision for Adaptive LLMs
liked
a model
17 days ago
el4/SIQ-1-35B-OPAL-GGUF
Text Generation
•
35B
•
Updated
14 days ago
•
16.5k
•
1
liked
a dataset
17 days ago
lemon07r/bartowski-imatrix-v5-semantic
Viewer
•
Updated
Feb 3
•
2.08k
•
158
•
5
liked
a model
25 days ago
el4/SIQ-1-35B-APEX-GGUF
Text Generation
•
35B
•
Updated
17 days ago
•
5.87k
•
3
Load more