Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
🔄
In a Training Loop
8.4
TFLOPS
Mr Munk
GODELEV
57
5
83
Follow
amadadai's profile picture
StentorLabs's profile picture
maxcurrent's profile picture
31 followers
·
31 following
AI & ML interests
High schooler by day, LLM builder by night. Driven by a deep love for both Physics and AI. Currently spending my runtime building on Hugging Face, experimenting with transformer architectures, and training custom LLMs.
Recent Activity
liked
a model
2 minutes ago
BananaMind/BananaMind-2.1-Unified
reacted
to
Banaxi-Tech
's
post
with 🔥
3 minutes ago
We're releasing BananaMind 2.1 Unified, a 35M three-tower relay model where the two output towers can only talk to each other through a silent middle tower that has no output head and no loss term. Tower B trains entirely on indirect gradient. It was never told what to predict. It woke up anyway. Jacobian lens shows it carrying the correct answer ("Paris", "oxygen", "blue") at its deepest layer. Its bridge gates grew 4-47x from init. Feed it from only one side and the representations collapse to junk — it needs both outer towers to become semantic. The PIQA result is the cleanest demonstration: Tower A alone scores 50.11 (chance). Tower C alone 52.12. Full system 61.75. All physical reasoning lives in the integration. The 35M three-tower beats the 50M single-tower BananaMind 2 Medium on PIQA. Trained on 38B tokens in ~7.5 hours on 8x RTX PRO 6000. Ships with 7 ablation modes so you can surgically cut the model apart without retraining. Full training logs, J-lens fits, and eval outputs for every mode included. This is a research model. It is interesting because you can take it apart. https://huggingface.co/BananaMind/BananaMind-2.1-Unified Follow us for more: https://huggingface.co/BananaMind @Banaxi-Tech
liked
a model
about 4 hours ago
SupraLabs/Supra2-Medium-Base
View all activity
Organizations
GODELEV
's datasets
9
Sort: Recently updated
GODELEV/Arithmetic-XL
Viewer
•
Updated
27 days ago
•
12M
•
124
•
4
GODELEV/BetterDataset-12M
Viewer
•
Updated
Jul 17
•
180k
•
624
•
2
GODELEV/D1-8-Lite
Viewer
•
Updated
Jun 14
•
6.3M
•
38
GODELEV/D1-32-Lite
Viewer
•
Updated
Jun 13
•
5.66M
•
104
GODELEV/Arithmetic
Viewer
•
Updated
Jun 5
•
96.1k
•
70
•
1
GODELEV/BetterDataset-2M
Viewer
•
Updated
Jun 1
•
2M
•
261
•
4
GODELEV/Trying-MakingBetterDataset-100K
Preview
•
Updated
May 8
•
14
GODELEV/Tiny-Stories-1500
Viewer
•
Updated
Dec 21, 2025
•
1.5k
•
14
GODELEV/Kishor_V2_53K_LLM_Prompt-Response_Pairs
Viewer
•
Updated
Jul 8, 2025
•
53k
•
12