Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
🔄
In a Training Loop
8.4
TFLOPS
Mr Munk
GODELEV
51
4
70
Follow
jaithree's profile picture
HannesVonEssen's profile picture
Jdudeo's profile picture
28 followers
·
29 following
AI & ML interests
High schooler by day, LLM builder by night. Driven by a deep love for both Physics and AI. Currently spending my runtime building on Hugging Face, experimenting with transformer architectures, and training custom LLMs.
Recent Activity
liked
a Space
about 3 hours ago
hugging-apps/charles-the-chess-bot
replied
to
Banaxi-Tech
's
post
about 5 hours ago
We're excited to release BananaMind 2 Micro, our smallest model yet. It fits a compact architecture in only 2.9M parameters achieving the highest parameter efficiency on BananaMind Base Bench against comparable models. It achieves comparable performance to GPT S2 5M and GPT S 5M at almost half the size while beating CMA 1M Mini. BananaMind 2 Micro achieved the #1 spot on the Open SLM Leaderboard for the sub 3M category (not added yet but it achieves #1) For the training we used Muon + the XSA Refresh Gate with a 5e-2 lr for Muon and 4e-3 for the 1D weights. Its score on our efficiency measure is 0.326 getting the first place with Syn 2.6M on the second place scoring 0.291 and GPT S 5M at 0.235* Check it out at https://huggingface.co/BananaMind/BananaMind-2-Micro and follow us at: @vovaRL @DedeProGames @Banaxi-Tech https://huggingface.co/BananaMind Our new releases aren't stopping 🚀 August 13-14 BananaMind 2 Pro
replied
to
Banaxi-Tech
's
post
about 5 hours ago
We're excited to release BananaMind 2 Micro, our smallest model yet. It fits a compact architecture in only 2.9M parameters achieving the highest parameter efficiency on BananaMind Base Bench against comparable models. It achieves comparable performance to GPT S2 5M and GPT S 5M at almost half the size while beating CMA 1M Mini. BananaMind 2 Micro achieved the #1 spot on the Open SLM Leaderboard for the sub 3M category (not added yet but it achieves #1) For the training we used Muon + the XSA Refresh Gate with a 5e-2 lr for Muon and 4e-3 for the 1D weights. Its score on our efficiency measure is 0.326 getting the first place with Syn 2.6M on the second place scoring 0.291 and GPT S 5M at 0.235* Check it out at https://huggingface.co/BananaMind/BananaMind-2-Micro and follow us at: @vovaRL @DedeProGames @Banaxi-Tech https://huggingface.co/BananaMind Our new releases aren't stopping 🚀 August 13-14 BananaMind 2 Pro
View all activity
Organizations
GODELEV
's datasets
9
Sort: Recently updated
GODELEV/Arithmetic-XL
Viewer
•
Updated
17 days ago
•
12M
•
109
•
3
GODELEV/BetterDataset-12M
Viewer
•
Updated
24 days ago
•
180k
•
652
•
1
GODELEV/D1-8-Lite
Viewer
•
Updated
Jun 14
•
6.3M
•
70
GODELEV/D1-32-Lite
Viewer
•
Updated
Jun 13
•
5.66M
•
52
GODELEV/Arithmetic
Viewer
•
Updated
Jun 5
•
96.1k
•
74
GODELEV/BetterDataset-2M
Viewer
•
Updated
Jun 1
•
2M
•
371
•
4
GODELEV/Trying-MakingBetterDataset-100K
Preview
•
Updated
May 8
•
16
GODELEV/Tiny-Stories-1500
Viewer
•
Updated
Dec 21, 2025
•
1.5k
•
15
GODELEV/Kishor_V2_53K_LLM_Prompt-Response_Pairs
Viewer
•
Updated
Jul 8, 2025
•
53k
•
12