Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
🔄
In a Training Loop
The Supreme GGUF-SafeTensors Guy
GGUFGuy
35
2
8
Follow
RexTRO111's profile picture
Harley-ml's profile picture
Dorman11's profile picture
13 followers
·
70 following
AI & ML interests
I like llms
Recent Activity
new
activity
33 minutes ago
CaptiveDreamer/CaraArchive:
🚩 Report: Other
replied
to
Banaxi-Tech
's
post
about 5 hours ago
We're releasing BananaMind 2.1 Unified, a 35M three-tower relay model where the two output towers can only talk to each other through a silent middle tower that has no output head and no loss term. Tower B trains entirely on indirect gradient. It was never told what to predict. It woke up anyway. Jacobian lens shows it carrying the correct answer ("Paris", "oxygen", "blue") at its deepest layer. Its bridge gates grew 4-47x from init. Feed it from only one side and the representations collapse to junk — it needs both outer towers to become semantic. The PIQA result is the cleanest demonstration: Tower A alone scores 50.11 (chance). Tower C alone 52.12. Full system 61.75. All physical reasoning lives in the integration. The 35M three-tower beats the 50M single-tower BananaMind 2 Medium on PIQA. Trained on 38B tokens in ~7.5 hours on 8x RTX PRO 6000. Ships with 7 ablation modes so you can surgically cut the model apart without retraining. Full training logs, J-lens fits, and eval outputs for every mode included. This is a research model. It is interesting because you can take it apart. https://huggingface.co/BananaMind/BananaMind-2.1-Unified Follow us for more: https://huggingface.co/BananaMind @Banaxi-Tech
updated
a model
about 6 hours ago
GGUFGuy/hgbn
View all activity
Organizations
GGUFGuy
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
liked
a model
6 days ago
Qwen/Qwen3.8-27B
Image-Text-to-Text
•
28B
•
Updated
6 days ago
•
1.37M
•
•
11.6k
liked
a model
7 days ago
kefir090/Lumen-118M-Base
Text Generation
•
0.1B
•
Updated
about 24 hours ago
•
1.46k
•
5
liked
2 models
8 days ago
black-forest-labs/FLUX.1-dev
Text-to-Image
•
12B
•
Updated
Jun 27, 2025
•
607k
•
•
14.2k
unsloth-jobs/why-can-i-make-this
Updated
8 days ago
•
1
liked
a Space
15 days ago
Running
on
Zero
Agents
221
Wan2.2 I2V OmniBlink
😉
221
Wan2.2 I2V Blink Effect
liked
a dataset
about 1 month ago
armand0e/claude-fable-5-claude-code
Traces
•
Updated
Jun 19
•
63
•
4.24k
•
361
liked
2 models
about 1 month ago
empero-ai/Qwythos-9B-v2-GGUF
Image-Text-to-Text
•
9B
•
Updated
Jul 12
•
544k
•
251
prism-ml/Bonsai-27B-gguf
Text Generation
•
27B
•
Updated
Jul 17
•
1.18M
•
788