Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
🔄
In a Training Loop
8.4
TFLOPS
Mr Munk
GODELEV
51
4
62
Follow
Quazim0t0's profile picture
preetsojitra's profile picture
blackbook-lm's profile picture
26 followers
·
29 following
AI & ML interests
High schooler by day, LLM builder by night. Driven by a deep love for both Physics and AI. Currently spending my runtime building on Hugging Face, experimenting with transformer architectures, and training custom LLMs.
Recent Activity
reacted
to
KlondikeDev
's
post
with 👀
about 1 hour ago
Boris-2 coming soon! The Models: Boris-2-75M: Trained on 26B tokens -- estimated to start training on August 12th. Boris-2-125M: Trained on 90B tokens -- Estimated to start training on August 20th. Boris-2-250M: Trained on 60B tokens -- Estimated to start training on September 10th. Why does 125M get more tokens than 250M? Well, the straight answer is time. It saves time, while still allowing the 250M model to exceed the 125M model. Furthermore, we are attempting a unique architecture and layering scheme to hopefully end up around the strength of SmolLM2-135M. Fingers crossed! We hope to end up in the ballpark of https://huggingface.co/AxiomicLabs/GPT-X2.5-135M or https://huggingface.co/BananaMind/BananaMind-2-Pro-Preview Following this, we will release the Pro, Instruct and Pro-Instruct variants. More info will be coming soon!
replied
to
Banaxi-Tech
's
post
about 2 hours ago
We're exited to announce BananaMind OS, our OS specically for running BananaMind models! Its able to run BananaMind 2 Nano at 4 bit on only 7-8MB of ram, the 2 bit on 6MB of ram and the 8 bit version on 14MB of RAM! It runs on a 486 or newer! Check out this video and image running BananaMind 2 Nano 4 Bit on 9 MB of RAM and a emulated 486 in QEMU at ~1TPS! We asked it: "What is the first letter of the alphabet?" The response is: "The first letter of the alphabet is: - A. " And if you're asking because of the video, yes I am a arch btw. Comment and like this post for a GitHub link and comment for adding other models!
replied
to
Banaxi-Tech
's
post
about 3 hours ago
We're exited to announce BananaMind OS, our OS specically for running BananaMind models! Its able to run BananaMind 2 Nano at 4 bit on only 7-8MB of ram, the 2 bit on 6MB of ram and the 8 bit version on 14MB of RAM! It runs on a 486 or newer! Check out this video and image running BananaMind 2 Nano 4 Bit on 9 MB of RAM and a emulated 486 in QEMU at ~1TPS! We asked it: "What is the first letter of the alphabet?" The response is: "The first letter of the alphabet is: - A. " And if you're asking because of the video, yes I am a arch btw. Comment and like this post for a GitHub link and comment for adding other models!
View all activity
Organizations
GODELEV
's models
14
Sort: Recently updated
GODELEV/Rose-Medium
Text Generation
•
97.8M
•
Updated
1 day ago
•
416
•
9
GODELEV/Rose-Mini
Text Generation
•
49.4M
•
Updated
8 days ago
•
623
•
13
GODELEV/Archaea-74M
Text Generation
•
74M
•
Updated
8 days ago
•
278
•
5
GODELEV/Archaea-74M-V1.1
Text Generation
•
74M
•
Updated
Jun 23
•
193
•
3
GODELEV/Exp-1
9.85M
•
Updated
Jun 20
•
55
•
1
GODELEV/TOK-16K
Updated
Jun 16
GODELEV/TOK-8K
Updated
Jun 14
GODELEV/TOK-32K
Updated
Jun 12
GODELEV/Ant-10M
Text Generation
•
9.9M
•
Updated
Jun 11
•
49
•
4
GODELEV/TOK-4K
Updated
Jun 7
GODELEV/Ant-5M
Text Generation
•
4.71M
•
Updated
Jun 6
•
85
•
2
GODELEV/Test-1-4000
Text Generation
•
0.2B
•
Updated
May 9
•
41
GODELEV/Test-1-3000
Text Generation
•
0.2B
•
Updated
May 9
•
11
GODELEV/Test-1-2000
Text Generation
•
0.2B
•
Updated
May 8
•
9