Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
🔄
tired of nightshade?
24.3
TFLOPS
Jackson Yeskie
8BitStudio
11
1
20
Follow
Nitesh45433's profile picture
StarkEater's profile picture
XBFDI's profile picture
5 followers
·
8 following
8-bitStudio
AI & ML interests
Text to image, Language models
Recent Activity
reacted
to
harshitkgupta
's
post
with 🚀
about 2 hours ago
Fine-tuned Qwen 2.5 (0.5B → 3B) on real coding-agent traces, 10 controlled runs, one 16GB Mac. Compared PyTorch MPS vs. Apple MLX for local LoRA SFT — and the honest answer is "it depends on what you're optimizing for": • PyTorch MPS: 2.2x–5.7x faster raw throughput, but hits a hard memory wall — can't load a 3B model in FP16 on 16GB. • Apple MLX: 4-bit QLoRA fits 3B+ models with almost flat memory scaling as context grows (+109 MB going from 1k→4k tokens). • 4-bit quantization doesn't cost you convergence — eval loss tracks closely across backends. • The bigger surprise: most of MLX's slowdown isn't the 4-bit dequant tax. Two of the 10 runs went unquantized to isolate it — dequant only explains 1.07x–1.4x of the gap. A ~4.1–4.6x framework-level gap remains either way. All 10 LoRA adapters + Trackio logs are public so the numbers are checkable, not just claimed. Full writeup: https://huggingface.co/blog/harshitkgupta/fine-tuning-coding-agents-on-mac-pytorch-mps-mlx
new
activity
about 6 hours ago
hugging-science/ripp:
Hugging Science is DEAD???
replied
to
Banaxi-Tech
's
post
2 days ago
We're excited to release BananaAll, our SLM Super App. It allows you to do EVERYTHING you need to do to trains SLMs in a single app, no terminal, no 30 chrome tabs. The train tab allows you to train models, select datasets from presets, and use other ones with auto mapping, model size slider, it automatically generates a training script for you. Then after you've trained the model or want to compare it to competitors, the evaluation tab, run ARC EASY, ARC Challenge, Hellaswag, PIQA, Arithmark 3, BananaMind Base Bench and more! Simple Results screen. And lastly the inference tab, run your trained models or others. Normally you would need seperate apps or scripts for that, but the BananaAll Super App lets you do all of that in a single app. We also trained a small 2.5M parameter model on 200M tokens of Fineweb edu, The results: BananaMind Base Bench 854 and 53% on PIQA. On only 200M tokens. Check it out at https://github.com/BananaMind/BananaAll.
View all activity
Organizations
None yet
models
2
Sort: Recently updated
8BitStudio/Aniimage-2
Text-to-Image
•
0.4B
•
Updated
14 days ago
•
199
•
5
8BitStudio/Aniimage-1
Text-to-Image
•
0.4B
•
Updated
18 days ago
•
130
•
2
datasets
2
Sort: Recently updated
8BitStudio/Roblox-luau-coding_L1
Viewer
•
Updated
Aug 1
•
12.3k
•
389
•
3
8BitStudio/TowerOfBabel
Viewer
•
Updated
Jul 27
•
274k
•
56