Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
61.4
TFLOPS
Ron
ronedgecomb
6
48
494
Follow
ariG23498's profile picture
Gargaz's profile picture
PhysiQuanty's profile picture
16 followers
·
100 following
https://ronedgecomb.site/
ronedgecomb
ronedgecomb
ronedgecomb.bsky.social
AI & ML interests
EVERYTHING TECHNOLOGY
Recent Activity
liked
a model
about 4 hours ago
distilbert/distilbert-base-uncased
reacted
to
owensong
's
post
with 🔥
about 5 hours ago
I just released Inflect v2, pushing complete TTS below 4M parameters. Inflect-Nano-v2 fits every learned inference component, including waveform generation, into 3.96M parameters (15.97 MB FP32), making it one of the smallest publicly available complete TTS models. Inflect-Micro-v2 stays under 10M at 9.36M parameters (37.53 MB), using the extra capacity for clearer, more natural speech. This was not about building a toy that technically produces audio. I wanted to see how much useful TTS could actually fit at this scale. Nano stayed surprisingly close to Micro across the reported release metrics, but if you want to hear Inflect at its best, start with Micro: https://huggingface.co/owensong/Inflect-Micro-v2 Text goes in and 24 kHz speech comes out without a second learned model. Both models have PyTorch and ONNX releases: PyTorch runs on CPU or CUDA, while the browser demo runs ONNX locally through WebGPU with a WASM fallback. The Inflect code and weights are released under Apache 2.0. Try Nano and Micro side by side: https://huggingface.co/spaces/owensong/Inflect-v2 Some people may remember the much rougher Inflect-Nano-v1 experiment. V2 was rebuilt and retrained to address its biggest weaknesses: metallic audio, unstable timing, weak prosody, and difficulty with unseen text. Nano-v2 is smaller while sounding substantially cleaner and more stable. In an anonymous community blind test across nine voices, Micro and Nano placed second and third by normalized preference rate. Full results, samples, limitations, and inference instructions are on the model pages. Inflect-Nano-v2: https://huggingface.co/owensong/Inflect-Nano-v2 I built Inflect independently with a limited training budget. If you test it, give it something difficult: unusual names, numbers, abbreviations, symbols, or awkward punctuation. Send me the exact text when something breaks. If you build something with it, I would genuinely love to see it.
liked
a model
about 20 hours ago
microsoft/Mage-VL
View all activity
Organizations
None yet
ronedgecomb
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
liked
a model
about 4 hours ago
distilbert/distilbert-base-uncased
Fill-Mask
•
67M
•
Updated
May 6, 2024
•
10.4M
•
•
922
liked
3 models
about 20 hours ago
microsoft/Mage-VL
Image-Text-to-Text
•
5B
•
Updated
about 4 hours ago
•
308
•
27
FunAudioLLM/Paraformer-GGUF
Automatic Speech Recognition
•
0.2B
•
Updated
1 day ago
•
660
•
2
OpenMOSS-Team/OmniVAE
Updated
about 22 hours ago
•
9
liked
2 models
3 days ago
inclusionAI/LLaDA2.2-flash
Text Generation
•
103B
•
Updated
3 days ago
•
328
•
57
nvidia/Cosmos-H-Dreams
Image-to-Video
•
Updated
3 days ago
•
116
•
5
liked
5 models
4 days ago
neuphonic/neutts-2e-q8-gguf
Text-to-Speech
•
0.2B
•
Updated
5 days ago
•
420
•
5
neuphonic/neutts-2e
Text-to-Speech
•
0.2B
•
Updated
5 days ago
•
4.89k
•
29
moonshotai/Kimi-K3
Image-Text-to-Text
•
2.8T
•
Updated
about 3 hours ago
•
2.85k
•
•
5.55k
Boogu/Boogu-Image-0.1-Edit-Turbo
Image-to-Image
•
10B
•
Updated
4 days ago
•
858
•
48
Boogu/Boogu-Image-0.1-Base
Text-to-Image
•
10B
•
Updated
4 days ago
•
827
•
75
liked
a Space
4 days ago
Running
30
Paperswithcode
📚
30
Mirror of paperswithcode.co served as a Hugging Face Space.
liked
2 models
4 days ago
mgwr/M87
Text-to-Image
•
Updated
20 days ago
•
6.55k
•
•
198
Kwaipilot/KAT-Coder-V2.5-Dev
Text Generation
•
35B
•
Updated
2 days ago
•
5.31k
•
231
liked
6 models
5 days ago
microsoft/Mage-Flow
Text-to-Image
•
4B
•
Updated
5 days ago
•
1.69k
•
379
danielhanchen/unsloth-blackwell-docker
Updated
1 day ago
•
7
microsoft/Fara1.5-9B
Image-Text-to-Text
•
9B
•
Updated
about 4 hours ago
•
1.48k
•
23
Tdamre/Fara1.5-9B-GGUF
Image-Text-to-Text
•
Updated
5 days ago
•
286
•
1
nvidia/MiniMax-M3-DSpark
Text Generation
•
3B
•
Updated
5 days ago
•
323
•
7
poolside/Laguna-S-2.1-FP8
Text Generation
•
118B
•
Updated
5 days ago
•
38.3k
•
22
Load more