Activity Feed

AI & ML interests

None defined yet.

Recent Activity

mrfakename 
posted an update 9 days ago
view post
Post
506
We’ve been working with LAION on a voice acting arena. You listen to two models doing the same scene and compare how well they pull it off.

It’s ready to try now - would love to hear what you think 🙂

TTS-AGI/voice-acting-arena
fffiloni 
posted an update 10 days ago
view post
Post
254
Been pushing Viggle Animate a bit further 👀

Viggle Mine is my take on longer, harder character replacement shots — crowds, distance changes, characters turning their back and reappearing.

Still experimental, but it’s getting surprisingly robust.

Try it : fffiloni/Viggle-Mine 🤗
AtAndDev 
posted an update 24 days ago
view post
Post
2849
SPECK 2 IS ALREADY OUT: specklabs/Speck2-140M

Pretrained on 4x more tokens than the previous releases (20b vs 5b).
Instruct tuned versions are coming soon.
Very interesting models are coming soon too (hint: super long context).

Thanks for everyone supporting!
  • 3 replies
·
AtAndDev 
posted an update 27 days ago
view post
Post
180
SPECK1.5 IS COMING SOON!
Same 5B token budget but much better corpus quality.

Also getting a ton of downloads, thanks for everyone downloading and liking <3

specklabs
AtAndDev 
posted an update 29 days ago
view post
Post
149
NEW SPECK UPDATES:

Just hit #14 and #15 with out FIRST models on Open SLM Leaderboard. The models were trained on 5B tokens, while competing with similarly sized models trained on more than 6-20x the data.

A new base model Speck1.5-140M being trained right now on a higher quality corpus and will be released soon.
SpeckChat3 is coming very soon with 1 million samples, specifically designed to post train small base models.

Also, just to clarify stuff, we will NOT release anything that is NOT MIT licensed EVER. Openness is needed in small language research.

Thanks to everyone supporting the project, and stay tuned for new releases!
AtAndDev 
posted an update about 1 month ago
view post
Post
1882
SPECK UPDATES:
1 New instruct model tuned on top of Speck1-140M: specklabs/Speck1-140M-Instruct
2 Instruction tuning datasets
2 GGUFs

Much more coming soon:
Speck1.1-140M-Instruct that is post trained on SpeckChat2 will be coming very soon
New base model Speck1.5-140M is coming with a much higher quality corpus

Thanks to everyone who is already supporting the project, and stay tuned for new releases!
  • 3 replies
·
fffiloni 
posted an update about 1 month ago
view post
Post
2164
If an agent can build the obvious demo, the obvious demo probably isn’t worth building anymore.
For years, turning a research repo into something people could actually try was valuable by itself.
That part is becoming automated — and that’s a good thing.

Which means the interesting work moves elsewhere: finding the weird use case, the right interaction, the unexpected model combination — or simply knowing which paper is worth anyone’s attention.

The demo used to be the product. Now it needs a point of view.
  • 1 reply
·
AtAndDev 
posted an update about 1 month ago
view post
Post
2120
FIRST SPECK MODEL RELEASED:
specklabs/Speck1-140M

new models coming very soon (both instruct and much better models), with much much higher training scale as i am getting marenostrum5 access soon!
we will be looking at 100b-2t token budgets :)
  • 4 replies
·
Nymbo 
posted an update about 2 months ago
view post
Post
2295
Anthropic gave me six months of Claude Max 20x through the Claude for Open Source program, granted based on my Hugging Face work. Thank you
Anthropic
for supporting open source.

So far I've been pointing it at Markdown Minimap, an Obsidian plugin that adds a scrollable IDE-style minimap to your notes. This week I've been clearing a backlog of user-reported issues on it, with Claude often handling them end to end.

https://github.com/Nymbo/Markdown-Minimap — issues and PRs welcome.
julien-c 
posted an update about 2 months ago
view post
Post
5572
who's working on an NVFP4 version of Kimi-K3?
  • 4 replies
·
Nymbo 
posted an update 2 months ago
view post
Post
6010
Introducing Inflect-v2, two exceptionally small, open-weight English TTS models at just 3.9M and 9.3M parameters. Both generate speech multiple times faster than real-time on CPU. Despite their size, Inflect-v2 delivers quality that is competitive with much larger lightweight TTS systems, including KittenTTS, Piper, and Supertonic-3.

CPU, CUDA, PyTorch, and ONNX are supported. Apache 2.0.

See it for yourselves:
owensong/Inflect-Micro-v2
owensong/Inflect-Nano-v2

Try the Demos:
Nymbo/Inflect-TTS (unlimited CPU usage)
owensong/Inflect-v2 (ultra-fast ZeroGPU usage)
  • 6 replies
·
fffiloni 
posted an update 3 months ago
fffiloni 
posted an update 3 months ago
view post
Post
1905
I made a Hugging Face Space for SCAIL-2 🤗

Reference character + driving motion → animated result.

A simple demo to explore the paper’s core workflow with curated examples.

👉 fffiloni/SCAIL-2
  • 1 reply
·
fffiloni 
posted an update 3 months ago
view post
Post
858
⏱️ Built a small Space for Visual Chronometer / Pulse of Motion.

Upload a video and estimate its Physical FPS: the frame rate implied by visual motion, independent of metadata.
Useful to inspect “chronometric hallucination” in generated videos: clips that look smooth, but move with the wrong physical time scale.

Try it here: fffiloni/Pulse-of-Motion
  • 1 reply
·