Activity Feed

AI & ML interests

None defined yet.

Recent Activity

sbrandeis 
posted an update about 20 hours ago
mrfakename 
posted an update 8 days ago
view post
Post
429
We’ve been working with LAION on a voice acting arena. You listen to two models doing the same scene and compare how well they pull it off.

It’s ready to try now - would love to hear what you think 🙂

TTS-AGI/voice-acting-arena
fffiloni 
posted an update 8 days ago
view post
Post
207
Been pushing Viggle Animate a bit further 👀

Viggle Mine is my take on longer, harder character replacement shots — crowds, distance changes, characters turning their back and reappearing.

Still experimental, but it’s getting surprisingly robust.

Try it : fffiloni/Viggle-Mine 🤗
fffiloni 
posted an update 30 days ago
view post
Post
2160
If an agent can build the obvious demo, the obvious demo probably isn’t worth building anymore.
For years, turning a research repo into something people could actually try was valuable by itself.
That part is becoming automated — and that’s a good thing.

Which means the interesting work moves elsewhere: finding the weird use case, the right interaction, the unexpected model combination — or simply knowing which paper is worth anyone’s attention.

The demo used to be the product. Now it needs a point of view.
  • 1 reply
·
Nymbo 
posted an update about 1 month ago
view post
Post
2281
Anthropic gave me six months of Claude Max 20x through the Claude for Open Source program, granted based on my Hugging Face work. Thank you
Anthropic
for supporting open source.

So far I've been pointing it at Markdown Minimap, an Obsidian plugin that adds a scrollable IDE-style minimap to your notes. This week I've been clearing a backlog of user-reported issues on it, with Claude often handling them end to end.

https://github.com/Nymbo/Markdown-Minimap — issues and PRs welcome.
julien-c 
posted an update about 2 months ago
view post
Post
5535
who's working on an NVFP4 version of Kimi-K3?
  • 4 replies
·
Nymbo 
posted an update about 2 months ago
view post
Post
5998
Introducing Inflect-v2, two exceptionally small, open-weight English TTS models at just 3.9M and 9.3M parameters. Both generate speech multiple times faster than real-time on CPU. Despite their size, Inflect-v2 delivers quality that is competitive with much larger lightweight TTS systems, including KittenTTS, Piper, and Supertonic-3.

CPU, CUDA, PyTorch, and ONNX are supported. Apache 2.0.

See it for yourselves:
owensong/Inflect-Micro-v2
owensong/Inflect-Nano-v2

Try the Demos:
Nymbo/Inflect-TTS (unlimited CPU usage)
owensong/Inflect-v2 (ultra-fast ZeroGPU usage)
  • 6 replies
·