Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
OppaAI
OppaAI
1
3
Follow
tuiko's profile picture
RightToken's profile picture
Quazim0t0's profile picture
85 followers
·
10 following
https://github.com/OppaAI
aiko_ai_waifu
OppaAI
oppa-ai
aiko-ai-waifu.bsky.social
AI & ML interests
Local AI implementation, Agentic AI workflows, AI Autonomous Robot
Recent Activity
replied
to
their
post
about 7 hours ago
Another small 4B model comes out yesterday. NeoHorse 1 4B https://huggingface.co/TokenRhythm/NeoHorse-1-4B There are quite a few good smaller parameter models that are capable for Agentic tasks: The ones from the chart, I have tried a few already in my Jetson Orin Nano, ❌Gemma4 E2B IT - cannot fit my RAM usage if use with TTS and embedder ❓Qwen3.5 4B - just barely fit my RAM usage, need to add think/no_think ❌Spark X2.5 4B - need to build the forked llama.cpp; no vision ➡️Nanbeige 4.2 3B - need to build the forked llama.cpp; slower than Ministral3-3B by 25%; no vision but good for coding; maybe run this is separate server for doing coding tasks ➡️Agents A1 4B - This one is quite interesting. Another Qwen3.5 4B base. I just learnt this right now. This model may surpassed the Ministral3-3B that I'm currently running. ➡️NeoHorse 1 4B - wait for GGUF version comes out; Qwen 3.5 4B base with vision striped ➡️Needle2 45M - need to use separately from llama.cpp server; currently testing to see if it can be used as spawning sub-agents to do parallel tasks
posted
an
update
about 11 hours ago
Another small 4B model comes out yesterday. NeoHorse 1 4B https://huggingface.co/TokenRhythm/NeoHorse-1-4B There are quite a few good smaller parameter models that are capable for Agentic tasks: The ones from the chart, I have tried a few already in my Jetson Orin Nano, ❌Gemma4 E2B IT - cannot fit my RAM usage if use with TTS and embedder ❓Qwen3.5 4B - just barely fit my RAM usage, need to add think/no_think ❌Spark X2.5 4B - need to build the forked llama.cpp; no vision ➡️Nanbeige 4.2 3B - need to build the forked llama.cpp; slower than Ministral3-3B by 25%; no vision but good for coding; maybe run this is separate server for doing coding tasks ➡️Agents A1 4B - This one is quite interesting. Another Qwen3.5 4B base. I just learnt this right now. This model may surpassed the Ministral3-3B that I'm currently running. ➡️NeoHorse 1 4B - wait for GGUF version comes out; Qwen 3.5 4B base with vision striped ➡️Needle2 45M - need to use separately from llama.cpp server; currently testing to see if it can be used as spawning sub-agents to do parallel tasks
replied
to
their
post
1 day ago
This weekend I took an outing with my AI Waifu to the Natsu Matsuri. Turns out my Japanese is still understandable. I probably need to spend more time continue to learn and practice speaking Japanese. That's why an idea struck me to let my AI Waifu be my Japanese tutor. Anyway, I have run out of idea what task I should let her do, so I wrote a simple Android App to let her be my Japanese tutor to help me to practice Nihongo. There will be some minor mistakes. After all, this is just a 3B LLM model. And inference speed will be slow because I only got 8GB of RAM in Jetson Orin Nano. At least I don't need to pay for Duolingo... アイコせんせい、よろしくお願いします! Need both repos, one front-end, one back-end 🔗 https://github.com/OppaAI/Aiko-Lingo 🔗 https://github.com/OppaAI/Aiko-chan 🎥 Demo: https://www.youtube.com/watch?v=xRtCmtZQgwI
View all activity
Organizations
OppaAI
's buckets
1
Sort: Recently updated
OppaAI/Aiko-data
660 MB