Wisp: small Llama-style base models trained from scratch on educational text and code.
DedeProGames PRO
AI & ML interests
Agentic & Coding finetune, Decoder-Onlys from Scratch, Image LoRAs, AI Researcher
Recent Activity
updated a dataset about 1 hour ago
DedeProGames/lm-tetris-arena-results liked a Space about 3 hours ago
DedeProGames/claude-code-mcp published a Space about 3 hours ago
DedeProGames/claude-code-mcpOrganizations
Kiyo
Kiyo: State-of-the-art models trained from scratch on vast amounts of web tokens, educational text, and code.
LowOnMind
LowOnMind: Small decoder-only models trained in small amount of tokens
NanoAndy
NTX
Chennus Series
Small and Efficient Chess Models
GPT-U
GPT-U: small Llama-style base models trained from scratch on web, educational text and code, with fully open scripts, logs and evals.
DynamicMind
DynamicMind: a family of lightweight models trained from scratch on diverse datasets.
Experimental Decoder-Onlys
Some experimental decoder-onlys made by me for research
NTX-2.1
medqwen
Experimental medical model based on Qwen2.5 Family
Wisp
Wisp: small Llama-style base models trained from scratch on educational text and code.
GPT-U
GPT-U: small Llama-style base models trained from scratch on web, educational text and code, with fully open scripts, logs and evals.
Kiyo
Kiyo: State-of-the-art models trained from scratch on vast amounts of web tokens, educational text, and code.
DynamicMind
DynamicMind: a family of lightweight models trained from scratch on diverse datasets.
LowOnMind
LowOnMind: Small decoder-only models trained in small amount of tokens
Experimental Decoder-Onlys
Some experimental decoder-onlys made by me for research
NanoAndy
NTX-2.1
NTX
medqwen
Experimental medical model based on Qwen2.5 Family
Chennus Series
Small and Efficient Chess Models