Slavik
SlavikF
AI & ML interests
None yet
Recent Activity
new activity 8 days ago
unsloth/Qwen3.8-27B-GGUF:Performance report on RTX 5090: 100 t/s with UD-Q6_K_XL new activity 10 days ago
unsloth/Qwen3.8-27B-GGUF:Make an SVG of The golden gate bridge new activity 10 days ago
unsloth/Qwen3.8-27B-GGUF:MTP?Organizations
None yet
Performance report on RTX 5090: 100 t/s with UD-Q6_K_XL
9
#14 opened 10 days ago
by
SlavikF
Make an SVG of The golden gate bridge
π 6
5
#16 opened 10 days ago
by
NjProVk
MTP?
15
#12 opened 10 days ago
by
TheWegemann
Qwen3.8-27B Serving Configs: RTX 4090 llama.cpp GGUF
π 1
8
#11 opened 10 days ago
by
erdal
decoding with DSpark is slower on llama
3
#31 opened 20 days ago
by
SlavikF
Does this includes DSpark?
5
#2 opened 24 days ago
by
SlavikF
Best inference engine to use at the moment for the full quant on 96gb vram
8
#4 opened 24 days ago
by
void009
Does it support image input?
π 4
1
#1 opened 25 days ago
by
SlavikF
Laguna updates July 26th 2026
β€οΈπ 15
10
#19 opened 29 days ago
by
danielhanchen
Do you plan to offer this model in GGUF?
π 1
3
#1 opened about 2 months ago
by
Nerdsking
Performance report UD-Q3_K_XL on RTX 5090 + 384GB DDR5: 4 t/s
#14 opened about 1 month ago
by
SlavikF
Performance report: 55 t/s on RTX 4060
π 1
1
#4 opened about 2 months ago
by
SlavikF
5-bit MiniMax M3 running locally on a single M3 Ultra 512GB via Unsloth!
ππ₯ 3
6
#3 opened 2 months ago
by
danielhanchen
Is MTP supported?
3
#5 opened 2 months ago
by
SlavikF
Performance report on RTX 4090D (48GB VRAM): 40 t/s
π 1
4
#11 opened 4 months ago
by
SlavikF
Report: 56 t/s on RTX 4090D (48GB VRAM) with UD-Q6_K_XL
4
#25 opened 3 months ago
by
SlavikF
Report: 30 t/s on RTX 4090D (48GB VRAM) with UD-Q6_K_XL
π€ 5
4
#7 opened 4 months ago
by
SlavikF
Q4_K_XL on RTX 4090D with 1M context: ~70 t/s
#14 opened 4 months ago
by
SlavikF
Apr 11: Updated with Google chat template fixes + more
πβ€οΈ 17
20
#16 opened 5 months ago
by
danielhanchen