Anjey Sapkovski
anjeysapkovski
AI & ML interests
None yet
Recent Activity
new activity 5 days ago
bottlecapai/ThinkingCap-Qwen3.6-27B:ThinkingCap-Qwen3.6-35b-a3b ? liked a model 17 days ago
ai-sage/GigaAM-Multilingual new activity 18 days ago
bottlecapai/ThinkingCap-Qwen3.6-27B:MTP version requestOrganizations
None yet
ThinkingCap-Qwen3.6-35b-a3b ?
❤️ 5
7
#6 opened about 1 month ago
by
Narutoouz
MTP version request
❤️🔥 3
2
#20 opened 18 days ago
by
anjeysapkovski
Request for UD quants of the model
🚀 1
#2 opened 2 months ago
by
anjeysapkovski
Same speed on 5060 Ti as llamacpp MTP model
➕ 1
#5 opened 2 months ago
by
anjeysapkovski
Starts with 50% speedup, but speed very fast decreases
2
#21 opened 3 months ago
by
seleznyov
draft with llama.cpp?
👀 2
3
#2 opened 4 months ago
by
Schnabulator
Thank you!!
🤗 3
2
#4 opened 3 months ago
by
zrfior
Inference broken with Jan
🚀👀 4
2
#22 opened 6 months ago
by
redaihf
The generation falls into constant repetition without any good result
🔥➕ 3
16
#2 opened 7 months ago
by
ddd2r2
cool model !!
👍 1
3
#3 opened 6 months ago
by
gopi87
Check in here for tok/s and benchmarks for local gguf models
👍 1
6
#1 opened 6 months ago
by
ykarout
Why does the KV cache occupy so much GPU memory?
13
#21 opened 7 months ago
by
yyg201708
1.5b?
🔥 4
7
#3 opened 8 months ago
by
cchance27
please make a 2.1 autoround model (NT)
❤️ 2
3
#1 opened 7 months ago
by
Khatvathiren
Intel AutoRound for best low-bit quantization
1
#5 opened 7 months ago
by
anjeysapkovski
tool calling not working as expected?
👍 2
15
#80 opened about 1 year ago
by
Spider-Jerusalem
2507 Thinking model release
11
#4 opened 10 months ago
by
anjeysapkovski
Not working runtime error
👍 10
3
#25 opened about 1 year ago
by
gokayai
Request for Qwen3-30B-A3B-Thinking-2507 q2ks autoround gguf
1
#1 opened 11 months ago
by
anjeysapkovski
Qwen3-30B-A3B Coder Instruct-2507-gguf-q2ks-mixed-AutoRound
#3 opened about 1 year ago
by
anjeysapkovski