Roman Ivanov
perelmanych
·
AI & ML interests
None yet
Recent Activity
liked a model about 18 hours ago
AngelSlim/Hy3-GGUF new activity 2 days ago
unsloth/DeepSeek-V4-Flash-0731-GGUF:KLD Benchmarks + why Q8_K_XL vs MXFP4 naming new activity 3 days ago
unsloth/DeepSeek-V4-Flash-0731-GGUF:Dual RTX3090 and 64GB DDR4 + 78GB Swap. Speed: ~1.78 TPSOrganizations
None yet
KLD Benchmarks + why Q8_K_XL vs MXFP4 naming
👍 6
6
#11 opened 3 days ago
by
danielhanchen
Dual RTX3090 and 64GB DDR4 + 78GB Swap. Speed: ~1.78 TPS
🤝👍 4
13
#8 opened 3 days ago
by
robert1968
Relative quant qualities
1
#2 opened 16 days ago
by
nalimc
Premature ending inside final answer after long thinking
#16 opened 11 days ago
by
perelmanych
Running with recently merged llama.cpp PR
👍 6
6
#16 opened about 1 month ago
by
ubergarm
Follow up question results in complete mess
#4 opened 29 days ago
by
perelmanych
Really looking forward for 9B or 12B variants
4
#20 opened about 1 month ago
by
perelmanych
Definitely interested in this one!
🚀 2
25
#1 opened 9 months ago
by
mtcl
Incorrect Model Uploaded
🤗👍 18
6
#8 opened 12 months ago
by
noteventhrice
Qwen3 coder version
#1 opened 12 months ago
by
perelmanych
Difference from other presets
#1 opened about 1 year ago
by
perelmanych
R1 32b is much worse than QwQ ...
22
#6 opened over 1 year ago
by
mirek190
SFT (Non-RL) distillation is this good on a sub-100B model?
3
#2 opened over 1 year ago
by
KrishnaKaasyap
IQ2_XS variant
1
#2 opened over 2 years ago
by
perelmanych
When we can expect vicuna variant of CodeLlama-2 34b model?
👍 1
#10 opened almost 3 years ago
by
perelmanych
Can't load q5_1 model
3
#1 opened about 3 years ago
by
perelmanych
Error when using with web-ui "KeyError: 'model.layers.39.self_attn.q_proj.wf1'"
❤️ 1
16
#7 opened over 3 years ago
by
TheFairyMan
Error using ooba-gooba
👍 1
39
#6 opened over 3 years ago
by
blueisbest
Error using ooba-gooba
👍 1
39
#6 opened over 3 years ago
by
blueisbest
Error using ooba-gooba
👍 1
39
#6 opened over 3 years ago
by
blueisbest