Vladimir
vstiff
AI & ML interests
None yet
Recent Activity
liked a model 3 days ago
autotrust/JEV-27B-VL liked a model 3 days ago
google/embeddinggemma-2 upvoted a paper 4 days ago
OmniReasoning: Pushing the Limits of Audio-Visual Joint ReasoningOrganizations
None yet
Does it still hallucinate?
β€οΈπ€ 3
6
#4 opened 10 days ago
by
vstiff
Q4_K_XL / llama-server / Hallucination after 150k of filled context
3
#42 opened about 1 month ago
by
manisab
Identified serious issue with Flash-Next hallucinating user instructions!
1
#57 opened 23 days ago
by
laserbeans
strix halo 128gb recommendations
21
#28 opened about 1 month ago
by
dilavni
Qwen3.8-Flash-Next on Strix Halo: ~40 tok/s code at 120K+ context
π 9
4
#45 opened about 1 month ago
by
Engardium
rocm support
3
#1 opened about 2 months ago
by
idchrono
Love your models, any chance of something between 35b and 397b?
π 5
4
#1 opened about 2 months ago
by
Shaunm89
Ryzen AI Max+ 395 results: 16.68 tok/s with UD-Q5_K_XL, Vulkan and MTP 4
β€οΈ 3
12
#41 opened about 2 months ago
by
erstmalreden
I think this model is not ready for use
ππ 7
20
#40 opened about 2 months ago
by
sa13ma
Same tool call infinite loop
1
#139 opened 2 months ago
by
vstiff
1-bit Kimi K3 vs Claude Opus 5 vs GPT 5.6
ππ 20
25
#12 opened 2 months ago
by
danielhanchen
Wrong Qwen3.6-35B-A3B Benchmark
π 1
2
#17 opened 2 months ago
by
XCurOS-35H
Prefill slowdown
πβ 2
4
#9 opened 3 months ago
by
vstiff
Inconsistent reasoning trigger on llama-server
π 4
9
#2 opened 3 months ago
by
Lowkey-Loki
Prefill slowdown
βπ 2
4
#9 opened 3 months ago
by
vstiff
Prefill slowdown
βπ 2
4
#9 opened 3 months ago
by
vstiff