Christian Otto Stelter PRO
stelterlab
AI & ML interests
None yet
Recent Activity
new activity 7 days ago
audreyt/Kolibri-1-NVFP4-W4A16:Works on RTX PRO 6000 new activity 7 days ago
Aleph-Alpha/Kolibri-1:DGX Spark recipe new activity 7 days ago
Aleph-Alpha/Kolibri-1:tool-eval-bench resultsOrganizations
None yet
Works on RTX PRO 6000
#1 opened 7 days ago
by
stelterlab
DGX Spark recipe
👍 2
#5 opened 7 days ago
by
stelterlab
tool-eval-bench results
👍 6
1
#4 opened 7 days ago
by
stelterlab
transformers support for Kolibri1ForCausalLM
2
#2 opened 8 days ago
by
stelterlab
llama.cpp support?
➕ 6
4
#3 opened 8 days ago
by
jacek2024
Add documentation on how to use with vLLM to README.md
🤝👍 6
1
#7 opened 7 months ago
by
stelterlab
Qwen3.5-35B-A3B AWQ quant planned?
2
#1 opened 7 months ago
by
amidwestnoob
Which transformer version did you use?
#3 opened 8 months ago
by
stelterlab
tokenizer_config.json missing chat_template field (tool calling broken without workaround)
1
#1 opened 8 months ago
by
seanthomaswilliams
Updated tokenizer_config.json now w/ chat_template included
#2 opened 8 months ago
by
stelterlab
NVFP4 / AWQ Quants or llm-compressor recipe
🤗 2
1
#1 opened 10 months ago
by
stelterlab
vLLM v0.11.1 seems to work, but v0.11.2 fails
👍❤️ 2
9
#3 opened 11 months ago
by
stelterlab
Error when running in VLLM
👍 2
21
#1 opened about 1 year ago
by
d8rt8v
Unable to run the model in VLLM: KeyError: 'layers.14.mlp.gate.qweight'
3
#1 opened about 1 year ago
by
fredericodeveloper
Rope Scaling pre-applied?
6
#1 opened about 1 year ago
by
the1dv
AWQ version
👍 14
13
#8 opened over 1 year ago
by
celsowm
How did you use auto-round to quantize?
3
#4 opened over 1 year ago
by
stelterlab
please update to Mistral-Small-3.2-24B-Instruct-2506
1
#5 opened over 1 year ago
by
celsowm
Tool Calling issue with stelterlab/Mistral-Small-24B-Instruct-2501-AWQ
1
#4 opened over 1 year ago
by
sbhatt765
Do you any plan to quantize the Qwen3-30B-A3B-AWQ model?
1
#2 opened over 1 year ago
by
Jeanxx