nm-testing/tinysmokeqwen3moe
2.93M • Updated • 22.4k
nm-testing/Meta-Llama-3-8B-Instruct-MXFP4
5B • Updated • 7
nm-testing/Llama-3.1-8B-Instruct-QKV-Cache-FP8
8B • Updated • 3
nm-testing/Llama3_2_1B_speculator.eagle3
0.4B • Updated • 61.9k
nm-testing/TinyLlama-1.1B-Chat-v1.0-NVFP4-test132
0.7B • Updated • 3
nm-testing/TinyLlama-1.1B-Chat-v1.0-awq-asym-test-awq-asym
1B • Updated • 2
nm-testing/TinyLlama-1.1B-Chat-v1.0-NVFP4-1105
Updated
nm-testing/TinyLlama-1.1B-Chat-v1.0-NVFP4-test011
Updated
nm-testing/TinyLlama-1.1B-Chat-v1.0-NVFP4-test
Updated
nm-testing/Kimi-Linear-48B-A3B-Instruct-FP8-DYNAMIC
49B • Updated • 486
nm-testing/llama2.c-stories42M-pruned2.4
Updated • 230
nm-testing/gpt-oss-20B.eagle3.unconverted-drafter
1B • Updated • 29
nm-testing/random-weights-llama3.1.8b-2layer-eagle3-unconverted
1B • Updated • 290
nm-testing/Qwen3-VL-235B-A22B-Instruct-FP8-BLOCK
Text Generation
• Updated nm-testing/Qwen3-30B-A3B-FP8-block
Text Generation
• 3B • Updated • 2.37k
nm-testing/granite-4.0-h-small-FP8-dynamic-test
Updated
nm-testing/tiny-testing-random-weights
584k • Updated • 3.58k
nm-testing/Llama4-Maverick-Eagle3-Speculators-64k-vocab
1B • Updated • 13
nm-testing/NVIDIA-Nemotron-Nano-9B-v2-quantized.w4a16
2B • Updated • 2
nm-testing/Qwen3-VL-8B-Instruct-W4A16
3B • Updated • 47
nm-testing/Qwen3-VL-8B-Instruct-NVFP4
6B • Updated • 504
• 3
nm-testing/Qwen3-VL-4B-Instruct-NVFP4
3B • Updated • 576
• 2
nm-testing/Llama-3.1-8B-Instruct-NVFP4-mse
5B • Updated • 2
nm-testing/Llama-3.1-8B-Instruct-NVFP4-static_minmax
5B • Updated • 3
nm-testing/EAGLE3-LLaMA3.1-Instruct-8B-sgl
0.4B • Updated • 10
nm-testing/Speculator-Qwen3-8B-Eagle3-converted-071-quantized-w4a16-sgl
1B • Updated • 13
nm-testing/SpeculatorLlama3-1-8B-Eagle3-sgl
1.0B • Updated • 9
nm-testing/Mockup-qwen235-eagle3-fp16-sgl
1B • Updated • 8