4 Commits

Author SHA1 Message Date
Eugene Rakhmatulin
9e089acf2b Updated Nemotron recipes to use VLLM CUTLASS 2026-03-22 23:03:24 -07:00
Eugene Rakhmatulin
b1eeefc0eb Changed Nemotron-3-Nano-NVFP4 to Marlin backend 2026-03-17 13:10:48 -07:00
Eugene Rakhmatulin
c6b245cfe8 Added prefix caching to nemotron recipe 2026-02-10 18:25:01 -08:00
Eugene Rakhmatulin
74876dd442 Added recipes for nemotron-nano-3 and qwen3-coder-next 2026-02-09 14:33:35 -08:00