Scott Leimroth
scottleimroth
ยท
AI & ML interests
Local LLM benchmarking and model tuning on NVIDIA DGX Spark. Quantisation, speculative decoding, abliteration.
Recent Activity
upvoted a paper 2 days ago
DFlash: Block Diffusion for Flash Speculative Decoding liked a model 2 days ago
nvidia/Qwen3.8-27B-NVFP4 liked a model 2 days ago
Qwen/Qwen3.8-27B