AI & ML interests

LLM quantization & model compression · NVFP4 / weight-only W4A16 · accuracy-gated certification (publish only on pass vs bf16) · efficient inference beyond Blackwell (Ada /Ampere / Hopper, FP4 Marlin) · reproducible, DOI-cited releases · reasoning & instruction-tuned models

uist 's models

None public yet