AI & ML interests

LLM quantization & model compression · NVFP4 / weight-only W4A16 · accuracy-gated certification (publish only on pass vs bf16) · efficient inference beyond Blackwell (Ada /Ampere / Hopper, FP4 Marlin) · reproducible, DOI-cited releases · reasoning & instruction-tuned models

Recent Activity

uist  updated a model 1 day ago
uist-labs/Qwen2.5-7B-Instruct-NVFP4A16
uist  updated a Space 4 days ago
uist-labs/README
View all activity