GeoNUSAF - SegNeXt-T - block split, fold 1

Kathmandu Valley land-use segmentation, 6 classes, ignore_index=255.

field value
architecture MSCAN-T encoder + LightHamHead decoder (SegNeXt, NeurIPS 2022)
encoder init ImageNet-1K, 100.0% of tensors loaded
params 4.23 M
head fuse stride 8 (stages [1, 2, 3])
NMF rank / steps R=16, 6 train / 7 eval
split mode block
fold 1 of 3
seed 42
input 512x512, ImageNet norm, effective GSD 0.586 m/px
classes Residential, Road, River, Forest, UnusedLand, Agricultural
lr (head/encoder) 0.0006 / 6e-05
regularization wd 0.01, drop_path 0.1, smooth 0.05, EMA True
best epoch 148
val mIoU 0.5701
val mF1 0.7091
val OA 0.8355
val kappa 0.6994

Per-class (validation)

class IoU F1
Residential 0.8573 0.9232
Road 0.4698 0.6392
River 0.4783 0.6471
Forest 0.7372 0.8487
UnusedLand 0.3054 0.4679
Agricultural 0.5727 0.7283

best.pt holds model_state (EMA weights when EMA is on), arch (the dict needed to rebuild the network), the run cfg, and metrics. Rebuild with segnext_model.py from this same repo. Model code derives from Visual-Attention-Network/SegNeXt (Apache-2.0).

Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support