GeoNUSAF - SegNeXt-T - block split, fold 2

Kathmandu Valley land-use segmentation, 6 classes, ignore_index=255.

field value
architecture MSCAN-T encoder + LightHamHead decoder (SegNeXt, NeurIPS 2022)
encoder init ImageNet-1K, 100.0% of tensors loaded
params 4.23 M
head fuse stride 8 (stages [1, 2, 3])
NMF rank / steps R=16, 6 train / 7 eval
split mode block
fold 2 of 3
seed 42
input 512x512, ImageNet norm, effective GSD 0.586 m/px
classes Residential, Road, River, Forest, UnusedLand, Agricultural
lr (head/encoder) 0.0006 / 6e-05
regularization wd 0.01, drop_path 0.1, smooth 0.05, EMA True
best epoch 155
val mIoU 0.5122
val mF1 0.6551
val OA 0.8652
val kappa 0.6737

Per-class (validation)

class IoU F1
Residential 0.8904 0.9420
Road 0.4595 0.6296
River 0.4129 0.5844
Forest 0.4875 0.6555
UnusedLand 0.2319 0.3764
Agricultural 0.5908 0.7428

best.pt holds model_state (EMA weights when EMA is on), arch (the dict needed to rebuild the network), the run cfg, and metrics. Rebuild with segnext_model.py from this same repo. Model code derives from Visual-Attention-Network/SegNeXt (Apache-2.0).

Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support