Views
No views yet
| field | value |
|---|---|
| architecture | MSCAN-T encoder + LightHamHead decoder (SegNeXt, NeurIPS 2022) |
| encoder init | ImageNet-1K, 100.0% of tensors loaded |
| params | 4.23 M |
| head fuse stride | 8 (stages [1, 2, 3]) |
| NMF rank / steps | R=16, 6 train / 7 eval |
| split mode | block |
| fold | 0 of 3 |
| seed | 42 |
| input | 512x512, ImageNet norm, effective GSD 0.586 m/px |
| classes | Residential, Road, River, Forest, UnusedLand, Agricultural |
| lr (head/encoder) | 0.0006 / 6e-05 |
| regularization | wd 0.01, drop_path 0.1, smooth 0.05, EMA True |
| best epoch | 178 |
| val mIoU | 0.3348 |
| val mF1 | 0.4616 |
| val OA | 0.6118 |
| val kappa | 0.4861 |
| class | IoU | F1 |
|---|---|---|
| Residential | 0.6210 | 0.7662 |
| Road | 0.2535 | 0.4045 |
| River | 0.0607 | 0.1145 |
| Forest | 0.5553 | 0.7140 |
| UnusedLand | 0.0980 | 0.1784 |
| Agricultural | 0.4201 | 0.5916 |
best.pt holds model_state (EMA weights when EMA is on), arch (the dict needed to rebuild the network), the run cfg, and metrics. Rebuild with segnext_model.py from this same repo. Model code derives from Visual-Attention-Network/SegNeXt (Apache-2.0).