Views
No views yet
| Property | Value |
|---|---|
| Backbone | DINOv3 ViT-L/16 |
| Input Resolution | 512x512 |
| Task | Semantic Segmentation |
| Dataset | ADE20K |
1@inproceedings{cavagnero2026pmt,
2 author = {Cavagnero, Niccolò and Norouzi, Narges and Dubbelman, Gijs and de Geus, Daan},
3 title = {PMT: Plain Mask Transformer for Image and Video Segmentation with Frozen Vision Encoders},
4 booktitle = {Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW)},
5 year = {2026},
6}