proomt

Search

Search posts, papers, and topics

compact models

RSS
  1. 1

    TAPe+ML: A Compact Structured Representation for Multi-Task Computer Vision

    TAPe+ML v3 is a compact multi‑task vision system that replaces raw‑pixel processing with a structured TAPe representation. Using <100 k parameters, it achieves 84.7 mAP50 (65.3 mAP50‑95) on COCO detection, 80.7 mask mAP50 (58.4 mask mAP50‑95) on COCO segmentation, 92 % top‑1 on Imagenette and 89.9 % on ImageNet‑Real, while also showing robustness to distribution shift in video scene detection.

    Hugging Face Daily Papersarxiv.org1 minpaper