Codesota · Models · ViT-Adapter-L (BEiT-3)Microsoft Research1 results · 1 benchmarks
Model card

ViT-Adapter-L (BEiT-3).

Microsoft Researchopen-sourceUnknown paramsViT-L with spatial prior adapter + BEiT-3 pre-training + Mask2Former head

ViT-Adapter bridges the gap between plain ViT and hierarchical backbones for dense predictions. BEiT-3 pre-trained ViT-L via ViT-Adapter sets SOTA of 62.8 MS mIoU on ADE20K val. ICLR 2023 Spotlight.

§ 02 · Benchmarks

Every benchmark ViT-Adapter-L (BEiT-3) has a recorded score for.

#BenchmarkArea · TaskMetricValueRankDateSource
01ADE20KVision & Documents · Semantic SegmentationmIoU62.8%#3/132026-04-20unverified
Rank column shows this model’s position vs all other models scored on the same benchmark + metric (competitors after the slash). #1 in red means current SOTA. Sorted by rank, then newest result.
§ 03 · Strengths by area

Where ViT-Adapter-L (BEiT-3) actually performs.

Vision & Documents
1
benchmark
avg rank #3.0
§ 05 · Related models

Other Microsoft Research models scored on Codesota.

Faster R-CNN
Unknown params · 19 results
Faster R-CNN (VGG-16)
~137M params · 2 results
DiT-L (Cascade R-CNN)
Unknown params · 1 result
LayoutLMv3-Large
Unknown params · 1 result
NaturalSpeech 3
Unknown params · 1 result
Swin-L (Cascade R-CNN)
1 result
NaturalSpeech
N/A params · 0 results
SwinV2-G
0 results
§ 06 · Sources & freshness

Where these numbers come from.

src
1
result
0 of 1 rows marked verified.