SIGN IN SIGN UP

feat: add CAPI imagenet benchmark (#1998)

* feat: add CAPI imagenet benchmark

* docs: note the encoder RoPE deviation in the CAPI benchmark

The encoder is a standard masked ViT; the reference also applies RoPE inside the encoder. Matches the deviation already documented for the CAPI examples (#1996).

* refactor: use vit-small encoder in CAPI benchmark

Align the encoder with the other vitb16 benchmarks (dino/dinov2/ibot) and size the predictor to it: num_heads=6 and depth=6 (half the encoder depth), following CAPI section 4.1. Addresses review.
L
Lőrincz-Molnár Szabolcs-Botond committed
413687d5c41f18d51cb25e40fd274660cc58981e
Parent: dba2914
Committed by GitHub <noreply@github.com> on 8/19/2026, 6:09:29 PM