arXivAI Research
One Block, Multiple Depths: Recurrent Vision Transformers with Depth-Programmed Experts
單一區塊實現多重深度:具備深度編程專家庫的循環 Vision Transformer
The paper introduces reViT, which reuses a single Transformer block recurrently by dynamically merging shared expert weights per depth, matching full-depth ViT accuracy with roughly 70% fewer stored parameters.
2 min read