Yaziwel/Awesome-RWKV-in-Vision

A curated list of papers on the applications of RWKV in computer vision. Please raise an issue if you suggest new qualified project.

246

19 commits

updated Jun 19, 2025

See the code

README

Awesome-RWKV-in-Vision Awesome

Image

RWKV4: StararXiv

RWKV5&6: StararXiv

RWKV7: StararXiv

RWKV Ecosystem: https://rwkv.cn/eco

Table of Contents

Vision Backbone

TitleCodeLink
Vision-RWKV: Efficient and Scalable Visual Perception with RWKV-Like ArchitecturesStararXiv

Image Restoration

TitleCodeLink
Restore-RWKV: Efficient and Effective Medical Image Restoration with RWKVStararXiv
GDSR: Global-Detail Integration through Dual-Branch Network with Wavelet Losses for Remote Sensing Image Super-ResolutionN/AarXiv
Exploring Linear Attention Alternative for Single Image Super-ResolutionStararXiv
URWKV: Unified RWKV Model with Multi-state Perspective for Low-light Image RestorationStararXiv
Multi-View Learning with Context-Guided Receptance for Image DenoisingStararXiv

Image Compression

TitleCodeLink
Linear Attention Modeling for Learned Image CompressionStararXiv

Image Generation

TitleCodeLink
Diffusion-RWKV: Scaling RWKV-Like Architectures for Diffusion ModelsStararXiv

Video Generation

TitleCodeLink
FEAT: Full-Dimensional Efficient Attention Transformer for Medical Video GenerationStararXiv

Image Segmentation

TitleCodeLink
Vision-RWKV: Efficient and Scalable Visual Perception with RWKV-Like ArchitecturesStararXiv
Mamba or RWKV: Exploring High-Quality and High-Efficiency Segment Anything ModelStararXiv
BSBP-RWKV: Background Suppression with Boundary Preservation for Efficient Medical Image SegmentationN/AACMMM
RWKV-UNet: Improving UNet with Long-Range Cooperation for Effective Medical Image SegmentationStararXiv
Zig-RiR: Zigzag RWKV-in-RWKV for Efficient Medical Image SegmentationStarIEEE-TMI
ScaleMatch: Multi-scale Consistency Enhancement for Semi-supervised Semantic SegmentationStarAAAI
RSRWKV: A Linear-Complexity 2D Attention Mechanism for Efficient Remote Sensing Vision TaskN/AarXiv

Vision-Language Model

TitleCodeLink
VisualRWKV: Visual Language model based on RWKVStararXiv
RWKV-CLIP: A Robust Vision-Language Representation LearnerStararXiv

3D Point Cloud Learning

TitleCodeLink
PointRWKV: Efficient RWKV-Like Model for Hierarchical Point Cloud LearningStararXiv
LION: Linear Group RNN for 3D Object Detection in Point CloudsStararXiv
OccRWKV: Rethinking Efficient 3D Semantic Occupancy Prediction with Linear ComplexityStararXiv

Video Understanding

TitleCodeLink
Video RWKV: Video Action Recognition Based RWKVN/AarXiv

Survey

TitleCodeLink
A Survey of RWKVStararXiv
From Transformers to the Future: An In-Depth Exploration of Modern Language Model ArchitecturesN/ALINK
The Evolution of RWKV: Advancements in Efficient Language ModelingN/AarXiv
computer-vision
mamba
rwkv
rwkv4
rwkv5
rwkv6
rwkv7
transformer

Yaziwel/Awesome-RWKV-in-Vision

A curated list of papers on the applications of RWKV in computer vision. Please raise an issue if you suggest new qualified project.

246

19 commits

updated Jun 19, 2025

See the code

README

Awesome-RWKV-in-Vision Awesome

Image

RWKV4: StararXiv

RWKV5&6: StararXiv

RWKV7: StararXiv

RWKV Ecosystem: https://rwkv.cn/eco

Table of Contents

Vision Backbone

TitleCodeLink
Vision-RWKV: Efficient and Scalable Visual Perception with RWKV-Like ArchitecturesStararXiv

Image Restoration

TitleCodeLink
Restore-RWKV: Efficient and Effective Medical Image Restoration with RWKVStararXiv
GDSR: Global-Detail Integration through Dual-Branch Network with Wavelet Losses for Remote Sensing Image Super-ResolutionN/AarXiv
Exploring Linear Attention Alternative for Single Image Super-ResolutionStararXiv
URWKV: Unified RWKV Model with Multi-state Perspective for Low-light Image RestorationStararXiv
Multi-View Learning with Context-Guided Receptance for Image DenoisingStararXiv

Image Compression

TitleCodeLink
Linear Attention Modeling for Learned Image CompressionStararXiv

Image Generation

TitleCodeLink
Diffusion-RWKV: Scaling RWKV-Like Architectures for Diffusion ModelsStararXiv

Video Generation

TitleCodeLink
FEAT: Full-Dimensional Efficient Attention Transformer for Medical Video GenerationStararXiv

Image Segmentation

TitleCodeLink
Vision-RWKV: Efficient and Scalable Visual Perception with RWKV-Like ArchitecturesStararXiv
Mamba or RWKV: Exploring High-Quality and High-Efficiency Segment Anything ModelStararXiv
BSBP-RWKV: Background Suppression with Boundary Preservation for Efficient Medical Image SegmentationN/AACMMM
RWKV-UNet: Improving UNet with Long-Range Cooperation for Effective Medical Image SegmentationStararXiv
Zig-RiR: Zigzag RWKV-in-RWKV for Efficient Medical Image SegmentationStarIEEE-TMI
ScaleMatch: Multi-scale Consistency Enhancement for Semi-supervised Semantic SegmentationStarAAAI
RSRWKV: A Linear-Complexity 2D Attention Mechanism for Efficient Remote Sensing Vision TaskN/AarXiv

Vision-Language Model

TitleCodeLink
VisualRWKV: Visual Language model based on RWKVStararXiv
RWKV-CLIP: A Robust Vision-Language Representation LearnerStararXiv

3D Point Cloud Learning

TitleCodeLink
PointRWKV: Efficient RWKV-Like Model for Hierarchical Point Cloud LearningStararXiv
LION: Linear Group RNN for 3D Object Detection in Point CloudsStararXiv
OccRWKV: Rethinking Efficient 3D Semantic Occupancy Prediction with Linear ComplexityStararXiv

Video Understanding

TitleCodeLink
Video RWKV: Video Action Recognition Based RWKVN/AarXiv

Survey

TitleCodeLink
A Survey of RWKVStararXiv
From Transformers to the Future: An In-Depth Exploration of Modern Language Model ArchitecturesN/ALINK
The Evolution of RWKV: Advancements in Efficient Language ModelingN/AarXiv
computer-vision
mamba
rwkv
rwkv4
rwkv5
rwkv6
rwkv7
transformer