Stars
[CVPR 2026 Findings] RecycleLoRA: Rank-Revealing QR-Based Dual-LoRA Subspace Adaptation for Domain Generalized Semantic Segmentation
[IEEE Access] Partial Large Kernel CNNs for Efficient Super-Resolution
IGConv: Implicit Grid Convolution for Multi-Scale Image Super-Resolution
[ICCV2025 Highlight] ESC: Emulating Self-attention with Convolution for Efficient Image Super-Resolution
Rank-Factorized Implicit Neural Bias: Scaling Super-Resolution Transformer with FlashAttention
[ACCV 2026] FFNet: MetaMixer-based Efficient Convolutional Mixer Design
NeurIPS 2025 Accepted Paper NoisyGRPO: Incentivizing Multimodal CoT Reasoning via Noise Injection and Bayesian Estimation
Open-source evaluation toolkit of large multi-modality models (LMMs), support 220+ LMMs, 80+ benchmarks
[CVPR 2025 Highlight] Your Large Vision-Language Model Only Needs A Few Attention Heads For Visual Grounding
[ICLR'25] Official code for the paper 'MLLMs Know Where to Look: Training-free Perception of Small Visual Details with Multimodal LLMs'
[CVPR 2024] SHViT: Single-Head Vision Transformer with Memory Efficient Macro Design
[CVPR 2025 Highlight] SoMA: Singular Value Decomposed Minor Components Adaptation for Domain Generalizable Representation Learning