🏛️ 顶级学术会议

会议 全称 级别 每年论文数 官网 开放获取
CVPR IEEE/CVF Computer Vision and Pattern Recognition CCF-A, 顶会 ~2400 https://cvpr.thecvf.com https://openaccess.thecvf.com/CVPR2025
ICCV IEEE/CVF International Conference on Computer Vision CCF-A, 顶会 ~1800 https://iccv2025.thecvf.com https://openaccess.thecvf.com/ICCV2025
ECCV European Conference on Computer Vision CCF-B ~1600 https://eccv2026.org 部分开放
NeurIPS Neural Information Processing Systems CCF-A ~3200 https://neurips.cc https://papers.nips.cc
ICLR International Conference on Learning Representations CCF-A ~2300 https://iclr.cc 开放
ACM MM ACM Multimedia CCF-A ~1300 https://www.acmmm.org 部分
AAAI AAAI Conference on Artificial Intelligence CCF-A ~2300 https://aaai.org 部分
BMVC British Machine Vision Conference CCF-C ~600 https://bmvc2025.org 开放
WACV Winter Conference on Applications of CV - ~800 https://wacv2025.thecvf.com 开放

📋 顶级期刊

期刊 影响因子 (2024) Publisher
TPAMI (IEEE Trans. on Pattern Analysis and Machine Intelligence) ~24.5 IEEE
IJCV (International Journal of Computer Vision) ~19.5 Springer
TIP (IEEE Trans. on Image Processing) ~10.8 IEEE
CVIU (Computer Vision and Image Understanding) ~7.6 Elsevier
PR (Pattern Recognition) ~8.5 Elsevier
MIA (Medical Image Analysis) ~13.8 Elsevier

🎓 顶级实验室

实验室 所属机构 知名教授 方向
SAIL (Stanford AI Lab) Stanford Fei-Fei Li, Leonidas Guibas CV + ML
CSAIL MIT Bill Freeman, Antonio Torralba CV + Graphics
BAIR UC Berkeley Alyosha Efros, Trevor Darrell CV + NLP
VGG / OII Oxford Andrew Zisserman, Andrea Vedaldi 识别/视频
INRIA (Willow) INRIA / ENS Jean Ponce, Josef Sivic 3D视觉
MPI Informatics Max Planck Bernt Schiele, Christian Theobalt 人体/3D
Facebook AI Research (FAIR) Meta Yann LeCun, Michael Tschannen 自监督/基础模型
Google DeepMind Google Nando de Freitas 多模态/RL
CVL ETH Zurich Luc Van Gool 3D/自动驾驶
CMU RI CMU Takeo Kanade, Martial Hebert 机器人视觉
商汤研究院 SenseTime 汤晓鸥 全栈CV
旷视研究院 Megvii 孙剑 检测+识别
腾讯优图 Tencent 贾佳亚 应用CV
阿里巴巴达摩院 Alibaba 团队 视觉AI
百度视觉 Baidu 团队 自动驾驶/OCR

📚 GitHub Awesome 系列

仓库 链接 内容
Awesome Computer Vision https://github.com/jbhuang0604/awesome-computer-vision 综合CV资源列表
Awesome Deep Vision https://github.com/kjw0612/awesome-deep-vision 深度学习CV
Awesome Object Detection https://github.com/amusi/awesome-object-detection 目标检测论文/代码
Awesome 3D Reconstruction https://github.com/openMVG/awesome_3DReconstruction 3D重建
Awesome SLAM SLAM相关 同步定位与建图
Awesome NeRF https://github.com/awesome-NeRF/awesome-NeRF NeRF资源大全
Awesome Vision Transformer Vision Transformer相关 ViT论文/代码
Awesome Face Recognition 人脸识别相关 人脸识别文献
Awesome Self-Supervised Learning https://github.com/jason718/awesome-self-supervised-learning 自监督学习
Awesome Diffusion Models https://github.com/hee9joon/Awesome-Diffusion-Models 扩散模型
Papers With Code https://paperswithcode.com/ 论文+代码+Benchmark
SOTA Models https://paperswithcode.com/sota 各任务SOTA排行榜

📈 关键数据集

图像分类

数据集 图片数 类别数 用途 链接
ImageNet-1K 1.2M 1000 通用分类 ImageNet
CIFAR-10/100 60K 10/100 小规模分类 https://www.cs.toronto.edu/~kriz/cifar.html
MNIST 70K 10 手写数字 Index of /exdb/mnist
Fashion-MNIST 70K 10 时尚分类 GitHub - zalandoresearch/fashion-mnist: A MNIST-like fashion product database. Benchmark · GitHub
Places365 1.8M 365 场景分类 http://places2.csail.mit.edu/

目标检测

数据集 图片数 类别数 标注方式 链接
COCO 330K 80 bbox+seg COCO - Common Objects in Context
PASCAL VOC 11.5K 20 bbox+seg The PASCAL Visual Object Classes Homepage
Open Images 9M 600 bbox+seg https://storage.googleapis.com/openimages/web/index.html
VisDrone 86K 10 无人机视角 https://github.com/VisDrone/VisDrone-Dataset
DOTA 280K 15 旋转框 https://captain-whu.github.io/DOTA/

图像分割

数据集 任务 链接
ADE20K 语义分割 https://groups.csail.mit.edu/vision/datasets/ADE20K/
Cityscapes 街景分割 https://www.cityscapes-dataset.com/
Pascal Context 场景理解 http://www.cs.stanford.edu/~roozbeh/pascal-context/
KITTI 自动驾驶 The KITTI Vision Benchmark Suite
COCO-Stuff 语义分割 GitHub - nightrome/cocostuff: The official homepage of the COCO-Stuff dataset. · GitHub

人脸

数据集 数量 链接
LFW 13K http://vis-www.cs.umass.edu/lfw/
MS-Celeb-1M 10M MS-Celeb-1M: Challenge of Recognizing One Million Celebrities in the Real World - Microsoft Research
VGGFace2 3.3M Visual Geometry Group - University of Oxford
MegaFace 1M MegaFace

3D视觉

数据集 用途 链接
ShapeNet 3D形状 ShapeNet
ModelNet 3D分类 Princeton ModelNet
ScanNet 室内3D ScanNet | Richly-annotated 3D Reconstructions of Indoor Scenes
KITTI 自动驾驶3D The KITTI Vision Benchmark Suite
Waymo Open 自动驾驶 https://waymo.com/open/
nuScenes 自动驾驶 https://www.nuscenes.org/
MegaDepth 3D重建 https://www.cs.cornell.edu/projects/megadepth/

🧰 开发工具与框架

核心框架

CV专用框架

框架 链接 特点
OpenCV OpenCV - Open Computer Vision Library 经典CV库
MMDetection GitHub - open-mmlab/mmdetection: OpenMMLab Detection Toolbox and Benchmark · GitHub 检测统一框架
MMSegmentation GitHub - open-mmlab/mmsegmentation: OpenMMLab Semantic Segmentation Toolbox and Benchmark. · GitHub 分割统一框架
MMPose GitHub - open-mmlab/mmpose: OpenMMLab Pose Estimation Toolbox and Benchmark. · GitHub 姿态估计
MMOCR https://github.com/open-mmlab/mmocr OCR工具箱
MMTracking https://github.com/open-mmlab/mmtracking 目标跟踪
Detectron2 https://github.com/facebookresearch/detectron2 Meta检测框架
Segmentation Models https://github.com/qubvel-org/segmentation_models.pytorch PyTorch分割
TIMM https://github.com/huggingface/pytorch-image-models 预训练模型库
PyTorch3D GitHub - facebookresearch/pytorch3d: PyTorch3D is FAIR's library of reusable components for deep learning with 3D data · GitHub 3D深度学习
Kornia GitHub - kornia/kornia: 🐍 Geometric Computer Vision Library for Spatial AI · GitHub 可微分CV库
FiftyOne GitHub - voxel51/fiftyone: Refine high-quality datasets and visual AI models · GitHub 数据集可视化工具

部署工具

工具 链接 用途
ONNX ONNX | Home 模型中间表示
TensorRT TensorRT SDK | NVIDIA Developer NVIDIA推理优化
OpenVINO https://docs.openvino.ai/ Intel推理引擎
TFLite https://www.tensorflow.org/lite 移动端部署
NCNN GitHub - Tencent/ncnn: ncnn is a high-performance neural network inference framework optimized for the mobile platform · GitHub 腾讯移动端推理
MNN GitHub - alibaba/MNN: MNN: A blazing-fast, lightweight inference engine battle-tested by Alibaba, powering high-performance on-device LLMs and Edge AI. · GitHub 阿里巴巴推理框架
TVM Apache TVM 端到端编译栈

🌐 学习平台与社区

平台 链接 特点
Papers With Code https://paperswithcode.com/ 论文+代码+排行榜
Hugging Face https://huggingface.co/models 模型库+数据集+Spaces
Kaggle https://www.kaggle.com/ CV竞赛+数据集+Notebook
GitHub https://github.com/ 源代码+复现
Google Colab https://colab.research.google.com/ 免费GPU
Roboflow https://roboflow.com/ CV数据管理与微调
CVF开放访问 https://openaccess.thecvf.com/ CVPR/ICCV全文
arXiv https://arxiv.org/list/cs.CV/recent 最新CV论文
Google Scholar https://scholar.google.com/ 引文搜索
知乎 https://www.zhihu.com/topic/19551392 中文CV讨论

🔗 论文下载与搜索

工具 链接
arXiv (cs.CV) https://arxiv.org/list/cs.CV/recent
arXiv Sanity https://arxiv-sanity-lite.com/ (Andrej Karpathy)
Papers With Code https://paperswithcode.com/
CVF Open Access https://openaccess.thecvf.com/
Semantic Scholar Semantic Scholar | AI-Powered Research Tool
Google Scholar https://scholar.google.com/
Connected Papers https://www.connectedpapers.com/
ResearchRabbit ResearchRabbit: AI Tool for Smarter, Faster Literature Reviews
Sci-Hub 非必要不使用
Aminer https://www.aminer.cn/

🔥 2024-2026 前沿趋势

方向 代表性工作 趋势说明
视觉基础模型 SAM 2, DINOv2, ViT-22B 大一统模型,few-shot/zero-shot
多模态大模型 GPT-4V, Gemini, LLaVA, Qwen-VL 视觉+语言融合
3D生成 3D Gaussian Splatting, Zero-1-to-3 3D AIGC爆发
视频生成 Sora, VideoPoet, Stable Video 文生视频/图生视频
可微渲染 NeRF, 3DGS, Flash3D 隐式/显式3D表示
自监督 DINOv2, I-JEPA, MAE 减少对标注的依赖
扩散模型 SD3, Flux, DDPM 图像生成主流范式
世界模型 UniSim, DayDreamer 视觉+物理仿真
平铺CNN/ConvNeXt ConvNeXt V2 CNN复兴,媲美ViT
端侧部署 MobileNetV4, EfficientViT 手机上跑大模型
对抗鲁棒性 PGD-AT, TRADES, Random Smoothing 让CV模型更安全
可解释性 Grad-CAM, LIME, SHAP, Attention Rollout 理解模型决策原因

📅 最后更新: 2026-06-17(新增对抗攻击与防御、可解释性模块)


附录:深层补充

一、2024-2026 最新顶会论文获取渠道

1.1 顶会论文搜索技巧
渠道 网址 使用技巧
CVF Open Access https://openaccess.thecvf.com 统一搜索 CVPR/ICCV/ECCV/WACV 全部论文全文
Dblp https://dblp.org 按会议年份过滤,如 dblp.org/db/conf/cvpr/cvpr2025.html
arXiv cs.CV Computer Vision and Pattern Recognition RSS 订阅 + 按更新时间追踪最新论文
Semantic Scholar https://www.semanticscholar.org API 批量搜索,结构化参考文献
Google Scholar https://scholar.google.com 引文追踪,作者主页(设置提醒,新论文自动邮件通知)
1.2 Papers with Code 的高级筛选

Papers with Code 不仅仅是论文+代码的配对,还提供强大的筛选功能:

按任务筛选

  • paperswithcode.com/task/object-detection 按不同 benchmark 查看 SOTA

  • 每个 benchmark 有 leaderboard 展示 Top 10 方法和数值

  • 可查看历史 SOTA 变迁(按时间线的精度提升曲线)

按数据集筛选

  • paperswithcode.com/datasets 查看数据集页面

  • 每个数据集展示其相关的 benchmark 和 SOTA 方法

按方法筛选

  • 支持搜索特定方法(如 'DETR')和对比不同方法的 benchmark 表现

RSS / Watchlist

  • 订阅特定任务/数据集,新 SOTA 出现时自动通知

1.3 Hugging Face Daily Papers
  • 每日前 5-10 篇最受关注的 arXiv 论文

  • 由社区共同投票(upvote)决定

  • 每天更新,覆盖 CV / NLP / 多模态 / 语音 / 强化学习

  • 配合 Hugging Face 的 model card 可直接试验

  • 链接:https://huggingface.co/papers


二、实用的研究工具

2.1 Connected Papers(论文关系图)

Connected Papers | Find and explore academic papers

输入一篇论文 → 生成论文影响力图谱

  • 节点:论文(越大代表引用越多或相关性越高)

  • :引用关系(共被引、相关引用的强引用关系)

  • 颜色:按时间(越深色越新)

  • 功能:Prior Works(该论文基于的早期工作)+ Derivative Works(后续发展)

用途:

  • 快速了解一个领域的边界和关键路径

  • 发现意料之外的跨领域联系

  • 为综述提供系统化的文献框架

2.2 Semantic Scholar(研究搜索)

https://www.semanticscholar.org/

AI 驱动的学术搜索引擎:

  • TLDR(Too Long; Didn't Read):AI 生成的一句话论文摘要

  • Citation Graph:引用网络的结构化展示

  • Influential Citations:标记出最有影响力的引用

  • API:支持批量论文搜索和元数据获取

  • Research Feed:基于你的阅读历史推荐新论文

2.3 Arxiv Sanity(论文推荐)

https://arxiv-sanity-lite.com/(Andrej Karpathy 维护)

  • 从 arXiv 最新论文中根据你的偏好推荐

  • 通过点赞和收藏学习你的兴趣

  • 每周邮件摘要

  • 比直接刷 arXiv 高效得多

2.4 科研 AI 助手
工具 功能 链接
Consensus 检索一句话的科学答案,附带引用 https://consensus.app
Galactica(开源) 面向科研的 LLM,支持公式/代码 https://galactica.org
Elicit 自动化文献综述搜索和对比 https://elicit.com
SciSpace 论文解读助手,解释公式/图表 https://typeset.io
Perplexity Pro 实时搜索+合成科研答案 https://perplexity.ai

三、中文社区资源

3.1 知乎 CV 话题
知乎专栏/话题 描述 链接
计算机视觉 话题 30 万关注,核心讨论区 zhihu.com/topic/19551392
AI 漫游指北 李沐的知乎专栏,高质量论文解读 zhihu.com/column/c_1251902395475275776
CV 前沿论文速递 每日/每周最新论文 zhihu.com/search?q=CVPR+2025+论文
计算机视觉与深度学习 黄浴老师的技术分析 知乎搜索"黄浴"
3.2 公众号推荐
公众号 特点
机器之心 每日 AI 新闻,覆盖全面
量子位 产业+技术趋势分析
CVer CV 入门到竞赛,中文教程最全
OpenMMLab OpenMMLab 官方,算法开源详解
AI 科技评论 业界深度访谈
Datawhale 公益学习组织,组团学 CV
极市平台 算法竞赛、工业应用
深度学习初学者(NewBeeNLP) 侧重 CV/DL 学习路径
3.3 国内技术博客
博客 地址
OpenMMLab 技术博客 https://github.com/open-mmlab/OpenMMLabBlog
Datawhale 教程 Datawhale · GitHub
动手学深度学习(李沐) https://d2l.ai/(中文版)
CS231n 中文笔记 ShowMeAI知识社区
GiantPandaCV 公众号 知乎/GitHub 搜 GiantPandaCV
B 站 CV 论文精讲 搜索"论文精讲" + "计算机视觉"
3.4 国内竞赛平台
平台 描述 链接
天池(阿里) CV 比赛 + 数据集 + GPU 算力 https://tianchi.aliyun.com
Kaggle CN(百度) Kaggle 中文版 https://kaggle.com(中文社区)
Biendata 国内数据竞赛 https://www.biendata.xyz
DataFountain 数据竞赛 https://www.datafountain.cn
Logo

Agent 垂直技术社区,欢迎活跃、内容共建。

更多推荐