OCR 与文档智能研究 Wiki

Tag: cv

12 items with this tag.

  • 2026年8月06日

    Learning Transferable Visual Models From NL Supervision (CLIP)

    • multimodal
    • cv
    • nlp
    • milestone
  • 2026年8月06日

    Masked Autoencoders Are Scalable Vision Learners (MAE)

    • ssl
    • cv
    • transformer
    • milestone
  • 2026年8月06日

    An Image is Worth 16x16 Words (Vision Transformer, ViT)

    • transformer
    • cv
    • milestone
  • 2026年7月25日

    ImageNet Classification with Deep CNNs (AlexNet)

    • cnn
    • cv
    • imagenet
    • milestone
  • 2026年7月23日

    Denoising Diffusion Probabilistic Models (DDPM)

    • generative
    • diffusion
    • cv
    • milestone
  • 2026年7月23日

    Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks

    • cv
    • detection
    • rpn
    • anchor
    • milestone
  • 2026年7月23日

    Going Deeper with Convolutions (GoogLeNet / Inception v1)

    • cnn
    • cv
    • imagenet
    • inception
    • milestone
  • 2026年7月23日

    Mask R-CNN

    • cv
    • instance-segmentation
    • detection
    • roialign
    • milestone
  • 2026年7月23日

    Deep Residual Learning for Image Recognition (ResNet)

    • resnet
    • cnn
    • cv
    • foundation
    • milestone
  • 2026年7月23日

    U-Net: Convolutional Networks for Biomedical Image Segmentation

    • cv
    • segmentation
    • biomedical
    • encoder-decoder
    • milestone
  • 2026年7月23日

    Very Deep Convolutional Networks for Large-Scale Image Recognition (VGG)

    • cnn
    • cv
    • imagenet
    • depth
    • milestone
  • 2026年7月23日

    You Only Look Once: Unified, Real-Time Object Detection (YOLO)

    • cv
    • detection
    • real-time
    • one-stage
    • milestone

12 items with this tag.

  • 2026年8月06日

    Learning Transferable Visual Models From NL Supervision (CLIP)

    • multimodal
    • cv
    • nlp
    • milestone
  • 2026年8月06日

    Masked Autoencoders Are Scalable Vision Learners (MAE)

    • ssl
    • cv
    • transformer
    • milestone
  • 2026年8月06日

    An Image is Worth 16x16 Words (Vision Transformer, ViT)

    • transformer
    • cv
    • milestone
  • 2026年7月25日

    ImageNet Classification with Deep CNNs (AlexNet)

    • cnn
    • cv
    • imagenet
    • milestone
  • 2026年7月23日

    Denoising Diffusion Probabilistic Models (DDPM)

    • generative
    • diffusion
    • cv
    • milestone
  • 2026年7月23日

    Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks

    • cv
    • detection
    • rpn
    • anchor
    • milestone
  • 2026年7月23日

    Going Deeper with Convolutions (GoogLeNet / Inception v1)

    • cnn
    • cv
    • imagenet
    • inception
    • milestone
  • 2026年7月23日

    Mask R-CNN

    • cv
    • instance-segmentation
    • detection
    • roialign
    • milestone
  • 2026年7月23日

    Deep Residual Learning for Image Recognition (ResNet)

    • resnet
    • cnn
    • cv
    • foundation
    • milestone
  • 2026年7月23日

    U-Net: Convolutional Networks for Biomedical Image Segmentation

    • cv
    • segmentation
    • biomedical
    • encoder-decoder
    • milestone
  • 2026年7月23日

    Very Deep Convolutional Networks for Large-Scale Image Recognition (VGG)

    • cnn
    • cv
    • imagenet
    • depth
    • milestone
  • 2026年7月23日

    You Only Look Once: Unified, Real-Time Object Detection (YOLO)

    • cv
    • detection
    • real-time
    • one-stage
    • milestone

Created with Quartz © 2026

  • GitHub