Swin Transformer: Hierarchical Vision Transformer using Shifted Windows
Explore this paper's citation graph
- Type
- article
- Published
- 2021-10-01
- Cited by
- 35,186
- References
- 86
- Access
- Open access
- OpenAlex
- https://openalex.org/W3138516171
- Semantic Scholar
- https://api.semanticscholar.org/CorpusID:232352874
Keywords
Transformer, Computer science, Segmentation, Artificial intelligence, Computation
References
- Very Deep Convolutional Networks for Large-Scale Image Recognition
- Cognitron: A self-organizing multilayered neural network
- Acceleration of stochastic approximation by averaging
- Going deeper with convolutions
- ImageNet: A large-scale hierarchical image database
- Gradient-based learning applied to document recognition
- Receptive fields, binocular interaction and functional architecture in the cat's visual cortex
- ImageNet classification with deep convolutional neural networks
- Deep Residual Learning for Image Recognition
- Semantic Understanding of Scenes Through the ADE20K Dataset
- Aggregated Residual Transformations for Deep Neural Networks
- Feature Pyramid Networks for Object Detection
- Deformable Convolutional Networks
- Soft-NMS — Improving Object Detection with One Line of Code
- An Analysis of Scale Invariance in Object Detection - SNIP
- Cascade R-CNN: Delving Into High Quality Object Detection
- OCNet: Object Context Network for Scene Parsing
- Deformable ConvNets V2: More Deformable, Better Results
- Hybrid Task Cascade for Instance Segmentation
- Attention Augmented Convolutional Networks
Cited by
- Deep-learning-based pyramid-transformer for localized porosity analysis of hot-press sintered ceramic paste
- Object Detection in 20 Years: A Survey
- Testing Robustness Against Unforeseen Adversaries
- LIP: Local Importance-Based Pooling
- Harmonic Convolutional Networks based on Discrete Cosine Transform
- An end‐to‐end joint learning scheme of image compression and quality enhancement with improved entropy minimization
- JCS: An Explainable COVID-19 Diagnosis System by Joint Classification and Segmentation
- Delving Deeper into Anti-Aliasing in ConvNets
- Efficient Transformers: A Survey
- Why I'm not Answering: Understanding Determinants of Classification of an Abstaining Classifier for Cancer Pathology Reports
- Long-tailed Recognition by Routing Diverse Distribution-Aware Experts
- Parameter-Efficient Person Re-Identification in the 3D Space
- End-to-End Object Detection with Adaptive Clustering Transformer
- Transformers in Vision: A Survey
- Salient Object Detection via Integrity Learning
- Active Boundary Loss for Semantic Segmentation
- Gaining Scale Invariance in UAV Bird’s Eye View Object Detection by Adaptive Resizing
- Synergy between Semantic Segmentation and Image Denoising via Alternate Boosting
- Multi-Scale Vision Longformer: A New Vision Transformer for High-Resolution Image Encoding
- Augmented Transformer with Adaptive Graph for Temporal Action Proposal Generation
Related papers
- Low complexity photo sensor dead pixel detection algorithm
- Sub-pixel mapping based on sub-pixel to sub-pixel spatial attraction model
- Bad pixel identification by means of principal components analysis
- Improved K-Pass Pixel Value Ordering Based Data Hiding
- Sub-pixel mapping of remotely sensed imagery with hybrid intra- and inter-pixel dependence
- Handling of Bad Pixels on Pixelated Solid State Detectors
- New discrepancy measures for segmentation evaluation