添加链接
link管理
链接快照平台
  • 输入网页链接,自动生成快照
  • 标签化管理网页链接
相关文章推荐
斯文的刺猬  ·  CVPR 2023 | 最全 AIGC ...·  2 周前    · 
鼻子大的酱牛肉  ·  mmpose · PyPI·  1 月前    · 
安静的八宝粥  ·  ICCV, ECCV, ...·  2 月前    · 
愤怒的伤疤  ·  PicGo-Core 和 Typora - ...·  2 周前    · 
重情义的西红柿  ·  bouncycastle.org·  1 月前    · 
爱玩的回锅肉  ·  Query Failing With ...·  3 月前    · 
Abstract:

In panorama understanding, the widely used equirectangular projection (ERP) entails boundary discontinuity and spatial distortion. It severely deteriorates the conventional CNNs and vision Transformers on panoramas. In this paper, we propose a simple yet effective architecture named PanoSwin to learn panorama representations with ERP. To deal with the challenges brought by equirectangular projection, we explore a pano-style shift windowing scheme and novel pitch attention to address the boundary discontinuity and the spatial distortion, respectively. Besides, based on spherical distance and Cartesian coordinates, we adapt absolute positional encodings and relative positional biases for panoramas to enhance panoramic geometry information. Realizing that planar image understanding might share some common knowledge with panorama understanding, we devise a novel two-stage learning framework to facilitate knowledge transfer from the planar images to panoramas. We conduct experiments against the state-of-the-art on various panoramic tasks, i.e., panoramic object detection, panoramic classification, and panoramic layout estimation. The experimental results demonstrate the effectiveness of PanoSwin in panorama understanding.