Pyramid Vision Transformer #17596

danielhoshizaki · 2022-06-08T02:26:58Z

Model description

I would like to add the Pyramid Vision Transformer model.

Paper Abstract

Pyramid Vision Transformer~(PVT), has several merits compared to prior arts. (1) Different from ViT that typically has low-resolution outputs and high computational and memory cost, PVT can be not only trained on dense partitions of the image to achieve high output resolution, which is important for dense predictions but also using a progressive shrinking pyramid to reduce computations of large feature maps. (2) PVT inherits the advantages from both CNN and Transformer, making it a unified backbone in various vision tasks without convolutions by simply replacing CNN backbones. (3) We validate PVT by conducting extensive experiments, showing that it boosts the performance of many downstream tasks

Open source status

The model implementation is available
The model weights are available

Provide useful links for the implementation

Model Implementation: https://github.com/whai362/PVT

Pretrained model weights for semantic segmentation: https://github.com/whai362/PVT/tree/v2/segmentation (based on ADE20K)

satpalsr · 2022-06-10T17:21:35Z

Hey @danielhoshizaki, Can I join you?

andy1122 · 2022-06-14T08:32:38Z

@danielhoshizaki , I would love to contribute for this new model.

danielhoshizaki · 2022-07-12T00:33:36Z

Thanks for offering to help and sorry about the late response.
I'm going to have a go at this model and if I need any help I will let you know.

satpalsr · 2022-07-22T04:20:45Z

Sure @danielhoshizaki

BakingBrains · 2023-01-15T06:56:26Z

anyone working on this?

danielhoshizaki added the New model label Jun 8, 2022

Xrenya mentioned this issue Mar 29, 2023

Add PVT(Pyramid Vision Transformer) #22445

Closed

5 tasks

Xrenya mentioned this issue Jul 8, 2023

Pvt model #24720

Merged

5 tasks

amyeroberts closed this as completed in #24720 Jul 24, 2023

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Pyramid Vision Transformer #17596

Pyramid Vision Transformer #17596

danielhoshizaki commented Jun 8, 2022 •

edited

Loading

satpalsr commented Jun 10, 2022

andy1122 commented Jun 14, 2022

danielhoshizaki commented Jul 12, 2022

satpalsr commented Jul 22, 2022

BakingBrains commented Jan 15, 2023

Pyramid Vision Transformer #17596

Pyramid Vision Transformer #17596

Comments

danielhoshizaki commented Jun 8, 2022 • edited Loading

Model description

Paper Abstract

Open source status

Provide useful links for the implementation

satpalsr commented Jun 10, 2022

andy1122 commented Jun 14, 2022

danielhoshizaki commented Jul 12, 2022

satpalsr commented Jul 22, 2022

BakingBrains commented Jan 15, 2023

danielhoshizaki commented Jun 8, 2022 •

edited

Loading