Grid Anchor based Image Cropping: A New Benchmark and An Efficient Model

Zeng, Hui; Li, Lida; Cao, Zisheng; Zhang, Lei

Computer Science > Computer Vision and Pattern Recognition

arXiv:1909.08989 (cs)

[Submitted on 18 Sep 2019]

Title:Grid Anchor based Image Cropping: A New Benchmark and An Efficient Model

Authors:Hui Zeng, Lida Li, Zisheng Cao, Lei Zhang

View PDF

Abstract:Image cropping aims to improve the composition as well as aesthetic quality of an image by removing extraneous content from it. Most of the existing image cropping databases provide only one or several human-annotated bounding boxes as the groundtruths, which can hardly reflect the non-uniqueness and flexibility of image cropping in practice. The employed evaluation metrics such as intersection-over-union cannot reliably reflect the real performance of a cropping model, either. This work revisits the problem of image cropping, and presents a grid anchor based formulation by considering the special properties and requirements (e.g., local redundancy, content preservation, aspect ratio) of image cropping. Our formulation reduces the searching space of candidate crops from millions to no more than ninety. Consequently, a grid anchor based cropping benchmark is constructed, where all crops of each image are annotated and more reliable evaluation metrics are defined. To meet the practical demands of robust performance and high efficiency, we also design an effective and lightweight cropping model. By simultaneously considering the region of interest and region of discard, and leveraging multi-scale information, our model can robustly output visually pleasing crops for images of different scenes. With less than 2.5M parameters, our model runs at a speed of 200 FPS on one single GTX 1080Ti GPU and 12 FPS on one i7-6800K CPU. The code is available at: \url{this https URL}.

Comments:	Extension of a CVPR 2019 paper. Dataset and PyTorch Code are released. arXiv admin note: substantial text overlap with arXiv:1904.04441
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:1909.08989 [cs.CV]
	(or arXiv:1909.08989v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.1909.08989

Submission history

From: Hui Zeng [view email]
[v1] Wed, 18 Sep 2019 14:41:42 UTC (4,911 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Grid Anchor based Image Cropping: A New Benchmark and An Efficient Model

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Grid Anchor based Image Cropping: A New Benchmark and An Efficient Model

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators