← 返回论文检索
ACM Multimedia 2024Technical Demonstrations

Unlimited Vision: Professional Composition by Yourself

Xin Jin 0015, Liaoruxing Zhang, Longteng Jiang, Dandan Li

PDF 由论文原始站点提供,PaperCompass 不保存论文文件。DOI 10.1145/3664647.3685001 ↗

摘要

This paper introduces a novel method for enhancing image composition guidance in photography. It utilizes advanced composition rules to guide a Real-Time Detection Transformer (RT-DETR) model in predicting aesthetically pleasing compositions for photographs. Unlike traditional methods constrained by original image boundaries, our approach allows the predicted framing to extend beyond these limits, offering dynamic, real-time guidance for image composition in photography. The system integrates multi-label composition classification and compositional element annotation, using YOLOv8 for key object detection and an enhanced Deep Hough Transform for compositional lines to guide photographers. It provides photographers with real-time guidance for optimal camera adjustments, transforming traditional post-processing tasks into an intuitive, interactive process. This method significantly enhances photographers' flexibility and effectiveness in capturing visually superior photographs.