Downloads · 30 days
0
LujianYao/PPC
PPC is a machine learning model from LujianYao. Use it for the machine learning task on the model card, and read the license before you ship it in a product.
<a href="https://arxiv.org/abs/2505.20655"<img src='https://img.shields.io/badge/arXiv-2505.20655-red?style=flat&logo=arXiv&logoColor=red' alt='arxiv'</a <a href="https://github.com/vivoCameraResearch/p-p-c"<img…
Downloads · 30 days
0
Access
Public
Updated Mar 3, 2026
Repo size
314 MB
Likes
1
Public
Click a slice to open those files.
.ckpt307 MB · 98%
From the Hugging Face model README
<a href="https://arxiv.org/abs/2505.20655"><img src='https://img.shields.io/badge/arXiv-2505.20655-red?style=flat&logo=arXiv&logoColor=red' alt='arxiv'></a> <a href="https://github.com/vivoCameraResearch/p-p-c"><img src='https://img.shields.io/badge/GitHub-Code-blue?logo=github' alt='GitHub'></a> <a href="https://vivocameraresearch.github.io/ppc/"><img src='https://img.shields.io/badge/Project-Page-Green' alt='Project Page'></a>
Traditional photography composition approaches are dominated by 2D cropping-based methods. However, these methods fall short when scenes contain poorly arranged subjects. Professional photographers often employ perspective adjustment as a form of 3D recomposition, modifying the projected 2D relationships between subjects while maintaining their actual spatial positions to achieve better compositional balance. Inspired by this artistic practice, we propose photography perspective composition (PPC), extending beyond traditional cropping-based methods. However, implementing the PPC faces significant challenges: the scarcity of perspective transformation datasets and undefined assessment criteria for perspective quality. To address these challenges, we present three key contributions: (1) An automated framework for building PPC datasets through expert photographs. (2) A video generation approach that demonstrates the transformation process from suboptimal to optimal perspectives. (3) A perspective quality assessment (PQA) model constructed based on human performance. Our approach is concise and requires no additional prompt instructions or camera trajectories, helping and guiding ordinary users to enhance their composition skills.