Screen Parsing: Towards Reverse Engineering of UI Models from Screenshots


연구 분야: Analysis



학회: UIST '21: The 34th Annual ACM Symposium on User Interface Software and Technology


초록

Automated understanding of user interfaces (UIs) from their pixels can improve accessibility, enable task automation, and facilitate interface design without relying on developers to comprehensively provide metadata. A first step is to infer what UI elements exist on a screen, but current approaches are limited in how they infer how those elements are semantically grouped into structured interface definitions. In this paper, we motivate the problem of screen parsing, the task of predicting UI elements and their relationships from a screenshot. We describe our implementation of screen parsing and provide an effective training procedure that optimizes its performance. In an evaluation comparing the accuracy of the generated output, we find that our implementation significantly outperforms current systems (up to 23%). Finally, we show three example applications that are facilitated by screen parsing: (i) UI similarity search, (ii) accessibility enhancement, and (iii) code generation from UI screenshots.


Author Profile
Jason Wu

Human-Computer Interaction Institute Carnegie Mellon University United States

United States
Author Profile
Xiaoyi Zhang

Apple United States

United States
Author Profile
Jeffrey Nichols

Apple United States

United States

📄 논문 정보

발행 연도 2021년
인용수 44
출판 국가 United States
사이트 ACM
좋아요 수 0

연관 논문 목록 (35건)