Identifying User Goals from UI Trajectories

Omri Berkovitch; Sapir Caduri; Noam Kahlon; Anatoly Efros; Avi; Caciularu; Ido Dagan

arXiv:2406.14314·cs.CL·March 4, 2025

Identifying User Goals from UI Trajectories

Omri Berkovitch, Sapir Caduri, Noam Kahlon, Anatoly Efros, Avi, Caciularu, Ido Dagan

PDF

Open Access

TL;DR

This paper introduces a new task for identifying user goals from UI trajectories, proposes an evaluation methodology, and benchmarks human and AI performance, revealing current models lag behind humans.

Contribution

It presents a novel goal identification task from UI trajectories, along with an evaluation method and benchmark results comparing humans and AI models.

Findings

01

GPT-4 and Gemini-1.5 Pro underperform humans

02

New evaluation metric for paraphrase detection in UI context

03

Significant room for improvement in AI goal inference

Abstract

Identifying underlying user goals and intents has been recognized as valuable in various personalization-oriented settings, such as personalized agents, improved search responses, advertising, user analytics, and more. In this paper, we propose a new task goal identification from observed UI trajectories aiming to infer the user's detailed intentions when performing a task within UI environments. To support this task, we also introduce a novel evaluation methodology designed to assess whether two intent descriptions can be considered paraphrases within a specific UI environment. Furthermore, we demonstrate how this task can leverage datasets designed for the inverse problem of UI automation, utilizing Android and web datasets for our experiments. To benchmark this task, we compare the performance of humans and state-of-the-art models, specifically GPT-4 and Gemini-1.5 Pro, using our…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsUsability and User Interface Design

MethodsRefunds@Expedia|||How do I get a full refund from Expedia? · Attention Is All You Need · Cosine Annealing · Byte Pair Encoding · Label Smoothing · Attention Dropout · Position-Wise Feed-Forward Layer · Dropout · Adam · Linear Warmup With Cosine Annealing