This dataset is a benchmark used in the experiments of the paper "Functional Semantics Embedding of GUI Screens for Knowledge-Augmented GUI Agents," submitted to ACM MM 2026.
This benchmark also includes MLLM prompts for step-level action decision, as well as prompts used to generate knowledge.
hf download user83kd9x/knowledge_agent_benchmark
--repo-type dataset
--local-dir… See the full description on the dataset page:
https://huggingface.co/datasets/user83kd9x/knowledge_agent_benchmark.