DevicesWorld: Benchmarking Cross-Device Agents in Heterogeneous Environments
LLM-based agents have rapidly improved at operating individual digital environments such as mobile applications, desktop systems, and smart homes. However, real-world user goals often span multiple devices: information may come from a phone, be processed on a desktop, and the result may need to appear on another device. Most existing benchmarks center on a single dominant execution environment, making it difficult to evaluate whether agents can acquire and integrate information across heterogene
Record details
Published: 15 July 2026
Source: arXiv cs.HC
Category: Research
Topics: Agents & autonomy · Environment
Retrieved: 16 July 2026
Related evidence
These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.
AgentSociety 2: An Integrated Research Environment for Executable Social Science
arXiv cs.CY · 15 July 2026
Agile perceptive multi-skill locomotion for quadrupedal robots in the wild
arXiv cs.AI · 15 July 2026
SAFETY SENTRY: Context-Aware Human Intervention via EXECUTE-ASK-REFUSE Routing
arXiv cs.AI · 15 July 2026
UESF-Bench: Benchmarking and Probing for Unified Embodied Seeking and Following
arXiv cs.AI · 15 July 2026
From Language to Navigation Goals: A Vision-Language Approach for Semantic Navigation of Mobile Robots Using RGB-D Perception
arXiv cs.AI · 15 July 2026
Object-centric diffusion policies for real-world robotic-arm imitation learning
Frontiers in Robotics and AI · 15 July 2026
How to cite this record
ethics.ai (15 July 2026), “DevicesWorld: Benchmarking Cross-Device Agents in Heterogeneous Environments,” evidence record 10926, https://ethics.ai/record/10926 (originally published by arXiv cs.HC).
Use and limitations
This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.