Skip to content
Review

StagedWorkspace: A Versioned Workspace for Knowledge-Work Agents

Aug 2026 · 0 citations · 39 references
Computer Science

TL;DR

StagedWorkspace is proposed, a versioned workspace for knowledge-work agents that binds parsed records and review diffs to content hashes of the native files as they change, and motivates benchmarks that score evidence, staged edits, and submitted artifacts as explicit state transitions.

Abstract

AI agents increasingly perform knowledge work (i.e., produce and modify persistent digital artifacts such as code repositories, documents, spreadsheets, slides, reports), yet the parsed views they search, the native files they edit, the changes they review, and the artifacts they submit can refer to different versions of the same work product. We formulate this as a workspace-state contract: every view should be explicitly tied to a version of the evolving workspace state. Coding agents partly address this need through repository contracts for search, diffs, and tests, whereas an analogous contract is less explicit for PDFs, spreadsheets, slides, notebooks, and mixed-format project folders. We propose StagedWorkspace, a versioned workspace for knowledge-work agents. The workspace binds parsed records and review diffs to content hashes of the native files as they change. In fixed-harness ablations on OfficeQA Pro and APEX-Agents, dual parsed/native access has the highest point estimate for every tested model; relative to the more limiting single view, it improves OfficeQA Pass@1 by 8.3-12.1 points and APEX mean rubric score by 4.7-9.2 points. SW-AGENT scores 63.9% with Gemini 3.1 Pro on OfficeQA and 42.1 with GPT-5.4 Nano on APEX, compared with published same-model scores of 29.3% and 25.5, respectively. A paired review-axis ablation on 57 file-editing tasks further finds higher observed scores when diffs are visible. These results identify workspace state as an experimental variable in knowledge-work agents and motivate benchmarks that score evidence, staged edits, and submitted artifacts as explicit state transitions.

View source

Similar papers

Open access Sep 2026

Reimagining research papers as interactive and reliable AI agents.

Here we introduce Paper2Agent, an automated framework that converts research papers into artificial intelligence (AI) agents. Paper2Agent transforms research output from passive artefacts into active systems that accelerate use and discovery. Conventional research papers require readers to understand and adapt the pape...

Jia-Cheng Miao, Joe R. Davis, Yaohui Zhang et al. · 2 citations
Review Aug 2026

Agentic Artifact Creation: Systems, Evaluation, Principles, and Opportunities

This survey examines agentic artifact creation, which is defined as stateful construction in which an AI system materially constructs or revises a deliverable and intermediate observations redirect later work, and formulate principles for keeping commitments and responsibility explicit, turning feedback into targeted r...

Tianfu Wang, Zhezheng Hao, Xinchi Xia et al. · 0 citations
#artificial intelligence Preprint Aug 2026

Dr. Claw: An AI Scientist Workspace for Vibe Research

Dr. Claw is presented, an open-source workspace that wraps existing coding-agent executors in a controllable and auditable human-in-the-loop workflow rather than introducing another autonomous agent.

D. Song, Han-Rong Zhang, Dawei Liu et al. · 0 citations
Preprint Aug 2026

ReproAgent: Contract-Guided Paper-to-Code Reproduction

ReproAgent is introduced, a four-stage Prepare--Plan--Generate--Repair pipeline built around a persistent implementation contract with two channels: an implementation-requirement channel that turns paper snippets into code obligations, and a reference-evidence channel that retrieves content and structure evidence from...

Xueyan Hu, Ze-Wei Pan, Zhongyuan Wang et al. · 0 citations
Jul 2026

CodeNib: A Multi-View Data System for Serving Repository Context to Coding Agents

Together, these results support multi-view repository-context serving with explicit, operation-specific validity boundaries with quality-cost frontiers across the repository-context lifecycle.

Zhongming Yu, Hengjia Yu, Boqin Yuan et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.