workflow traces for AI agents

Computer-use agent evals from real workflows.

Screenpipe captures how people actually complete desktop work, then turns that trace into SOPs, prompts, expected outcomes, edge cases, and evals for computer-use agents.

01

Capture the human run

Observe the workflow across browser, desktop apps, meetings, ERP, CRM, spreadsheets, and internal tools.

02

Convert to SOP and spec

Extract steps, inputs, decision points, expected output, edge cases, and acceptance criteria.

03

Build the eval case

Package the trace into a task an agent can attempt and a grader can judge against a real outcome.

04

Improve the dataset

Keep examples, failures, variants, and confidence notes tied to the workflow owner and deployment context.

eval artifact

Build an eval from a recorded workflow.

TaskUpdate a CRM opportunity after a customer call.
InputsMeeting transcript, browser state, CRM fields, Slack thread, previous account notes.
Expected outcomeOpportunity stage, next step, owner, call summary, and follow-up date are correct.
Failure modesWrong account, hallucinated next step, missed pricing objection, duplicate record.
Privacy scopeOnly approved fields and redacted excerpts are included in the eval package.

deployment modes

Local-first does not mean one data path.

Screenpipe can run as a personal assistant or a scoped team deployment. The important question for buyers is not a slogan; it is which data flow they approve.

Local-only

What stays local
Screen capture, accessibility text, OCR output, audio files, transcripts, and the local database.
What may leave the device
Nothing is required to leave the device for core capture and search.
Buyer decision
Best for self-serve use, regulated pilots, and proving value before any cloud path is enabled.

Local + optional cloud AI

What stays local
The raw capture store remains on the endpoint unless the user or organization enables export or sync.
What may leave the device
Selected prompts, summaries, or context snippets may be sent to the chosen AI provider or confidential route.
Buyer decision
Buyer chooses model, provider, retention posture, redaction, and whether local models are required.

Team / enterprise

What stays local
Endpoint capture and local history can stay on managed devices under admin policy.
What may leave the device
Team reports, sync, admin workflows, exports, connectors, and agent outputs depend on deployment scope.
Buyer decision
Buyer defines consent, retention, employee controls, report contents, and admin visibility.

point of view

Use capture to decide what to automate.

Screenpipe helps teams identify repeated work, assess what to automate, and define tasks for agents. The report also records which data paths the buyer approves.

Start with one department

Deploy to up to 10 machines in one team. Capture their work to build a workflow map.

The useful data lives between systems

ERP, CRM, and ticketing logs can miss the spreadsheets, tabs, messages, meetings, and decisions between systems. Those steps can reveal automation opportunities.

Use recorded workflows to specify agent tasks

A usable computer-use agent spec needs real inputs, expected outcomes, edge cases, failure modes, and a way to grade the result.

Privacy is part of the deliverable

A workflow report should say what was captured, excluded, redacted, retained, exported, and shared before the team expands deployment.

Privacy is part of the eval design.

Decide what to capture, exclude, and redact, and how long to keep it. Also define what can leave the device: raw examples, derived SOPs, or only acceptance criteria. These choices belong in the deployment plan.