Skip to content

Product vision ​

Your agent's copilot. Your team's workspace.

Cully has two product modes around the coding agents people already use. They share the same CLI, terminal and services; these are workflows, not new --solo or --team setup flags.

Solo mode ​

A copilot for your coding agent and work environment. Understand session health, notice loops and unchecked edits, recover private context, replay what happened and continue in another agent. The journal stays local; saved memory stays owner-scoped. Immediate advisor warnings work offline. Start with Solo mode.

Team mode ​

Keep your team's agent work moving, whoever picks it up next. A task outlives any session, person or agent: its outcome, accountable person, attempts, decisions, checkpoint, review evidence and useful lessons remain connected. Members claim work and hand it off; a noncontributing maintainer accepts delivery and adopts explicitly published guidance. Private solo notes and journals remain separate.

The first team workflow uses CLI/MCP and manual agent launch. It implements issuer-scoped OAuth identity, explicit project roles, atomic versioned claims, dependencies, board/inbox, checkpoints, reported review evidence and project lessons/playbooks. Managed execution and automatic learning remain planned.

The adoption hypothesis is that connected tasks reduce repeated investigation and explicit evidence improves review confidence. Validate those outcomes with real teams before claiming savings or fixing pricing. Measure handoff time, review waiting time, reopened tasks and lesson usefulness, with counts and limitations.

This page separates shipped work from proposals. Proposals are not promises and carry no dates. For the exact list of what exists today, see current capabilities.

text
Terminal and hooks ──► Session journal ──► Session intelligence ──► Advisor and review
 (observe)              (understand)         (decide)                 (advise)
                                                  │
                                                  └──► Project memory (remember) ──► Improve

Shipped ​

CapabilityWhat it gives you
One terminal for every agentClaude Code, Codex, Cursor CLI and any executable share one wrapper
Agent health barAgent, project, time, context, loops and unchecked edits at a glance
Session journalA private, coarse record every other feature reads
Loop detection"The same command failed 3 times after edits"
cully replay and cully timelineWhat the agent did, step by step, and which files it touched
cully handoffA structured handoff for the next agent
cully rescueEvidence and recovery steps when a session is stuck
cully statusMeasured session health and a rule-based risk label
Session forensicsFlight-recorder postmortems, audit-grade review and replayable playbooks from the same record
Project workflow hint"You usually run checks after editing here"
Team workspace through CLI/MCPAuthorized project roles, stable tasks, board/inbox, claims and selected checkpoints
Task acceptanceReported evidence for each criterion and revision/version-pinned maintainer approval
Explicit project learningPrivate drafts, publication/withdrawal and maintained source-linked playbooks

Next ​

Work that builds on the journal and needs moderate additions.

Guardrails ​

User problemAn agent can run a destructive command or touch a protected path before you notice.
ExperienceA short rule file in the project. When a tool call matches, Cully warns in the health bar, and for agents with pre-tool hooks it can ask before the call runs.
ReusesThe hook path, the journal, the advisor panel.
MissingPre-tool hooks for each agent, a rule format, an approval prompt in the terminal.
ComplexityMedium.
DifferentiationHigh. Observation becomes control.

Cully Bench ​

User problemCully cannot yet say whether it saves time or tokens.
ExperienceA published benchmark of multi-session tasks with and without Cully.
ReusesThe journal for time to first useful action, repeated searches and repeated failures.
MissingA task suite, a harness, a way to score continuation correctness.
ComplexityMedium to high.
DifferentiationCredibility. Replaces "ship faster" with numbers.

A lighter local install ​

User problemTrying Cully starts PostgreSQL, Mem0, a data API and MCP in Docker.
ExperienceA single-binary mode that needs no Docker for the terminal, journal, replay, handoff and rescue, with memory added when you want it.
ReusesEverything in the terminal and journal already runs without the memory stack.
MissingA lightweight memory store, or a clear "terminal only" install path.
ComplexityMedium.
DifferentiationLowers the first-run cost.

Session report ​

User problemAfter a long session you want a summary for a pull request.
Experiencecully report renders the journal as a short review note: what changed, what was verified, what is open.
ReusesReplay, handoff, health.
MissingA format and a place to attach it.
ComplexityLow to medium.
DifferentiationMedium.

Team workspace target ​

The manual-launch workflow is the shipped Team slice. Existing private sessions and task links remain owner-scoped. Team work uses separate records and explicit authorization; Cully does not intercept model traffic.

Demo and commercial validation ​

Demonstrate one task: one developer starts it, another continues in a different agent, the reviewer sees missing evidence, the checks are supplied, and an accepted lesson improves a later task. Include a private note that never appears to the teammate and a membership removal that stops new reads.

Measure time to first useful action after handoff, repeated investigation, time waiting for review, reopened tasks and playbook usefulness. Compare like-for-like work with and without Cully. Show counts and denominators; no arbitrary productivity score. Validate willingness to adopt before fixing packaging or pricing.

Planned team stages ​

StageDeliverableGate
Managed executionCapability-aware adapters, isolated worktrees, bounded runs, cancellation and recoveryDuplicate events do not duplicate runs; unsupported controls are explicit
Shared learning extensionsTeam-wide visibility, semantic shared retrieval, ranking and optional background draft suggestionsPrivate notes never enter shared retrieval; edits/withdrawals invalidate derived tips
Browser and trackersWeb board/inbox and GitHub/Linear-style artifact syncSigned events, delivery deduplication and external permission checks before bidirectional writes
Measured improvementTeam workflow suggestions and agent comparisons from verified outcomesShow cohort size and confounders before any automatic routing

Later ​

Ideas that need new data, new instrumentation or a lot of history first.

Learned team workflow ​

User problem"We always run a security review after authorization changes" lives only in people's heads.
ExperienceWhen a session type usually ends with certain steps and some are missing, Cully says so.
ReusesThe journal, project memory, the existing workflow hint.
MissingClassifying work without reading prompts and collecting enough verified outcomes to support useful patterns.
ComplexityHigh.
DifferentiationHigh, if it is accurate.

Shared learning workers ​

User problemUseful lessons stay private drafts or drown in a long project history.
ExperienceOpt-in preparation suggests draft lessons and ranks currently authorized tips; publishing and playbook adoption stay explicit.
ReusesProject lessons, playbooks and live authorization checks from Team mode.
MissingBounded job IDs, cancellation, membership-aware caches and Mem0 hydration that never exposes private notes.
ComplexityHigh.
DifferentiationHigh if withdrawal and access removal stay correct.

Agent performance intelligence ​

User problemWhich agent should take this task?
ExperienceA comparison across work types from your own history.
ReusesThe journal records which agent did what.
MissingMeasured outcomes: task completion, rework, human intervention and test results, over hundreds of sessions.
ComplexityHigh.
DifferentiationHigh.
PrincipleNo arbitrary scores. Any comparison must come from signals Cully measures, with the sample size shown.

An agent cockpit ​

User problemSeveral agents run at once and you lose track.
ExperienceOne view of every running session, its health and its open risks.
ReusesThe terminal, the health bar, the journal.
MissingA multi-session view, in the terminal or the browser.
ComplexityHigh.
DifferentiationOthers are building cockpits. Cully's edge would be understanding the work, not the panes.

Principles ​

  • Do not make a commodity feature the identity. Memory alone, a terminal alone, an MCP server alone or an advisor alone is not Cully. The combination is.
  • Measure before you claim. Anything shown as a number comes from a signal Cully has.
  • Stay around the agent. Cully is not another coding agent and does not intercept your agent's model traffic.
  • Keep it private by default. New data must justify itself in privacy.

Built in the open. Apache 2.0 licensed.