pr_01m47d15m3e54sn21z27rpy5n9/docs/PLAN.md

896 lines50,167 bytesCodeBlame
1# g1t plan
2
3g1t is a git forge for agents, built on Cloudflare Workers and Artifacts for
4the "Build the Next-Gen Git Platform on Cloudflare" competition.
5
6- Submission closes **October 14, 2026, 11:59 PM PDT**: a 5–10 minute demo
7 video, this repository (MIT) and run instructions.
8- Judging: 50% originality and quality of the prototype for agent-oriented
9 collaboration; 25% multi-agent concurrency, coordination, context
10 preservation, review and conflict handling; 25% ease of use.
11
12## The point
13
14**GitHub is where people keep code. g1t is where a team of agents ships it.**
15
16Hosting git is table stakes, and g1t does it the way GitHub does: issues,
17branches, pull requests, review, protected branches, people working by hand.
18None of that is the selling point. The selling point is the layer above it,
19which no forge has: **you hand g1t an outcome, and a fleet of agents converges
20it onto `main`, coordinating with each other and with you, with every decision
21on the record.**
22
23Three things only g1t does, and every feature should serve one of them:
24
251. **Outcomes, not pull requests.** The unit people work in is "make onboarding
26 work offline", not branch #4012. A brief becomes a plan of issues with
27 dependencies; agents take them as they unblock; people steer the outcome
28 and see it converge. GitHub, Origin and Entire all stop at the pull request.
292. **Agents that work as a team.** Agents know what the others are doing, file
30 what they find instead of widening their change, ask and answer each other
31 through the forge, and defer to people. Many agents on one codebase without
32 a human refereeing collisions.
333. **`main` that only ever moves forward.** Every change lands through checks
34 in the combination it will live in (the merge queue), failures go back to the
35 agent that wrote them, and any line can answer "why is this here?"
36
37### Against the others
38
39| | GitHub | Cursor Origin | Entire | g1t |
40| --- | --- | --- | --- | --- |
41| Core idea | Code hosting with Copilot bolted on | A forge for Cursor's cloud agents | Store every agent session with the code | Agents converge an outcome onto `main` |
42| Unit of work | Pull request | Pull request, stacked | Commit plus session | Outcome → plan → issues → pull requests |
43| Agent context | In the Copilot app | In Cursor | In the repo, per commit | Per commit, plus why-blame on any line and what agents told each other |
44| Many agents at once | Compare outputs by hand | Agents can review agents | Not the focus | Plan with dependencies, overlap awareness, coordination tools, queue |
45| Landing | Merge queue (paid) | Stacks | Not the focus | Speculative queue testing combinations; failures return to their agent |
46| Which agents | Copilot, some others | Cursor's | Any (CLI) | Hosted agents plus any MCP client |
47
48Entire's insight, that the session belongs with the code, is one g1t shares
49and already ships (sessions, why-blame). Origin's, that agents should live in
50the forge, too. Neither coordinates a team of agents towards an outcome; that
51is the gap g1t is built for.
52
53## Product model
54
55g1t keeps the two things every engineer already knows, issues and pull
56requests, and changes the assumption underneath them. A forge built for
57people expects a few changes in flight, each watched by its author. g1t
58expects dozens of agents working at once across a project, each on its own
59issue, all of which have to land on `main`. A person assigns an issue to
60the g1t agent and chooses nothing else: not how many agents, and not which
61model. An issue can still collect more than one pull request (a second
62attempt, or someone's own agent alongside g1t's), and when it does the
63issue records which one was taken.
64
65| Concept | What it is |
66| --- | --- |
67| **Issue** | What should change in a repo: a bug, a feature, a question. Opened by a person, an agent or an integration such as an error tracker. Carries labels, acceptance checks (commands that must pass), comments, and every pull request made for it. |
68| **Pull request** | A proposed change in its own Artifacts fork, made by an agent or a person, usually for an issue. Any number can be open for one issue. Starts as a draft; marked ready; merged or closed. |
69| **Session** | The agent's full context for a pull request: prompt, messages, tool calls, cost. Stored with the pull request and linked from every commit it produced. |
70| **Compare view** | Every pull request for an issue side by side with diff, check results, conflicts against main and against each other, and a reviewer agent's summary. |
71| **Merge** | A person or a policy picks a pull request. A per-repo merge queue lands it. The issue closes, recording which pull request resolved it; the others for that issue close as superseded, or are rebased by their agents when the issue is kept open. |
72
73Issues and pull requests share one sequence of numbers per repository, so
74`#12` names exactly one of them.
75
76Features that fall out of the model:
77
78- **Why-blame.** Click a line and see the prompt and reasoning that produced
79 it, not only the commit.
80- **Overlap radar.** Pull requests that touch the same files are flagged
81 while the agents are still working, and the agents are told.
82- **Live lanes.** Watch every pull request progress in real time.
83
84## Why issues and pull requests, not something new
85
86An earlier version of this plan merged the two into one new object, an
87"issue" holding "pull requests". That was wrong, for three reasons.
88
89- **Issues come from everywhere.** People file them, agents file them, and
90 Sentry files them. Most are never worked on by whoever opened them. They
91 need their own life: labels, triage, discussion, closing as not planned.
92- **"Which change did we take?" needs two objects.** When five agents each
93 propose a change, the answer has to be recorded somewhere other than the
94 five proposals. On g1t it is on the issue: `resolved by #14`.
95- **Nobody should have to learn a word to use the product.** An engineer who
96 has used any forge can use g1t on the first day, and finds the agent
97 features where they would look for them.
98
99What g1t adds to the familiar pair:
100
101- **Several pull requests per issue is supported**, not an accident. It
102 is the exception, for a second attempt or a competing one, but when it
103 happens the issue's page lists them with their state, and merging one
104 closes the issue with that pull request recorded and the others marked
105 superseded.
106- **A pull request can be part of the work.** Merging with "keep the issue
107 open" leaves the issue and its other pull requests alone.
108- **Every pull request has a fork and a session.** See
109 [forks and branches](https://docs.g1t.sh/concepts/forks/).
110- **Labels need no setup.** A repository starts with `bug`, `feature`,
111 `docs`, `chore` and `question`; any other name becomes a label the first
112 time it is used, so an integration can tag what it files.
113- **The developer path is unchanged.** Push a branch, open a pull request
114 from it, get review, merge. Agents get a fork per pull request instead.
115- **Both paths meet at `main`.** The same landing rules apply to a person's
116 pull request and an agent's.
117
118## Converging on main
119
120Twelve issues started together will finish at different times and touch
121overlapping code. Getting them all into `main` without a person refereeing
122is the hard part, and it is handled in four places.
123
1241. **Before work starts: plan the overlap away.** A project is a graph of
125 issues. A planner agent can split a large goal into issues, predict
126 which files each will touch, and add a dependency where two would collide,
127 so one starts from the other's result instead of from `main`.
1282. **While agents work: overlap radar.** Each pull request's changed files and
129 symbols are tracked as it pushes. When two pull requests from different issues
130 enter the same area, both agents are told what the other is doing there.
1313. **When `main` moves: the author resolves.** Every open pull request is
132 trial-merged against the new `main`. A clean merge updates the pull request
133 silently. A conflict resumes that pull request's agent with its original
134 session and the incoming change, so the conflict is resolved by the agent
135 that wrote the code and still knows why.
1364. **At landing: a speculative queue.** Approved pull requests enter the
137 repo's queue. g1t builds the combined states (`main`+A, `main`+A+B, …) and runs
138 their checks in parallel. Pull requests land in order as their combined state
139 passes; one that fails is ejected back to its agent and the states behind
140 it are rebuilt. `main` only ever receives a state that passed.
141
142Landing can be fully automatic: a repo policy such as "checks pass and the
143reviewer agent approves" merges without a person.
144
145## Agents aware of each other
146
147Each repo keeps a live **work registry**: for every running pull request, its
148issue, a running summary of what it has done, and the files and symbols it
149has touched or plans to touch. Agents use it through MCP tools; g1t also
150acts on it without being asked.
151
152- **Before starting.** When an issue is opened, or an agent is about to
153 begin a task, g1t searches open issues and running pull requests for the same
154 goal (by meaning, not wording) and for the same area of code. If a match
155 exists the agent is told who is on it and how far along, and chooses: join
156 as a deliberate racer, wait for the result, or drop the task. Duplicate
157 issues are offered for merging.
158- **Finding out-of-scope work.** An agent that discovers something outside
159 its issue asks the registry who works there. If another pull request owns that
160 area, it **hands off**: a note, the relevant excerpt of its session, and
161 optionally commits the receiver can take. If nobody does, it opens a child
162 issue instead of widening its own change.
163- **Asking.** An agent can put a question or a request to another pull request.
164 The receiver gets it at its next turn.
165- **Waiting.** An agent that needs another pull request's result parks itself.
166 Its sandbox sleeps, spend stops, and it resumes from the new state when
167 that pull request merges.
168- **Agents that do not cooperate.** For pushes from tools that never call
169 these tools, g1t compares the pushed change against running pull requests and
170 flags near-duplicates itself.
171
172Every handoff, question and wait has a state (offered, accepted, declined,
173done), appears in the timeline, and is visible to people. A handoff declined
174twice, or two agents passing work back and forth, goes to the "needs you"
175inbox.
176
177## Review at scale
178
179Cloudflare's brief asks "how do you review everything they produce?". With
180hundreds of agents, a person cannot read every diff, so review is by
181exception.
182
183- **Evidence, not diffs.** Every pull request carries a proof bundle: checks run
184 and their output, a preview URL, a plain-language summary, and the
185 behaviour that changed.
186- **Two agent reviewers.** One reviews the change against the issue. A
187 second is adversarial: it tries to break the change and reports what it
188 found.
189- **Risk tiers.** Each change is scored from what it touches, how large it
190 is, and how the reviewers ruled. Low risk merges on policy; high risk goes
191 to a person with the evidence already assembled.
192- **Trust is earned.** An agent's record on a path (merged, reverted, caught
193 by review) raises or lowers the tier its changes land in.
194- **Sampling.** A share of auto-merged changes is sent to a person anyway,
195 to keep the policy honest.
196
197## Rethinking the git primitives
198
199- **No branches for agents.** A pull request is a fork; `main` is the only
200 long-lived line. There is nothing to name, clean up or go stale.
201- **Projected main.** New pull requests start from `main` plus everything already
202 in the landing queue, so they are built on the state they will land on.
203- **Structural merge.** The merge engine merges by syntax tree, not by line,
204 for supported languages. Two agents adding different functions to the same
205 file do not conflict.
206- **Forkable sessions.** A session can be forked at any turn: the code as it
207 was at that moment plus the conversation up to it, continued with a
208 different instruction. Branching applies to the reasoning as well as the
209 code.
210- **Provenance in history.** Every commit records its issue, session,
211 agent, model and cost, and is signed with a key issued to that pull request. The
212 history can be audited by machine.
213
214## People in the loop
215
216### Code that arrives from outside
217
218People will keep pushing with plain git, their editor, or another tool. Every
219push goes through g1t's git front end, so none of it bypasses the model.
220
221- **A push to a branch becomes a pull request.** g1t adopts it with the pusher as
222 author. A reviewer agent writes the issue it appears to serve and offers
223 to attach it to an open issue it matches. From there it gets the same
224 checks, compare view and queue as agent work.
225- **A push to `main` follows repo policy.** Protected: refused with a message
226 saying which ref to push to instead, so it enters the queue. Open: accepted
227 and treated as "`main` moved", which re-verifies the queue and triggers
228 resolve-on-move for every open pull request.
229- **Context is an open format.** A commit trailer names the session that
230 produced it, so any tool can attach its transcript. Commits without one are
231 shown in why-blame as "pushed by a person, no session".
232- **Approval rules.** Per repo and per path: merge automatically, require a
233 named person, or require a person when the change is large or the reviewer
234 agent is unsure.
235
236### Joining work that is already running
237
238- **Every session has a live page** that works on a phone: the transcript as
239 it streams, the current diff, check results.
240- **Steer.** Send a message, pause, or redirect. Hosted agents receive it
241 immediately; a person's own Claude Code receives it at its next turn
242 through the CLI hooks.
243- **Answer.** When an agent is blocked on a question, it appears in a "needs
244 you" inbox and as a notification. The answer resumes the agent.
245- **Take over and hand back.** Check out the pull request's fork, commit by hand,
246 push, and let the agent continue from there.
247
248### Planning by writing
249
250- **Brief.** Write the outcome in prose on the site, or commit it as a
251 markdown file. A planner agent turns it into a project: issues, acceptance
252 checks, dependencies. The person edits the graph before anything starts.
253- **Plan from their own agent.** The same operations are MCP tools, so a
254 person can plan in their own Claude Code session and create the project
255 from there.
256- **The brief stays the source of truth.** Editing it later re-plans: new
257 issues are added, obsolete ones are closed.
258
259### Seeing what moved
260
261- **Project page.** The outcome, the issue graph coloured by state, and how
262 many acceptance checks pass now compared with when the project started.
263- **Digest.** An agent-written summary per project and per person: what
264 merged, what is blocked on whom, which conflicts were resolved, what it
265 cost.
266- **Timeline.** Every event (push, steer, check, conflict, merge) in order,
267 each linked to the session and the person or agent behind it.
268
269## One session, any surface
270
271A session belongs to g1t, not to the device it started on. The browser, a
272phone and Claude Code are views of the same session.
273
274- **Browser and phone.** The site is a responsive, installable web app with
275 push notifications. Everything a person does (brief, steer, answer,
276 approve, merge) works there.
277- **Claude Code.** Through `mcp.g1t.sh` and the CLI hooks, a local session is
278 a g1t session: its transcript syncs as it runs and it appears in mission
279 control like any other.
280- **Moving a session.** A local session can be sent to the cloud: a hosted
281 agent takes over the fork and the transcript and continues, so the laptop
282 can close. A hosted session can be pulled down: the CLI checks out the fork
283 and resumes it in local Claude Code with its history.
284- **Limit.** A session running only on a laptop stops when the laptop does.
285 It can be steered between turns but not continued until it is moved or the
286 laptop is back.
287
288## For people who do not write code
289
290- **Documents are first-class.** Specs, guides, policies and decisions live
291 in repos as markdown, shown in a Docs view: rendered pages, edited in the
292 browser like a document, with inline comments. "Suggest a change" is an
293 pull request and "publish" is merge, without git vocabulary.
294- **Document issues.** "Write the onboarding guide for the billing API" is
295 an issue. Its acceptance checks are a checklist judged by a reviewer agent
296 instead of commands. Agents draft and revise; people comment and approve.
297- **Templates.** Product brief, RFC, decision record. A filled-in template is
298 a brief the planner can turn into a project.
299- **Explain.** Ask about any repo, project or change in plain language and
300 get an answer with links to the code and sessions behind it.
301- **Living documentation.** g1t generates "how this works" pages from the
302 code and keeps them current. When a merged change contradicts a document,
303 an issue opens to update it.
304- **See it, don't read it.** Every pull request on a deployable repo gets a
305 preview URL (Workers Builds from the pull request's fork), so an approver clicks
306 through the result instead of reading a diff. Changes are also summarised
307 in plain language.
308- **Roles.** Viewer, commenter, planner, approver: a person can plan and
309 approve work without ever cloning a repo.
310
311## The macro view
312
313The hierarchy above a single repo:
314
315| Level | What it is |
316| --- | --- |
317| **Workspace** | A company or team: its people, repos, agents, budget and policies. |
318| **Initiative** | A business outcome with an owner and measurable results, e.g. "move billing to usage-based pricing". Spans any number of repos. |
319| **Project** | One deliverable inside an initiative: a brief and its graph of issues. |
320| **Issue / Pull request** | As above. An issue may touch several repos; a pull request for it then holds one fork per repo and they land together. |
321
322How it feeds up, and what is built:
323
324- **A workspace is the unit everything belongs to.** Repositories, people,
325 access tokens, and later projects, budgets and policies are the
326 workspace's, never a person's. An account owns nothing; its first step
327 after confirming its email is creating a workspace, and the site sends it
328 there from wherever it was going.
329- **One namespace.** Usernames and workspaces share one set of names, as on
330 Docker Hub and npm. A username is reserved for its owner's workspace, so
331 `g1t.sh/<name>` never means two things.
332- **The workspace page is the roll-up.** `g1t.sh/<workspace>` shows its
333 repositories with their open issues and pull requests, and the pull
334 requests in progress across all of them. Projects and initiatives will
335 roll up to the same page. Its own pages live under `/<workspace>/-/`
336 (people, access tokens, settings), which no repository can be named.
337- **Workspace access tokens instead of service accounts.** A workspace has
338 tokens of its own, in the same table and code path as personal ones. One
339 acts as the workspace, with a member's rights in that workspace only,
340 records who made it and when it was last used, and keeps working when
341 that person leaves. CI, integrations and automations use these.
342
343### Portfolio
344
345One page answers "where is the business" across every initiative:
346
347- **Health** per initiative: on track, at risk, or blocked, derived from
348 facts (checks passing, issues stalled, questions waiting on a person),
349 not self-reported.
350- **Progress** as measurable results: acceptance checks passing, issues
351 merged out of planned, and the trend since the start.
352- **Forecast** from actual throughput: at the current rate, when the
353 remaining issues land.
354- **Spend** in tokens and dollars against a budget, per initiative.
355- **Waiting on people**: every decision or approval a person owes, by name.
356- **Roadmap**: initiatives laid out as now, next, later, with optional
357 time-boxed cycles for teams that work in sprints.
358
359### Status without asking
360
361- **Standup.** An agent writes a daily report per initiative and one for the
362 whole workspace: what merged, what changed direction, what is at risk and
363 why, what needs a person. Delivered by email or webhook.
364- **Ask.** A question box over the full event log and all sessions: "what
365 happened on the billing migration since Monday?" answers with links to the
366 sessions and commits behind each claim.
367
368### Long-running agents
369
370Work that runs for days needs supervision that does not depend on someone
371watching.
372
373- **Checkpoints.** A long pull request reports milestones against its issue, so
374 progress is visible before anything merges.
375- **Stall and drift detection.** A pull request with no meaningful progress, or
376 whose changes have wandered away from its issue, is flagged and can be
377 stopped or re-briefed automatically.
378- **Budgets.** Hard limits on spend and time per pull request, project and
379 initiative.
380
381### Context hub
382
383Agents working across repos and days need context that outlives any one
384session and reaches beyond the code. The context hub is one place an agent
385asks, whatever the source.
386
387| Source | What it holds | How it gets there |
388| --- | --- | --- |
389| **Memory** | Decisions, conventions, gotchas, facts about systems | Written by agents and people in g1t |
390| **Code and sessions** | The repos, and the reasoning behind every change | Already in g1t |
391| **Connected sources** | Jira and Linear tickets, Notion and Confluence pages, Google Drive documents, Slack threads, Sentry issues | Connectors, authorised per workspace |
392
393How it behaves:
394
395- **One search.** An agent asks a question and gets ranked results across
396 all sources, each labelled with where it came from, who wrote it, and how
397 fresh it is.
398- **Connected sources stay where they are.** g1t indexes them for search and
399 fetches the current version when an agent opens one. The external system
400 remains the source of truth, and a link placed on an issue ("see
401 JIRA-482", a Notion URL) is pulled into the agent's starting context.
402- **Permissions carry over.** A connector only exposes what the connecting
403 account can see, and a workspace admin chooses which spaces, projects or
404 channels are included.
405- **External content is untrusted.** A ticket or page can contain text meant
406 to manipulate an agent. It is marked as reference material, never treated
407 as instructions.
408- **Documentation is separate.** Context is what agents know; documentation
409 is what people read, and it is generated from context and code.
410
411Memory is the part of the hub that g1t owns and agents write to:
412
413- **Memory is written freely.** Any agent or person adds an entry with one
414 call: a decision, a convention, a gotcha, a fact about a system. No review
415 gate. Each entry records who wrote it, from which session, and when.
416- **It is still a repository.** Each workspace has a memory repo in
417 Artifacts, so every write is a commit: versioned, attributable, and
418 revertible.
419- **It is kept healthy by an agent.** A consolidation agent merges
420 duplicates, retires entries that newer ones contradict, and flags
421 conflicts it cannot settle. People can pin an entry (agents may not change
422 it), correct it, or retract it.
423- **Agents read it.** Every session starts with the context relevant to its
424 issue, found by search, and can query more through MCP.
425- **Updates are events.** A memory write or a change in a connected source
426 is an event, so "when context changes, update the affected docs" is an
427 automation, on by default.
428- **It is scoped inside the workspace.** Some context applies to the whole
429 workspace, some to one initiative, project or repo, so an agent gets what
430 applies to its work.
431- **It never crosses workspaces.** A workspace is the isolation boundary: its
432 context, sessions and private repos are invisible to every other
433 workspace, and an agent's token is bound to one workspace.
434
435## Working in g1t
436
437- **Mission control.** The signed-in home page: every running session, every
438 issue waiting on a decision, and what merged, across all repos.
439- **Projects.** Group issues across repos toward one outcome and track how
440 many are open, racing, or merged.
441- **Steering.** Send a message to a running pull request, or to all pull requests on an
442 issue at once, without stopping them.
443- **Automations.** Rules that start work without a person (next section).
444
445## Automations and integrations
446
447> **2026-10-03:** g1t's own `.g1t/automations` format was built and then set
448> aside at the user's request ("let's just copy GitHub Actions on that for the
449> time being"). Automation on g1t is GitHub Actions workflows in
450> `.g1t/workflows/`; the format below is kept for later.
451
452An automation is **when** an event happens, **if** conditions hold, **do**
453something. They are defined as files in the repo (`.g1t/automations/`), the
454way GitHub Actions workflows are, and can also be built in the UI.
455
456### Events that can trigger one
457
458| Source | Examples |
459| --- | --- |
460| Git | push, merge, check failed, `main` moved |
461| g1t | issue opened, pull request stalled, context updated, handoff declined, budget reached |
462| Time | cron schedule |
463| Integrations | Sentry issue, PagerDuty incident, Linear or Jira ticket, Slack message or mention, GitHub issue, Stripe event |
464| Anything else | a signed generic webhook, or an email to a per-repo address |
465
466**Actions**: open an issue (optionally assigning N agents to it), message
467a running pull request, update documentation, notify, call a
468webhook, write back to the source system.
469
470**Example: Sentry.** A new production error arrives. The automation opens an
471issue labelled `bug`, with the stack trace, release and frequency in its
472description. Why-blame
473finds the session that wrote the failing line, so the fixing agent starts
474with the original reasoning. When the fix merges, g1t comments on the Sentry
475issue and resolves it.
476
477### Rules every automation obeys
478
479- **Deduplication.** The same Sentry issue firing 500 times maps to one
480 issue.
481- **Limits.** Concurrency and budget caps per automation.
482- **Loop protection.** Work started by an automation cannot retrigger the
483 same automation without a person in between.
484- **External input is untrusted.** A webhook payload can contain text written
485 by an attacker. Agents started by external events run with reduced
486 permissions and cannot merge without the repo's approval rule passing.
487
488**Checks** are the other half of what GitHub Actions does: build and test
489commands declared in `.g1t/checks.yaml`, run in sandboxes on every pull request
490and on every combined state in the landing queue.
491
492Agents can also reach integrations directly: an agent definition lists MCP
493servers (Sentry, Linear and so on) it may use while working.
494
495## A repository that maintains itself
496
497> **2026-10-04:** the user asked for Dependabot, GitHub Advanced Security and
498> Vercel-style deployments, "so you're not having to maintain shit and you're
499> just pushing up agents that are delivering work consistently".
500
501GitHub reports problems and leaves the fix to you. In g1t, an agent opens an
502issue for each problem, writes the fix, runs its checks, links a preview and
503lands it through the queue. People only decide.
504
505### Upkeep agents
506
507- **Dependency updates.** A scheduled scan reads the lockfiles (npm, Cargo,
508 Go, pip), finds outdated and vulnerable packages and opens one issue per
509 update or group, assigned to g1t-agent. The agent upgrades the package,
510 fixes what the upgrade broke and lands it through the queue. A repository
511 sets how often it scans, which packages it groups and what lands without
512 review (`.g1t/upkeep.yml`, shaped like `dependabot.yml`).
513- **Secret scanning.** Pushes are scanned for known token formats. A push
514 that adds a secret is refused with the file and line; one already in
515 history opens an issue to rotate it and remove it.
516- **Vulnerability alerts.** Dependencies are matched against the OSV
517 database. Every alert links to the issue and pull request fixing it.
518- **Code scanning.** A reviewer agent reads each pull request's diff for
519 security problems and leaves findings as review comments with a
520 suggested fix. Findings on `main` open issues.
521- **A security page per repository** lists alerts, secrets and findings,
522 with the agent work on each, like GitHub's Security tab.
523
524All of these are event sources for the existing issue → agent → checks →
525queue pipeline; they need no new kind of work.
526
527### Deployments
528
529- **A preview for every pull request**, at
530 `<pr>--<repo>--<owner>.g1t.page`, linked on the pull request and updated
531 on each push. `main` deploys to `<repo>--<owner>.g1t.page`, and a
532 repository can add its own domain.
533- **On g1t.page, not g1t.sh,** so customer code never shares cookies or an
534 origin with the site people sign in to.
535- **Built on Workers for Platforms.** Each deployment is a user Worker in a
536 dispatch namespace; one dispatch Worker on `*.g1t.page` routes to it. The
537 build runs in the same runners as Actions. Static sites and Workers apps
538 first; container apps and databases later.
539- **Agents use the preview.** The reviewer agent opens the preview in a
540 browser, takes screenshots of what changed and attaches them to its
541 review, so an approver sees the result without reading the diff.
542- **Environments.** Preview, production and their secrets; deploy history
543 and one-click rollback.
544- **Scale to zero.** An idle branch costs neither g1t nor the customer
545 anything: a Worker runs, and is billed, only while it answers a request.
546 A preview is deleted when its pull request closes or merges, and after
547 a set number of idle days. Container apps, later, sleep when idle.
548- **Billed to the customer, never free.** The user (2026-10-04): "we
549 should not be giving any of this available for free". Every deployment
550 is metered per workspace (requests, CPU time, deployed apps, stored
551 data), priced at Cloudflare's cost + 20% like agent usage, and drawn
552 from prepaid credit. `FREE_WHILE_BUILDING` and the free model allowance
553 do not cover deployments: a workspace without credit cannot deploy, and
554 turning deployments on says so first.
555- **Turning them on is a paid plan, the way Cloudflare's is.** "Including
556 things like them even enabling the feature should have that pay like
557 Cloudflare does." Enabling deployments for a workspace starts a monthly
558 fee that includes an allowance of requests, CPU time and deployed apps;
559 usage past it is billed per unit, as Workers for Platforms bills g1t.
560 The fee and allowance are the user's to set.
561- **Recommended, off in one click.** Because enabling costs money, it is
562 never switched on without the workspace agreeing: new repositories
563 recommend it prominently. A repository can turn it off, keep only production, or
564 deploy somewhere else from its own workflows. Apps built for Cloudflare
565 (Workers, static assets, D1, KV, R2) deploy without configuration.
566
567Later, toward GitLab's DevOps breadth: environment protection rules,
568package and container registries, releases, container hosting.
569
570## Agents and models
571
572### Defining an agent
573
574An agent is a file (`.g1t/agents/<name>.md`, or in the workspace library):
575instructions, the harness and model to run, the tools and MCP servers it may
576use, its sandbox image, permissions and budget. Agents take roles: planner,
577implementer, reviewer, conflict resolver, documenter, memory consolidator.
578Each role has a default that a repo can replace.
579
580### Where it runs, and on whose model
581
582| Option | How it works | Fits |
583| --- | --- | --- |
584| Hosted, g1t's model | g1t runs the sandbox and bills usage | Getting started; no keys to manage |
585| Hosted, your API key | Same sandbox, your Anthropic, OpenAI or Google key | Teams with existing contracts |
586| Hosted, your endpoint | Any OpenAI-compatible URL: Bedrock, Vertex, Azure, a self-hosted model | Private or fine-tuned models |
587| Your runner | A g1t runner daemon on your own machines picks up pull requests | Code or models that may not leave your network |
588| Your own session | Local Claude Code, Cursor or any MCP client joins through `mcp.g1t.sh` | Individuals; subscription plans |
589
590Decisions behind this:
591
592- **g1t does not build its own agent loop.** It runs existing harnesses
593 (Claude Code first, through its headless mode) behind a small runner
594 contract: a container image, an entry command, and session events reported
595 through the CLI. Other harnesses plug in by meeting the contract.
596- **All hosted model traffic goes through Cloudflare AI Gateway.** That gives
597 one place for spend tracking, budgets, rate limits, fallback and logs,
598 whichever provider or endpoint is behind it.
599- **Nobody picks a model.** A person assigns work to `g1t-agent`, as they
600 would assign an issue to Copilot, and g1t routes it. Today the kind of
601 work decides (implementing, reviewing, catching up), from one setting on
602 the runner, and each request is tagged at the gateway with that kind, the
603 repository and the pull request. The session records which model ran.
604 The gateway's own dynamic routes cannot make the choice yet: they work
605 only on its OpenAI-compatible endpoint, and the harness speaks
606 Anthropic's.
607- **Subscriptions stay local.** A Claude subscription cannot be used by a
608 hosted sandbox; it needs an API key. People on subscriptions use their own
609 Claude Code session, which is a full participant.
610- **The workspace pays.** A workspace buys credit by card and each agent
611 run deducts what the model cost plus a margin. The billing service asks
612 nothing of the others: the runner asks it before starting a sandbox and
613 is refused when there is no credit, and the sandbox reports what its run
614 cost with a token only it holds. Where no card processor is configured
615 nothing is charged and agents stay limited to listed accounts.
616- **Keys are secrets.** Stored in Cloudflare Secrets Store, injected into the
617 sandbox for one pull request, never shown again.
618
619### Seeing a pull request through
620
621Assigning an issue is the only thing a person does until there is something
622to merge. A pull request made by a g1t agent goes through checks, a review
623by another agent, revision when either finds something, and catching up
624when `main` moves, without anyone pressing a button. It ends as ready to
625merge, or as "needs you" with the reason: the checks still fail after two
626revisions, a review could not be written, or a conflict could not be
627resolved.
628
629The work service decides the next step from the pull request's state and
630claims it in one statement, so a step is taken once. The runner asks on
631every event that could change the answer (ready, pushed, checks finished,
632review finished, `main` moved), and on a five-minute sweep for anything
633missed, and carries the step out in a sandbox.
634
635A pull request does not have to be up to date with `main` to merge,
636unless the repository's settings require it, as on GitHub. Merging one that
637is behind brings it up to date first (a clean merge needs no model; an
638agent resolves a conflict) and lands it when that push arrives. With the
639requirement on, catching up is a step of its own and the checks run again
640on the result.
641
642Each repository sets its own rules, on one settings page: whether its
643default branch takes pushes at all, how many approvals a merge needs and
644whether an agent's counts, whether failed checks can be overridden, whether
645a second agent reviews, and how often an agent is sent back before a person
646is asked. A g1t agent's pull request follows the same rules as anyone's.
647Pushes to a protected branch are refused in the git front end, with the
648reason shown by git beside the branch.
649
650Merging is a person's decision unless the repository says otherwise. With
651"merge automatically when ready" turned on in its settings, a ready pull
652request lands by itself, attributed to `g1t`. That is the whole path from
653an assigned issue to a commit on `main` with nobody in between. Required
654human approval per path, and risk tiers, are still to come.
655
656### Choosing the right agent automatically
657
658Because several agents can work on the same issue, every issue with more
659than one pull request is an evaluation on real work. g1t records, per repo and per kind of issue, each
660agent's win rate, cost and time. That produces a leaderboard, and a routing
661policy: send each new issue to the agent that wins that kind most often,
662start with the cheapest that is good enough, and escalate to a stronger one
663when checks fail.
664
665## What GitHub ships today, and where g1t differs
666
667GitHub's Agent HQ and Copilot app give each agent session its own git
668worktree and branch, list sessions in a mission-control view grouped by
669project, and let a task be assigned to several agents so their output can be
670compared. Underneath, the unit of work is still a branch and a pull request.
671
672| | GitHub | g1t |
673| --- | --- | --- |
674| Where a session works | A worktree on one developer's machine, or a cloud sandbox | A server-side fork that any agent on any machine can join and anyone can open |
675| Agent context | Lives in the app's session view | Stored with the repository and linked from each commit (why-blame) |
676| Several agents on one task | Separate pull requests to compare by hand | One issue holding every pull request made for it, compared side by side, with the merged one recorded on the issue |
677| Collisions between agents | Found as merge conflicts at the end | Flagged during the work (overlap radar) |
678| Landing changes | One pull request at a time | A merge queue that lands the chosen one and closes the rest as superseded |
679| Which agents | Those offered through a Copilot subscription | Any MCP client, plus hosted agents |
680
681## How agents connect
682
6831. **Bring your own agent.** A remote MCP server at `mcp.g1t.sh` lets Claude
684 Code (or any MCP client) list issues, claim one, get a clone URL and
685 token, report progress and submit. Adding it is one command; sign-in is a
686 browser OAuth flow with no token to paste. The `g1t` CLI installs Claude
687 Code hooks that upload the session transcript as the agent works.
6882. **g1t agents.** Assign an issue to g1t's own agent, or many issues at
689 once, each to an agent of its own. g1t starts a sandbox for each
690 (Cloudflare Containers), running a coding agent headless against its
691 own pull request and fork.
6923. **API and CLI.** Everything above is available at `api.g1t.sh` and
693 through `g1t`.
694
695## Public surfaces
696
697| Host | What it serves |
698| --- | --- |
699| `g1t.sh` | The site, git over HTTPS, git over SSH |
700| `api.g1t.sh` | REST API, with no version in its paths, with a published OpenAPI document, cursor pagination, rate-limit headers, idempotency keys on writes, server-sent events for live pull request state, and signed webhooks |
701| `mcp.g1t.sh` | Remote MCP server over streamable HTTP |
702
703g1t is its own OAuth 2.1 authorization server: authorization code with PKCE,
704dynamic client registration that stores nothing (a client id encodes its
705own registration, so the open endpoint cannot be used to fill a database),
706discovery metadata and rotating refresh tokens. Still to come: scopes
707per resource (`repo:read`, `repo:write`, `issue:write`, `pull:write`).
708MCP clients, the CLI (device flow) and third-party apps all use it. Access
709tokens and SSH keys remain for git itself.
710
711## Architecture
712
713| Component | Language | Runs on | Responsibility |
714| --- | --- | --- | --- |
715| `crates/contracts`, `packages/contracts` | Rust, TypeScript | — | The interface of every service, the event catalogue, shared types. Services and clients depend on this, never on each other's code. |
716| `services/identity` | Rust | Worker + D1 | Accounts, workspaces and memberships, sessions, SSH keys, access tokens, device sign-in, OAuth codes and grants |
717| `services/repos` | Rust | Worker + D1 + Artifacts | Repository registry, contents, forks, diffs, landing, git over HTTPS. Storage sits behind a `GitStore` port with an Artifacts adapter. |
718| `services/work` | Rust | Worker + D1 | Issues, pull requests, comments, sessions; later a Durable Object per repo for the landing queue and live state |
719| `services/billing` | Rust | Worker + D1 + Stripe | Each workspace's agent credit: payments, the ledger of every run, and the gate on starting one |
720| `services/events` | Rust | Worker + Queues + D1 | The event bus: durable log, and one queue per subscribing service |
721| `services/runner`, `crates/runner` | TypeScript, Rust | Worker + Containers | Starts a sandbox per g1t agent; the program inside runs the agent harness and reports through the public API |
722| `apps/web` | TypeScript | Worker | Server-rendered site. Holds no data; calls services over RPC. |
723| `apps/docs` | TypeScript | Worker (static) | Documentation and the API explorer |
724| `apps/api` | Rust | Worker | REST API (`api.g1t.sh`), MCP server (`mcp.g1t.sh`) and OpenAPI document, all generated from one list of operations; the OAuth endpoints |
725| `crates/sshd` | Rust | Container | Git over SSH, bridged to Artifacts |
726| `crates/merged` | Rust | Container | Trial merges, conflict matrix, landing merges (needs real git; the Artifacts binding is read-only) |
727| `crates/core` | Rust | native and WASM | pkt-line, packfile and diff code shared by the above and by the Worker |
728| `crates/g1t` | Rust | user's machine | CLI: auth, SSH proxy, Claude Code hooks, issues and pull requests |
729
730Storage: Artifacts for repositories (one fork per pull request), D1 for accounts
731and metadata, R2 for session transcripts and logs, Durable Object SQLite for
732per-repo coordination state.
733
734How the services fit together:
735
736- **Each service is its own Worker with its own database.** It deploys,
737 scales and fails on its own. Callers reach it through a typed RPC binding
738 to the interface in `packages/contracts`.
739- **Expected failures are values.** Every call returns a `Result`, so "not
740 found" or "forbidden" crosses a service boundary as data.
741- **Side effects travel as events.** A service publishes what happened
742 (`git.push`, `issue.opened`, `pull.merged`, …) to the bus and does not
743 call other services to react. Each subscriber consumes from its own queue.
744 Timelines, webhooks and automations read the same stream, which is what
745 lets something like GitHub Actions be built on top.
746- **Every read takes the viewer.** Authorization is decided inside the
747 service that owns the data, not by its callers.
748
749The Workers runtime scales request handling on its own, so the edge layer
750stays in TypeScript. Rust is used where there is real computation or a real
751protocol to implement.
752
753## Languages
754
755The site is TypeScript. Everything behind it is Rust, compiled to
756WebAssembly for Workers and natively for containers and the CLI. Identity, repos, work, events and the API are all Rust. The one exception
757is the Worker that starts sandboxes, because Cloudflare's Containers
758library is TypeScript. Rust services speak a
759small JSON protocol over service bindings (`POST /rpc/<method>`), with the
760types in `crates/contracts`.
761
762## Identifiers
763
764Every id is a [TypeID](https://github.com/jetify-com/typeid): a prefix naming
765the kind of thing, then a UUIDv7 in lowercase base32, such as
766`pr_01jb2k7x9hfq0b3zj0f5s2m8ra`.
767
768- The prefix makes an id self-describing and stops ids of different kinds
769 being mixed up.
770- Ids sort by creation time as plain strings. In SQLite (D1 and Durable
771 Objects) that keeps inserts at the end of the primary-key index instead of
772 scattering them, and gives time-ordered paging for free.
773- The suffix decodes to a standard UUIDv7 for any system that wants one.
774- Ids are made by the service that creates the record, not by the database,
775 so they work across services and can be assigned before a write.
776
777## Events at scale, and audit
778
779The current event log is a single D1 database. That is fine for a
780prototype and wrong for the target: D1 is one writer and 10 GB. The design
781for volume splits storage by how the data is read.
782
783| Tier | Store | Holds | Read by |
784| --- | --- | --- | --- |
785| Hot | A Durable Object per repository, with SQLite | Recent events for that repo | Timelines, live pages over WebSocket |
786| Complete | Cloudflare Pipelines into R2 as Apache Iceberg | Every event, forever, partitioned by day and workspace | Analytics, standups, "ask", export |
787| Audit | The same R2 store, under object lock | Who did what, from where, with which credential | Compliance, investigation |
788
789- **No single hot database.** Each repository's recent events live with that
790 repository, so load spreads across as many objects as there are repos.
791- **The complete record is files, not rows.** Iceberg on R2 has no practical
792 size limit and is queried with SQL.
793- **Audit is a property of every event.** The envelope carries the actor
794 (person, agent, token or system), the credential used, the request id and
795 the source address. Audit entries for a workspace are hash-chained, so a
796 removed or altered entry is detectable, and are written under a retention
797 lock.
798- **Delivery is at least once.** Consumers are idempotent on the event id.
799
800## Accounts and forge basics
801
802- Registration with email verification, sign-in, forgot password (Cloudflare
803 Email Sending), Turnstile on public forms.
804- GitHub sign-in, SSH keys, access tokens, active sessions.
805- Profiles, public and private repositories, repository search (D1 full-text).
806- Rendered README, syntax highlighting, commit history, diffs.
807
808## Built on Cloudflare
809
810| Need | Product |
811| --- | --- |
812| Repositories; a fork per pull request; data residency per workspace | Artifacts (forks, jurisdictions) |
813| Reacting to pushes | Artifacts event subscriptions on Queues |
814| Preview URL per pull request; deploy on merge | Workers for Platforms on `g1t.page` |
815| Site, API, MCP, git front end | Workers |
816| Per-repo coordination, live updates | Durable Objects |
817| Pull request lifecycles, automations | Workflows, Cron Triggers |
818| Agent sandboxes, SSH server, merge engine | Sandbox SDK and Containers |
819| Fast starts on large repos | ArtifactFS |
820| Model traffic, spend, budgets | AI Gateway |
821| Summaries, embeddings | Workers AI |
822| Context hub search | Vectorize |
823| Accounts and metadata | D1 |
824| Transcripts and logs | R2 |
825| Email, bot protection, keys | Email Sending, Turnstile, Secrets Store |
826
827## The submission
828
829- **g1t is built on g1t.** This repository is hosted on g1t.sh, its features
830 are opened as issues and built by racing agents, and it deploys from
831 Artifacts through Workers Builds. The history is the proof.
832- **The demo follows one story.** A brief becomes a project; twelve issues
833 fan out to dozens of agents; agents notice each other, hand off, and
834 resolve a conflict; reviewers triage; the queue lands everything on
835 `main`; why-blame explains a line; the portfolio shows where it all
836 stands. Then the same thing at a thousand agents.
837- **Judges can try it in a minute.** Open registration on g1t.sh, one-click
838 import of a GitHub repo, one command to connect Claude Code, a seeded demo
839 workspace, and a single deploy command for running their own copy.
840- **The formats are open.** The commit trailers, session format and runner
841 contract are published so other tools can interoperate.
842
843## Build order
844
845What is left is ordered by how much it shows the point above, not by forge
846parity. Forge basics are done well enough; each item below should make the
847demo's story stronger.
848
849Done from this list: the outcome page (a plan's issues as a live graph with
850cost and a feed of what happened), coordination you can see (agents' issues
851and comments stand out in the feed; agents hold g1t's tools through a token
852scoped to one repository), steering a running agent (messages delivered
853between steps, and at the end), and recording sessions from anyone's own
854Claude Code (`curl -fsSL https://g1t.sh/install/claude.sh | sh`). Integrations are
855in: a workspace's own model provider (Anthropic or any Anthropic-compatible
856endpoint, reached through a model proxy so no sandbox holds a key, for a
857flat orchestration fee), alerts from Sentry, Datadog and signed webhooks
858that open one issue per problem and can start an agent, and Jira and Linear
859tickets that agents read, people import, and that hear back. Racing a
860set number of agents on one issue is dropped: choosing how many agents to
861use is not something people should have to do.
862
8631. ~~**Handoffs and questions between agents** as states on the outcome
864 page.~~ Done: questions and handoffs show as waiting, read, answered,
865 taken on or declined; since 2026-10-03 an agent asked while it is not at
866 work is woken to answer, where before the question waited forever.
8672. **Upkeep agents** (above): dependency updates, secret scanning,
868 vulnerability alerts, code scanning, the security page. No new
869 infrastructure.
8703. **Deployments on `g1t.page`** (above): previews per pull request,
871 production on merge, the reviewer agent checking the preview.
8724. **The large run, building 2 and 3.** About 20–30 issues on g1t itself,
873 built by agents and landed through the queue, for the video: g1t built
874 on g1t.
8755. **Polish for judges trying it in a minute:** a seeded demo workspace, the
876 empty states, and the first-run path from sign-up to an outcome landing.
877
878Earlier items still open, after those:
879
8801. Deleting a branch once its pull request merges; approval rules per
881 path; risk tiers.
8822. Scopes on OAuth grants and access tokens.
8833. Event storage per the design above: per-repo hot log, Iceberg on R2,
884 hash-chained audit.
8854. CLI with Claude Code hooks to record sessions automatically.
8865. Reviewing and catching up automatically, by policy; required reviews;
887 risk tiers.
8886. Compare view, proof bundles; handoff between agents.
8897. Projects, mission control, steering; why-blame, digest, timeline.
8908. Context hub, portfolio; automations and integrations (Sentry first).
8919. SSH; bot protection; own keys, endpoints and runners.
89210. Large run (100+ agents across many issues), hardening, demo.
893
894Later: code search, mirroring to GitHub, passkeys, SSH
895on port 22 without the CLI proxy (needs the Workers inbound TCP private
896beta).