When to use this profile
Give a coding agent a task, a source checkout and a bounded set of tools. Run baseline checks before edits so existing failures are distinguishable from regressions. After each meaningful change, return the relevant test output to the model. Export the final diff for review; publication is a separate action with separate authorization.
Prepare a repeatable environment
- Record the base commit and remove unrelated working-tree changes from the input archive.
- Set a command timeout, output budget and maximum number of agent iterations.
- Install the project's pinned dependencies and record baseline failures.
git -C /work/project diff --statReturn results the agent can use
- A unified diff and a list of files changed.
- Validation results for the relevant tests and build.
- A concise explanation linking each change to the task.
Use a seeded bug with an existing regression test. Confirm the agent fixes it, preserves unrelated files and returns the patch without pushing it.
Resources and boundaries
Start with 4 vCPU and 8 GB of RAM, then measure peak memory and task duration on a representative fixture. These are workload planning values, not a benchmark or a provisioned configuration. Use the sandbox cost calculator to estimate running time and retained snapshots.
- An agent reading repository instructions can be steered into running unsafe commands.
- Repository access should be read-only unless the workflow explicitly authorizes publication.
- An unchanged test suite is evidence about tested behavior, not proof that every change is correct.
Keep model inference separate from this execution profile. A hosted model or gpuOS can decide the next action while the CPU environment runs it. The quickstart describes the account workflow and the proposed runtime contract.