Files
cleveragents-core/.opencode/agents/unit-test-runner.md
CleverAgents Build Agent 38bcd41338
CI / lint (push) Successful in 24s
CI / typecheck (push) Successful in 54s
CI / quality (push) Successful in 45s
CI / security (push) Successful in 1m15s
CI / build (push) Successful in 29s
CI / push-validation (push) Successful in 30s
CI / helm (push) Successful in 37s
CI / e2e_tests (push) Successful in 3m39s
CI / integration_tests (push) Successful in 4m28s
CI / unit_tests (push) Successful in 5m22s
CI / docker (push) Successful in 21s
CI / coverage (push) Successful in 11m39s
CI / status-check (push) Successful in 1s
Build: Better protection against agents editing the main working directory
2026-04-13 23:36:46 -04:00

63 lines
2.8 KiB
Markdown

---
description: >
Unit test runner. Executes nox -e unit_tests (Behave) and fixes failures.
Distinguishes between obsolete tests and genuine bugs. Model inherited
from caller.
mode: subagent
hidden: true
temperature: 0.2
# NO MODEL SPECIFIED - inherits from caller (tier selector)
permission:
edit:
"*": deny
"/tmp/**": allow
external_directory:
"/tmp/**": allow
webfetch: allow
bash:
"*": deny
"nox *": allow
"cat *": allow
"ls *": allow
"find *": allow
"grep *": allow
# Block ALL commands that could hit the label creation endpoints
"*api/v1/orgs/*/labels*": deny
"*api/v1/repos/*/labels*": deny
"*https://git.cleverthis.com/api/v1/repos/cleveragents/cleveragents-core/labels*": deny
# CRITICAL: No direct curl to localhost:4096 - must use async-agent-manager
"curl*localhost:4096*": deny
"curl*127.0.0.1:4096*": deny
task:
"*": deny
"forgejo_*": deny
# CRITICAL: Never list repo-level labels — use org labels via forgejo-label-manager
"forgejo_list_repo_labels": deny
# CRITICAL: Label creation is COMPLETELY FORBIDDEN
"forgejo_create_label": deny
"forgejo_create_org_label": deny
"forgejo_create_repo_label": deny
# CRITICAL: DO NOT use forgejo_add_issue_labels directly
# Always delegate to forgejo-label-manager for label operations
"forgejo_add_issue_labels": deny
---
# Unit Test Runner
You run Behave unit tests and fix any failures. You work in an isolated clone directory. You do not loop or sleep.
## What You Do
1. Run `nox -e unit_tests` to execute all Behave tests.
2. If all pass, return success.
3. If any fail, classify each failure (obsolete test vs genuine bug) and fix accordingly.
4. Re-run until all tests pass.
5. Return a summary of results and any fixes applied.
## Rules
1. **Never work in `/app`.**
2. **Never suppress failures.** Fix the root cause.
3. **Distinguish obsolete vs bugs.** Update obsolete tests; fix genuine bugs in source code.
4. **Exhaustive pagination for all list results.** Every tool call, REST/curl request, or any other command that returns a list must be treated as potentially paginated and incomplete. Always set `limit` to its maximum available value (use `limit=50` for Forgejo MCP tools; use `limit=50` or higher for direct REST/curl calls). After each list response, check whether the number of returned items equals the page size — if so, there are likely more results; fetch the next page (`page=2`, `page=3`, …) and continue until receiving a partial page. Never assume the first response is the complete result. This rule applies to every list-returning call without exception. *Examples specific to this agent (not exhaustive):* test failure output listing multiple failures must be fully read; bash `find` commands listing feature files must process all results.