Add new agents and skills for enhanced project orchestration and review processes
- Introduced `critic`, an independent adversarial reviewer for security and correctness. - Added `fable-orchestrator` to manage task routing and verification. - Implemented `gauntlet-critic` for fresh-context evaluation of gauntlet rounds. - Created `planner` for generating executable implementation plans with dependencies. - Developed `security-auditor` for application security reviews and audits. - Established `system-steward` to improve agent prompts and skills based on verified failures. - Added `dev-loop` skill for autonomous development loops over repositories. - Implemented `gauntlet-loop` skill for iterative quality benchmarking against reference standards. - Updated project settings to utilize the new orchestrator agent. - Created documentation for `GAUNTLET.md`, `PROGRESS.md`, and `REFERENCE_BAR.md` to track project status and quality benchmarks. - Added detailed prompting style guide to enhance understanding of prompt patterns and agentic loops.
This commit is contained in:
23
_to_delete/replaced-20260806-fableflip/docs/GAUNTLET.md
Normal file
23
_to_delete/replaced-20260806-fableflip/docs/GAUNTLET.md
Normal file
@@ -0,0 +1,23 @@
|
||||
# Gauntlet board
|
||||
|
||||
> Loop state for reference-benchmarked work. One row per part; one line per round. Move finished gauntlets to `docs/archive/`. Statuses: `not started` · `looping` · `parity — stopped` · `diminishing returns — stopped` · `budget exhausted` · `parked (decision-ready)` · `integrated`.
|
||||
>
|
||||
> Parts mirror `REFERENCE_BAR.md` (2026-08-06). **The bar is not yet concrete and budgets are unset** — both are parked decision-ready in `PROGRESS.md`; no gauntlet starts until they're resolved. Exception: "Replace reliability" can start once the owner approves the site matrix (its bar is behavioral).
|
||||
|
||||
## Parts
|
||||
|
||||
| Part | Bar (REFERENCE_BAR.md row) | Rounds | Last verdict | Biggest open gap | Budget left | Status |
|
||||
| --- | --- | --- | --- | --- | --- | --- |
|
||||
| Floating toolbar + result modal | Floating toolbar + result modal | 0 | — | — | [set] | not started — awaiting bar artifacts |
|
||||
| Options page | Options page | 0 | — | — | [set] | not started — awaiting bar artifacts |
|
||||
| Popup + Prompt Builder | Popup + Prompt Builder | 0 | — | — | [set] | not started — awaiting bar artifacts |
|
||||
| Replace reliability | Replace reliability (behavioral) | 0 | — | — | [set] | not started — awaiting site-matrix approval |
|
||||
| Store listing | Store listing | 0 | — | — | [set] | not started — blocked by T-01 |
|
||||
|
||||
## Round history
|
||||
|
||||
- _None yet._
|
||||
|
||||
## Final verdicts
|
||||
|
||||
- _None yet._
|
||||
29
_to_delete/replaced-20260806-fableflip/docs/PROGRESS.md
Normal file
29
_to_delete/replaced-20260806-fableflip/docs/PROGRESS.md
Normal file
@@ -0,0 +1,29 @@
|
||||
# Progress board
|
||||
|
||||
> For the owner. What works, how to see it, and what's waiting on you — plain language, no agent jargon. Refreshed at every phase seal and session end. `HANDOFF.md` speaks to the next agent; this page speaks to you.
|
||||
|
||||
**Updated:** 2026-08-06 · **Overall:** working MV3 extension (Phase 1 + the 2026-07 fix wave); operating system upgraded to the gauntlet-loop/opus kit today.
|
||||
|
||||
## What works now
|
||||
|
||||
- The extension itself: selection → floating toolbar → fix/rephrase/shorten/expand/explain/prompt → Replace or Copy; four providers (OpenAI/Anthropic/Groq/OpenRouter); encrypted BYO key; Options with live model listing; Prompt Builder; 58/58 unit tests, typecheck and build green (2026-07-23).
|
||||
- The agent operating system: upgraded from the older fable kit — 13 specialists (incl. your custom `lexai-extension-dev`, kept and modernized) + 4 new ones (ux-ui-designer, ux-psychologist, and the fresh-eyes `gauntlet-critic` referee), 12 skills, all your lessons and security-auditor memory preserved. Lead is now `claude --agent opus-orchestrator`.
|
||||
|
||||
## See it yourself
|
||||
|
||||
- `npm run build` → `chrome://extensions` → Load unpacked → `.output/chrome-mv3` → select text on any page.
|
||||
- Open `CLAUDE.md` — your repo rules and 9 codebase invariants are carried over intact; the gauntlet protocol is new in §3.
|
||||
|
||||
## Waiting on you — each item blocks ONLY its own lane
|
||||
|
||||
| # | Decision | Options (recommended bold) | What it unblocks |
|
||||
| --- | --- | --- | --- |
|
||||
| 1 | Groq-key re-entry check (from 2026-07-23 handoff): reload unpacked, re-enter Groq key in Options, ↻ Load → model → Save, confirm a real-page action | **do the 5-min check** / report it already done | closes the key-mismatch fix loop |
|
||||
| 2 | Supply reference-bar artifacts (screenshots/recording of Grammarly or your chosen benchmark → `docs/reference/`) | **Grammarly toolbar + card screenshots** / pick another benchmark / defer gauntlets | UI gauntlet rounds |
|
||||
| 3 | Approve the Replace-reliability site matrix in `docs/REFERENCE_BAR.md` (Gmail, GitHub, X, LinkedIn, Google Docs?, Reddit, Notion) | **approve as listed (Docs out of scope)** / edit the list | the behavioral gauntlet — can start without screenshots |
|
||||
| 4 | Set gauntlet budgets on `docs/GAUNTLET.md` | **modest budget on one part first** / several at once | looping |
|
||||
| 5 | Delete `_to_delete\` in the repo (replaced kit files + transfer archive parked there) | delete now / leave for later | nothing — housekeeping |
|
||||
|
||||
## Next up — proceeds without you
|
||||
|
||||
- T-01 (`<all_urls>` narrowing) and T-02 (real key encryption) remain the ranked pre-release risks from `HANDOFF.md` — routable to security-auditor + lexai-extension-dev any time.
|
||||
25
_to_delete/replaced-20260806-fableflip/docs/REFERENCE_BAR.md
Normal file
25
_to_delete/replaced-20260806-fableflip/docs/REFERENCE_BAR.md
Normal file
@@ -0,0 +1,25 @@
|
||||
# Reference bar
|
||||
|
||||
> The concrete quality bar for gauntlet work. Every entry must point at something a referee can open, run, or look at — an adjective is not a bar. Changing a bar mid-gauntlet is an owner decision recorded in `DECISIONS.md`.
|
||||
>
|
||||
> **Status: NOT YET CONCRETE — decision-ready.** LexAI's repo contains no reference artifacts, so the rows below are *proposals*: the parts are real, but each needs owner-supplied artifacts (screenshots/recordings into `docs/reference/`, or a named competitor install to compare live) before a gauntlet can start. Behavioral rows can start sooner — their bar is a checkable matrix, not an artifact.
|
||||
|
||||
## Bars by part (proposed)
|
||||
|
||||
| Part | Reference artifact(s) — TO SUPPLY | How to compare | Minimum parity |
|
||||
| --- | --- | --- | --- |
|
||||
| Floating toolbar + result modal (in-page UI) | screenshots/screen-recording of Grammarly's selection toolbar + suggestion card (or another benchmark extension the owner picks) → `docs/reference/` | load unpacked, select text on a real page at the same spots, screenshot side-by-side | placement, legibility, non-intrusiveness, and interaction states read as polished as the reference |
|
||||
| Options page | reference settings page screenshots (Grammarly / a best-in-class extension options UI) | side-by-side render | clarity of provider→key→model flow; error/rejected-key states as discoverable as the reference |
|
||||
| Popup + Prompt Builder | reference popup/composer screenshots | side-by-side render + walk the compose flow | task flow completable as directly as the reference |
|
||||
| Replace reliability (behavioral) | site matrix the owner approves (e.g. Gmail compose, GitHub textarea/PR comment, X/Twitter composer, LinkedIn, Google Docs*, Reddit, Notion) | run fix→Replace on each; record works / partial / fails | Replace works on every approved matrix site; no self-triggering; no host-page breakage (*Google Docs may be declared out of scope — record it) |
|
||||
| Store listing | top-ranked writing-assistant CWS listings (live pages) | side-by-side read of `store-assets/` vs the live listings | screenshots, copy, and permission justification at parity before any CWS push (blocked by T-01 `<all_urls>` anyway) |
|
||||
|
||||
## Reference sources
|
||||
|
||||
- `docs/reference/` — **empty until the owner supplies artifacts** (screenshots, recordings)
|
||||
- A named competitor extension installed locally for live blind A/B, if preferred over screenshots
|
||||
|
||||
## Out of scope for the bar
|
||||
|
||||
- Grammarly's backend features (tone rewriting service, plagiarism, team features) — LexAI is BYO-key by design; the bar is UI/UX and reliability parity, not feature parity.
|
||||
- Anything postponed in `docs/TASKS.md` or blocked by open security items (T-01 `<all_urls>`, T-02 key encryption) — those gate release, not gauntlet rounds.
|
||||
Reference in New Issue
Block a user