Danilo Alessandro Vinciprova

Danilo Alessandro Vinciprova

AI-Assisted Web Developer · React · Next.js · Claude Code

I learned to think of building as parts that must fit together.

AI Systems · multi-agent workflow

FWEA — AI agent workflow

Builder → Reviewer → Fixer → Verifier, as I designed it. The review cycles shown here are read from its own log.

Release candidateBuilt and reviewed; final A/B validation not yet completed

Claude Code · multi-agent design, review and verification

Multi-agent system designed and orchestrated by me.

Problem

An AI agent that grades its own work is not a reliable check: the same view that made a mistake is the one looking for it. FWEA separates building, reviewing, fixing and verifying into different agents, and records everything with hashes.

One agent alone

  1. builds the work
  2. grades its own work
  3. approves it

Nobody else looked. A blind spot in the build is also a blind spot in the check.

FWEA: separate roles

  1. a builder builds
  2. two isolated critics review
  3. a fixer repairs; a checker re-reads
  4. two fresh verifiers confirm
  5. the result is hashed

No role grades its own output, and every round leaves a record.

How I work

“I use AI agents as part of a structured development workflow: build → review → fix → verify, while keeping responsibility for requirements, decisions and final quality.”

ILF

Designed ILF, a personal framework of structured AI “labs” for web design, film, food and product work

Stated by Danilo

Architecture

The roles of the real FWEA run on 28 September 2026, as its review log records them. Each critic worked in isolation, and the tree that passed is pinned by its SHA-256 digest.

Builder → critics → fixer → verifiers

Evidence tier: Existing evidence · roles from the review log

isolated · read-onlywhile a round fails: fix, then fresh criticsBuilderbuilds the skill treeCritic · Lens 1correctness &decidabilityCritic · Lens 2scope & economyFixer+ internal checkerVerifier 1checks every findingVerifier 2checks every findingPASScheckpoint hashedsha256 8551e9a4…isolated · read-onlynext roundBuilderbuilds the skill treeCritic · Lens 1correctness &decidabilityCritic · Lens 2scope & economyFixer+ internal checkerVerifier 1checks every findingVerifier 2checks every findingPASScheckpoint hashedsha256 8551e9a4…
FWEA roles: a builder; two isolated critics (lens 1: correctness & decidability; lens 2: scope & economy); a fixer with an internal checker, looping back to fresh critics while a round fails; two verifiers; PASS, with the skill-tree digest sha256 8551e9a4a99b… recorded in the checkpoint.

The sequence, as text

  1. Builder builds the skill tree.
  2. Two critics, isolated and read-only, review it: lens 1 for correctness and decidability, lens 2 for scope and economy. Each reports BLOCKER, MAJOR and MINOR findings.
  3. Fixer corrects the findings; an internal checker re-reads the changes, and a fix-up closes what it finds.
  4. While a round fails, fresh critics review the corrected tree: rounds 0, 1 and 2, then the final critics.
  5. Two verifiers check that every serious finding of the final critics is closed, and look for regressions.
  6. PASS: the tree is pinned by its digest, sha256 8551e9a4a99b7b6b8e67c1530022adc2bfa15274545f57b50d1c842ab491593b.

The fix rounds (fixer, internal checker, fix-up) are recorded in the same log; their lines are not shipped here. The rounds, the verifiers and the digest are in the excerpts under Evidence.

Review cycles

Each critic reported its findings as BLOCKER, MAJOR or MINOR. 4 critic rounds ended FAIL; after the last fix round, two verifiers checked every serious finding and step 2 passed. The numbers are read from the log lines themselves when this site is built.

Findings per round and reviewer

Evidence tier: Existing evidence · parsed from the log

  • BLOCKER
  • MAJOR
  • MINOR
  • solid, hatched, hollow · numbers: BLOCKER · MAJOR · MINOR
Round 0FAILLens 1 correctness & decidability3 · 12 · 11Lens 2 scope & economy2 · 10 · 5Round 1FAILLens 1 correctness & decidability2 · 15 · 7Lens 2 scope & economy0 · 4 · 8Round 2FAILLens 1 correctness & decidability1 · 11 · 10Lens 2 scope & economy0 · 3 · 9Final critics · stricter isolation18 unique serious findings, all fixedLens 1 correctness & decidability2 · 13 · 5Lens 2 scope & economy0 · 4 · 7Targeted verification · after the last fix round✓ PASSV1 regression18/18 closed · 0 · 0 · 5V2 regression19/19 closed · 0 · 0 · 4Round 0FAILLens 13 · 12 · 11Lens 22 · 10 · 5Round 1FAILLens 12 · 15 · 7Lens 20 · 4 · 8Round 2FAILLens 11 · 11 · 10Lens 20 · 3 · 9Final critics18 unique seriousLens 12 · 13 · 5Lens 20 · 4 · 7Targeted verification✓ PASSV118/18 closed · 0 · 0 · 5V219/19 closed · 0 · 0 · 4
Findings per round and reviewer, from the review log
RoundReviewerBLOCKERMAJORMINORResultLog
Round 0Critic · Lens 131211FAILline 17 of the review log, Round 0
Critic · Lens 22105FAILRound: FAILline 18 of the review log, Round 0
Round 1Critic · Lens 12157FAILline 28 of the review log, Round 1
Critic · Lens 2048FAILRound: FAILline 29 of the review log, Round 1
Round 2Critic · Lens 111110FAILline 39 of the review log, Round 2
Critic · Lens 2039FAILRound: FAILline 40 of the review log, Round 2
Final criticsstricter isolationCritic · Lens 12135FAILline 56 of the review log, Final critics
Critic · Lens 2047FAILRound: 18 unique serious findings, all fixedline 57 of the review log, Final critics
Targeted verificationafter the last fix roundVerifier 1regression findings005CLEAN · 18/18 closedline 67 of the review log, Targeted verification
Verifier 2regression findings004CLEAN · 19/19 closedRound: step 2 PASSline 68 of the review log, Targeted verification

Real results from the review log, 28 September 2026. Stacked bars of BLOCKER, MAJOR and MINOR findings for each reviewer, in 5 rounds; for the verifiers, the findings closed and the new (regression) findings. The same numbers are in the table; each “line” link opens that round’s verbatim excerpt.

Demo

A sample claim goes through the same four roles. It is a simulation: the sentence is invented and none of these findings come from the real log (the real numbers are in Review cycles).

A sample claim, step by step

Evidence tier: Simulation · illustrative

  1. 01 Builder: Writes a first draft

    The builder turns the brief into a first version of the sentence.

    Sample claim (draft)

    Expert in all modern web technologies, with flawless results.

    Draft

  2. 02 Critics: Two isolated critics review it

    Each critic reads only the draft and the brief, with read-only tools, and reports findings by severity.

    Sample claim (draft)

    Expert in all modern web technologies, with flawless results.

    • BLOCKER · Lens 1 · correctness & decidability: “all modern web technologies” cannot be decided: no evidence could ever prove it.
    • MINOR · Lens 2 · scope & economy: “flawless results” is a vague phrase that adds nothing checkable.

    FAIL · 1 BLOCKER, 0 MAJOR, 1 MINOR

  3. 03 Fixer: Fixes what the critics found

    The fixer changes only what the findings name; an internal checker re-reads the changed lines.

    Sample claim (after the fix)
    Expert in all modern web technologies, with flawless results.Builds responsive pages in HTML, CSS and JavaScript; each skill links to its evidence.
    • F1 (BLOCKER): the claim now names the skills, each with its evidence.
    • F2 (MINOR): the vague phrase is removed.
    • Checker: no new finding in the changed lines.

    Fixed · waiting for verification

  4. 04 Verifier: A fresh verifier checks every finding

    The verifier did not take part in the review or the fix. It checks each finding against the new text, then looks for regressions.

    Sample claim (after the fix)

    Builds responsive pages in HTML, CSS and JavaScript; each skill links to its evidence.

    • F1: CLOSED
    • F2: CLOSED
    • Regression: 0 BLOCKER, 0 MAJOR, 0 MINOR

    PASS · sealed with its SHA-256

    sha256 of the fixed sentence: f0841dcaff317314bdd6e489d685ffa95469bac55e73e95c7411d50892da5a60

Evidence

The status of the skill, the checkpoint digest and manifest, and the verbatim lines of the review log that this page is built from.

Status

Evidence tier: Existing evidence

Release candidateBuilt and reviewed; final A/B validation not yet completed

Release candidate. Static review passed with two independent verifiers; the A/B benchmark against no-skill is designed but not run yet.

Multi-agent system designed and orchestrated by me.

Checkpoint

Evidence tier: Existing evidence

Skill-tree digest (the tree that passed step 2)
8551e9a4a99b7b6b8e67c1530022adc2bfa15274545f57b50d1c842ab491593b
MANIFEST.sha256
684 lines · sha256 39f269514ec828c05051c5bbcb5b8b2536c5b2f4511c441250efd561c8bf96c7

Computed locally on 2 October 2026; the skill files are not published. Open the record

A multi-agent AI workflow I designed and directed with Claude Code: it builds and tests a reusable web-design skill (builder, reviewer, fixer and verifier agents; checkpoints and SHA-256 hashes). The first version of this CV was built with it.

See proof
  1. Claim

    A multi-agent AI workflow I designed and directed with Claude Code: it builds and tests a reusable web-design skill (builder, reviewer, fixer and verifier agents; checkpoints and SHA-256 hashes). The first version of this CV was built with it.

    Stated by Danilo

  2. Evidence

  3. Result

    Verbatim excerpts of the review log (28/09/2026): four rounds of two isolated critics each ended FAIL; after the last fix round, two verifiers closed every serious finding and step 2 passed. The skill is a release candidate: the A/B benchmark against no-skill is designed but not run yet. That Danilo designed and directed the workflow, and that the first version of this CV was built with it, are his statements.

AI-assisted development with Claude Code; multi-agent workflows with builder, reviewer, fixer and verifier roles; designing and testing Claude Code skills

See proof
  1. Claim

    AI-assisted development with Claude Code; multi-agent workflows with builder, reviewer, fixer and verifier roles; designing and testing Claude Code skills

    Stated by Danilo

  2. Evidence

  3. Result

    The FWEA review log (28/09/2026) and checkpoint record show a multi-agent workflow testing a skill: isolated critic agents with read-only tools, then two verifier agents that closed every serious finding. That Danilo designed the workflow, and the rest of this claim, are his statement.

The review log, verbatim

6 line ranges of the log, copied unchanged; the file itself is not published.

Evidence tier: Existing evidence

tools/step2/STEP2_RUN_LOG.md lines 10–14 without 11

Raw text of tools/step2/STEP2_RUN_LOG.md, lines 10–14 without 11
Isolation check at launch 1 (from the stream-json init events):
line 11 left out
- MCP servers: none;
- tools: Read/Grep/Glob plus session utilities; no Agent, Task, Skill, Bash, Edit, Write, WebFetch or WebSearch;
- permission mode: default.

Evidence tier: Existing evidence

FWEA review log: the critics’ isolation check · copied 2 October 2026 · sha256 b18466513995…

tools/step2/STEP2_RUN_LOG.md lines 16–19

Raw text of tools/step2/STEP2_RUN_LOG.md, lines 16–19
## Round 0 result (launch 3, both critics valid: 0 void tool calls)
- L1 (082f4db933bb): FAIL — 3 BLOCKER, 12 MAJOR, 11 MINOR (17 tool calls, all inside packet)
- L2 (7807a7e16c55): FAIL — 2 BLOCKER, 10 MAJOR, 5 MINOR (18 tool calls, all inside packet; 279 s)
- Step 2 = FAIL on the pre-RC1 tree def5e846…. Correction round 1 of max 2 starts (Spec §13 Correction rounds).

Evidence tier: Existing evidence

FWEA review log: round 0 result · copied 2 October 2026 · sha256 290557a856df…

tools/step2/STEP2_RUN_LOG.md lines 27–30

Raw text of tools/step2/STEP2_RUN_LOG.md, lines 27–30
## Round 1 result (fresh critics, both valid: 0 void tool calls)
- L1 (0db2ad0aa774): FAIL — 2 BLOCKER, 15 MAJOR, 7 MINOR (17 tool calls, all inside packet)
- L2 (12694d35367d): FAIL — 0 BLOCKER, 4 MAJOR, 8 MINOR (17 tool calls, all inside packet)
- Step 2 = FAIL on tree e2f346bc…. Correction round 2 of max 2 (last) starts; snapshot fwea-workspace/snapshots_pre_round2_skill. Fixer instructed with explicit anti-overfitting rules (critics can read the test spec).

Evidence tier: Existing evidence

FWEA review log: round 1 result · copied 2 October 2026 · sha256 75d069d1619f…

tools/step2/STEP2_RUN_LOG.md lines 38–42

Raw text of tools/step2/STEP2_RUN_LOG.md, lines 38–42
## Round 2 result (fresh critics, both valid: 0 void tool calls)
- L1 (1bab4fd91aff): FAIL — 1 BLOCKER, 11 MAJOR, 10 MINOR (18 tool calls)
- L2 (17b649e28abc): FAIL — 0 BLOCKER, 3 MAJOR, 9 MINOR (19 tool calls)
- Trend L1 B/M: 3/12 → 2/15 → 1/11; L2 B/M: 2/10 → 0/4 → 0/3.
- Step 2 = FAIL after the 2 allowed correction rounds → OWNER DECISION required (Spec §13).

Evidence tier: Existing evidence

FWEA review log: round 2 result · copied 2 October 2026 · sha256 6925e87c1316…

tools/step2/STEP2_RUN_LOG.md lines 55–58

Raw text of tools/step2/STEP2_RUN_LOG.md, lines 55–58
## Final critics (v1.1 isolation) — both valid, 0 void tool calls
- L1 (4d95e6389f06): FAIL — 2 BLOCKER, 13 MAJOR, 5 MINOR (14 tool calls)
- L2 (9bd3b623e9f7): FAIL — 0 BLOCKER, 4 MAJOR, 7 MINOR (15 tool calls)
- No finding argues from a test item (the critics had no test material). Residual serious findings: 2 BLOCKER + 17 MAJOR (L1-01 and L2-04 overlap, so 18 unique). They go to the owner one by one (D-FWEA-1); nothing is accepted automatically.

Evidence tier: Existing evidence

FWEA review log: final critics result · copied 2 October 2026 · sha256 21552fc94357…

tools/step2/STEP2_RUN_LOG.md lines 66–70

Raw text of tools/step2/STEP2_RUN_LOG.md, lines 66–70
## Targeted verification result (D-FWEA-2) — STEP 2 PASS
- V1 (80df4d8e9a39): CLEAN — 18/18 findings CLOSED; regression 0 BLOCKER, 0 MAJOR, 5 MINOR; scripts sizes PASS, dead refs PASS, 0 duplicate groups (15 tool calls, 0 void).
- V2 (4fc38e62dc7a): CLEAN — 19/19 entries CLOSED (L2-04 counted separately); regression 0 BLOCKER, 0 MAJOR, 4 MINOR; scripts PASS (15 tool calls, 0 void).
- No finding among the 18 was accepted implicitly: all were fixed and independently verified closed.
- STEP 2 = PASS on tree 8551e9a4a99b7b6b8e67c1530022adc2bfa15274545f57b50d1c842ab491593b (2026-09-28).

Evidence tier: Existing evidence

FWEA review log: targeted verification result · copied 2 October 2026 · sha256 5f7e8e932477…