KSKS Security Research
Learning map / Session 4 / Scenes 01–13

WORKSHOP 01 · BUILD

Build it. Attack it.Decide whether it may ship.

A practical build sequence from bounded task and non-goals through skill, structured output, narrow tools, controls, attack tests, and ship decision.

workshopsecure agentarchitecture reviewbuild guide
Learning guide
Level
Start here
Reading time
15 min
Presentation
Session 4
Progress
1 of 3

01 · Mental model

The deliverable is a controlled system, not a clever prompt

Start with a bounded, repeated task and an expert reviewer. Write non-goals before the system instruction. Threat-model the workflow before selecting tools. Define structured output before model orchestration. Package the expert method as a skill. Expose narrow read-only verbs. Add budgets and stop states before more intelligence.

This order makes the system testable at each boundary and prevents the prompt from becoming an informal substitute for architecture, authorization, or evidence.

02 · Visual explanation

01Taskbounded outcome
02Threat modelassets + abuse paths
03Contractschema + non-goals
04Toolsnarrow permissions
05Evaluatequality + safety
The secure build sequenceEach step creates an artifact that the next step can validate.

03 · Compare and decide

A demo versus a releasable capability

Decision lensDemo provesRelease evidence proves
PossibilityOne compelling run can workRepresentative cases pass repeatedly
ControlThe happy path looks boundedAttacks and failures stop safely
QualityOutput looks usefulRubrics, schemas, and reviewers agree
OperationsThe UI respondsCost, latency, logs, rollback, and ownership are ready

04 · Cybersecurity example

First agent: architecture review assistant

The agent reviews a supplied design and produces findings without changing any environment.

01

Define required input and non-goals.

02

Package the review method and findings schema.

03

Use retrieval for official guidance and read-only metadata tools.

04

Attack the inputs, score results, and require expert acceptance.

Outcome: The first release is useful, bounded, observable, and intentionally unable to deploy changes.

05 · What to remember

The 60-second recall

01

Choose a task with a clear expert reviewer and definition of done.

02

Write non-goals and output contracts before prompts and tools.

03

Ship only what can be bounded, tested, observed, and stopped.

Teach-back prompt: Explain this concept to a teammate using the diagram, then name one failure mode and the control that stops it.

06 · Questions people ask

FAQ

A repeated, read-only, evidence-based task with bounded inputs, structured outputs, low side effects, and an available expert reviewer.

07 · Primary sources

Continue with authoritative guidance