The Harness Around the Model

A thread you can test

Prompt Security

8 notes move from the word to a real choice at work — understand it first, then decide whether to use it.

READING THREADOPEN
8notes
HOW TO READStart where you are stuck, then follow the evidence and trade-offs

Each note stands alone, or becomes the next step in this thread.

The Harness Around the ModelNo login

THE QUESTION THIS PAGE ANSWERS

ANSWER FIRST

What is Prompt Security, and which AI decisions does it change?

SQL injection analogy → message list essence → lack of parameterization → overview of 5 attack types This page keeps the related concepts, common mistakes, and practical notes in one reading thread.

DECISION RULE

First decide whether you are blocked by a definition, a choice, or verification; then choose the closest of the 8 notes below.

TRY NEXT

Start with “Prompt Injection: Why Attacks Work,” then restate the conclusion using your own task.

WATCH FOR

Do not treat every method in a topic as interchangeable. The answer changes with the input, risk, and acceptance bar.

THIS QUESTION THREAD

Put the word back inside the choice it changes.

8 notes
Security

Prompt Injection: Why Attacks Work

SQL injection analogy → message list essence → lack of parameterization → overview of 5 attack types

The Harness Around the Model 3 min →
Security

Prompt Injection: 12 Attack Cases

Privilege escalation / role-play / Few-Shot / structural injection / metaphor disguise — vulnerable vs defended versions

The Harness Around the Model 3 min →
Hands-on

Prompt Defense: Three-Layer Interception

Input-layer regex → prompt-layer constraints → output-layer leak detection → secondary review; simulate the full attack chain

The Harness Around the Model 5 min →
Security

AI Safety Red Lines: Four Boundaries

What must not be done, consequences, and the four types of safety boundaries every PM must uphold

The Harness Around the Model 3 min →
Security

Risk Classification & Accountability

AI output risk classification model, role-based responsibility assignment and governance framework

The Harness Around the Model 3 min →
Concept

How Much Freedom Should AI Have?

Fully autonomous vs step-by-step approval — five permission modes and their use cases

From Working Demo to Useful Product 3 min →
Hands-on

Too Many Popups Annoy Users, None Is Unsafe

The Human-in-the-loop balance point: a risk-tier approach

From Working Demo to Useful Product 3 min →
Architecture

Do You Know What the Agent Did?

Event streams and Token tracking. Without logs, you'll never know what went wrong

From Working Demo to Useful Product 3 min →