evaluation 4
- AI Agent Security (Part 9): From Firing to Density — Chains, Gates, and the Per-K Frontier
- AI Agent Security (9편): 작동 여부에서 점수 밀도로 — 연쇄 호출, 선택 기준, 그리고 K별 성능 한계
- AI Agent Security (Part 8): The Evaluation Reset — Partial Banking and the Search for a Discrete Lever
- AI Agent Security (8편): 평가 체계 재설정 — 부분 점수 보존과 불연속적 도약의 원인 찾기