최상위 10_Wiki/Topic_*였던 4개 카테고리 폴더를 10_Wiki/Topics/Topic_* 로 재배치.
콘텐츠 변경 없음(순수 폴더 이동) — Topics/ 하위 나머지 폴더는 이미 지난 커밋에서
전부 정리된 상태(잔존 항목은 에이전트 운영 상태 및 사용자가 보존을 요청한
업데이트0615/무제 3.canvas 뿐).
"매 reliability 의 feature 의 — 매 first feature 의". SRE (Site Reliability Engineering) 의 Google-originated discipline 의 software engineering 의 ops 의 applying. 핵심: SLOs 의 define, error budgets 의 enforce, toil 의 eliminate, blameless postmortems.
매 핵심
매 SRE 의 핵심 의 concepts
SLI: 매 measurement (e.g., 200-OK rate over 5min).
SLO: 매 target (e.g., 99.9% over 28d rolling).
SLA: 매 customer contract (with $ penalty).
Error budget: 매 100% - SLO. 매 budget 의 burn 시 release freeze.
언제: postmortem drafting from timeline, log anomaly summarization, runbook generation, oncall question answering.
언제 X: auto-remediation 의 LLM-only — 매 hallucinated kubectl 의 prod 의 destroy.
❌ 안티패턴
No SLO: 매 alert noise — 매 every blip 의 page.
100% uptime goal: 매 unattainable, 매 budget 0 = no innovation.
Blame culture: postmortem 의 finger-pointing — engineers 의 hide incidents.
Toil unbounded: SREs 의 burned out — quit within 12mo.