OnProud
Module 07 Production, Recovery & What Comes Next

Launch & Operate · Module 07

Production, Recovery & What Comes Next

เราจะเปิด product ให้คนใช้จริง รู้เมื่อมันพัง กู้กลับ และดูแลให้เปลี่ยนต่อได้อย่างไร?

Production readinessDeployMonitoringBackup/restoreRollbackHandoff

Readiness review

Deploy สำเร็จยังไม่เท่ากับพร้อมให้คนใช้

กลับมาตรวจ Useful, Correct, Safe enough, Operable, Recoverable และ Changeable จาก Product Plan, tests, beta, migration และ Agent boundaries ก่อนกำหนด launch scope

Production plan

รู้ version, data change, secrets และทางถอยก่อนปล่อย

แผนต้องระบุ release commit, environment, migrations, backups, feature flags, smoke tests, monitoring, owner, stop condition และ rollback

Prompt · Draft the production launch plan
อ่าน production readiness evidence และ repository ปัจจุบัน
เสนอ launch plan: release commit, secrets/config, database migration,
backup, deploy order, smoke tests, monitoring, rollback, owner และ stop conditions
แยก experimental Agent write หลัง feature flag
รอ Human approval ห้าม deploy

Deploy and smoke

ปล่อยทีละจุดแล้วตรวจ Primary Path ทันที

หลัง deploy ตรวจ health, sign in, add book, Search/ISBN, autocomplete fallback และ user isolation ด้วย test account โดยไม่ใช้ข้อมูลส่วนตัวจริง

Observe

ถ้าไม่มี signal เราจะรู้ปัญหาจากผู้ใช้

ตั้ง uptime/error monitoring และ release markers ให้ตอบได้ว่าเริ่มพังเมื่อไร, กระทบ flow ใด และสัมพันธ์กับ release ไหน โดยไม่ log secret หรือข้อมูลเกินจำเป็น

Backup and restore

Backup ที่ไม่เคย restore ยังไม่ใช่หลักฐาน

ระบุข้อมูลที่ backup, retention, owner และวิธี restore ไปพื้นที่ปลอดภัย จากนั้นซ้อม restore และตรวจ record/critical flow

Rollback drill

ตอนระบบพังไม่ใช่เวลาคิดทางถอยจากศูนย์

ซ้อม detect → contain → rollback/disable → verify → communicate พร้อมแยกกรณี code rollback ได้แต่ data migration ย้อนตรง ๆ ไม่ได้

Detect Contain Roll back or restore Verify Learn

Incident runbook

เขียน safe first actions และจุดที่ต้องหยุดเรียกผู้เชี่ยวชาญ

Runbook ระบุ signal, impact, owner, credential revoke, rollback/restore, verification, communication และ escalation สำหรับ money, privacy หรือ data-loss risk

Final demonstration

Demo problem, decisions, evidence และ failure—not just happy path

เล่า Intake → Product/Design/Build Plan → Core Path → Beta Change → Data/API evolution → Agent boundary → Production/Recovery พร้อมหลักฐานแต่ละ gate

Continue

Product ที่ live แล้วคือจุดเริ่มต้นของรอบถัดไป

วางหนึ่ง outcome 30 วัน, หนึ่ง risk 90 วัน, maintenance rhythm สำหรับ feedback, cost, errors, dependencies และ backup พร้อมสิ่งที่ตั้งใจยังไม่ทำ

Graduation Gate

เจ้าของ product ต้องรู้ทั้งก้าวต่อไปและขอบเขตตัวเอง

จบเมื่อผู้เรียนดูแล release เล็ก, อ่าน signal, ใช้ runbook, อธิบาย recovery และรู้ว่าเมื่อไรต้องหยุดเรียก security, legal, data หรือ operations expert

Useful. Correct. Safe enough. Operable. Recoverable. Changeable.

Module checkpoint

สิ่งที่ต้องนำออกจากห้อง

MyShelf milestone

ProudVault อยู่ production พร้อม monitoring, recovery evidence และแผน 30/90 วัน

Deliverable

Production URL, launch checklist, incident runbook, restore/rollback evidence และ continuation plan

Exit gate

Critical flow ใช้งานจริง, monitoring เห็น failure, rollback/restore ซ้อมได้ และเจ้าของรู้จังหวะดูแลหรือ escalate

จบ module นี้แล้ว — ไปต่อหรือกลับไปดูภาพรวม

← Module 06 ดูทุก module