Skip to main content
reopt Handbook
reopt Handbook
LLMOps and AgentOps in Production

Architecture and Release

Ch1. System ArchitectureCh2. Versioning and ReleaseCh3. Evaluation Framework

Operational Reliability

Ch4. Online GuardrailsCh5. Observability and SLOsCh6. Cost and Latency Optimization

Growth and Response

Ch7. Experiment OperationsCh8. Incident Management Runbook

Verification

Verification ReportVerification ArchiveUpdates
Handbook›LLMOps and AgentOps›Updates
한국어English

Updates

Changelog for LLMOps and AgentOps in Production

Last Updated

September 5, 2026

Change History

2026-09-05

Change Summary

  • Added a current Anthropic lineup table to cost-latency.mdx, checked 2026-09-05: Fable 5.1 (10/10/10/50 with 0.25cachereads),Opus5(0.25 cache reads), Opus 5 (0.25cachereads),Opus5(5/25),Sonnet5(25), Sonnet 5 (25),Sonnet5(2/10),andHaiku4.5(10), and Haiku 4.5 (10),andHaiku4.5(1/$5).
  • Kept the 2026-05-17 snapshot as the budget basis of its own date and pointed new routing policies at the current table instead.
  • Noted that Fable 5.1 charges 2.5% of the input price for cache reads, versus 10% on other Claude models, which changes the cache-reuse economics of long agentic sessions.

Verification Scope

Only the Anthropic pricing and lineup rows were rechecked. The OpenAI and DeepSeek rows and the rest of the chapters remain at their 2026-06-13 verification state.

Sources

  • Claude models overview: https://platform.claude.com/docs/en/about-claude/models/overview
  • Claude Fable 5.1 overview: https://platform.claude.com/docs/en/models/fable-5-1/overview

2026-06-13

Change Summary

  • Normalized a non-standard Callout type in the "Security First" note (warning → warn) so it renders as a supported variant.
  • Minor review-pass cleanup; no changes to guidance, formulas, or source baselines.

Affected Chapters

  • versioning-release.mdx

2026-05-18

Change Summary

  • Created the English edition from the Korean llmops-agentops handbook.
  • Localized chapter titles, navigation links, formulas, callouts, evidence tables, and verification records.
  • Preserved source baselines and freshness markers from the 2026-05-17 Korean 4th verification.

Affected Chapters

  • index.mdx
  • system-architecture.mdx
  • versioning-release.mdx
  • evaluation.mdx
  • online-guardrails.mdx
  • observability-slo.mdx
  • cost-latency.mdx
  • experimentation.mdx
  • incident-management.mdx
  • verification.mdx
  • verification-archive.mdx
  • updates.mdx

2026-05-17

Source Edition Change Summary

  • Corrected A2A from v0.3.0 Draft to latest v1.0.0 (Ch1)
  • Added MCP 2025-11-25 security requirements: OAuth 2.1, Protected Resource Metadata, Client ID Metadata Documents, token audience binding, token passthrough prohibition (Ch1)
  • Reflected OpenAI Agents SDK operating surfaces: guardrails, human review, resumable approval state, hosted/private MCP, tracing, agent evals, voice agents (Ch3-Ch5)
  • Added supply-chain, permission, and telemetry controls from OWASP MCP Top 10 and Agentic Skills Top 10 (Ch4, Ch8)
  • Corrected OTel GenAI to Development and OWASP AOS to work-in-progress (Ch5)
  • Refreshed model pricing around GPT-5.5/GPT-5.4/GPT-5.4 mini, Claude Opus/Sonnet/Haiku 4.x, and DeepSeek V4 Flash/Pro (Ch6)
  • Fixed the λ_risk mismatch in the experiment decision formula (Ch7)
  • Added MCP/Skill compromise, A2A abuse, and voice/realtime degradation incident types (Ch8)
  • Added chapter-level source baselines, check:freshness, operational examples, and verification archive separation

Primary Sources

  • https://developers.openai.com/api/docs/guides/agents
  • https://developers.openai.com/api/docs/guides/agents/guardrails-approvals
  • https://developers.openai.com/api/docs/guides/agents/integrations-observability
  • https://developers.openai.com/api/docs/guides/agent-evals
  • https://openai.com/api/pricing/
  • https://modelcontextprotocol.io/specification/2025-11-25
  • https://a2a-protocol.org/latest/specification/
  • https://opentelemetry.io/docs/specs/semconv/gen-ai/
  • https://aos.owasp.org/aos/
  • https://owasp.org/www-project-mcp-top-10/
  • https://owasp.org/www-project-agentic-skills-top-10/
  • https://docs.claude.com/en/docs/test-and-evaluate/strengthen-guardrails/handle-streaming-refusals
  • https://platform.claude.com/docs/en/about-claude/pricing
  • https://api-docs.deepseek.com/quick_start/pricing/
  • https://www.langchain.com/blog/introducing-langsmith-fleet
  • https://www.braintrust.dev/docs/loop
  • https://www.pagerduty.com/newsroom/pagerduty-expands-ai-ecosystem-to-supercharge-ai-agents/

Template

### YYYY-MM-DD

**Change Summary**

- ...

**Affected Chapters**

- `...`

**Primary Sources**

- ...

Related docs

Updates

Harness Engineering · Change log for the Harness Engineering handbook

Verification Report

Structure, link, metric, and logic verification for LLMOps and AgentOps in Production

Verification Report

Claude Code Complete Guide · Verification checklist for the English Claude Code handbook locale.

Earlier verification records

Claude Code Command Master · Claude Code: Earlier review dates, version baselines, and findings, with a link to the current verification scope.

Updates

AI-Era GTM · Change log and source refresh notes for the AI-Era GTM handbook.

Verification Archive

Historical verification records and correction history for LLMOps and AgentOps in Production

On this page

Change History2026-09-052026-06-132026-05-182026-05-17Template