Skip to content

All Content tagged with Operational Excellence

The operational excellence pillar focuses on running and monitoring systems, and continually improving processes and procedures. Key topics include automating changes, responding to events, and defining standards to manage daily operations.

Content language: English

Filter content
Select tags to filter
Sort by
Sort by most recent
40 results
AWS Agent Registry is now generally available, but most teams meet it mid-sprawl: dozens of agents and MCP tools scattered across accounts, with no shared inventory, no cross-team discovery, and no au...
How to install Agent Toolkit for AWS including AWS MCP server for Claude Desktop app and Claude Code CLI
Learn operational best practices for running InfluxDB v2 workloads on AWS — covering both Amazon Timestream for InfluxDB (managed) and self-managed InfluxDB on Amazon EC2. This article addresses commo...
This article explains how S3 Account Regional Namespaces work in practice and provides two ready-to-use engineering patterns for integrating them into your applications. It's written for developers an...
How to install Agent Toolkit for AWS for ChatGPT app
CloudWatch alarms fire when the graph looks clean. They take minutes to react to obvious spikes. They get stuck in INSUFFICIENT_DATA for no apparent reason. These are among the most common questions o...
How to install Kiro IDE and Agent Toolkit for AWS including AWS MCP server to build, deploy and manage your AWS environment with natural language prompts. Include optional steps to install Open Source...
UK Organisations running Amazon WorkSpaces in AUTO_STOP mode cannot receive patches when powered down. The built-in monthly maintenance window does not meet the Cyber Essentials 14-day patching requir...
This article explains how to use the structured data in AWS Health Planned Lifecycle Events to create and manage change requests in ServiceNow, with practical approaches for organizing change tasks ba...
This article shows how AWS Unified Operations helps financial institutions enhance their overall operational excellence to meet Digital Operational Resilience Act (DORA) requirements.
Learn how to integrate Dynatrace with AWS Incident Detection and Response to automate incident response and create context-rich support cases that expedite issue resolution.
This article addresses a common knowledge gap among cloud architects and developers who often misunderstand how Service Level Agreements (SLAs) work in distributed systems.
  • 1
  • 2
  • 3
  • 4
  • Page size
    12 / page