DevOps
Fundamentals, automation strategy, incident practice and the platform-versus-toolchain question.
What does the DevOps cluster cover?
The DevOps cluster on the DevOpsArk blog collects 5 articles on fundamentals, automation strategy, incident practice and the platform-versus-toolchain question.
Everything in DevOps
Blameless postmortems: the format that actually gets filled in
Most postmortem templates go unused after the first few sections. Why that happens, and what a blameless postmortem format looks like when people actually complete it.
DevOps platform or toolchain? An honest comparison
The real trade-off between assembling best-of-breed tools and adopting an integrated platform, including the costs of each that vendors on both sides tend not to mention.
How to automate Docker builds without hand-writing Dockerfiles
Multi-stage builds, layer caching that actually works, hardening defaults, and how to generate and maintain container definitions across a large service estate.
DevOps automation: what to automate, in what order
A practical sequence for automating a delivery path, which step gives the most back first, which automation tends to be regretted, and how to tell when a stage is genuinely done.
What is DevOps? A definition that survives contact with practice
What DevOps actually means, where the definition came from, what the lifecycle looks like in practice, and the common misreadings that turn it into a job title instead of a way of working.
What the platform does about this
Log Management
Centralised logs with structure and retention
AI Log Analysis
Find the line that matters
Monitoring
Infrastructure and application monitoring
Alerting
Alerts that are worth waking up for
ArkApps
Application inventory and lifecycle
Pipelines
CI/CD pipelines with policy and provenance
ArkCD
Continuous delivery and progressive rollout
360 DITE
Delivery, infrastructure, testing and experience in one score
Elsewhere on the blog
Kubernetes
Architecture, monitoring, deployment strategy, multi-cluster operations and troubleshooting.
Observability
Metrics, logs, traces, alerting design and what observability actually means.
AI DevOps
Agentic DevOps, AI log analysis, incident response and where automation should stop.
DevSecOps
Container and Kubernetes security, vulnerability management, secrets handling, audit trails and compliance evidence.
Cloud and cost
Multi-cloud operations, cost optimisation, infrastructure drift and infrastructure hygiene.
Want a structured route through this?
Learning tracks arrange these articles into an ordered path with the glossary terms they depend on.