> ## Documentation Index
> Fetch the complete documentation index at: https://notes.kodekloud.com/llms.txt
> Use this file to discover all available pages before exploring further.

# DevOps Practices in Platform Engineering

> Overview of DevOps and SRE practices for platform engineering, emphasizing flow, feedback, automation, SLOs, observability, and building platform teams to improve developer velocity and reliability

Welcome. This lesson continues the fundamentals of platform engineering by focusing on DevOps practices that are essential for building and operating a platform that scales.

Agenda

* The Three Ways of DevOps: flow, feedback, and experimentation
* Key SRE principles: reduction of toil, incident management, SLOs/SLIs, observability
* Building a collaborative platform culture
* How infrastructure teams transform into platform product teams

<Frame>
  <img src="https://mintcdn.com/kodekloud-c4ac6d9a/og-mfTVvAAl8u5l1/images/Prep-Course-Certified-Cloud-Native-Platform-Engineering-Associate-CNPA/Domain-1-Platform-Engineering-Core-Fundamentals/DevOps-Practices-in-Platform-Engineering/agenda-devops-sre-platform-product-teams.jpg?fit=max&auto=format&n=og-mfTVvAAl8u5l1&q=85&s=c3484e376ea67c4ba27666d85729aa71" alt="A presentation slide titled &#x22;Agenda&#x22; with a blue gradient panel on the left and four colorful numbered markers down the middle. The right side lists agenda items about DevOps, SRE principles, collaborative platform culture, and transforming infrastructure teams into product teams." width="1920" height="1080" data-path="images/Prep-Course-Certified-Cloud-Native-Platform-Engineering-Associate-CNPA/Domain-1-Platform-Engineering-Core-Fundamentals/DevOps-Practices-in-Platform-Engineering/agenda-devops-sre-platform-product-teams.jpg" />
</Frame>

Why platform engineering and DevOps belong together

Platform engineering does not replace DevOps — it amplifies and operationalizes it across the company. A successful platform product embeds developer ergonomics, SRE targets, and automation so teams can safely move faster.

Four key connections between platform engineering and DevOps:

* Shared mission across teams — align on developer experience and reliability goals.
* Break down silos — enable cross-functional collaboration (dev, ops, QA, security).
* Scale through automation — treat repetitive work as code and automate it.
* Drive self-service success — enable internal consumers to do more without asking.

At Sparkle Pony Ranch, the platform team combines complementary skills:

* Swati: deep DevOps and CI/CD automation
* Alan: infrastructure and IaC (infrastructure-as-code)
* Phuong: developer workflows and cloud-native platform integration

<Frame>
  <img src="https://mintcdn.com/kodekloud-c4ac6d9a/og-mfTVvAAl8u5l1/images/Prep-Course-Certified-Cloud-Native-Platform-Engineering-Associate-CNPA/Domain-1-Platform-Engineering-Core-Fundamentals/DevOps-Practices-in-Platform-Engineering/platform-engineering-devops-foundations-pillars.jpg?fit=max&auto=format&n=og-mfTVvAAl8u5l1&q=85&s=fb7a0f4660dd01dd52728961a32ef7fb" alt="An infographic titled &#x22;Platform Engineering Built on DevOps Foundations&#x22; showing four colored pillars—Shared Mission, Internal Enablement, Scale Through Automation, and Self‑Service Success—each with an icon and brief description. The slide is branded © KodeKloud." width="1920" height="1080" data-path="images/Prep-Course-Certified-Cloud-Native-Platform-Engineering-Associate-CNPA/Domain-1-Platform-Engineering-Core-Fundamentals/DevOps-Practices-in-Platform-Engineering/platform-engineering-devops-foundations-pillars.jpg" />
</Frame>

<Callout icon="lightbulb" color="#1CB2FE">
  Platform teams succeed when they treat the platform as a product: define key user journeys, instrument reliability targets, and continuously measure developer experience.
</Callout>

Three Ways of DevOps — foundation for platform design

The “Three Ways” (from The Phoenix Project) are practical principles to embed in platform architecture: Flow, Feedback, and Continuous Experimentation.

1. First Way — Flow

* Goal: optimize frictionless workflows from code to production.
* Approach: automate verification, CI/CD, and promote continuous delivery so commits move rapidly through build, test, and deployment.

Typical flow: developer commits → CI builds the artifact → a GitOps operator (e.g., Argo CD) applies manifests → production in minutes.

<Frame>
  <img src="https://mintcdn.com/kodekloud-c4ac6d9a/og-mfTVvAAl8u5l1/images/Prep-Course-Certified-Cloud-Native-Platform-Engineering-Associate-CNPA/Domain-1-Platform-Engineering-Core-Fundamentals/DevOps-Practices-in-Platform-Engineering/cicd-flow-developer-github-argocd-production.jpg?fit=max&auto=format&n=og-mfTVvAAl8u5l1&q=85&s=5e2bb38071b536b94ab44a36b619f40e" alt="A slide titled &#x22;First Way – Flow&#x22; that visually shows a four-step CI/CD pipeline. The steps are: Developer commits code → GitHub Actions builds → ArgoCD deploys → Production in minutes, each shown in blue chevrons with simple icons." width="1920" height="1080" data-path="images/Prep-Course-Certified-Cloud-Native-Platform-Engineering-Associate-CNPA/Domain-1-Platform-Engineering-Core-Fundamentals/DevOps-Practices-in-Platform-Engineering/cicd-flow-developer-github-argocd-production.jpg" />
</Frame>

Infrastructure automation follows the same pattern: changes to IaC (for example, Terraform) should trigger plan, tests/validation, peer review, and safe promotion to production-like environments with verification gates.

2. Second Way — Feedback

* Goal: provide rapid, continuous feedback at every stage of the delivery pipeline.
* Approach: shorten the time from change to verification and notification so engineers quickly know success or failure and can iterate.

Fast feedback helps teams detect regressions early, rely on observability, and maintain production readiness.

<Frame>
  <img src="https://mintcdn.com/kodekloud-c4ac6d9a/og-mfTVvAAl8u5l1/images/Prep-Course-Certified-Cloud-Native-Platform-Engineering-Associate-CNPA/Domain-1-Platform-Engineering-Core-Fundamentals/DevOps-Practices-in-Platform-Engineering/second-way-feedback-flow-30s-notify.jpg?fit=max&auto=format&n=og-mfTVvAAl8u5l1&q=85&s=52d5667b2dd62f2e383a6da4f5eadda7" alt="An infographic titled &#x22;Second Way – Feedback&#x22; showing a three-step flow: 1) Code pushed, 2) Tests fail, and 3) Developer notified in 30 seconds." width="1920" height="1080" data-path="images/Prep-Course-Certified-Cloud-Native-Platform-Engineering-Associate-CNPA/Domain-1-Platform-Engineering-Core-Fundamentals/DevOps-Practices-in-Platform-Engineering/second-way-feedback-flow-30s-notify.jpg" />
</Frame>

3. Third Way — Continuous Learning and Experimentation

* Goal: create a culture of experimentation and learning from failures.
* Approach: run blameless postmortems, capture root causes, and harden the platform to reduce repeat incidents.

Platform teams should convert incidents into product improvements so platform reliability and developer productivity both increase.

<Frame>
  <img src="https://mintcdn.com/kodekloud-c4ac6d9a/og-mfTVvAAl8u5l1/images/Prep-Course-Certified-Cloud-Native-Platform-Engineering-Associate-CNPA/Domain-1-Platform-Engineering-Core-Fundamentals/DevOps-Practices-in-Platform-Engineering/third-way-continuous-learning-outage.jpg?fit=max&auto=format&n=og-mfTVvAAl8u5l1&q=85&s=298cf226b544cebcb4f37f223d442cec" alt="A slide titled &#x22;Third Way – Continuous Learning&#x22; showing three colored avatar icons labeled Swati, Alan, and Phuong and a tag &#x22;Sparkle Pony Ranch.&#x22; In the center is a computer warning labeled &#x22;Outage occurs&#x22; with two steps: &#x22;1. Team investigates root cause&#x22; and &#x22;2. Platform is improved for all.&#x22;" width="1920" height="1080" data-path="images/Prep-Course-Certified-Cloud-Native-Platform-Engineering-Associate-CNPA/Domain-1-Platform-Engineering-Core-Fundamentals/DevOps-Practices-in-Platform-Engineering/third-way-continuous-learning-outage.jpg" />
</Frame>

Key operating practices derived from the Three Ways

* Defer irreversible decisions: only commit to provider-specific or heavyweight capabilities when they solve a real need.
* Amplify feedback loops: instrument and automate checks so verification is fast and reliable.
* Work in small batches: keep changes small and frequent to reduce risk and simplify rollbacks.

<Frame>
  <img src="https://mintcdn.com/kodekloud-c4ac6d9a/og-mfTVvAAl8u5l1/images/Prep-Course-Certified-Cloud-Native-Platform-Engineering-Associate-CNPA/Domain-1-Platform-Engineering-Core-Fundamentals/DevOps-Practices-in-Platform-Engineering/platform-engineering-three-principles.jpg?fit=max&auto=format&n=og-mfTVvAAl8u5l1&q=85&s=ff214dec69b5a6995d56fef2be151a49" alt="A presentation slide titled &#x22;Three Key Principles Applied to Platform Engineering&#x22; with three colored boxes: &#x22;Make Decisions Late,&#x22; &#x22;Amplify Feedback Loops,&#x22; and &#x22;Work in Small Batches.&#x22; Each box has a short note: enable flexible architectures for pivoting; shorten time from change to validation; and use incremental changes to reduce risk." width="1920" height="1080" data-path="images/Prep-Course-Certified-Cloud-Native-Platform-Engineering-Associate-CNPA/Domain-1-Platform-Engineering-Core-Fundamentals/DevOps-Practices-in-Platform-Engineering/platform-engineering-three-principles.jpg" />
</Frame>

Concrete role examples (Sparkle Pony Ranch)

* Alan: builds declarative, reusable Terraform modules to provision cloud infrastructure.
* Swati: automates CI/CD tests and reliability checks, integrating quick feedback into pipelines.
* Phuong: delivers small feature increments to developer-facing tools (for example, a developer portal like Backstage) through frequent, well-scoped pull requests.

<Frame>
  <img src="https://mintcdn.com/kodekloud-c4ac6d9a/og-mfTVvAAl8u5l1/images/Prep-Course-Certified-Cloud-Native-Platform-Engineering-Associate-CNPA/Domain-1-Platform-Engineering-Core-Fundamentals/DevOps-Practices-in-Platform-Engineering/platform-engineering-three-key-principles-avatars.jpg?fit=max&auto=format&n=og-mfTVvAAl8u5l1&q=85&s=16fc4744f2f575633fd1bed50c86d2ea" alt="A presentation slide titled &#x22;Three Key Principles Applied to Platform Engineering&#x22; showing three colorful avatar icons labeled Swati, Alan, and Phuong with short role descriptions (Swati: automates testing and monitoring; Alan: creates reusable Terraform templates for AWS and Azure; Phuong: works in small batches with clear pull requests). A &#x22;Sparkle Pony Ranch&#x22; tag and KodeKloud copyright appear on the slide." width="1920" height="1080" data-path="images/Prep-Course-Certified-Cloud-Native-Platform-Engineering-Associate-CNPA/Domain-1-Platform-Engineering-Core-Fundamentals/DevOps-Practices-in-Platform-Engineering/platform-engineering-three-key-principles-avatars.jpg" />
</Frame>

Outcomes: faster validation, smaller blast radius

These practices create a platform that is reliable, evolves quickly, and enables developers to ship with confidence: smaller changes, faster reviews, easier rollbacks, and reduced blast radius. Short feedback cycles let teams validate in minutes rather than days.

<Frame>
  <img src="https://mintcdn.com/kodekloud-c4ac6d9a/og-mfTVvAAl8u5l1/images/Prep-Course-Certified-Cloud-Native-Platform-Engineering-Associate-CNPA/Domain-1-Platform-Engineering-Core-Fundamentals/DevOps-Practices-in-Platform-Engineering/minutes-not-days-accelerated-learning-feedback.jpg?fit=max&auto=format&n=og-mfTVvAAl8u5l1&q=85&s=f22d1a59389c65672cd5bf47ac14fc75" alt="A presentation slide titled &#x22;Minutes, Not Days – Accelerated Learning&#x22; showing a left column of user avatars and names (Sparkle Pony Ranch, Swati, Alan, Phuong) and a right-side illustration of a computer screen with &#x22;Feedback&#x22; and &#x22;Pull Request&#x22; buttons." width="1920" height="1080" data-path="images/Prep-Course-Certified-Cloud-Native-Platform-Engineering-Associate-CNPA/Domain-1-Platform-Engineering-Core-Fundamentals/DevOps-Practices-in-Platform-Engineering/minutes-not-days-accelerated-learning-feedback.jpg" />
</Frame>

Flexible architecture and parameterized templates

Parameterized IaC templates let you delay irreversible decisions and simplify migration planning. For example, reusable Terraform modules capture services, configurations, and constraints in code so you can reason about architecture and map requirements to another provider later. Note that provider-specific APIs will require adaptation during a cloud migration.

<Frame>
  <img src="https://mintcdn.com/kodekloud-c4ac6d9a/og-mfTVvAAl8u5l1/images/Prep-Course-Certified-Cloud-Native-Platform-Engineering-Associate-CNPA/Domain-1-Platform-Engineering-Core-Fundamentals/DevOps-Practices-in-Platform-Engineering/flexible-architecture-parameterized-templates-aws-azure.jpg?fit=max&auto=format&n=og-mfTVvAAl8u5l1&q=85&s=65a11a4bd5d99432581c1a7ebf43416f" alt="A presentation slide titled &#x22;Flexible Architecture – Delaying Irreversible Decisions&#x22; showing a user avatar labeled &#x22;Alan&#x22; linked to &#x22;Parameterized Templates.&#x22; Below are three blue icons illustrating a flow: &#x22;Built for AWS&#x22; → &#x22;Requirements change&#x22; → &#x22;Seamless migration to Azure.&#x22;" width="1920" height="1080" data-path="images/Prep-Course-Certified-Cloud-Native-Platform-Engineering-Associate-CNPA/Domain-1-Platform-Engineering-Core-Fundamentals/DevOps-Practices-in-Platform-Engineering/flexible-architecture-parameterized-templates-aws-azure.jpg" />
</Frame>

Site Reliability Engineering (SRE) fundamentals

Embed SRE practices directly into the platform so reliability becomes measurable and actionable:

| SRE Concept | Purpose | Example |
| - | - | - |
| SLO / SLI | Define and measure acceptable reliability | `99.5%` availability for platform APIs |
| Error budget | Balance feature velocity vs. reliability | Allow limited risk for rapid delivery until budget exhausted |
| Observability | Diagnose and verify behavior across stack | Metrics, logs, traces for CI/CD, control plane, and services |
| Incident practices | Runbooks, blameless postmortems, on-call rotation | Fast incident triage and continuous improvement |

Sparkle Pony Ranch sets SLOs for CI/CD throughput and platform availability; when an SLO is breached, the team treats incidents as product gaps and iterates.

<Frame>
  <img src="https://mintcdn.com/kodekloud-c4ac6d9a/og-mfTVvAAl8u5l1/images/Prep-Course-Certified-Cloud-Native-Platform-Engineering-Associate-CNPA/Domain-1-Platform-Engineering-Core-Fundamentals/DevOps-Practices-in-Platform-Engineering/sre-platform-slos-gitops-60s-observability.jpg?fit=max&auto=format&n=og-mfTVvAAl8u5l1&q=85&s=5015f3be99657fa2c4573eb9a625d909" alt="A presentation slide titled &#x22;Site Reliability Engineering – Built Into the Platform&#x22; showing an SLO target of 99.5% and a note that the GitOps operator syncs within 60 seconds. It also lists key SRE practices (SLOs, error budgets, observability, blameless culture) and observability pillars (metrics, logs, traces)." width="1920" height="1080" data-path="images/Prep-Course-Certified-Cloud-Native-Platform-Engineering-Associate-CNPA/Domain-1-Platform-Engineering-Core-Fundamentals/DevOps-Practices-in-Platform-Engineering/sre-platform-slos-gitops-60s-observability.jpg" />
</Frame>

<Callout icon="warning" color="#FF6B6B">
  SLOs are tools, not targets to game. Use them to guide trade-offs between velocity and stability, and adjust error budgets responsibly to protect users and developer productivity.
</Callout>

Culture and organizational practices

Tools are necessary, but culture determines success. Platform engineering benefits from:

* Shared ownership — developers, operators, QA, DBAs, and architects collaborate on platform outcomes.
* Developer rotation through platform on-call — increases empathy for running services in production.
* Peer reviews and automation — senior engineers maintain quality through design reviews and automated checks.
* Adoption through collaboration — demonstrate value and enable teams rather than enforce usage.

<Frame>
  <img src="https://mintcdn.com/kodekloud-c4ac6d9a/og-mfTVvAAl8u5l1/images/Prep-Course-Certified-Cloud-Native-Platform-Engineering-Associate-CNPA/Domain-1-Platform-Engineering-Core-Fundamentals/DevOps-Practices-in-Platform-Engineering/platform-teams-shared-ownership-rotation-reviews.jpg?fit=max&auto=format&n=og-mfTVvAAl8u5l1&q=85&s=d99f7c69bdf508c9aa70fa0f51edde68" alt="A presentation slide titled &#x22;Culture Beats Tools – Building Platform Teams&#x22; with three colored icons and a segmented bar. It highlights three practices: Shared Ownership, Developer Rotation, and Peer Reviews." width="1920" height="1080" data-path="images/Prep-Course-Certified-Cloud-Native-Platform-Engineering-Associate-CNPA/Domain-1-Platform-Engineering-Core-Fundamentals/DevOps-Practices-in-Platform-Engineering/platform-teams-shared-ownership-rotation-reviews.jpg" />
</Frame>

Summary — practical checklist

* Design for flow: automate CI/CD and GitOps for repeatable deployments.
* Build fast feedback loops: tests, observability, and notifications that surface failures quickly.
* Embrace continuous experimentation: blameless postmortems and platform hardening after incidents.
* Use parameterized templates: delay irreversible choices and keep migration options open.
* Embed SRE: define SLOs/SLIs, manage error budgets, and instrument metrics/logs/traces.
* Foster shared ownership and developer empathy through rotations and collaboration.

These foundations let a platform product team (like Sparkle Pony Ranch — Swati, Alan, and Phuong) scale DevOps practices across an organization and deliver better developer velocity with predictable reliability.

<Frame>
  <img src="https://mintcdn.com/kodekloud-c4ac6d9a/og-mfTVvAAl8u5l1/images/Prep-Course-Certified-Cloud-Native-Platform-Engineering-Associate-CNPA/Domain-1-Platform-Engineering-Core-Fundamentals/DevOps-Practices-in-Platform-Engineering/devops-practices-platform-foundation.jpg?fit=max&auto=format&n=og-mfTVvAAl8u5l1&q=85&s=4f569c10c369757ed44c00e0224b3a27" alt="A presentation slide titled &#x22;DevOps Practices – The Foundation of Effective Platforms&#x22; showing a &#x22;Sparkle Pony Ranch&#x22; team with avatar icons for Swati, Alan, and Phuong on the left. On the right are three key points: &#x22;Building a platform, not just infrastructure,&#x22; &#x22;Embedding DevOps principles by design,&#x22; and &#x22;Scaling DevOps practices across the engineering org.&#x22;" width="1920" height="1080" data-path="images/Prep-Course-Certified-Cloud-Native-Platform-Engineering-Associate-CNPA/Domain-1-Platform-Engineering-Core-Fundamentals/DevOps-Practices-in-Platform-Engineering/devops-practices-platform-foundation.jpg" />
</Frame>

Links and references

* The Phoenix Project overview: [https://en.wikipedia.org/wiki/The\_Phoenix\_Project](https://en.wikipedia.org/wiki/The_Phoenix_Project)
* GitOps with Argo CD: [https://learn.kodekloud.com/user/courses/gitops-with-argocd](https://learn.kodekloud.com/user/courses/gitops-with-argocd)
* Terraform basics course: [https://learn.kodekloud.com/user/courses/terraform-basics-training-course](https://learn.kodekloud.com/user/courses/terraform-basics-training-course)
* Backstage developer portal: [https://learn.kodekloud.com/user/courses/certified-backstage-associate-cba](https://learn.kodekloud.com/user/courses/certified-backstage-associate-cba)

Keep the Three Ways — Flow, Feedback, and Continuous Experimentation — at the center of platform design to increase developer velocity and organizational reliability.

<CardGroup>
  <Card title="Watch Video" icon="video" cta="Learn more" href="https://learn.kodekloud.com/user/courses/certified-cloud-native-platform-engineering-associate-cnpa/module/2a91f7db-45c5-4944-a2b2-15da9f74f4d5/lesson/0a104e6e-85f9-410d-8452-7c7af49924a9" />
</CardGroup>


This documentation is built and hosted on [Mintlify](https://mintlify.com), a developer documentation platform.