Neutron OVN Southbound DB Bloat & Compaction Debug Prompt
Diagnose a bloated or slow OVN Southbound/Northbound database in a Neutron OVN deployment — runaway size, slow ovsdb-server, chassis churn — and compact it safely without disrupting the dataplane.
- Target user
- OpenStack operators running Neutron with the OVN ML2 backend
- Difficulty
- Advanced
- Tools
- Claude, ChatGPT
The prompt
You are a senior OpenStack networking engineer who has rescued OVN control planes where the Southbound DB grew until ovsdb-server crawled and port binding stalled. I will provide: - OVN topology (number of chassis/compute nodes, NB/SB DB clustering — RAFT or active/standby) - DB sizing (`ovsdb-server` file sizes, memory use, `ovn-sbctl show` scale) - Symptoms (slow port binding, ovn-controller lag, high CPU on ovsdb-server, DB file growth) - Logs from ovsdb-server, ovn-northd, and ovn-controller Your job: 1. **Measure the bloat** — distinguish on-disk transaction-log growth (which compaction fixes) from genuine data-volume growth (too many ports/logical flows). Check the DB file size vs `ovsdb-tool db-version`/cluster status. 2. **Find the churn source** — identify stale Chassis/Port_Binding/MAC_Binding rows, look for flapping chassis re-registering, and check MAC_Binding table growth from ARP/ND learning that never ages out. 3. **Compact safely** — explain RAFT-cluster compaction (it compacts automatically, but `ovsdb-server/compact` can be triggered) versus standalone DBs; warn against `ovsdb-tool compact` on a running clustered DB and give the correct online method. 4. **Logical-flow scale** — check ovn-northd logical-flow counts and whether features like ACLs or distributed routing are multiplying flows; recommend `ovn-nbctl --print-wait-time` and northd timing to spot the bottleneck. 5. **Stale-data cleanup** — safely remove orphaned Chassis entries for decommissioned nodes and clear stale MAC_Binding rows, confirming nothing live references them first. 6. **Validate** — confirm DB size dropped, port-binding latency recovered, ovn-controller reconnected on all chassis, and no logical switch/router lost connectivity. Output as: (a) bloat root-cause statement (log growth vs data growth), (b) the safe compaction procedure for the actual DB topology, (c) stale-row cleanup commands with pre-checks, (d) a scale/flow-count assessment, (e) validation and rollback steps. Never run offline ovsdb-tool compact against a live clustered DB — use the online server command, and snapshot the DB first.
Run this prompt with AI
Test it, get an AI-improved version, or compare models — live in the Prompt Workspace. No copy-paste.
Related prompts
-
OVN Control Plane Deep Dive Prompt
Debug OVN control plane — Northbound/Southbound databases, ovn-northd, ovn-controller, logical flows, raft cluster health.
-
Neutron OVN Metadata Agent HA & Scaling Design Prompt
Design a highly available, horizontally scaled ovn-metadata-agent layer so instance cloud-init metadata requests never fail or hang, across a large OVN deployment with many chassis and high instance churn.
-
Neutron OVN BGP EVPN L3 Gateway Design Prompt
Design and validate an OVN-BGP-Agent EVPN/VRF deployment so tenant networks are advertised into a fabric without overlapping VLAN sprawl.
-
Neutron OVN northd Sync Lag Debug Prompt
Diagnose why logical resources in the OVN Northbound DB are not propagating to the Southbound DB or chassis, causing ports that never go ACTIVE or traffic that never programs.
More OpenStack prompts & error guides
Browse every OpenStack prompt and troubleshooting guide in one place.
Reading prompts? Get all 500 in one free PDF
500 battle-tested, copy-paste AI prompts engineered by a senior systems engineer — every one with fill-in placeholders and safety/back-out notes. Drop your email and it's yours.
- 500 prompts: Linux · Kubernetes · Terraform · OpenStack · GitLab · Docker · Monitoring · Incident Response
- Instant PDF download — yours free, forever
- Plus one practical AI-workflow email a week (no spam)
Single opt-in · unsubscribe anytime · no spam.