Skip to content
New releaseNOS 1.2 · Aurora is now shipping — conversational NOC in 24 languages.HardwareCluster-aware installer for Core X100 available today.AppsField Engineer & Subscriber apps re-published to iOS and Android stores.AI EnginePer-tenant AI budget controls now live for multi-tenant operators.PerformanceRADIUS backend picks up a 30% throughput lift over 1.1 LTS.
NetXol
AI Engine

AI auto-fix playbook for stuck ACS workers

5
LLea Marinoasked · Jun 30, 2026, 03:28 PM
Sharing a playbook we've been running for six weeks. Symptom: ACS worker processes on our X20 occasionally stall after ONT firmware pushes — usually 2-3 workers out of 32. Manual fix was systemctl restart nos-acs@N. We wrote a playbook in NOS that watches worker health, restarts the affected process, and logs the event to the ticket. AI catches it in about 40 seconds. Down from a 20-minute pager cycle. Happy to share the playbook YAML if there's interest.
#ai#auto-fix#acs#playbook
4 replies·2842 views
0
AAli Bashirreplied · Jul 1, 2026, 09:22 AM
This would be amazing — please share the YAML. We've been fighting the same ACS-worker stalls for months.
3
LLea Marinoreplied · Jul 1, 2026, 02:41 PM
Sharing here for anyone who wants it. Save as `~/.nos/playbooks/acs-worker-heal.yaml`, restart the AI engine, and enable via Settings → AI → Playbooks. Note the safe-fallback timeout is 60s — I set it deliberately conservative.
0
AAndré Silvareplied · Jul 8, 2026, 07:52 AM
Deployed this on our X5 yesterday, caught two stalled workers overnight, filed the ticket, restarted, closed. Slept through it. Thank you.
0
NNabil Zaidireplied · Jul 15, 2026, 11:33 AM
Might be worth folding this into the default playbook library on the next release — it's a common enough failure mode.
Sign in to reply. New here? Create an account or sign in.