Hello @D-Roc,
As promised, we’ve completed a full review of the diagnostic bundle from your ROCK. It covers June 27 through July 1 - so it predates your July 4-5 episodes - but the pattern across it is remarkably consistent, and it answers most of the questions you asked. Here’s what we found.
What the logs show. Every single Roon Server session in the bundle - 19 of them - ends the same way: the log stops mid-write on a routine background line, with no shutdown message and no error. There are zero application crashes anywhere, even though crash reporting is active in Roon Server. Separately, your ROCK’s system log shows the external Seagate backup drive coming up with a “Volume was not properly unmounted” warning - something that only happens when the entire machine goes down without a chance to flush its filesystems.
To answer your direct question - does Roon Server crash, or does the network go first: neither, in the sense you mean. Roon Server writes its log to the local disk, which needs no network at all - if only the network dropped, the log would keep going for hours. It doesn’t. It stops mid-sentence, which is the signature of the whole machine stopping: a power loss or a hard freeze.
The most important detail: the restarts in the logs come in two rhythms. Some are long gaps of 6-24 hours, which match your description of finding the box dead and pressing the button. But others are clusters of restarts spaced 35-90 seconds apart - far too fast to be someone walking over and pressing the front button. During those bursts, the machine was power-cycling on its own.
What this rules out. It’s not a software crash, it’s not the boot drive (no storage errors anywhere, including after your M.2 swap), and it’s not the NIC or DHCP - the captured boot shows the network coming up cleanly, with an IP lease obtained in about 4 seconds. It also can’t be fully explained by the router IGMP/multicast theory: we do see brief LAN blips in the logs (your streamer and a Windows machine dropping and reconnecting), but they don’t line up with the moments the logs die. Discovery problems can hide a healthy server from your remotes - they cannot stop a machine from writing to its own disk.
About the HDMI screen: a frozen frame looks identical to a live one at a glance, so the screen “staying up” unfortunately doesn’t prove the system was still running. A simple trick for next time: keep a USB keyboard plugged into the NUC, and when the ROCK becomes unreachable, tap the Num Lock or Caps Lock key. If the keyboard LED doesn’t toggle, the machine is frozen solid - and that’s our answer. Or you can click enter and type something in the console like ‘resetnetwork’ and check if the ROCK is still accessible
So the evidence points consistently at power delivery or a hardware-level freeze on the NUC itself. With that in mind, here’s the refined plan - please take it one step at a time:
Step 1 - swap the power adapter. This is now the single most valuable test, and it’s inexpensive. A marginal power brick is the classic cause of exactly this signature: clean logs, no errors, machine just stops.
Step 2 - change the power path. Plug the NUC into a different wall outlet, ideally on a different circuit, and bypass any power strip, surge protector, or UPS it currently goes through.
Step 3 - one final clean reinstall, done differently. If the failures continue even on the new power setup, reinstall Roon OS fresh one more time - but this time, run it for a few day or three without restoring your backup. Every reinstall so far has ended with your archive being restored on top, so a truly clean baseline has never actually been tested. If a bare install survives where the restored one didn’t, that tells us a lot; if it still fails, hardware is confirmed beyond reasonable doubt.
Step 4 - keep the capture discipline. Beyond the steps above, freeze the setup (no other changes), use a short press of the power button for clean shutdowns - never hold it - and when the next failure happens, note the exact time, try the keyboard LED test, and post the results here. The current bundle ends July 1, so a fresh set of diagnostics after the next captured failure will confirm whether the pattern held through your recent episodes.
We know this has been a long road, but the picture is genuinely narrowing: healthy software, healthy drive, healthy network stack - and a machine that keeps stopping mid-write. Let’s catch it in the act. We’ll stay on this with you until it’s stable.