-
A small custom device that turns a Raspberry Pi into a physical alarm panel for UPS events. It uses the UPS network mana...
(Page 15)
-
Resolving a server boot failure.
(Page 130)
-
Replacing a faulty power supply unit under load.
(Page 129)
-
Owning the root-cause-analysis lifecycle so incidents turn into durable lessons - on time, with quality, and with the record intact.
(Page 56)
-
Repairing a RAID controller battery failure.
(Page 127)
-
Replacing a failed server motherboard.
(Page 128)
-
The power strip inside the cabinet that feeds individual servers.
(Page 22)
-
Separating signal from noise in large covariance matrices using the Marchenko-Pastur distribution.
(Page 78)
-
A small, ARM-based Linux computer that is cheap enough to deploy by the dozen.
(Page 61)
-
Using cheap, Linux-capable single-board computers as monitoring, bridging, and control nodes in a data center.
(Page 60)
-
N, N+1, 2N and what "concurrently maintainable" really means.
(Page 40)
-
Detecting changes in market behavior, and why knowing the regime can matter more than knowing the forecast.
(Page 79)
-
Applying Weibull distributions, degradation models, and survival analysis to estimate when data center equipment will need replacement.
(Page 115)
-
A breaker panel that extends a PDU's circuits closer to the racks.
(Page 48)
-
Allocating capital by risk contribution rather than dollar weight, and why this avoids concentration in equities.
(Page 80)
-
A temperature sensor that looked wrong was actually right; the problem was the rack it was sitting in.
(Page 39)
16 items under this folder.