Too many power outages in Baker Lab, and other Chem buildings! |
Date |
Outage duration |
Cause |
Official link |
ChemIT notes |
|---|---|---|---|---|
2/27/2014 |
2 minutes |
Per 3/3/14 email: The outage was caused as a result of routine maintenance activities which were being conducted at the Campus' main substation which takes power from the NYSEG transmission system and provides it to campus. This work has been conducted many times before without incident but in this case caused a major disruption of electricity supply. Staff from Utilities and external technical resources are investigating the root cause of this unexpected event. |
No link to info in 3/3/14 email? |
Michael led our effort to initially evaluate and restore systems, with Oliver adding to documentation and to-do's. Lulu completed the cluster restoration efforts. |
1/27/2014 |
17-19 minutes |
? |
Lulu, Michael, and Oliver shut down headnodes and other systems which were on UPS. (Those systems non UPS shut down hard, per usual.) |
|
12/23/2013 |
2 minutes |
Human error? |
Terrible timing, right before the longest staff holiday of the year. |
|
7/17/13 |
Half a morning (~2 hours) |
|
|
Question: When power is initially restored, do you trust it? Or might it simply kick back off in some circumstances?
Assuming protection for 1-3 minutes MAXIMUM:
Do all headnodes and stand-alone computers in 248 Baker Lab
CCB Headnodes' UPS status:
Cluster |
Done |
Not done |
Notes |
|---|---|---|---|
Collum |
X |
|
|
Lancaster, with Crane (new) |
X |
|
Funded by Crane. |
Hoffmann |
X |
|
|
Scheraga |
X |
|
See below chart for s4 tand-alone computational computers |
Loring |
|
X |
Unique: Need to do ASAP |
Abruna |
|
X |
Unique: Need to do ASAP |
C4 Headnode: pilot |
X |
|
Provisioned on the margin, since still a pilot. |
Widom |
|
X |
See "C4", above |
Stand-alone computers' UPS status:
Computer |
Done |
Note done |
Notes |
|---|---|---|---|
Scheraga's 4 GPU rack-mounted computational computers |
|
X |
Need to protect? |
NMR web-based scheduler |
X |
|
|
Coates: MS SQL Server |
|
X |
Unique: Need to do ASAP |
Do all switches: Maybe ~$340 ($170*2), every ~4 years.
Do all compute nodes: ~$18K every ~4 years.
Compute node counts, for UPS pricing estimates. Does not include head node:
Cluster |
Compute node count |
Cost estimate |
Notes |
|---|---|---|---|
Collum |
8? |
|
Sprin |
Lancaster, with Crane (new) |
12? |
|
Funded by Crane. |
Hoffmann |
14? |
|
|
Scheraga |
92? |
|
See below chart for s4 tand-alone computational computers |
Loring |
6? |
X |
Unique: Need to do ASAP |
Abruna |
6? |
X |
Unique: Need to do ASAP |
C4 Headnode: pilot |
N/A |
|
This CCB Community headnode pilot is provisioned on the margin. |
Widom |
2 |
X |
See "C4", above |
Menu => List Machine / User Selections = > List Group or selected criteria
If nodes done show up, consider: