Changes between Version 20 and Version 21 of Production_Cluster_Status
- Timestamp:
- Feb 18, 2010, 10:42:48 AM (16 years ago)
Legend:
- Unmodified
- Added
- Removed
- Modified
-
Production_Cluster_Status
v20 v21 7 7 === Nodes down === 8 8 9 None. 9 * ipp008 : motherboard problems -- CPUs were changed on 2010.02.17, but still no boot 10 * ipp018 : motherboard problems? power supply problems? 11 * ipp037 : long history of return, etc. 10 12 11 13 === Notes === 12 14 13 * ipp008 raid is being completely rebuilt 14 * ipp005 is rebuilding disk !#16 15 15 * ipp005 - The memory was completely swapped and crashes continued. CPUs were replaced (2010.02.09). In production since 2010.02.16. 16 16 * ipp014 - using the chassis from ipp017 with disks from ipp014 (chassis had been ipp037, but was sent back to ASA for repair, ASA shipped back with replacement motherboard (Tyan S2912G2NR-E) Tyan S2912G2NR is EOL - kernel 2.6.31.5 is required.) 17 17 * ipp017 - using the chassis from ipp037 with disks from ipp017 … … 21 21 * ipp037 - using the chassis from ipp014 with disks from ipp037 22 22 23 * ippdb00 - '''IPP needs to stress-test this machine'''23 * ippdb00 - in production as nebulous server 24 24 * ippdb02 - memtest86 4 passes completed no errors. (2009-12-30) 25 25 26 26 === Known issues === 27 27 28 * ipp005 - occasional crashes -- we suspect the memory may be bad. '''note: as of 2009.10.30, the memory has been replaced; we need to stress this machine and see if it continues to crash'''29 28 * ipp008 - complete raid failure 2009.10.28 -- disks have been replaced and raid needs to be re-built. ipp008 has not been used for storage for some time, so we have not lost any vital data. 30 29 * ipp009 - several recent crashes with coincident complaints from CPU !#3. '''note: we will swap CPU 1 and 3 and then stress-test this machine to see if the effect moves with the CPU'''
