<?xml version="1.0" encoding="UTF-8"?>
<rss xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:rdf="http://www.w3.org/1999/02/22-rdf-syntax-ns#" xmlns:taxo="http://purl.org/rss/1.0/modules/taxonomy/" version="2.0">
  <channel>
    <title>topic Problem with HP Cluster in Windows Server 2003</title>
    <link>https://community.hpe.com/t5/windows-server-2003/problem-with-hp-cluster/m-p/3426720#M1297</link>
    <description>Hi,&lt;BR /&gt;&lt;BR /&gt;Our Hp Cluster failed this week for no apparent reason.  We’re running Windows 2003 Enterprise with Microsoft clustering on two HP DL740’s using HP Secure Path back to a MSA1000.&lt;BR /&gt;&lt;BR /&gt;According to the cluster log the quorum drive went off line and was inaccessible for approximately 9 minutes.  Then several other drives on the SAN went offline as well.&lt;BR /&gt;&lt;BR /&gt;Whatever happened caused the SCSI buses not to close down properly and our SQL database was corrupted.&lt;BR /&gt;&lt;BR /&gt;The Quorum drive consists of 4 36.4gb 10000rpm hard drives.  These 4 drives are in a RAID 1 setup and then striped.  All disks are spilt over different controllers.&lt;BR /&gt;&lt;BR /&gt;There is no way that this should have failed unless the entire SAN failed (unlikely).  As the other drives went off line over a period of 9 minutes this doesn’t make sense either.&lt;BR /&gt;&lt;BR /&gt;As all the drives, SAN and controllers appear to be ok and the cluster is now performing correctly I have no idea of how to diagnose what happened.  &lt;BR /&gt;&lt;BR /&gt;Any advice or comments would be apprecia</description>
    <pubDate>Sat, 20 Nov 2004 07:35:24 GMT</pubDate>
    <dc:creator>Stephen McCann</dc:creator>
    <dc:date>2004-11-20T07:35:24Z</dc:date>
    <item>
      <title>Problem with HP Cluster</title>
      <link>https://community.hpe.com/t5/windows-server-2003/problem-with-hp-cluster/m-p/3426720#M1297</link>
      <description>Hi,&lt;BR /&gt;&lt;BR /&gt;Our Hp Cluster failed this week for no apparent reason.  We’re running Windows 2003 Enterprise with Microsoft clustering on two HP DL740’s using HP Secure Path back to a MSA1000.&lt;BR /&gt;&lt;BR /&gt;According to the cluster log the quorum drive went off line and was inaccessible for approximately 9 minutes.  Then several other drives on the SAN went offline as well.&lt;BR /&gt;&lt;BR /&gt;Whatever happened caused the SCSI buses not to close down properly and our SQL database was corrupted.&lt;BR /&gt;&lt;BR /&gt;The Quorum drive consists of 4 36.4gb 10000rpm hard drives.  These 4 drives are in a RAID 1 setup and then striped.  All disks are spilt over different controllers.&lt;BR /&gt;&lt;BR /&gt;There is no way that this should have failed unless the entire SAN failed (unlikely).  As the other drives went off line over a period of 9 minutes this doesn’t make sense either.&lt;BR /&gt;&lt;BR /&gt;As all the drives, SAN and controllers appear to be ok and the cluster is now performing correctly I have no idea of how to diagnose what happened.  &lt;BR /&gt;&lt;BR /&gt;Any advice or comments would be apprecia</description>
      <pubDate>Sat, 20 Nov 2004 07:35:24 GMT</pubDate>
      <guid>https://community.hpe.com/t5/windows-server-2003/problem-with-hp-cluster/m-p/3426720#M1297</guid>
      <dc:creator>Stephen McCann</dc:creator>
      <dc:date>2004-11-20T07:35:24Z</dc:date>
    </item>
    <item>
      <title>Re: Problem with HP Cluster</title>
      <link>https://community.hpe.com/t5/windows-server-2003/problem-with-hp-cluster/m-p/3426721#M1298</link>
      <description>the best way to fix something like that is to eliminate the impossibilities and fix whatever remains. scientific method. find ways to disprove possibilities based on the evidence. &lt;BR /&gt;&lt;BR /&gt;if all the LUNS when of at once, then it indicates the problem can't be based in a single LUN. we're left with a problem with the server, fibre channel NIC(s), RAID manager, switches, or fabric. &lt;BR /&gt;&lt;BR /&gt;multiple servers had the same problem? that eliminates a software or server based problem. &lt;BR /&gt;&lt;BR /&gt;redundant fabrics? then it can't be a switch problem, or a lost connection. &lt;BR /&gt;&lt;BR /&gt;redundant array managers? then it can't be the control board in the disk array, but it still could be the backplane. &lt;BR /&gt;&lt;BR /&gt;check the SAN switches for module disconnects or power events. possibility of a "cleaning crew unplugged it" problem? &lt;BR /&gt;&lt;BR /&gt;a word of caution, don't eliminate anything unless the evidence shows it isn't possible. don't eliminate that new shiny GBIC because it was just installed last week.  &lt;BR /&gt;</description>
      <pubDate>Mon, 22 Nov 2004 10:25:06 GMT</pubDate>
      <guid>https://community.hpe.com/t5/windows-server-2003/problem-with-hp-cluster/m-p/3426721#M1298</guid>
      <dc:creator>Thomas Bianco</dc:creator>
      <dc:date>2004-11-22T10:25:06Z</dc:date>
    </item>
    <item>
      <title>Re: Problem with HP Cluster</title>
      <link>https://community.hpe.com/t5/windows-server-2003/problem-with-hp-cluster/m-p/3426722#M1299</link>
      <description>Stephen&lt;BR /&gt;   why do you have such behemoth quorum disk? Where is your data? Please post information about your disk and cluster groups.&lt;BR /&gt;&lt;BR /&gt;Have you reviewed the  System and Application log? Read them and find references to Securepath events at the time of the problems.&lt;BR /&gt;&lt;BR /&gt;How do you know "the SCSI buses didn't close down properly" ?&lt;BR /&gt;&lt;BR /&gt;Regards</description>
      <pubDate>Mon, 22 Nov 2004 14:09:17 GMT</pubDate>
      <guid>https://community.hpe.com/t5/windows-server-2003/problem-with-hp-cluster/m-p/3426722#M1299</guid>
      <dc:creator>Richard Perez</dc:creator>
      <dc:date>2004-11-22T14:09:17Z</dc:date>
    </item>
  </channel>
</rss>

