* [BUG] Lockup using ALi SATA controller (sata_uli)
@ 2005-03-21 22:44 Markus Dahms
2005-03-23 6:26 ` Andrew Morton
0 siblings, 1 reply; 3+ messages in thread
From: Markus Dahms @ 2005-03-21 22:44 UTC (permalink / raw)
To: linux-kernel
Hi folks,
I have a reproducable lockup of my system using an ALi SATA controller
and writing some 100 MB to the attached disk.
kernel: 2.6.11 SMP (w/ and w/o preemption tested)
system: 2 x Pentium III on Asus CUV266-D
controller: ALi M5283 based PCI card (1xPATA, 2xSATA)
disk: 160GB Hitachi
The Linux system is on another disk (PATA, onboard controller).
The same controller with the same disk in another system (AMD) show
the same errors (lockup, BadCRC as seen below), the disk with another
controller (SII) works.
Thanks to netconsole I could get the following "last words":
| ata2: command 0x35 timeout, stat 0xd0 host_stat 0x81
| ata2: status=0xd0 { Busy }
| SCSI error : <1 0 0 0> return code = 0x8000002
| sda: Current: sense key: Aborted Command
| Additional sense: Scsi parity error
| end_request: I/O error, dev sda, sector 2087934
| Buffer I/O error on device sda5, logical block 260976
| lost page write due to I/O error on sda5
| ATA: abnormal status 0xD0 on port 0x9007
| ATA: abnormal status 0xD0 on port 0x9007
| ATA: abnormal status 0xD0 on port 0x9007
After a hard reset the following lines appear 4 to 10 times in the
boot messages and the filesystem recovery (XFS in my case) fails.
A second attempt to mount the filesystem succeeds.
| ata2: status=0x51 { DriveReady SeekComplete Error }
| ata2: error=0x84 { DriveStatusError BadCRC }
The relevant lines from the boot log:
| libata version 1.10 loaded.
| ACPI: PCI interrupt 0000:00:0e.0[A] -> GSI 17 (level, low) -> IRQ 17
| ata1: SATA max UDMA/133 cmd 0x9800 ctl 0x9402 bmdma 0x8400 irq 17
| ata2: SATA max UDMA/133 cmd 0x9000 ctl 0x8802 bmdma 0x8408 irq 17
| ata1: no device found (phy stat 00000000)
| scsi0 : sata_uli
| ata2: dev 0 cfg 49:2f00 82:74eb 83:7fea 84:4023 85:74e8 86:3c02 87:4023 88:003f
| ata2: dev 0 ATA, max UDMA/100, 321672960 sectors: lba48
| ata2: dev 0 configured for UDMA/100
| scsi1 : sata_uli
| Vendor: ATA Model: HDS722516VLSA80 Rev: V34O
| Type: Direct-Access ANSI SCSI revision: 05
| SCSI device sda: 321672960 512-byte hdwr sectors (164697 MB)
| SCSI device sda: drive cache: write back
| SCSI device sda: 321672960 512-byte hdwr sectors (164697 MB)
| SCSI device sda: drive cache: write back
| sda: sda1 < sda5 >
| Attached scsi disk sda at scsi1, channel 0, id 0, lun 0
Do you have some hints?
Greetings,
Markus
--
Do bl Sp ce is a v ry saf me hod of driv compr s ion
^ permalink raw reply [flat|nested] 3+ messages in thread
* Re: [BUG] Lockup using ALi SATA controller (sata_uli)
2005-03-21 22:44 [BUG] Lockup using ALi SATA controller (sata_uli) Markus Dahms
@ 2005-03-23 6:26 ` Andrew Morton
2005-03-24 17:00 ` Markus Dahms
0 siblings, 1 reply; 3+ messages in thread
From: Andrew Morton @ 2005-03-23 6:26 UTC (permalink / raw)
To: Markus Dahms; +Cc: linux-kernel
Markus Dahms <dahms@fh-brandenburg.de> wrote:
>
> I have a reproducable lockup of my system using an ALi SATA controller
> and writing some 100 MB to the attached disk.
>
> ...
> Do you have some hints?
As a test you might like to try an uniprocessor kernel - we do have a
deadlock on the sata error recovery paths at present.
Or ensure that you've enabled the io-apci in Kconfig and boot with
nmi_watchdog=1.
^ permalink raw reply [flat|nested] 3+ messages in thread
* Re: [BUG] Lockup using ALi SATA controller (sata_uli)
2005-03-23 6:26 ` Andrew Morton
@ 2005-03-24 17:00 ` Markus Dahms
0 siblings, 0 replies; 3+ messages in thread
From: Markus Dahms @ 2005-03-24 17:00 UTC (permalink / raw)
To: Andrew Morton; +Cc: linux-kernel
Hello Andrew Morton,
>> I have a reproducable lockup of my system using an ALi SATA controller
>> and writing some 100 MB to the attached disk.
>> ...
>> Do you have some hints?
> As a test you might like to try an uniprocessor kernel - we do have a
> deadlock on the sata error recovery paths at present.
The UP kernel didn't change anything, same error on console. In addition
this kernel crashes every second time or so immediately after boot
(some pci stuff in backtrace), I'll try catch some error messages....
> Or ensure that you've enabled the io-apci in Kconfig and boot with
> nmi_watchdog=1.
Except the line "testing NMI watchdog ... OK." in the boot messages
I don't see any difference. :-(
thanks anyway,
Markus
--
A CRAY is the only computer that runs an endless loop in just 4 hours...
^ permalink raw reply [flat|nested] 3+ messages in thread
end of thread, other threads:[~2005-03-24 17:00 UTC | newest]
Thread overview: 3+ messages (download: mbox.gz / follow: Atom feed)
-- links below jump to the message on this page --
2005-03-21 22:44 [BUG] Lockup using ALi SATA controller (sata_uli) Markus Dahms
2005-03-23 6:26 ` Andrew Morton
2005-03-24 17:00 ` Markus Dahms
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox
all inboxes | Powered by JetHome®