07-31-2015 04:07 AM
Hi,
I have a new Linux machine with two DC S3610 1.6TB SSDs. It's Debian jessie so kernel 3.6.17. Since around one month after installation these errors started appearing:
Jul 30 16:30:59 snaps kernel: [186914.249429] ata1.00: exception Emask 0x0 SAct 0x3 SErr 0x0 action 0x6 frozen
Jul 30 16:30:59 snaps kernel: [186914.250465] ata1.00: failed command: WRITE FPDMA QUEUED
Jul 30 16:30:59 snaps kernel: [186914.251505] ata1.00: cmd 61/08:00:39:db:8e/00:00:09:00:00/40 tag 0 ncq 4096 out
Jul 30 16:30:59 snaps kernel: [186914.251505] res 40/00:01:00:00:00/00:00:00:00:00/00 Emask 0x4 (timeout)
Jul 30 16:30:59 snaps kernel: [186914.253613] ata1.00: status: { DRDY }
Jul 30 16:30:59 snaps kernel: [186914.254781] ata1.00: failed command: WRITE FPDMA QUEUED
Jul 30 16:30:59 snaps kernel: [186914.255810] ata1.00: cmd 61/08:08:71:fc:4e/00:00:66:00:00/40 tag 1 ncq 4096 out
Jul 30 16:30:59 snaps kernel: [186914.255810] res 40/00:01:00:00:00/00:00:00:00:00/00 Emask 0x4 (timeout)
Jul 30 16:30:59 snaps kernel: [186914.257940] ata1.00: status: { DRDY }
Jul 30 16:30:59 snaps kernel: [186914.259086] ata1: hard resetting link
Jul 30 16:31:00 snaps kernel: [186914.577366] ata1: SATA link up 6.0 Gbps (SStatus 133 SControl 300)
Jul 30 16:31:00 snaps kernel: [186914.578307] ata1.00: configured for UDMA/133
Jul 30 16:31:00 snaps kernel: [186914.578310] ata1.00: device reported invalid CHS sector 0
Jul 30 16:31:00 snaps kernel: [186914.578311] ata1.00: device reported invalid CHS sector 0
Jul 30 16:31:00 snaps kernel: [186914.578316] ata1: EH complete
The error is always the same, and the only thing on ata1.00 is one of the SSDs. I switched the two SSDs around and the problem followed the same SSD.
I can't force the error to happen on demand, it just seems to happen every other day or so, though not at the same time of day. All IO is held up briefly while the link is reset. The drive passes a SMART long self-test.
So is this drive faulty? If not, what can I try to fix this? If so, is there an easy way to prove it for RMA purposes?
Jul 27 05:59:30 snaps kernel: [ 33.054376] ata1.00: ATA-9: INTEL SSDSC2BX016T4, G2010110, max UDMA/133
Jul 27 05:59:30 snaps kernel: [ 33.054474] ata1.00: 3125627568 sectors, multi 1: LBA48 NCQ (depth 31/32)
Jul 27 05:59:30 snaps kernel: [ 33.054567] ata2.00: ATA-9: INTEL SSDSC2BX016T4, G2010110, max UDMA/133
Jul 27 05:59:30 snaps kernel: [ 33.054657] ata2.00: 3125627568 sectors, multi 1: LBA48 NCQ (depth 31/32)
$ sudo smartctl -i /dev/sda
smartctl 6.4 2014-10-07 r4002 [x86_64-linux-3.16.0-4-amd64] (local build)
Copyright (C) 2002-14, Bruce Allen, Christian Franke, www.smartmontools.org
=== START OF INFORMATION SECTION ===
Device Model: INTEL SSDSC2BX016T4
Serial Number: BTHC511604V41P6PGN
LU WWN Device Id: 5 5cd2e4 04b7b1bfa
Firmware Version: G2010110
User Capacity: 1,600,321,314,816 bytes [1.60 TB]
Sector Sizes: 512 bytes logical, 4096 bytes physical
Rotation Rate: Solid State Device
Form Factor: 2.5 inches
Device is: Not in smartctl database [for details use: -P showall]
ATA Version is: ACS-2 T13/2015-D revision 3
SATA Version is: SATA 2.6, 6.0 Gb/s (current: 6.0 Gb/s)
Local Time is: Fri Jul 31 11:04:09 2015 UTC
SMART support is: Available - device has SMART capability.
SMART support is: Enabled
$ sudo smartctl -i /dev/sdb
smartctl 6.4 2014-10-07 r4002 [x86_64-linux-3.16.0-4-amd64] (local build)
Copyright (C) 2002-14, Bruce Allen, Christian Franke, www.smartmontools.org
=== START OF INFORMATION SECTION ===
Device Model: INTEL SSDSC2BX016T4
Serial Number: BTHC511604SD1P6PGN
LU WWN Device Id: 5 5cd2e4 04b7b1ba2
Firmware Version: G2010110
User Capacity: 1,600,321,314,816 bytes [1.60 TB]
Sector Sizes: 512 bytes logical, 4096 bytes physical
Rotation Rate: Solid State Device
Form Factor: 2.5 inches
Device is: Not in smartctl database [for details use: -P showall]
ATA Version is: ACS-2 T13/2015-D revision 3
SATA Version is: SATA 2.6, 6.0 Gb/s (current: 6.0 Gb/s)
Local Time is: Fri Jul 31 11:04:35 2015 UTC
SMART support is: Available - device has SMART capability.
SMART support is: Enabled
Message was edited by: Andy Smith Now seeing same problems with other SSD, so this is not restricted to a single drive.
08-24-2015 05:17 PM
Hello Jmsr,
As you mentioned, this change is not listed in the Release Notes. However, we have confirmed that the new firmware contained in Intel® Solid-State Drive Toolbox version 3.3.1 includes a fix for the condition reported in this thread.
Also, we would like to inform that the new firmware will also be integrated into the upcoming release of Intel® SSD Data Center Tool. We expect the new version to be available soon.
08-25-2015 06:35 AM
Hi Jonathan,
Thanks. Can you please give us full details on what's known regarding this 'condition', i.e. at least what models are affected and what circumstances are required for this 'condition' to be triggered?
Having suffered a 100% failure rate so far, we've a number of orders / builds which are on hold pending further information and either a confirmed resolution or workaround.
I'm wondering if there's another way to avoid it in the meantime while waiting for a firmware update that doesn't require windows?
Regards
James
08-26-2015 01:02 AM
This thread is very interesting. I've been reporting a problem affecting both of the Intel DC S3610 drives that I have where they are occasionally dropped by my RAID controller. I suspect its a drive firmware problem (and could easily be the same one) but so far customer support have not acknowledged a problem with the drives and have basically just said that they can't guarantee compatibility with any specific motherboard. Curiously, they haven't mentioned that since I first reported the problem this new firmware has become available. I really hope it fixes my problem though.
08-27-2015 03:18 PM
I am surprised to not see it mentioned here, but today my support case was updated to say that a new version of ISDCT is available, and that it contains the firmware update which fixes this issue.
https://downloadcenter.intel.com/download/23931/Intel-Solid-State-Drive- Intel® Download Center
I haven't yet had chance to test it but I expect to do so within the next 24 hours. It will then take several days for me to be confident that things are improved.
I'd be interested in knowing others' experiences.
Cheers,
Andy
08-27-2015 04:16 PM
Hello,
As you mentioned, the https://downloadcenter.intel.com/download/23931/Intel-Solid-State-Drive-Data-Center-Tool Intel® Solid-State Drive Data Center Tool, version 2.2.4, was released for download today. Please let us know if the issues are resolved with the update, or what is the behavior with the new firmware.