Troubleshoot Amazon EC2 Linux instances with failed status checks
The following information can help you troubleshoot issues if your Linux instance fails a status check. First determine whether your applications are exhibiting any problems. If you verify that the instance is not running your applications as expected, review the status check information and the system logs.
For examples of problems that can cause status checks to fail, see Status checks for Amazon EC2 instances.
Contents
I/O ERROR: neither local nor remote disk (Broken distributed block device)
request_module: runaway loop modprobe (Looping legacy kernel modprobe on older Linux versions)
"FATAL: Could not load /lib/modules" or "BusyBox" (Missing kernel modules)
fsck: No such file or directory while trying to open... (File system not found)
VFS: Unable to mount root fs on unknown-block (Root filesystem mismatch)
Error: Unable to determine major/minor number of root device... (Root file system/device mismatch)
... days without being checked, check forced (File system check required)
Unable to load SELinux Policy. Machine is in enforcing mode. Halting now. (SELinux misconfiguration)
Review status check information
To investigate impaired instances using the Amazon EC2 console
Open the Amazon EC2 console at https://console.aws.amazon.com/ec2/
. -
In the navigation pane, choose Instances, and then select your instance.
-
Select the Status and alarms tab to see the individual results for all System status checks, Instance status checks, and Attached EBS status checks.
If a status check has failed, you can try one of the following options:
-
Create an alarm to recover the instance in response to the failed status check. For more information, see Create alarms that stop, terminate, reboot, or recover an instance.
-
(Instance status checks) If you changed the instance type to a Nitro-based instance, status checks fail if you migrated from an instance that does not have the required ENA and NVMe drivers. For more information, see Compatibility for changing the instance type.
-
For an instance with an EBS root volume, stop and restart the instance. For more information, see Stop and start Amazon EC2 instances.
-
For an instance with an instance store root volume, terminate the instance and launch a replacement instance. For more information, see Terminate Amazon EC2 instances.
-
Wait for Amazon EC2 to resolve the issue.
-
Contact Support or post your issue to AWS re:Post
. -
If your instance is in an Auto Scaling group:
-
(System status checks and instance status checks) By default, Amazon EC2 Auto Scaling automatically launches a replacement instance. For more information, see Health checks for instances in an Auto Scaling group in the Amazon EC2 Auto Scaling User Guide.
-
(Attached EBS status checks) You must configure Amazon EC2 Auto Scaling to automatically launch a replacement instance. For more information, see Monitor and replace Auto Scaling instances with impaired Amazon EBS volumes in the Amazon EC2 Auto Scaling User Guide.
-
-
Retrieve the system log and look for errors. For more information, see Retrieve the system logs.
Retrieve the system logs
If an instance status check fails, you can reboot the instance and retrieve the system logs. The logs may reveal an error that can help you troubleshoot the issue. Rebooting clears unnecessary information from the logs.
To reboot an instance and retrieve the system log
Open the Amazon EC2 console at https://console.aws.amazon.com/ec2/
. -
In the navigation pane, choose Instances, and select your instance.
-
Choose Instance state, Reboot instance. It might take a few minutes for your instance to reboot.
-
Verify that the problem still exists; in some cases, rebooting may resolve the problem.
-
When the instance is in the
runningstate, choose Actions, Monitor and troubleshoot, Get system log. -
Review the log that appears on the screen, and use the list of known system log error statements below to troubleshoot your issue.
-
If your issue is not resolved, you can post your issue to AWS re:Post
.
Troubleshoot system log errors for Linux instances
For Linux instances that have failed an instance status check, such as the instance reachability check, verify that you followed the steps above to retrieve the system log. The following list contains some common system log errors and suggested actions you can take to resolve the issue for each error.
Memory Errors
Device Errors
Kernel Errors
File System Errors
-
fsck: No such file or directory while trying to open... (File system not found)
-
VFS: Unable to mount root fs on unknown-block (Root filesystem mismatch)
-
Error: Unable to determine major/minor number of root device... (Root file system/device mismatch)
-
... days without being checked, check forced (File system check required)
Operating System Errors
Out of memory: kill process
An out-of-memory error is indicated by a system log entry similar to the one shown below.
[115879.769795] Out of memory: kill process 20273 (httpd) score 1285879
or a child
[115879.769795] Killed process 1917 (php-cgi) vsz:467184kB, anon-
rss:101196kB, file-rss:204kB
Potential cause
Exhausted memory
Suggested actions
| For this instance type | Do this |
|---|---|
|
Amazon EBS-backed |
Do one of the following:
|
|
Instance store-backed |
Do one of the following:
|
ERROR: mmu_update failed (Memory management update failed)
Memory management update failures are indicated by a system log entry similar to the following:
...
Press `ESC' to enter the menu... 0 [H[J Booting 'Amazon Linux 2011.09 (2.6.35.14-95.38.amzn1.i686)'
root (hd0)
Filesystem type is ext2fs, using whole disk
kernel /boot/vmlinuz-2.6.35.14-95.38.amzn1.i686 root=LABEL=/ console=hvc0 LANG=
en_US.UTF-8 KEYTABLE=us
initrd /boot/initramfs-2.6.35.14-95.38.amzn1.i686.img
ERROR: mmu_update failed with rc=-22
Potential cause
Issue with Amazon Linux
Suggested action
Post your issue to AWS re:Post
I/O error (block device failure)
An input/output error is indicated by a system log entry similar to the following example:
[9943662.053217] end_request: I/O error, dev sde, sector 52428288
[9943664.191262] end_request: I/O error, dev sde, sector 52428168
[9943664.191285] Buffer I/O error on device md0, logical block 209713024
[9943664.191297] Buffer I/O error on device md0, logical block 209713025
[9943664.191304] Buffer I/O error on device md0, logical block 209713026
[9943664.191310] Buffer I/O error on device md0, logical block 209713027
[9943664.191317] Buffer I/O error on device md0, logical block 209713028
[9943664.191324] Buffer I/O error on device md0, logical block 209713029
[9943664.191332] Buffer I/O error on device md0, logical block 209713030
[9943664.191339] Buffer I/O error on device md0, logical block 209713031
[9943664.191581] end_request: I/O error, dev sde, sector 52428280
[9943664.191590] Buffer I/O error on device md0, logical block 209713136
[9943664.191597] Buffer I/O error on device md0, logical block 209713137
[9943664.191767] end_request: I/O error, dev sde, sector 52428288
[9943664.191970] end_request: I/O error, dev sde, sector 52428288
[9943664.192143] end_request: I/O error, dev sde, sector 52428288
[9943664.192949] end_request: I/O error, dev sde, sector 52428288
[9943664.193112] end_request: I/O error, dev sde, sector 52428288
[9943664.193266] end_request: I/O error, dev sde, sector 52428288
...
Potential causes
| Instance type | Potential cause |
|---|---|
|
Amazon EBS-backed |
A failed Amazon EBS volume |
|
Instance store-backed |
A failed physical drive |
Suggested actions
| For this instance type | Do this |
|---|---|
|
Amazon EBS-backed |
Use the following procedure:
|
|
Instance store-backed |
Terminate the instance and launch a new instance. NoteData cannot be recovered. Recover from backups. NoteIt's a good practice to use either Amazon S3 or Amazon EBS for backups. Instance store volumes are directly tied to single host and single disk failures. |
I/O ERROR: neither local nor remote disk (Broken distributed block device)
An input/output error on the device is indicated by a system log entry similar to the following example:
...
block drbd1: Local IO failed in request_timer_fn. Detaching...
Aborting journal on device drbd1-8.
block drbd1: IO ERROR: neither local nor remote disk
Buffer I/O error on device drbd1, logical block 557056
lost page write due to I/O error on drbd1
JBD2: I/O error detected when updating journal superblock for drbd1-8.
Potential causes
| Instance type | Potential cause |
|---|---|
|
Amazon EBS-backed |
A failed Amazon EBS volume |
|
Instance store-backed |
A failed physical drive |
Suggested action
Terminate the instance and launch a new instance.
For an Amazon EBS-backed instance you can recover data from a recent snapshot by creating an image from it. Any data added after the snapshot cannot be recovered.
request_module: runaway loop modprobe (Looping legacy kernel modprobe on older Linux versions)
This condition is indicated by a system log similar to the one shown below. Using an unstable or old Linux kernel (for example, 2.6.16-xenU) can cause an interminable loop condition at startup.
Linux version 2.6.16-xenU (builder@xenbat.amazonsa) (gcc version 4.0.1
20050727 (Red Hat 4.0.1-5)) #1 SMP Mon May 28 03:41:49 SAST 2007
BIOS-provided physical RAM map:
Xen: 0000000000000000 - 0000000026700000 (usable)
0MB HIGHMEM available.
...
request_module: runaway loop modprobe binfmt-464c
request_module: runaway loop modprobe binfmt-464c
request_module: runaway loop modprobe binfmt-464c
request_module: runaway loop modprobe binfmt-464c
request_module: runaway loop modprobe binfmt-464c
Suggested actions
| For this instance type | Do this |
|---|---|
|
Amazon EBS-backed |
Use a newer kernel, either GRUB-based or static, using one of the following options: Option 1: Terminate the instance and launch a new instance, specifying the
Option 2:
|
|
Instance store-backed |
Terminate the instance and launch a new instance, specifying the |
"FATAL: kernel too old" and "fsck: No such file or directory while trying to open /dev" (Kernel and AMI mismatch)
This condition is indicated by a system log similar to the one shown below.
Linux version 2.6.16.33-xenU (root@dom0-0-50-45-1-a4-ee.z-2.aes0.internal)
(gcc version 4.1.1 20070105 (Red Hat 4.1.1-52)) #2 SMP Wed Aug 15 17:27:36 SAST 2007
...
FATAL: kernel too old
Kernel panic - not syncing: Attempted to kill init!
Potential causes
Incompatible kernel and userland
Suggested actions
| For this instance type | Do this |
|---|---|
|
Amazon EBS-backed |
Use the following procedure:
|
|
Instance store-backed |
Use the following procedure:
|
"FATAL: Could not load /lib/modules" or "BusyBox" (Missing kernel modules)
This condition is indicated by a system log similar to the one shown below.
[ 0.370415] Freeing unused kernel memory: 1716k freed
Loading, please wait...
WARNING: Couldn't open directory /lib/modules/2.6.34-4-virtual: No such file or directory
FATAL: Could not open /lib/modules/2.6.34-4-virtual/modules.dep.temp for writing: No such file or directory
FATAL: Could not load /lib/modules/2.6.34-4-virtual/modules.dep: No such file or directory
Couldn't get a file descriptor referring to the console
Begin: Loading essential drivers... ...
FATAL: Could not load /lib/modules/2.6.34-4-virtual/modules.dep: No such file or directory
FATAL: Could not load /lib/modules/2.6.34-4-virtual/modules.dep: No such file or directory
Done.
Begin: Running /scripts/init-premount ...
Done.
Begin: Mounting root file system... ...
Begin: Running /scripts/local-top ...
Done.
Begin: Waiting for root file system... ...
Done.
Gave up waiting for root device. Common problems:
- Boot args (cat /proc/cmdline)
- Check rootdelay= (did the system wait long enough?)
- Check root= (did the system wait for the right device?)
- Missing modules (cat /proc/modules; ls /dev)
FATAL: Could not load /lib/modules/2.6.34-4-virtual/modules.dep: No such file or directory
FATAL: Could not load /lib/modules/2.6.34-4-virtual/modules.dep: No such file or directory
ALERT! /dev/sda1 does not exist. Dropping to a shell!
BusyBox v1.13.3 (Ubuntu 1:1.13.3-1ubuntu5) built-in shell (ash)
Enter 'help' for a list of built-in commands.
(initramfs)
Potential causes
One or more of the following conditions can cause this problem:
-
Missing ramdisk
-
Missing correct modules from ramdisk
-
Amazon EBS root volume not correctly attached as
/dev/sda1
Suggested actions
| For this instance type | Do this |
|---|---|
|
Amazon EBS-backed |
Use the following procedure:
|
|
Instance store-backed |
Use the following procedure:
|
ERROR Invalid kernel (EC2 incompatible kernel)
This condition is indicated by a system log similar to the one shown below.
...
root (hd0)
Filesystem type is ext2fs, using whole disk
kernel /vmlinuz root=/dev/sda1 ro
initrd /initrd.img
ERROR Invalid kernel: elf_xen_note_check: ERROR: Will only load images
built for the generic loader or Linux images
xc_dom_parse_image returned -1
Error 9: Unknown boot failure
Booting 'Fallback'
root (hd0)
Filesystem type is ext2fs, using whole disk
kernel /vmlinuz.old root=/dev/sda1 ro
Error 15: File not found
Potential causes
One or both of the following conditions can cause this problem:
-
Supplied kernel is not supported by GRUB
-
Fallback kernel does not exist
Suggested actions
| For this instance type | Do this |
|---|---|
|
Amazon EBS-backed |
Use the following procedure:
|
|
Instance store-backed |
Use the following procedure:
|
fsck: No such file or directory while trying to open... (File system not found)
This condition is indicated by a system log similar to the one shown below.
Welcome to Fedora
Press 'I' to enter interactive startup.
Setting clock : Wed Oct 26 05:52:05 EDT 2011 [ OK ]
Starting udev: [ OK ]
Setting hostname localhost: [ OK ]
No devices found
Setting up Logical Volume Management: File descriptor 7 left open
No volume groups found
[ OK ]
Checking filesystems
Checking all file systems.
[/sbin/fsck.ext3 (1) -- /] fsck.ext3 -a /dev/sda1
/dev/sda1: clean, 82081/1310720 files, 2141116/2621440 blocks
[/sbin/fsck.ext3 (1) -- /mnt/dbbackups] fsck.ext3 -a /dev/sdh
fsck.ext3: No such file or directory while trying to open /dev/sdh
/dev/sdh:
The superblock could not be read or does not describe a correct ext2
filesystem. If the device is valid and it really contains an ext2
filesystem (and not swap or ufs or something else), then the superblock
is corrupt, and you might try running e2fsck with an alternate superblock:
e2fsck -b 8193 <device>
[FAILED]
*** An error occurred during the file system check.
*** Dropping you to a shell; the system will reboot
*** when you leave the shell.
Give root password for maintenance
(or type Control-D to continue):
Potential causes
-
A bug exists in ramdisk filesystem definitions /etc/fstab
-
Misconfigured filesystem definitions in /etc/fstab
-
Missing/failed drive
Suggested actions
| For this instance type | Do this |
|---|---|
|
Amazon EBS-backed |
Use the following procedure:
The sixth field in the fstab defines availability requirements of the mount – a nonzero value implies that an fsck will be done on that volume and must succeed. Using this field can be problematic in Amazon EC2 because a failure typically results in an interactive console prompt that is not currently available in Amazon EC2. Use care with this feature and read the Linux man page for fstab. |
|
Instance store-backed |
Use the following procedure:
|
General error mounting filesystems (failed mount)
This condition is indicated by a system log similar to the one shown below.
Loading xenblk.ko module
xen-vbd: registered block device major 8
Loading ehci-hcd.ko module
Loading ohci-hcd.ko module
Loading uhci-hcd.ko module
USB Universal Host Controller Interface driver v3.0
Loading mbcache.ko module
Loading jbd.ko module
Loading ext3.ko module
Creating root device.
Mounting root filesystem.
kjournald starting. Commit interval 5 seconds
EXT3-fs: mounted filesystem with ordered data mode.
Setting up other filesystems.
Setting up new root fs
no fstab.sys, mounting internal defaults
Switching to new root and running init.
unmounting old /dev
unmounting old /proc
unmounting old /sys
mountall:/proc: unable to mount: Device or resource busy
mountall:/proc/self/mountinfo: No such file or directory
mountall: root filesystem isn't mounted
init: mountall main process (221) terminated with status 1
General error mounting filesystems.
A maintenance shell will now be started.
CONTROL-D will terminate this shell and re-try.
Press enter for maintenance
(or type Control-D to continue):
Potential causes
| Instance type | Potential cause |
|---|---|
|
Amazon EBS-backed |
|
|
Instance store-backed |
|
Suggested actions
| For this instance type | Do this |
|---|---|
|
Amazon EBS-backed |
Use the following procedure:
|
|
Instance store-backed |
Try one of the following:
|
VFS: Unable to mount root fs on unknown-block (Root filesystem mismatch)
This condition is indicated by a system log similar to the one shown below.