mdadm

Commit Graph

Author	SHA1	Message	Date
NeilBrown	2e48e34945	test: udev-settle before testing device. I think we sometime get way ahead of udev and devices disappear and appear almost at random. So add some settling. Signed-off-by: NeilBrown <neilb@suse.de>	2009-10-19 16:56:13 +11:00
Mike Frysinger	d16c7af6d8	mdadm(8): fix spurious space after -e header Signed-off-by: Mike Frysinger <vapier@gentoo.org> Signed-off-by: NeilBrown <neilb@suse.de>	2009-10-19 13:15:48 +11:00
Zdenek Behan	9a36a9b713	Monitor: add option to specify rebuild increments ie. the percent increments after which RebuildNN event is generated This is particulary useful when using --program option, rather than (only) syslog for alerts. Signed-off-by: Zdenek Behan <rain@matfyz.cz> Signed-off-by: NeilBrown <neilb@suse.de>	2009-10-19 13:13:58 +11:00
NeilBrown	1373b07d75	mdmon: lock current memory as well as future memory. mlockall(MCL_FUTURE) only locks mappings that have not yet been created. To lock all memory used by the process, we need MCL_CURRENT \| MCL_FUTURE Signed-off-by: NeilBrown <neilb@suse.de>	2009-10-19 13:04:16 +11:00
NeilBrown	5d504f4278	Merge git://github.com/djbw/mdadm	2009-10-19 12:52:58 +11:00
NeilBrown	6636f0efb3	tests/imsm: allow for rounding of array size. IMSM rounds array size to a multiple of 1024K, so our tests must assume this. Signed-off-by: NeilBrown <neilb@suse.de>	2009-10-16 17:57:28 +11:00
NeilBrown	ba6241244b	Test different r5/r6 layouts. Make sure kernel and restripe agree on all different layouts. Signed-off-by: NeilBrown <neilb@suse.de>	2009-10-16 17:50:07 +11:00
NeilBrown	1eac9f8454	restripe: fix assignment of raid6 blocks for syndrome calculation. Particularly for the _6 style. Signed-off-by: NeilBrown <neilb@suse.de>	2009-10-16 17:50:06 +11:00
NeilBrown	4180aa4d4e	Handle negative delta_disks in super0 and super1. Signed-off-by: NeilBrown <neilb@suse.de>	2009-10-16 17:43:54 +11:00
NeilBrown	82f2d6abf0	Grow_restart to handle reducing number of devices in an array. FIXME this is wrong . what direction does reshape_position move? If the device count in an array is shrinking, the critical region is different so the tests need to be different when restarting. Signed-off-by: NeilBrown <neilb@suse.de>	2009-10-16 17:43:51 +11:00
NeilBrown	eba7152931	Grow: don't make 'blocks' too large during in-place reshape. On small (test) arrays, multiplying by 16 can make the 'chunk' size larger than half the array, which is a problem. Signed-off-by: NeilBrown <neilb@suse.de>	2009-10-16 17:02:34 +11:00
Dan Williams	9f1da82421	mdmon: preserve socket over chroot Connect to the monitor in the old namespace and use that connection for WaitClean requests when stopping the victim mdmon instance. This allows ping_monitor() to work post chroot(). Cc: Hans de Goede <hdegoede@redhat.com> Signed-off-by: Dan Williams <dan.j.williams@intel.com>	2009-10-13 17:41:58 -07:00
Dan Williams	b928b5a038	mdmon: exec(2) when the switchroot argument is not "/" Try to execute mdmon from the target namespace. When used for initramfs handovers we need to drop all references to the initramfs filesystem for that memory to be freed. Cc: Hans de Goede <hdegoede@redhat.com> Signed-off-by: Dan Williams <dan.j.williams@intel.com>	2009-10-13 17:41:58 -07:00
Dan Williams	96a8270d46	mdmon: avoid writes in the startup path for mdmon on root arrays When killing a previous monitor be careful not to cause writes to the filesystem until the reads necessary to get the monitor operational have completed. The code is already prepared for errors creating the pid and socket files, so simply defer creation of these files until after the first call to manage(). Cc: Hans de Goede <hdegoede@redhat.com> Signed-off-by: Dan Williams <dan.j.williams@intel.com>	2009-10-13 17:41:57 -07:00
Dan Williams	aae5a11207	Detail: export MD_UUID from mapfile The load_super() from an mdadm --detail call may race against an mdmon update. When this happens the load_super sees an inconsistent metadata block and returns an error. The fallback path to use the map file contents lacks uuid reporting, so provide __fname_from_uuid for generically printing a uuid. Reported-by: Hans de Goede <hdegoede@redhat.com> Signed-off-by: Dan Williams <dan.j.williams@intel.com>	2009-10-13 17:41:57 -07:00
Dan Williams	d2b9eb5993	imsm: regression test for prodigal array member scenario Provide a test to sanity check assembly and reassembly in the presence of conflicting family number information. Signed-off-by: Dan Williams <dan.j.williams@intel.com>	2009-10-13 17:41:53 -07:00
Dan Williams	6e46bf344b	imsm: add --update=uuid support When disks have conflicting container memberships (same container ids but incompatible member arrays) --update=uuid can be used to move offenders to a new container id by changing 'orig_family_num'. Note that this only supports random updates of the uuid as the actual uuid is synthesized. We also need to communicate the new 'orig_family_num' value to all disks involved in the update. A new field 'update_private' is added to struct mdinfo to allow this information to be transmitted. Signed-off-by: Dan Williams <dan.j.williams@intel.com>	2009-10-13 17:41:53 -07:00
Dan Williams	955e9ea139	ddf: prevent superblock being zeroed on --update The full fix would be to support updating ddf metadata, but this minimal fix just prevents the superblock from being zeroed when someone inadvertently passes an unsupported --update option during assembly. Reported-by: Hans de Goede <hdegoede@redhat.com> Signed-off-by: Dan Williams <dan.j.williams@intel.com>	2009-10-13 17:41:53 -07:00
Dan Williams	e683ca88ac	imsm: fix/support --update Fix init_super_imsm() to return an empty mpb when info == NULL, and teach store_super_imsm() to simply write out the passed in mpb. Fixes: https://bugzilla.redhat.com/show_bug.cgi?id=523320 Reported-by: Hans de Goede <hdegoede@redhat.com> Signed-off-by: Dan Williams <dan.j.williams@intel.com>	2009-10-13 17:41:53 -07:00
Dan Williams	f796af5d5e	imsm: fix spare record writeout race imsm_activate_spare() in the manager thread may race against write_super_imsm_spares() in the monitor thread. Give write_super_imsm_spares() its own private mpb buffer to prevent confusing the manager. This change uncovered cases where spares were not being assembled due to a failed metadata version number check. Spares can freely associate across metadata version number, so reduce the scope of the version check in the spare assembly case. Signed-off-by: Dan Williams <dan.j.williams@intel.com>	2009-10-13 17:41:53 -07:00
NeilBrown	521f349cb0	restripe: fix compile warning. Just a type cast... Signed-off-by: NeilBrown <neilb@suse.de>	2009-10-12 17:00:23 +11:00
NeilBrown	471ac41e46	test changelevel: add tests for changing degraded arrays. Signed-off-by: NeilBrown <neilb@suse.de>	2009-10-12 16:57:55 +11:00
NeilBrown	cc50ccdc29	restripe : various fixed for RAID6 2-failure recovery. Signed-off-by: NeilBrown <neilb@suse.de>	2009-10-12 16:57:22 +11:00
NeilBrown	487e48afab	Test level changes and related reshaping. Signed-off-by: NeilBrown <neilb@suse.de>	2009-10-12 16:57:18 +11:00
NeilBrown	725cac4c56	Grow: ignore error from final wait_backup The last time wait_backup is called, it might see reshape finish and so return an error indicator. But this is not an error, and we must go ahead and prepare the array for full access. Signed-off-by: NeilBrown <neilb@suse.de>	2009-10-12 16:55:19 +11:00
NeilBrown	5fdf37e357	Grow: make sure bsb2 is properly aligned We do O_DIRECT io in bsb2, so it must be aligned properly. Easiest if it is static. Signed-off-by: NeilBrown <neilb@suse.de>	2009-10-12 16:55:12 +11:00
NeilBrown	249887eb76	testreshape5 - add tests for RAID6 .. to make sure our raid6 calculations are working. Signed-off-by: NeilBrown <neilb@suse.de>	2009-10-12 16:55:05 +11:00
NeilBrown	ca4f89a3b7	Merge branch 'master' into devel-3.1 Conflicts: mdadm.8	2009-10-01 16:58:40 +10:00
NeilBrown	2b9aa337af	Fix null-dereference in set_member_info set_member_info would try to dereference ->metadata_version, without checking that it isn't NULL. Signed-off-by: NeilBrown <neilb@suse.de>	2009-10-01 12:51:04 +10:00
NeilBrown	0e90271e53	Add missing space in "--detail --brief" output. We need a space between the device name and the word "level".. Signed-off-by: NeilBrown <neilb@suse.de>	2009-10-01 12:38:31 +10:00
Dan Williams	a2b9798159	imsm: disambiguate family_num This is a result of trawling through the Windows implementation to learn the mechanism of how it disambiguates family_num. It is a continuation of commit `148acb7b` "imsm: fix family number handling" which introduced a regression when reassembling a container with stale disks and rebuilt members. When rebuilding, a new family number is assigned to protect against the "prodigal array member" problem. It prevents a former family member from returning to the system and causing a rebuild to go the wrong direction. However, this invalidates looking at the generation number to determine the most up-to-date disk when comparing across family numbers. Instead the assembly logic looks for agreement between a disk's local family membership compared against a global list of all families in the system. Whenever a disk's local metadata does not match a family number on the global list that family number is marked offline. It is possible that this logic results in multiple incompatible but valid family numbers existing in a container. In this case mdadm.conf cannot be consulted because it only records the uuid which is generated from static fields in the metadata. The metadata lacks the data needed to disambiguate "local" versus "foreign". The "foreign" array in this case requires updating to change its container-id information (orig_family_num), and possibly the member array names. Signed-off-by: Dan Williams <dan.j.williams@intel.com>	2009-09-30 11:45:41 -07:00
Dan Williams	51725a7c25	imsm: kill close() of component device None of the other formats close the passed in fd at load, and this becomes a problem when trying to support --update where we need O_EXCL protection across the entire operation. Signed-off-by: Dan Williams <dan.j.williams@intel.com>	2009-09-30 11:44:38 -07:00
Dan Williams	25ed7e5924	imsm: cleanup disk status tests Add is_failed(), is_configured(), and is_spare() helpers to clean up disk status flag testing. Signed-off-by: Dan Williams <dan.j.williams@intel.com>	2009-09-28 14:40:59 -07:00
NeilBrown	58ad57f684	Release mdadm-3.0.2 Just one bugfix.	2009-09-25 18:19:07 +10:00
NeilBrown	40d28f0d1b	super0: fix crash on assemble if homehost is not set. If homehost is not set - typically during early boot, and assemble of v0.90 metadata arrays will crash. Reported-by: Paweł Sikora <pluto@agmk.net> Signed-off-by: NeilBrown <neilb@suse.de>	2009-09-25 17:56:22 +10:00
NeilBrown	e38cc2d87b	Fix raid6 error recovery in 'restripe' code. Thanks to Matthias Urlichs for discovering and reporting this. Signed-off-by: NeilBrown <neilb@suse.de>	2009-09-25 17:23:33 +10:00
NeilBrown	d8419fe9e9	Release mdadm-3.0.1 Just bugfixes. Signed-off-by: NeilBrown <neilb@suse.de>	2009-09-25 17:08:19 +10:00
NeilBrown	2de8abd572	testreshape5 - flush devices between tests. We need to flush the block devices before reading different data. Signed-off-by: NeilBrown <neilb@suse.de>	2009-09-25 16:57:01 +10:00
NeilBrown	ae80545ac3	Merge branch 'master' of git://github.com/djbw/mdadm	2009-09-25 14:11:11 +10:00
Hans de Goede	f5df5d69a7	mdmon: fix freeing unallocated memory mdmon was creating a supertype struct with malloc, and thus not necessarily getting zero-d memory. This was causing it to segfault when called like this from the initrd: /sbin/mdmon /proc/mdstat /sysroot The problem was that load_super_imsm would get called on the non-zero'd super struct, whcih in turn calls free_super_imsm, which checks st->sb, which should be zero but isn't and then starts freeing bogus memory. Signed-off-by: Dan Williams <dan.j.williams@intel.com>	2009-09-24 06:52:06 -07:00
Dan Williams	cf53434e5c	imsm: clear CONFIGURED_DISK for failed drives Synchronizing with what the Windows driver does. Signed-off-by: Dan Williams <dan.j.williams@intel.com>	2009-09-15 11:35:28 -07:00
Dan Williams	ee5aad5ae2	imsm: kill USABLE_DISK flag 'USABLE_DISK' is not a 'persistent' status flag it is an internal status flag used for the in memory representation of the disk in the Windows driver. Signed-off-by: Dan Williams <dan.j.williams@intel.com>	2009-09-15 11:35:28 -07:00
Dan Williams	ed57a7e8ba	Examine: don't count containers as spares mdadm -Ebs will include containers in the scanned device list. Examine() falsely thinks they are spares when MD_DISK_SYNC is not set. This could be fixed by forcing all formats to set this flag for container devices, but this flag is currently used by imsm to identify free-floating spares. Signed-off-by: Dan Williams <dan.j.williams@intel.com>	2009-09-15 11:35:28 -07:00
Dan Williams	436305c690	Detail: fix for an imsm container with a spare Spares for imsm arrays do not have any info about the container in their metadata records. If Detail() inadvertantly picks such a device for ->get_array_info() it will end up with less than useful info for the container. So, continue to read from the disks until a non-spare device is found. This bug was found by timeouts waiting for udev to create the user-friendly container name. To detect future UUID reporting problems and a debug print to the timeout case in wait_for(). Signed-off-by: Dan Williams <dan.j.williams@intel.com>	2009-09-15 11:34:20 -07:00
Dan Williams	ee836c39b5	Examine: fixup output in the presence of containers with spares If we dump any 'spare' or 'device' information for a container in the 'brief' case then we need a newline before printing member array info. Signed-off-by: Dan Williams <dan.j.williams@intel.com>	2009-09-15 11:34:20 -07:00
Dan Williams	709743c554	imsm: fix spare promotion 1/ Fix an off by one error when detecting whether the device allocation loop succeeded or not 2/ Update ->num_raid_devs before copying to avoid a segmentation fault Signed-off-by: Dan Williams <dan.j.williams@intel.com>	2009-09-15 11:34:20 -07:00
NeilBrown	2a17c77bdb	Add a missing 'closedir'. Thanks to David Binderman for finding and reporting it. Signed-off-by: NeilBrown <neilb@suse.de>	2009-09-11 16:10:24 +10:00
NeilBrown	7cbeb80e90	super1: remove fd leak when opening /dev/urandom As reported in https://bugzilla.novell.com/show_bug.cgi?id=527722 I forgot to close the fd after reading the random number. Signed-off-by: NeilBrown <neilb@suse.de>	2009-08-13 15:02:39 +10:00
NeilBrown	f24e2d6c06	mdadm.8 : update documentation for new --grow modes	2009-08-13 11:41:40 +10:00
NeilBrown	e9e43ec367	Grow: support restart of new migrations.	2009-08-13 11:12:54 +10:00

... 3 4 5 6 7 ...

1302 Commits All Branches Search

1302 Commits

All Branches