minio

Commit Graph

Author	SHA1	Message	Date
Krishnan Parthasarathi	ca64b86112	Return possible states a heal operation (#4045 )	8 years ago
Krishnan Parthasarathi	2bd694dbc8	Add disksUnavailable healStatus const (#3990 ) `disksUnavailable` healStatus constant indicates that a given object needs healing but one or more of disks requiring heal are offline. This can be used by admin heal API consumers to distinguish between a successful heal and a no-op since the outdated disks were offline.	8 years ago
Krishnan Parthasarathi	c27ece409b	heal: Check if all parts are available and valid (#3967 ) In the algorithm to check if an object requires healing, in addition to checking if all disks have xl.json present we should check if all parts of the object are present and have valid blake2b checksums. Also fixed a minor compilation error in heal-objects-list.go.	8 years ago
Krishnan Parthasarathi	c192e5c9b2	Implement heal-upload admin API (#3914 ) This API is meant for administrative tools like mc-admin to heal an ongoing multipart upload on a Minio server. N B This set of admin APIs apply only for Minio servers. `github.com/minio/minio/pkg/madmin` provides a go SDK for this (and other admin) operations. Specifically, func HealUpload(bucket, object, uploadID string, dryRun bool) error Sample admin API request: POST /?heal&bucket=mybucket&object=myobject&upload-id=myuploadID&dry-run - Header(s): ["x-minio-operation"] = "upload" Notes: - bucket, object and upload-id are mandatory query parameters - if dry-run is set, API returns success if all parameters passed are valid.	8 years ago
Harshavardhana	e49efcb9d9	xl: quickHeal heal bucket only when needed. (#3854 ) This improves the startup time significantly for clusters which have lot of buckets. Also fixes a bug where `.minio.sys` is created on disks which do not have `format.json`	8 years ago
Krishnan Parthasarathi	e3fd4c0dd6	XL: Make listOnlineDisks and outDatedDisks consistent w/ each other. (#3808 )	8 years ago
Harshavardhana	bcc5b6e1ef	xl: Rename getOrderedDisks as shuffleDisks appropriately. (#3796 ) This PR is for readability cleanup - getOrderedDisks as shuffleDisks - getOrderedPartsMetadata as shufflePartsMetadata Distribution is now a second argument instead being the primary input argument for brevity. Also change the usage of type casted int64(0), instead rely on direct type reference as `var variable int64` everywhere.	8 years ago
Harshavardhana	6a6c930f5b	xl: Abort multipart upload should honor quorum properly. (#3670 ) Current implementation didn't honor quorum properly and didn't handle the errors generated properly. This patch addresses that and also moves common code `cleanupMultipartUploads` into xl specific private function. Fixes #3665	8 years ago
Krishnan Parthasarathi	864b8795aa	heal: Should delete stale object parts before healing (#3649 )	8 years ago
Anis Elleuch	0715032598	heal: Add ListBucketsHeal object API (#3563 ) ListBucketsHeal will list which buckets that need to be healed: * ListBucketsHeal() (buckets []BucketInfo, err error)	8 years ago
Krishnan Parthasarathi	c194b9f5f1	Implement mgmt REST APIs for heal subcommands (#3533 ) The heal APIs supported in this change are, - listing of objects to be healed. - healing a bucket. - healing an object.	8 years ago
Harshavardhana	1c699d8d3f	fs: Re-implement object layer to remember the fd (#3509 ) This patch re-writes FS backend to support shared backend sharing locks for safe concurrent access across multiple servers.	8 years ago
Harshavardhana	2d6f8153fa	format: Check properly for disks in valid formats. (#3427 ) There was an error in how we validated disk formats, if one of the disk was formatted and was formatted with FS would cause confusion and object layer would never initialize essentially go into an infinite loop. Validate pre-emptively and also check for FS format properly.	8 years ago
Harshavardhana	4daa0d2cee	lock: Moving locking to handler layer. (#3381 ) This is implemented so that the issues like in the following flow don't affect the behavior of operation. ``` GetObjectInfo() .... --> Time window for mutation (no lock held) .... --> Time window for mutation (no lock held) GetObject() ``` This happens when two simultaneous uploads are made to the same object the object has returned wrong info to the client. Another classic example is "CopyObject" API itself which reads from a source object and copies to destination object. Fixes #3370 Fixes #2912	8 years ago
Harshavardhana	ff4ce0ee14	fs/xl: Combine input checks into re-usable functions. (#3383 ) Repeated code around both object layers are moved and combined into simple re-usable functions.	8 years ago
Bala FA	0f2e493c9a	Use isErrIgnored() function wherever applicable. (#3343 )	8 years ago
Bala FA	1d4ac4b084	Rename getUUID() into mustGetUUID() (#3320 ) In case of UUID generation failure mustGetUUID() will panic than infinitely trying in for loop.	8 years ago
Harshavardhana	5197649081	utils: reduceErrs returns and validates quorum errors. (#3300 ) This is needed as explained by @krisis Lets say we have following errors. ``` []error{nil, errFileNotFound, errDiskAccessDenied, errDiskAccesDenied} ``` Since the last two errors are filtered, the maximum is nil, depending on map order. Let's say we get nil from reduceErr. Clearly at this point we don't have quorum nodes agreeing about the data and since GetObject only requires N/2 (Read quorum) and isDiskQuorum would have returned true. This is problematic and can lead to undersiable consequences. Fixes #3298	8 years ago
Krishnan Parthasarathi	eed9ab0464	XL: pickValidXLMeta should return error instead of panic'ing (#3277 )	8 years ago
Harshavardhana	0b9f0d14a1	auth/rpc: Take remote disk offline after maximum allowed attempts. (#3288 ) Disks when are offline for a long period of time, we should ignore the disk after trying Login upto 5 times. This is to reduce the network chattiness, this also reduces the overall time spent on `net.Dial`. Fixes #3286	8 years ago
Anis Elleuch	ffbee70e04	Avoid removing 'tmp' directory inside '.minio.sys' (#3294 )	8 years ago
Harshavardhana	1c47365445	xl/bootup: Upon bootup handle errors loading bucket and event configs. (#3287 ) In a situation when we have lots of buckets the bootup time might have slowed down a bit but during this situation the servers quickly going up and down would be an in-transit state. Certain calls which do not use quorum like `readXLMetaStat` might return an error saying `errDiskNotFound` this is returned in place of expected `errFileNotFound` which leads to an issue where server doesn't start. To avoid this situation we need to ignore them as safe values to be ignored, for the most part these are network related errors. Fixes #3275	8 years ago
Harshavardhana	c91d3791f9	heal: Add healing support for bucket, bucket metadata files. (#3252 ) This patch implements healing in general but it is only used as part of quickHeal(). Fixes #3237	8 years ago
Aditya Manthramurthy	dd0698d14c	Improve namespace lock API: (#3203 ) - abstract out instrumentation information. - use separate lockInstance type that encapsulates the nsMutex, volume, path and opsID as the frontend or top-level lock object.	8 years ago
Harshavardhana	39331b6b4e	xl: GetCheckSumInfo() shouldn't fail if hash not available. (#2984 ) In a multipart upload scenario disks going down and coming backup can lead to certain parts missing on the disk/server which was going down. This is a valid case since these blocks can be missing and should be healed through heal operation. But we are not supposed to fail prematurely since we have enough data on the other disks as well within read-quorum. This fix relaxes previous assumption, fixes a major corruption issue reproduced by @vadmeste. Fixes #2976	8 years ago
Harshavardhana	fee3f99a6e	xl: heal bucket should validate if bucket exists first. (#2953 ) Fixes #2944	8 years ago
Krishna Srinivas	f5f007e183	Test: Add test case for xl.HealObject() (#2884 ) fixes #2842	8 years ago
Harshavardhana	1e6d67b16d	server: Remove deadcode. (#2699 )	8 years ago
Krishna Srinivas	7cc77eba45	XL/Healing: errDiskNotFound is the only pardonable error in xlShouldHeal. (#2586 ) This is so that we try to heal a file for all the "bad" cases except when the disk is down.	8 years ago
Anis Elleuch	200d327737	List only objects that need healing (#2546 )	8 years ago
Harshavardhana	bccf549463	server: Move all the top level files into cmd folder. (#2490 ) This change brings a change which was done for the 'mc' package to allow for clean repo and have a cleaner github drop in experience.	8 years ago
Krishna Srinivas	e2498edb45	contoller: Implement controlled healing and trigger (#2381 ) This patch introduces new command line 'control' - minio control TO manage minio server connecting through GoRPC API frontend. - minio control heal Is implemented for healing objects.	8 years ago
karthic rao	5fe72cf205	Removing readAllMeta from xl-v1-healing.go and placing it in xl-v1-utils.go (#2296 )	8 years ago
Krishna Srinivas	b090c7112e	Refactor of xl.PutObjectPart and erasureCreateFile. (#2193 ) * XL: Refactor of xl.PutObjectPart and erasureCreateFile. * GetCheckSum and AddCheckSum methods for xlMetaV1 * Simple unit test case for erasureCreateFile()	8 years ago
Harshavardhana	3b69b4ada4	server: Change server startup message. (#2195 ) This change brings in the new agreed startup message for the server. Adds additional links point to Minio SDKs as well.	8 years ago
Harshavardhana	623e0f9243	XL: listOnlineDisks should use modTime instead of version. (#2166 ) This change is needed to make reading from objects future proof in-terms of handling online disks. Our current counter is not based on affirmative knowledge and relies on arithmetic sequence which can lead to bugs. Using modTime simplifies the understanding of `xl.json` and future tooling / debugging of the format.	8 years ago
Krishnan Parthasarathi	bef72f26db	xl: Make locking more granular for PutObjectPart requests (#2168 )	8 years ago
Harshavardhana	2e1f66c37d	XL: Handle quorum situations properly for write operations. (#1986 ) Adds two test cases one for - PutObject when write quorum is not available. - PutObjectPart when write quorum is not available. Fixes #1951	9 years ago
Harshavardhana	e8990e42c2	XL: Make allocations simpler avoid redundant allocs. (#1961 ) - Reduce 10MiB buffers for loopy calls to use 128KiB. - start using 128KiB buffer where needed.	9 years ago
Harshavardhana	8c0942bf0d	XL: Remove usage of reduceErr and make it isQuorum verification. (#1909 ) Fixes #1908	9 years ago
Krishna Srinivas	de1c7d33eb	XL: appendFile should return error if quorum is not met. (#1898 ) Fixes #1890	9 years ago
Harshavardhana	fb95c1fad3	XL: Bring in some modularity into format verification and healing. (#1832 )	9 years ago
Harshavardhana	c493ab5d0d	XL: Bring in sha512 checksum support. (#1797 )	9 years ago
Harshavardhana	a4a0ea605b	XL: Fix GetObject erasure decode issues. (#1793 )	9 years ago
Harshavardhana	feb337098d	XL: bring in new storage API. (#1780 ) Fixes #1771	9 years ago
Harshavardhana	553fdb9211	XL: Bring in support for object versions written during writeQuorum. (#1762 ) Erasure is initialized as needed depending on the quorum and onlineDisks. This way we can manage the quorum at the object layer.	9 years ago

31 Commits (6f4862659f7f86e1637fe579343bc878f499d9bf)