redict

mirror of https://codeberg.org/redict/redict.git synced 2025-01-22 16:18:28 -05:00

Author	SHA1	Message	Date
antirez	3f92e05637	Clarify a comment in slaveTryPartialResynchronization().	2014-01-08 14:28:13 +01:00
antirez	fdf50e1e3d	Log disconnection with slave only when ip:port is available.	2013-12-25 18:41:53 +01:00
antirez	2041882286	anetPeerToString / SockName: port can be NULL on errors too.	2013-12-25 18:41:49 +01:00
antirez	a2a900356e	anetTcpGenericConnect() bug introduced in 9d19977 fixed. Durign a refactoring I mispelled _port for port. This is one of the reasons I never used _varname myself.	2013-12-25 18:41:45 +01:00
antirez	cb23d510f4	Remove useless goto from anetTcpGenericConnect().	2013-12-25 18:41:41 +01:00
antirez	491f681088	anetTcpGenericConnect() code improved + 1 bug fix. Now the socket is closed if anetNonBlock() fails, and in general the code structure makes it harder to introduce this kind of bugs in the future. Reference: pull request #1059.	2013-12-25 18:15:28 +01:00
antirez	f510549044	Cluster: clusterProcessPacket() was not 80 cols friendly. The function actually needs to be split into sub-functions at some point in the future.	2013-12-25 17:57:36 +01:00
antirez	e789384255	Fix CONFIG REWRITE handling of unknown options. There were two problems with the implementation. 1) "save" was not correctly processed when no save point was configured, as reported in issue #1416. 2) The way the code checked if an option existed in the "processed" dictionary was wrong, as we add the element with as a key associated with a NULL value, so dictFetchValue() can't be used to check for existance, but dictFind() must be used, that returns NULL only if the entry does not exist at all.	2013-12-23 12:50:27 +01:00
antirez	7e9433cee1	Configuring port to 0 disables IP socket as specified. This was no longer the case with 2.8 becuase of a bug introduced with the IPv6 support. Now it is fixed. This fixes issue #1287 and #1477.	2013-12-23 11:31:35 +01:00
antirez	94e8c9e77e	Make new masters inherit replication offsets. Currently replication offsets could be used into a limited way in order to understand, out of a set of slaves, what is the one with the most updated data. For example this comparison is possible of N slaves were replicating all with the same master. However the replication offset was not transferred from master to slaves (that are later promoted as masters) in any way, so for instance if there were three instances A, B, C, with A master and B and C replication from A, the following could happen: C disconnects from A. B is turned into master. A is switched to master of B. B receives some write. In this context there was no way to compare the offset of A and C, because B would use its own local master replication offset as replication offset to initialize the replication with A. With this commit what happens is that when B is turned into master it inherits the replication offset from A, making A and C comparable. In the above case assuming no inconsistencies are created during the disconnection and failover process, A will show to have a replication offset greater than C. Note that this does not mean offsets are always comparable to understand what is, in a set of instances, since in more complex examples the replica with the higher replication offset could be partitioned away when picking the instance to elect as new master. However this in general improves the ability of a system to try to pick a good replica to promote to master.	2013-12-22 11:43:25 +01:00
antirez	ba5eb44d14	Slave disconnection is an event worth logging.	2013-12-22 10:15:35 +01:00
antirez	66ec1412fe	Redis Cluster: add repl_ping_slave_period to slave data validity time. When the configured node timeout is very small, the data validity time (maximum data age for a slave to try a failover) is too little (ten times the configured node timeout) when the replication link with the master is mostly idle. In this case we'll receive some data from the master only every server.repl_ping_slave_period to refresh the last interaction with the master. This commit adds to the max data validity time the slave ping period to avoid this problem of slaves sensing too old data without a good reason. However this max data validity time is likely a setting that should be configurable by the Redis Cluster user in a way completely independent from the node timeout.	2013-12-22 10:05:16 +01:00
antirez	b2dedd9da8	Log when a slave lose the connection with its master.	2013-12-21 00:23:37 +01:00
antirez	658aff9d29	Redis Cluster: move node failure reports logging from VERBOSE to NOTICE level.	2013-12-21 00:04:53 +01:00
antirez	5a404c87c1	Redis Cluster: remove no longer relevant comment.	2013-12-20 14:40:11 +01:00
antirez	fda4cba912	Redis Cluster: reconfigure replication when master changes address.	2013-12-20 12:47:22 +01:00
antirez	d7374032c0	Redis Cluster: handshake code refactoring + Gossip IP switch detection. This commit makes it simple to start an handshake with a specific node address, and uses this in order to detect a node IP change and start a new handshake in order to fix the IP if possible.	2013-12-20 12:38:03 +01:00
antirez	a2c938c834	Redis Cluster: delay state change when in the majority again. As specified in the Redis Cluster specification, when a node can reach the majority again after a period in which it was partitioend away with the minorty of masters, wait some time before accepting queries, to provide a reasonable amount of time for other nodes to upgrade its configuration. This lowers the probabilities of both a client and a master with not updated configuration to rejoin the cluster at the same time, with a stale master accepting writes.	2013-12-20 09:56:18 +01:00
antirez	305d7f29f3	Clarify include directive behavior in example redis.conf.	2013-12-19 16:02:31 +01:00
antirez	b3632319a4	CONFIG REWRITE: no special handling or include and rename-command. CONFIG REWRITE is now wiser and does not touch what it does not understand inside redis.conf.	2013-12-19 15:57:11 +01:00
Yubao Liu	7da423f79f	CONFIG REWRITE: don't throw some options on config rewrite Those options will be thrown without this patch: include, rename-command, min-slaves-to-write, min-slaves-max-lag, appendfilename.	2013-12-19 15:56:48 +01:00
antirez	3b9cf3ed3a	CONFIG REWRITE: old development comments removed.	2013-12-19 15:30:06 +01:00
antirez	b221e13dac	CONFIG REWRITE: don't wipe unknown options. With this commit options not explicitly rewritten by CONFIG REWRITE are not touched at all. These include new options that may not have support for REWRITE, and other special cases like rename-command and include.	2013-12-19 15:25:45 +01:00
antirez	6d184e02be	Example redis.conf formatted to better show appendfilename option.	2013-12-19 10:18:45 +01:00
antirez	7a666ac419	Cluster: set n->slaves to NULL in clusterNodeResetSlaves(). The value was otherwise undefined, so next time the node was promoted again from slave to master, adding a slave to the list of slaves would likely crash the server or result into undefined behavior.	2013-12-17 14:50:24 +01:00
antirez	fda91dbde3	Cluster: check link is valid before sending UPDATE.	2013-12-17 12:28:37 +01:00
antirez	f57bb36ce7	Cluster: initialize todo_before_sleep flags to 0.	2013-12-17 12:22:02 +01:00
antirez	c70c0c6db7	Cluster: use proper type mstime_t for ping delay var.	2013-12-17 10:27:36 +01:00
antirez	7c1cbdceb2	Cluster: use an hardcoded 60 sec timeout in redis-trib connections. Later this should be configurable from the command line but at least now we use something more appropriate for our use case compared to the redis-rb default timeout.	2013-12-17 10:00:33 +01:00
antirez	47815d38e0	Fixed clearNodeFailureIfNeeded() time type to mstime_t. This prevented 32bit cluster instances from clearing the FAIL flag when needed.	2013-12-17 09:45:52 +01:00
antirez	e88e6a6334	Cluster: use long long for timestamps in clusterGenNodesDescription(). Ping sent and pong received fields need to be casted to long long to be printed correctly into 32 bit systems.	2013-12-17 09:38:11 +01:00
antirez	2dfc5e35a9	Makefile.dep updated.	2013-12-13 13:10:05 +01:00
antirez	b1ba58f341	SDIFF iterator misuse bug regression test added. See commit `c00453d` for more info about the bug.	2013-12-13 11:37:13 +01:00
antirez	c00453da1d	SDIFF iterator misuse fixed in diff algorithm #1 . The bug could be easily triggered by: SADD foo a b c 1 2 3 4 5 6 SDIFF foo foo When the key was the same in two sets, an unsafe iterator was used to check existence of elements in the same set we were iterating. Usually this would just result into a wrong output, however with the dict.c API misuse protection we have in place, the result was actually an assertion failed that was triggered by the CI test, while creating random datasets for the "MASTER and SLAVE consistency" test.	2013-12-13 11:34:21 +01:00
antirez	5320148883	Sentinel: dead code removed.	2013-12-13 11:01:13 +01:00
antirez	452dea30f6	Makefile: remove odd syntax not compatible with some make versions. See issue #1448.	2013-12-12 15:19:39 +01:00
Salvatore Sanfilippo	62e4956936	Merge pull request #1415 from Dieken/fix-typo fix typo in redis.conf and sentinel.conf	2013-12-12 02:30:11 -08:00
Salvatore Sanfilippo	a99c751d6c	Merge pull request #1460 from codeeply/simplify2 comment mistake fixed	2013-12-12 02:23:44 -08:00
codeeply	0f06f8df07	comment mistake fixed	2013-12-12 16:33:29 +08:00
antirez	a5ec247f13	Replication: publish the slave_repl_offset when disconnected from master. When a slave was disconnected from its master the replication offset was reported as -1. Now it is reported as the replication offset of the previous master, so that failover can be performed using this value in order to try to select a slave with more processed data from a set of slaves of the old master.	2013-12-11 15:23:15 +01:00
Salvatore Sanfilippo	0a89d9a0b1	Merge pull request #1451 from yossigo/unbalanced-quotes-fix Return proper error on requests with an unbalanced number of quotes.	2013-12-11 03:06:18 -08:00
Yossi Gottlieb	88a5cede88	Fix wrong repldboff type which causes dropped replication in rare cases.	2013-12-11 11:38:02 +01:00
Yubao Liu	6d5fa2e06c	fix typo in redis.conf and sentinel.conf	2013-12-11 15:46:42 +08:00
antirez	11120689c4	Slaves heartbeats during sync improved. The previous fix for false positive timeout detected by master was not complete. There is another blocking stage while loading data for the first synchronization with the master, that is, flushing away the current data from the DB memory. This commit uses the newly introduced dict.c callback in order to make some incremental work (to send "\n" heartbeats to the master) while flushing the old data from memory. It is hard to write a regression test for this issue unfortunately. More support for debugging in the Redis core would be needed in terms of functionalities to simulate a slow DB loading / deletion.	2013-12-10 18:47:31 +01:00
antirez	2eb781b35b	dict.c: added optional callback to dictEmpty(). Redis hash table implementation has many non-blocking features like incremental rehashing, however while deleting a large hash table there was no way to have a callback called to do some incremental work. This commit adds this support, as an optiona callback argument to dictEmpty() that is currently called at a fixed interval (one time every 65k deletions).	2013-12-10 18:46:24 +01:00
antirez	2c4ab8a534	Log empty DB + Loading data into two separated messages.	2013-12-10 18:43:25 +01:00
antirez	7c531eb5ad	Don't send more than 1 newline/sec while loading RDB.	2013-12-10 18:43:19 +01:00
antirez	27db38d069	Slaves heartbeat while loading RDB files. Starting with Redis 2.8 masters are able to detect timed out slaves, while before 2.8 only slaves were able to detect a timed out master. Now that timeout detection is bi-directional the following problem happens as described "in the field" by issue #1449: 1) Master and slave setup with big dataset. 2) Slave performs the first synchronization, or a full sync after a failed partial resync. 3) Master sends the RDB payload to the slave. 4) Slave loads this payload. 5) Master detects the slave as timed out since does not receive back the REPLCONF ACK acknowledges. Here the problem is that the master has no way to know how much the slave will take to load the RDB file in memory. The obvious solution is to use a greater replication timeout setting, but this is a shame since for the 0.1% of operation time we are forced to use a timeout that is not what is suited for 99.9% of operation time. This commit tries to fix this problem with a solution that is a bit of an hack, but that modifies little of the replication internals, in order to be back ported to 2.8 safely. During the RDB loading time, we send the master newlines to avoid being sensed as timed out. This is the same that the master already does while saving the RDB file to still signal its presence to the slave. The single newline is used because: 1) It can't desync the protocol, as it is only transmitted all or nothing. 2) It can be safely sent while we don't have a client structure for the master or in similar situations just with write(2).	2013-12-09 20:26:00 +01:00
antirez	eaf1bfb88b	Handle inline requested terminated with just \n.	2013-12-09 13:28:39 +01:00
Yossi Gottlieb	6e70c01148	Return proper error on requests with an unbalanced number of quotes.	2013-12-08 12:58:12 +02:00

1 2 3 4 5 ...

3753 Commits