redict

mirror of https://codeberg.org/redict/redict.git synced 2025-01-23 00:28:26 -05:00

Author	SHA1	Message	Date
Oran Agra	d31c0c5264	fix loading race in psync2 tests	2020-04-28 09:18:01 +03:00
Oran Agra	4447ddc8bb	Keep track of meaningful replication offset in replicas too Now both master and replicas keep track of the last replication offset that contains meaningful data (ignoring the tailing pings), and both trim that tail from the replication backlog, and the offset with which they try to use for psync. the implication is that if someone missed some pings, or even have excessive pings that the promoted replica has, it'll still be able to psync (avoid full sync). the downside (which was already committed) is that replicas running old code may fail to psync, since the promoted replica trims pings form it's backlog. This commit adds a test that reproduces several cases of promotions and demotions with stale and non-stale pings Background: The mearningful offset on the master was added recently to solve a problem were the master is left all alone, injecting PINGs into it's backlog when no one is listening and then gets demoted and tries to replicate from a replica that didn't have any of the PINGs (or at least not the last ones). however, consider this case: master A has two replicas (B and C) replicating directly from it. there's no traffic at all, and also no network issues, just many pings in the tail of the backlog. now B gets promoted, A becomes a replica of B, and C remains a replica of A. when A gets demoted, it trims the pings from its backlog, and successfully replicate from B. however, C is still aware of these PINGs, when it'll disconnect and re-connect to A, it'll ask for something that's not in the backlog anymore (since A trimmed the tail of it's backlog), and be forced to do a full sync (something it didn't have to do before the meaningful offset fix). Besides that, the psync2 test was always failing randomly here and there, it turns out the reason were PINGs. Investigating it shows the following scenario: cycle 1: redis #1 is master, and all the rest are direct replicas of #1 cycle 2: redis #2 is promoted to master, #1 is a replica of #2 and #3 is replica of #1 now we see that when #1 is demoted it prints: 17339:S 21 Apr 2020 11:16:38.523 * Using the meaningful offset 3929963 instead of 3929977 to exclude the final PINGs (14 bytes difference) 17339:S 21 Apr 2020 11:16:39.391 * Trying a partial resynchronization (request e2b3f8817735fdfe5fa4626766daa938b61419e5:3929964). 17339:S 21 Apr 2020 11:16:39.392 * Successful partial resynchronization with master. and when #3 connects to the demoted #2, #2 says: 17339:S 21 Apr 2020 11:16:40.084 * Partial resynchronization not accepted: Requested offset for secondary ID was 3929978, but I can reply up to 3929964 so the issue here is that the meaningful offset feature saved the day for the demoted master (since it needs to sync from a replica that didn't get the last ping), but it didn't help one of the other replicas which did get the last ping.	2020-04-27 15:52:23 +02:00
antirez	022f09447b	Merge branch 'unstable' of github.com:/antirez/redis into unstable	2020-04-24 16:59:56 +02:00
antirez	8a7f255cd0	LCS -> STRALGO LCS. STRALGO should be a container for mostly read-only string algorithms in Redis. The algorithms should have two main characteristics: 1. They should be non trivial to compute, and often not part of programming language standard libraries. 2. They should be fast enough that it is a good idea to have optimized C implementations. Next thing I would love to see? A small strings compression algorithm.	2020-04-24 16:54:32 +02:00
Salvatore Sanfilippo	42d309fffc	Merge pull request #7114 from guybe7/stream_tag_xsetid Add the stream tag to XSETID tests	2020-04-23 16:29:46 +02:00
Salvatore Sanfilippo	72f0751905	Merge pull request #7123 from fayadexinqing/optimizeClusterSlots Optimize the command of cluster slots	2020-04-23 16:18:22 +02:00
antirez	8d67211450	Tracking: test expired keys notifications.	2020-04-22 11:45:34 +02:00
antirez	58d61dd639	Tracking: NOLOOP tests.	2020-04-22 11:24:19 +02:00
yanhui13	6b547c3956	add tcl test for cluster slots	2020-04-21 16:56:10 +08:00
Guy Benoish	1bc557c9c5	Add the stream tag to XSETID tests	2020-04-19 15:59:58 +03:00
antirez	002052f8de	A few comments and name changes for #7103 .	2020-04-17 10:51:12 +02:00
Oran Agra	b9fa42a197	testsuite run the defrag latency test solo this test is time sensitive and it sometimes fail to pass below the latency threshold, even on strong machines. this test was the reson we're running just 2 parallel tests in the github actions CI, revering this.	2020-04-16 18:09:22 +03:00
antirez	121c51f4f3	Merge branch 'lcs' into unstable	2020-04-06 13:51:55 +02:00
antirez	af3c722fec	LCS: more tests.	2020-04-06 13:51:49 +02:00
antirez	8dc28b6c75	LCS tests.	2020-04-06 13:45:37 +02:00
Oran Agra	cf3789f045	diffrent fix for runtest --host --port	2020-04-06 09:41:14 +03:00
Guy Benoish	1b0d30aeb7	Try to fix time-sensitive tests in blockonkey.tcl There is an inherent race between the deferring client and the "main" client of the test: While the deferring client issues a blocking command, we can't know for sure that by the time the "main" client tries to issue another command (Usually one that unblocks the deferring client) the deferring client is even blocked... For lack of a better choice this commit uses TCL's 'after' in order to give some time for the deferring client to issues its blocking command before the "main" client does its thing. This problem probably exists in many other tests but this commit tries to fix blockonkeys.tcl	2020-04-03 14:51:45 +03:00
Salvatore Sanfilippo	cbf212f981	Merge pull request #7030 from valentinogeron/xread-in-lua XREAD and XREADGROUP should not be allowed from scripts when BLOCK op…	2020-04-03 11:14:13 +02:00
Guy Benoish	4665b3ebfb	Fix no-negative-zero test	2020-04-02 18:41:29 +03:00
Salvatore Sanfilippo	10b626b3d5	Merge pull request #6546 from guybe7/fix_neg_zero Make sure Redis does not reply with negative zero	2020-04-02 16:26:57 +02:00
Salvatore Sanfilippo	dfef407499	Merge pull request #7029 from valentinogeron/fix-xack XACK should be executed in a "all or nothing" fashion.	2020-04-02 11:23:23 +02:00
Guy Benoish	c4dc5b80b2	Fix memory corruption in moduleHandleBlockedClients By using a "circular BRPOPLPUSH"-like scenario it was possible the get the same client on db->blocking_keys twice (See comment in moduleTryServeClientBlockedOnKey) The fix was actually already implememnted in moduleTryServeClientBlockedOnKey but it had a bug: the funxction should return 0 or 1 (not OK or ERR) Other changes: 1. Added two commands to blockonkeys.c test module (To reproduce the case described above) 2. Simplify blockonkeys.c in order to make testing easier 3. cast raxSize() to avoid warning with format spec	2020-04-01 12:53:26 +03:00
Salvatore Sanfilippo	0c52ce6c8e	Merge pull request #7037 from guybe7/fix_module_replicate_multi Modules: Test MULTI/EXEC replication of RM_Replicate	2020-03-31 17:00:57 +02:00
Guy Benoish	6c8221580c	RENAME can unblock XREADGROUP Other changes: Support stream in serverLogObjectDebugInfo	2020-03-31 17:41:10 +03:00
Guy Benoish	d6eb3afd13	Modules: Test MULTI/EXEC replication of RM_Replicate Makse sure call() doesn't wrap replicated commands with a redundant MULTI/EXEC Other, unrelated changes: 1. Formatting compiler warning in INFO CLIENTS 2. Use CLIENT_ID_AOF instead of UINT64_MAX	2020-03-31 13:55:51 +03:00
antirez	4379b8b411	Fix the propagate Tcl test after module changes.	2020-03-31 12:09:38 +02:00
antirez	95f154985c	Modify the propagate unit test to show more cases.	2020-03-31 12:04:06 +02:00
antirez	9dcf878f1b	Fix module commands propagation double MULTI bug. `37a10cef` introduced automatic wrapping of MULTI/EXEC for the alsoPropagate API. However this collides with the built-in mechanism already present in module.c. To avoid complex changes near Redis 6 GA this commit introduces the ability to exclude call() MUTLI/EXEC wrapping for also propagate in order to continue to use the old code paths in module.c.	2020-03-31 11:00:45 +02:00
Valentino Geron	9a1843ef2d	XREAD and XREADGROUP should not be allowed from scripts when BLOCK option is being used	2020-03-26 15:46:31 +02:00
Valentino Geron	1547d72cf3	XACK should be executed in a "all or nothing" fashion. First, we must parse the IDs, so that we abort ASAP. The return value of this command cannot be an error if the client successfully acknowledged some messages, so it should be executed in a "all or nothing" fashion.	2020-03-26 15:40:23 +02:00
Salvatore Sanfilippo	2ea7f0ecad	Merge pull request #6644 from oranagra/stream_aofrw AOFRW on an empty stream created with MKSTREAM loads badkly	2020-03-26 11:12:44 +01:00
Oran Agra	3b29556a0c	AOFRW on an empty stream created with MKSTREAM loads badkly the AOF will be loaded successfully, but the stream will be missing, i.e inconsistencies with the original db. this was because XADD with id of 0-0 would error. add a test to reproduce.	2020-03-25 21:47:57 +02:00
antirez	c4d7f30e25	PSYNC2: meaningful offset test.	2020-03-25 15:43:34 +01:00
Oran Agra	ec007559ff	MULTI/EXEC during LUA script timeout are messed up Redis refusing to run MULTI or EXEC during script timeout may cause partial transactions to run. 1) if the client sends MULTI+commands+EXEC in pipeline without waiting for response, but these arrive to the shards partially while there's a busy script, and partially after it eventually finishes: we'll end up running only part of the transaction (since multi was ignored, and exec would fail). 2) similar to the above if EXEC arrives during busy script, it'll be ignored and the client state remains in a transaction. the 3rd test which i added for a case where MULTI and EXEC are ok, and only the body arrives during busy script was already handled correctly since processCommand calls flagTransaction	2020-03-23 20:45:32 +02:00
antirez	61de1c1146	Fix BITFIELD_RO test.	2020-03-23 12:02:12 +01:00
Salvatore Sanfilippo	493a7f9823	Merge pull request #6951 from yangbodong22011/feature-bitfield-ro Added BITFIELD_RO variants for read-only operations.	2020-03-23 11:23:21 +01:00
antirez	1e16b9384d	Merge branch 'unstable' of github.com:/antirez/redis into unstable	2020-03-20 13:21:28 +01:00
antirez	5497a44037	Regression test for #7011 .	2020-03-20 12:52:06 +01:00
WuYunlong	af5167b7f3	Add 14-consistency-check.tcl to prove there is a data consistency issue.	2020-03-18 16:17:46 +08:00
bodong.ybd	336458d4b5	Fix bug of tcl test using external server	2020-03-11 21:01:27 +08:00
Oran Agra	27641ee490	fix for flaky psync2 test *** [err]: PSYNC2: total sum of full synchronizations is exactly 4 in tests/integration/psync2.tcl Expected 5 == 4 (context: type eval line 6 cmd {assert {$sum == 4}} proc ::test) issue was that sometime the test got an unexpected full sync since it tried to switch to the replica before it was in sync with it's master.	2020-03-05 16:55:14 +02:00
bodong.ybd	94376f46ad	Added BITFIELD_RO variants for read-only operations.	2020-03-04 20:51:45 +08:00
Salvatore Sanfilippo	d2c5f80e2e	Merge pull request #6926 from oranagra/fork-test-fix fix race in module api test for fork	2020-02-27 09:58:04 +01:00
Oran Agra	2f1a1c3835	fix github actions failing latency test for active defrag - part 2 it seems that running two clients at a time is ok too, resuces action time from 20 minutes to 10. we'll use this for now, and if one day it won't be enough we'll have to run just the sensitive tests one by one separately from the others. this commit also fixes an issue with the defrag test that appears to be very rare.	2020-02-27 08:34:53 +02:00
Oran Agra	537893420b	fix github actions failing latency test for active defrag seems that github actions are slow, using just one client to reduce false positives. also adding verbose, testing only on latest ubuntu, and building on older one. when doing that, i can reduce the test threshold back to something saner	2020-02-25 17:53:23 +02:00
Salvatore Sanfilippo	3fbb41ecc9	Merge pull request #6920 from oranagra/defrag-test-latency-fix Fix latency sensitivity of new defrag test	2020-02-24 11:53:32 +01:00
antirez	73305861f5	Test engine: experimental change to avoid busy port problems.	2020-02-24 10:46:23 +01:00
Oran Agra	0a643efa0c	fix race in module api test for fork in some cases we were trying to kill the fork before it got created	2020-02-23 16:48:37 +02:00
Oran Agra	62adabd0e0	Fix latency sensitivity of new defrag test I saw that the new defag test for list was failing in CI recently, so i reduce it's threshold from 12 to 60. besides that, i add / improve the latency test for that other two defrag tests (add a sensitive latency and digest / save checks) and fix bad usage of debug populate (can't overrides existing keys). this was the original intention, which creates higher fragmentation.	2020-02-23 13:05:52 +02:00
antirez	e78c4e813c	Test engine: detect timeout when checking for Redis startup.	2020-02-21 18:55:56 +01:00

1 2 3 4 5 ...

1059 Commits