redict

mirror of https://codeberg.org/redict/redict.git synced 2025-01-23 08:38:27 -05:00

Author	SHA1	Message	Date
Oran Agra	8ea131fc80	Fix leak in new blockedclient module API test (#7780 )	2020-09-10 10:22:16 +03:00
Yossi Gottlieb	b2a73c404b	Tests: fix oom-score-adj false positives. (#7772 ) The key save delay is too short and on certain systems the child process is gone before we have a chance to inspect it.	2020-09-09 18:58:06 +03:00
杨博东	0666267d27	Tests: Add aclfile load and save tests (#7765 ) improves test coverage	2020-09-09 17:13:35 +03:00
Roi Lipman	042189fd87	RM_ThreadSafeContextTryLock a non-blocking method for acquiring GIL (#7738 ) Co-authored-by: Yossi Gottlieb <yossigo@gmail.com> Co-authored-by: Oran Agra <oran@redislabs.com>	2020-09-09 16:01:16 +03:00
Yossi Gottlieb	a8b7268911	Tests: validate CONFIG REWRITE for all params. (#7764 ) This is a catch-all test to confirm that that rewrite produces a valid output for all parameters and that this process does not introduce undesired configuration changes.	2020-09-09 15:43:11 +03:00
Yossi Gottlieb	818a746e32	Fix default/explicit "save" parameter loading. (#7767 ) Save parameters should either be default or whatever specified in the config file. This fixes an issue introduced in #7092 which causes configuration file settings to be applied on top of the defaults.	2020-09-09 15:12:57 +03:00
Yossi Gottlieb	918abd7276	Tests: clean up stale .cli files. (#7768 )	2020-09-09 12:30:43 +03:00
Eran Liberty	b120366d48	Allow exec with read commands on readonly replica in cluster (#7766 ) There was a bug. Although cluster replicas would allow read commands, they would not allow a MULTI-EXEC that's composed solely of read commands. Adds tests for coverage. Co-authored-by: Oran Agra <oran@redislabs.com> Co-authored-by: Eran Liberty <eranl@amazon.com>	2020-09-09 09:35:42 +03:00
Oran Agra	0a1e734193	handle cur_test for nested tests if there are nested tests and nested servers, we need to restore the previous value of cur_test when a test exist. example: ``` test{test 1} { start_server { test{test 1.1 - master only} { } start_server { test{test 1.2 - with replication} { } } } } ``` when `test 1.1 - master only exists`, we're still inside `test 1`	2020-09-08 14:12:03 +03:00
bodong.ybd	f22fa9594d	Tests: Some fixes for macOS 1) cur_test: when restart_server, "no such variable" error occurs ./runtest --single integration/rdb test {client freed during loading} SET ::cur_test restart_server kill_server test "Check for memory leaks (pid $pid)" SET ::cur_test UNSET ::cur_test UNSET ::cur_test // This global variable has been unset. 2) `ps --ppid` not available on macOS platform, can be replaced with `pgrep -P pid`.	2020-09-08 14:27:53 +08:00
Oran Agra	b491d477c3	Fix cluster consistency-check test (#7754 ) This test was failing from time to time see discussion at the bottom of #7635 This was probably due to timing, the DEBUG SLEEP executed by redis-cli didn't sleep for enough time. This commit changes: 1) use SET-ACTIVE-EXPIRE instead of DEBUG SLEEP 2) reduce many `after` sleeps with retry loops to speed up the test. 3) add many comment explaining the different steps of the test and it's purpose. 4) config appendonly before populating the volatile keys, so that they'll be part of the AOF command stream rather than the preamble RDB portion. other complications: recently kill_instance switched from SIGKILL to SIGTERM, and this would sometimes fail since there was an AOFRW running in the background. now we wait for it to end before attempting the kill.	2020-09-07 18:06:25 +03:00
Yossi Gottlieb	2df4cb93ac	Tests: fix unmonitored servers. (#7756 ) There is an inherent race condition in port allocation for spawned servers. If a server fails to start because a port is taken, a new port is allocated. This fixes a problem where the logs are not truncated and as a result a large number of unmonitored servers are started.	2020-09-07 17:30:36 +03:00
Oran Agra	42ba7a1b75	fix broken cluster/sentinel tests by recent commit (#7752 ) `2b998de46` added a file for stderr to keep valgrind log but i forgot to add a similar thing when valgrind isn't being used. the result is that `glob */err.txt` fails.	2020-09-07 16:26:11 +03:00
Oran Agra	573246f73c	if diskless repl child is killed, make sure to reap the pid (#7742 ) Starting redis 6.0 and the changes we made to the diskless master to be suitable for TLS, I made the master avoid reaping (wait3) the pid of the child until we know all replicas are done reading their rdb. I did that in order to avoid a state where the rdb_child_pid is -1 but we don't yet want to start another fork (still busy serving that data to replicas). It turns out that the solution used so far was problematic in case the fork child was being killed (e.g. by the kernel OOM killer), in that case there's a chance that we currently disabled the read event on the rdb pipe, since we're waiting for a replica to become writable again. and in that scenario the master would have never realized the child exited, and the replica will remain hung too. Note that there's no mechanism to detect a hung replica while it's in rdb transfer state. The solution here is to add another pipe which is used by the parent to tell the child it is safe to exit. this mean that when the child exits, for whatever reason, it is safe to reap it. Besides that, i'm re-introducing an adjustment to REPLCONF ACK which was part of #6271 (Accelerate diskless master connections) but was dropped when that PR was rebased after the TLS fork/pipe changes (`5a47794`). Now that RdbPipeCleanup no longer calls checkChildrenDone, and the ACK has chance to detect that the child exited, it should be the one to call it so that we don't have to wait for cron (server.hz) to do that.	2020-09-06 16:43:57 +03:00
Oran Agra	2b998de460	Improve valgrind support for cluster tests (#7725 ) - redirect valgrind reports to a dedicated file rather than console - try to avoid killing instances with SIGKILL so that we get the memory leak report (killing with SIGTERM before resorting to SIGKILL) - search for valgrind reports when done, print them and fail the tests - add --dont-clean option to keep the logs on exit - fix exit error code when crash is found (would have exited with 0) changes that affect the normal redis test suite: - refactor check_valgrind_errors into two functions one to search and one to report - move the search half into util.tcl to serve the cluster tests too - ignore "address range perms" valgrind warnings which seem non relevant.	2020-09-06 11:11:49 +03:00
Oran Agra	fe5da2e60d	test infra - add durable mode to work around test suite crashing in some cases a command that returns an error possibly due to a timing issue causes the tcl code to crash and thus prevents the rest of the tests from running. this adds an option to make the test proceed despite the crash. maybe it should be the default mode some day.	2020-09-06 09:59:19 +03:00
Oran Agra	1b7ba44e79	test infra - wait_done_loading reduce code duplication in aof.tcl. move creation of clients into the test so that it can be skipped	2020-09-06 09:59:19 +03:00
Oran Agra	b65e5aca86	test infra - flushall between tests in external mode	2020-09-06 09:59:19 +03:00
Oran Agra	677d14c213	test infra - improve test skipping ability - skip full units - skip a single test (not just a list of tests) - when skipping tag, skip spinning up servers, not just the tests - skip tags when running against an external server too - allow using multiple tags (split them)	2020-09-06 09:59:19 +03:00
Oran Agra	e3e69c25fd	test infra - reduce disk space usage this is important when running a test with --loop	2020-09-06 09:59:19 +03:00
Oran Agra	9d527d076b	test infra - write test name to logfile	2020-09-06 09:59:19 +03:00
Oran Agra	9ef8d2f671	Run active defrag while blocked / loading (#7726 ) During long running scripts or loading RDB/AOF, we may need to do some defragging. Since processEventsWhileBlocked is called periodically at unknown intervals, and many cron jobs either depend on run_with_period (including active defrag), or rely on being called at server.hz rate (i.e. active defrag knows ho much time to run by looking at server.hz), the whileBlockedCron may have to run a loop triggering the cron jobs in it (currently only active defrag) several times. Other changes: - Adding a test for defrag during aof loading. - Changing key-load-delay config to take negative values for fractions of a microsecond sleep	2020-09-03 08:47:29 +03:00
Yossi Gottlieb	b61b663895	Fix oom-score-adj on older distros. (#7724 ) Don't assume `ps` handles `-h` to display output without headers and manually trim headers line from output.	2020-08-30 12:23:47 +03:00
valentinogeron	b7289e912c	EXEC with only read commands should not be rejected when OOM (#7696 ) If the server gets MULTI command followed by only read commands, and right before it gets the EXEC it reaches OOM, the client will get OOM response. So, from now on, it will get OOM response only if there was at least one command that was tagged with `use-memory` flag	2020-08-27 09:19:24 +03:00
Oran Agra	daef1f00c2	Add test coverage for CLIENT UNBLOCK (#7712 ) plus minor other fixes to list.tcl	2020-08-27 08:09:39 +03:00
Valentino Geron	9204a9b2c2	Fix LPOS command when RANK is greater than matches When calling to LPOS command when RANK is higher than matches, the return value is non valid response. For example: ``` LPUSH l a :1 LPOS l b RANK 5 COUNT 10 -4 ``` It may break client-side parser. Now, we count how many replies were replied in the array. ``` LPUSH l a :1 LPOS l b RANK 5 COUNT 10 0 ```	2020-08-23 16:03:30 +03:00
Yossi Gottlieb	f80f3f492a	Tests: fix redis-cli with remote hosts. (#7693 )	2020-08-23 10:17:43 +03:00
杨博东	cbaf3c5bba	Fix flock cluster config may cause failure to restart after kill -9 (#7674 ) After fork, the child process(redis-aof-rewrite) will get the fd opened by the parent process(redis), when redis killed by kill -9, it will not graceful exit(call prepareForShutdown()), so redis-aof-rewrite thread may still alive, the fd(lock) will still be held by redis-aof-rewrite thread, and redis restart will fail to get lock, means fail to start. This issue was causing failures in the cluster tests in github actions. Co-authored-by: Oran Agra <oran@redislabs.com>	2020-08-20 08:59:02 +03:00
Yossi Gottlieb	64c360c515	Module API: fix missing RM_CLIENTINFO_FLAG_SSL. (#7666 ) The `REDISMODULE_CLIENTINFO_FLAG_SSL` flag was already a part of the `RedisModuleClientInfo` structure but was not implemented.	2020-08-17 17:46:54 +03:00
Yossi Gottlieb	2530dc0ebd	Add oom-score-adj configuration option to control Linux OOM killer. (#1690 ) Add Linux kernel OOM killer control option. This adds the ability to control the Linux OOM killer oom_score_adj parameter for all Redis processes, depending on the process role (i.e. master, replica, background child). A oom-score-adj global boolean flag control this feature. In addition, specific values can be configured using oom-score-adj-values if additional tuning is required.	2020-08-12 17:58:56 +03:00
Mota	ddcbb628a1	Test:Fix invalid cases in hash.tcl and dump.tcl (#4611 )	2020-08-12 10:25:24 +08:00
Tyson Andre	6f11acbd67	Implement SMISMEMBER key member [member ...] (#7615 ) This is a rebased version of #3078 originally by shaharmor with the following patches by TysonAndre made after rebasing to work with the updated C API: 1. Add 2 more unit tests (wrong argument count error message, integer over 64 bits) 2. Use addReplyArrayLen instead of addReplyMultiBulkLen. 3. Undo changes to src/help.h - for the ZMSCORE PR, I heard those should instead be automatically generated from the redis-doc repo if it gets updated Motivations: - Example use case: Client code to efficiently check if each element of a set of 1000 items is a member of a set of 10 million items. (Similar to reasons for working on #7593) - HMGET and ZMSCORE already exist. This may lead to developers deciding to implement functionality that's best suited to a regular set with a data type of sorted set or hash map instead, for the multi-get support. Currently, multi commands or lua scripting to call sismember multiple times would almost definitely be less efficient than a native smismember for the following reasons: - Need to fetch the set from the string every time instead of reusing the C pointer. - Using pipelining or multi-commands would result in more bytes sent and received by the client for the repeated SISMEMBER KEY sections. - Need to specially encode the data and decode it from the client for lua-based solutions. - Proposed solutions using Lua or SADD/SDIFF could trigger writes to memory, which is undesirable on a redis replica server or when commands get replicated to replicas. Co-Authored-By: Shahar Mor <shahar@peer5.com> Co-Authored-By: Tyson Andre <tysonandre775@hotmail.com>	2020-08-11 11:55:06 +03:00
Meir Shpilraien (Spielrein)	3f494cc49d	see #7544 , added RedisModule_HoldString api. (#7577 ) Added RedisModule_HoldString that either returns a shallow copy of the given String (by increasing the String ref count) or a new deep copy of String in case its not possible to get a shallow copy. Co-authored-by: Itamar Haber <itamar@redislabs.com>	2020-08-09 06:11:47 +03:00
Oran Agra	e2d64485b8	Reduce the probability of failure when start redis in runtest-cluster #7554 (#7635 ) When runtest-cluster, at first, we need to create a cluster use spawn_instance, a port which is not used is choosen, however sometimes we can't run server on the port. possibley due to a race with another process taking it first. such as redis/redis/runs/896537490. It may be due to the machine problem or In order to reduce the probability of failure when start redis in runtest-cluster, we attemp to use another port when find server do not start up. Co-authored-by: Oran Agra <oran@redislabs.com> Co-authored-by: yanhui13 <yanhui13@meituan.com>	2020-08-09 06:08:00 +03:00
Oran Agra	c17e597d05	Accelerate diskless master connections, and general re-connections (#6271 ) Diskless master has some inherent latencies. 1) fork starts with delay from cron rather than immediately 2) replica is put online only after an ACK. but the ACK was sent only once a second. 3) but even if it would arrive immediately, it will not register in case cron didn't yet detect that the fork is done. Besides that, when a replica disconnects, it doesn't immediately attempts to re-connect, it waits for replication cron (one per second). in case it was already online, it may be important to try to re-connect as soon as possible, so that the backlog at the master doesn't vanish. In case it disconnected during rdb transfer, one can argue that it's not very important to re-connect immediately, but this is needed for the "diskless loading short read" test to be able to run 100 iterations in 5 seconds, rather than 3 (waiting for replication cron re-connection) changes in this commit: 1) sync command starts a fork immediately if no sync_delay is configured 2) replica sends REPLCONF ACK when done reading the rdb (rather than on 1s cron) 3) when a replica unexpectedly disconnets, it immediately tries to re-connect rather than waiting 1s 4) when when a child exits, if there is another replica waiting, we spawn a new one right away, instead of waiting for 1s replicationCron. 5) added a call to connectWithMaster from replicationSetMaster. which is called from the REPLICAOF command but also in 3 places in cluster.c, in all of these the connection attempt will now be immediate instead of delayed by 1 second. side note: we can add a call to rdbPipeReadHandler in replconfCommand when getting a REPLCONF ACK from the replica to solve a race where the replica got the entire rdb and EOF marker before we detected that the pipe was closed. in the test i did see this race happens in one about of some 300 runs, but i concluded that this race is unlikely in real life (where the replica is on another host and we're more likely to first detect the pipe was closed. the test runs 100 iterations in 3 seconds, so in some cases it'll take 4 seconds instead (waiting for another REPLCONF ACK). Removing unneeded startBgsaveForReplication from updateSlavesWaitingForBgsave Now that CheckChildrenDone is calling the new replicationStartPendingFork (extracted from serverCron) there's actually no need to call startBgsaveForReplication from updateSlavesWaitingForBgsave anymore, since as soon as updateSlavesWaitingForBgsave returns, CheckChildrenDone is calling replicationStartPendingFork that handles that anyway. The code in updateSlavesWaitingForBgsave had a bug in which it ignored repl-diskless-sync-delay, but removing that code shows that this bug was hiding another bug, which is that the max_idle should have used >= and not >, this one second delay has a big impact on my new test.	2020-08-06 16:53:06 +03:00
WuYunlong	40c7628fc8	Fix tests/cluster/cluster.tcl about wrong usage of lrange. (#6702 )	2020-08-04 18:00:58 +03:00
Tyson Andre	f11f26cc53	Add a ZMSCORE command returning an array of scores. (#7593 ) Syntax: `ZMSCORE KEY MEMBER [MEMBER ...]` This is an extension of #2359 amended by Tyson Andre to work with the changed unstable API, add more tests, and consistently return an array. - It seemed as if it would be more likely to get reviewed after updating the implementation. Currently, multi commands or lua scripting to call zscore multiple times would almost definitely be less efficient than a native ZMSCORE for the following reasons: - Need to fetch the set from the string every time instead of reusing the C pointer. - Using pipelining or multi-commands would result in more bytes sent by the client for the repeated `ZMSCORE KEY` sections. - Need to specially encode the data and decode it from the client for lua-based solutions. - The fastest solution I've seen for large sets(thousands or millions) involves lua and a variadic ZADD, then a ZINTERSECT, then a ZRANGE 0 -1, then UNLINK of a temporary set (or lua). This is still inefficient. Co-authored-by: Tyson Andre <tysonandre775@hotmail.com>	2020-08-04 17:49:33 +03:00
Oran Agra	824bd2ac11	fix new rdb test failing on timing issues (#7604 ) apparenlty on github actions sometimes 500ms is not enough	2020-08-04 08:53:50 +03:00
Oran Agra	f7e7775990	module hook for master link up missing on successful psync (#7584 ) besides, hooks test was time sensitive. when the replica managed to reconnect quickly after the client kill, the test would fail	2020-07-31 13:14:29 +03:00
WuYunlong	f3352daf4f	Fix running single test 14-consistency-check.tcl (#7587 )	2020-07-30 08:56:21 +03:00
Yossi Gottlieb	bedf1b2126	Fix TLS cluster tests. (#7578 ) Fix consistency test added in `af5167b7f` without considering TLS redis-cli configuration.	2020-07-28 14:04:06 +03:00
Oran Agra	109b5ccdcd	Fix failing tests due to issues with wait_for_log_message (#7572 ) - the test now waits for specific set of log messages rather than wait for timeout looking for just one message. - we don't wanna sample the current length of the log after an action, due to a race, we need to start the search from the line number of the last message we where waiting for. - when attempting to trigger a full sync, use multi-exec to avoid a race where the replica manages to re-connect before we completed the set of actions that should force a full sync. - fix verify_log_message which was broken and unused	2020-07-28 11:15:29 +03:00
Jiayuan Chen	f31260b044	Add optional tls verification (#7502 ) Adds an `optional` value to the previously boolean `tls-auth-clients` configuration keyword. Co-authored-by: Yossi Gottlieb <yossigo@gmail.com>	2020-07-28 10:45:21 +03:00
Oran Agra	8a57969fd7	Stabilize bgsave test that sometimes fails with valgrind (#7559 ) on ci.redis.io the test fails a lot, reporting that bgsave didn't end. increaseing the timeout we wait for that bgsave to get aborted. in addition to that, i also verify that it indeed got aborted by checking that the save counter wasn't reset. add another test to verify that a successful bgsave indeed resets the change counter.	2020-07-23 13:06:24 +03:00
Meir Shpilraien (Spielrein)	8d82639319	This PR introduces a new loaded keyspace event (#7536 ) Co-authored-by: Oran Agra <oran@redislabs.com> Co-authored-by: Itamar Haber <itamar@redislabs.com>	2020-07-23 12:38:51 +03:00
Oran Agra	36b9494385	testsuite may leave servers alive on error (#7549 ) in cases where you have test name { start_server { start_server { assert } } } the exception will be thrown to the test proc, and the servers are supposed to be killed on the way out. but it seems there was always a bug of not cleaning the server stack, and recently (#7404) we started relying on that stack in order to kill them, so with that bug sometimes we would have tried to kill the same server twice, and leave one alive. luckly, in most cases the pattern is: start_server { test name { } }	2020-07-21 16:56:19 +03:00
Yossi Gottlieb	f57e844b2e	Tests: drop TCL 8.6 dependency. (#7548 ) This re-implements the redis-cli --pipe test so it no longer depends on a close feature available only in TCL 8.6. Basically what this test does is run redis-cli --pipe, generates a bunch of commands and pipes them through redis-cli, and inspects the result in both Redis and the redis-cli output. To do that, we need to close stdin for redis-cli to indicate we're done so it can flush its buffers and exit. TCL has bi-directional channels can only offers a way to "one-way close" a channel with TCL 8.6. To work around that, we now generate the commands into a file and feed that file to redis-cli directly. As we're writing to an actual file, the number of commands is now reduced.	2020-07-21 14:17:14 +03:00
Remi Collet	3f2fbc4c61	Fix deprecated tail syntax in tests (#7543 )	2020-07-21 09:07:54 +03:00
WuYunlong	93bdbf5aa4	Fix command help for unexpected options (#7476 )	2020-07-15 12:38:22 +03:00
Oran Agra	254c962554	redis-cli tests, fix valgrind timing issue (#7519 ) this test when run with valgrind on github actions takes 160 seconds	2020-07-14 18:04:08 +03:00
WuYunlong	8128d39737	Fix out of update help info in tcl tests. (#7516 ) Before this commit, the output of "./runtest-cluster --help" is incorrect. After this commit, the format of the following 3 output is consistent: ./runtest --help ./runtest-cluster --help ./runtest-sentinel --help	2020-07-14 11:35:04 +03:00
Oran Agra	e5227aab89	fix recently added time sensitive tests failing with valgrind (#7512 ) interestingly the latency monitor test fails because valgrind is slow enough so that the time inside PEXPIREAT command from the moment of the first mstime() call to get the basetime until checkAlreadyExpired calls mstime() again is more than 1ms, and that test was too sensitive. using this opportunity to speed up the test (unrelated to the failure) the fix is just the longer time passed to PEXPIRE.	2020-07-13 16:40:03 +03:00
Oran Agra	02ef355f98	runtest --stop pause stops before terminating the redis server (#7513 ) in the majority of the cases (on this rarely used feature) we want to stop and be able to connect to the shard with redis-cli. since these are two different processes interracting with the tty we need to stop both, and we'll have to hit enter twice, but it's not that bad considering it is rarely used.	2020-07-13 16:09:08 +03:00
WuYunlong	d792db7948	Add missing latency-monitor tcl test to test_helper.tcl. (#6782 )	2020-07-10 11:41:48 +03:00
Yossi Gottlieb	3e6f2b1a45	TLS: Session caching configuration support. (#7420 ) * TLS: Session caching configuration support. * TLS: Remove redundant config initialization.	2020-07-10 11:33:47 +03:00
Yossi Gottlieb	d9f970d8d3	TLS: Add missing redis-cli options. (#7456 ) * Tests: fix and reintroduce redis-cli tests. These tests have been broken and disabled for 10 years now! * TLS: add remaining redis-cli support. This adds support for the redis-cli --pipe, --rdb and --replica options previously unsupported in --tls mode. * Fix writeConn().	2020-07-10 10:25:55 +03:00
Oran Agra	5977a94842	RESTORE ABSTTL won't store expired keys into the db (#7472 ) Similarly to EXPIREAT with TTL in the past, which implicitly deletes the key and return success, RESTORE should not store key that are already expired into the db. When used together with REPLACE it should emit a DEL to keyspace notification and replication stream.	2020-07-10 10:02:37 +03:00
Oran Agra	909bc97c52	skip a test that uses +inf on valgrind (#7440 ) On some platforms strtold("+inf") with valgrind returns a non-inf result [err]: INCRBYFLOAT does not allow NaN or Infinity in tests/unit/type/incr.tcl Expected 'ERRwould produce' to equal or match '1189731495357231765085759.....'	2020-07-10 08:29:02 +03:00
Oran Agra	8e76e13472	stabilize tests that look for log lines (#7367 ) tests were sensitive to additional log lines appearing in the log causing the search to come empty handed. instead of just looking for the n last log lines, capture the log lines before performing the action, and then search from that offset.	2020-07-10 08:28:22 +03:00
Oran Agra	69ade87325	tests/valgrind: don't use debug restart (#7404 ) * tests/valgrind: don't use debug restart DEBUG REATART causes two issues: 1. it uses execve which replaces the original process and valgrind doesn't have a chance to check for errors, so leaks go unreported. 2. valgrind report invalid calls to close() which we're unable to resolve. So now the tests use restart_server mechanism in the tests, that terminates the old server and starts a new one, new PID, but same stdout, stderr. since the stderr can contain two or more valgrind report, it is not enough to just check for the absence of leaks, we also need to check for some known errors, we do both, and fail if we either find an error, or can't find a report saying there are no leaks. other changes: - when killing a server that was already terminated we check for leaks too. - adding DEBUG LEAK which was used to test it. - adding --trace-children to valgrind, although no longer needed. - since the stdout contains two or more runs, we need slightly different way of checking if the new process is up (explicitly looking for the new PID) - move the code that handles --wait-server to happen earlier (before watching the startup message in the log), and serve the restarted server too. * squashme - CR fixes	2020-07-10 08:26:52 +03:00
antirez	959cdb358b	Merge branch 'unstable' of github.com:/antirez/redis into unstable	2020-06-24 09:09:59 +02:00
antirez	a5a3a7bbc6	LPOS: option FIRST renamed RANK.	2020-06-24 09:09:43 +02:00
Salvatore Sanfilippo	6bbbdd26f4	Merge pull request #7390 from oranagra/exec_fails_abort EXEC always fails with EXECABORT and multi-state is cleared	2020-06-23 13:12:52 +02:00
Oran Agra	65a3307bc9	EXEC always fails with EXECABORT and multi-state is cleared In order to support the use of multi-exec in pipeline, it is important that MULTI and EXEC are never rejected and it is easy for the client to know if the connection is still in multi state. It was easy to make sure MULTI and DISCARD never fail (done by previous commits) since these only change the client state and don't do any actual change in the server, but EXEC is a different story. Since in the past, it was possible for clients to handle some EXEC errors and retry the EXEC, we now can't affort to return any error on EXEC other than EXECABORT, which now carries with it the real reason for the abort too. Other fixes in this commit: - Some checks that where performed at the time of queuing need to be re- validated when EXEC runs, for instance if the transaction contains writes commands, it needs to be aborted. there was one check that was already done in execCommand (-READONLY), but other checks where missing: -OOM, -MISCONF, -NOREPLICAS, -MASTERDOWN - When a command is rejected by processCommand it was rejected with addReply, which was not recognized as an error in case the bad command came from the master. this will enable to count or MONITOR these errors in the future. - make it easier for tests to create additional (non deferred) clients. - add tests for the fixes of this commit.	2020-06-23 12:01:33 +03:00
meir@redislabs.com	a89bf734a9	Fix RM_ScanKey module api not to return int encoded strings The scan key module API provides the scan callback with the current field name and value (if it exists). Those arguments are RedisModuleString* which means it supposes to point to robj which is encoded as a string. Using createStringObjectFromLongLong function might return robj that points to an integer and so break a module that tries for example to use RedisModule_StringPtrLen on the given field/value. The PR introduces a fix that uses the createObject function and sdsfromlonglong function. Using those function promise that the field and value pass to the to the scan callback will be Strings. The PR also changes the Scan test module to use RedisModule_StringPtrLen to catch the issue. without this, the issue is hidden because RedisModule_ReplyWithString knows to handle integer encoding of the given robj (RedisModuleString). The PR also introduces a new test to verify the issue is solved.	2020-06-14 11:20:15 +03:00
antirez	0091125cae	LPOS: tests + crash fix.	2020-06-11 12:39:06 +02:00
antirez	2ebcd63d6a	Adapt EVAL+busy script test to new behavior.	2020-06-09 12:19:14 +02:00
zhaozhao.zz	5d0774d62f	AOF: append origin SET if no expire option	2020-06-03 17:55:18 +08:00
Oran Agra	c480af9007	fix pingoff test race	2020-05-31 15:51:52 +03:00
antirez	23f2b4d0a8	Test: take PSYNC2 test master timeout high during switch. This will likely avoid false positives due to trailing pings.	2020-05-28 10:47:30 +02:00
antirez	9d38a270db	Test: add the tracking unit as default.	2020-05-28 10:22:22 +02:00
Salvatore Sanfilippo	bb06a8c870	Merge pull request #7327 from oranagra/test-port-ranges tests: each test client work on a distinct port range	2020-05-28 09:52:10 +02:00
Salvatore Sanfilippo	28e574bc4b	Merge pull request #7336 from oranagra/modules_ci_32bit 32bit CI needs to build modules correctly	2020-05-28 09:51:58 +02:00
Oran Agra	2a8af8e675	adjust revived meaningful offset tests these tests create several edge cases that are otherwise uncovered (at least not consistently) by the test suite, so although they're no longer testing what they were meant to test, it's still a good idea to keep them in hope that they'll expose some issue in the future.	2020-05-28 09:10:51 +03:00
Oran Agra	90f3856fd5	revive meaningful offset tests	2020-05-28 08:21:24 +03:00
Oran Agra	15bcb813d4	32bit CI needs to build modules correctly	2020-05-27 18:19:30 +03:00
Oran Agra	1cf33a46d5	tests: find_available_port start search from next port i.e. don't start the search from scratch hitting the used ones again. this will also reduce the likelihood of collisions (if there are any left) by increasing the time until we re-use a port we did use in the past.	2020-05-27 16:12:35 +03:00
antirez	484cfc3d76	Another meaningful offset test removed.	2020-05-27 12:50:02 +02:00
antirez	32d0df0c1f	Remove the PSYNC2 meaningful offset test.	2020-05-27 12:47:34 +02:00
Oran Agra	e258a1c087	tests: each test client work on a distinct port range apparently when running tests in parallel (the default of --clients 16), there's a chance for two tests to use the same port. specifically, one test might shutdown a master and still have the replica up, and then another test will re-use the port number of master for another master, and then that replica will connect to the master of the other test. this can cause a master to count too many full syncs and fail a test if we run the tests with --single integration/psync2 --loop --stop see Probmem 2 in #7314	2020-05-26 11:17:08 +03:00
antirez	091fb64681	Test: PSYNC2 test can now show server logs.	2020-05-25 20:26:29 +02:00
Qu Chen	42f5da5d2d	Disconnect chained replicas when the replica performs PSYNC with the master always to avoid replication offset mismatch between master and chained replicas.	2020-05-21 18:42:10 -07:00
Oran Agra	88d71f4793	fix a rare active defrag edge case bug leading to stagnation There's a rare case which leads to stagnation in the defragger, causing it to keep scanning the keyspace and do nothing (not moving any allocation), this happens when all the allocator slabs of a certain bin have the same % utilization, but the slab from which new allocations are made have a lower utilization. this commit fixes it by removing the current slab from the overall average utilization of the bin, and also eliminate any precision loss in the utilization calculation and move the decision about the defrag to reside inside jemalloc. and also add a test that consistently reproduce this issue.	2020-05-20 16:04:42 +03:00
Salvatore Sanfilippo	9fba05f758	Merge pull request #7232 from trevor211/handleHashTagWhenComputingHashSlot Tcl client support hash tagged keys.	2020-05-19 09:23:44 +02:00
Oran Agra	75c11d7fec	fix valgrind test failure in replication test in `b4416280c` i added more keys to that test to make it run longer but in valgrind this now means the test times out, give valgrind more time.	2020-05-18 10:26:53 +03:00
antirez	0ca2f4f824	Merge branch 'unstable' of github.com:/antirez/redis into unstable	2020-05-17 18:24:48 +02:00
antirez	96bb0c9471	Improve the PSYNC2 test reliability.	2020-05-17 18:24:34 +02:00
Oran Agra	357aace895	add regression test for the race in #7205 with the original version of 6.0.0, this test detects an excessive full sync. with the fix in `1a7cd2c0e`, this test detects memory corruption, especially when using libc allocator with or without valgrind.	2020-05-17 18:26:02 +03:00
Salvatore Sanfilippo	8f5c2bc8aa	Merge pull request #7229 from yossigo/tls-fails-on-recent-debian TLS: Fix test failures on recent Debian/Ubuntu.	2020-05-14 18:15:17 +02:00
antirez	c38fd1f661	Merge branch 'free_clients_during_loading' into unstable	2020-05-14 11:28:08 +02:00
antirez	dec6fd3adc	Regression test for #7249 .	2020-05-14 11:27:31 +02:00
Oran Agra	b4416280cf	fix unstable replication test this test which has coverage for varoius flows of diskless master was failing randomly from time to time. the failure was: [err]: diskless all replicas drop during rdb pipe in tests/integration/replication.tcl log message of 'Diskless rdb transfer, last replica dropped, killing fork child' not found what seemed to have happened is that the master didn't detect that all replicas dropped by the time the replication ended, it thought that one replica is still connected. now the test takes a few seconds longer but it seems stable.	2020-05-12 08:59:09 +03:00
Oran Agra	905e28ee87	fix redis 6.0 not freeing closed connections during loading. This bug was introduced by a recent change in which readQueryFromClient is using freeClientAsync, and despite the fact that now freeClientsInAsyncFreeQueue is in beforeSleep, that's not enough since it's not called during loading in processEventsWhileBlocked. furthermore, afterSleep was called in that case but beforeSleep wasn't. This bug also caused slowness sine the level-triggered mode of epoll kept signaling these connections as readable causing us to keep doing connRead again and again for ll of these, which keep accumulating. now both before and after sleep are called, but not all of their actions are performed during loading, some are only reserved for the main loop. fixes issue #7215	2020-05-11 11:33:46 +03:00
WuYunlong	70c4851f2c	Handle keys with hash tag when computing hash slot using tcl cluster client.	2020-05-11 13:14:18 +08:00
WuYunlong	bc9efb577b	Add a test to prove current tcl cluster client can not handle keys with hash tag.	2020-05-11 13:14:18 +08:00
Yossi Gottlieb	4d1178cc24	TLS: Fix test failures on recent Debian/Ubuntu. Seems like on some systems choosing specific TLS v1/v1.1 versions no longer works as expected. Test is reduced for v1.2 now which is still good enough to test the mechansim, and matters most anyway.	2020-05-10 17:38:04 +03:00
antirez	e49d97298a	Test: --dont-clean should do first cleanup.	2020-05-05 13:18:53 +02:00
Salvatore Sanfilippo	acf566b291	Merge pull request #7179 from bytedance/cpu-affinity Support setcpuaffinity on linux/bsd	2020-05-04 10:56:20 +02:00
Oran Agra	deee2c1ef2	add daily github actions with libc malloc and valgrind * fix memlry leaks with diskless replica short read. * fix a few timing issues with valgrind runs * fix issue with valgrind and watchdog schedule signal about the valgrind WD issue: the stack trace test in logging.tcl, has issues with valgrind: ==28808== Can't extend stack to 0x1ffeffdb38 during signal delivery for thread 1: ==28808== too small or bad protection modes it seems to be some valgrind bug with SA_ONSTACK. SA_ONSTACK seems unneeded since WD is not recursive (SA_NODEFER was removed), also, not sure if it's even valid without a call to sigaltstack()	2020-05-04 09:52:20 +03:00
zhenwei pi	1a0deab2a5	Support setcpuaffinity on linux/bsd Currently, there are several types of threads/child processes of a redis server. Sometimes we need deeply optimise the performance of redis, so we would like to isolate threads/processes. There were some discussion about cpu affinity cases in the issue: https://github.com/antirez/redis/issues/2863 So implement cpu affinity setting by redis.conf in this patch, then we can config server_cpulist/bio_cpulist/aof_rewrite_cpulist/ bgsave_cpulist by cpu list. Examples of cpulist in redis.conf: server_cpulist 0-7:2 means cpu affinity 0,2,4,6 bio_cpulist 1,3 means cpu affinity 1,3 aof_rewrite_cpulist 8-11 means cpu affinity 8,9,10,11 bgsave_cpulist 1,10-11 means cpu affinity 1,10,11 Test on linux/freebsd, both work fine. Signed-off-by: zhenwei pi <pizhenwei@bytedance.com>	2020-05-02 21:19:47 +08:00
Salvatore Sanfilippo	e0de7c0852	Merge pull request #7134 from guybe7/xstate_command Extend XINFO STREAM output	2020-04-28 16:31:00 +02:00
Guy Benoish	1e2aee3919	Extend XINFO STREAM output Introducing XINFO STREAM <key> FULL	2020-04-28 13:03:43 +03:00
Oran Agra	d31c0c5264	fix loading race in psync2 tests	2020-04-28 09:18:01 +03:00
Oran Agra	4447ddc8bb	Keep track of meaningful replication offset in replicas too Now both master and replicas keep track of the last replication offset that contains meaningful data (ignoring the tailing pings), and both trim that tail from the replication backlog, and the offset with which they try to use for psync. the implication is that if someone missed some pings, or even have excessive pings that the promoted replica has, it'll still be able to psync (avoid full sync). the downside (which was already committed) is that replicas running old code may fail to psync, since the promoted replica trims pings form it's backlog. This commit adds a test that reproduces several cases of promotions and demotions with stale and non-stale pings Background: The mearningful offset on the master was added recently to solve a problem were the master is left all alone, injecting PINGs into it's backlog when no one is listening and then gets demoted and tries to replicate from a replica that didn't have any of the PINGs (or at least not the last ones). however, consider this case: master A has two replicas (B and C) replicating directly from it. there's no traffic at all, and also no network issues, just many pings in the tail of the backlog. now B gets promoted, A becomes a replica of B, and C remains a replica of A. when A gets demoted, it trims the pings from its backlog, and successfully replicate from B. however, C is still aware of these PINGs, when it'll disconnect and re-connect to A, it'll ask for something that's not in the backlog anymore (since A trimmed the tail of it's backlog), and be forced to do a full sync (something it didn't have to do before the meaningful offset fix). Besides that, the psync2 test was always failing randomly here and there, it turns out the reason were PINGs. Investigating it shows the following scenario: cycle 1: redis #1 is master, and all the rest are direct replicas of #1 cycle 2: redis #2 is promoted to master, #1 is a replica of #2 and #3 is replica of #1 now we see that when #1 is demoted it prints: 17339:S 21 Apr 2020 11:16:38.523 * Using the meaningful offset 3929963 instead of 3929977 to exclude the final PINGs (14 bytes difference) 17339:S 21 Apr 2020 11:16:39.391 * Trying a partial resynchronization (request e2b3f8817735fdfe5fa4626766daa938b61419e5:3929964). 17339:S 21 Apr 2020 11:16:39.392 * Successful partial resynchronization with master. and when #3 connects to the demoted #2, #2 says: 17339:S 21 Apr 2020 11:16:40.084 * Partial resynchronization not accepted: Requested offset for secondary ID was 3929978, but I can reply up to 3929964 so the issue here is that the meaningful offset feature saved the day for the demoted master (since it needs to sync from a replica that didn't get the last ping), but it didn't help one of the other replicas which did get the last ping.	2020-04-27 15:52:23 +02:00
antirez	022f09447b	Merge branch 'unstable' of github.com:/antirez/redis into unstable	2020-04-24 16:59:56 +02:00
antirez	8a7f255cd0	LCS -> STRALGO LCS. STRALGO should be a container for mostly read-only string algorithms in Redis. The algorithms should have two main characteristics: 1. They should be non trivial to compute, and often not part of programming language standard libraries. 2. They should be fast enough that it is a good idea to have optimized C implementations. Next thing I would love to see? A small strings compression algorithm.	2020-04-24 16:54:32 +02:00
Salvatore Sanfilippo	42d309fffc	Merge pull request #7114 from guybe7/stream_tag_xsetid Add the stream tag to XSETID tests	2020-04-23 16:29:46 +02:00
Salvatore Sanfilippo	72f0751905	Merge pull request #7123 from fayadexinqing/optimizeClusterSlots Optimize the command of cluster slots	2020-04-23 16:18:22 +02:00
antirez	8d67211450	Tracking: test expired keys notifications.	2020-04-22 11:45:34 +02:00
antirez	58d61dd639	Tracking: NOLOOP tests.	2020-04-22 11:24:19 +02:00
yanhui13	6b547c3956	add tcl test for cluster slots	2020-04-21 16:56:10 +08:00
Guy Benoish	1bc557c9c5	Add the stream tag to XSETID tests	2020-04-19 15:59:58 +03:00
antirez	002052f8de	A few comments and name changes for #7103 .	2020-04-17 10:51:12 +02:00
Oran Agra	b9fa42a197	testsuite run the defrag latency test solo this test is time sensitive and it sometimes fail to pass below the latency threshold, even on strong machines. this test was the reson we're running just 2 parallel tests in the github actions CI, revering this.	2020-04-16 18:09:22 +03:00
antirez	121c51f4f3	Merge branch 'lcs' into unstable	2020-04-06 13:51:55 +02:00
antirez	af3c722fec	LCS: more tests.	2020-04-06 13:51:49 +02:00
antirez	8dc28b6c75	LCS tests.	2020-04-06 13:45:37 +02:00
Oran Agra	cf3789f045	diffrent fix for runtest --host --port	2020-04-06 09:41:14 +03:00
Guy Benoish	1b0d30aeb7	Try to fix time-sensitive tests in blockonkey.tcl There is an inherent race between the deferring client and the "main" client of the test: While the deferring client issues a blocking command, we can't know for sure that by the time the "main" client tries to issue another command (Usually one that unblocks the deferring client) the deferring client is even blocked... For lack of a better choice this commit uses TCL's 'after' in order to give some time for the deferring client to issues its blocking command before the "main" client does its thing. This problem probably exists in many other tests but this commit tries to fix blockonkeys.tcl	2020-04-03 14:51:45 +03:00
Salvatore Sanfilippo	cbf212f981	Merge pull request #7030 from valentinogeron/xread-in-lua XREAD and XREADGROUP should not be allowed from scripts when BLOCK op…	2020-04-03 11:14:13 +02:00
Guy Benoish	4665b3ebfb	Fix no-negative-zero test	2020-04-02 18:41:29 +03:00
Salvatore Sanfilippo	10b626b3d5	Merge pull request #6546 from guybe7/fix_neg_zero Make sure Redis does not reply with negative zero	2020-04-02 16:26:57 +02:00
Salvatore Sanfilippo	dfef407499	Merge pull request #7029 from valentinogeron/fix-xack XACK should be executed in a "all or nothing" fashion.	2020-04-02 11:23:23 +02:00
Guy Benoish	c4dc5b80b2	Fix memory corruption in moduleHandleBlockedClients By using a "circular BRPOPLPUSH"-like scenario it was possible the get the same client on db->blocking_keys twice (See comment in moduleTryServeClientBlockedOnKey) The fix was actually already implememnted in moduleTryServeClientBlockedOnKey but it had a bug: the funxction should return 0 or 1 (not OK or ERR) Other changes: 1. Added two commands to blockonkeys.c test module (To reproduce the case described above) 2. Simplify blockonkeys.c in order to make testing easier 3. cast raxSize() to avoid warning with format spec	2020-04-01 12:53:26 +03:00
Salvatore Sanfilippo	0c52ce6c8e	Merge pull request #7037 from guybe7/fix_module_replicate_multi Modules: Test MULTI/EXEC replication of RM_Replicate	2020-03-31 17:00:57 +02:00
Guy Benoish	6c8221580c	RENAME can unblock XREADGROUP Other changes: Support stream in serverLogObjectDebugInfo	2020-03-31 17:41:10 +03:00
Guy Benoish	d6eb3afd13	Modules: Test MULTI/EXEC replication of RM_Replicate Makse sure call() doesn't wrap replicated commands with a redundant MULTI/EXEC Other, unrelated changes: 1. Formatting compiler warning in INFO CLIENTS 2. Use CLIENT_ID_AOF instead of UINT64_MAX	2020-03-31 13:55:51 +03:00
antirez	4379b8b411	Fix the propagate Tcl test after module changes.	2020-03-31 12:09:38 +02:00
antirez	95f154985c	Modify the propagate unit test to show more cases.	2020-03-31 12:04:06 +02:00
antirez	9dcf878f1b	Fix module commands propagation double MULTI bug. `37a10cef` introduced automatic wrapping of MULTI/EXEC for the alsoPropagate API. However this collides with the built-in mechanism already present in module.c. To avoid complex changes near Redis 6 GA this commit introduces the ability to exclude call() MUTLI/EXEC wrapping for also propagate in order to continue to use the old code paths in module.c.	2020-03-31 11:00:45 +02:00
Valentino Geron	9a1843ef2d	XREAD and XREADGROUP should not be allowed from scripts when BLOCK option is being used	2020-03-26 15:46:31 +02:00
Valentino Geron	1547d72cf3	XACK should be executed in a "all or nothing" fashion. First, we must parse the IDs, so that we abort ASAP. The return value of this command cannot be an error if the client successfully acknowledged some messages, so it should be executed in a "all or nothing" fashion.	2020-03-26 15:40:23 +02:00
Salvatore Sanfilippo	2ea7f0ecad	Merge pull request #6644 from oranagra/stream_aofrw AOFRW on an empty stream created with MKSTREAM loads badkly	2020-03-26 11:12:44 +01:00
Oran Agra	3b29556a0c	AOFRW on an empty stream created with MKSTREAM loads badkly the AOF will be loaded successfully, but the stream will be missing, i.e inconsistencies with the original db. this was because XADD with id of 0-0 would error. add a test to reproduce.	2020-03-25 21:47:57 +02:00
antirez	c4d7f30e25	PSYNC2: meaningful offset test.	2020-03-25 15:43:34 +01:00
Oran Agra	ec007559ff	MULTI/EXEC during LUA script timeout are messed up Redis refusing to run MULTI or EXEC during script timeout may cause partial transactions to run. 1) if the client sends MULTI+commands+EXEC in pipeline without waiting for response, but these arrive to the shards partially while there's a busy script, and partially after it eventually finishes: we'll end up running only part of the transaction (since multi was ignored, and exec would fail). 2) similar to the above if EXEC arrives during busy script, it'll be ignored and the client state remains in a transaction. the 3rd test which i added for a case where MULTI and EXEC are ok, and only the body arrives during busy script was already handled correctly since processCommand calls flagTransaction	2020-03-23 20:45:32 +02:00
antirez	61de1c1146	Fix BITFIELD_RO test.	2020-03-23 12:02:12 +01:00
Salvatore Sanfilippo	493a7f9823	Merge pull request #6951 from yangbodong22011/feature-bitfield-ro Added BITFIELD_RO variants for read-only operations.	2020-03-23 11:23:21 +01:00
antirez	1e16b9384d	Merge branch 'unstable' of github.com:/antirez/redis into unstable	2020-03-20 13:21:28 +01:00
antirez	5497a44037	Regression test for #7011 .	2020-03-20 12:52:06 +01:00
WuYunlong	af5167b7f3	Add 14-consistency-check.tcl to prove there is a data consistency issue.	2020-03-18 16:17:46 +08:00
bodong.ybd	336458d4b5	Fix bug of tcl test using external server	2020-03-11 21:01:27 +08:00
Oran Agra	27641ee490	fix for flaky psync2 test *** [err]: PSYNC2: total sum of full synchronizations is exactly 4 in tests/integration/psync2.tcl Expected 5 == 4 (context: type eval line 6 cmd {assert {$sum == 4}} proc ::test) issue was that sometime the test got an unexpected full sync since it tried to switch to the replica before it was in sync with it's master.	2020-03-05 16:55:14 +02:00
bodong.ybd	94376f46ad	Added BITFIELD_RO variants for read-only operations.	2020-03-04 20:51:45 +08:00
Salvatore Sanfilippo	d2c5f80e2e	Merge pull request #6926 from oranagra/fork-test-fix fix race in module api test for fork	2020-02-27 09:58:04 +01:00
Oran Agra	2f1a1c3835	fix github actions failing latency test for active defrag - part 2 it seems that running two clients at a time is ok too, resuces action time from 20 minutes to 10. we'll use this for now, and if one day it won't be enough we'll have to run just the sensitive tests one by one separately from the others. this commit also fixes an issue with the defrag test that appears to be very rare.	2020-02-27 08:34:53 +02:00
Oran Agra	537893420b	fix github actions failing latency test for active defrag seems that github actions are slow, using just one client to reduce false positives. also adding verbose, testing only on latest ubuntu, and building on older one. when doing that, i can reduce the test threshold back to something saner	2020-02-25 17:53:23 +02:00
Salvatore Sanfilippo	3fbb41ecc9	Merge pull request #6920 from oranagra/defrag-test-latency-fix Fix latency sensitivity of new defrag test	2020-02-24 11:53:32 +01:00
antirez	73305861f5	Test engine: experimental change to avoid busy port problems.	2020-02-24 10:46:23 +01:00
Oran Agra	0a643efa0c	fix race in module api test for fork in some cases we were trying to kill the fork before it got created	2020-02-23 16:48:37 +02:00
Oran Agra	62adabd0e0	Fix latency sensitivity of new defrag test I saw that the new defag test for list was failing in CI recently, so i reduce it's threshold from 12 to 60. besides that, i add / improve the latency test for that other two defrag tests (add a sensitive latency and digest / save checks) and fix bad usage of debug populate (can't overrides existing keys). this was the original intention, which creates higher fragmentation.	2020-02-23 13:05:52 +02:00
antirez	e78c4e813c	Test engine: detect timeout when checking for Redis startup.	2020-02-21 18:55:56 +01:00
antirez	c6954de3ea	Test engine: better tracking of what workers are doing.	2020-02-21 17:08:45 +01:00
antirez	8a14fff545	Merge branch 'unstable' of github.com:/antirez/redis into unstable	2020-02-21 13:48:52 +01:00
antirez	a8d70ac568	Test is more complex now, increase default timeout.	2020-02-21 13:48:43 +01:00
Salvatore Sanfilippo	c552fad6d4	Merge pull request #6864 from guybe7/fix_memleak_in_test_ld_conv Fix memory leak in test_ld_conv	2020-02-20 13:08:31 +01:00
Salvatore Sanfilippo	e741b0c257	Merge pull request #6903 from oranagra/defrag_lists Defrag big lists in portions to avoid latency and freeze	2020-02-20 13:00:39 +01:00
Guy Benoish	770cb0ba97	XGROUP DESTROY should unblock XREADGROUP with -NOGROUP	2020-02-19 08:25:31 +05:30
Oran Agra	485425cec7	Defrag big lists in portions to avoid latency and freeze When active defrag kicks in and finds a big list, it will create a bookmark to a node so that it is able to resume iteration from that node later. The quicklist manages that bookmark, and updates it in case that node is deleted. This will increase memory usage only on lists of over 1000 (see active-defrag-max-scan-fields) quicklist nodes (1000 ziplists, not 1000 items) by 16 bytes. In 32 bit build, this change reduces the maximum effective config of list-compress-depth and list-max-ziplist-size (from 32767 to 8191)	2020-02-18 17:22:32 +02:00
antirez	8ea7a3ee68	Tracking: first set of tests for the feature.	2020-02-14 14:29:00 +01:00
Guy Benoish	5c73a6e206	Fix memory leak in test_ld_conv	2020-02-06 18:36:21 +05:30
antirez	d5c6a833c8	Merge branch 'acl-log' into unstable	2020-02-06 11:24:16 +01:00
antirez	90fae58b49	ACL LOG: make max log entries configurable.	2020-02-04 13:19:40 +01:00
antirez	64a73e9293	ACL LOG: test for AUTH reason.	2020-02-04 12:58:48 +01:00
WuYunlong	01eaf53bb3	Add tcl regression test in scripting.tcl to reproduce memory leak.	2020-02-04 16:34:11 +08:00
Guy Benoish	2deb55512f	ld2string should fail if string contains \0 in the middle This bug affected RM_StringToLongDouble and HINCRBYFLOAT. I added tests for both cases. Main changes: 1. Fixed string2ld to fail if string contains \0 in the middle 2. Use string2ld in getLongDoubleFromObject - No point of having duplicated code here The two changes above broke RM_SaveLongDouble/RM_LoadLongDouble because the long double string was saved with length+1 (An innocent mistake, but it's actually a bug - The length passed to RM_SaveLongDouble should not include the last \0).	2020-01-30 18:15:17 +05:30
antirez	b189a21974	ACL LOG: implement a few basic tests.	2020-01-30 11:14:13 +01:00
Salvatore Sanfilippo	bf53f9280a	Merge pull request #6699 from guybe7/module_blocked_on_key_timeout_memleak Modules: Fix blocked-client-related memory leak	2020-01-29 12:06:14 +01:00
Salvatore Sanfilippo	bb93686754	Merge pull request #6703 from guybe7/blocking_xread_empty_reply Blocking XREAD[GROUP] should always reply with valid data (or timeout)	2020-01-09 17:32:14 +01:00
Guy Benoish	d7d13721d3	Modules: Fix blocked-client-related memory leak If a blocked module client times-out (or disconnects, unblocked by CLIENT command, etc.) we need to call moduleUnblockClient in order to free memory allocated by the module sub-system and blocked-client private data Other changes: Made blockedonkeys.tcl tests a bit more aggressive in order to smoke-out potential memory leaks	2019-12-30 10:10:59 +05:30
Guy Benoish	a351e74fe9	Blocking XREAD[GROUP] should always reply with valid data (or timeout) This commit solves the following bug: 127.0.0.1:6379> XGROUP CREATE x grp $ MKSTREAM OK 127.0.0.1:6379> XADD x 666 f v "666-0" 127.0.0.1:6379> XREADGROUP GROUP grp Alice BLOCK 0 STREAMS x > 1) 1) "x" 2) 1) 1) "666-0" 2) 1) "f" 2) "v" 127.0.0.1:6379> XADD x 667 f v "667-0" 127.0.0.1:6379> XDEL x 667 (integer) 1 127.0.0.1:6379> XREADGROUP GROUP grp Alice BLOCK 0 STREAMS x > 1) 1) "x" 2) (empty array) The root cause is that we use s->last_id in streamCompareID while we should use the last valid ID	2019-12-30 10:06:01 +05:30
Salvatore Sanfilippo	3ff95d9074	Merge pull request #6706 from guybe7/stream_id_edge_cases Stream: Handle streamID-related edge cases	2019-12-29 14:53:06 +01:00
Oran Agra	0c3fe52ef7	config.c adjust config limits and mutable - make lua-replicate-commands mutable (it never was, but i don't see why) - make tcp-backlog immutable (fix a recent refactory mistake) - increase the max limit of a few configs to match what they were before the recent refactory	2019-12-26 15:16:15 +02:00
Guy Benoish	1f75ce30df	Stream: Handle streamID-related edge cases This commit solves several edge cases that are related to exhausting the streamID limits: We should correctly calculate the succeeding streamID instead of blindly incrementing 'seq' This affects both XREAD and XADD. Other (unrelated) changes: Reply with a better error message when trying to add an entry to a stream that has exhausted last_id	2019-12-26 15:31:37 +05:30
antirez	e6e58e455c	Revert "Geo: output 10 chars of geohash, not 11." This reverts commit `009862ab7e`.	2019-12-18 12:54:46 +01:00
zhaozhao.zz	58554396d6	incrbyfloat: fix issue #5256 ttl lost after propagate	2019-12-18 15:44:51 +08:00
zhaozhao.zz	24044f3356	add a new SET option KEEPTTL that doesn't remove expire time	2019-12-18 15:20:36 +08:00
Madelyn Olson	a511a37bb7	Fixed some documentation	2019-12-17 07:49:21 +00:00
Madelyn Olson	67aa527b22	Added some documentation and fixed a test	2019-12-17 07:15:04 +00:00
Madelyn Olson	034dcf185c	Add module APIs for custom authentication	2019-12-17 06:59:59 +00:00
Yossi Gottlieb	0283db5883	Improve RM_ModuleTypeReplaceValue() API. With the previous API, a NULL return value was ambiguous and could represent either an old value of NULL or an error condition. The new API returns a status code and allows the old value to be returned by-reference. This commit also includes test coverage based on tests/modules/datatype.c which did not exist at the time of the original commit.	2019-12-12 18:50:11 +02:00
Oran Agra	5661b19005	fix crash in module short read test	2019-12-02 19:17:35 +02:00
Salvatore Sanfilippo	ce7ec725e3	Merge pull request #6624 from oranagra/config_c_step_3 Additional config.c refractory and bugfixes	2019-12-02 08:59:36 +01:00
Oran Agra	07b365b7d7	revert an accidental test code change done as part of the tls project it seems that commit `b087dd1db6` accidentially changed gen_write_load to not use deferred client, which causes them to be slower and not generate high load which they should, making some tests less effecitive	2019-12-01 16:10:09 +02:00
Oran Agra	18e72c5cc7	Converting more configs to use generic infra, and moving defaults to config.c Changes in behavior: - Change server.stream_node_max_entries from int64_t to long long, so that it can be used by the generic infra - standard error reply instead of "repl-backlog-size must be 1 or greater" and such - tls-port and a few TLS booleans were readable (config get) even when USE_OPENSSL was off (now they aren't) - syslog-enabled, syslog-ident, cluster-enabled, appendfilename, and supervised didn't have a get (now they do) - pidfile was initialized to NULL in InitServerConfig but had CONFIG_DEFAULT_PID_FILE in rewriteConfig (so the real default was "", but rewrite would cause it to be set), fixed the rewrite. - TLS config in server.h was uninitialized (if no tls config args were provided) Adding test for sanity and coverage	2019-11-28 11:24:57 +02:00
Salvatore Sanfilippo	a1b654819c	Merge pull request #6598 from oranagra/module-hook-test try to fix an unstable test (module hook for loading progress)	2019-11-25 17:54:21 +01:00
Salvatore Sanfilippo	64c2508ee3	Merge branch 'unstable' into rm_get_server_info	2019-11-21 10:06:15 +01:00
Salvatore Sanfilippo	808394b77d	Merge pull request #6603 from daidaotong/typofix fix typo in scripting.acl	2019-11-20 10:06:33 +01:00
Daniel Dai	99b5696390	fix typo	2019-11-19 20:14:59 -05:00
Oran Agra	ed2269762b	try to fix an unstable test (module hook for loading progress) there were two lssues, one is taht BGREWRITEAOF failed since the initial one was still in progress the solution for this one is to enable appendonly from the server startup so there's no initial aofrw. the other problem was 0 loading progress events, theory is that on some platforms a sleep of 1 will cause a much greater delay due to the context switch, but on other platform it doesn't. in theory a sleep of 100 micro for 1k keys whould take 100ms, and with hz of 500 we should be gettering 50 events (one every 2ms). in practise it doesn't work like that, so trying to find a sleep that would be long enough but still not cause the test to take too long.	2019-11-19 15:01:51 +02:00
Salvatore Sanfilippo	e7144fbed8	Merge branch 'unstable' into module-long-double	2019-11-19 12:15:45 +01:00
Salvatore Sanfilippo	e916058f0b	Merge pull request #6557 from oranagra/rm_lru_lfu_revized rename RN_SetLRUOrLFU -> RM_SetLRU and RN_SetLFU	2019-11-19 11:58:07 +01:00
antirez	446f24e66d	Merge branch 'unstable' of github.com:/antirez/redis into unstable	2019-11-19 11:52:40 +01:00
Salvatore Sanfilippo	b311a368e0	Merge pull request #6558 from oranagra/module_testrdb_leak fix leak in module api rdb test	2019-11-19 11:49:43 +01:00
antirez	936e01e5bb	Fix stream test after addition of 0-0 ID test.	2019-11-19 11:49:05 +01:00
Salvatore Sanfilippo	397a8b57cc	Merge pull request #6513 from oranagra/test_assertions test infra: improve prints on failed assertions	2019-11-19 11:34:11 +01:00
Salvatore Sanfilippo	656e40eed2	Merge branch 'unstable' into scan_module_impl	2019-11-19 11:08:02 +01:00
Salvatore Sanfilippo	3d89210477	Merge pull request #3383 from yossigo/datatype_load_save Redis Module API calls to allow re-use of data type RDB save/load.	2019-11-19 10:55:42 +01:00
Guy Benoish	4a12047c61	XADD with ID 0-0 stores an empty key Calling XADD with 0-0 or 0 would result in creating an empty key and storing it in the database. Even worse, because XADD will reply with error the action will not be replicated, creating a master-replica inconsistency	2019-11-13 16:47:30 +05:30
Oran Agra	0f8692b464	Add RM_ScanKey to scan hash, set, zset, changes to RM_Scan API - Adding RM_ScanKey - Adding tests for RM_ScanKey - Refactoring RM_Scan API Changes in RM_Scan - cleanup in docs and coding convention - Moving out of experimantal Api - Adding ctx to scan callback - Dont use cursor of -1 as an indication of done (can be a valid cursor) - Set errno when returning 0 for various reasons - Rename Cursor to ScanCursor - Test filters key that are not strings, and opens a key if NULL	2019-11-11 16:05:55 +02:00

... 2 3 4 5 6 ...

1311 Commits