futriix

Author	SHA1	Message	Date
Oran Agra	eca1014fe3	Run active defrag while blocked / loading (#7726 ) During long running scripts or loading RDB/AOF, we may need to do some defragging. Since processEventsWhileBlocked is called periodically at unknown intervals, and many cron jobs either depend on run_with_period (including active defrag), or rely on being called at server.hz rate (i.e. active defrag knows ho much time to run by looking at server.hz), the whileBlockedCron may have to run a loop triggering the cron jobs in it (currently only active defrag) several times. Other changes: - Adding a test for defrag during aof loading. - Changing key-load-delay config to take negative values for fractions of a microsecond sleep	2020-09-03 08:47:29 +03:00
Oran Agra	6041fc99b5	Reduce the probability of failure when start redis in runtest-cluster #7554 (#7635 ) When runtest-cluster, at first, we need to create a cluster use spawn_instance, a port which is not used is choosen, however sometimes we can't run server on the port. possibley due to a race with another process taking it first. such as redis/redis/runs/896537490. It may be due to the machine problem or In order to reduce the probability of failure when start redis in runtest-cluster, we attemp to use another port when find server do not start up. Co-authored-by: Oran Agra <oran@redislabs.com> Co-authored-by: yanhui13 <yanhui13@meituan.com> (cherry picked from commit 1deaad884c38e92e5b691f36b253ef4ee2201ca4)	2020-09-01 09:27:58 +03:00
Yossi Gottlieb	ba1da77a3d	Fix oom-score-adj on older distros. (#7724 ) Don't assume `ps` handles `-h` to display output without headers and manually trim headers line from output. (cherry picked from commit ae8420298cacc2737e8e3ffa3c5acc038cd27849)	2020-09-01 09:27:58 +03:00
Yossi Gottlieb	f0e28abc07	Add oom-score-adj configuration option to control Linux OOM killer. (#1690 ) Add Linux kernel OOM killer control option. This adds the ability to control the Linux OOM killer oom_score_adj parameter for all Redis processes, depending on the process role (i.e. master, replica, background child). A oom-score-adj global boolean flag control this feature. In addition, specific values can be configured using oom-score-adj-values if additional tuning is required. (cherry picked from commit 70c823a64e800f22ac68f0172acdd1da82d7be32)	2020-09-01 09:27:58 +03:00
Meir Shpilraien (Spielrein)	63e3f1e449	see #7544 , added RedisModule_HoldString api. (#7577 ) Added RedisModule_HoldString that either returns a shallow copy of the given String (by increasing the String ref count) or a new deep copy of String in case its not possible to get a shallow copy. Co-authored-by: Itamar Haber <itamar@redislabs.com> (cherry picked from commit 4f99b22118ca91e3a7fe9c1c68c19dd717dfdbb5)	2020-09-01 09:27:58 +03:00
Meir Shpilraien (Spielrein)	f63e428e5b	This PR introduces a new loaded keyspace event (#7536 ) Co-authored-by: Oran Agra <oran@redislabs.com> Co-authored-by: Itamar Haber <itamar@redislabs.com> (cherry picked from commit 73198c50194cbf0254afd4cc5245f9274a538d13)	2020-09-01 09:27:58 +03:00
valentinogeron	3c136a7777	EXEC with only read commands should not be rejected when OOM (#7696 ) If the server gets MULTI command followed by only read commands, and right before it gets the EXEC it reaches OOM, the client will get OOM response. So, from now on, it will get OOM response only if there was at least one command that was tagged with `use-memory` flag (cherry picked from commit 0292720ccb0a189d3ed49d7bf912602360a4ecdd)	2020-09-01 09:27:58 +03:00
Valentino Geron	34124fff88	Fix LPOS command when RANK is greater than matches When calling to LPOS command when RANK is higher than matches, the return value is non valid response. For example: ``` LPUSH l a :1 LPOS l b RANK 5 COUNT 10 -4 ``` It may break client-side parser. Now, we count how many replies were replied in the array. ``` LPUSH l a :1 LPOS l b RANK 5 COUNT 10 0 ``` (cherry picked from commit 7a555da64f56a4fb2f300d84a35778bee8f471ca)	2020-09-01 09:27:58 +03:00
Yossi Gottlieb	c5675c66bc	Tests: fix redis-cli with remote hosts. (#7693 ) (cherry picked from commit 257f9f462f7782dcaecf7bbf35f4701b20b88a45)	2020-09-01 09:27:58 +03:00
杨博东	b42976bd56	Fix flock cluster config may cause failure to restart after kill -9 (#7674 ) After fork, the child process(redis-aof-rewrite) will get the fd opened by the parent process(redis), when redis killed by kill -9, it will not graceful exit(call prepareForShutdown()), so redis-aof-rewrite thread may still alive, the fd(lock) will still be held by redis-aof-rewrite thread, and redis restart will fail to get lock, means fail to start. This issue was causing failures in the cluster tests in github actions. Co-authored-by: Oran Agra <oran@redislabs.com> (cherry picked from commit 5e6212e087c4696abc682b64079202c9ade8666c)	2020-09-01 09:27:58 +03:00
Yossi Gottlieb	d13c44583c	Module API: fix missing RM_CLIENTINFO_FLAG_SSL. (#7666 ) The `REDISMODULE_CLIENTINFO_FLAG_SSL` flag was already a part of the `RedisModuleClientInfo` structure but was not implemented. (cherry picked from commit 2ec11f941ae41188e517670fc3224b12c7666541)	2020-09-01 09:27:58 +03:00
Oran Agra	b4a6b4f28d	fix new rdb test failing on timing issues (#7604 ) apparenlty on github actions sometimes 500ms is not enough (cherry picked from commit 191b1181023b0860ec60afde7a41bd4f03c55097)	2020-09-01 09:27:58 +03:00
Oran Agra	3a4ee4b6d6	module hook for master link up missing on successful psync (#7584 ) besides, hooks test was time sensitive. when the replica managed to reconnect quickly after the client kill, the test would fail (cherry picked from commit c5d85c69c75438f98f84e549877c2999a2e450a8)	2020-09-01 09:27:58 +03:00
WuYunlong	37fba8f4d8	Fix running single test 14-consistency-check.tcl (#7587 ) (cherry picked from commit be11e1b5eaf0d6ab5e68f86c1346570531eee766)	2020-09-01 09:27:58 +03:00
Yossi Gottlieb	2e6563cdcf	Fix TLS cluster tests. (#7578 ) Fix consistency test added in 0c9916d00 without considering TLS redis-cli configuration. (cherry picked from commit 675b00c7e0b7d68bafa11fcc7f66a394c3c3cd36)	2020-09-01 09:27:58 +03:00
Oran Agra	10a8407a4f	Fix failing tests due to issues with wait_for_log_message (#7572 ) - the test now waits for specific set of log messages rather than wait for timeout looking for just one message. - we don't wanna sample the current length of the log after an action, due to a race, we need to start the search from the line number of the last message we where waiting for. - when attempting to trigger a full sync, use multi-exec to avoid a race where the replica manages to re-connect before we completed the set of actions that should force a full sync. - fix verify_log_message which was broken and unused (cherry picked from commit 06aaeabaea9d9b248e8a790dde352cd14d66628a)	2020-09-01 09:27:58 +03:00
Jiayuan Chen	7b2af98316	Add optional tls verification (#7502 ) Adds an `optional` value to the previously boolean `tls-auth-clients` configuration keyword. Co-authored-by: Yossi Gottlieb <yossigo@gmail.com> (cherry picked from commit 198770751fdc4c46eb4971ead9b5787fd6ce39fd)	2020-09-01 09:27:58 +03:00
Oran Agra	558a343b3c	Stabilize bgsave test that sometimes fails with valgrind (#7559 ) on ci.redis.io the test fails a lot, reporting that bgsave didn't end. increaseing the timeout we wait for that bgsave to get aborted. in addition to that, i also verify that it indeed got aborted by checking that the save counter wasn't reset. add another test to verify that a successful bgsave indeed resets the change counter. (cherry picked from commit 49d4aebce0a0b94cd2b302d276be95d1a1ce8610)	2020-09-01 09:27:58 +03:00
Oran Agra	2b45c88a6a	testsuite may leave servers alive on error (#7549 ) in cases where you have test name { start_server { start_server { assert } } } the exception will be thrown to the test proc, and the servers are supposed to be killed on the way out. but it seems there was always a bug of not cleaning the server stack, and recently (#7404) we started relying on that stack in order to kill them, so with that bug sometimes we would have tried to kill the same server twice, and leave one alive. luckly, in most cases the pattern is: start_server { test name { } } (cherry picked from commit bb170fa06e5909dd816b6530121952d57c8209a0)	2020-09-01 09:27:58 +03:00
Yossi Gottlieb	6d80011e73	Tests: drop TCL 8.6 dependency. (#7548 ) This re-implements the redis-cli --pipe test so it no longer depends on a close feature available only in TCL 8.6. Basically what this test does is run redis-cli --pipe, generates a bunch of commands and pipes them through redis-cli, and inspects the result in both Redis and the redis-cli output. To do that, we need to close stdin for redis-cli to indicate we're done so it can flush its buffers and exit. TCL has bi-directional channels can only offers a way to "one-way close" a channel with TCL 8.6. To work around that, we now generate the commands into a file and feed that file to redis-cli directly. As we're writing to an actual file, the number of commands is now reduced. (cherry picked from commit dbc0a64843ccd07515ac41ca80497a9e5ffd107a)	2020-09-01 09:27:58 +03:00
Remi Collet	443e57b08e	Fix deprecated tail syntax in tests (#7543 ) (cherry picked from commit 7853d8410b12c3ffac699c8a2e06f2a8e6df26b0)	2020-09-01 09:27:58 +03:00
Yossi Gottlieb	ae8420298c	Fix oom-score-adj on older distros. (#7724 ) Don't assume `ps` handles `-h` to display output without headers and manually trim headers line from output.	2020-08-30 12:23:47 +03:00
valentinogeron	0292720ccb	EXEC with only read commands should not be rejected when OOM (#7696 ) If the server gets MULTI command followed by only read commands, and right before it gets the EXEC it reaches OOM, the client will get OOM response. So, from now on, it will get OOM response only if there was at least one command that was tagged with `use-memory` flag	2020-08-27 09:19:24 +03:00
Oran Agra	b01816ca6e	Add test coverage for CLIENT UNBLOCK (#7712 ) plus minor other fixes to list.tcl	2020-08-27 08:09:39 +03:00
John Sully	ff9df842d8	Implement use-fork config (fails with diskless repl) Former-commit-id: f2d5c2bca22e9fd506db123c47b7f60cdded7e2c	2020-08-24 03:17:59 +00:00
Valentino Geron	7a555da64f	Fix LPOS command when RANK is greater than matches When calling to LPOS command when RANK is higher than matches, the return value is non valid response. For example: ``` LPUSH l a :1 LPOS l b RANK 5 COUNT 10 -4 ``` It may break client-side parser. Now, we count how many replies were replied in the array. ``` LPUSH l a :1 LPOS l b RANK 5 COUNT 10 0 ```	2020-08-23 16:03:30 +03:00
Yossi Gottlieb	257f9f462f	Tests: fix redis-cli with remote hosts. (#7693 )	2020-08-23 10:17:43 +03:00
杨博东	5e6212e087	Fix flock cluster config may cause failure to restart after kill -9 (#7674 ) After fork, the child process(redis-aof-rewrite) will get the fd opened by the parent process(redis), when redis killed by kill -9, it will not graceful exit(call prepareForShutdown()), so redis-aof-rewrite thread may still alive, the fd(lock) will still be held by redis-aof-rewrite thread, and redis restart will fail to get lock, means fail to start. This issue was causing failures in the cluster tests in github actions. Co-authored-by: Oran Agra <oran@redislabs.com>	2020-08-20 08:59:02 +03:00
Yossi Gottlieb	2ec11f941a	Module API: fix missing RM_CLIENTINFO_FLAG_SSL. (#7666 ) The `REDISMODULE_CLIENTINFO_FLAG_SSL` flag was already a part of the `RedisModuleClientInfo` structure but was not implemented.	2020-08-17 17:46:54 +03:00
Yossi Gottlieb	70c823a64e	Add oom-score-adj configuration option to control Linux OOM killer. (#1690 ) Add Linux kernel OOM killer control option. This adds the ability to control the Linux OOM killer oom_score_adj parameter for all Redis processes, depending on the process role (i.e. master, replica, background child). A oom-score-adj global boolean flag control this feature. In addition, specific values can be configured using oom-score-adj-values if additional tuning is required.	2020-08-12 17:58:56 +03:00
Mota	e637e1f02b	Test:Fix invalid cases in hash.tcl and dump.tcl (#4611 )	2020-08-12 10:25:24 +08:00
Tyson Andre	6e17afa80a	Implement SMISMEMBER key member [member ...] (#7615 ) This is a rebased version of #3078 originally by shaharmor with the following patches by TysonAndre made after rebasing to work with the updated C API: 1. Add 2 more unit tests (wrong argument count error message, integer over 64 bits) 2. Use addReplyArrayLen instead of addReplyMultiBulkLen. 3. Undo changes to src/help.h - for the ZMSCORE PR, I heard those should instead be automatically generated from the redis-doc repo if it gets updated Motivations: - Example use case: Client code to efficiently check if each element of a set of 1000 items is a member of a set of 10 million items. (Similar to reasons for working on #7593) - HMGET and ZMSCORE already exist. This may lead to developers deciding to implement functionality that's best suited to a regular set with a data type of sorted set or hash map instead, for the multi-get support. Currently, multi commands or lua scripting to call sismember multiple times would almost definitely be less efficient than a native smismember for the following reasons: - Need to fetch the set from the string every time instead of reusing the C pointer. - Using pipelining or multi-commands would result in more bytes sent and received by the client for the repeated SISMEMBER KEY sections. - Need to specially encode the data and decode it from the client for lua-based solutions. - Proposed solutions using Lua or SADD/SDIFF could trigger writes to memory, which is undesirable on a redis replica server or when commands get replicated to replicas. Co-Authored-By: Shahar Mor <shahar@peer5.com> Co-Authored-By: Tyson Andre <tysonandre775@hotmail.com>	2020-08-11 11:55:06 +03:00
Meir Shpilraien (Spielrein)	4f99b22118	see #7544 , added RedisModule_HoldString api. (#7577 ) Added RedisModule_HoldString that either returns a shallow copy of the given String (by increasing the String ref count) or a new deep copy of String in case its not possible to get a shallow copy. Co-authored-by: Itamar Haber <itamar@redislabs.com>	2020-08-09 06:11:47 +03:00
Oran Agra	1deaad884c	Reduce the probability of failure when start redis in runtest-cluster #7554 (#7635 ) When runtest-cluster, at first, we need to create a cluster use spawn_instance, a port which is not used is choosen, however sometimes we can't run server on the port. possibley due to a race with another process taking it first. such as redis/redis/runs/896537490. It may be due to the machine problem or In order to reduce the probability of failure when start redis in runtest-cluster, we attemp to use another port when find server do not start up. Co-authored-by: Oran Agra <oran@redislabs.com> Co-authored-by: yanhui13 <yanhui13@meituan.com>	2020-08-09 06:08:00 +03:00
Oran Agra	cad93ed273	Accelerate diskless master connections, and general re-connections (#6271 ) Diskless master has some inherent latencies. 1) fork starts with delay from cron rather than immediately 2) replica is put online only after an ACK. but the ACK was sent only once a second. 3) but even if it would arrive immediately, it will not register in case cron didn't yet detect that the fork is done. Besides that, when a replica disconnects, it doesn't immediately attempts to re-connect, it waits for replication cron (one per second). in case it was already online, it may be important to try to re-connect as soon as possible, so that the backlog at the master doesn't vanish. In case it disconnected during rdb transfer, one can argue that it's not very important to re-connect immediately, but this is needed for the "diskless loading short read" test to be able to run 100 iterations in 5 seconds, rather than 3 (waiting for replication cron re-connection) changes in this commit: 1) sync command starts a fork immediately if no sync_delay is configured 2) replica sends REPLCONF ACK when done reading the rdb (rather than on 1s cron) 3) when a replica unexpectedly disconnets, it immediately tries to re-connect rather than waiting 1s 4) when when a child exits, if there is another replica waiting, we spawn a new one right away, instead of waiting for 1s replicationCron. 5) added a call to connectWithMaster from replicationSetMaster. which is called from the REPLICAOF command but also in 3 places in cluster.c, in all of these the connection attempt will now be immediate instead of delayed by 1 second. side note: we can add a call to rdbPipeReadHandler in replconfCommand when getting a REPLCONF ACK from the replica to solve a race where the replica got the entire rdb and EOF marker before we detected that the pipe was closed. in the test i did see this race happens in one about of some 300 runs, but i concluded that this race is unlikely in real life (where the replica is on another host and we're more likely to first detect the pipe was closed. the test runs 100 iterations in 3 seconds, so in some cases it'll take 4 seconds instead (waiting for another REPLCONF ACK). Removing unneeded startBgsaveForReplication from updateSlavesWaitingForBgsave Now that CheckChildrenDone is calling the new replicationStartPendingFork (extracted from serverCron) there's actually no need to call startBgsaveForReplication from updateSlavesWaitingForBgsave anymore, since as soon as updateSlavesWaitingForBgsave returns, CheckChildrenDone is calling replicationStartPendingFork that handles that anyway. The code in updateSlavesWaitingForBgsave had a bug in which it ignored repl-diskless-sync-delay, but removing that code shows that this bug was hiding another bug, which is that the max_idle should have used >= and not >, this one second delay has a big impact on my new test.	2020-08-06 16:53:06 +03:00
WuYunlong	2c68d2d5c7	Fix tests/cluster/cluster.tcl about wrong usage of lrange. (#6702 )	2020-08-04 18:00:58 +03:00
Tyson Andre	486e39e86e	Add a ZMSCORE command returning an array of scores. (#7593 ) Syntax: `ZMSCORE KEY MEMBER [MEMBER ...]` This is an extension of #2359 amended by Tyson Andre to work with the changed unstable API, add more tests, and consistently return an array. - It seemed as if it would be more likely to get reviewed after updating the implementation. Currently, multi commands or lua scripting to call zscore multiple times would almost definitely be less efficient than a native ZMSCORE for the following reasons: - Need to fetch the set from the string every time instead of reusing the C pointer. - Using pipelining or multi-commands would result in more bytes sent by the client for the repeated `ZMSCORE KEY` sections. - Need to specially encode the data and decode it from the client for lua-based solutions. - The fastest solution I've seen for large sets(thousands or millions) involves lua and a variadic ZADD, then a ZINTERSECT, then a ZRANGE 0 -1, then UNLINK of a temporary set (or lua). This is still inefficient. Co-authored-by: Tyson Andre <tysonandre775@hotmail.com>	2020-08-04 17:49:33 +03:00
Oran Agra	191b118102	fix new rdb test failing on timing issues (#7604 ) apparenlty on github actions sometimes 500ms is not enough	2020-08-04 08:53:50 +03:00
Oran Agra	c5d85c69c7	module hook for master link up missing on successful psync (#7584 ) besides, hooks test was time sensitive. when the replica managed to reconnect quickly after the client kill, the test would fail	2020-07-31 13:14:29 +03:00
WuYunlong	be11e1b5ea	Fix running single test 14-consistency-check.tcl (#7587 )	2020-07-30 08:56:21 +03:00
Yossi Gottlieb	675b00c7e0	Fix TLS cluster tests. (#7578 ) Fix consistency test added in 0c9916d00 without considering TLS redis-cli configuration.	2020-07-28 14:04:06 +03:00
Oran Agra	06aaeabaea	Fix failing tests due to issues with wait_for_log_message (#7572 ) - the test now waits for specific set of log messages rather than wait for timeout looking for just one message. - we don't wanna sample the current length of the log after an action, due to a race, we need to start the search from the line number of the last message we where waiting for. - when attempting to trigger a full sync, use multi-exec to avoid a race where the replica manages to re-connect before we completed the set of actions that should force a full sync. - fix verify_log_message which was broken and unused	2020-07-28 11:15:29 +03:00
Jiayuan Chen	198770751f	Add optional tls verification (#7502 ) Adds an `optional` value to the previously boolean `tls-auth-clients` configuration keyword. Co-authored-by: Yossi Gottlieb <yossigo@gmail.com>	2020-07-28 10:45:21 +03:00
Oran Agra	49d4aebce0	Stabilize bgsave test that sometimes fails with valgrind (#7559 ) on ci.redis.io the test fails a lot, reporting that bgsave didn't end. increaseing the timeout we wait for that bgsave to get aborted. in addition to that, i also verify that it indeed got aborted by checking that the save counter wasn't reset. add another test to verify that a successful bgsave indeed resets the change counter.	2020-07-23 13:06:24 +03:00
Meir Shpilraien (Spielrein)	73198c5019	This PR introduces a new loaded keyspace event (#7536 ) Co-authored-by: Oran Agra <oran@redislabs.com> Co-authored-by: Itamar Haber <itamar@redislabs.com>	2020-07-23 12:38:51 +03:00
Oran Agra	bb170fa06e	testsuite may leave servers alive on error (#7549 ) in cases where you have test name { start_server { start_server { assert } } } the exception will be thrown to the test proc, and the servers are supposed to be killed on the way out. but it seems there was always a bug of not cleaning the server stack, and recently (#7404) we started relying on that stack in order to kill them, so with that bug sometimes we would have tried to kill the same server twice, and leave one alive. luckly, in most cases the pattern is: start_server { test name { } }	2020-07-21 16:56:19 +03:00
Yossi Gottlieb	dbc0a64843	Tests: drop TCL 8.6 dependency. (#7548 ) This re-implements the redis-cli --pipe test so it no longer depends on a close feature available only in TCL 8.6. Basically what this test does is run redis-cli --pipe, generates a bunch of commands and pipes them through redis-cli, and inspects the result in both Redis and the redis-cli output. To do that, we need to close stdin for redis-cli to indicate we're done so it can flush its buffers and exit. TCL has bi-directional channels can only offers a way to "one-way close" a channel with TCL 8.6. To work around that, we now generate the commands into a file and feed that file to redis-cli directly. As we're writing to an actual file, the number of commands is now reduced.	2020-07-21 14:17:14 +03:00
Remi Collet	7853d8410b	Fix deprecated tail syntax in tests (#7543 )	2020-07-21 09:07:54 +03:00
WuYunlong	8b20802a09	Fix command help for unexpected options (#7476 ) (cherry picked from commit e5166eccee3396a24dfd3a79d3211943e5a3d25e)	2020-07-20 21:08:26 +03:00
Oran Agra	6bdc5a4a08	redis-cli tests, fix valgrind timing issue (#7519 ) this test when run with valgrind on github actions takes 160 seconds (cherry picked from commit 8a14ce8634c49d992aa929cf0f98e96f03bccba4)	2020-07-20 21:08:26 +03:00

... 8 9 10 11 12 ...

1842 Commits