futriix

Author	SHA1	Message	Date
Oran Agra	05f8975d21	redis-cli tests, fix valgrind timing issue (#7519 ) this test when run with valgrind on github actions takes 160 seconds (cherry picked from commit 254c96255420e950bcad1a46bc4f8617b4373797)	2020-07-20 21:08:26 +03:00
WuYunlong	4780cc5efa	Fix out of update help info in tcl tests. (#7516 ) Before this commit, the output of "./runtest-cluster --help" is incorrect. After this commit, the format of the following 3 output is consistent: ./runtest --help ./runtest-cluster --help ./runtest-sentinel --help (cherry picked from commit 8128d39737adaf7092c9a367f71fbe9e0a2b33a2)	2020-07-20 21:08:26 +03:00
Oran Agra	aea4db2f5a	fix recently added time sensitive tests failing with valgrind (#7512 ) interestingly the latency monitor test fails because valgrind is slow enough so that the time inside PEXPIREAT command from the moment of the first mstime() call to get the basetime until checkAlreadyExpired calls mstime() again is more than 1ms, and that test was too sensitive. using this opportunity to speed up the test (unrelated to the failure) the fix is just the longer time passed to PEXPIRE. (cherry picked from commit e5227aab899628653285478a9d1083e8e8f51b57)	2020-07-20 21:08:26 +03:00
Oran Agra	b5c5f870a4	runtest --stop pause stops before terminating the redis server (#7513 ) in the majority of the cases (on this rarely used feature) we want to stop and be able to connect to the shard with redis-cli. since these are two different processes interracting with the tty we need to stop both, and we'll have to hit enter twice, but it's not that bad considering it is rarely used. (cherry picked from commit 02ef355f98691adba4126bbdab0d4d2bfe475701)	2020-07-20 21:08:26 +03:00
WuYunlong	f838df92d2	Add missing latency-monitor tcl test to test_helper.tcl. (#6782 ) (cherry picked from commit d792db7948c9de9c4ac3b7669fac2dbc9eb7b173)	2020-07-20 21:08:26 +03:00
Yossi Gottlieb	7a536c2912	TLS: Session caching configuration support. (#7420 ) * TLS: Session caching configuration support. * TLS: Remove redundant config initialization. (cherry picked from commit 3e6f2b1a45176ac3d81b95cb6025f30d7aaa1393)	2020-07-20 21:08:26 +03:00
Yossi Gottlieb	b057ff81ee	TLS: Add missing redis-cli options. (#7456 ) * Tests: fix and reintroduce redis-cli tests. These tests have been broken and disabled for 10 years now! * TLS: add remaining redis-cli support. This adds support for the redis-cli --pipe, --rdb and --replica options previously unsupported in --tls mode. * Fix writeConn(). (cherry picked from commit d9f970d8d3f0b694f1e8915cab6d4eab9cfb2ef1)	2020-07-20 21:08:26 +03:00
Oran Agra	95ba01b538	RESTORE ABSTTL won't store expired keys into the db (#7472 ) Similarly to EXPIREAT with TTL in the past, which implicitly deletes the key and return success, RESTORE should not store key that are already expired into the db. When used together with REPLACE it should emit a DEL to keyspace notification and replication stream. (cherry picked from commit 5977a94842a25140520297fe4bfda15e0e4de711)	2020-07-20 21:08:26 +03:00
Oran Agra	33ca884cc5	skip a test that uses +inf on valgrind (#7440 ) On some platforms strtold("+inf") with valgrind returns a non-inf result [err]: INCRBYFLOAT does not allow NaN or Infinity in tests/unit/type/incr.tcl Expected 'ERRwould produce' to equal or match '1189731495357231765085759.....' (cherry picked from commit 909bc97c526db757a3d022b29911ff6d08eba59c)	2020-07-20 21:08:26 +03:00
Oran Agra	2b5f23197c	stabilize tests that look for log lines (#7367 ) tests were sensitive to additional log lines appearing in the log causing the search to come empty handed. instead of just looking for the n last log lines, capture the log lines before performing the action, and then search from that offset. (cherry picked from commit 8e76e13472b7d277af78691775c2cf845f68ab90)	2020-07-20 21:08:26 +03:00
Oran Agra	1104113c07	tests/valgrind: don't use debug restart (#7404 ) * tests/valgrind: don't use debug restart DEBUG REATART causes two issues: 1. it uses execve which replaces the original process and valgrind doesn't have a chance to check for errors, so leaks go unreported. 2. valgrind report invalid calls to close() which we're unable to resolve. So now the tests use restart_server mechanism in the tests, that terminates the old server and starts a new one, new PID, but same stdout, stderr. since the stderr can contain two or more valgrind report, it is not enough to just check for the absence of leaks, we also need to check for some known errors, we do both, and fail if we either find an error, or can't find a report saying there are no leaks. other changes: - when killing a server that was already terminated we check for leaks too. - adding DEBUG LEAK which was used to test it. - adding --trace-children to valgrind, although no longer needed. - since the stdout contains two or more runs, we need slightly different way of checking if the new process is up (explicitly looking for the new PID) - move the code that handles --wait-server to happen earlier (before watching the startup message in the log), and serve the restarted server too. * squashme - CR fixes (cherry picked from commit 69ade87325eedebdb44760af9a8c28e15381888e)	2020-07-20 21:08:26 +03:00
antirez	17aaf5ec97	LPOS: option FIRST renamed RANK. (cherry picked from commit a5a3a7bbc61203398ecc1d5b52c76214f5672776)	2020-07-20 21:08:26 +03:00
Oran Agra	05e483cbb3	EXEC always fails with EXECABORT and multi-state is cleared In order to support the use of multi-exec in pipeline, it is important that MULTI and EXEC are never rejected and it is easy for the client to know if the connection is still in multi state. It was easy to make sure MULTI and DISCARD never fail (done by previous commits) since these only change the client state and don't do any actual change in the server, but EXEC is a different story. Since in the past, it was possible for clients to handle some EXEC errors and retry the EXEC, we now can't affort to return any error on EXEC other than EXECABORT, which now carries with it the real reason for the abort too. Other fixes in this commit: - Some checks that where performed at the time of queuing need to be re- validated when EXEC runs, for instance if the transaction contains writes commands, it needs to be aborted. there was one check that was already done in execCommand (-READONLY), but other checks where missing: -OOM, -MISCONF, -NOREPLICAS, -MASTERDOWN - When a command is rejected by processCommand it was rejected with addReply, which was not recognized as an error in case the bad command came from the master. this will enable to count or MONITOR these errors in the future. - make it easier for tests to create additional (non deferred) clients. - add tests for the fixes of this commit. (cherry picked from commit 65a3307bc95aadbc91d85cdf9dfbe1b3493222ca)	2020-07-20 21:08:26 +03:00
meir@redislabs.com	51e178454d	Fix RM_ScanKey module api not to return int encoded strings The scan key module API provides the scan callback with the current field name and value (if it exists). Those arguments are RedisModuleString* which means it supposes to point to robj which is encoded as a string. Using createStringObjectFromLongLong function might return robj that points to an integer and so break a module that tries for example to use RedisModule_StringPtrLen on the given field/value. The PR introduces a fix that uses the createObject function and sdsfromlonglong function. Using those function promise that the field and value pass to the to the scan callback will be Strings. The PR also changes the Scan test module to use RedisModule_StringPtrLen to catch the issue. without this, the issue is hidden because RedisModule_ReplyWithString knows to handle integer encoding of the given robj (RedisModuleString). The PR also introduces a new test to verify the issue is solved. (cherry picked from commit a89bf734a933e45b9dd3ae85ef4c3b62bd6891d8)	2020-07-20 21:08:26 +03:00
antirez	e0022d8cfe	LPOS: tests + crash fix.	2020-06-12 12:08:06 +02:00
antirez	a7d3670da5	Adapt EVAL+busy script test to new behavior.	2020-06-09 12:19:30 +02:00
zhaozhao.zz	1d7bf208ce	AOF: append origin SET if no expire option	2020-06-09 11:53:01 +02:00
Oran Agra	f33de403ed	fix pingoff test race	2020-06-06 11:44:21 +02:00
antirez	59cd4c9f65	Test: take PSYNC2 test master timeout high during switch. This will likely avoid false positives due to trailing pings.	2020-05-28 10:56:14 +02:00
antirez	6c1bb7b194	Test: add the tracking unit as default.	2020-05-28 10:22:29 +02:00
Oran Agra	1aee695e52	tests: find_available_port start search from next port i.e. don't start the search from scratch hitting the used ones again. this will also reduce the likelihood of collisions (if there are any left) by increasing the time until we re-use a port we did use in the past.	2020-05-28 10:09:51 +02:00
Oran Agra	a2ae463520	tests: each test client work on a distinct port range apparently when running tests in parallel (the default of --clients 16), there's a chance for two tests to use the same port. specifically, one test might shutdown a master and still have the replica up, and then another test will re-use the port number of master for another master, and then that replica will connect to the master of the other test. this can cause a master to count too many full syncs and fail a test if we run the tests with --single integration/psync2 --loop --stop see Probmem 2 in #7314	2020-05-28 10:09:51 +02:00
Oran Agra	86e562d691	32bit CI needs to build modules correctly	2020-05-28 10:09:51 +02:00
Oran Agra	ab2984b1e2	adjust revived meaningful offset tests these tests create several edge cases that are otherwise uncovered (at least not consistently) by the test suite, so although they're no longer testing what they were meant to test, it's still a good idea to keep them in hope that they'll expose some issue in the future.	2020-05-28 10:09:51 +02:00
Oran Agra	1ff5a222de	revive meaningful offset tests	2020-05-28 10:09:51 +02:00
antirez	3f8d113f1b	Another meaningful offset test removed.	2020-05-28 10:09:51 +02:00
antirez	d4541349dc	Remove the PSYNC2 meaningful offset test.	2020-05-28 10:09:51 +02:00
antirez	8f10137227	Test: PSYNC2 test can now show server logs.	2020-05-28 10:09:51 +02:00
Qu Chen	58fc456cbd	Disconnect chained replicas when the replica performs PSYNC with the master always to avoid replication offset mismatch between master and chained replicas.	2020-05-22 12:37:59 +02:00
Oran Agra	436be34986	fix a rare active defrag edge case bug leading to stagnation There's a rare case which leads to stagnation in the defragger, causing it to keep scanning the keyspace and do nothing (not moving any allocation), this happens when all the allocator slabs of a certain bin have the same % utilization, but the slab from which new allocations are made have a lower utilization. this commit fixes it by removing the current slab from the overall average utilization of the bin, and also eliminate any precision loss in the utilization calculation and move the decision about the defrag to reside inside jemalloc. and also add a test that consistently reproduce this issue.	2020-05-22 12:37:49 +02:00
WuYunlong	2e41827435	Handle keys with hash tag when computing hash slot using tcl cluster client.	2020-05-22 12:37:49 +02:00
WuYunlong	eb2c8b2c61	Add a test to prove current tcl cluster client can not handle keys with hash tag.	2020-05-22 12:37:49 +02:00
Oran Agra	00d8b92b89	fix valgrind test failure in replication test in b4416280c i added more keys to that test to make it run longer but in valgrind this now means the test times out, give valgrind more time.	2020-05-22 12:37:49 +02:00
Oran Agra	5e17e6276c	add regression test for the race in #7205 with the original version of 6.0.0, this test detects an excessive full sync. with the fix in 1a7cd2c0e, this test detects memory corruption, especially when using libc allocator with or without valgrind.	2020-05-22 12:37:49 +02:00
antirez	96e7c011e2	Improve the PSYNC2 test reliability.	2020-05-22 12:37:49 +02:00
Yossi Gottlieb	16ba33c05b	TLS: Fix test failures on recent Debian/Ubuntu. Seems like on some systems choosing specific TLS v1/v1.1 versions no longer works as expected. Test is reduced for v1.2 now which is still good enough to test the mechansim, and matters most anyway.	2020-05-15 22:23:24 +02:00
Oran Agra	9da134cd88	fix redis 6.0 not freeing closed connections during loading. This bug was introduced by a recent change in which readQueryFromClient is using freeClientAsync, and despite the fact that now freeClientsInAsyncFreeQueue is in beforeSleep, that's not enough since it's not called during loading in processEventsWhileBlocked. furthermore, afterSleep was called in that case but beforeSleep wasn't. This bug also caused slowness sine the level-triggered mode of epoll kept signaling these connections as readable causing us to keep doing connRead again and again for ll of these, which keep accumulating. now both before and after sleep are called, but not all of their actions are performed during loading, some are only reserved for the main loop. fixes issue #7215	2020-05-14 11:29:43 +02:00
antirez	f7f219a137	Regression test for #7249 .	2020-05-14 11:29:43 +02:00
Oran Agra	5c41802d55	fix unstable replication test this test which has coverage for varoius flows of diskless master was failing randomly from time to time. the failure was: [err]: diskless all replicas drop during rdb pipe in tests/integration/replication.tcl log message of 'Diskless rdb transfer, last replica dropped, killing fork child' not found what seemed to have happened is that the master didn't detect that all replicas dropped by the time the replication ended, it thought that one replica is still connected. now the test takes a few seconds longer but it seems stable.	2020-05-14 11:29:43 +02:00
antirez	e48c37316e	Test: --dont-clean should do first cleanup.	2020-05-08 10:37:36 +02:00
zhenwei pi	d6436eb7cf	Support setcpuaffinity on linux/bsd Currently, there are several types of threads/child processes of a redis server. Sometimes we need deeply optimise the performance of redis, so we would like to isolate threads/processes. There were some discussion about cpu affinity cases in the issue: https://github.com/antirez/redis/issues/2863 So implement cpu affinity setting by redis.conf in this patch, then we can config server_cpulist/bio_cpulist/aof_rewrite_cpulist/ bgsave_cpulist by cpu list. Examples of cpulist in redis.conf: server_cpulist 0-7:2 means cpu affinity 0,2,4,6 bio_cpulist 1,3 means cpu affinity 1,3 aof_rewrite_cpulist 8-11 means cpu affinity 8,9,10,11 bgsave_cpulist 1,10-11 means cpu affinity 1,10,11 Test on linux/freebsd, both work fine. Signed-off-by: zhenwei pi <pizhenwei@bytedance.com>	2020-05-08 10:37:35 +02:00
Oran Agra	3d3861dd88	add daily github actions with libc malloc and valgrind * fix memlry leaks with diskless replica short read. * fix a few timing issues with valgrind runs * fix issue with valgrind and watchdog schedule signal about the valgrind WD issue: the stack trace test in logging.tcl, has issues with valgrind: ==28808== Can't extend stack to 0x1ffeffdb38 during signal delivery for thread 1: ==28808== too small or bad protection modes it seems to be some valgrind bug with SA_ONSTACK. SA_ONSTACK seems unneeded since WD is not recursive (SA_NODEFER was removed), also, not sure if it's even valid without a call to sigaltstack()	2020-05-08 10:37:35 +02:00
Guy Benoish	6c0bc608a1	Extend XINFO STREAM output Introducing XINFO STREAM <key> FULL	2020-04-30 13:02:58 +02:00
Oran Agra	ea63aea72d	fix loading race in psync2 tests	2020-04-28 11:20:15 +02:00
Oran Agra	e4d2bb62b2	Keep track of meaningful replication offset in replicas too Now both master and replicas keep track of the last replication offset that contains meaningful data (ignoring the tailing pings), and both trim that tail from the replication backlog, and the offset with which they try to use for psync. the implication is that if someone missed some pings, or even have excessive pings that the promoted replica has, it'll still be able to psync (avoid full sync). the downside (which was already committed) is that replicas running old code may fail to psync, since the promoted replica trims pings form it's backlog. This commit adds a test that reproduces several cases of promotions and demotions with stale and non-stale pings Background: The mearningful offset on the master was added recently to solve a problem were the master is left all alone, injecting PINGs into it's backlog when no one is listening and then gets demoted and tries to replicate from a replica that didn't have any of the PINGs (or at least not the last ones). however, consider this case: master A has two replicas (B and C) replicating directly from it. there's no traffic at all, and also no network issues, just many pings in the tail of the backlog. now B gets promoted, A becomes a replica of B, and C remains a replica of A. when A gets demoted, it trims the pings from its backlog, and successfully replicate from B. however, C is still aware of these PINGs, when it'll disconnect and re-connect to A, it'll ask for something that's not in the backlog anymore (since A trimmed the tail of it's backlog), and be forced to do a full sync (something it didn't have to do before the meaningful offset fix). Besides that, the psync2 test was always failing randomly here and there, it turns out the reason were PINGs. Investigating it shows the following scenario: cycle 1: redis #1 is master, and all the rest are direct replicas of #1 cycle 2: redis #2 is promoted to master, #1 is a replica of #2 and #3 is replica of #1 now we see that when #1 is demoted it prints: 17339:S 21 Apr 2020 11:16:38.523 * Using the meaningful offset 3929963 instead of 3929977 to exclude the final PINGs (14 bytes difference) 17339:S 21 Apr 2020 11:16:39.391 * Trying a partial resynchronization (request e2b3f8817735fdfe5fa4626766daa938b61419e5:3929964). 17339:S 21 Apr 2020 11:16:39.392 * Successful partial resynchronization with master. and when #3 connects to the demoted #2, #2 says: 17339:S 21 Apr 2020 11:16:40.084 * Partial resynchronization not accepted: Requested offset for secondary ID was 3929978, but I can reply up to 3929964 so the issue here is that the meaningful offset feature saved the day for the demoted master (since it needs to sync from a replica that didn't get the last ping), but it didn't help one of the other replicas which did get the last ping.	2020-04-27 15:52:49 +02:00
Guy Benoish	43329c9b64	Add the stream tag to XSETID tests	2020-04-27 15:52:49 +02:00
antirez	3722f89f49	LCS -> STRALGO LCS. STRALGO should be a container for mostly read-only string algorithms in Redis. The algorithms should have two main characteristics: 1. They should be non trivial to compute, and often not part of programming language standard libraries. 2. They should be fast enough that it is a good idea to have optimized C implementations. Next thing I would love to see? A small strings compression algorithm.	2020-04-24 16:49:27 +02:00
yanhui13	f03f1fad67	add tcl test for cluster slots	2020-04-24 10:15:04 +02:00
antirez	06917e581c	Tracking: test expired keys notifications.	2020-04-24 10:14:48 +02:00
antirez	e434b2ce4f	Tracking: NOLOOP tests.	2020-04-24 10:14:48 +02:00

1 2 3 4 5 ...

1082 Commits