futriix

Author	SHA1	Message	Date
zhaozhao.zz	28c5a17edf	replica redirect read&write to primary in standalone mode (#325 ) To implement #319 1. replica is able to redirect read and write commands to it's primary in standalone mode * reply with "-REDIRECT primary-ip:port" 2. add a subcommand `CLIENT CAPA redirect`, a client can announce the capability to handle redirection * if a client can handle redirection, the data access commands (read and write) will be redirected 3. allow `readonly` and `readwrite` command in standalone mode, may be a breaking change * a client with redirect capability cannot process read commands on a replica by default * use READONLY command can allow read commands on a replica --------- Signed-off-by: zhaozhao.zz <zhaozhao.zz@alibaba-inc.com>	2024-06-27 19:00:45 +08:00
Ouri Half	ab3873011a	Replacing REDIS_STATIC with static (#691 ) As discussed, we want to remove the old `REDIS_STATIC` flag which is no longer relevant. When moving from Redis to Valkey we renamed all REDIS flags in Makefile. The REDIS_STATIC flag was renamed to SERVER_STATIC, but this change was not updated in some of the files. After discussing it with @madolson and @ranshid, we decided that since this was introduced 10 years ago, and in many places in the code base we simply use `static`, we should simplify and remove the flag entirely. --------- Signed-off-by: Ouri Half <ourih@amazon.com>	2024-06-26 09:47:59 -07:00
Pierre	495c35d918	Add check in CLUSTERLINK KILL cmd to avoid freeing links to myself (#689 ) Add check in CLUSTERLINK KILL cmd to avoid freeing cluster bus links to myself. Also add an assert in `freeClusterLink()`. Testing: ``` 127.0.0.1:6379> debug clusterlink kill all c0404ee68574c6aa1048aaebfe90283afe51d2fc (error) ERR Cannot free cluster link(s) to myself ``` Signed-off-by: Pierre Turin <pieturin@amazon.com>	2024-06-25 15:18:30 -07:00
Kyle Kim (kimkyle@)	b49eaad367	Introduce a minimal debugger for .tcl integration test suite. (#683 ) Introduce a break-point function called `bp`, based on the tcl wiki's minimal debugger. ```tcl proc bp {{s {}}} { if ![info exists ::bp_skip] { set ::bp_skip [list] } elseif {[lsearch -exact $::bp_skip $s]>=0} return if [catch {info level -1} who] {set who ::} while 1 { puts -nonewline "$who/$s> "; flush stdout gets stdin line if {$line=="c"} {puts "continuing.."; break} if {$line=="i"} {set line "info locals"} catch {uplevel 1 $line} res puts $res } } ``` ``` ... your test code before break-point bp 1 ... your test code after break-point ``` The `bp 1` will give back the tcl interpreter to the developer, and allow you to interactively print local variables (through `puts`), run functions and so forth. Source: https://wiki.tcl-lang.org/page/A+minimal+debugger --------- Signed-off-by: Kyle Kim <kimkyle@amazon.com> Signed-off-by: Madelyn Olson <madelyneolson@gmail.com> Co-authored-by: Madelyn Olson <madelyneolson@gmail.com>	2024-06-25 10:24:53 -07:00
ranshid	3df9d42794	Fix bad memory accounting for sds when no malloc_size available (#694 ) Issue Introduced by #453. When we check the SDS _TYPE_5 allocation size we mistakenly used zmalloc_size which DOES take the PREFIX size into account when no malloc_size support. Later when we free we add the PREFIX_SIZE again which leads to negative memory accounting on some tests. Example test failure: https://github.com/valkey-io/valkey/actions/runs/9654170962/job/26627901497 Signed-off-by: ranshid <ranshid@amazon.com>	2024-06-25 08:18:07 -07:00
Lipeng Zhu	4d3d6c06a1	Reduce redundant call of prepareClientToWrite when call addReply* continuously. (#670 ) ## Description While exploring hotspots with profiling some benchmark workloads, we noticed the high cycles ratio of `prepareClientToWrite`, taking about 9% of the CPU of `smembers`, `lrange` commands. After deep dive the code logic, we thought we can gain the performance by reducing the redundant call of `prepareClientToWrite` when call addReply* continuously. For example: In https://github.com/valkey-io/valkey/blob/unstable/src/networking.c#L1080-L1082, `prepareClientToWrite` is called three times in a row. --------- Signed-off-by: Lipeng Zhu <lipeng.zhu@intel.com> Co-authored-by: Wangyang Guo <wangyang.guo@intel.com>	2024-06-24 18:33:30 -07:00
Ping Xie	32ca6e5b38	Improve `CLUSTER SETSLOT` replication handling to support older replica versions. (#686 )	2024-06-23 22:08:52 -07:00
Madelyn Olson	ce79539047	Fail tests immediately if the server is no longer running (#678 ) Fix a minor inconvenience I have when writing tests. If I have a typo or forget to generate the tls certificates, the start_server handle will just loop for 2 minutes before printing the error. This just fails and prints as soon as it sees the error. Signed-off-by: Madelyn Olson <madelyneolson@gmail.com>	2024-06-21 15:29:05 +08:00
Binbin	bf1fb1fd36	Fix copy-paste error in scripts eviction test (#671 ) The test needs to test "return 2" but mistakenly uses "return 1". Also remove a extra debug print. Signed-off-by: Binbin <binloveplay1314@qq.com>	2024-06-20 10:28:47 +08:00
poiuj	0143b7c9dd	Add zfree_with_size to optimize sdsfree since we can get zmalloc_size from the header (#453 ) ### Description ### zfree updates memory statistics. It gets the size of the buffer from jemalloc by calling zmalloc_size. This operation is costly. We can avoid it if we know the buffer size. For example, we can calculate size of sds from the data we have in its header. This commit introduces zfree_with_size function that accepts both pointer to a buffer, and its size. zfree is refactored to call zfree_with_size. sdsfree uses the new interface for all but SDS_TYPE_5. ### Benchmark ### Dataset is 3 million strings. Each benchmark run uses its own value size (8192, 512, and 120). The benchmark is 100% write load for 5 minutes. ``` value size new tps old tps % new us/call old us/call % 8k 272088.53 269971.75 0.78 1.83 1.92 -4.69 512 356881.91 352856.72 1.14 1.27 1.35 -5.93 120 377523.81 368774.78 2.37 1.14 1.19 -4.20 ``` --------- Signed-off-by: Vadym Khoptynets <vadymkh@amazon.com> Signed-off-by: Madelyn Olson <madelyneolson@gmail.com> Co-authored-by: Madelyn Olson <madelyneolson@gmail.com>	2024-06-19 16:13:55 -07:00
Wen Hui	5d2348cee2	Update json file and sentinelCommand function for Valkey Sentinel (#675 ) In this PR, we update master keyword to primary keyword several in sentinel command json file and sentinelCommand function. And there is no update for configurable parameters in sentinel.conf file Signed-off-by: hwware <wen.hui.ware@gmail.com>	2024-06-19 17:16:49 -04:00
Wen Hui	e84eda9092	Remove useless code in sentinel source code (#676 ) Just remove them. Signed-off-by: hwware <wen.hui.ware@gmail.com>	2024-06-19 17:16:35 -04:00
kukey	ae2d4217e1	Add new SCRIPT SHOW subcommand to dump script via sha1 (#617 ) In some scenarios, the business may not be able to find the previously used Lua script and only have a SHA signature. Or there are multiple identical evalsha's args in monitor/slowlog, and admin is not able to distinguish the script body. Add a new script subcommmand to show the contents of script given the scripts sha1. Returns a NOSCRIPT error if the script is not present in the cache. Usage: `SCRIPT SHOW sha1` Complexity: `O(1)` Closes #604. Doc PR: https://github.com/valkey-io/valkey-doc/pull/143 --------- Signed-off-by: wei.kukey <wei.kukey@gmail.com> Signed-off-by: Madelyn Olson <madelyneolson@gmail.com> Co-authored-by: Madelyn Olson <madelyneolson@gmail.com> Co-authored-by: Binbin <binloveplay1314@qq.com> Co-authored-by: Viktor Söderqvist <viktor.soderqvist@est.tech>	2024-06-18 17:48:58 -07:00
ranshid	be2c321682	Support RDB compatability with Redis 7.2.4 RDB format (#665 ) This PR makes our current RDB format compatible with the Redis 7.2.4 RDB format. there are 2 changes introduced in this PR: 1. Move back the RDB version to 11 2. Make slot info section persist as AUX data instead of dedicated section. We have introduced slot-info as part of the work to replace cluster metadata with slot specific dictionaries. This caused us to bump the RDB version and thus we prevent downgrade (which is conceptualy O.K but better be prevented). We do not require the slot-info section to exist, so making it an AUX section will help suppport version downgrade from Valkey 8. fixes: [#645](https://github.com/valkey-io/valkey/issues/645) NOTE: tested manually by: 1. connecting Redis 7.2.4 replica to a Valkey 8(RC) 2. upgrade/downgrade Redis 7.2.4 cluster and Valkey 8(RC) cluster --------- Signed-off-by: ranshid <ranshid@amazon.com> Co-authored-by: Viktor Söderqvist <viktor.soderqvist@est.tech>	2024-06-18 12:04:06 -07:00
Binbin	a2cc2fe26d	Fix memory leak when loading slot migrations states fails (#658 ) When we goto eoferr, we need to release the auxkey and auxval, this is a cleanup, also explicitly check that decoder return value is C_ERR. Introduced in #586. Signed-off-by: Binbin <binloveplay1314@qq.com>	2024-06-17 21:18:57 -07:00
Madelyn Olson	b33f932c56	Add missing commas from debug command (#662 ) The missing commas caused the `DEBUG HELP` to be compressed onto a single line. Signed-off-by: Madelyn Olson <madelyneolson@gmail.com>	2024-06-18 12:08:08 +08:00
Ping Xie	4135894a5d	Update remaining `master` references to `primary` (#660 ) Signed-off-by: Ping Xie <pingxie@google.com>	2024-06-17 20:31:15 -07:00
Binbin	495a121f19	Adjust the log level of some logs in the cluster (#633 ) I think the log of pfail status changes will be very useful. The other parts were scanned and found that it can be modified. Changes: 1. Changing pfail status releated logs from VERBOSE to NOTICE. 2. Changing configEpoch collision log from VERBOSE(warning) to NOTICE. 3. Changing some logs from DEBUG to NOTICE. Signed-off-by: Binbin <binloveplay1314@qq.com> Co-authored-by: Madelyn Olson <madelyneolson@gmail.com>	2024-06-18 10:46:56 +08:00
Andy Pan	5a51bf5045	Combine events to eliminate redundant kevent(2) calls (#638 ) Combine events to eliminate redundant kevent(2) calls to improve performance. --------- Signed-off-by: Andy Pan <i@andypan.me>	2024-06-16 21:18:20 -07:00
Binbin	db6d3c1138	Only primary with slots has the right to mark a node as failed (#634 ) In markNodeAsFailingIfNeeded we will count needed_quorum and failures, needed_quorum is the half the cluster->size and plus one, and cluster-size is the size of primary node which contain slots, but when counting failures, we dit not check if primary has slots. Only the primary has slots that has the rights to vote, adding a new clusterNodeIsVotingPrimary to formalize this concept. Release notes: bugfix where nodes not in the quorum group might spuriously mark nodes as failed --------- Signed-off-by: Binbin <binloveplay1314@qq.com> Co-authored-by: Ping Xie <pingxie@outlook.com>	2024-06-16 20:46:08 -07:00
Sankar	a81c32079c	Make cluster meet reliable under link failures (#461 ) When there is a link failure while an ongoing MEET request is sent the sending node stops sending anymore MEET and starts sending PINGs. Since every node responds to PINGs from unknown nodes with a PONG, the receiving node never adds the sending node. But the sending node adds the receiving node when it sees a PONG. This can lead to asymmetry in cluster membership. This changes makes the sender keep sending MEET until it sees a PONG, avoiding the asymmetry. --------- Signed-off-by: Sankar <1890648+srgsanky@users.noreply.github.com>	2024-06-16 20:37:09 -07:00
Samuel Adetunji	93123f97a0	Format yaml files (#615 ) Closes #533 --------- Signed-off-by: adetunjii <adetunjithomas1@outlook.com>	2024-06-14 13:40:06 -07:00
Madelyn Olson	6faa48a358	Don't initialize the key buffer in getKeysResult (#631 ) getKeysResults is typically initialized with 2kb of zeros (16 * 256), which isn't strictly necessary since the only thing we have to initialize is some of the metadata fields. The rest of the data can remain junk as long as we don't access it. This was a bit of a regression in 7.0 with the keyspecs, since we doubled the size of the zeros, but hopefully this recovers a lot of the performance drop. I saw a modest performance bump for deep pipeline of cluster mode (~8%). I think we would see some comparable improvements in the other places where we are using it such as tracking and ACLs. --------- Signed-off-by: Madelyn Olson <matolson@amazon.com>	2024-06-14 08:42:00 -07:00
Binbin	d5496e42bc	Lua scripts promoted from eval to script load to avoid evict (#637 ) In ad28d222edcef9d4496fd7a94656013f07dd08e5, we added a Lua eval scripts eviction. If the script was previously added via EVAL, we promote it to SCRIPT LOAD, prevent it from being evicted later. Signed-off-by: Binbin <binloveplay1314@qq.com>	2024-06-14 08:32:19 -07:00
Wenwen-Chen	4c6bf30f58	Latency report: Rebranding and refine Dave dialog (#644 ) This patch try to correct the latency report. 1. Rename Redis to Valkey. 2. Remove redundant Dave dialog, and refine the output message. --------- Signed-off-by: Wenwen Chen <wenwen.chen@samsung.com> Signed-off-by: hwware <wen.hui.ware@gmail.com>	2024-06-14 12:56:59 +02:00
Ping Xie	8a776c3509	Fix potential infinite loop in `clusterNodeGetPrimary` (#651 )	2024-06-13 23:43:36 -07:00
Ping Xie	5d9d41868d	Replace `DEBUG RESTART` with `pause_server` and `resume_server` (#652 )	2024-06-13 17:52:50 -07:00
Binbin	d309b9b235	Make configs dir/dbfilename/cluster-config-file reject empty string (#636 ) Until now, these configuration items allowed typing empty strings, but empty strings behave strangely. Empty dir will fail in chdir with No such file or directory: ``` ./src/valkey-server --dir "" * FATAL CONFIG FILE ERROR (Version 255.255.255) * Reading the configuration file, at line 2 >>> 'dir ""' No such file or directory ``` Empty dbfilename will cause shutdown to fail since it will always fail in rdb save: ``` ./src/valkey-server --dbfilename "" * User requested shutdown... * Saving the final RDB snapshot before exiting. # Error moving temp DB file temp-19530.rdb on the final destination (in server root dir /xxx/xxx/valkey): No such file or directory # Error trying to save the DB, can't exit. # Errors trying to shut down the server. Check the logs for more information. ``` Empty cluster-config-file will fail in clusterLockConfig: ``` ./src/valkey-server --cluster-enabled yes --cluster-config-file "" Can't open in order to acquire a lock: No such file or directory ``` With this patch, now we will just reject it in config set like: ``` * FATAL CONFIG FILE ERROR (Version 255.255.255) * Reading the configuration file, at line 2 >>> 'xxx ""' xxx can't be empty ``` Signed-off-by: Binbin <binloveplay1314@qq.com>	2024-06-14 01:47:20 +02:00
Harkrishn Patro	76fc041685	represent cluster node flags with bitwise shift value (#642 ) While debugging a cluster bus issue, found the cluster node flags were represented in numbers. I generally find it easy when these are represented as bitwise shift operation. It improves readability a bit. Signed-off-by: Harkrishn Patro <harkrisp@amazon.com>	2024-06-14 00:58:03 +02:00
uriyage	d211078a27	Fix query buffer resized test flakiness (#646 ) Added a wait_for_condition to avoid the timing issue. ``` * [err]: query buffer resized correctly in tests/unit/querybuf.tcl Expected 11 >= 16384 && 11 <= 32770 (context: type eval line 24 cmd {assert {$orig_test_client_qbuf >= 16384 && $orig_test_client_qbuf <= $MAX_QUERY_BUFFER_SIZE}} proc ::test) * [err]: query buffer resized correctly when not idle in tests/unit/querybuf.tcl Expected 11 > 32768 (context: type eval line 14 cmd {assert {$orig_test_client_qbuf > 32768}} proc ::test) *** [err]: query buffer resized correctly with fat argv in tests/unit/querybuf.tcl query buffer should not be resized when client idle time smaller than 2s ``` Signed-off-by: Uri Yagelnik <uriy@amazon.com>	2024-06-13 18:07:07 +08:00
Harkrishn Patro	b546dd26e5	Allow CLUSTER NODES/INFO/MYID/MYSHARDID during loading state (#596 ) Allow CLUSTER subcommands NODES, INFO, MYID, MYSHARDID while loading data. It's safe to allow them and it's helpful for clients to get cluster nodes/info information during a node failover and while loading data to monitor the state of the cluster. --------- Signed-off-by: Harkrishn Patro <harkrisp@amazon.com>	2024-06-13 06:09:01 +02:00
Madelyn Olson	627d387ad8	Improve reliability of querybuf test (#639 ) We've been seeing some pretty consistent failures from `test-valgrind-test` and `test-sanitizer-address` because of the querybuf test periodically failing. I tracked it down to the test periodically taking too long and the client cron getting triggered. A simple solution is to just disable the cron during the key race condition. I was able to run this locally for 100 iterations without seeing a failure. Example: https://github.com/valkey-io/valkey/actions/runs/9474458354/job/26104103514 and https://github.com/valkey-io/valkey/actions/runs/9474458354/job/26104106830. Signed-off-by: Madelyn Olson <matolson@amazon.com>	2024-06-12 14:27:42 -07:00
Viktor Söderqvist	4bb7cc471a	Remove unnecessary clang-format off annotations (#628 ) We added some clang-format off comments before we had decided on the format configuration. Now, it turns out that turning formatting off is often not necessary. --------- Signed-off-by: Viktor Söderqvist <viktor.soderqvist@est.tech>	2024-06-12 12:52:18 +02:00
Shivshankar	e65b2d235c	Update rewriteConfigSaveOption function code to rewrite multiple save in one line. (#583 ) Currently, "config rewrite" writes some default value in the config file incase of empty config file specified. But it adds multiple "save" config entries as follows: ``` save 3600 1 save 300 100 save 60 10000 ``` After the fix the save will look like: ``` save 3600 1 300 100 60 10000 ``` --------- Signed-off-by: Shivshankar-Reddy <shiva.sheri.github@gmail.com>	2024-06-10 16:24:04 -04:00
Madelyn Olson	a3f1535b57	Fix misuse of safe iterators (#612 ) Safe iterators must call resetIterators when they are done being used. Fix one issue where a safe iterator was not correctly calling reset during cluster slot caching and fixed a second issue where reset iterator was being called twice. For the double release case, kvstoreIteratorNextDict is responsible for patching up the iterator, but we were calling it a second time in kvstoreIteratorNext. In addition, I added some documentation around initializing iterators, added an assert to prevent double initialization, and remove a function from the public interface which isn't needed and might lead to incorrect usage of the safe iterators. Bumping srgsanky for finding it here: `c4782066e7 (r142867004)`. --------- Signed-off-by: Madelyn Olson <madelyneolson@gmail.com>	2024-06-10 12:30:57 -07:00
Wen Hui	95a753b18d	Add BSD license explicitly (#620 ) Add "BSD 3-Clause License" in License 1 and License 2 part --------- Signed-off-by: hwware <wen.hui.ware@gmail.com> Co-authored-by: Madelyn Olson <madelyneolson@gmail.com>	2024-06-10 13:14:37 -04:00
Neal Gompa (ニール・ゴンパ)	71dd85dc5a	src/Makefile: Link libatomic on POWER systems (#607 ) This ensures that fallbacks for unsupported atomic operations are available for POWER systems. Signed-off-by: Neal Gompa <neal@gompa.dev>	2024-06-09 15:09:08 -07:00
skyfirelee	09b5825b26	Moving client->authenticated to a flag instead of an int (#592 ) Moving client->authenticated to a flag Fix #589 Signed-off-by: artikell <739609084@qq.com>	2024-06-09 11:49:05 -07:00
flowerysong	d28ae52004	Remove redundant function `nextPingExt()` (#613 ) Functionally identical to the older, documented `getNextPingExt()`. Fixes #610. Signed-off-by: Paul Arthur <paul.arthur@flowerysong.com>	2024-06-08 21:55:58 -07:00
Ping Xie	aad6769a80	Replicate slot migration states via RDB aux fields (#586 )	2024-06-07 20:32:27 -07:00
Ping Xie	54c9747935	Remove `master` and `slave` from source code (#591 ) External facing interfaces are not affected. --------- Signed-off-by: Ping Xie <pingxie@google.com>	2024-06-07 14:21:33 -07:00
Madelyn Olson	bce240eab7	Replace masteruser and masterauth with primaryuser and primaryauth (#598 ) Make the one backwards compatible config change we are allowed to replace for removing master from our API. `masterauth` and `masteruser` are still used as an alias, but aren't explicitly referenced. As an addendum to https://github.com/valkey-io/valkey/pull/591, it would be good to have this in 8. Given the related PR for updated other references for master, I just updated the ones around this specific change. Signed-off-by: Madelyn Olson <madelyneolson@gmail.com>	2024-06-07 00:46:52 -07:00
Viktor Söderqvist	ad5fd5b95c	More rebranding (#606 ) More rebranding of * Log messages (#252) * The DENIED error reply * Internal function names and comments, mainly Lua API --------- Signed-off-by: Viktor Söderqvist <viktor.soderqvist@est.tech>	2024-06-07 01:40:55 +02:00
Viktor Söderqvist	278ce0cae0	Rebrand the Lua debugger (#603 ) Signed-off-by: Viktor Söderqvist <viktor.soderqvist@est.tech>	2024-06-06 19:53:17 +02:00
Eran Liberty	60c10a5a4d	Remove valdup from BenchmarkDictType (#600 ) makes SERVER_CFLAGS='-DSERVER_TEST' compile as well Introduced in #443. Signed-off-by: Eran Liberty <eranl@amazon.com> Co-authored-by: Eran Liberty <eranl@amazon.com>	2024-06-05 11:00:53 +08:00
Shivshankar	9319f7aeca	Replace valkey in log and panic messages (#550 ) Part of #207 --------- Signed-off-by: Shivshankar-Reddy <shiva.sheri.github@gmail.com>	2024-06-04 20:46:59 +02:00
Eran Liberty	0700c441c6	Remove unused valDup (#443 ) Remove the unused value duplicate API from dict. It's unused in the codebase and introduces unnecessary overhead. --------- Signed-off-by: Eran Liberty <eran.liberty@gmail.com>	2024-06-03 12:22:06 -07:00
Madelyn Olson	b95e7c384f	Skip tls for xgroup read regression since it doesn't matter (#595 ) "Client blocked on XREADGROUP while stream's slot is migrated" uses the migrate command, which requires special handling for TLS and non-tls. This was not being handled, so was throwing an error. Signed-off-by: Madelyn Olson <madelyneolson@gmail.com>	2024-06-03 11:49:15 -07:00
uriyage	b72e43ed16	Adjust query buffer resized correctly test to non-jemalloc allocators. (#593 ) Test `query buffer resized correctly` start to fail (https://github.com/valkey-io/valkey/actions/runs/9278013807) with non-jemalloc allocators after https://github.com/valkey-io/valkey/pull/258 PR. With Jemalloc we allocate ~20K for the query buffer, in the test we read 1 byte in the first read, in the second read we make sure we have at least 16KB free place in the query buffer and we have as Jemalloc allocated 20KB, But with non jemalloc we allocate in the first read exactly 16KB. in the second read we check and see that we don't have 16KB free space as we already read 1 byte hence we reallocate this time greedly (*2 of the requested size of 16KB+1) hence the test condition that the querybuf size is < 32KB is no longer true The `query buffer resized correctly test` starts [failing](https://github.com/valkey-io/valkey/actions/runs/9278013807) with non-jemalloc allocators after PR #258 . With jemalloc, we allocate ~20KB for the query buffer. In the test, we read 1 byte initially and then ensure there is at least 16KB of free space in the buffer for the second read, which is satisfied by jemalloc's 20KB allocation. However, with non-jemalloc allocators, the first read allocates exactly 16KB. When we check again, we don't have 16KB free due to the 1 byte already read. This triggers a greedy reallocation (doubling the requested size of 16KB+1), causing the query buffer size to exceed the 32KB limit, thus failing the test condition. This PR adjusted the test query buffer upper limit to be 32KB +2. Signed-off-by: Uri Yagelnik <uriy@amazon.com>	2024-06-03 11:15:28 -07:00
poiuj	417660449f	Adjust sds types (#502 ) sds type should be determined based on the size of the underlying buffer, not the logical length of the sds. Currently we truncate the alloc field in case buffer is larger than we can handle. It leads to a mismatch between alloc field and the actual size of the buffer. Even considering that alloc doesn't include header size and the null terminator. It also leads to a waste of memory with jemalloc. For example, let's consider creation of sds of length 253. According to the length, the appropriate type is SDS_TYPE_8. But we allocate `253 + sizeof(struct sdshdr8) + 1` bytes, which sums to 257 bytes. In this case jemalloc allocates buffer from the next size bucket. With current configuration on Linux it's 320 bytes. So we end up with 320 bytes buffer, while we can't address more than 255. The same happens with other types and length close enough to the appropriate powers of 2. The downside of the adjustment is that with allocators that do not allocate larger than requested chunks (like GNU allocator), we switch to a larger type "too early". It leads to small waste of memory. Specifically: sds of length 31 takes 35 bytes instead of 33 (2 bytes wasted) sds of length 255 takes 261 bytes instead of 259 (2 bytes wasted) sds of length 65,535 takes 65,545 bytes instead of 65,541 (4 bytes wasted) sds of length 4,294,967,295 takes 4,294,967,313 bytes instead of 4,294,967,305 (8 bytes wasted) --------- Signed-off-by: Vadym Khoptynets <vadymkh@amazon.com>	2024-06-02 20:55:54 -07:00

... 2 3 4 5 6 ...

12567 Commits