AuroraMiddleware/zstd - zstd - Gitea: Git with a cup of tea

Author	SHA1	Message	Date
W. Felix Handte	f1b428fdac	Rename enableDedicatedDictSearch to dedicatedDictSearch in MatchState This makes it clear that not only is the feature allowed here, we're actually using it, as opposed to the CCtxParam field, in which it's enabled, but we may or may not be using it.	2020-09-10 18:51:52 -04:00
W. Felix Handte	41012193ad	Always Init CDict's enableDedicatedDictSearch Field	2020-09-10 18:51:52 -04:00
W. Felix Handte	34b545acb0	Add a ZSTD_dedicatedDictSearch ZSTD_dictMode_e to Allow Const Propagation Speed +1.5%.	2020-09-10 18:51:52 -04:00
W. Felix Handte	beefdb0d3d	Fix ZSTD_c_forceAttachDict Bounds	2020-09-10 18:51:52 -04:00
W. Felix Handte	def62e2d3e	Fix Compilation Warnings	2020-09-10 18:51:52 -04:00
Bimba Shrestha	9c628238d3	creating ZSTD_createCDict_advanced_internal	2020-09-10 18:51:52 -04:00
Bimba Shrestha	0a9787c3e1	changing to int for consistency	2020-09-10 18:51:52 -04:00
Bimba Shrestha	e29bc3a009	using dict mls instead of src mls	2020-09-10 18:51:52 -04:00
Bimba Shrestha	145c2d12f9	add hashtable head prefetching	2020-09-10 18:51:52 -04:00
Bimba Shrestha	5d5507788d	change method name for consistency	2020-09-10 18:51:52 -04:00
Bimba Shrestha	b30f71becf	pass correct cparams	2020-09-10 18:51:52 -04:00
Bimba Shrestha	71fda0362f	making cctxParams a pointer	2020-09-10 18:51:52 -04:00
Bimba Shrestha	628559d0e4	loading dict using new algorithm	2020-09-10 18:51:52 -04:00
Bimba Shrestha	22705f0c93	adding dedicatedDictSearch algorithm	2020-09-10 18:51:52 -04:00
Bimba Shrestha	31e581bf65	adding enableDedicatedDictSearch to matchState_t	2020-09-10 18:51:52 -04:00
Bimba Shrestha	50550a14ad	adding dedicated dict load method to lazy	2020-09-10 18:51:52 -04:00
Bimba Shrestha	75b6360036	adding ZSTD_createCDict_advanced2 to zstd.h	2020-09-10 18:51:52 -04:00
Bimba Shrestha	b7dddbe89b	always attach dict when using dedicatedDictSearch	2020-09-10 18:51:52 -04:00
Bimba Shrestha	e36a373df4	adding dedicatedDictSearch cParams helper methods	2020-09-10 18:51:52 -04:00
Bimba Shrestha	f10d4e313c	adding ZSTD_dedicatedDictSearch_defaultCParameters variable	2020-09-10 18:51:52 -04:00
Bimba Shrestha	c497cb6716	Add ZSTD_c_enableDedicatedDictSearch Param	2020-09-10 18:51:52 -04:00
Nick Terrell	79ded1b4a9	[lib] Add ZSTD_NO_UNUSED_FUNCTIONS macro to hide unused functions The unused function definitions are hidden behind a `#ifndef ZSTD_NO_UNUSED_FUNCTIONS` check. Initially hiding all functions which are unused and take up more than 2KB of stack space, because these will show up as warnings in the Linux Kernel build system.	2020-09-09 14:35:39 -07:00
Nick Terrell	ac3a136b0a	[lib] Replace 64-bit divisions with ZSTD_div64()	2020-09-09 14:35:39 -07:00
Nick Terrell	a90779397a	[lib] Reduce zstd stack usage by 1KB	2020-09-09 14:35:39 -07:00
Nick Terrell	046aca190f	Fix ZSTD_initCStream_advanced() with no dictionary and static allocation	2020-09-09 14:35:39 -07:00
Nick Terrell	f91ed5c766	[lib] s/current/curr because it collides with Linux Kernel macro	2020-09-09 14:35:39 -07:00
Nick Terrell	5e4efd22d4	Merge pull request #2291 from i-do-cpp/fix-compression-level-default Fix setParameter not falling back to default compression level	2020-09-08 16:42:34 -07:00
Nick Terrell	6da8acd231	Merge pull request #2293 from allanjude/coverity Resolve Coverity 1432392 Unintentional integer overflow	2020-09-03 13:58:45 -07:00
Allan Jude	8665793164	Resolve Coverity 1432392 Unintentional integer overflow Unintentional integer overflow (OVERFLOW_BEFORE_WIDEN) overflow_before_widen: Potentially overflowing expression: cdict->dictContentSize * 6U with type unsigned int (32 bits, unsigned) is evaluated using 32-bit arithmetic, and then used in a context that expects an expression of type U64 (64 bits, unsigned).	2020-09-03 19:31:50 +00:00
i-do-cpp	aec8b27fff	Update zstd_compress.c	2020-08-31 09:34:08 +02:00
i-do-cpp	d514281e73	Fix setParameter not falling back to default compression level on 0 value See documentation for `ZSTD_c_compressionLevel`: `Special: value 0 means default, which is controlled by ZSTD_CLEVEL_DEFAULT`	2020-08-31 09:25:43 +02:00
Nick Terrell	c465f24457	ZSTD_ prefix mem{cpy,move,set},malloc,calloc,free	2020-08-26 12:26:03 -07:00
Nick Terrell	a686d306d2	Rename ZSTD_{malloc,calloc,free} to ZSTD_custom{Malloc,Calloc,Free}	2020-08-26 12:25:08 -07:00
Nick Terrell	80f577baa2	Move standard includes to zstd_deps.h	2020-08-26 12:25:08 -07:00
Nick Terrell	614e446000	Merge pull request #2271 from terrelln/small-blocks Small block optimizations	2020-08-24 18:54:33 -07:00
Nick Terrell	8def0e5fd3	Fix up code after reading through	2020-08-24 12:24:45 -07:00
Nick Terrell	575731b6db	Use ncount=1 when < 4096 symbols	2020-08-18 16:47:53 -07:00
Nick Terrell	ba1fd17a9f	speed up literal header decoding	2020-08-17 12:17:53 -07:00
Nick Terrell	6004c1117f	speed up small blocks	2020-08-16 23:03:38 -07:00
Yann Collet	38e38546a4	Merge pull request #2258 from Niadb/dev Added STATIC_BMI2 for compile time detection of BMI2 on MSVC, when enabled various intrinsics are used	2020-08-04 09:43:59 -07:00
Niadb	216a63dcf7	Add files via upload	2020-07-28 02:52:52 -06:00
Yann Collet	8b9cdd2597	fixed overlapping count & workspace special case	2020-07-26 22:40:21 -07:00
Yann Collet	051232223f	optimized histogram new version easier to vectorize leads to smaller code and faster execution notably at the last recombination stage (basically, fixed cost per block). Assembly inspected with godbolt On my laptop, with `clang` and `-mavx2` : 2K block : 1280 MB/s -> 1550 MB/s 8K block : 1750 MB/s -> 1860 MB/s	2020-07-26 22:24:22 -07:00
Yann Collet	c224367ede	ensure workspace is large enough even when MAX_TABLELOG is reduced	2020-07-16 20:33:50 -07:00
Nick Terrell	1047097dad	[superblock] Add defensive assert and bounds check The bound check condition should always be met because we selected `set_basic` as our encoding type. But that code is very far away, so assert it is true so if it is ever false we can catch it, and add a bounds check. Fixes #2213.	2020-06-22 10:21:38 -07:00
Nick Terrell	08981d2638	[lib] Allow compression dictionaries with missing symbols Allow compression to use dictionaries with missing symbols in their entropy tables. We set the FSE repeat mode to check when there are missing symbols, and set the FSE repeat mode to valid when all symbols are present. Note that when not all symbols are present, the heuristics which favor dictionary tables for lower compression levels won't activate. Tested by manually creating a dictionary with missing symbols of every type, and validing that the compressor rejects it before this change, and accepts it after this change. Also, I ran the `dictionary_loader` fuzzer for >1 hour of CPU time without running into cases where compression succeeds, but decompression fails. Fixes #2174.	2020-06-12 17:57:19 -07:00
Felix Handte	2af4e07326	Merge pull request #2133 from felixhandte/single-size-calculation Consolidate CCtx Size Estimation Code	2020-05-28 13:07:18 -04:00
Nick Terrell	3cc227e90e	[ldm][mt] Fix loadedDictEnd	2020-05-19 15:55:03 -07:00
Yann Collet	fdc56baa42	fix 22294 (#2151 )	2020-05-18 21:05:10 -07:00
Nick Terrell	b2092c6dc4	[ldm] Reset loadedDictEnd when the context is reset	2020-05-18 12:35:44 -07:00
Nick Terrell	add7ed2d4a	[lib] Fix bug in loading LDM dictionary in MT mode Exposed when loading a dictionary < LDM minMatch bytes in MT mode. Test Plan: ``` CC=clang make -j zstreamtest MOREFLAGS="-O0 -fsanitize=address" ./zstreamtest -vv -i100000000 -t1 --newapi -s7065 -t3925297 ``` TODO: Add an explicit test that loads a small dictionary in MT mode	2020-05-14 11:52:28 -07:00
W. Felix Handte	3bb7992350	Fix Size Estimate for LDM Seq Space	2020-05-14 13:50:53 -04:00
Nick Terrell	70c80e19e6	[greedy] Fix performance instability	2020-05-12 17:51:16 -07:00
Nick Terrell	c3e921c639	Merge pull request #2131 from terrelln/raw-dict-fuzzer Fix rare scenario with lazy parser, dictionary, and repcodes	2020-05-12 17:44:31 -07:00
W. Felix Handte	d9a1e37aec	Nit: Fix Size Type for 32-bit	2020-05-12 18:03:31 -04:00
W. Felix Handte	1aa6c7ccce	Assert We Allocated Approximately What We Expected To	2020-05-12 16:55:03 -04:00
W. Felix Handte	27e2482217	Minor Refactor	2020-05-12 16:55:03 -04:00
W. Felix Handte	afc2488973	Handle Non-Static CCtxes in Estimation	2020-05-12 16:54:33 -04:00
W. Felix Handte	7ed996f5a0	Consolidate CCtx Size Estimation Code This commit pulls out the internals of `ZSTD_estimateCCtxSize_usingCCtxParams` into a helper. It then migrates two other callsites to use that helper, a small optimization for `ZSTD_estimateCStreamSize_usingCCtxParams`, which folds the buffer sizing into the helper, and then `ZSTD_resetCCtx_internal`, which is more invasive. This attempts to guarantee that the estimates returned to users are always correct.	2020-05-12 16:26:53 -04:00
Nick Terrell	3c1eba4d99	[lib] Fix lazy repcode validity checks	2020-05-12 12:25:06 -07:00
Nick Terrell	4e0515916d	[lib] Fix repcode validation in no dict mode	2020-05-12 11:57:15 -07:00
Nick Terrell	6d687a8816	[lib] Fix dictionary + repcodes + optimal parser	2020-05-12 10:36:53 -07:00
Nick Terrell	4b88bd3ee0	[lib][fuzz] Assert sequences are valid in round trip tests	2020-05-11 20:38:49 -07:00
Nick Terrell	80d3585e31	[lib] Fix lazy parser with dictionary + repcodes	2020-05-11 19:04:30 -07:00
Yann Collet	608f1bfc4c	fixed context downsize with initStatic When context is created using initStatic, no resize is possible. fix : only bump oversizeDuration when !initStatic	2020-05-11 18:16:38 -07:00
W. Felix Handte	c6636afbbb	Fix ZSTD_estimateCCtxSize() Under ASAN `ZSTD_estimateCCtxSize()` provides estimates for one-shot compression, which is guaranteed not to buffer inputs or outputs. So it ignores the sizes of the buffers, assuming they'll be zero. However, the actual workspace allocation logic always allocates those buffers, and when running under ASAN, the workspace surrounds every allocation with 256 bytes of redzone. So the 0-sized buffers end up consuming 512 bytes of space, which is accounted for in the actual allocation path through the use of `ZSTD_cwksp_alloc_size()` but isn't in the estimation path, since it ignores the buffers entirely. This commit fixes this.	2020-05-11 18:58:19 -04:00
Yann Collet	54144285fd	small speed improvement for strategy fast gcc 9.3.0 : kennedy : 459 -> 466 silesia : 360 -> 365 enwik8 : 267 -> 269 clang 10.0.0 : kennedy : 436 -> 441 silesia : 364 -> 366 enwik8 : 271 -> 272	2020-05-07 06:15:58 -07:00
Felix Handte	ad8dbae1b7	Merge pull request #2103 from felixhandte/relative-includes Migrate Includes to Relative Paths	2020-05-06 09:42:23 -07:00
Yann Collet	c29fd7cd8b	some more conversion warnings hunting down some static analyzer warnings	2020-05-05 10:16:59 -07:00
Yann Collet	c1b836f4c3	fix minor conversion warnings	2020-05-04 14:43:09 -07:00
W. Felix Handte	6028827fee	Rewrite Include Paths to be Relative Addresses #1998.	2020-05-04 15:20:26 -04:00
Felix Handte	7e9aabd652	Merge pull request #2099 from felixhandte/compile-under-pedantic Compile Under `-pedantic -Werror` and `-std=c90`	2020-05-04 10:07:13 -07:00
Felix Handte	816ed80774	Merge pull request #1984 from MeghnaM/1636-Reduce-stack-usage-of-HUF_sort Reduce stack usage of HUF_sort()	2020-05-04 08:15:31 -07:00
W. Felix Handte	c7da66c9cf	Purge C++-Style Comments (`// ...`), Make Compilation Succeed Under C90	2020-05-04 10:59:15 -04:00
W. Felix Handte	6696933b32	Make All Invocations Start With Literal Format String	2020-05-04 10:59:15 -04:00
W. Felix Handte	5e5f262612	Add (Possibly Empty) Info Strings to All Variadic Error Handling Macro Invocations	2020-05-04 10:58:55 -04:00
Nick Terrell	e103d7b4a6	Fix superblock mode (#2100 ) Fixes: Enable RLE blocks for superblock mode Fix the limitation that the literals block must shrink. Instead, when we're within 200 bytes of the next header byte size, we will just use the next one up. That way we should (almost?) always have space for the table. Remove the limitation that the first sub-block MUST have compressed literals and be compressed. Now one sub-block MUST be compressed (otherwise we fall back to raw block which is okay, since that is streamable). If no block has compressed literals that is okay, we will fix up the next Huffman table. Handle the case where the last sub-block is uncompressed (maybe it is very small). Before it would skip superblock in this case, now we allow the last sub-block to be uncompressed. To do this we need to regenerate the correct repcodes. Respect disableLiteralsCompression in superblock mode Fix superblock mode to handle a block consisting of only compressed literals Fix a off by 1 error in superblock mode that disabled it whenever there were last literals Fix superblock mode with long literals/matches (> 0xFFFF) Allow superblock mode to repeat Huffman tables Respect ZSTD_minGain(). Tests: Simple check for the condition in #2096. When the simple_round_trip fuzzer enables superblock mode, it checks that the compressed size isn't expanded too much. Remaining limitations: O(targetCBlockSize^2) because we recompute statistics every sequence Unable to split literals of length > targetCBlockSize into multiple sequences Refuses to generate sub-blocks that don't shrink the compressed data, so we could end up with large sub-blocks. We should emit those sections as uncompressed blocks instead. ... Fixes #2096	2020-05-01 16:11:47 -07:00
Meghna Malhotra	0adfc8dfce	Fix broken CI; make changes in response to the comments	2020-05-01 13:45:48 -07:00
Meghna Malhotra	53d76dc20f	Remove magic constant and made other changes addressing the comments	2020-05-01 13:45:48 -07:00
Meghna Malhotra	fe8402b522	WIP: Still getting an error	2020-05-01 13:45:48 -07:00
Meghna Malhotra	a084d959bd	WIP: Increased wksp size, but it's segfaulting	2020-05-01 13:45:48 -07:00
Meghna Malhotra	fdb2780c47	Move rank table into HUF_buildCTable_wksp()	2020-05-01 13:45:48 -07:00
Bimba Shrestha	1875f616ce	passing dictContentType instead of rawContent every time	2020-04-21 22:29:35 -07:00
Bimba Shrestha	5b0a452cac	Adding --long support for --patch-from (#1959 ) * adding long support for patch-from * adding refPrefix to dictionary_decompress * adding refPrefix to dictionary_loader * conversion nit * triggering log mode on chainLog < fileLog and removing old threshold * adding refPrefix to dictionary_round_trip * adding docs * adding enableldm + forceWindow test for dict * separate patch-from logic into FIO_adjustParamsForPatchFromMode * moving memLimit adjustment to outside ifdefs (need for decomp) * removing refPrefix gate on dictionary_round_trip * rebase on top of dev refPrefix change * making sure refPrefx + ldm is < 1% of srcSize * combining notes for patch-from * moving memlimit logic inside fileio.c * adding display for optimal parser and long mode trigger * conversion nit * fuzzer found heap-overflow fix * another conversion nit * moving FIO_adjustMemLimitForPatchFromMode outside ifndef * making params immutable * moving memLimit update before createDictBuffer call * making maxSrcSize unsigned long long * making dictSize and maxSrcSize params unsigned long long * error on files larger than 4gb * extend refPrefix test to include round trip * conversion to size_t * making sure ldm is at least 10x better * removing break * including zstd_compress_internal and removing redundant macros * exposing ZSTD_cycleLog() * using cycleLog instead of chainLog * add some more docs about user optimizations * formatting	2020-04-17 15:58:53 -05:00
Nick Terrell	5fcbc484c8	Merge pull request #2040 from caoyzh/dev-2 Optimize by prefetching on aarch64	2020-04-08 13:14:47 -07:00
Bimba Shrestha	c0d4b2b5a3	Merge pull request #2075 from bimbashrestha/dict_fuzzer_ref [bug] handling case where prefix is NULL or 0 sized in refPrefix_advanced	2020-04-07 17:37:19 -05:00
Bimba Shrestha	1658ae75cd	handling nil case for refprefix	2020-04-07 14:41:53 -07:00
Carl Woffenden	a93fadfcd9	Further replication removed `CHECK_F` is now in `error_private.h`. Minor tidy.	2020-04-07 11:25:16 +02:00
Carl Woffenden	7af7735fa3	Merge remote-tracking branch 'upstream/dev' into single-file-lib	2020-04-07 11:13:02 +02:00
Carl Woffenden	edd9a07322	Code replicated in compression and decompression moved to shared headers `CHECK_F` macro moved to `error_private.h` (shared between `fse_compress.c` and `fse_decompress.c`). `ZSTD_limitCopy()` moved to `zstd_internal.h` (shared between `zstd_compress.c` and `zstd_decompress.c`). Erroneous build artefact `zstd.h` removed from repo.	2020-04-07 11:02:06 +02:00
Bimba Shrestha	0154866749	moving consts to zstd_internal and reusing them	2020-04-03 14:26:15 -07:00
Carl Woffenden	7c420344d2	Single-file decoder script can now (optionally) create an encoder To complement the single-file decoder a new script was added to create an amalgamated single-file of all of the Zstd source, along with examples and (simple) tests.	2020-04-03 19:07:46 +02:00
Nick Terrell	ac58c8d720	Fix copyright and license lines * All copyright lines now have -2020 instead of -present * All copyright lines include "Facebook, Inc" * All licenses are now standardized The copyright in `threading.{h,c}` is not changed because it comes from zstdmt. The copyright and license of `divsufsort.{h,c}` is not changed.	2020-03-26 17:02:06 -07:00
Nick Terrell	d34204a7b7	Merge pull request #2029 from terrelln/minor-opt [opt] Update repcodes less often	2020-03-23 18:12:32 -07:00
caoyzh	7201980650	Optimize by prefetching on aarch64	2020-03-14 15:25:59 +08:00
Bimba Shrestha	66607d0eac	Merge pull request #2033 from bimbashrestha/icc [opt] Small icc level 1 compression speed gain using #pragma vector	2020-03-10 20:42:19 -05:00
Bimba Shrestha	a89c45bdbd	Typo	2020-03-10 15:19:48 -05:00
Bimba Shrestha	43fc88f443	Adding comment and remvoing ivdep	2020-03-10 14:57:27 -05:00
Bimba Shrestha	dba3abc95a	Missed returns	2020-03-05 12:20:59 -08:00
Bimba Shrestha	a75e5f2ffc	bitscan add undef check	2020-03-05 11:52:15 -08:00

1 2 3 4 5 ...

1596 Commits