AuroraMiddleware/zstd - zstd - Gitea: Git with a cup of tea

Author	SHA1	Message	Date
Yann Collet	a7ef3a219c	zstdmt : fixed last job size	2018-01-19 18:19:09 -08:00
Yann Collet	3ad7d4951c	zstdmt : finally vanquished an elusive and rare race condition	2018-01-19 17:35:08 -08:00
Yann Collet	940634a610	zstdmt : simplify job creation job will not be created when not enough room within job Table	2018-01-19 13:25:06 -08:00
Yann Collet	dc69623453	zstdmt: fixed corruption issue in ZSTDMT_endStream() when invoked directly.	2018-01-19 12:41:56 -08:00
Yann Collet	70f81d6030	zstdmt uses POOL_tryAdd() to call a new worker so that it's no longer a blocking call. This makes it possible to stream out data gradually, while waiting for a worker to become available.	2018-01-19 10:01:40 -08:00
Yann Collet	6f7280fb33	fixed frame checksum issue and race conditions	2018-01-18 16:20:26 -08:00
Yann Collet	997e4d0ccd	added POOL_tryAdd()	2018-01-18 14:39:51 -08:00
Yann Collet	ef97d5a287	Merge branch 'progressiveMT' into progressiveFlush	2018-01-18 13:35:24 -08:00
Yann Collet	b6ab232f2d	Merge branch 'dev' into progressiveMT	2018-01-18 13:34:56 -08:00
Nick Terrell	9d96761520	Set repcodes for empty ZSTD_CDict When the dictionary is <= 8 bytes, no data is loaded from the dictionary. In this case the repcodes weren't set, because they were inserted after the size check. Fix this problem in general by first setting the cdict state to a clean state of an empty dictionary, then filling the state from there.	2018-01-18 13:28:30 -08:00
Yann Collet	c7190c69cc	fixes for @terrelln comments	2018-01-18 11:15:23 -08:00
Yann Collet	1b5d80d633	zstdmt: added ability to flush current job before it's completed however, zstdmt may still wait on next available worker, so it's not smooth yet.	2018-01-18 11:03:27 -08:00
Yann Collet	aa79c18e3f	fixed a few access contention passes thread sanitizer test	2018-01-17 17:18:19 -08:00
Yann Collet	394eec697b	Introduce ZSTD_getFrameProgression() Produces 3 statistics for ongoing frame compression : - ingested - consumed (effectively compressed) - produced Ingested can be larger than consumed due to buffering effect. For the time being, this patch mostly fixes the % ratio issue, since it computes consumed / produced, instead of ingested / produced. That being said, update is not "smooth", because on a slow enough setting, fileio spends most of its time waiting for a worker to complete its job. This could be improved thanks to more granular flushing i.e. start flushing before ongoing job is fully completed.	2018-01-17 16:39:02 -08:00
Yann Collet	b86865323a	Merge branch 'dev' into progressiveMT fixed minor conflict on cdict	2018-01-17 13:51:03 -08:00
Yann Collet	d14cc881b0	zstdmt : fixed very large window sizes would create too large buffers, since default job size == window size * 4. This would crash on 32-bit systems. Also : jobSize being a 32-bit unsigned, it cannot be >= 4 GB, so the formula was failing for large window sizes >= 1 GB. Fixed now : max job Size is 2 GB, whatever the window size.	2018-01-17 12:39:58 -08:00
Yann Collet	58dd7de640	zstdmt: fixed an endless loop on allocation failure this happened on 32-bits build when requiring a too large input buffer, typically on wlog=29, creating jobs of 2 GB size. also : zstd32 now compiles with multithread support enabled by default (can be disabled with HAVE_THREAD=0)	2018-01-17 12:10:15 -08:00
Nick Terrell	16bd0fd4df	Reduce size of ZSTD_CDict Shaves 492,076 B off of the `ZSTD_CDict`. The size of a `ZSTD_CDict` created from a 112,640 B dictionary is: \| Level \| Before (B) \| After (B) \| \|-------\|------------\|-----------\| \| 1 \| 648,448 \| 156,412 \| \| 3 \| 1,140,008 \| 647,932 \|	2018-01-17 11:50:49 -08:00
Yann Collet	cb57c107ff	zstdmt: minor variable renaming, for clarity	2018-01-17 11:39:07 -08:00
Yann Collet	1dba98d563	introduced parameter ZSTD_p_nonBlockingMode This new parameter makes it possible to call streaming ZSTDMT with a single thread set which is non blocking. It makes it possible for the main thread to do other tasks in parallel while the worker thread does compression. Typically, for zstd cli, it means it can do I/O stuff. Applied within fileio.c, this patch provides non-negligible gains during compression. Tested on my laptop, with enwik9 (1000000000 bytes) : time zstd -f enwik9 With traditional single-thread blocking mode : real 0m9.557s user 0m8.861s sys 0m0.538s With new single-worker non blocking mode : real 0m7.938s user 0m8.049s sys 0m0.514s => 20% faster	2018-01-16 16:15:47 -08:00
Yann Collet	6025465e42	ZSTDMT : minor CCtx memory optimization can be useful when a compression job only has small amount of data to compress.	2018-01-16 15:34:41 -08:00
Yann Collet	2e23333094	ZSTDMT can now work in non-blocking mode with 1 thread it still fallbacks to single-thread blocking invocation when input is small (<1job) or when invoking ZSTDMT_compress(), which is blocking. Also : fixed a bug in new block-granular compression routine.	2018-01-16 15:28:43 -08:00
Yann Collet	8e83c5c910	Merge branch 'dev' into progressiveMT	2018-01-16 12:54:33 -08:00
Nick Terrell	aae267a2e1	Reorganize block state	2018-01-16 11:17:50 -08:00
Nick Terrell	887cd4e35e	Split ZSTD_CCtx into smaller sub-structures	2018-01-16 11:17:50 -08:00
Yann Collet	9477f6529d	Merge pull request #984 from terrelln/dict-load Load more dictionary positions into table if empty	2018-01-13 13:20:42 -08:00
Yann Collet	58ecf13e02	zstdmt : can compress at block granularity offering perspective of more accurate progression report.	2018-01-13 13:18:57 -08:00
Nick Terrell	9a211d1f05	Load more dictionary positions into table if empty If the hash table is empty load positions into the hash table that we would otherwise skip. \| Level \| Data Set \| Improvement \| \|-------\|--------------\|-------------\| \| 1 \| github \| 0.44% \| \| 1 \| hg-changelog \| 0.13% \| \| 1 \| hg-commands \| 1.28% \| \| 1 \| hg-manifest \| 0.70% \| \| 3 \| github \| 0.74% \| \| 3 \| hg-changelog \| 0.87% \| \| 3 \| hg-commands \| 1.74% \| \| 3 \| hg-manifest \| 0.23% \|	2018-01-12 16:17:22 -08:00
Yann Collet	863b2f8db4	Merge pull request #983 from terrelln/dict-wlog Increase windowLog from CDict based on the srcSize when known	2018-01-12 07:47:43 -08:00
Nick Terrell	b610b777d3	Increase windowLog from CDict based on the srcSize when known	2018-01-11 16:23:21 -08:00
Yann Collet	cacf47cbee	Merge branch 'dev' into dubtlazy and fixed conflicts	2018-01-11 13:25:08 -08:00
Yann Collet	04c00f9388	Merge pull request #982 from facebook/fix304 Fix for #304 and #977 : error during dictionary creation	2018-01-11 13:20:59 -08:00
Yann Collet	b9a14900ff	changed function name to ZSTD_DUBT_findBestMatch()	2018-01-11 12:38:31 -08:00
Yann Collet	752bae4a48	added warning message when pathological dataset is detected (note : cover_optimize needs -v to display the warning)	2018-01-11 11:29:28 -08:00
Yann Collet	e8093dde09	fixed #304 Pathological samples may result in literal section being incompressible. This case is now detected, and literal distribution is replaced by one that can be written into the dictionary.	2018-01-11 11:16:32 -08:00
Yann Collet	218e9fe0fc	added a test case for dictBuilder failure cyclic data set makes the entropy stage fails now, onto a fix for #304 ...	2018-01-11 09:42:38 -08:00
Yann Collet	ff795580f2	fixed bug #976 , reported by @indygreg constants in zstd.h should not depend on MIN() macro which existence is not guaranteed. Added a test to check the specific constants. The test is a bit too specific. But I have found no way to control a more generic "are all macro already defined" condition, especially as this is a valid construction (the missing macro might be defined later, intentionnally).	2018-01-10 20:33:45 -08:00
Yann Collet	292eeb672f	api doc : grouped all ZSTD_create*_advanced() functions together in a new "custom memory allocator" paragraph which is itself part of "memory management" category. This makes it simpler to see the relation between the type and its usages.	2018-01-10 09:07:47 -08:00
Yann Collet	3ea156368c	API doc : grouped ZSTD_initStatic*() together within "memory management" category.	2018-01-10 08:49:50 -08:00
Yann Collet	b17fb488b0	fixed msan test a pointer calculation was wrong in a corner case	2018-01-06 20:50:36 +01:00
Yann Collet	658d6b8588	Merge branch 'dev' into dubtlazy	2018-01-06 12:40:58 +01:00
Yann Collet	a927fae2a1	fixed ZSTD_reduceIndex() following suggestions from @terrelln. Also added some comments to present logic behind ZSTD_preserveUnsortedMark().	2018-01-06 12:31:26 +01:00
Yann Collet	2eff217136	updated /lib documentation	2017-12-31 15:50:00 +01:00
Yann Collet	00db4dbbb3	fixed minor argument property for Visual	2017-12-30 15:42:28 +01:00
Yann Collet	f597f55675	improved btlazy2 : list of unsorted candidates can reach extDict It used to stop on reaching extDict, for simplification. As a consequence, there was a small loss of performance each time the round buffer would restart from beginning. It's not a large difference though, just several hundreds of bytes on silesia. This patch fixes it.	2017-12-30 15:12:59 +01:00
Yann Collet	a68b76afef	updated compression level table for btlazy2 now selected for levels 13, 14 and 15. Also : dropped the requirement for monotonic memory budget increase of compression levels,, which was required for ZSTD_estimateCCtxSize() in order to ensure that a memory budget for level L is large enough for any level <= L. This condition is now ensured at run time inside ZSTD_estimateCCtxSize().	2017-12-30 11:40:35 +01:00
Yann Collet	eb52e2f45e	simplify ZSTD_preserveUnsortedMark() implementation since no compiler attempts to auto-vectorize it.	2017-12-30 11:13:52 +01:00
Yann Collet	d228b6b0d0	btlazy2 : optimization for dictionary compression we want the dictionary table to be fully sorted, not just lazily filled. Dictionary loading is a bit more intensive, but it saves cpu cycles for match search during compression.	2017-12-29 19:14:18 +01:00
Yann Collet	02f64ef955	btlazy2: fixed interaction between unsortedMark and reduceTable	2017-12-29 19:08:51 +01:00
Yann Collet	64482c2c97	fixed bug in dubt the chain of unsorted candidates could grow beyond lowLimit.	2017-12-29 17:04:37 +01:00

1 2 3 4 5 ...

2032 Commits