lz4.git - LZ4 is lossless compression algorithm, providing compression speed > 500 MB/s per core, scalable with multi-cores CPU. It features an extremely fast decoder, with speed in multiple GB/s per core, typically reaching RAM speed limits on multi-core systems.

	Commit message (Collapse)	Author	Age	Files	Lines
*	Merge pull request #542 from wbx-github/dev	Yann Collet	2018-05-29	1	-3/+4
\|\ \| \| \| \|	allow to override uname when cross-compiling
\| *	allow to override uname when cross-compiling	Waldemar Brodkorb	2018-05-22	1	-3/+4
\| \| \| \| \| \| \| \| \| \| \| \|	When cross-compiling for example from Darwin to Linux it might be useful to override uname output to force Linux and create Linux libraries instead of Darwin libraries.
* \|	Also Fix Appveyor Cast Warning	W. Felix Handte	2018-05-22	1	-1/+1
\| \|
* \|	Add `extern "C"` Guards Around Experimental HC Declarations	W. Felix Handte	2018-05-22	1	-0/+8
\| \|
* \|	Remove #define-rename of `LZ4_decompress_safe_forceExtDict`	W. Felix Handte	2018-05-22	1	-8/+8
\| \|
* \|	Test Linking C-Compiled Library and C++-Compiled Tests	W. Felix Handte	2018-05-22	1	-0/+15
\|/
*	Add Haiku as a validated target.	fbrosson	2018-05-17	1	-1/+1
\| \| \| \|	lz4 1.8.2 works fine on Haiku and passes all tests.
*	Merge pull request #537 from lz4/xpHCmf2	Yann Collet	2018-05-07	1	-48/+81
\|\ \| \| \| \|	Speed optimization for optimal parser
\| *	renamed variable for clarity	Yann Collet	2018-05-07	1	-12/+12
\| \|
\| *	fixed minor conversion warning	Yann Collet	2018-05-07	1	-1/+2
\| \|
\| *	small PA optimization	Yann Collet	2018-05-06	1	-11/+18
\| \| \| \| \| \| \| \| \| \|	which measurably improves speed on levels 9+
\| *	lz4hc: fixed PA / SC parameter order	Yann Collet	2018-05-05	1	-5/+5
\| \| \| \| \| \| \| \| \| \| \| \|	also : reserved PA for levels 9+ (instead of 8+). In most cases, speed is lower, and compression benefit is not worth.
\| *	lz4hc: SC only enabled for opt parser	Yann Collet	2018-05-05	1	-7/+7
\| \| \| \| \| \| \| \| \| \|	the trade off is not good for regular HC parser : compression is a little bit better, but speed cost is too large in comparison.
\| *	fixed SC.opt integration with regular HC parser	Yann Collet	2018-05-05	1	-4/+4
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	Only enabled when searching forward. note : it slighly improves compression ratio, but measurably decreases speed. Trade-off to analyse.
\| *	lz4hc: fixed performance issue	Yann Collet	2018-05-05	1	-114/+20
\| \| \| \| \| \| \| \|	when combining both PA and CS optimizations
\| *	integrated chain swapper into HC match finder	Yann Collet	2018-05-05	1	-45/+76
\| \| \| \| \| \| \| \| \| \| \| \| \| \|	slower than expected Pattern analyzer and Chain Swapper work slower when both activated. Reasons unclear.
\| *	implemented search accelerator	Yann Collet	2018-05-03	1	-2/+18
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	greatly improves speed compared to non-accelerated, especially for slower files. On my laptop, -b12 : ``` calgary.tar : 4.3 MB/s => 9.0 MB/s enwik7 : 10.2 MB/s => 13.3 MB/s silesia.tar : 4.0 MB/s => 8.7 MB/s ``` Note : this is the simplified version, without handling dictionaries, external buffer, nor pattern analyzer. Current `dev` branch on these samples gives : ``` calgary.tar : 4.2 MB/s enwik7 : 9.7 MB/s silesia.tar : 3.5 MB/s ``` interestingly, it's slower, presumably due to handling of dictionaries.
\| *	created LZ4HC_FindLongestMatch()	Yann Collet	2018-05-03	1	-16/+88
\| \| \| \| \| \| \| \| \| \| \| \|	simplified match finder only searching forward and within current buffer, for easier testing of optimizations.
* \|	Merge pull request #538 from lz4/frameTestError	Yann Collet	2018-05-07	2	-3/+19
\|\ \ \| \| \| \| \| \|	Fix frametest error
\| * \|	small extDict : fixed side-effect	Yann Collet	2018-05-06	2	-3/+7
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	don't fix dictionaries of size 0. setting dictEnd == source triggers prefix mode, thus removing possibility to use CDict.
\| * \|	fixed frametest error	Yann Collet	2018-05-06	2	-1/+13
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	The error can be reproduced using following command : ./frametest -v -i100000000 -s1659 -t31096808 It's actually a bug in the stream LZ4 API, when starting a new stream and providing a first chunk to complete with size < MINMATCH. In which case, the chunk becomes a dictionary. No hash was generated and stored, but the chunk is accessible as default position 0 points to dictStart, and position 0 is still within MAX_DISTANCE. Then, next attempt to read 32-bits from position 0 fails. The issue would have been mitigated by starting from index 64 KB, effectively eliminating position 0 as too far away. The proper fix is to eliminate such "dictionary" as too small. Which is what this patch does.
* \| \|	Fix make install	Nick Terrell	2018-05-04	1	-34/+36
\|/ / \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	* Uninstall didn't remove the pkg-config correctly. * Fix `mandir` * Allow overriding either upper- or lower-case location variables, but always use the lower case variables. * Add test case that ensures overriding both upper- and lower-case variables is the same, and that the directory is empty after uninstall.
* \|	Merge pull request #529 from felixhandte/lz4f-fast-reset-for-streaming-only	Yann Collet	2018-05-03	2	-10/+33
\|\ \ \| \|/ \|/\|	LZ4F: Only Reset the LZ4_stream_t when Init'ing a Streaming Block
\| *	Only Reset the LZ4 Stream when Init'ing a Streaming Block	W. Felix Handte	2018-05-03	2	-10/+33
\| \|
* \|	Merge branch 'dev' into lz4fRingBuffer	Yann Collet	2018-05-03	1	-10/+8
\|\ \
\| * \	Merge pull request #528 from lz4/complexShortcut	Yann Collet	2018-05-03	2	-49/+108
\| \|\ \ \| \| \| \| \| \| \| \|	Faster decoding speed
\| \| * \|	fix comments / indentation	Cyan4973	2018-05-03	1	-10/+8
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	as requested by @terrelln
* \| \| \|	random lz4f clarifications	Yann Collet	2018-05-02	1	-29/+47
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	the initial intention was to update lz4f ring buffer strategy, but lz4f doesn't use ring buffer. Instead, it uses the destination buffer as much as possible, and merely copies just what's required to preserve history into its own buffer, at the end. Pretty efficient. This patch just clarifies a few comments and add some assert(). It's built on top of #528. It also updates doc.
* \| \| \|	Merge branch 'dev' into lz4fRingBuffer	Yann Collet	2018-05-02	1	-13/+13
\|\ \ \ \ \| \|/ / / \| \| / / \| \|/ / \|/\| \|
\| * \|	increased nbAttempts for lz4 -12	Yann Collet	2018-05-02	1	-13/+13
\| \| \| \| \| \| \| \| \| \| \| \|	shaves one more kilobyte from silesia.tar
* \| \|	introduce LZ4_decoderRingBufferSize()	Yann Collet	2018-05-02	2	-26/+63
\| \| \| \| \| \| \| \| \| \| \| \|	fuzzer : fix and robustify ring buffer tests
* \| \|	simplify shortcut	Yann Collet	2018-05-02	1	-55/+22
\| \| \|
* \| \|	Merge branch 'dev' into complexShortcut	Yann Collet	2018-05-02	2	-53/+55
\|\ \ \ \| \|/ /
\| * \|	Merge pull request #521 from lz4/BD_deterministic	Yann Collet	2018-05-01	1	-48/+46
\| \|\ \ \| \| \| \| \| \| \| \|	fix lz4hc -BD non-determinism
\| \| * \|	renamed variable for clarity	Cyan4973	2018-05-01	1	-7/+7
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	lowLimit -> lowestMatchIndex
\| \| * \|	lz4hc changed variable	Yann Collet	2018-04-30	1	-2/+2
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	to reduce confusion dictLowLimit => dictStart
\| \| * \|	Merge branch 'dev' into BD_deterministic	Yann Collet	2018-04-27	5	-38/+68
\| \| \|\ \
\| \| * \| \|	fix lz4hc -BD non-determinism	Yann Collet	2018-04-27	1	-3/+2
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	related to chain table update
\| \| * \| \|	lz4hc : minor editions for clarity	Yann Collet	2018-04-27	1	-38/+37
\| \| \| \| \|
\| * \| \| \|	clarified streaming decompression function	Yann Collet	2018-04-30	1	-5/+9
\| \| \|_\|/ \| \|/\| \| \| \| \| \| \| \| \| \|	restrictions for ring buffer
* \| \| \|	Merge pull request #527 from svpv/fastDec	Yann Collet	2018-04-30	1	-25/+82
\|\ \ \ \ \| \|/ / / \|/\| \| \|	lz4.c: two-stage shortcut for LZ4_decompress_generic
\| * \| \|	lz4.c: two-stage shortcut for LZ4_decompress_generic	Alexey Tourbin	2018-04-28	1	-25/+82
\| \| \| \|
* \| \| \|	Merge pull request #523 from svpv/makeV1	Yann Collet	2018-04-29	1	-28/+34
\|\ \ \ \ \| \| \| \| \| \| \| \| \| \|	lib/Makefile: show commands with V=1
\| * \| \| \|	lib/Makefile: show commands with V=1	Alexey Tourbin	2018-04-28	1	-28/+34
\| \|/ / / \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	`make V=1` will now show the commands executed to build the library. A similar technique is used in e.g. linux/Makefile. The bulk of this change is produced with the following vim command: :g!/^\t@echo\>/s/^\t@/\t\$(Q)/
* \| \| \|	Merge pull request #515 from svpv/refactorDec	Yann Collet	2018-04-29	1	-49/+114
\|\ \ \ \ \| \|/ / / \|/\| \| \|	lz4.c: refactor the decoding routines
\| * \| \|	lz4.c: fixed the LZ4_decompress_fast_continue case	Alexey Tourbin	2018-04-27	1	-2/+22
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	The change is very similar to that of the LZ4_decompress_safe_continue case. The only reason a make this a separate change is to ensure that the fuzzer, after it's been enhanced, can detect the flaw in LZ4_decompress_fast_continue, and that the change indeed fixes the flaw.
\| * \| \|	lz4.c: fixed the LZ4_decompress_safe_continue case	Alexey Tourbin	2018-04-26	1	-17/+36
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	The previous change broke decoding with a ring buffer. That's because I didn't realize that the "double dictionary mode" was possible, i.e. that the decoding routine can look both at the first part of the dictionary passed as prefix and the second part passed via dictStart+dictSize. So this change introduces the LZ4_decompress_safe_doubleDict helper, which handles this "double dictionary" situation. (This is a bit of a misnomer, there is only one dictionary, but I can't think of a better name, and perhaps the designation is not all too bad.) The helper is used only once, in LZ4_decompress_safe_continue, it should be inlined with LZ4_FORCE_O2_GCC_PPC64LE attached to LZ4_decompress_safe_continue. (Also, in the helper functions, I change the dictStart parameter type to "const void", to avoid a cast when calling helpers. In the helpers, the upcast to "BYTE" is still required, for compatibility with C++.) So this fixes the case of LZ4_decompress_safe_continue, and I'm surprised by the fact that the fuzzer is now happy and does not detect a similar problem with LZ4_decompress_fast_continue. So before fixing LZ4_decompress_fast_continue, the next logical step is to enhance the fuzzer.
\| * \| \|	lz4.c: refactor the decoding routines	Alexey Tourbin	2018-04-25	1	-53/+79
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	I noticed that LZ4_decompress_generic is sometimes instantiated with identical set of parameters, or (what's worse) with a subtly different sets of parameters. For example, LZ4_decompress_fast_withPrefix64k is instantiated as follows: return LZ4_decompress_generic(source, dest, 0, originalSize, endOnOutputSize, full, 0, withPrefix64k, (BYTE)dest - 64 KB, NULL, 64 KB); while the equivalent withPrefix64k call in LZ4_decompress_usingDict_generic passes 0 for the last argument instead of 64 KB. It turns out that there is no difference in this case: if you change 64 KB to 0 KB in LZ4_decompress_fast_withPrefix64k, you get the same binary code. Moreover, because it's been clarified that LZ4_decompress_fast doesn't check match offsets, it is now obvious that both of these fast/withPrefix64k instantiations are simply redundant. Exactly because LZ4_decompress_fast doesn't check offsets, it serves well with any prefixed dictionary. There's a difference, though, with LZ4_decompress_safe_withPrefix64k. It also passes 64 KB as the last argument, and if you change that to 0, as in LZ4_decompress_usingDict_generic, you get a completely different binary code. It seems that passing 0 enables offset checking: const int checkOffset = ((safeDecode) && (dictSize < (int)(64 KB))); However, the resulting code seems to run a bit faster. How come enabling extra checks can make the code run faster? Curiouser and curiouser! This needs extra study. Currently I take the view that the dictSize should be set to non-zero when nothing else will do, i.e. when passing the external dictionary via dictStart. Otherwise, lowPrefix betrays just enough information about the dictionary. * * Anyway, with this change, I instantiate all the necessary cases as functions with distinctive names, which also take fewer arguments and are therefore less error-prone. I also make the functions non-inline. (The compiler won't inline the functions because they are used more than once. Hence I attach LZ4_FORCE_O2_GCC_PPC64LE to the instances while removing from the callers.) The number of instances is now is reduced from 18 (safe+fast+partial+4continue+4prefix+4dict+2prefix64+forceExtDict) down to 7 (safe+fast+partial+2prefix+2dict). The size of the code is not the only issue here. Separate helper function are much more amenable to profile-guided optimization: it is enough to profile only a few basic functions, while the other less-often used functions, such as LZ4_decompress__continue, will benefit automatically. This is the list of LZ4_decompress functions in liblz4.so, sorted by size. Exported functions are marked with a capital T. $ nm -S lib/liblz4.so \|grep -wi T \|grep LZ4_decompress \|sort -k2 0000000000016260 0000000000000005 T LZ4_decompress_fast_withPrefix64k 0000000000016dc0 0000000000000025 T LZ4_decompress_fast_usingDict 0000000000016d80 0000000000000040 T LZ4_decompress_safe_usingDict 0000000000016d10 000000000000006b T LZ4_decompress_fast_continue 0000000000016c70 000000000000009f T LZ4_decompress_safe_continue 00000000000156c0 000000000000059c T LZ4_decompress_fast 0000000000014a90 00000000000005fa T LZ4_decompress_safe 0000000000015c60 00000000000005fa T LZ4_decompress_safe_withPrefix64k 0000000000002280 00000000000005fa t LZ4_decompress_safe_withSmallPrefix 0000000000015090 000000000000062f T LZ4_decompress_safe_partial 0000000000002880 00000000000008ea t LZ4_decompress_fast_extDict 0000000000016270 0000000000000993 t LZ4_decompress_safe_forceExtDict
* \| \| \|	Merge pull request #520 from felixhandte/frame-dict-nits	Yann Collet	2018-04-27	2	-8/+15
\|\ \ \ \ \| \|_\|_\|/ \|/\| \| \|	Minor Fixes to Dictionary Preparation in LZ4 Frame
\| * \| \|	Avoid Possibly Redundant Table Clears When Loading HC Dict	W. Felix Handte	2018-04-27	2	-2/+2
\| \| \| \|