lz4.git - LZ4 is lossless compression algorithm, providing compression speed > 500 MB/s per core, scalable with multi-cores CPU. It features an extremely fast decoder, with speed in multiple GB/s per core, typically reaching RAM speed limits on multi-core systems.

	Commit message (Collapse)	Author	Age	Files	Lines
*	lz4.c: fixed the LZ4_decompress_fast_continue case	Alexey Tourbin	2018-04-27	1	-2/+22
\| \| \| \| \| \| \|	The change is very similar to that of the LZ4_decompress_safe_continue case. The only reason a make this a separate change is to ensure that the fuzzer, after it's been enhanced, can detect the flaw in LZ4_decompress_fast_continue, and that the change indeed fixes the flaw.
*	fuzzer.c: enabled ring buffer tests for decompress_fast	Alexey Tourbin	2018-04-27	1	-40/+81
\| \| \| \| \| \| \| \| \| \| \| \| \|	Ring buffer tests were performed only with LZ4_decompress_safe_continue, leaving my buggy changes to LZ4_decompress_safe_continue undetected. The tests are now replicated and performed in a similar manner for both LZ4_decompress_safe_continue and LZ4_decompress_safe_continue (except for the small buffer case where only one function can be tested, because part of the dictionary is overwritten with the output). I also updated function names in the messages (changed them to the actual ones). The error was reported for LZ4_decompress_safe(), which I found misleading.
*	lz4.c: fixed the LZ4_decompress_safe_continue case	Alexey Tourbin	2018-04-26	2	-18/+37
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	The previous change broke decoding with a ring buffer. That's because I didn't realize that the "double dictionary mode" was possible, i.e. that the decoding routine can look both at the first part of the dictionary passed as prefix and the second part passed via dictStart+dictSize. So this change introduces the LZ4_decompress_safe_doubleDict helper, which handles this "double dictionary" situation. (This is a bit of a misnomer, there is only one dictionary, but I can't think of a better name, and perhaps the designation is not all too bad.) The helper is used only once, in LZ4_decompress_safe_continue, it should be inlined with LZ4_FORCE_O2_GCC_PPC64LE attached to LZ4_decompress_safe_continue. (Also, in the helper functions, I change the dictStart parameter type to "const void", to avoid a cast when calling helpers. In the helpers, the upcast to "BYTE" is still required, for compatibility with C++.) So this fixes the case of LZ4_decompress_safe_continue, and I'm surprised by the fact that the fuzzer is now happy and does not detect a similar problem with LZ4_decompress_fast_continue. So before fixing LZ4_decompress_fast_continue, the next logical step is to enhance the fuzzer.
*	lz4.c: refactor the decoding routines	Alexey Tourbin	2018-04-25	1	-53/+79
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	I noticed that LZ4_decompress_generic is sometimes instantiated with identical set of parameters, or (what's worse) with a subtly different sets of parameters. For example, LZ4_decompress_fast_withPrefix64k is instantiated as follows: return LZ4_decompress_generic(source, dest, 0, originalSize, endOnOutputSize, full, 0, withPrefix64k, (BYTE)dest - 64 KB, NULL, 64 KB); while the equivalent withPrefix64k call in LZ4_decompress_usingDict_generic passes 0 for the last argument instead of 64 KB. It turns out that there is no difference in this case: if you change 64 KB to 0 KB in LZ4_decompress_fast_withPrefix64k, you get the same binary code. Moreover, because it's been clarified that LZ4_decompress_fast doesn't check match offsets, it is now obvious that both of these fast/withPrefix64k instantiations are simply redundant. Exactly because LZ4_decompress_fast doesn't check offsets, it serves well with any prefixed dictionary. There's a difference, though, with LZ4_decompress_safe_withPrefix64k. It also passes 64 KB as the last argument, and if you change that to 0, as in LZ4_decompress_usingDict_generic, you get a completely different binary code. It seems that passing 0 enables offset checking: const int checkOffset = ((safeDecode) && (dictSize < (int)(64 KB))); However, the resulting code seems to run a bit faster. How come enabling extra checks can make the code run faster? Curiouser and curiouser! This needs extra study. Currently I take the view that the dictSize should be set to non-zero when nothing else will do, i.e. when passing the external dictionary via dictStart. Otherwise, lowPrefix betrays just enough information about the dictionary. * * Anyway, with this change, I instantiate all the necessary cases as functions with distinctive names, which also take fewer arguments and are therefore less error-prone. I also make the functions non-inline. (The compiler won't inline the functions because they are used more than once. Hence I attach LZ4_FORCE_O2_GCC_PPC64LE to the instances while removing from the callers.) The number of instances is now is reduced from 18 (safe+fast+partial+4continue+4prefix+4dict+2prefix64+forceExtDict) down to 7 (safe+fast+partial+2prefix+2dict). The size of the code is not the only issue here. Separate helper function are much more amenable to profile-guided optimization: it is enough to profile only a few basic functions, while the other less-often used functions, such as LZ4_decompress__continue, will benefit automatically. This is the list of LZ4_decompress functions in liblz4.so, sorted by size. Exported functions are marked with a capital T. $ nm -S lib/liblz4.so \|grep -wi T \|grep LZ4_decompress \|sort -k2 0000000000016260 0000000000000005 T LZ4_decompress_fast_withPrefix64k 0000000000016dc0 0000000000000025 T LZ4_decompress_fast_usingDict 0000000000016d80 0000000000000040 T LZ4_decompress_safe_usingDict 0000000000016d10 000000000000006b T LZ4_decompress_fast_continue 0000000000016c70 000000000000009f T LZ4_decompress_safe_continue 00000000000156c0 000000000000059c T LZ4_decompress_fast 0000000000014a90 00000000000005fa T LZ4_decompress_safe 0000000000015c60 00000000000005fa T LZ4_decompress_safe_withPrefix64k 0000000000002280 00000000000005fa t LZ4_decompress_safe_withSmallPrefix 0000000000015090 000000000000062f T LZ4_decompress_safe_partial 0000000000002880 00000000000008ea t LZ4_decompress_fast_extDict 0000000000016270 0000000000000993 t LZ4_decompress_safe_forceExtDict
*	Merge pull request #503 from lz4/l120	Yann Collet	2018-04-19	4	-66/+132
\|\ \| \| \| \|	minor length reduction of several large lines
\| *	modified indentation for consistency	Yann Collet	2018-04-19	1	-17/+33
\| \|
\| *	minor length reduction of several large lines	Yann Collet	2018-04-18	4	-65/+115
\| \|
* \|	Merge pull request #502 from lhacc1/dev	Yann Collet	2018-04-19	1	-0/+4
\|\ \ \| \|/ \|/\|	Wrap likely/unlikely macroses with #ifndef
\| *	Wrap likely/unlikely macroses with #ifndef	Dmitrii Rodionov	2018-04-18	1	-0/+4
\| \| \| \| \| \| \| \| \| \|	It prevent redefine error when project using lz4 has its own likely/unlikely macroses.
* \|	Merge pull request #497 from lz4/lowAddr	Yann Collet	2018-04-18	7	-229/+443
\|\ \ \| \|/ \|/\|	Compatibility with low memory addresses
\| *	fixed LZ4_compress_fast_extState_fastReset() in 32-bit mode	Yann Collet	2018-04-17	1	-2/+2
\| \|
\| *	fix dictDelta setting error	Yann Collet	2018-04-17	1	-1/+1
\| \| \| \| \| \| \| \|	wrong test
\| *	fix matchIndex overflow	Yann Collet	2018-04-17	1	-12/+4
\| \| \| \| \| \| \| \|	can happen with dictCtx
\| *	Merge branch 'dev' into lowAddr	Yann Collet	2018-04-17	1	-2/+9
\| \|\ \| \|/ \|/\|
* \|	Merge pull request #501 from felixhandte/fix-dict-load-offset	Yann Collet	2018-04-17	1	-2/+9
\|\ \ \| \| \| \| \| \|	Always Bump Offset by 64 KB in LZ4_loadDict()
\| * \|	Always Bump Offset by 64 KB in LZ4_loadDict()	W. Felix Handte	2018-04-17	1	-2/+9
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	This actually ensures the guarantee referred to in the comment in LZ4_compress_fast_continue().
\| \| *	fixed dictCtx compression	Yann Collet	2018-04-17	1	-7/+12
\| \| \|
\| \| *	edited a few traces for debugging	Yann Collet	2018-04-17	2	-10/+10
\| \| \|
\| \| *	fixed minor format warnings	Yann Collet	2018-04-16	1	-3/+3
\| \| \|
\| \| *	fixed fuzzer tests	Yann Collet	2018-04-16	1	-15/+17
\| \| \| \| \| \| \| \| \| \| \| \|	which were modified in parallel within branc `dev`
\| \| *	Merge branch 'dev' into lowAddr	Yann Collet	2018-04-16	4	-19/+115
\| \| \|\ \| \|_\|/ \|/\| \|
* \| \|	Merge pull request #499 from felixhandte/lz4-attach-dict-tests	Yann Collet	2018-04-13	1	-2/+99
\|\ \ \ \| \|/ / \|/\| \|	Test LZ4_attach_dictionary() and Friends
\| * \|	Further Test that ExtDictCtx Mode Produces the Exact Same Output	W. Felix Handte	2018-04-13	1	-2/+21
\| \| \|
\| * \|	Add Tests for LZ4_attach_dictionary and Friends	W. Felix Handte	2018-04-13	1	-0/+78
\|/ /
* \|	Merge pull request #496 from lz4/circleci	Yann Collet	2018-04-12	3	-17/+16
\|\ \ \| \| \| \| \| \|	Reduced LZ4 test time on circle-ci
\| * \|	modified versionsTest	Yann Collet	2018-04-12	1	-1/+1
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	to use MOREFLAGS rather CPPFLAGS as some older versions of LZ4 overwrite CPPFLAGS environment variable.
\| * \|	allow system-defined CPPFLAGS in /tests	Yann Collet	2018-04-11	1	-1/+1
\| \| \|
\| * \|	reduced test time on circle-ci	Yann Collet	2018-04-11	3	-16/+15
\|/ /
\| *	fixed gcc performance regression	Yann Collet	2018-04-16	1	-2/+4
\| \|
\| *	fixed minor unused variable warning	Yann Collet	2018-04-13	1	-3/+0
\| \|
\| *	added comment on variables required after _next_match	Yann Collet	2018-04-13	1	-0/+8
\| \|
\| *	fixed potential ptrdiff_t overflow (32-bits mode)	Yann Collet	2018-04-13	1	-14/+11
\| \| \| \| \| \| \| \|	Also removed pointer comparison, which should solve #485
\| *	compatibility with gcc-4.4 string.h version	Cyan4973	2018-04-13	2	-25/+75
\| \| \| \| \| \| \| \| \| \| \| \| \| \|	Someone found it would be a great idea to define there a global variable under the very generic name "index". Cause problem with shadow warnings, so no variable can be named "index" now ... Also : automatically update API manual
\| *	added sudo rights for low-mem-address tests	Cyan4973	2018-04-13	1	-1/+1
\| \|
\| *	fixed : counting matches which overlap extDict and prefix	test4973	2018-04-12	1	-10/+17
\| \|
\| *	modified a few traces for debug	test4973	2018-04-12	3	-7/+9
\| \|
\| *	fixed LZ4_compress_fast_extState_fastReset()	test4973	2018-04-11	1	-8/+7
\| \|
\| *	Merge branch 'dev' into lowAddr	test4973	2018-04-11	3	-67/+121
\| \|\ \| \|/ \|/\|
* \|	Merge pull request #492 from felixhandte/avoid-prepare-in-continue	Yann Collet	2018-04-11	3	-62/+115
\|\ \ \| \| \| \| \| \|	Several Changes Concerning Table Preparation in LZ4 Fast
\| * \|	Fix Silly Warning (const-ness in declaration has no effect on value types!)	W. Felix Handte	2018-04-11	1	-1/+1
\| \| \|
\| * \|	Minor Fixes	W. Felix Handte	2018-04-11	2	-11/+13
\| \| \|
\| * \|	Add a LZ4_STATIC_LINKING_ONLY Macro to Guard Experimental APIs	W. Felix Handte	2018-04-11	3	-0/+4
\| \| \|
\| * \|	Expose dictCtx Functionality in LZ4	W. Felix Handte	2018-04-11	3	-2/+33
\| \| \|
\| * \|	Rename _extState_noReset -> _extState_fastReset and Edit Comments	W. Felix Handte	2018-04-11	3	-27/+41
\| \| \|
\| * \|	Remove Extraneous Assignment (clearedTable == 0)	W. Felix Handte	2018-04-11	1	-1/+0
\| \| \|
\| * \|	Expose a Faster Stream Reset Function	W. Felix Handte	2018-04-10	3	-28/+37
\| \| \|
\| * \|	Avoid Calling LZ4_prepareTable() in LZ4_compress_fast_continue()	W. Felix Handte	2018-04-09	1	-20/+14
\|/ /
\| *	fix minor conversion warning	test4973	2018-04-10	1	-1/+1
\| \| \| \| \| \| \| \|	cast from void not implicit for C++
\| *	fixed minor conversion warning	test4973	2018-04-10	1	-1/+2
\| \| \| \| \| \| \| \|	ptr diff -> U32
\| *	Merge branch 'dev' into lowAddr	test4973	2018-04-09	2	-24/+23
\| \|\ \| \|/ \|/\|