summaryrefslogtreecommitdiffstats
path: root/lib
Commit message (Collapse)AuthorAgeFilesLines
...
| * | | | clarified streaming decompression functionYann Collet2018-04-301-5/+9
| | |_|/ | |/| | | | | | | | | | restrictions for ring buffer
* | | | Merge pull request #527 from svpv/fastDecYann Collet2018-04-301-25/+82
|\ \ \ \ | |/ / / |/| | | lz4.c: two-stage shortcut for LZ4_decompress_generic
| * | | lz4.c: two-stage shortcut for LZ4_decompress_genericAlexey Tourbin2018-04-281-25/+82
| | | |
* | | | Merge pull request #523 from svpv/makeV1Yann Collet2018-04-291-28/+34
|\ \ \ \ | | | | | | | | | | lib/Makefile: show commands with V=1
| * | | | lib/Makefile: show commands with V=1Alexey Tourbin2018-04-281-28/+34
| |/ / / | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | `make V=1` will now show the commands executed to build the library. A similar technique is used in e.g. linux/Makefile. The bulk of this change is produced with the following vim command: :g!/^\t@echo\>/s/^\t@/\t\$(Q)/
* | | | Merge pull request #515 from svpv/refactorDecYann Collet2018-04-291-49/+114
|\ \ \ \ | |/ / / |/| | | lz4.c: refactor the decoding routines
| * | | lz4.c: fixed the LZ4_decompress_fast_continue caseAlexey Tourbin2018-04-271-2/+22
| | | | | | | | | | | | | | | | | | | | | | | | | | | | The change is very similar to that of the LZ4_decompress_safe_continue case. The only reason a make this a separate change is to ensure that the fuzzer, after it's been enhanced, can detect the flaw in LZ4_decompress_fast_continue, and that the change indeed fixes the flaw.
| * | | lz4.c: fixed the LZ4_decompress_safe_continue caseAlexey Tourbin2018-04-261-17/+36
| | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | The previous change broke decoding with a ring buffer. That's because I didn't realize that the "double dictionary mode" was possible, i.e. that the decoding routine can look both at the first part of the dictionary passed as prefix and the second part passed via dictStart+dictSize. So this change introduces the LZ4_decompress_safe_doubleDict helper, which handles this "double dictionary" situation. (This is a bit of a misnomer, there is only one dictionary, but I can't think of a better name, and perhaps the designation is not all too bad.) The helper is used only once, in LZ4_decompress_safe_continue, it should be inlined with LZ4_FORCE_O2_GCC_PPC64LE attached to LZ4_decompress_safe_continue. (Also, in the helper functions, I change the dictStart parameter type to "const void*", to avoid a cast when calling helpers. In the helpers, the upcast to "BYTE*" is still required, for compatibility with C++.) So this fixes the case of LZ4_decompress_safe_continue, and I'm surprised by the fact that the fuzzer is now happy and does not detect a similar problem with LZ4_decompress_fast_continue. So before fixing LZ4_decompress_fast_continue, the next logical step is to enhance the fuzzer.
| * | | lz4.c: refactor the decoding routinesAlexey Tourbin2018-04-251-53/+79
| | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | I noticed that LZ4_decompress_generic is sometimes instantiated with identical set of parameters, or (what's worse) with a subtly different sets of parameters. For example, LZ4_decompress_fast_withPrefix64k is instantiated as follows: return LZ4_decompress_generic(source, dest, 0, originalSize, endOnOutputSize, full, 0, withPrefix64k, (BYTE*)dest - 64 KB, NULL, 64 KB); while the equivalent withPrefix64k call in LZ4_decompress_usingDict_generic passes 0 for the last argument instead of 64 KB. It turns out that there is no difference in this case: if you change 64 KB to 0 KB in LZ4_decompress_fast_withPrefix64k, you get the same binary code. Moreover, because it's been clarified that LZ4_decompress_fast doesn't check match offsets, it is now obvious that both of these fast/withPrefix64k instantiations are simply redundant. Exactly because LZ4_decompress_fast doesn't check offsets, it serves well with any prefixed dictionary. There's a difference, though, with LZ4_decompress_safe_withPrefix64k. It also passes 64 KB as the last argument, and if you change that to 0, as in LZ4_decompress_usingDict_generic, you get a completely different binary code. It seems that passing 0 enables offset checking: const int checkOffset = ((safeDecode) && (dictSize < (int)(64 KB))); However, the resulting code seems to run a bit faster. How come enabling extra checks can make the code run faster? Curiouser and curiouser! This needs extra study. Currently I take the view that the dictSize should be set to non-zero when nothing else will do, i.e. when passing the external dictionary via dictStart. Otherwise, lowPrefix betrays just enough information about the dictionary. * * * Anyway, with this change, I instantiate all the necessary cases as functions with distinctive names, which also take fewer arguments and are therefore less error-prone. I also make the functions non-inline. (The compiler won't inline the functions because they are used more than once. Hence I attach LZ4_FORCE_O2_GCC_PPC64LE to the instances while removing from the callers.) The number of instances is now is reduced from 18 (safe+fast+partial+4*continue+4*prefix+4*dict+2*prefix64+forceExtDict) down to 7 (safe+fast+partial+2*prefix+2*dict). The size of the code is not the only issue here. Separate helper function are much more amenable to profile-guided optimization: it is enough to profile only a few basic functions, while the other less-often used functions, such as LZ4_decompress_*_continue, will benefit automatically. This is the list of LZ4_decompress* functions in liblz4.so, sorted by size. Exported functions are marked with a capital T. $ nm -S lib/liblz4.so |grep -wi T |grep LZ4_decompress |sort -k2 0000000000016260 0000000000000005 T LZ4_decompress_fast_withPrefix64k 0000000000016dc0 0000000000000025 T LZ4_decompress_fast_usingDict 0000000000016d80 0000000000000040 T LZ4_decompress_safe_usingDict 0000000000016d10 000000000000006b T LZ4_decompress_fast_continue 0000000000016c70 000000000000009f T LZ4_decompress_safe_continue 00000000000156c0 000000000000059c T LZ4_decompress_fast 0000000000014a90 00000000000005fa T LZ4_decompress_safe 0000000000015c60 00000000000005fa T LZ4_decompress_safe_withPrefix64k 0000000000002280 00000000000005fa t LZ4_decompress_safe_withSmallPrefix 0000000000015090 000000000000062f T LZ4_decompress_safe_partial 0000000000002880 00000000000008ea t LZ4_decompress_fast_extDict 0000000000016270 0000000000000993 t LZ4_decompress_safe_forceExtDict
* | | | Merge pull request #520 from felixhandte/frame-dict-nitsYann Collet2018-04-272-8/+15
|\ \ \ \ | |_|_|/ |/| | | Minor Fixes to Dictionary Preparation in LZ4 Frame
| * | | Avoid Possibly Redundant Table Clears When Loading HC DictW. Felix Handte2018-04-272-2/+2
| | | |
| * | | Remove Redundant LZ4_resetStream() CallW. Felix Handte2018-04-271-2/+1
| | | |
| * | | Rename LZ4F_applyCDict() -> LZ4F_initStream()W. Felix Handte2018-04-271-4/+12
| | |/ | |/|
* | | Merge pull request #519 from lz4/fdParserYann Collet2018-04-275-38/+68
|\ \ \ | |/ / |/| | Faster decoding speed
| * | ensure favorDecSpeed is properly initializedYann Collet2018-04-272-8/+8
| | | | | | | | | | | | | | | | | | | | | also : - fix a potential malloc error - proper use of ALLOC macro inside lz4hc - update html API doc
| * | fixed a number of minor cast warningsYann Collet2018-04-272-6/+5
| | |
| * | fasterDecSpeed can be triggered from cli with --favor-decSpeedYann Collet2018-04-262-2/+2
| | |
| * | favorDecSpeed feature can be triggered from lz4frameYann Collet2018-04-264-11/+27
| | | | | | | | | | | | and lz4hc.
| * | introduced ability to parse for decompression speedYann Collet2018-04-261-19/+34
| | | | | | | | | | | | | | | | | | | | | triggered through an enum. Now, it's still necessary to properly expose this capability all the way up to the cli.
* | | Merge pull request #518 from felixhandte/fix-517-dict-size-truncationYann Collet2018-04-261-4/+21
|\ \ \ | | | | | | | | Limit Dictionary Size During LZ4F Decompression
| * | | Limit Dictionary Size During LZ4F DecompressionW. Felix Handte2018-04-261-4/+21
| |/ / | | | | | | | | | Fixes lz4/lz4#517.
* | | Merge _destSize Compress Variant into LZ4_compress_generic()W. Felix Handte2018-04-261-190/+66
|/ /
* | Change Over Includes in the ProjectW. Felix Handte2018-04-241-1/+2
| |
* | Integrate lz4frame_static.h Declarations into lz4frame.hW. Felix Handte2018-04-242-113/+124
| |
* | Merge pull request #512 from lz4/HC_dictYann Collet2018-04-243-52/+266
|\ \ | | | | | | In-place unmutable dictionaries for LZ4HC
| * | Remove Debug Log StatementsW. Felix Handte2018-04-242-30/+5
| | |
| * | Revert Stream Size Const to Correct ValueW. Felix Handte2018-04-241-1/+1
| | |
| * | Change vLimit CalculationW. Felix Handte2018-04-211-1/+1
| | |
| * | Remove Redundant Static AssertW. Felix Handte2018-04-211-1/+0
| | |
| * | Simpler loadDict() ResetW. Felix Handte2018-04-201-1/+1
| | |
| * | Tolerate Base Pointer UnderflowW. Felix Handte2018-04-201-2/+2
| | |
| * | Don't Segfault on Malloc FailureW. Felix Handte2018-04-201-3/+5
| | |
| * | Sign-Extend -1 to Pointer WidthW. Felix Handte2018-04-201-3/+2
| | |
| * | Fix Constant ValueW. Felix Handte2018-04-201-1/+1
| | |
| * | Handle Index Underflows SafelyW. Felix Handte2018-04-201-8/+7
| | |
| * | Consts and Asserts and Other Minor NitsW. Felix Handte2018-04-201-6/+8
| | |
| * | Add Comments on New Public APIsW. Felix Handte2018-04-201-0/+29
| | |
| * | Add API for Attaching DictionariesW. Felix Handte2018-04-203-2/+31
| | |
| * | Also Reset the Chain TableW. Felix Handte2018-04-201-1/+1
| | |
| * | Remove inputBuffer from Context, Work Around its AbsenceW. Felix Handte2018-04-202-9/+15
| | |
| * | Remove Commented Out Support for Match Continuation over Segment BoundaryW. Felix Handte2018-04-201-5/+0
| | |
| * | Fix Signedness of ComparisonW. Felix Handte2018-04-201-1/+1
| | |
| * | Don't Clear the Dictionary Context Until No Longer UsefulW. Felix Handte2018-04-201-2/+5
| | |
| * | Copy DictCtx into Working Context on Inputs Larger than 4 KBW. Felix Handte2018-04-201-1/+12
| | |
| * | Force Inline on HashChainW. Felix Handte2018-04-201-1/+1
| | |
| * | Split DictCtx-using Code Into Separate Inlining ChainW. Felix Handte2018-04-201-20/+74
| | |
| * | Use Fast Reset in LZ4F AgainW. Felix Handte2018-04-201-1/+1
| | |
| * | Use Fast Reset API in LZ4FW. Felix Handte2018-04-201-1/+1
| | |
| * | Add Fast Reset PathsW. Felix Handte2018-04-202-2/+21
| | |
| * | Remove Match Upper Bounds CheckW. Felix Handte2018-04-201-7/+6
| | |