d6/xz-analysis-mirror

Commit Graph

Author	SHA1	Message	Date
Lasse Collin	46007049cd	liblzma: Fix compilation of fastpos_tablegen.c. The macro lzma_attr_visibility_hidden has to be defined to make fastpos.h usable. The visibility attribute is irrelevant to fastpos_tablegen.c so simply #define the macro to an empty value. fastpos_tablegen.c is never built by the included build systems and so the problem wasn't noticed earlier. It's just a standalone program for generating fastpos_table.c. Fixes: https://github.com/tukaani-project/xz/pull/69 Thanks to GitHub user Jamaika1.	2023-10-31 21:41:09 +02:00
Lasse Collin	41113fe30a	liblzma: Use lzma_attr_visibility_hidden on private extern declarations. These variables are internal to liblzma and not exposed in the API.	2023-10-30 18:06:25 +02:00
Dimitri Papadopoulos Orfanos	42df7c7aa1	Docs: Fix typos found by codespell	2023-07-31 20:02:21 +08:00
Jia Tan	8f23657498	liblzma: Exports lzma_mt_block_size() as an API function. The lzma_mt_block_size() was previously just an internal function for the multithreaded .xz encoder. It is used to provide a recommended Block size for a given filter chain. This function is helpful to determine the maximum Block size for the multithreaded .xz encoder when one wants to change the filters between blocks. Then, this determined Block size can be provided to lzma_stream_encoder_mt() in the lzma_mt options parameter when intializing the coder. This requires one to know all the filter chains they are using before starting to encode (or at least the filter chain that will need the largest Block size), but that isn't a bad limitation.	2023-05-11 23:54:44 +08:00
Jia Tan	116e81f002	Build: Removes redundant check for LZMA1 filter support.	2023-03-23 21:48:52 +08:00
Lasse Collin	33b8a24b66	liblzma: Add LZMA_FILTER_LZMA1EXT to support LZMA1 without end marker. Some file formats need support for LZMA1 streams that don't use the end of payload marker (EOPM) alias end of stream (EOS) marker. So far liblzma API has supported decompressing such streams via lzma_alone_decoder() when .lzma header specifies a known uncompressed size. Encoding support hasn't been available in the API. Instead of adding a new LZMA1-only API for this purpose, this commit adds a new filter ID for use with raw encoder and decoder. The main benefit of this approach is that then also filter chains are possible, for example, if someone wants to implement support for .7z files that use the x86 BCJ filter with LZMA1 (not BCJ2 as that isn't supported in liblzma).	2022-11-27 23:16:21 +02:00
Lasse Collin	9a304bf1e4	liblzma: Avoid unneeded use of void pointer in LZMA decoder.	2022-11-27 18:43:07 +02:00
Lasse Collin	218394958c	liblzma: Pass the Filter ID to LZ encoder and decoder. This allows using two Filter IDs with the same initialization function and data structures.	2022-11-27 18:20:33 +02:00
Lasse Collin	3be88ae071	liblzma: Allow nice_len 2 and 3 even if match finder requires 3 or 4. That is, if the specified nice_len is smaller than the minimum of the match finder, silently use the match finder's minimum value instead of reporting an error. The old behavior is annoying to users and it complicates xz options handling too.	2022-11-24 23:23:55 +02:00
Lasse Collin	c392bf8ccb	liblzma: Fix infinite loop in LZMA encoder init with dict_size >= 2 GiB. The encoder doesn't support dictionary sizes larger than 1536 MiB. This is validated, for example, when calculating the memory usage via lzma_raw_encoder_memusage(). It is also enforced by the LZ part of the encoder initialization. However, LZMA encoder with LZMA_MODE_NORMAL did an unsafe calculation with dict_size before such validation and that results in an infinite loop if dict_size was 2 << 30 or greater.	2022-11-22 11:23:23 +02:00
Lasse Collin	107c93ee5c	liblzma: Rename a variable and improve a comment.	2022-07-14 18:12:38 +03:00
Lasse Collin	9595a3119b	liblzma: Add optional autodetection of LZMA end marker. Turns out that this is needed for .lzma files as the spec in LZMA SDK says that end marker may be present even if the size is stored in the header. Such files are rare but exist in the real world. The code in liblzma is so old that the spec didn't exist in LZMA SDK back then and I had understood that such files weren't possible (the lzma tool in LZMA SDK didn't create such files). This modifies the internal API so that LZMA decoder can be told if EOPM is allowed even when the uncompressed size is known. It's allowed with .lzma and not with other uses. Thanks to Karl Beldan for reporting the problem.	2022-07-13 22:24:07 +03:00
jiat75	6468f7e41a	liblzma: Add NULL checks to LZMA and LZMA2 properties encoders. Previously lzma_lzma_props_encode() and lzma_lzma2_props_encode() assumed that the options pointers must be non-NULL because the with these filters the API says it must never be NULL. It is good to do these checks anyway.	2022-02-07 00:20:01 +02:00
Lasse Collin	6c6f0db340	liblzma: Fix unitialized variable. This was introduced two weeks ago in the commit `625f4c7c99`. Thanks to Nathan Moinvaziri.	2021-01-29 21:19:08 +02:00
Lasse Collin	625f4c7c99	liblzma: Add rough support for output-size-limited encoding in LZMA1. With this it is possible to encode LZMA1 data without EOPM so that the encoder will encode as much input as it can without exceeding the specified output size limit. The resulting LZMA1 stream will be a normal LZMA1 stream without EOPM. The actual uncompressed size will be available to the caller via the uncomp_size pointer. One missing thing is that the LZMA layer doesn't inform the LZ layer when the encoding is finished and thus the LZ may read more input when it won't be used. However, this doesn't matter if encoding is done with a single call (which is the planned use case for now). For proper multi-call encoding this should be improved. This commit only adds the functionality for internal use. Nothing uses it yet.	2021-01-14 18:58:13 +02:00
Lasse Collin	b3ed19a55f	liblzma: Remove unneeded <sys/types.h> from fastpos_tablegen.c. This file only generates fastpos_table.c. It isn't built as a part of liblzma.	2020-02-24 23:23:18 +02:00
Lasse Collin	43dfe04e62	liblzma: Add more uses of lzma_memcmplen() to the normal mode of LZMA. This gives a tiny encoder speed improvement. This could have been done in 2014 after the commit `544aaa3d13` but it was forgotten.	2020-02-21 17:40:02 +02:00
Lasse Collin	7136f1735c	Rename unaligned_read32ne to read32ne, and similarly for the others.	2019-12-31 00:47:49 +02:00
Lasse Collin	dfac2c9a1d	liblzma: Fix warnings from -Wsign-conversion. Also, more parentheses were added to the literal_subcoder macro in lzma_comon.h (better style but no functional change in the current usage).	2019-06-23 21:38:56 +03:00
Lasse Collin	33773c6f2a	liblzma: Use unaligned_readXXne functions instead of type punning. Now gcc -fsanitize=undefined should be clean. Thanks to Jeffrey Walton.	2019-06-01 19:01:21 +03:00
Lasse Collin	94e3f986aa	Fix or hide warnings from GCC 7's -Wimplicit-fallthrough.	2017-08-14 20:08:33 +03:00
Lasse Collin	d4a0462abe	liblzma: Avoid multiple definitions of lzma_coder structures. Only one definition was visible in a translation unit. It avoided a few casts and temp variables but seems that this hack doesn't work with link-time optimizations in compilers as it's not C99/C11 compliant. Fixes: http://www.mail-archive.com/xz-devel@tukaani.org/msg00279.html	2016-11-21 20:24:50 +02:00
Lasse Collin	f4c95ba94b	liblzma: Rename lzma_presets.c back to lzma_encoder_presets.c. It would be too annoying to update other build systems just because of this.	2015-11-03 20:55:45 +02:00
Lasse Collin	4cc584985c	Build: Build LZMA1/2 presets also when only decoder is wanted. People shouldn't rely on the presets when decoding raw streams, but xz uses the presets as the starting point for raw decoder options anyway. lzma_encocder_presets.c was renamed to lzma_presets.c to make it clear it's not used solely by the encoder code.	2015-11-03 18:06:40 +02:00
Lasse Collin	f243f5f44c	liblzma: Silence more uint32_t vs. size_t warnings.	2015-03-07 22:01:00 +02:00
Lasse Collin	117d962685	liblzma: Fix a compression-ratio regression in LZMA1/2 in fast mode. The bug was added in the commit `f48fce093b` and thus affected 5.1.4beta and 5.2.0. Luckily the bug cannot cause data corruption or other nasty things.	2015-02-21 23:40:26 +02:00
Lasse Collin	544aaa3d13	liblzma: Use lzma_memcmplen() in normal mode of LZMA. Two locations were not changed yet because the simplest change assumes that the initial "len" may be greater than "limit".	2014-07-25 22:38:28 +03:00
Lasse Collin	f48fce093b	liblzma: Simplify LZMA fast mode code by using memcmp().	2014-07-25 22:30:38 +03:00
Lasse Collin	6bf5308e34	liblzma: Use lzma_memcmplen() in fast mode of LZMA.	2014-07-25 22:29:49 +03:00
Lasse Collin	a19d9e8575	liblzma: Avoid C99 compound literal arrays. MSVC 2013 doesn't like them. Maybe they aren't so good for readability either since many aren't used to them.	2014-01-12 16:44:52 +02:00
Lasse Collin	3778db1be5	liblzma: Make the use of lzma_allocator const-correct. There is a tiny risk of causing breakage: If an application assigns lzma_stream.allocator to a non-const pointer, such code won't compile anymore. I don't know why anyone would do such a thing though, so in practice this shouldn't cause trouble. Thanks to Jan Kratochvil for the patch.	2012-07-17 18:19:59 +03:00
Lasse Collin	1403707fc6	liblzma: Check that the first byte of range encoded data is 0x00. It is just to be more pedantic and thus perhaps catch broken files slightly earlier.	2012-06-28 10:47:49 +03:00
Lasse Collin	3e321a3acd	Remove doubled words from documentation and comments. Spot candidates by running these commands: git ls-files \|xargs perl -0777 -n \ -e 'while (/\b(then?\|[iao]n\|i[fst]\|but\|f?or\|at\|and\|[dt]o)\s+\1\b/gims)' \ -e '{$n=($` =~ tr/\n/\n/ + 1); ($v=$&)=~s/\n/\\n/g; print "$ARGV:$n:$v\n"}' Thanks to Jim Meyering for the original patch.	2011-04-12 11:59:49 +03:00
Lasse Collin	25fe729532	liblzma: Add the forgotten lzma_lzma2_block_size(). This should have been in `5eefc0086d`.	2011-04-11 21:15:07 +03:00
Lasse Collin	0d21f49a80	liblzma: Fix decoding of LZMA2 streams having no uncompressed data. The decoder considered empty LZMA2 streams to be corrupt. This shouldn't matter much with .xz files, because no encoder creates empty LZMA2 streams in .xz. This bug is more likely to cause problems in applications that use raw LZMA2 streams.	2011-03-31 11:54:48 +03:00
Lasse Collin	974ebe6349	liblzma: Rename a few variables and constants. This has no semantic changes. I find the new names slightly more logical and they match the names that are already used in XZ Embedded. The name fastpos wasn't changed (not worth the hassle).	2010-10-26 10:36:41 +03:00
Lasse Collin	0076e03641	Clean up a few FIXMEs and TODOs. lzma_chunk_size() was commented out because it is currently useless.	2010-10-19 11:44:37 +03:00
Lasse Collin	075257ab04	Fix the preset -3e. depth=0 was missing.	2010-09-26 18:10:31 +03:00
Lasse Collin	8fd3ac046d	Don't set lc=4 with --extreme. This should reduce the cases where --extreme makes compression worse. On the other hand, some other files may now benefit slightly less from --extreme.	2010-09-04 22:16:28 +03:00
Lasse Collin	b4b1cbcb53	Tweak the compression presets -0 .. -5. "Extreme" mode might need some further tweaking still. Docs were not updated yet.	2010-09-03 15:13:12 +03:00
Lasse Collin	920a69a8d8	Rename MIN() and MAX() to my_min() and my_max(). This should avoid some minor portability issues.	2010-05-26 10:36:46 +03:00
Lasse Collin	eb7d51a3fa	Collection of language fixes to comments and docs. Thanks to Jonathan Nieder.	2010-02-12 13:16:15 +02:00
Lasse Collin	0733f4c999	Make fastpos.h use tuklib_integer.h instead of bsr.h when --enable-small has been specified.	2009-11-22 11:55:03 +02:00
Lasse Collin	e330fb7e6b	Fix wrong indentation caused by incorrect settings in the text editor.	2009-11-15 12:54:45 +02:00
Lasse Collin	418d64a32e	Fix a design error in liblzma API. Originally the idea was that using LZMA_FULL_FLUSH with Stream encoder would read the filter chain from the same array that was used to intialize the Stream encoder. Since most apps wouldn't use LZMA_FULL_FLUSH, most apps wouldn't need to keep the filter chain available after initializing the Stream encoder. However, due to my mistake, it actually required keeping the array always available. Since setting the new filter chain via the array used at initialization time is not a nice way to do it for a couple of reasons, this commit ditches it and introduces lzma_filters_update(). This new function replaces also the "persistent" flag used by LZMA2 (and to-be-designed Subblock filter), which was also an ugly thing to do. Thanks to Alexey Tourbin for reminding me about the problem that Stream encoder used to require keeping the filter chain allocated.	2009-11-14 18:59:19 +02:00
Lasse Collin	ebfb2c5e1f	Use a tuklib module for integer handling. This replaces bswap.h and integer.h. The tuklib module uses <byteswap.h> on GNU, <sys/endian.h> on *BSDs and <sys/byteorder.h> on Solaris, which may contain optimized code like inline assembly.	2009-10-04 22:57:12 +03:00
Lasse Collin	18a4233a53	Fix a couple of warnings.	2009-09-11 09:25:09 +03:00
Lasse Collin	f42ee98166	Build system fixes Don't use libtool convenience libraries to avoid recently discovered long-standing subtle but somewhat severe bugs in libtool (at least 1.5.22 and 2.2.6 are affected). It was found when porting XZ Utils to Windows <http://lists.gnu.org/archive/html/libtool/2009-06/msg00070.html> but the problem is significant also e.g. on GNU/Linux. Unless --disable-shared is passed to configure, static library built from a set of convenience libraries will contain PIC objects. That is, while libtool builds non-PIC objects too, only PIC objects will be used from the convenience libraries. On 32-bit x86 (tested on mobile XP2400+), using PIC instead of non-PIC makes the decompressor 10 % slower with the default CFLAGS. So while xz was linked against static liblzma by default, it got the slower PIC objects unless --disable-shared was used. I tend develop and benchmark with --disable-shared due to faster build time, so I hadn't noticed the problem in benchmarks earlier. This commit also adds support for building Windows resources into liblzma and executables.	2009-06-30 17:09:57 +03:00
Lasse Collin	1c9360b7d1	Fix @variables@ to $(variables) in Makefile.am files. Fix the ordering of libgnu.a and LTLIBINTL on the linker command line and added missing LTLIBINTL to tests/Makefile.am.	2009-06-26 14:47:31 +03:00
Lasse Collin	02ddf09bc3	Put the interesting parts of XZ Utils into the public domain. Some minor documentation cleanups were made at the same time.	2009-04-13 11:27:40 +03:00

1 2

96 Commits