mupdf - MuPDF PDF reader and library

Age	Commit message (Collapse)	Author
2017-04-27	Typedef function pointers consistently.	Tor Andersson

2017-04-27	Move fz_outline and pdf_xref debug printing to pdfshow.c	Tor Andersson
	That's where it's actually being used.
2017-04-27	Use FZ_SEEK_SET macros for fz_seek.	Tor Andersson
	Don't depend on stdio.h for our own I/O functions.
2017-04-27	Include required system headers.	Tor Andersson

2017-04-13	Move extension/mimetype detection to common function.	Sebastian Rasmussen
	A document handler normally only exposes a list of extensions and mimetypes. Only formats that use some kind of extra detection mechnism need to supply a recognize() callback, such as xps that can handle .xps-files unpacked into a directory.
2017-03-28	Return fz_document from all document handlers.	Sebastian Rasmussen
	To make it possible to avoid casting in most cases.
2017-03-23	Introduce fz_new_derived_...	Robin Watts
	Instead of having fz_new_XXXX(ctx, type, ...) macros that call fz_new_XXXX_of_size etc, use fz_new_derived_... Clearer naming, and doesn't clash with fz_new_document_writer.
2017-03-01	Add page lookup cache for faster link destination lookups in outlines.	Tor Andersson
	Loading outlines wants to look up all link destinations, and doing the normal link destination lookups triggers loading all page objects used. This means we need to parse a lot of objects, which can be quite slow. We can load the page tree faster by only looking at intermediate page tree nodes. If we load the page tree and create a reverse lookup table for use when loading the outline, we can speed up the time to run 'mutool show pdfref17.pdf outline' from 900ms to 100ms.
2017-01-17	Fix typos.	Sebastian Rasmussen

2016-12-27	Strip extraneous blank lines.	Tor Andersson

2016-12-23	Don't add bogus entries when pdf_update_object is called with NULL.	Tor Andersson
	Treat such calls as deleting the object, as per pdf_delete_object.
2016-12-19	Make pdf_trailer() return NULL if there is no xref.	Sebastian Rasmussen

2016-12-16	Also repair object streams when repairing on the fly.	Tor Andersson

2016-12-16	Bug 697412: When repairing, forget the previous xref.	Tor Andersson

2016-12-12	PDF Portfolio support.	Robin Watts
	New PDF Portfolio manipulation API. Simple mutool 'portfolio' tool for listing/extracting/embedding files.
2016-11-16	pdf: Add 'compressed/raw' flag to pdf_add_stream.	Tor Andersson
	Also expose the argument to JS and JNI.
2016-11-14	Make fz_buffer structure private to fitz.	Robin Watts
	Move the definition of the structure contents into new fitz-imp.h file. Make all code outside of fitz access the buffer through the defined API. Add a convenience API for people that want to get buffers as null terminated C strings.
2016-11-14	Add optional 'object' argument to pdf_add_stream.	Tor Andersson

2016-11-11	Add pdf_layer configuration API.	Robin Watts
	Add API to: * allow enumeration of layer configs (OCCDs) within PDF files. * allow selection of layer configs. * allow enumeration of the "UI" (or "Human readable") form of layer configs. * allow selection/toggling of entries in the UI.
2016-11-03	Fix signed/unsigned and size_t/int/fz_off_t warnings.	Robin Watts
	All seen in MSVC, mostly in 64bit builds.
2016-11-01	When hinted object is not found, avoid return from within fz_try().	Sebastian Rasmussen

2016-10-28	Clean up link destination handling.	Tor Andersson
	All link destinations should be URIs, and a document specific function can be called to resolve them to actual page numbers. Outlines have cached page numbers as well as string URIs.
2016-10-18	Internal drop functions don't need to check for NULL.	Sebastian Rasmussen

2016-10-18	Avoid checking argument to fz_drop_*()/fz_free().	Sebastian Rasmussen
	As fz_drop_*()/fz_free() all must handle NULL.
2016-10-12	Always call fz_drop_document() to drop the document.	Sebastian Rasmussen

2016-10-11	Free document in fz_drop_document(), not in subclassing documents.	Sebastian Rasmussen

2016-10-07	pdf: Remove unneccessary document argument to pdf_to_utf8 etc.	Tor Andersson

2016-10-06	Update Xref reading code to cope with 19 byte entries.	Robin Watts
	The spec says entries should be 20 bytes long. In practise we see 19 byte long ones more often than we like. This is due to the use of a single EOL char rather than 2. The PCLm files I've seen use 19 byte ones, so update the code to cope with these.
2016-09-23	Fix leak in error case of pdf_add_stream.	Robin Watts

2016-09-22	Bug 697015: Avoid object references vanishing during repair.	Robin Watts
	A PDF repair can be triggered 'just in time', when we encounter a problem in the file. The idea is that this can happen without the enclosing code being aware of it. Thus the enclosing code may be holding 'borrowed' references (such as those returned by pdf_dict_get()) at the time when the repair is triggered. We are therefore at pains to ensure that the repair does not replace any objects that exist already, so that the calling code will not have these references unexpectedly invalidated. The sole exception to this is when we replace the 'Length' fields in stream dictionaries with the actual lengths. Bug 697015 shows exactly this situation causing a reference to become invalid. The solution implemented here is to add an 'orphan list' to the document, where we put these (hopefully few, small) objects. These orphans are kept around until the document is closed.
2016-09-19	fz_store: Reap passes.	Robin Watts
	A few commits back, we introduced the fz_key_storable concept to allow us to cope with objects that were used both as values within the store and as parts of keys within the store. This commit worked, but showed up performance problems; when the store has several million PDF objects in it, bulk changes (such as dropping a display list or document) could trigger many passes across the store. We therefore introduce a mechanism to ameliorate this. These passes, now known as "reap passes", can be batched together using fz_defer_reap_start and fz_defer_reap_end. We trigger this start/end around display list dropping, and around PDF content stream processing. This should be fine, as deferral will be interrupted if we ever run our of memory during mallocing.
2016-09-16	Remove unused variable.	Robin Watts

2016-09-16	Silence some warnings.	Robin Watts

2016-09-16	Tweak store handling of PDF document destroy.	Robin Watts
	When we destroy a PDF document, currently we bin everything from the store. Instead, drop just the objects that are specifically tied to that document. Any object tied to the document has a pdf_obj with the required document pointer in it as the key.
2016-09-01	pdf: Load/open streams by indirect reference object when possible.	Tor Andersson

2016-08-30	Don't try to copy a NULL dictionary.	Tor Andersson

2016-07-22	Bug 696941: Fix use after free.	Robin Watts
	The file is HORRIBLY corrupt, and triggers Sophos to think it's PDF malware (which it isn't). It does however trigger a use after free, worked around here.
2016-07-13	Bug 696892: PDF annotation appearance stream synthesis SEGV	Robin Watts
	The code would SEGV if we were trying to synthesise an appearance stream for an annotation, and the docs pdf resources table had not been initialised. We now intialise the pdf resource tables when we initialise a pdf device. This is the earliest point we know we are going to need them, and covers all cases.
2016-07-13	Use fz_malloc_struct rather than fz_calloc.	Robin Watts
	This helps with Memento debugging, and looks neater.
2016-07-08	Separate close and drop functionality for devices and writers.	Tor Andersson
	Closing a device or writer may throw exceptions, but much of the foreign language bindings (JNI and JS) depend on drop to never throw an exception (exceptions in finalizers are bad).
2016-07-06	Start slimming pdf_page.	Tor Andersson
	We want to turn pdf_page into a thin wrapper around a pdf_obj, so that any updates to the underlying PDF objects will be reflected without having to reload the pdf_page.
2016-07-06	Add fitz to pdf downcasting functions for pages and annotations.	Tor Andersson

2016-07-06	Fix garbage collection and page grafting for indirect reference chains.	Tor Andersson
	The mark & sweep pass of garbage collection, and resolving indirect objects when grafting objects was following the full chain of indirect references. In the unusual case where a numbered object is itself only an indirect reference to another object, this intermediate numbered object would be missed both when marking for garbage collection, and when copying objects for grafting. Add a function to resolve only one step for these two uses. The following is an example of a file that would break during garbage collection if we follow full indirect reference chains: %PDF-1.3 1 0 obj <</Type/Catalog /Foo[2 0 R 3 0 R]>> endobj 2 0 obj 4 0 R endobj 3 0 obj 5 0 R endobj 4 0 obj <</Length 1>> stream A endstream endobj 5 0 obj <</Length 1>> stream B endstream endobj
2016-07-06	pdf: Drop generation number from public interfaces.	Tor Andersson
	The generation number is only needed for decryption, and is assumed to be zero or irrelevant for all other uses. Store the original object number and generation in the xref slot, so that we can decrypt them even when the objects have been renumbered, without needing to pass the original object number around through the stream loading APIs.
2016-07-06	pdf: Increment generation number in the xref when deleting an object.	Tor Andersson

2016-07-06	pdf: Check ownership when adding objects to a document.	Tor Andersson

2016-06-17	Use 'size_t' instead of int as appropriate.	Robin Watts
	This silences the many warnings we get when building for x64 in windows. This does not address any of the warnings we get in thirdparty libraries - in particular harfbuzz. These look (at a quick glance) harmless though.
2016-06-14	Fix typos in various parts of the code.	Sebastian Rasmussen

2016-04-27	Remove useless try/catch/rethrows.	Tor Andersson

2016-04-27	Fix 696649: remove fz_rethrow_message calls.	Tor Andersson