pdf.js

Author	SHA1	Message	Date
Jonas Jenwald	e09ad99973	Merge pull request #15916 from Snuffleupagus/fetch-transfer [api-minor] Enabling transferring of data fetched with the `PDFFetchStream` implementation	2023-01-13 13:28:12 +01:00
Jonas Jenwald	1362cd91d0	Improve input validation in `PDFDataTransportStream._onReceiveData` (PR 15908 follow-up) The mozilla-central [method `PdfDataListener.readData`](https://searchfox.org/mozilla-central/rev/893a8f062ec6144c84403fbfb0a57234418b89cf/toolkit/components/pdfjs/content/PdfStreamConverter.jsm#207-210) can return `null`, hence it seems like a very good idea to update `PDFDataTransportStream._onReceiveData` to handle that gracefully since the current code will throw in that case. Also, improves the JSDocs for the `PDFDataRangeTransport` class in the API.	2023-01-12 15:24:59 +01:00
Jonas Jenwald	cee97fcd15	[api-minor] Enabling transferring of data fetched with the `PDFFetchStream` implementation Note how in the API we're transferring the PDF data that's fetched over the network[1]: - `f28bf23a31/src/display/api.js (L2467-L2480)` - `f28bf23a31/src/display/api.js (L2553-L2564)` To support that functionality we have the `PDFDataTransportStream`, `PDFFetchStream`, `PDFNetworkStream`, and `PDFNodeStream` implementations. Here these stream-implementations vary slightly in how they handle `ArrayBuffer`s internally, w.r.t. transferring or copying the data: - In `PDFDataTransportStream` we optionally, after PR 15908, allow transferring of the PDF data as provided externally (used e.g. in the Firefox PDF Viewer). - In `PDFFetchStream` we're currenly always copying the PDF data returned by the Fetch API, which seems unnecessary. As discussed in PR 15908, it'd seem very weird if this sort of browser API didn't allow transferring of the returned data. - In `PDFNetworkStream` we're already, since many years, transferring the PDF data returned by the `XMLHttpRequest` functionality. Note how the `getArrayBuffer` helper function simply returns an `ArrayBuffer` response as-is. - In `PDFNodeStream` we're currently copying the PDF data, however this is unfortunately necessary since Node.js returns data as a `Buffer` object[2]. Given that the `PDFNetworkStream` has been, indirectly, supporting transferring of PDF data for years it would seem really strange if this didn't also apply to the `PDFFetchStream`-implementation. Hence this patch simply enables transferring of PDF data, when accessed using the Fetch API, unconditionally to help reduced main-thread memory usage since the `PDFFetchStream`-implementation is used by default in browsers (for the GENERIC build). --- [1] As opposed to PDF data being provided as e.g. a TypedArray when calling `getDocument` in the API. [2] This is a "special" Node.js object, see https://nodejs.org/api/buffer.html#buffer, which doesn't exist in browsers.	2023-01-12 13:59:21 +01:00
Jonas Jenwald	bbe629018d	[api-minor] Add a new `transferPdfData` option to allow transferring more data to the worker-thread (bug 1809164) Also, removes the `initialData`-parameter JSDocs for the `getDocument`-function given that this parameter has been completely unused since PR 8982 (over five years ago). Note that the `initialData`-parameter is, and always was, intended to be provided when initializing a `PDFDataRangeTransport`-instance.	2023-01-10 21:03:44 +01:00
Tim van der Meij	69113f08f2	Merge pull request #15887 from Snuffleupagus/rm-setPDFNetworkStreamFactory Inline the `setPDFNetworkStreamFactory` functionality in `src/display/api.js`	2023-01-07 13:16:23 +01:00
Tim van der Meij	b428824269	Merge pull request #15879 from Snuffleupagus/useWorkerFetch-defaults [api-minor] Improve the `useWorkerFetch` default value checks	2023-01-07 13:13:25 +01:00
Jonas Jenwald	1d5de9f4f4	Inline the `setPDFNetworkStreamFactory` functionality in `src/display/api.js` Given that this is internal functionality, not exposed in the official API, it's not entirely clear (at least to me) why we can't just initialize this directly in `src/display/api.js` instead. When testing both the development viewer and all the ways in which we run tests, everthing still appears to work just fine with this patch.	2023-01-06 13:23:07 +01:00
Jonas Jenwald	1a69d537c1	[api-minor] Limit the `PDFDocumentLoadingTask.onUnsupportedFeature` functionality to GENERIC builds (PR 15758 follow-up) This was deprecated in PR 15758 but it's unfortunately quite difficult to tell if third-party users are depending on this, e.g. to implement custom error reporting, and if so to what extent. However, thanks to the pre-processor we can limit most of this code to GENERIC builds which still seem like a worthwhile change. These changes reduce the bundle size of the Firefox PDF Viewer by 3.8 kB in total.	2023-01-01 17:53:12 +01:00
Jonas Jenwald	0c1fb4e740	[api-minor] Remove the `PDFDocumentProxy.stats` getter (PR 15758 follow-up) This was deprecated in PR 15758 and given that it's quite unlikely that any third-party users are relying on this functionality, since it was only ever added to support telemetry reporting in the Firefox PDF Viewer, it should hopefully be fine to remove this fairly quickly. These changes reduce the bundle size of the Firefox PDF Viewer by 4.5 kB in total.	2023-01-01 17:06:47 +01:00
Jonas Jenwald	2c57a4232c	[api-minor] Improve the `useWorkerFetch` default value checks Given that the Fetch API only supports the http/https protocols, worker-thread fetching of CMaps and Standard-fonts may thus fail in certain cases. To improve the default behaviour we'll now also check that the `cMapUrl` and `standardFontDataUrl` options are appropriate, except in Firefox where this should always work.	2023-01-01 14:48:28 +01:00
Jonas Jenwald	3110d1f29a	Merge pull request #15869 from Snuffleupagus/_abortOperatorList-clearTimeout Always abort a pending `streamReader` cancel timeout in `PDFPageProxy._abortOperatorList` (PR 15825 follow-up)	2022-12-27 13:26:43 +01:00
Jonas Jenwald	841abb53e6	Remove `PDFPageProxy.getJSActions` caching, since it's unused, in the API Note how, in the scripting initialization in the viewer, we only ever invoke `PDFPageProxy.getJSActions` once per page in order to improve overall performance; see `a575aa13b9/web/pdf_scripting_manager.js (L372-L375)` Hence it really shouldn't be necessary to cache its result in the API, especially when that is done manually rather than using something like `shadow`.	2022-12-27 10:39:33 +01:00
Jonas Jenwald	ae24dbd064	Always abort a pending `streamReader` cancel timeout in `PDFPageProxy._abortOperatorList` (PR 15825 follow-up) When we're destroying a `PDFPageProxy`-instance, during full document destruction, we'll force-abort any worker-thread parsing of operatorLists. Hence we should make sure that any pending cancel timeout is always aborted, since a later `PDFPageProxy._abortOperatorList` call should always "replace" a previous one. Please note: Technically this was always wrong, but with the changes in PR 15825 it became ever so slightly easier to trigger this thanks to the potentially longer timeout.	2022-12-27 10:19:39 +01:00
Jonas Jenwald	ded02941f2	[api-minor] Move, most of, the `isPureXfa`-handling from `PDFViewer` and into `PDFPageView` By moving this code the "pageviewer"-component example will become slightly more usable on its own, it may simplify a future addition of XFA Foreground document support, and finally also serves as preparation for the following patches.	2022-12-18 13:10:23 +01:00
Calixte Denizet	a84d14b382	[Editor] Avoid to scroll when an annotation is commited (fixes issue #15744 )	2022-12-17 13:48:19 +01:00
Jonas Jenwald	506bbb7283	Merge pull request #15825 from Snuffleupagus/cancel-extraDelay [api-minor] Allow specifying an extra-delay, in `RenderTask.cancel`, for worker-thread aborting of operatorList parsing	2022-12-14 19:26:39 +01:00
Jonas Jenwald	91524d1a60	[api-minor] Allow specifying an extra-delay, in `RenderTask.cancel`, for worker-thread aborting of operatorList parsing This is done to support upcoming viewer-changes, and in order to prevent third-party users from outright breaking things we'll simply ignore too large values.	2022-12-14 12:34:16 +01:00
Jonas Jenwald	dcf9ff2182	Handle possibly undefined parameters once per `AnnotationLayer.render` invocation There's no reason to repeat this for every single annotation. Also, adds a couple of missing JSDoc-parameters.	2022-12-14 12:23:24 +01:00
Calixte Denizet	2ebf8745a2	[JS] Run the named actions before running the format when the file is open (issue #15818 ) It's a follow-up of #14950: some format actions are ran when the document is open but we must be sure we've everything ready for that, hence we have to run some named actions before runnig the global format. In playing with the form, I discovered that the blur event wasn't triggered when JS called `setFocus` (because in such a case the mouse was never down). So I removed the mouseState thing to just use the correct commitKey when blur is triggered by a TAB key.	2022-12-13 21:12:32 +01:00
Calixte Denizet	0c1ec946aa	[JS] Handle correctly choice widgets where the display and the export values are different (issue #15815 )	2022-12-13 19:08:26 +01:00
Calixte Denizet	1a397681fe	The annotation layer dimensions must be set before adding some elements (follow-up of #15770 ) In order to move the annotations in the DOM to have something which corresponds to the visual order, we need to have their dimensions/positions which means that the parent must have some dimensions.	2022-12-13 14:54:45 +01:00
Jonas Jenwald	cafdc48147	[api-minor] Add a new `PageViewport`-getter to access the original, un-scaled, viewport dimensions While reviewing recent patches, I couldn't help but noticing that we now have a lot of call-sites that manually access the `PageViewport.viewBox`-property. Rather than repeating that verbatim all over the code-base, this patch adds a lazily computed and cached getter for this data instead.	2022-12-11 18:37:35 +01:00
Jonas Jenwald	9b6d0d994d	Remove the API-caching of annotation-data This was essentially done only to compensate for the viewer calling `PDFPageProxy.getAnnotations` unconditionally on every annotationLayer-rendering invocation. With the previous patch that's no longer happening, and this API-caching should thus no longer be necessary.	2022-12-11 18:12:10 +01:00
Calixte Denizet	a989b5a879	Set the dimensions of the various layers at their creation - Use a unique helper function in display/display_utils.js; - Move those dimensions in css' side.	2022-12-10 14:35:06 +01:00
Calixte Denizet	4f0bfabe7a	Take all the viewBox into account when computing the coordinates of an annotation in the page (fixes #15789 )	2022-12-08 15:02:20 +01:00
Calixte Denizet	b93bf9f654	[Editor] Don't use the editor parent which can be null. An annotation editor layer can be destroyed when it's invisible, hence some annotations can have a null parent but when printing/saving or when changing font size, color, ... of all added annotations (when selected with ctrl+a) we still need to have some parent properties especially the page dimensions, global scale factor and global rotation angle. This patch aims to remove all the references to the parent in the editor instances except in some cases where an editor should obviously have one. It fixes #15780.	2022-12-08 14:06:06 +01:00
Calixte Denizet	9af89381cd	[Editor] Add a very basic and incomplete workaround for issue #15780 The main issue is due to the fact that an editor's parent can be null when we want to serialize it and that lead to an exception which break all the saving/printing process. So this incomplete patch fixes only the saving/printing issue but not the underlying problem (i.e. having a null parent) and doesn't bring that much complexity, so it should help to uplift it the next Firefox release.	2022-12-06 16:22:24 +01:00
Jonas Jenwald	cdd39ec69e	Merge pull request #15778 from Snuffleupagus/keep-structTree Don't re-create the `structTreeLayer` on zooming and rotation	2022-12-06 10:02:20 +01:00
Jonas Jenwald	0274245e90	Remove the unused `TextLayerRenderTask._renderingDone` property (PR 15259 follow-up) This is yet another property that I forgot to remove in PR 15259.	2022-12-05 11:49:14 +01:00
Jonas Jenwald	fe8fded23b	[api-minor] Combine the `textContent`/`textContentStream` parameters Rather than handling these parameters separately, which is a left-over from back when streaming of textContent was originally added, we can simply pass either data directly to the `TextLayer` and let it handle things accordingly. Also, improves a few JSDoc comments and `typedef`-imports.	2022-12-04 21:22:14 +01:00
Jonas Jenwald	da0e6bc590	Don't re-create the `structTreeLayer` on zooming and rotation Compared to the recent PR 15722 for the `textLayer` this one should be a (comparatively) much a smaller win overall, since most documents don't have any structTree-data and the required parsing should be cheaper. However, it seems to me that it cannot hurt to improve this nonetheless. Note that by moving the `structTreeLayer` initialization we remove the need for the "textlayerrendered" event listener, which thus simplifies the code a little bit. Also, removes the API-caching of the structTree-data since this was basically done to offset the lack of caching in the viewer.	2022-12-04 10:18:58 +01:00
Tim van der Meij	67e1c37e0f	Merge pull request #15773 from Snuffleupagus/view-worker-normalize [api-minor] Normalize the `view`-getter on the worker-thread	2022-12-02 19:52:44 +01:00
Tim van der Meij	99cfef882f	Merge pull request #15752 from Snuffleupagus/no-typeof-undefined Enable the `no-typeof-undefined` ESLint plugin rule	2022-12-02 19:48:16 +01:00
Jonas Jenwald	5f8598abb7	[api-minor] Normalize the `view`-getter on the worker-thread Please note: I don't really expect that this is will be an observable change, since virtually all PDF documents already order e.g. /MediaBox and /CropBox entries correctly. By normalizing boundingBoxes already on the worker-thread, we can be sure that even a corrupt document won't cause issues. Note how we're passing the `view`-getter to the `PartialEvaluator.getTextContent` method, in order to detect textContent which is outside of the page, hence it makes sense to ensure that it's formatted as expected. Furthermore, by normalizing this once on the worker-tread we should no longer have to worry about a possibly negative width/height in the `PageViewport` constructor. Finally, the patch also simplifies the `view`-getter a little bit.	2022-12-02 15:46:39 +01:00
Calixte Denizet	eed9bf71c5	Refactor the text layer code in order to avoid to recompute it on each draw The idea is just to resuse what we got on the first draw. Now, we only update the scaleX of the different spans and the other values are dependant of --scale-factor. Move some properties in the CSS in order to avoid any updates in JS.	2022-12-01 18:42:43 +01:00
Jonas Jenwald	47dbfc4ade	Enable the `no-typeof-undefined` ESLint plugin rule Please see https://github.com/sindresorhus/eslint-plugin-unicorn/blob/main/docs/rules/no-typeof-undefined.md	2022-12-01 18:20:39 +01:00
Jonas Jenwald	fa54a58790	Merge pull request #15765 from Snuffleupagus/rm-textLayer-timeout [api-minor] Remove the TextLayer `timeout` parameter (PR 15742 follow-up)	2022-11-29 21:21:45 +01:00
calixteman	f3206b351f	Merge pull request #15764 from calixteman/15753 [Annotation] Send correctly the updated values to the JS sandbox	2022-11-29 20:04:12 +01:00
Jonas Jenwald	7c25b1b455	[api-minor] Remove the TextLayer `timeout` parameter (PR 15742 follow-up) The deprecation is included in the current release, i.e. version `3.1.81`, and given the edge-case nature of this option I really don't think that we need to keep it deprecated for multiple releases.	2022-11-29 19:57:38 +01:00
Calixte Denizet	20fd9099f8	[Annotation] Send correctly the updated values to the JS sandbox	2022-11-29 17:34:06 +01:00
Jonas Jenwald	82d127883d	Stop duplicating the `platform` getter in multiple files Currently both of the `AnnotationElement` and `KeyboardManager` classes contain identical `platform` getters, which seems like unnecessary duplication. With the pre-processor we can also limit the feature-testing to only GENERIC builds, since `navigator` should always be available in browsers.	2022-11-29 12:14:40 +01:00
Calixte Denizet	b9cb651c44	[api-minor] Remove all the useless telemetry stuff in the viewer (bug 1802468) Add a deprecation notification for PDFDocumentLoadingTask.onUnsupportedFeature and PDFDocumentProxy.stats which are likely useless. The unsupported feature stuff have initially been added in (#4048) in order to be able to display a warning bar and to help to have some numbers to know how a feature was used. Those data are no more used in Firefox.	2022-11-28 20:55:15 +01:00
Jonas Jenwald	85f03c0ea4	Slightly modernize the `FontLoader.isSyncFontLoadingSupported` getter This is very old code, which is unused (by default) in browsers nowadays since the Font Loading API will always be preferred. For Node.js environments we use the same constant as elsewhere throughout the code-base, and we can also simplify the Firefox-specific check given that the lowest supported version is `102` (as of this writing). Finally the old TODO is removed, since the general availability of the Font Loading API has made it redundant.	2022-11-27 12:19:11 +01:00
Jonas Jenwald	aa5b678f94	Add default icons for FileAttachment annotations (bug 1230933) Please note: This "borrows" the icons from Thunderbird. According to the PDF specification, see https://web.archive.org/web/20220309040754if_/https://www.adobe.com/content/dam/acom/en/devnet/pdf/pdfs/PDF32000_2008.pdf#G11.2096626, we should be providing default icons for FileAttachment annotations without appearances.	2022-11-26 11:24:59 +01:00
Jonas Jenwald	b3e161c328	[api-minor] Deprecate the TextLayer `timeout` parameter This has never really been used anywhere within the PDF.js library[1], and when streaming of textContent was introduced this parameter was effectively made redundant. Note that when streaming of textContent is used, all text-layout has already happened by the time that this `timeout`-functionality is actually invoked (thus making it pointless). While the `timeout`-functionality may still "work" when the textContent is provided upfront, although it's never been used/tested, streaming will generally perform better (in e.g. a viewer setting). Please note: While unrelated here, also removes a now unused property that I forgot in PR 15259. --- [1] At least not since the code was moved into its current file, which happened in PR 6619 and landed seven years ago.	2022-11-24 23:08:39 +01:00
Jonas Jenwald	47682985d3	Add support for Optional Content in TilingPatterns (issue 15716) This can't be a particularly common feature, since we've supported Optional Content for over two years and this is the very first TilingPattern-case we've seen.	2022-11-23 12:58:00 +01:00
Jonas Jenwald	f3e0f86641	Simplify the `getFilenameFromUrl` helper function	2022-11-23 11:48:08 +01:00
Jonas Jenwald	0ba242ea4a	Support FileAttachments with hash-signs in the filename (issue 15729) The reason for the issue is that we use the generic `getFilenameFromUrl` helper function, which was originally intended for regular URLs. For the filenames we're dealing with in FileAttachments, we really only want to strip the path when one exists[1]. --- [1] See [bug 1230933](https://bugzilla.mozilla.org/show_bug.cgi?id=1230933) for an example of such a case.	2022-11-23 10:47:33 +01:00
Jonas Jenwald	3e4caf2e13	Take the mask-offset into account when rendering repeated image masks (bug 1799927) Please note: As usual when I'm working with the `src/display/canvas.js` code I don't really know what I'm doing, but it at least appears to work.	2022-11-13 16:15:30 +01:00
Jonas Jenwald	bab1097db3	Remove the constructor in the `StatTimer` class With modern EcmaScript features, we can define these fields directly instead. Please note that for backwards compatibility purposes they are still public as before, however note that this functionality is disabled by default (see the `pdfBug` API option). Also, we can (slightly) simplify the two loops used in the `toString` method.	2022-11-11 12:31:04 +01:00

1 2 3 4 5 ...

1728 Commits