Skip to content

Reuse preprocessed EPUB HTML when reopening files - #416

Merged
cary-rowen merged 1 commit into
blindpandas:developfrom
cary-rowen:perf/reuse-preprocessed-epub-html
Jul 31, 2026
Merged

Reuse preprocessed EPUB HTML when reopening files#416
cary-rowen merged 1 commit into
blindpandas:developfrom
cary-rowen:perf/reuse-preprocessed-epub-html

Conversation

@cary-rowen

Copy link
Copy Markdown
Collaborator

Link to issue number:

N/A

Summary of the issue:

Reopening an unchanged EPUB repeats the expensive HTML preprocessing step, which slows down file opening.

Description of how this pull request fixes the issue:

The existing EPUB cache now stores preprocessed HTML under a versioned key. Cache misses keep the current behavior, while cache hits skip the repeated preprocessing. Existing file-change detection and fallback handling remain unchanged.

A regression test verifies that repeated opens reuse the cached result without changing the document text.

Testing performed:

Manually opened the provided EPUB samples once to populate the cache and reopened them from the cache. The document text, storage text, and anchor positions matched between both opens. Cached opens were approximately 23–40% faster, saving about 0.3–1.6 seconds depending on the sample.

Known issues with pull request:

The improvement applies to reopening unchanged EPUB files; first-open time is unchanged. No known functional issues.

@cary-rowen
cary-rowen marked this pull request as ready for review July 31, 2026 10:59
@cary-rowen
cary-rowen merged commit 499251b into blindpandas:develop Jul 31, 2026
5 checks passed
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant