jsoup 1.14.3 is out now, adding native XPath selector support, improved \<template> support, and also includes a bunch of bug fixes, improvements, and performance enhancements.
Improvement: added native XPath support in Element#selectXpath(String)
jhy/jsoup#1629
Improvement: added full support for the <template> tag to the HTML5 parser spec.
jhy/jsoup#1634
Improvement: added support in CharacterReader to track newlines, so that parse errors can be reported more
intuitively.
jhy/jsoup#1624
Improvement: tracked parse errors now have more details, including the erroneous token, to help clarify the errors.
Improvement: speed and memory optimizations for the :has(subquery) selector.
Improvement: the :contains(text) and :containsOwn(text) selectors are now whitespace normalized, aligning to the
document text that they are matching against.
jhy/jsoup#876
Improvement: in Element, speed optimized adopting all of an element's child nodes into a currently empty element.
Improves the HTML adoption agency algorithm when adopting elements with many children.
jhy/jsoup#1638
Improvement: increased the parse speed when in RCData (e.g. ) and unescaped tokens are found, by
memoizing the scan and reducing GC.
jhy/jsoup#1644
Improvement: when parsing custom tags (in HTML or XML), added a flyweight cache on Tag.valueOf(name) to reduce
memory overhead when many tags are repeated. Also tuned other areas of the parser when many very deeply stacked
custom elements were present.
jhy/jsoup#1646
Bugfix: when tracking errors or checking for validity in the Cleaner, errors were incorrectly raised for missing
optional closing tags.
Bugfix: the OSGi bundle meta-data incorrectly set a version on the import of javax.annotation (used as a build-time
dependency for nullability assertions).
jhy/jsoup#1616
Bugfix: the Attributes::equals() method was sensitive to the order of its contents, but it should not be.
jhy/jsoup#1492
Bugfix: when the HTML parser was configured to preserve case, Element text methods would miss adding whitespace for
"BR" tags.
Bugfix: attribute names are now normalized & validated correctly for the specific output syntax (HTML or XML).
Previously, syntactically invalid attribute names could be output by the html() methods. Such attributes are still
available in the DOM, and will be normalized if possible on output.
Dependabot will resolve any conflicts with this PR as long as you don't alter it yourself. You can also trigger a rebase manually by commenting @dependabot rebase.
Dependabot commands and options
You can trigger Dependabot actions by commenting on this PR:
- `@dependabot rebase` will rebase this PR
- `@dependabot recreate` will recreate this PR, overwriting any edits that have been made to it
- `@dependabot merge` will merge this PR after your CI passes on it
- `@dependabot squash and merge` will squash and merge this PR after your CI passes on it
- `@dependabot cancel merge` will cancel a previously requested merge and block automerging
- `@dependabot reopen` will reopen this PR if it is closed
- `@dependabot close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually
- `@dependabot ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself)
- `@dependabot ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself)
- `@dependabot ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself)
Bumps jsoup from 1.14.2 to 1.14.3.
Release notes
Sourced from jsoup's releases.
Changelog
Sourced from jsoup's changelog.
... (truncated)
Commits
0006162
[maven-release-plugin] prepare release jsoup-1.14.380a9396
Javadoc update for XPath0d1f04a
Javadoc update to add@since
1.14.389de796
Bump junit-jupiter from 5.8.0 to 5.8.1 (#1645)b14eb2a
Test case and change note for parser improvements incl tag flyweight4b46397
Short-circuit tag scans for custom tagsd3f4e31
Flyweight Tag.valueOf in TreeBuildera8df71b
Limit the stack depth we scan looking for mis-closed DD / DT tagse4ae6fa
Per spec, only foster incoming nodes if current node is a table foster target41932fe
JDK 17 changelogDependabot will resolve any conflicts with this PR as long as you don't alter it yourself. You can also trigger a rebase manually by commenting
@dependabot rebase
.Dependabot commands and options
You can trigger Dependabot actions by commenting on this PR: - `@dependabot rebase` will rebase this PR - `@dependabot recreate` will recreate this PR, overwriting any edits that have been made to it - `@dependabot merge` will merge this PR after your CI passes on it - `@dependabot squash and merge` will squash and merge this PR after your CI passes on it - `@dependabot cancel merge` will cancel a previously requested merge and block automerging - `@dependabot reopen` will reopen this PR if it is closed - `@dependabot close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - `@dependabot ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - `@dependabot ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - `@dependabot ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself)