Improvement: added the ability to track the position (line, column, index) in the original input source from where
a given node was parsed. Accessible via Node.sourceRange() and Element.endSourceRange().
jhy/jsoup#1790
Improvement: added Element.firstElementChild(), Element.lastElementChild(), Node.firstChild(), Node.lastChild(),
as convenient accessors to those child nodes and elements.
Improvement: added Element.expectFirst(cssQuery), which is just like Element.selectFirst(), but instead of returning
a null if there is no match, will throw an IllegalArgumentException. This is useful if you want to simply abort
processing if an expected match is not found.
Improvement: when pretty-printing HTML, doctypes are emitted on a newline if there is a preceding comment.
jhy/jsoup#1664
Improvement: when pretty-printing, trim the leading and trailing spaces of textnodes in block tags when possible,
so that they are indented correctly.
jhy/jsoup#1798
Improvement: in Element#selectXpath(), disable namespace awareness. This makes it possible to always select elements
by their simple local name, regardless of whether an xmlns attribute was set.
jhy/jsoup#1801
Bugfix: when using the readToByteBuffer method, such as in Connection.Response.body(), if the document has not
already been parsed and must be read fully, and there is any maximum buffer size being applied, only the default
internal buffer size is read.
jhy/jsoup#1774
Bugfix: when serializing HTML, newlines in elements descending from a pre tag were incorrectly skipped. That caused
what should have been preformatted output to instead be a run of text.
jhy/jsoup#1776
Bugfix: when pretty-print serializing HTML, newlines separating phrasing content (e.g. a tag within a tag
would be incorrectly skipped, instead of normalized to a space. Additionally, improved space normalization between
other end of line occurences, and whitespace handling after a closing
jhy/jsoup#1787
*** Release 1.15.1 [2022-May-15]
Change: removed previously deprecated methods and classes (including org.jsoup.safety.Whitelist; use
org.jsoup.safety.Safelist instead).
Improvement: when converting jsoup Documents to W3C Documents in W3CDom, preserve HTML valid attribute names if the
input document is using the HTML syntax. (Previously, would always coerce using the more restrictive XML syntax.)
jhy/jsoup#1648
Improvement: added the :containsWholeText(text) selector, to match against non-normalized Element text. That can be
useful when elements can only be distinguished by e.g. specific case, or leading whitespace, etc.
jhy/jsoup#1636
Dependabot will resolve any conflicts with this PR as long as you don't alter it yourself. You can also trigger a rebase manually by commenting @dependabot rebase.
Dependabot commands and options
You can trigger Dependabot actions by commenting on this PR:
- `@dependabot rebase` will rebase this PR
- `@dependabot recreate` will recreate this PR, overwriting any edits that have been made to it
- `@dependabot merge` will merge this PR after your CI passes on it
- `@dependabot squash and merge` will squash and merge this PR after your CI passes on it
- `@dependabot cancel merge` will cancel a previously requested merge and block automerging
- `@dependabot reopen` will reopen this PR if it is closed
- `@dependabot close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually
- `@dependabot ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself)
- `@dependabot ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself)
- `@dependabot ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself)
Bumps jsoup from 1.14.3 to 1.15.2.
Release notes
Sourced from jsoup's releases.
Changelog
Sourced from jsoup's changelog.
... (truncated)
Commits
d9566b5
[maven-release-plugin] prepare release jsoup-1.15.21541765
Javadoc tweak7fb6d02
Keep the W3CBuilder static2b573de
Disable namespaces in Element#selectXpathb873e21
Use Charset.forname, to better cache charset lookups38b3224
Correct javadoc and add@WillClose
annotationsfc41ec9
Trim leading and trailing spaces in blocks when appropriate67b48dd
Pretty-print doctypes on a newline8733445
Fixed an OOB in TreeBuilder when getting the body Elemente714ef1
Improved newline and whitespace normalizationDependabot will resolve any conflicts with this PR as long as you don't alter it yourself. You can also trigger a rebase manually by commenting
@dependabot rebase
.Dependabot commands and options
You can trigger Dependabot actions by commenting on this PR: - `@dependabot rebase` will rebase this PR - `@dependabot recreate` will recreate this PR, overwriting any edits that have been made to it - `@dependabot merge` will merge this PR after your CI passes on it - `@dependabot squash and merge` will squash and merge this PR after your CI passes on it - `@dependabot cancel merge` will cancel a previously requested merge and block automerging - `@dependabot reopen` will reopen this PR if it is closed - `@dependabot close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - `@dependabot ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - `@dependabot ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - `@dependabot ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself)