3.39.1
- #48943 - From Quarkus 3.22, Database dev services not reused if the unit test is using a
@QuarkusTestResource - #51130 - Remote Dev fails when using hibernate-reactive extension
- #55389 - Fix
@RunOnVirtualThreadannotation in combination with blocking endpoints - #55907 - Quarkus 3.38 Date serialization issue
- #55958 - container-image-jib: no progress output while the image is pushed, because ProgressEvent is never subscribed to
- #55959 - Jib container push progress log
- #56009 - Dev UI Dev MCP endpoint hangs indefinitely (no response, no error) when a JSON-RPC request's id is a string
- #56064 -
quarkus-rest-jackson: reflection-free (de)serializers treat aMapsubclass as a bean and drop every map entry - #56070 - quarkus-spring-security may not fail the build when
@PostAuthorize,@PreFilteror@PostFilterare used - #56079 - Skip Hibernate dev integrator in remote server-side dev mode
- #56080 - MongoDB Panache: Session-commit on transaction rollback
- #56086 - [3.39] Do not set an invalid country code when generating the Quarkus Dev CA
- #56087 - [3.39] Handle string JSON-RPC ids and always respond on the Dev MCP endpoint (3.39 backport)
- #56088 - Fail the build when unsupported Spring Security annotations are used
- #56092 - [3.39] Revert "Disable the Quarkus Dev CA generation test on Windows"
- #56101 - Document that CORS cannot be configured both programmatically and with properties
- #56104 - quarkus-config-doc-maven-plugin picks up things from git worktrees that it should ignore
- #56105 - Don't descend into nested git checkouts when scanning for config-doc target directories
- #56130 - Register terminal provider SPI and defer FFM init for native images
- #56131 - docs: clean up independently maintained CORS guide
- #56132 - Redis replication with topology=static does not respect host ordering - Set loses master/replica order
- #56158 - REST Client:
@RestClienton the interface fails the build with a duplicate annotation since 3.32 - #56159 - Ignore a
@RestClientqualifier placed on the REST Client interface - #56166 - Fix MongoDB Panache committing sessions after a JTA transaction timeout
- #56177 - Bump io.micrometer:micrometer-bom from 1.17.0 to 1.17.1
- #56183 - Redis: make sure the configured hosts are ordered
- #56192 - Do not rebuild the application for every launch of a QuarkusMainTest
- #56197 - Upgrade aesh, aesh-readline, fix ffm downcall and customizers
- #56204 - Switch to JavaParser for generating build items doc
jsoup 1.23.2
jsoup 1.23.2 is out now. This release focuses largely on bug fixes, specification correctness, and performance improvements.
It brings closer alignment with the HTML, XML, URL, and form submission specifications; improves XML and W3C DOM conversion; and makes HTTP workloads more efficient through streamed request bodies and broader JDK HttpClient reuse.
This version also includes new node insertion methods for Elements, and DOM mutations now reject operations that would create a cycle.
jsoup is a Java library for working with real-world HTML and XML. It provides a very convenient API for extracting and manipulating data, using the best of HTML5 DOM methods and CSS selectors.
Download jsoup now.
- Improved consecutive
StreamParser.selectFirst()calls during progressive parsing, so later matches are returned with their parsed contents when earlier selections had left them as parser lookahead. E.g., given<title>One</title><p id=hit>Full</p><p>Next</p>, selectingtitleand then#hitnow advances the partial lookahead and returns<p id="hit">Full</p>, rather than returning an empty<p id="hit"></p>before its content is parsed. The updated readiness tracking follows StreamParser's normal emission order across implicit HTML structure and parser recovery. #2551 - Improved XML parser performance and memory use for documents with many nested namespace declarations by recording namespace changes within each element scope. #2556
- Improved
W3CDomconversion performance for documents with many nested namespace declarations. The W3C converter now uses the same optimized namespace tracking as the XML parser. #2559 - Improved
W3CDomXML conversion to retain processing instructions, comments outside the root element, and CDATA sections, which were previously dropped or converted to text. #2572 - DOM mutation methods, including child insertion and replacement, now reject operations that would create a cycle, such as making a node its own child or moving an ancestor beneath a descendant. #2552
- Added
Elements#before(Node),after(Node),prepend(Node), andappend(Node)to match the existing HTML string methods. #953 - Large file-backed uploads through
Connection.requestBodyStream(InputStream)now stream directly with the JDKHttpClienton Java 11+, rather than being loaded fully into memory first. #2575 - Extended Java 11+ HTTP client reuse from requests sharing a
Jsoup.newSession()to ordinaryJsoup.connect()calls, reducing transport thread and connection setup churn under sustained request loads. Sessions with custom authentication or SSL contexts continue to use their own client. #2584
- Aligned the XML parser stack depth and lookups to the configured maximum, which now defaults to 512 for both HTML and XML. Use
Parser#setMaxDepth(int)to configure. #2570
- Fixed
W3CDomnamespace conversion in several cases: #2559- Namespace declarations and prefixed attributes now carry the correct namespace URI, so namespace-aware DOM lookups work as expected.
- Attributes added after parsing, or included through subtree conversion, now use inherited prefix declarations.
- Namespace declarations now apply regardless of attribute order, and an empty declaration shadows an inherited binding only within its scope.
- With namespace awareness disabled, inherited and undeclared prefixes now receive the declarations needed for XML serialization.
- Valid HTML names that are not XML QNames, such as
a:b:c, are normalized. Attributes that still cannot be represented are skipped, and unrepresentable elements no longer change the surrounding tree.
- Fixed
W3CDomconversion of programmatically created or renamed elements whose names can be represented in a jsoup HTML DOM but are not valid XML names, such as1abc. These names are now normalized (e.g._1abc) instead of causing aNullPointerException. #2560 - Fixed XML doctype serialization when a system identifier contains a double quote, which could otherwise produce invalid XML. #2571
- XML serialization now repairs element and attribute names that start with an invalid character, rather than outputting
nullelements or dropping attributes. For example, an attribute named1ais written as_1a. Additional leading underscores keep repaired attribute names unique if they conflict with another attribute. #2573 - Supplementary Unicode characters are now escaped correctly when serializing with non-UTF, non-ASCII output charsets such as ISO-8859-1. Previously, characters could be emitted unescaped when their low 16-bit value was representable by the configured charset, causing replacement or corruption when the output was encoded. #2578
- Fixed the JDK
HttpClientimplementation to accept responses missing aContent-Typeheader, matching theHttpURLConnectionimplementation. #2549 - Fixed HTTP response content-type matching to handle media types case-insensitively and recognize structured
+xmlsuffixes, including vendor-specific media types. #2550 - HTTP request URL normalization now percent-encodes ASCII control characters, DEL, and embedded fragment delimiters, keeping normalized URLs valid for HTTP requests while preserving existing escapes. #2585
- Corrected multipart form encoding to percent-escape CR and LF in field names and filenames, matching the HTML form submission specification. Multipart file content-types containing CR or LF are now rejected with a
ValidationException. #2555 - Aligned trailing comment placement with the HTML specification: comments after
</body>remain children of thehtmlelement, while comments after</html>remain children of the document. #2557 - When using the optional
re2jregular expression engine, memory allocation errors caused by complex selector patterns at match time are now normalized to aValidationExceptionwith aPattern complexity errormessage. - Fixed parsing of malformed SVG and MathML content so that breakout HTML tags are placed according to the HTML specification. #2562
- Fixed deeply nested malformed HTML parsing that could lose the document body because stack lookups did not align to the configured maximum parser depth. #2569
- Aligned RCDATA, RAWTEXT, and script-data parsing with the HTML specification: malformed end tags no longer consume following markup, unclosed
title/textareacontent stays text through EOF, and custom text tags match exact names. #2577 - Improved URL validation during HTTP/HTTPS URL resolution and cleaning; resolved URLs without a host are now rejected instead of being accepted based only on their scheme prefix, aligning to RFC 9110. Valid relative links and non-HTTP(S) schemes are unchanged. #2579
- Redirects with malformed single-slash HTTP locations now use standard URL resolution to align with browsers. #2580
- Template fragment parsing now handles unmatched
</template>tags without throwing aValidationException. #2581 - Improved source tracking for adopted formatting elements and malformed markup ending at EOF. #2582
My sincere thanks to everyone who contributed to this release! If you have any suggestions for the next release, I would love to hear them; please get in touch via jsoup discussions, or with me directly.
You can also follow me (@jhy@tilde.zone) on Mastodon (Fediverse) to receive occasional notes about jsoup releases.
Nightly
- 5b3666d: [py] update new BiDi layer generation to conform to latest proposed ADR (#17942) (Titus Fortner) #17942
- 083869c: [dotnet] [bidi] Throw in case of unknown discriminator (#17948) (Nikolay Borisenko) #17948
- 7bfaedb: [rb] reject an inbound BiDi scalar outside its union's declared arms (#17947) (Titus Fortner) #17947
- 498c0e7: [build] alert Slack when a CDP update lands on trunk (#17950) (Titus Fortner) #17950