jsoup 1.9.1 发布,HTML 解析器_html/css_WEB-ITnose
更新日志:
改进:
-
Added support for HTTP and SOCKS request proxies, specifiable per connection. See Connection.proxy(String, int).
-
Added support for sending plain HTTP request bodies in POST and PUT requests, with Connection.requestBody(String).
-
Added support in Jsoup.Connect() for HEAD, OPTIONS, and TRACE.
-
Added support for HTTP 307 Temporary Redirect (replays posts, if applicable).
-
Performance improvements when parsing HTML, particularly on Android Dalvik.
-
Added support for writing HTML into Appendable objects (like OutputStreamWriter), to enable stream serialization. See Node.html(T)
-
Added support for XML namespaces when converting jsoup documents to W3C documents.
-
Added support for UTF-16 and UTF-32 character set detection from byte-order-marks (BOM).
-
Added support for tags with non-ascii (unicode) letters.
-
Added Connection.data(String) to retrieve a data KeyVal by its key. Useful to update form data before submission.
Bug 修复
-
Fixed an issue in the Parent selector where it would not match against the root element it was applied to.
-
Fix an issue where Elements.select(String) would not return every matching element if they had the same content.
-
Added not-null validators to Element.appendText() and Element.prependText()
-
Fixed an issue when moving moving nodes using Element.insert(int, Collection) where the sibling index would be set incorrectly, leading to the original loads being lost.
-
Reverted Node.equals() and Node.hashCode() back to identity (object) comparisons, as deep content inspection had negative performance impacts and hashkey stability problems. Functionality replaced with Node.hasSameValue().
-
In Connection, if the same header key is seen multiple times, combine their values with a comma per the HTTP RFC, instead of keeping just one value. Also fixes an issue where header values could be out of order.
下载地址:
-
Source code (zip)
-
Source code (tar.gz)
jsoup 是一款 Java 的HTML 解析器,可直接解析某个URL地址、HTML文本内容。它提供了一套非常省力的API,可通过DOM,CSS以及类似于 JQuery 的操作方法来取出和操作数据。
jsoup的主要功能如下:
-
从一个URL,文件或字符串中解析HTML;
-
使用DOM或CSS选择器来查找、取出数据;
-
可操作HTML元素、属性、文本;
jsoup是基于MIT协议发布的,可放心使用于商业项目。
下一篇: 可视化框架设计-序
推荐阅读
-
文本编辑软件 Atom 1.5.0 已经发布_html/css_WEB-ITnose
-
Java web求一个新闻发布页面_html/css_WEB-ITnose
-
jsoup解析HTML_html/css_WEB-ITnose
-
Jsoup 爬取页面的数据和 理解HTTP消息头_html/css_WEB-ITnose
-
_html/css_WEB-ITnose">
【Jsoup】doc.getElementsByTag("img");无法获得验证码图片_html/css_WEB-ITnose
-
WebFont 智能压缩工具——字蛛 1.0.0 正式版发布_html/css_WEB-ITnose
-
jfinal cms 2.6.0 发布 添加栏目相关模块_html/css_WEB-ITnose
-
jfinal cms 2.6.0 发布 添加栏目相关模块_html/css_WEB-ITnose
-
MVC发布网站后,客户端访问默认是兼容模式.80分求解决._html/css_WEB-ITnose
-
有趣的CSS盒子模型--【牛腩新闻发布系统】_html/css_WEB-ITnose