jsoup 1.9.1 发布,HTML 解析器_html/css_WEB-ITnose

php中文网
发布: 2016-06-24 11:20:33
原创
1206人浏览过

jsoup 1.9.1 发布。

更新日志:

改进:

  • Added support for HTTP and SOCKS request proxies, specifiable per connection. See Connection.proxy(String, int).

  • Added support for sending plain HTTP request bodies in POST and PUT requests, with Connection.requestBody(String).

  • Added support in Jsoup.Connect() for HEAD, OPTIONS, and TRACE.

  • Added support for HTTP 307 Temporary Redirect (replays posts, if applicable).

  • Performance improvements when parsing HTML, particularly on Android Dalvik.

  • Added support for writing HTML into Appendable objects (like OutputStreamWriter), to enable stream serialization. See Node.html(T)

  • Added support for XML namespaces when converting jsoup documents to W3C documents.

  • Added support for UTF-16 and UTF-32 character set detection from byte-order-marks (BOM).

  • Added support for tags with non-ascii (unicode) letters.

  • Added Connection.data(String) to retrieve a data KeyVal by its key. Useful to update form data before submission.

Bug 修复

  • Fixed an issue in the Parent selector where it would not match against the root element it was applied to.

  • Fix an issue where Elements.select(String) would not return every matching element if they had the same content.

  • Added not-null validators to Element.appendText() and Element.prependText()

  • Fixed an issue when moving moving nodes using Element.insert(int, Collection) where the sibling index would be set incorrectly, leading to the original loads being lost.

  • Reverted Node.equals() and Node.hashCode() back to identity (object) comparisons, as deep content inspection had negative performance impacts and hashkey stability problems. Functionality replaced with Node.hasSameValue().

  • In Connection, if the same header key is seen multiple times, combine their values with a comma per the HTTP RFC, instead of keeping just one value. Also fixes an issue where header values could be out of order.

下载地址:

  • Source code (zip)

  • Source code (tar.gz)

jsoup 是一款 Java 的HTML 解析器,可直接解析某个URL地址、HTML文本内容。它提供了一套非常省力的API,可通过DOM,CSS以及类似于 JQuery 的操作方法来取出和操作数据。

自由画布
自由画布

百度文库和百度网盘联合开发的AI创作工具类智能体

自由画布 73
查看详情 自由画布

jsoup的主要功能如下:

立即学习前端免费学习笔记(深入)”;

  1. 从一个URL,文件或字符串中解析HTML;

  2. 使用DOM或CSS选择器来查找、取出数据;

  3. 可操作HTML元素、属性、文本;

jsoup是基于MIT协议发布的,可放心使用于商业项目。

HTML速学教程(入门课程)
HTML速学教程(入门课程)

HTML怎么学习?HTML怎么入门?HTML在哪学?HTML怎么学才快?不用担心,这里为大家提供了HTML速学教程(入门课程),有需要的小伙伴保存下载就能学习啦!

下载
来源:php中文网
本文内容由网友自发贡献,版权归原作者所有,本站不承担相应法律责任。如您发现有涉嫌抄袭侵权的内容,请联系admin@php.cn
最新问题
开源免费商场系统广告
热门教程
更多>
最新下载
更多>
网站特效
网站源码
网站素材
前端模板
关于我们 免责申明 举报中心 意见反馈 讲师合作 广告合作 最新更新 English
php中文网:公益在线php培训,帮助PHP学习者快速成长!
关注服务号 技术交流群
PHP中文网订阅号
每天精选资源文章推送
PHP中文网APP
随时随地碎片化学习

Copyright 2014-2025 https://www.php.cn/ All Rights Reserved | php.cn | 湘ICP备2023035733号