【发布时间】:2021-04-13 04:11:49
【问题描述】:
假设我有一个这样的 html 片段:
<p>Generally speaking, in the U.S., if you want to <a href="\"https://web.archive.org/web/20210408195204/https://www.wsj.com/articles/investors-big-and-small-are-driving-stock-gains-with-borrowed-money-11617799940\"">
borrow money from your broker to buy stocks</a>, you are capped at 2-to-1 leverage. If you have $100, you can buy $200 worth of stock. Back in the olden days, you could have bought $300 or $500 or $1,000 of stock with your $100, borrowing the rest from your broker, but then a Great Depression happened and regulators clamped down on margin lending. </p>
我已经用jsoup 对其进行了解析,并将其表示为Element。
我希望能够将元素拆分为:
- 一段文字:
"Generally speaking, in the U.S., if you want to " - 一个元素
<a href="\&quot;https://web.archive.org/web/20210408195204/https://www.wsj.com/articles/investors-big-and-small-are-driving-stock-gains-with-borrowed-money-11617799940\&quot;"> - 另一段文字:
"borrow money from your broker to buy stocks</a>, you are capped at 2-to-1 leverage. If you have $100, you can buy $200 worth of stock. Back in the olden days, you could have bought&nbsp;$300 or $500 or $1,000 of stock with your $100, borrowing the rest from your broker, but then a Great Depression happened and regulators clamped down on margin lending"
并在保持这些部分的顺序的同时做到这一点。
到目前为止我看过的东西:
- getAllElements() 仅返回索引 0 处的 p 标签本身,然后返回 a 标签的元素
- children() 只为 a 标签返回 1 个元素。
【问题讨论】: