提取html xml,分析XML以使用htmlparser2提取特定标记的文本

最新推荐文章于 2024-03-21 09:54:14 发布

weixin_39936134

最新推荐文章于 2024-03-21 09:54:14 发布

阅读量200

点赞数

文章标签：提取html xml

我正在尝试node-htmlparser2,一开始就被卡住了。我有数千个这样的XML文件:

â¦

我想要里面的一切

作为单个字符串。我下面的代码有效,但在我看来这不是正确的方法

let isFoo = false;

let txt = '';

const p = new htmlparser.Parser({

onopentag: function(name, attribs){

if (name === 'foo') {

isFoo = true;

}

ontext: function(text){

if (isFoo) {

txt += text;

}

onclosetag: function(tagname){

if (tagname === 'foo') {

isFoo = false;

return txt;

}

}, {decodeEntities: true, xmlMode: true});

let data = [];

for (let file in files) {

let record = {

filename: file,

filetext: p.write(file)

}

data.push(record);

p.end();

}

有没有更好的方法可以在没有这种愚蠢的情况下使用htmlparser2?

isFoo

旗帜?

确定要放弃本次机会？

福利倒计时

: :

立减 ¥

普通VIP年卡可用

关注关注