python修改xml文件_如何使用python将.txt文件转换为xml文件?

Latitude :23.1100348

Longitude:72.5364922

date&time :30:August:2014 05:04:31 PM

gsm cell id: 4993

Neighboring List- Lac : Cid : RSSI

15000 : 7072 : 25 dBm

15000 : 7073 : 23 dBm

15000 : 6102 : 24 dBm

15000 : 6101 : 24 dBm

15000 : 6103 : 17 dBm

Latitude :23.1120549

Longitude:72.5397988

date&time :30:August:2014 05:04:34 PM

gsm cell id: 4993

Neighboring List- Lac : Cid : RSSI

15000 : 7072 : 24 dBm

15000 : 7073 : 22 dBm

15000 : 6102 : 23 dBm

15000 : 6101 : 23 dBm

15000 : 2552 : 16 dBm

这是my.txt文件,我想将其转换为xml文件,例如

我试图列出所有组件,但没有得到o / p.我想将所有纬度,经度,gsm单元格ID,时间值存储在列表中,这将在xml文件中添加类似内容.

我写下面的代码.

import re

pa = 'Longitude|Latitude|gsm cell id|Neighboring List- Lac : Cid : RSSI'

with open('cell.txt','rw') as file:

for line in file:

line.strip()

if re.search(pa, line):

lineInfo = line.split(':')

title = lineInfo[0]

value = lineInfo[1]

解决方法:

尝试以下代码作为入门:

#!python3

import re

import xml.etree.ElementTree as ET

rex = re.compile(r'''(?P

Longitude

|Latitude

|date&time

|gsm\s+cell\s+id

)

\s*:?\s*

(?P.*)

''', re.VERBOSE)

root = ET.Element('root')

root.text = '\n' # newline before the celldata element

with open('cell.txt') as f:

celldata = ET.SubElement(root, 'celldata')

celldata.text = '\n' # newline before the collected element

celldata.tail = '\n\n' # empty line after the celldata element

for line in f:

# Empty line starts new celldata element (hack style, uggly)

if line.isspace():

celldata = ET.SubElement(root, 'celldata')

celldata.text = '\n'

celldata.tail = '\n\n'

# If the line contains the wanted data, process it.

m = rex.search(line)

if m:

# Fix some problems with the title as it will be used

# as the tag name.

title = m.group('title')

title = title.replace('&', '')

title = title.replace(' ', '')

e = ET.SubElement(celldata, title.lower())

e.text = m.group('value')

e.tail = '\n'

# Display for debugging

ET.dump(root)

# Include the root element to the tree and write the tree

# to the file.

tree = ET.ElementTree(root)

tree.write('cell.xml', encoding='utf-8', xml_declaration=True)

它显示您的示例数据:

23.1100348

72.5364922

30:August:2014 05:04:31 PM

4993

23.1120549

72.5397988

30:August:2014 05:04:34 PM

4993

所需近邻列表的更新:

#!python3

import re

import xml.etree.ElementTree as ET

rex = re.compile(r'''(?P

Longitude

|Latitude

|date&time

|gsm\s+cell\s+id

|Neighboring\s+List-\s+Lac\s+:\s+Cid\s+:\s+RSSI

)

\s*:?\s*

(?P.*)

''', re.VERBOSE)

root = ET.Element('root')

root.text = '\n' # newline before the celldata element

with open('cell.txt') as f:

celldata = ET.SubElement(root, 'celldata')

celldata.text = '\n' # newline before the collected element

celldata.tail = '\n\n' # empty line after the celldata element

for line in f:

# Empty line starts new celldata element (hack style, uggly)

if line.isspace():

celldata = ET.SubElement(root, 'celldata')

celldata.text = '\n'

celldata.tail = '\n\n'

else:

# If the line contains the wanted data, process it.

m = rex.search(line)

if m:

# Fix some problems with the title as it will be used

# as the tag name.

title = m.group('title')

title = title.replace('&', '')

title = title.replace(' ', '')

if line.startswith('Neighboring'):

neighbours = ET.SubElement(celldata, 'neighbours')

neighbours.text = '\n'

neighbours.tail = '\n'

else:

e = ET.SubElement(celldata, title.lower())

e.text = m.group('value')

e.tail = '\n'

else:

# This is the neighbour item. Split it by colon,

# and set the attributes of the item element.

item = ET.SubElement(neighbours, 'item')

item.tail = '\n'

lac, cid, rssi = (a.strip() for a in line.split(':'))

item.attrib['lac'] = lac

item.attrib['cid'] = cid

item.attrib['rssi'] = rssi.split()[0] # dBm removed

# Include the root element to the tree and write the tree

# to the file.

tree = ET.ElementTree(root)

tree.write('cell.xml', encoding='utf-8', xml_declaration=True)

更新以在邻居之前接受空行-更好的通用实现:

#!python3

import re

import xml.etree.ElementTree as ET

rex = re.compile(r'''(?P

Longitude

|Latitude

|date&time

|gsm\s+cell\s+id

|Neighboring\s+List-\s+Lac\s+:\s+Cid\s+:\s+RSSI

)

\s*:?\s*

(?P.*)

''', re.VERBOSE)

root = ET.Element('root')

root.text = '\n' # newline before the celldata element

with open('cell.txt') as f:

celldata = ET.SubElement(root, 'celldata')

celldata.text = '\n' # newline before the collected element

celldata.tail = '\n\n' # empty line after the celldata element

status = 0 # init status of the finite automaton

for line in f:

if status == 0: # lines of the heading expected

# If the line contains the wanted data, process it.

m = rex.search(line)

if m:

# Fix some problems with the title as it will be used

# as the tag name.

title = m.group('title')

title = title.replace('&', '')

title = title.replace(' ', '')

if line.startswith('Neighboring'):

neighbours = ET.SubElement(celldata, 'neighbours')

neighbours.text = '\n'

neighbours.tail = '\n'

status = 1 # empty line and then list of neighbours expected

else:

e = ET.SubElement(celldata, title.lower())

e.text = m.group('value')

e.tail = '\n'

# keep the same status

elif status == 1: # empty line expected

if line.isspace():

status = 2 # list of neighbours must follow

else:

raise RuntimeError('Empty line expected. (status == {})'.format(status))

status = 999 # error status

elif status == 2: # neighbour or the empty line as final separator

if line.isspace():

celldata = ET.SubElement(root, 'celldata')

celldata.text = '\n'

celldata.tail = '\n\n'

status = 0 # go to the initial status

else:

# This is the neighbour item. Split it by colon,

# and set the attributes of the item element.

item = ET.SubElement(neighbours, 'item')

item.tail = '\n'

lac, cid, rssi = (a.strip() for a in line.split(':'))

item.attrib['lac'] = lac

item.attrib['cid'] = cid

item.attrib['rssi'] = rssi.split()[0] # dBm removed

# keep the same status

elif status == 999: # error status -- break the loop

break

else:

raise LogicError('Unexpected status {}.'.format(status))

break

# Display for debugging

ET.dump(root)

# Include the root element to the tree and write the tree

# to the file.

tree = ET.ElementTree(root)

tree.write('cell.xml', encoding='utf-8', xml_declaration=True)

该代码实现了所谓的有限自动机,其中状态变量表示其当前状态.您可以使用铅笔和纸来可视化它-用内部状态数字绘制小圆圈(在图论中称为节点).处于状态时,您仅允许某种输入(行).识别输入后,您将箭头(图论中的定向边)绘制到另一种状态(可能是同一状态,作为返回到同一节点的循环).箭头标有“条件|行动’.

一开始的结果可能看起来很复杂;但是,从某种意义上讲,您始终可以将精力集中在属于特定状态的代码部分上,这很容易.而且,可以轻松修改代码.但是,有限自动机的功能有限.但是它们只是解决此类问题的理想之选.

标签:python-3-x,python-2-7,xml,python

来源: https://codeday.me/bug/20191121/2050118.html

  • 0
    点赞
  • 1
    收藏
    觉得还不错? 一键收藏
  • 0
    评论
评论
添加红包

请填写红包祝福语或标题

红包个数最小为10个

红包金额最低5元

当前余额3.43前往充值 >
需支付:10.00
成就一亿技术人!
领取后你会自动成为博主和红包主的粉丝 规则
hope_wisdom
发出的红包
实付
使用余额支付
点击重新获取
扫码支付
钱包余额 0

抵扣说明:

1.余额是钱包充值的虚拟货币,按照1:1的比例进行支付金额的抵扣。
2.余额无法直接购买下载,可以购买VIP、付费专栏及课程。

余额充值