xml - 如何在excel中使一个xml标签成为一行？-6ren

xml - 如何在excel中使一个xml标签成为一行？

转载作者：行者123 更新时间：2023-12-04 21:06:10

我是 xml 文件的新手，所以我需要一些帮助。我需要将 xml 文件导入到 excel 中，但我无法正确格式化。 “P”标签中的内容我需要在一行中，属性值(n 值)是列标题。 “D”标签中的内容在我正在使用的文件中多次出现，“D”的 id 上升到 31。所有这些标签都在“B”标签中。其中有敏感信息，所以我不得不用东西替换文本，对不起。我希望这一切都是有道理的，任何信息或指向正确的方向都会受到赞赏。谢谢你。

<D id="1">
    <V n="stuff1">stuff1</V>
    <V n="stuff2">stuff2</V>
    <V n="stuff3">stuff3</V>

    <P id="stuff11">
        <V n="stuff111">stuff111</V>
        <V n="stuff112">stuff112</V>
        <V n="stuff113">stuff113</V>
        <V n="stuff114">stuff114</V>
        <V n="stuff115">stuff115</V>
        <V n="stuff116">stuff116</V>
    </P>
</D>

<D id="2">
    <V n="stuff1">stuff1</V>
    <V n="stuff2">stuff2</V>
    <V n="stuff3">stuff3</V>

    <P id="stuff21">
        <V n="stuff111">stuff211</V>
        <V n="stuff112">stuff212</V>
        <V n="stuff113">stuff213</V>
        <V n="stuff114">stuff214</V>
        <V n="stuff115">stuff215</V>
        <V n="stuff116">stuff216</V>
    </P>
</D>

最佳答案

不幸的是，我之前发现 Excel 会导入元素值，但不会导入这些元素中的属性。例如“stuff211”导入 stuff211 但不导入 stuff111。我认为这只是功能在 Excel 中工作方式的一个限制。导入必须在 Excel 中，还是可以使用 Python 等编程语言？我之前已经编写了一个程序来从 xml 文件中提取特定元素和属性值，如果需要，我很乐意在明天挖掘并分享？

更新

这是我之前编写并用于将数据从 xml 文件剥离到 csv 文件中的 python 脚本。请注意，我还没有设置它来获取所有属性和元素，因为我最初的目的是从文件中获取特定数据。您需要将 search_items 全局列表编辑为要搜索的项目。

您可以使用 xml 文件路径的单个参数从命令行调用脚本，也可以不使用 arg，系统将提示您选择目录。请让我知道，如果你有任何问题:

#!/usr/bin/python

# Change Ideas
# ------------
# Add option to get all xml elements / attributes
# Add support for json?

import sys, os, Tkinter, tkFileDialog as fd, traceback

# stop tinker shell from opening as only needed for file dialog
root = Tkinter.Tk()
root.withdraw()

# globals
debug_on = False
#get_all_elements = False
#get_all_attributes = False

# search items to be defined each time you run to identify values to search for.
# each item should have a search text, a type and optionally a heading e.g.
# search_items = ['exact_serach_text', 'item_type', 'column_heading(optional)']
# note: search items are case sensitive.
#
##############################   E X A M P L E   ##############################
search_items = [
    ['policyno=',                       'xml_attribute',    'Policy No'    ],
    ['transid=',                        'xml_attribute',    'Trans ID'     ],
    ['policyPremium=',                  'xml_attribute',    'Pol Prem'     ],
    ['outstandingBalance=',             'xml_attribute',    'Balance'      ],
    ['APRCharge=',                      'xml_attribute',    'APR Chrg'     ],
    ['PayByAnnualDD=',                  'xml_attribute',    'Annual DD'    ],
    ['PayByDD=',                        'xml_attribute',    'Pay by DD'    ],
    ['mtaDebitAmount=',                 'xml_attribute',    'MTA Amt'      ],
    ['paymentMethod=',                  'xml_attribute',    'Pmt Meth'     ],
    ['ddFirstPaymentAmount=',           'xml_attribute',    '1st Amt'      ],
    ['ddRemainingPaymentsAmount=',      'xml_attribute',    'Other Amt'    ],
    ['ddNumberOfPaymentsRemaining=',    'xml_attribute',    'Instl Rem'    ],
    ]
item_types = ['xml_attribute', 'xml_element']

def get_heads():
    heads = []
    for i in search_items:
        try:
            # raise error if i[2] does not exist or is empty
            assert len(i[2]) > 0, "No value in heading, use search text."
        except:
            heads.append(i[0]) # use search item as not heading is given
        else:
            heads.append(i[2])
    return heads

def write_csv_file(path, heads, data):
    """
    Writes data to file, use None for heads param if no headers required.
    """
    with open(path, 'wb') as fileout:
        writer = csv.writer(fileout)
        if heads:
            writer.writerow(heads)
        for row in data:
            try:
                writer.writerow(row)
            except:
                print '...row failed in write to file:', row
                exc_type, exc_value, exc_traceback = sys.exc_info()
                lines = traceback.format_exception(exc_type, exc_value, exc_traceback)
                for line in lines:
                    print '!!', line
    print 'Data written to:', path, '\n'

def find_val_in_line(item, item_type, line):
    if item_type.lower() == 'xml_element':
        print 'Testing still in progress for xml elements, please check output carefully'
        b1, b2 = ">", "<"
        tmp = line.find(item) # find the starting point of the element value
        x = line.find(b1, tmp+1) + len(boundary) # find next boundary after item
        y = line.find(b2, x) # find subsequent boundary to mark end of element value
        return line[x:y]
    elif item_type.lower() == 'xml_attribute':
        b = '"'
        tmp = line.find(item) # find the starting point of the attribute
        x = line.find(b, tmp+1) + len(b) # find next boundary after item
        y = line.find(b, x) # find subsequent boundary to mark end of attribute
        return line[x:y] # return value between start and end boundaries
    else:
        print 'This program does not currently support type:', item_type
        print 'Returning null'
        return None

def find_vals_in_file(file_path):
    with open(file_path, "r") as f:
        buf = f.readlines()
        f.seek(0)
        data, row = [], []
        found_in_row, pos = 0, 0
        l = len(search_items)
        if debug_on: print '\nsearch_items set to:\n' + str(search_items) + '\n'

        # loop through the lines in the file...
        for line in buf:
            if debug_on: print '\n..line set to:\n  ' + line

            # loop through search items on each line...
            for i in search_items:
                if debug_on: print '...item set to:\n  ' + i[0]

                # if the search item is found in the line...
                if i[0] in line:

                    val = find_val_in_line(i[0], i[1], line)

                    # only count as another item found if not already in that row
                    try:
                        # do not increment cnt if this works as item already exists
                        row[pos] = val
                        if debug_on: print '.....repeat item found:- ' + i[0] + ':' + val
                    except IndexError:
                        found_in_row += 1 # Index does not exist, count as new
                        row.append(val)
                        if debug_on: print '....item found, row set to:\n    ' + str(row)

                    # if we have a full row then add row to data and start row again...
                    if found_in_row == l:
                        if debug_on: print '......row complete, appending to data\n'
                        data.append(row)
                        row, found_in_row = [], 0
                pos += 1 # start at 0 and increment 1 at end of each search item
            pos = 0
        f.close()
    return data

def main():
    path, matches = None, []
    os.chdir(os.getenv('userprofile'))

    # check cmd line args provided...
    if len(sys.argv) > 1:
        path = sys.argv[1]
    else:
        while not path:
            try:
                print 'Please select a file to be parsed...'
                path = fd.askopenfilename()
            except:
                print 'Error selecting file, please try again.'

    # search for string in each file...
    try:
        matches = find_vals_in_file(path)
    except:
        exc_type, exc_value, exc_traceback = sys.exc_info()
        lines = traceback.format_exception(exc_type, exc_value, exc_traceback)
        print "An error occurred checking file:", path
        print ''.join('!! ' + line for line in lines)

    # write output to file...
    if len(matches) == 0:
        print "No matches were found in the files reviewed."
    else:
        heads = get_heads()            
        output_path = os.path.join(os.getcwd(),'tmp_matches.csv')
        write_csv_file(output_path, heads, matches)
        print "Please rename the file if you wish to keep it as it will be overwritten..."
        print "\nOpening file..."
        os.startfile(output_path)

        if debug_on:
            print "\nWriting output to screen\n", "-"*24
            print heads
            for row in matches:
                print row

if __name__ == '__main__':
    main()

希望这对你有用。到目前为止，我只测试了几个不同的 xml 文件，但对我来说还可以。

关于xml - 如何在excel中使一个xml标签成为一行？，我们在Stack Overflow上找到一个类似的问题： https://stackoverflow.com/questions/17305902/

文章推荐： r - 基于交替值的快速排序/过滤

文章推荐： vba - Excel VBA中颜色不同但颜色索引相同

xml - 如何在没有源 xml 文件根节点的情况下将一个 xml 文件包含在另一个 xml 中？
正如标题中所问，我有两个如下结构的 XML 文件 A.xml //here I want to include B.xml
c# - 如何将等 xml 标签格式更改为
我有一个 xml 文件。根据我的要求，我需要更新空标签，例如我需要更改 to .是否可以像那样更改标签.. 谢谢... 最佳答案 var xmlString=" "; var properStri
xml - Golang : get inner xml from xml with xml.解码
我有这样简单的 XML: Song Playing 09:41:18 Frederic Delius Violin Son
xml - XML 阅读器是否应该忽略 XML 文件中的连续空格？
在我的工作中，我们有自己的 XML 类来构建 DOM，但我不确定应该如何处理连续的空格？例如 Hello World 当它被读入 DOM 时，文本节点应该包含 Hello 和 World
xml - 比较来自不同 XML 文件的元素值并附加到第一个 XML
我有以下 2 个 xml 文件，我必须通过比较 wd:Task_Name_ID 和 TaskID 的 XML 文件 2。例如，Main XML File-1 wd:Task_Name_ID 具有以下
xml - 使 XML 构建器从字符串中插入 XML
我在 Rails 应用程序中有一个 XML View ，需要从另一个文件插入 XML 以进行测试。我想说“构建器，只需盲目地填充这个字符串，因为它已经是 xml”，但我在文档中看不到这样做的任何内容
xml - XML 数据和 XML 元数据之间有什么区别？
我正在重建一些 XML 提要，因此我正在研究何时使用元素以及何时使用带有 XML 的属性。一些网站说“数据在元素中，元数据在属性中。” 那么，两者有什么区别呢？让我们以 W3Schools 为例:
xml - 文档中的多个 XML 声明是否为格式正确的 XML？
在同一个文档中有两个 XML 声明是否是格式正确的 XML？ hello 我相信不是，但是我找不到支持我的消息来源。来自 Extensible Markup Language
xml - 在 XML 中包装任意 XML
我需要在包装器 XML 文档中嵌入任意(语法上有效的)XML 文档。嵌入式文档被视为纯文本，在解析包装文档时不需要可解析。我知道“CDATA trick”，但如果内部 XML 文档本身包含 CDAT
xml - XML 解析器和 XML 处理器是否相同？
XML 解析器和 XML 处理器是两个不同的东西吗？他们是两个不同的工作吗？最佳答案 XML 解析器和 XML 处理器是一样的。它不适用于其他语言。 XML 是通用数据标记语言。解析 XML 文件已
xml - 在保留格式的同时从文件读取 XML 和从文件读取 XML
我使用这个 perl 代码从一个文件中读取 XML，然后写入另一个文件(我的完整脚本有添加属性的代码): #!usr/bin/perl -w use strict; use XML::DOM; use
xml - 使用 PowerShell 将 system.xml.xml 元素转换为 system.xml.xml 文档
我正在编写一个我了解有限的历史脚本。对象 A 的类型为 system.xml.xmlelement，我需要将其转换为类型 system.xml.xmldocument 以与对象 B 进行比较(类型
xml - 如何将子节点结构从一个 XML 文件复制到另一个 XML 文件(合并两个 XML 文件)？
我有以下两个 XML 文件: 文件1 101 102 103 501 502 503
xml - 如何将子节点结构从一个 XML 文件复制到另一个 XML 文件(合并两个 XML 文件)？
我有以下两个 XML 文件: 文件1 101 102 103 501 502 503
java - 转换性能 XML>XSL>XML 与 XML>JAXB>XML
我有一个案例，其中一个 xml 作为输入，另一个 xml 作为输出:我可以选择使用 XSL 和通过 JAXB 进行 Unmarshalling 编码。性能方面，有什么真正的区别吗？最佳答案首先，程
java - 从 XML 元素获取 XML 时的标签顺序(XML 包含 XML)？
我有包含 XML 的 XML，我想使用 JAXB 解析它 qwqweqwezxcasdasd eee 解析器 public static NotificationRequest parse(Strin
xml - 无法使用 XML 架构和 Perl (XML::LibXML) 验证 XML
xml: mario de2f15d014d40b93578d255e6221fd60 Mario F 23 maria maria
java.net.MalformedURLException : no protocol: [c:\XML\file. xml，c :\XML\file2. xml，c :\XML\file3. xml]
尝试更新 xml 文件数组时出现以下错误。代码片段: File dir = new File("c:\\XML"); File[] files = dir.listFiles(new Filenam
xml - 如何使用 ConvertTo-Xml 和 Select-Xml 加载或读取 XML 文件？
我怎样才能完成这样的事情: PS /home/nicholas/powershell> PS /home/nicholas/powershell> $date=(Get-Date | ConvertT
xml - 删除 XML 节点以将 XML 日志文件的大小减小到给定大小
我在从 xml 文件中删除节点时遇到一些困难。我发现很多其他人通过各种方式在 powershell 中执行此操作的示例，下面的代码似乎与我见过的许多其他示例相同，但我没有得到所需的行为。我的目标是将

行者123

个人简介

我是一名优秀的程序员,十分优秀！

作者热门文章

滴滴打车优惠券免费领取

全站热门文章

首页

博学

6Ren·AI

商城

xml - 如何在excel中使一个xml标签成为一行？