layout: global title: XML Files displayTitle: XML Files license: | Licensed to the Apache Software Foundation (ASF) under one or more contributor license agreements. See the NOTICE file distributed with this work for additional information regarding copyright ownership. The ASF licenses this file to You under the Apache License, Version 2.0 (the “License”); you may not use this file except in compliance with the License. You may obtain a copy of the License at
http://www.apache.org/licenses/LICENSE-2.0
Spark SQL provides spark.read().xml("file_1_path","file_2_path")
to read a file or directory of files in XML format into a Spark DataFrame, and dataframe.write().xml("path")
to write to a xml file. When reading a XML file, the rowTag
option must be specified to indicate the XML element that maps to a DataFrame row
. The option() function can be used to customize the behavior of reading or writing, such as controlling behavior of the XML attributes, XSD validation, compression, and so on.
Data source options of XML can be set via:
.option
/.options
methods ofDataFrameReader
DataFrameWriter
DataStreamReader
DataStreamWriter
from_xml
to_xml
schema_of_xml
OPTIONS
clause at CREATE TABLE USING DATA_SOURCE