Please use this identifier to cite or link to this item: http://hdl.handle.net/1942/609
Title: DTDs versus XML Schema: A Practical Study.
Authors: NEVEN, Frank 
BEX, Geert Jan 
VAN DEN BUSSCHE, Jan 
Issue Date: 2004
Publisher: ACM
Source: ACM International Conference Proceeding Series; Vol. 67 archive Proceedings of the 7th International Workshop on the Web and Databases: colocated with ACM SIGMOD/PODS 2004. p. 79-84.
Abstract: Among the various proposals answering the shortcomings of Document Type Definitions (DTDs), XML Schema is the most widely used. Although DTDs and XML Schema Defintions (XSDs) differ syntactically, they are still quite related on an abstract level. Indeed, freed from all syntactic sugar, XML Schemas can be seen as an extension of DTDs with a restricted form of specialization. In the present paper, we inspect a number of DTDs and XSDs harvested from the web and try to answer the following questions: (1) which of the extra features/expressiveness of XML Schema not allowed by DTDs are effectively used in practice; and, (2) how sophisticated are the structural properties (i.e. the nature of regular expressions) of the two formalisms. It turns out that at present real-world XSDs only sparingly use the new features introduced by XML Schema: on a structural level the vast majority of them can already be defined by DTDs. Further, we introduce a class of simple regular expressions and obtain that a surprisingly high fraction of the content models belong to this class. The latter result sheds light on the justification of simplifying assumptions that sometimes have to be made in XML research.
Document URI: http://hdl.handle.net/1942/609
Link to publication/dataset: http://doi.acm.org/10.1145/1017074.1017095
Category: C2
Type: Proceedings Paper
Appears in Collections:Research publications

Files in This Item:
File Description SizeFormat 
33 6-1.pdf127.57 kBAdobe PDFView/Open
Show full item record

Page view(s)

38
checked on Nov 7, 2023

Download(s)

126
checked on Nov 7, 2023

Google ScholarTM

Check


Items in DSpace are protected by copyright, with all rights reserved, unless otherwise indicated.