1 / 17

XML and General Dutch Dictionary (ANW)

  . XML and General Dutch Dictionary (ANW). Peter van der Kamp www.inl.nl kamp@inl.nl. Van der Kamp, Lexical databases and digital tools, april 29 th , 2005, 1.   . Topics. Characteristics Schema XML Dictionary Editor Problems to be solved.

upton
Download Presentation

XML and General Dutch Dictionary (ANW)

An Image/Link below is provided (as is) to download presentation Download Policy: Content on the Website is provided to you AS IS for your information and personal use and may not be sold / licensed / shared on other websites without getting consent from its author. Content is provided to you AS IS for your information and personal use only. Download presentation by click this link. While downloading, if for some reason you are not able to download a presentation, the publisher may have deleted the file from their server. During download, if you can't get a presentation, the file might be deleted by the publisher.

E N D

Presentation Transcript


  1.   XML and General Dutch Dictionary (ANW) Peter van der Kamp www.inl.nl kamp@inl.nl Van der Kamp, Lexical databases and digital tools, april 29th, 2005, 1

  2.   Topics • Characteristics • Schema • XML Dictionary Editor • Problems to be solved Van der Kamp, Lexical databases and digital tools, april 29th, 2005, 2

  3.   Characteristics Online dictionary, no printed version Dutch language (incl. Flanders) from 1970 - 2018 Based on a corpus of 100 mio words Elaborated microstructure XML Van der Kamp, Lexical databases and digital tools, april 29th, 2005, 3

  4.   Schema characteristics • Divided into 12 subschemas • Currently all elements: zero or more occurrences except headword • Currently 186 atomic elements • Many enumerations (378, to be used as controlled vocabulary) • Some elements allowed at different levels Van der Kamp, Lexical databases and digital tools, april 29th, 2005, 4

  5.   Schema Entry Entry PoS Sense PoS Van der Kamp, Lexical databases and digital tools, april 29th, 2005, 5

  6.   XML Dictionary Editor • User requirements: • Don’t want to work with tags • Tags invisible Van der Kamp, Lexical databases and digital tools, april 29th, 2005, 6

  7.   XML Dictionary Editor (cont’d) • User requirements: • Form like input • Use of predefined lists (controlled vocabulary) Van der Kamp, Lexical databases and digital tools, april 29th, 2005, 7

  8.   XML Dictionary Editor (cont’d) • User requirements: • Insert, add and remove elements must be easy Van der Kamp, Lexical databases and digital tools, april 29th, 2005, 8

  9.   XML Dictionary Editor (cont’d) • User requirements: • Hide/show elements • Technical requirements • Subschema enabled Van der Kamp, Lexical databases and digital tools, april 29th, 2005, 9

  10.   XML Dictionary Editor (cont’d) XML editor, but… …which one? XMLWriter Van der Kamp, Lexical databases and digital tools, april 29th, 2005, 10

  11.   XML Dictionary Editor (cont’d) Currently the best possible solution: Authentic (free XML content editor from Altova) StyleVision (e-forms and stylesheet designer from Altova) (http://www.altova.com) Van der Kamp, Lexical databases and digital tools, april 29th, 2005, 11

  12.   XML Dictionary Editor: problems Problem: hide element = delete element Hide element important due to size of entry • Solution (to be implemented): • Extra element <hide> in schema • Checkbox as ‘data entry device’ • When unchecked: perform hide Disadvantage: <hide> is noise in dictionary entry Van der Kamp, Lexical databases and digital tools, april 29th, 2005, 13

  13.   XML Dictionary Editor: problems Problem: visualize difference between container elements and atomic elements. Current implementation requires some schema knowledge Van der Kamp, Lexical databases and digital tools, april 29th, 2005, 16

  14.   Conclusion / future work Developing forms easy Current implementation satisfying Database solution (relational vs. xml) Retrieval Easy use of (X)query language Van der Kamp, Lexical databases and digital tools, april 29th, 2005, 17

More Related