Both sides previous revision
Previous revision
Next revision
|
Previous revision
Next revision
Both sides next revision
|
user:zeman:treebanks:hi [2011/12/06 17:39] zeman Hindi development data sample. |
user:zeman:treebanks:hi [2011/12/06 22:35] zeman Inside. |
==== Size ==== | ==== Size ==== |
| |
HyDT-Bangla shows dependencies between chunks, not words. The node/tree ratio is thus much lower than in other treebanks. The ICON 2009 version came with a data split into three parts: training, development and test: | HyDT-Hindi contains dependencies on two levels: between chunks and inside chunks. The ICON 2009 CoNLL-formatted version contained only dependencies between chunks, thus the node/tree ratio was much lower than in other treebanks. The ICON 2009 version came with a data split into three parts: training, development and test: |
| |
^ Part ^ Sentences ^ Chunks ^ Ratio ^ | ^ Part ^ Sentences ^ Chunks ^ Ratio ^ |
| Training | 980 | 6449 | 6.58 | | | Training | 1501 | 13779 | 9.18 | |
| Development | 150 | 811 | 5.41 | | | Development | 150 | 1250 | 8.33 | |
| Test | 150 | 961 | 6.41 | | | Test | 150 | 1156 | 7.71 | |
| TOTAL | 1280 | 8221 | 6.42 | | | TOTAL | 1801 | 16185 | 8.99 | |
| |
The ICON 2010 version came with a data split into three parts: training, development and test: | The ICON 2010 version came with a data split into three parts: training, development and test. The intra-chunk dependencies have been added: |
| |
^ Part ^ Sentences ^ Chunks ^ Ratio ^ Words ^ Ratio ^ | ^ Part ^ Sentences ^ Chunks ^ Ratio ^ Words ^ Ratio ^ |
| Training | 979 | 6440 | 6.58 | 10305 | 10.52 | | | Training | 2972 | | | 64452 | 21.69 | |
| Development | 150 | 812 | 5.41 | 1196 | 7.97 | | | Development | 543 | | | 12616 | 23.23 | |
| Test | 150 | 961 | 6.41 | 1350 | 9.00 | | | Test | 321 | | | 6588 | 20.52 | |
| TOTAL | 1279 | 8213 | 6.42 | 12851 | 10.04 | | | TOTAL | 3836 | | | 83656 | 21.81 | |
| |
I have counted the sentences and chunks. The number of words comes from (Husain et al., 2010). Note that the paper gives the number of training sentences as 980 (instead of 979), which is a mistake. The last training sentence has the id 980 but there is no sentence with id 418. | I have counted the sentences and tokens (words) on the ''.conll'' files; there are slight differences from the statistics presented in (Husain et al., 2010). |
| |
Apparently the training-development-test data split was more or less identical in both years, except for the minor discrepancies (number of training sentences and development chunks). | |
| |
==== Inside ==== | ==== Inside ==== |
| |
* Broken characters (''\x{FFFD} REPLACEMENT CHARACTER'') in the WX encoding. | The text uses the [[http://ltrc.iiit.ac.in/nlptools2010/files/documents/map.pdf|WX encoding]] of Indian letters. If we know what the original script is (Devanagari in this case) we can map the WX encoding to the original characters in UTF-8. WX uses English letters so if there was embedded English (or other string using Latin letters) it will probably get lost during the conversion. Note that there are (not infrequent) broken characters (''\x{FFFD} REPLACEMENT CHARACTER'') in the WX encoding and the correct characters cannot be recovered automatically. |
| |
-- | |
| |
The text uses the [[http://ltrc.iiit.ac.in/nlptools2010/files/documents/map.pdf|WX encoding]] of Indian letters. If we know what the original script is (Bengali in this case) we can map the WX encoding to the original characters in UTF-8. WX uses English letters so if there was embedded English (or other string using Latin letters) it will probably get lost during the conversion. | |
| |
The CoNLL format contains only the chunk heads. The native SSF format shows the other words in the chunk, too, but it does not capture intra-chunk dependency relations. This is an example of a multi-word chunk: | |
| |
<code>3 (( NP <fs af='rumAla,n,,sg,,d,0,0' head="rumAla" drel=k2:VGF name=NP3> | |
3.1 ekatA QC <fs af='eka,num,,,,,,'> | |
3.2 ledisa JJ <fs af='ledisa,unk,,,,,,'> | |
3.3 rumAla NN <fs af='rumAla,n,,sg,,d,0,0' name="rumAla"> | |
))</code> | |
| |
In the CoNLL format, the CPOS column contains the [[http://ltrc.iiit.ac.in/nlptools2010/files/documents/Chunk-Tag-List.pdf|chunk label]] (e.g. ''NP'' = //noun phrase//) and the POS column contains the [[http://ltrc.iiit.ac.in/nlptools2010/files/documents/POS-Tag-List.pdf|part of speech]] of the chunk head. | |
| |
Occasionally there are ''NULL'' nodes that do not correspond to any surface chunk or token. They represent ellided participants. | Occasionally there are ''NULL'' nodes that do not correspond to any surface chunk or token. They represent ellided participants. |
The [[http://ltrc.iiit.ac.in/nlptools2010/files/documents/dep-tagset.pdf|syntactic tags]] (dependency relation labels) are //karaka// relations, i.e. deep syntactic roles according to the Pāṇinian grammar. There are separate versions of the treebank with [[http://ltrc.iiit.ac.in/nlptools2010/files/documents/mapping_fine-to-coarse.pdf|fine-grained and coarse-grained]] syntactic tags. | The [[http://ltrc.iiit.ac.in/nlptools2010/files/documents/dep-tagset.pdf|syntactic tags]] (dependency relation labels) are //karaka// relations, i.e. deep syntactic roles according to the Pāṇinian grammar. There are separate versions of the treebank with [[http://ltrc.iiit.ac.in/nlptools2010/files/documents/mapping_fine-to-coarse.pdf|fine-grained and coarse-grained]] syntactic tags. |
| |
According to [[http://ltrc.iiit.ac.in/nlptools2010/files/documents/toolscontest10-workshoppaper-final.pdf|(Husain et al., 2010)]], in the ICON 2010 version, the chunk tags, POS tags and inter-chunk dependencies (topology + tags) were annotated manually. The rest (lemma, morphosyntactic features, headword of chunk) was marked automatically. | According to [[http://ltrc.iiit.ac.in/nlptools2010/files/documents/toolscontest10-workshoppaper-final.pdf|(Husain et al., 2010)]], in the ICON 2010 version, the chunk tags, POS tags, lemma, morphosyntactic features and inter-chunk dependencies (topology + tags) were annotated manually. The rest (intra-chunk dependencies, headword of chunk) was marked automatically. The tool for intra-chunk dependency parsing achieves about 96% accuracy. |
| |
Note: There have been cycles in the Hindi part of HyDT but no such problem occurs in the Bengali part. | Note: There have been cycles in the Hindi part of HyDT. |
| |
==== Sample ==== | ==== Sample ==== |
The first sentence of the ICON 2010 test data (with fine-grained syntactic tags) in the Shakti format: | The first sentence of the ICON 2010 test data (with fine-grained syntactic tags) in the Shakti format: |
| |
<code xml><document id=""> | <code xml><document docid="fullnews_id_2484368"> |
<head> | <head> |
<annotated-resource name="HyDT-Bangla" version="0.5" type="dep-interchunk-only" layers="morph,pos,chunk,dep-interchunk-only" language="ben" date-of-release="20101013"> | <caption>elaosI Kolane para hogI bAwacIwa pre isalAmAbAxa.</caption> |
| <language>Hindi </language> |
| <domain_name>News Articles </domain_name> |
| <word_count>313</word_count> |
| <byte_count>37563</byte_count> |
| <availability> |
| <format>CML/SSF</format> |
| <sentence_marker>.</sentence_marker> |
| <normalization>No</normalization> |
| </availability> |
| <encoding_description> |
| <original_encoding>ISO 8859</format> |
| <new_encoding>Unicode UTF8</new_encoding> |
| </encoding_description> |
| <distributor>LTRC, IIIT Hyderabad</distributor> |
| <project_description>NSF Hindi/Urdu Dependency Treebanking Project</place> |
| <creation> |
| </raw_corpus creation_date="" institute_name="IIIT Hyderabad"> |
| </annotated_corpus creation_date="06/01/2009" institute_name="IIIT Hyderabad"> |
| <edition_number>1.0</edition_number> |
| </creation> |
| <publication> |
| <place>New Delhi</place> |
| <date>28/5/2004</date> |
| <type>Newspaper</type> |
| <publisher> |
| <name>Amar Ujala</name> |
| <url>http://www.amarujala.com</url> |
| </publisher> |
| </publication> |
| |
| <annotated-resource name="HyDT-Hindi" version="2.0" type="dep-words" layers="morph,pos,chunk,dep-word" language="hin" date-of-release="20101013"> |
<annotation-standard> | <annotation-standard> |
<morph-standard name="Anncorra-morph" version="1.31" date="20080920" /> | <morph-standard name="Anncorra-morph" version="1.31" date="20080920" /> |
<pos-standard name="Anncorra-pos" version="" date="20061215" /> | <pos-standard name="Anncorra-pos" version="" date="20061215" /> |
<chunk-standard name="Anncorra-chunk" version="" date="20061215" /> | <chunk-standard name="Anncorra-chunk" version="" date="20061215" /> |
<dependency-standard name="Anncorra-dep" version="2.0" date="" dep-tagset-granularity="6" /> | <intrachunk-dependency-standard name="Anncorra-intrachunk-dep" version="1.0" date="" dep-tagset-granularity="5" /> |
| <dependency-standard name="Anncorra-dep" version="2.0" date="" dep-tagset-granularity="6" /> |
</annotation-standard> | </annotation-standard> |
<annotated-resource> | </annotated-resource> |
</head> | </head> |
| <body> |
| <tb number="1" segment="no" bullet="no"> |
| <foreign language="select" writingsystem="LTR"></foreign> |
| <text> |
<Sentence id="1"> | <Sentence id="1"> |
1 (( NP <fs af='mAXabIlawA,n,,sg,,d,0,0' head="mAXabIlawA" drel=k1:VGF name=NP> | 1 pAkiswAna XC <fs af='pAkiswAna,n,m,sg,3,d,0,0' posn='10' drel='mod:kaSmIra' chunkType='child:NP' name='pAkiswAna'> |
1.1 mAXabIlawA NNP <fs af='mAXabIlawA,n,,sg,,d,0,0' name="mAXabIlawA"> | 2 aXikqwa XC <fs af='aXikqwa,adj,any,any,,o,,' posn='20' drel='mod:kaSmIra' chunkType='child:NP' name='aXikqwa'> |
)) | 3 kaSmIra NNP <fs af='kaSmIra,n,m,sg,3,o,0_meM,0' drel='k7p:Ae' posn='30' vpos='vib_4' name='kaSmIra' chunkId='NP' chunkType='head:NP'> |
2 (( NP <fs af='waKana,pn,,,,d,0,0' head="waKana" drel=k7t:VGF name=NP2> | 4 meM PSP <fs af='meM,psp,,,,,,' posn='40' drel='lwg__psp:kaSmIra' chunkType='child:NP' name='meM'> |
2.1 waKana PRP <fs af='waKana,pn,,,,d,0,0' name="waKana"> | 5 � XC <fs af='�,num,m,sg,3,d,,' posn='50' drel='mod:akwUbara' chunkType='child:NP2' name='�'> |
)) | 6 akwUbara NNP <fs af='akwUbara,n,m,sg,3,o,0_ko,0' drel='k7t:Ae' posn='60' vpos='vib_3' name='akwUbara' chunkId='NP2' chunkType='head:NP2'> |
3 (( NP <fs af='hAwa,n,,sg,,o,era,era' head="hAwera" drel=r6:NP4 name=NP3> | 7 ko PSP <fs af='ko,psp,,,,,,' posn='70' drel='lwg__psp:akwUbara' chunkType='child:NP2' name='ko'> |
3.1 hAwera NN <fs af='hAwa,n,,sg,,o,era,era' name="hAwera"> | 8 Ae VM <fs af='A,v,m,sg,any,,yA,yA' drel='nmod__k1inv:BUkaMpa' posn='80' name='Ae' chunkId='VGNF' chunkType='head:VGNF'> |
)) | 9 BUkaMpa NN <fs af='BUkaMpa,n,m,sg,3,o,0_se,0' drel='rh:macI' posn='90' vpos='vib_2' name='BUkaMpa' chunkId='NP3' chunkType='head:NP3'> |
4 (( NP <fs af='GadZi,unk,,,,,,' head="GadZi" drel=k2:VGNF name=NP4> | 10 se PSP <fs af='se,psp,,,,,,' posn='100' drel='lwg__psp:BUkaMpa' chunkType='child:NP3' name='se'> |
4.1 GadZi NN <fs af='GadZi,unk,,,,,,' name="GadZi"> | 11 macI VM <fs af='maca,v,f,sg,any,,yA,yA' drel='nmod__k1inv:wabAhI' posn='110' name='macI' chunkId='VGNF2' chunkType='head:VGNF2'> |
)) | 12 wabAhI NN <fs af='wabAhI,n,f,sg,3,o,0_kA_bAxa,0' drel='k7t:kareMge' posn='120' vpos='vib_2_3' name='wabAhI' chunkId='NP4' chunkType='head:NP4'> |
5 (( VGNF <fs af='Kul,v,,,5,,ne,ne' head="Kule" drel=vmod:VGF name=VGNF> | 13 ke PSP <fs af='kA,psp,m,sg,3,o,,' posn='130' drel='lwg__psp:wabAhI' chunkType='child:NP4' name='ke'> |
5.1 Kule VM <fs af='Kul,v,,,5,,ne,ne' name="Kule"> | 14 bAxa NST <fs af='bAxa,n,,,,,,' posn='140' drel='lwg__psp:wabAhI' chunkType='child:NP4' name='bAxa'> |
)) | 15 BArawa NNP <fs af='BArawa,n,m,sg,3,d,0,0' drel='ccof:Ora' posn='150' name='BArawa' chunkId='NP5' chunkType='head:NP5'> |
6 (( NP <fs af='tebila,n,,sg,,d,me,me' head="tebile" drel=k7p:VGF name=NP5> | 16 Ora CC <fs af='Ora,avy,,,,,,' drel='k1:kareMge' posn='160' name='Ora' chunkId='CCP' chunkType='head:CCP'> |
6.1 tebile NN <fs af='tebila,n,,sg,,d,me,me' name="tebile"> | 17 pAkiswAna NNP <fs af='pAkiswAna,n,m,sg,3,d,0,0' drel='ccof:Ora' posn='170' name='pAkiswAna2' chunkId='NP6' chunkType='head:NP6'> |
)) | 18 mAnavIya JJ <fs af='mAnavIya,adj,any,any,,o,,' posn='180' drel='nmod__adj:xqRtikoNa' chunkType='child:NP7' name='mAnavIya'> |
7 (( VGF <fs af='rAK,v,,,5,,Cila,Cila' head="rAKaCila" name=VGF> | 19 xqRtikoNa NN <fs af='xqRtikoNa,n,m,sg,3,d,0,0' drel='k2:apanAwe' posn='190' name='xqRtikoNa' chunkId='NP7' chunkType='head:NP7'> |
7.1 rAKaCila VM <fs af='rAK,v,,,5,,Cila,Cila' name="rAKaCila"> | 20 apanAwe VM <fs af='apanA,v,m,pl,any,,wA_ho+yA,wA' drel='vmod:kareMge' posn='200' vpos='tam_2' name='apanAwe' chunkId='VGNF3' chunkType='head:VGNF3'> |
7.2 । SYM | 21 hue VAUX <fs af='ho,v,m,pl,any,,yA,yA' posn='210' drel='lwg__vaux:apanAwe' chunkType='child:VGNF3' name='hue'> |
)) | 22 SanivAra NNP <fs af='SanivAra,n,m,sg,3,o,0_ko,0' drel='k7t:kareMge' posn='220' vpos='vib_2' name='SanivAra' chunkId='NP8' chunkType='head:NP8'> |
| 23 ko PSP <fs af='ko,psp,,,,,,' posn='230' drel='lwg__psp:SanivAra' chunkType='child:NP8' name='ko2'> |
| 24 islAmAbAxa NNP <fs af='isalAmAbAxa,n,m,sg,3,d,0_meM,0' drel='k7p:kareMge' posn='240' vpos='vib_2' name='islAmAbAxa' chunkId='NP9' chunkType='head:NP9'> |
| 25 meM PSP <fs af='meM,psp,,,,,,' posn='250' drel='lwg__psp:islAmAbAxa' chunkType='child:NP9' name='meM2'> |
| 26 niyaMwraNa XC <fs af='niyaMwraNa,n,m,sg,3,d,0,0' posn='260' drel='mod:reKA' chunkType='child:NP10' name='niyaMwraNa'> |
| 27 reKA NN <fs af='reKA,n,f,sg,3,d,0,0' drel='k2:Kolane' posn='270' name='reKA' chunkId='NP10' chunkType='head:NP10'> |
| 28 ( SYM <fs af=',punc,,,,,,' posn='280' drel='rsym:elaosI' chunkType='child:NP11' name='('> |
| 29 elaosI NN <fs af='elaosI,n,m,sg,3,d,0,0' drel='nmod:reKA' posn='290' name='elaosI' chunkId='NP11' chunkType='head:NP11'> |
| 30 ) SYM <fs af=',punc,,,,,,' posn='300' drel='rsym:elaosI' chunkType='child:NP11' name=')'> |
| 31 Kolane VM <fs af='Kola,v,any,sg,any,o,nA_kA,nA' drel='r6:masale' posn='310' vpos='tam_2' name='Kolane' chunkId='VGNN' chunkType='head:VGNN'> |
| 32 ke PSP <fs af='kA,psp,m,sg,,o,,' posn='320' drel='lwg__psp:Kolane' chunkType='child:VGNN' name='ke2'> |
| 33 masale NN <fs af='masalA,n,m,sg,3,o,0_para,0' drel='k7:kareMge' posn='330' vpos='vib_2' name='masale' chunkId='NP12' chunkType='head:NP12'> |
| 34 para PSP <fs af='para,psp,,,,,,' posn='340' drel='lwg__psp:masale' chunkType='child:NP12' name='para'> |
| 35 bAwacIwa NN <fs af='bAwacIwa,n,f,sg,3,d,0,0' drel='pof:kareMge' posn='350' name='bAwacIwa' chunkId='NP13' chunkType='head:NP13'> |
| 36 kareMge VM <fs af='kara,v,m,pl,3,,gA,gA' posn='360' name='kareMge' chunkId='VGF' chunkType='head:VGF'> |
| 37 . SYM <fs af='.,punc,,,,,,' posn='370' drel='rsym:kareMge' chunkType='child:VGF' name='.'> |
</Sentence></code> | </Sentence></code> |
| |
And in the CoNLL format: | And in the CoNLL format: |
| |
| 1 | mAXabIlawA | mAXabIlawA | NP | NNP | lex-mAXabIlawA<nowiki>|</nowiki>cat-n<nowiki>|</nowiki>gend-<nowiki>|</nowiki>num-sg<nowiki>|</nowiki>pers-<nowiki>|</nowiki>case-d<nowiki>|</nowiki>vib-0<nowiki>|</nowiki>tam-0<nowiki>|</nowiki>head-mAXabIlawA<nowiki>|</nowiki>name-NP | 7 | k1 | _ | _ | | | <nowiki>1</nowiki> | <nowiki>pAkiswAna</nowiki> | <nowiki>pAkiswAna</nowiki> | <nowiki>XC</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-pAkiswAna|cat-n|gend-m|num-sg|pers-3|case-d|vib-0|tam-0|posn-10|chunkType-child:NP|name-pAkiswAna</nowiki> | <nowiki>3</nowiki> | <nowiki>mod</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| 2 | waKana | waKana | NP | PRP | lex-waKana<nowiki>|</nowiki>cat-pn<nowiki>|</nowiki>gend-<nowiki>|</nowiki>num-<nowiki>|</nowiki>pers-<nowiki>|</nowiki>case-d<nowiki>|</nowiki>vib-0<nowiki>|</nowiki>tam-0<nowiki>|</nowiki>head-waKana<nowiki>|</nowiki>name-NP2 | 7 | k7t | _ | _ | | | <nowiki>2</nowiki> | <nowiki>aXikqwa</nowiki> | <nowiki>aXikqwa</nowiki> | <nowiki>XC</nowiki> | <nowiki>adj</nowiki> | <nowiki>lex-aXikqwa|cat-adj|gend-any|num-any|pers-|case-o|vib-|tam-|posn-20|chunkType-child:NP|name-aXikqwa</nowiki> | <nowiki>3</nowiki> | <nowiki>mod</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| 3 | hAwera | hAwa | NP | NN | lex-hAwa<nowiki>|</nowiki>cat-n<nowiki>|</nowiki>gend-<nowiki>|</nowiki>num-sg<nowiki>|</nowiki>pers-<nowiki>|</nowiki>case-o<nowiki>|</nowiki>vib-era<nowiki>|</nowiki>tam-era<nowiki>|</nowiki>head-hAwera<nowiki>|</nowiki>name-NP3 | 4 | r6 | _ | _ | | | <nowiki>3</nowiki> | <nowiki>kaSmIra</nowiki> | <nowiki>kaSmIra</nowiki> | <nowiki>NNP</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-kaSmIra|cat-n|gend-m|num-sg|pers-3|case-o|vib-0_meM|tam-0|posn-30|vpos-vib_4|name-kaSmIra|chunkId-NP|chunkType-head:NP</nowiki> | <nowiki>8</nowiki> | <nowiki>k7p</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| 4 | GadZi | GadZi | NP | NN | lex-GadZi<nowiki>|</nowiki>cat-unk<nowiki>|</nowiki>gend-<nowiki>|</nowiki>num-<nowiki>|</nowiki>pers-<nowiki>|</nowiki>case-<nowiki>|</nowiki>vib-<nowiki>|</nowiki>tam-<nowiki>|</nowiki>head-GadZi<nowiki>|</nowiki>name-NP4 | 5 | k2 | _ | _ | | | <nowiki>4</nowiki> | <nowiki>meM</nowiki> | <nowiki>meM</nowiki> | <nowiki>PSP</nowiki> | <nowiki>psp</nowiki> | <nowiki>lex-meM|cat-psp|gend-|num-|pers-|case-|vib-|tam-|posn-40|chunkType-child:NP|name-meM</nowiki> | <nowiki>3</nowiki> | <nowiki>lwg__psp</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| 5 | Kule | Kul | VGNF | VM | lex-Kul<nowiki>|</nowiki>cat-v<nowiki>|</nowiki>gend-<nowiki>|</nowiki>num-<nowiki>|</nowiki>pers-5<nowiki>|</nowiki>case-<nowiki>|</nowiki>vib-ne<nowiki>|</nowiki>tam-ne<nowiki>|</nowiki>head-Kule<nowiki>|</nowiki>name-VGNF | 7 | vmod | _ | _ | | | <nowiki>5</nowiki> | <nowiki>�</nowiki> | <nowiki>�</nowiki> | <nowiki>XC</nowiki> | <nowiki>num</nowiki> | <nowiki>lex-�|cat-num|gend-m|num-sg|pers-3|case-d|vib-|tam-|posn-50|chunkType-child:NP2|name-�</nowiki> | <nowiki>6</nowiki> | <nowiki>mod</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| 6 | tebile | tebila | NP | NN | lex-tebila<nowiki>|</nowiki>cat-n<nowiki>|</nowiki>gend-<nowiki>|</nowiki>num-sg<nowiki>|</nowiki>pers-<nowiki>|</nowiki>case-d<nowiki>|</nowiki>vib-me<nowiki>|</nowiki>tam-me<nowiki>|</nowiki>head-tebile<nowiki>|</nowiki>name-NP5 | 7 | k7p | _ | _ | | | <nowiki>6</nowiki> | <nowiki>akwUbara</nowiki> | <nowiki>akwUbara</nowiki> | <nowiki>NNP</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-akwUbara|cat-n|gend-m|num-sg|pers-3|case-o|vib-0_ko|tam-0|posn-60|vpos-vib_3|name-akwUbara|chunkId-NP2|chunkType-head:NP2</nowiki> | <nowiki>8</nowiki> | <nowiki>k7t</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| 7 | rAKaCila | rAK | VGF | VM | lex-rAK<nowiki>|</nowiki>cat-v<nowiki>|</nowiki>gend-<nowiki>|</nowiki>num-<nowiki>|</nowiki>pers-5<nowiki>|</nowiki>case-<nowiki>|</nowiki>vib-Cila<nowiki>|</nowiki>tam-Cila<nowiki>|</nowiki>head-rAKaCila<nowiki>|</nowiki>name-VGF | 0 | main | _ | _ | | | <nowiki>7</nowiki> | <nowiki>ko</nowiki> | <nowiki>ko</nowiki> | <nowiki>PSP</nowiki> | <nowiki>psp</nowiki> | <nowiki>lex-ko|cat-psp|gend-|num-|pers-|case-|vib-|tam-|posn-70|chunkType-child:NP2|name-ko</nowiki> | <nowiki>6</nowiki> | <nowiki>lwg__psp</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>8</nowiki> | <nowiki>Ae</nowiki> | <nowiki>A</nowiki> | <nowiki>VM</nowiki> | <nowiki>v</nowiki> | <nowiki>lex-A|cat-v|gend-m|num-sg|pers-any|case-|vib-yA|tam-yA|posn-80|name-Ae|chunkId-VGNF|chunkType-head:VGNF</nowiki> | <nowiki>9</nowiki> | <nowiki>nmod__k1inv</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>9</nowiki> | <nowiki>BUkaMpa</nowiki> | <nowiki>BUkaMpa</nowiki> | <nowiki>NN</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-BUkaMpa|cat-n|gend-m|num-sg|pers-3|case-o|vib-0_se|tam-0|posn-90|vpos-vib_2|name-BUkaMpa|chunkId-NP3|chunkType-head:NP3</nowiki> | <nowiki>11</nowiki> | <nowiki>rh</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>10</nowiki> | <nowiki>se</nowiki> | <nowiki>se</nowiki> | <nowiki>PSP</nowiki> | <nowiki>psp</nowiki> | <nowiki>lex-se|cat-psp|gend-|num-|pers-|case-|vib-|tam-|posn-100|chunkType-child:NP3|name-se</nowiki> | <nowiki>9</nowiki> | <nowiki>lwg__psp</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>11</nowiki> | <nowiki>macI</nowiki> | <nowiki>maca</nowiki> | <nowiki>VM</nowiki> | <nowiki>v</nowiki> | <nowiki>lex-maca|cat-v|gend-f|num-sg|pers-any|case-|vib-yA|tam-yA|posn-110|name-macI|chunkId-VGNF2|chunkType-head:VGNF2</nowiki> | <nowiki>12</nowiki> | <nowiki>nmod__k1inv</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>12</nowiki> | <nowiki>wabAhI</nowiki> | <nowiki>wabAhI</nowiki> | <nowiki>NN</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-wabAhI|cat-n|gend-f|num-sg|pers-3|case-o|vib-0_kA_bAxa|tam-0|posn-120|vpos-vib_2_3|name-wabAhI|chunkId-NP4|chunkType-head:NP4</nowiki> | <nowiki>36</nowiki> | <nowiki>k7t</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>13</nowiki> | <nowiki>ke</nowiki> | <nowiki>kA</nowiki> | <nowiki>PSP</nowiki> | <nowiki>psp</nowiki> | <nowiki>lex-kA|cat-psp|gend-m|num-sg|pers-3|case-o|vib-|tam-|posn-130|chunkType-child:NP4|name-ke</nowiki> | <nowiki>12</nowiki> | <nowiki>lwg__psp</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>14</nowiki> | <nowiki>bAxa</nowiki> | <nowiki>bAxa</nowiki> | <nowiki>NST</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-bAxa|cat-n|gend-|num-|pers-|case-|vib-|tam-|posn-140|chunkType-child:NP4|name-bAxa</nowiki> | <nowiki>12</nowiki> | <nowiki>lwg__psp</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>15</nowiki> | <nowiki>BArawa</nowiki> | <nowiki>BArawa</nowiki> | <nowiki>NNP</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-BArawa|cat-n|gend-m|num-sg|pers-3|case-d|vib-0|tam-0|posn-150|name-BArawa|chunkId-NP5|chunkType-head:NP5</nowiki> | <nowiki>16</nowiki> | <nowiki>ccof</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>16</nowiki> | <nowiki>Ora</nowiki> | <nowiki>Ora</nowiki> | <nowiki>CC</nowiki> | <nowiki>avy</nowiki> | <nowiki>lex-Ora|cat-avy|gend-|num-|pers-|case-|vib-|tam-|posn-160|name-Ora|chunkId-CCP|chunkType-head:CCP</nowiki> | <nowiki>36</nowiki> | <nowiki>k1</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>17</nowiki> | <nowiki>pAkiswAna</nowiki> | <nowiki>pAkiswAna</nowiki> | <nowiki>NNP</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-pAkiswAna|cat-n|gend-m|num-sg|pers-3|case-d|vib-0|tam-0|posn-170|name-pAkiswAna2|chunkId-NP6|chunkType-head:NP6</nowiki> | <nowiki>16</nowiki> | <nowiki>ccof</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>18</nowiki> | <nowiki>mAnavIya</nowiki> | <nowiki>mAnavIya</nowiki> | <nowiki>JJ</nowiki> | <nowiki>adj</nowiki> | <nowiki>lex-mAnavIya|cat-adj|gend-any|num-any|pers-|case-o|vib-|tam-|posn-180|chunkType-child:NP7|name-mAnavIya</nowiki> | <nowiki>19</nowiki> | <nowiki>nmod__adj</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>19</nowiki> | <nowiki>xqRtikoNa</nowiki> | <nowiki>xqRtikoNa</nowiki> | <nowiki>NN</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-xqRtikoNa|cat-n|gend-m|num-sg|pers-3|case-d|vib-0|tam-0|posn-190|name-xqRtikoNa|chunkId-NP7|chunkType-head:NP7</nowiki> | <nowiki>20</nowiki> | <nowiki>k2</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>20</nowiki> | <nowiki>apanAwe</nowiki> | <nowiki>apanA</nowiki> | <nowiki>VM</nowiki> | <nowiki>v</nowiki> | <nowiki>lex-apanA|cat-v|gend-m|num-pl|pers-any|case-|vib-wA_ho+yA|tam-wA|posn-200|vpos-tam_2|name-apanAwe|chunkId-VGNF3|chunkType-head:VGNF3</nowiki> | <nowiki>36</nowiki> | <nowiki>vmod</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>21</nowiki> | <nowiki>hue</nowiki> | <nowiki>ho</nowiki> | <nowiki>VAUX</nowiki> | <nowiki>v</nowiki> | <nowiki>lex-ho|cat-v|gend-m|num-pl|pers-any|case-|vib-yA|tam-yA|posn-210|chunkType-child:VGNF3|name-hue</nowiki> | <nowiki>20</nowiki> | <nowiki>lwg__vaux</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>22</nowiki> | <nowiki>SanivAra</nowiki> | <nowiki>SanivAra</nowiki> | <nowiki>NNP</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-SanivAra|cat-n|gend-m|num-sg|pers-3|case-o|vib-0_ko|tam-0|posn-220|vpos-vib_2|name-SanivAra|chunkId-NP8|chunkType-head:NP8</nowiki> | <nowiki>36</nowiki> | <nowiki>k7t</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>23</nowiki> | <nowiki>ko</nowiki> | <nowiki>ko</nowiki> | <nowiki>PSP</nowiki> | <nowiki>psp</nowiki> | <nowiki>lex-ko|cat-psp|gend-|num-|pers-|case-|vib-|tam-|posn-230|chunkType-child:NP8|name-ko2</nowiki> | <nowiki>22</nowiki> | <nowiki>lwg__psp</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>24</nowiki> | <nowiki>islAmAbAxa</nowiki> | <nowiki>isalAmAbAxa</nowiki> | <nowiki>NNP</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-isalAmAbAxa|cat-n|gend-m|num-sg|pers-3|case-d|vib-0_meM|tam-0|posn-240|vpos-vib_2|name-islAmAbAxa|chunkId-NP9|chunkType-head:NP9</nowiki> | <nowiki>36</nowiki> | <nowiki>k7p</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>25</nowiki> | <nowiki>meM</nowiki> | <nowiki>meM</nowiki> | <nowiki>PSP</nowiki> | <nowiki>psp</nowiki> | <nowiki>lex-meM|cat-psp|gend-|num-|pers-|case-|vib-|tam-|posn-250|chunkType-child:NP9|name-meM2</nowiki> | <nowiki>24</nowiki> | <nowiki>lwg__psp</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>26</nowiki> | <nowiki>niyaMwraNa</nowiki> | <nowiki>niyaMwraNa</nowiki> | <nowiki>XC</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-niyaMwraNa|cat-n|gend-m|num-sg|pers-3|case-d|vib-0|tam-0|posn-260|chunkType-child:NP10|name-niyaMwraNa</nowiki> | <nowiki>27</nowiki> | <nowiki>mod</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>27</nowiki> | <nowiki>reKA</nowiki> | <nowiki>reKA</nowiki> | <nowiki>NN</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-reKA|cat-n|gend-f|num-sg|pers-3|case-d|vib-0|tam-0|posn-270|name-reKA|chunkId-NP10|chunkType-head:NP10</nowiki> | <nowiki>31</nowiki> | <nowiki>k2</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>28</nowiki> | <nowiki>(</nowiki> | <nowiki>(</nowiki> | <nowiki>SYM</nowiki> | <nowiki>punc</nowiki> | <nowiki>lex-|cat-punc|gend-|num-|pers-|case-|vib-|tam-|posn-280|chunkType-child:NP11|name-(</nowiki> | <nowiki>29</nowiki> | <nowiki>rsym</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>29</nowiki> | <nowiki>elaosI</nowiki> | <nowiki>elaosI</nowiki> | <nowiki>NN</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-elaosI|cat-n|gend-m|num-sg|pers-3|case-d|vib-0|tam-0|posn-290|name-elaosI|chunkId-NP11|chunkType-head:NP11</nowiki> | <nowiki>27</nowiki> | <nowiki>nmod</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>30</nowiki> | <nowiki>)</nowiki> | <nowiki>)</nowiki> | <nowiki>SYM</nowiki> | <nowiki>punc</nowiki> | <nowiki>lex-|cat-punc|gend-|num-|pers-|case-|vib-|tam-|posn-300|chunkType-child:NP11|name-)</nowiki> | <nowiki>29</nowiki> | <nowiki>rsym</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>31</nowiki> | <nowiki>Kolane</nowiki> | <nowiki>Kola</nowiki> | <nowiki>VM</nowiki> | <nowiki>v</nowiki> | <nowiki>lex-Kola|cat-v|gend-any|num-sg|pers-any|case-o|vib-nA_kA|tam-nA|posn-310|vpos-tam_2|name-Kolane|chunkId-VGNN|chunkType-head:VGNN</nowiki> | <nowiki>33</nowiki> | <nowiki>r6</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>32</nowiki> | <nowiki>ke</nowiki> | <nowiki>kA</nowiki> | <nowiki>PSP</nowiki> | <nowiki>psp</nowiki> | <nowiki>lex-kA|cat-psp|gend-m|num-sg|pers-|case-o|vib-|tam-|posn-320|chunkType-child:VGNN|name-ke2</nowiki> | <nowiki>31</nowiki> | <nowiki>lwg__psp</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>33</nowiki> | <nowiki>masale</nowiki> | <nowiki>masalA</nowiki> | <nowiki>NN</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-masalA|cat-n|gend-m|num-sg|pers-3|case-o|vib-0_para|tam-0|posn-330|vpos-vib_2|name-masale|chunkId-NP12|chunkType-head:NP12</nowiki> | <nowiki>36</nowiki> | <nowiki>k7</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>34</nowiki> | <nowiki>para</nowiki> | <nowiki>para</nowiki> | <nowiki>PSP</nowiki> | <nowiki>psp</nowiki> | <nowiki>lex-para|cat-psp|gend-|num-|pers-|case-|vib-|tam-|posn-340|chunkType-child:NP12|name-para</nowiki> | <nowiki>33</nowiki> | <nowiki>lwg__psp</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>35</nowiki> | <nowiki>bAwacIwa</nowiki> | <nowiki>bAwacIwa</nowiki> | <nowiki>NN</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-bAwacIwa|cat-n|gend-f|num-sg|pers-3|case-d|vib-0|tam-0|posn-350|name-bAwacIwa|chunkId-NP13|chunkType-head:NP13</nowiki> | <nowiki>36</nowiki> | <nowiki>pof</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>36</nowiki> | <nowiki>kareMge</nowiki> | <nowiki>kara</nowiki> | <nowiki>VM</nowiki> | <nowiki>v</nowiki> | <nowiki>lex-kara|cat-v|gend-m|num-pl|pers-3|case-|vib-gA|tam-gA|posn-360|name-kareMge|chunkId-VGF|chunkType-head:VGF</nowiki> | <nowiki>0</nowiki> | <nowiki>main</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>37</nowiki> | <nowiki>.</nowiki> | <nowiki>.</nowiki> | <nowiki>SYM</nowiki> | <nowiki>punc</nowiki> | <nowiki>lex-.|cat-punc|gend-|num-|pers-|case-|vib-|tam-|posn-370|chunkType-child:VGF|name-.</nowiki> | <nowiki>36</nowiki> | <nowiki>rsym</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| |
And after conversion of the WX encoding to the Bengali script in UTF-8: | And after conversion of the WX encoding to the Devanagari script in UTF-8: |
| |
| 1 | মাধবীলতা | মাধবীলতা | NP | NNP | lex-mAXabIlawA<nowiki>|</nowiki>cat-n<nowiki>|</nowiki>gend-<nowiki>|</nowiki>num-sg<nowiki>|</nowiki>pers-<nowiki>|</nowiki>case-d<nowiki>|</nowiki>vib-0<nowiki>|</nowiki>tam-0<nowiki>|</nowiki>head-mAXabIlawA<nowiki>|</nowiki>name-NP | 7 | k1 | _ | _ | | | <nowiki>1</nowiki> | <nowiki>पाकिस्तान</nowiki> | <nowiki>पाकिस्तान</nowiki> | <nowiki>XC</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-pAkiswAna|cat-n|gend-m|num-sg|pers-3|case-d|vib-0|tam-0|posn-10|chunkType-child:NP|name-pAkiswAna</nowiki> | <nowiki>3</nowiki> | <nowiki>mod</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| 2 | তখন | তখন | NP | PRP | lex-waKana<nowiki>|</nowiki>cat-pn<nowiki>|</nowiki>gend-<nowiki>|</nowiki>num-<nowiki>|</nowiki>pers-<nowiki>|</nowiki>case-d<nowiki>|</nowiki>vib-0<nowiki>|</nowiki>tam-0<nowiki>|</nowiki>head-waKana<nowiki>|</nowiki>name-NP2 | 7 | k7t | _ | _ | | | <nowiki>2</nowiki> | <nowiki>अधिकृत</nowiki> | <nowiki>अधिकृत</nowiki> | <nowiki>XC</nowiki> | <nowiki>adj</nowiki> | <nowiki>lex-aXikqwa|cat-adj|gend-any|num-any|pers-|case-o|vib-|tam-|posn-20|chunkType-child:NP|name-aXikqwa</nowiki> | <nowiki>3</nowiki> | <nowiki>mod</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| 3 | হাতের | হাত | NP | NN | lex-hAwa<nowiki>|</nowiki>cat-n<nowiki>|</nowiki>gend-<nowiki>|</nowiki>num-sg<nowiki>|</nowiki>pers-<nowiki>|</nowiki>case-o<nowiki>|</nowiki>vib-era<nowiki>|</nowiki>tam-era<nowiki>|</nowiki>head-hAwera<nowiki>|</nowiki>name-NP3 | 4 | r6 | _ | _ | | | <nowiki>3</nowiki> | <nowiki>कश्मीर</nowiki> | <nowiki>कश्मीर</nowiki> | <nowiki>NNP</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-kaSmIra|cat-n|gend-m|num-sg|pers-3|case-o|vib-0_meM|tam-0|posn-30|vpos-vib_4|name-kaSmIra|chunkId-NP|chunkType-head:NP</nowiki> | <nowiki>8</nowiki> | <nowiki>k7p</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| 4 | ঘড়ি | ঘড়ি | NP | NN | lex-GadZi<nowiki>|</nowiki>cat-unk<nowiki>|</nowiki>gend-<nowiki>|</nowiki>num-<nowiki>|</nowiki>pers-<nowiki>|</nowiki>case-<nowiki>|</nowiki>vib-<nowiki>|</nowiki>tam-<nowiki>|</nowiki>head-GadZi<nowiki>|</nowiki>name-NP4 | 5 | k2 | _ | _ | | | <nowiki>4</nowiki> | <nowiki>में</nowiki> | <nowiki>में</nowiki> | <nowiki>PSP</nowiki> | <nowiki>psp</nowiki> | <nowiki>lex-meM|cat-psp|gend-|num-|pers-|case-|vib-|tam-|posn-40|chunkType-child:NP|name-meM</nowiki> | <nowiki>3</nowiki> | <nowiki>lwg__psp</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| 5 | খুলে | খুল্ | VGNF | VM | lex-Kul<nowiki>|</nowiki>cat-v<nowiki>|</nowiki>gend-<nowiki>|</nowiki>num-<nowiki>|</nowiki>pers-5<nowiki>|</nowiki>case-<nowiki>|</nowiki>vib-ne<nowiki>|</nowiki>tam-ne<nowiki>|</nowiki>head-Kule<nowiki>|</nowiki>name-VGNF | 7 | vmod | _ | _ | | | <nowiki>5</nowiki> | <nowiki>�</nowiki> | <nowiki>�</nowiki> | <nowiki>XC</nowiki> | <nowiki>num</nowiki> | <nowiki>lex-�|cat-num|gend-m|num-sg|pers-3|case-d|vib-|tam-|posn-50|chunkType-child:NP2|name-�</nowiki> | <nowiki>6</nowiki> | <nowiki>mod</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| 6 | টেবিলে | টেবিল | NP | NN | lex-tebila<nowiki>|</nowiki>cat-n<nowiki>|</nowiki>gend-<nowiki>|</nowiki>num-sg<nowiki>|</nowiki>pers-<nowiki>|</nowiki>case-d<nowiki>|</nowiki>vib-me<nowiki>|</nowiki>tam-me<nowiki>|</nowiki>head-tebile<nowiki>|</nowiki>name-NP5 | 7 | k7p | _ | _ | | | <nowiki>6</nowiki> | <nowiki>अक्तूबर</nowiki> | <nowiki>अक्तूबर</nowiki> | <nowiki>NNP</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-akwUbara|cat-n|gend-m|num-sg|pers-3|case-o|vib-0_ko|tam-0|posn-60|vpos-vib_3|name-akwUbara|chunkId-NP2|chunkType-head:NP2</nowiki> | <nowiki>8</nowiki> | <nowiki>k7t</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| 7 | রাখছিল | রাখ্ | VGF | VM | lex-rAK<nowiki>|</nowiki>cat-v<nowiki>|</nowiki>gend-<nowiki>|</nowiki>num-<nowiki>|</nowiki>pers-5<nowiki>|</nowiki>case-<nowiki>|</nowiki>vib-Cila<nowiki>|</nowiki>tam-Cila<nowiki>|</nowiki>head-rAKaCila<nowiki>|</nowiki>name-VGF | 0 | main | _ | _ | | | <nowiki>7</nowiki> | <nowiki>को</nowiki> | <nowiki>को</nowiki> | <nowiki>PSP</nowiki> | <nowiki>psp</nowiki> | <nowiki>lex-ko|cat-psp|gend-|num-|pers-|case-|vib-|tam-|posn-70|chunkType-child:NP2|name-ko</nowiki> | <nowiki>6</nowiki> | <nowiki>lwg__psp</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>8</nowiki> | <nowiki>आए</nowiki> | <nowiki>आ</nowiki> | <nowiki>VM</nowiki> | <nowiki>v</nowiki> | <nowiki>lex-A|cat-v|gend-m|num-sg|pers-any|case-|vib-yA|tam-yA|posn-80|name-Ae|chunkId-VGNF|chunkType-head:VGNF</nowiki> | <nowiki>9</nowiki> | <nowiki>nmod__k1inv</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>9</nowiki> | <nowiki>भूकंप</nowiki> | <nowiki>भूकंप</nowiki> | <nowiki>NN</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-BUkaMpa|cat-n|gend-m|num-sg|pers-3|case-o|vib-0_se|tam-0|posn-90|vpos-vib_2|name-BUkaMpa|chunkId-NP3|chunkType-head:NP3</nowiki> | <nowiki>11</nowiki> | <nowiki>rh</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>10</nowiki> | <nowiki>से</nowiki> | <nowiki>से</nowiki> | <nowiki>PSP</nowiki> | <nowiki>psp</nowiki> | <nowiki>lex-se|cat-psp|gend-|num-|pers-|case-|vib-|tam-|posn-100|chunkType-child:NP3|name-se</nowiki> | <nowiki>9</nowiki> | <nowiki>lwg__psp</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>11</nowiki> | <nowiki>मची</nowiki> | <nowiki>मच</nowiki> | <nowiki>VM</nowiki> | <nowiki>v</nowiki> | <nowiki>lex-maca|cat-v|gend-f|num-sg|pers-any|case-|vib-yA|tam-yA|posn-110|name-macI|chunkId-VGNF2|chunkType-head:VGNF2</nowiki> | <nowiki>12</nowiki> | <nowiki>nmod__k1inv</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>12</nowiki> | <nowiki>तबाही</nowiki> | <nowiki>तबाही</nowiki> | <nowiki>NN</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-wabAhI|cat-n|gend-f|num-sg|pers-3|case-o|vib-0_kA_bAxa|tam-0|posn-120|vpos-vib_2_3|name-wabAhI|chunkId-NP4|chunkType-head:NP4</nowiki> | <nowiki>36</nowiki> | <nowiki>k7t</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>13</nowiki> | <nowiki>के</nowiki> | <nowiki>का</nowiki> | <nowiki>PSP</nowiki> | <nowiki>psp</nowiki> | <nowiki>lex-kA|cat-psp|gend-m|num-sg|pers-3|case-o|vib-|tam-|posn-130|chunkType-child:NP4|name-ke</nowiki> | <nowiki>12</nowiki> | <nowiki>lwg__psp</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>14</nowiki> | <nowiki>बाद</nowiki> | <nowiki>बाद</nowiki> | <nowiki>NST</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-bAxa|cat-n|gend-|num-|pers-|case-|vib-|tam-|posn-140|chunkType-child:NP4|name-bAxa</nowiki> | <nowiki>12</nowiki> | <nowiki>lwg__psp</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>15</nowiki> | <nowiki>भारत</nowiki> | <nowiki>भारत</nowiki> | <nowiki>NNP</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-BArawa|cat-n|gend-m|num-sg|pers-3|case-d|vib-0|tam-0|posn-150|name-BArawa|chunkId-NP5|chunkType-head:NP5</nowiki> | <nowiki>16</nowiki> | <nowiki>ccof</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>16</nowiki> | <nowiki>और</nowiki> | <nowiki>और</nowiki> | <nowiki>CC</nowiki> | <nowiki>avy</nowiki> | <nowiki>lex-Ora|cat-avy|gend-|num-|pers-|case-|vib-|tam-|posn-160|name-Ora|chunkId-CCP|chunkType-head:CCP</nowiki> | <nowiki>36</nowiki> | <nowiki>k1</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>17</nowiki> | <nowiki>पाकिस्तान</nowiki> | <nowiki>पाकिस्तान</nowiki> | <nowiki>NNP</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-pAkiswAna|cat-n|gend-m|num-sg|pers-3|case-d|vib-0|tam-0|posn-170|name-pAkiswAna2|chunkId-NP6|chunkType-head:NP6</nowiki> | <nowiki>16</nowiki> | <nowiki>ccof</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>18</nowiki> | <nowiki>मानवीय</nowiki> | <nowiki>मानवीय</nowiki> | <nowiki>JJ</nowiki> | <nowiki>adj</nowiki> | <nowiki>lex-mAnavIya|cat-adj|gend-any|num-any|pers-|case-o|vib-|tam-|posn-180|chunkType-child:NP7|name-mAnavIya</nowiki> | <nowiki>19</nowiki> | <nowiki>nmod__adj</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>19</nowiki> | <nowiki>दृष्टिकोण</nowiki> | <nowiki>दृष्टिकोण</nowiki> | <nowiki>NN</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-xqRtikoNa|cat-n|gend-m|num-sg|pers-3|case-d|vib-0|tam-0|posn-190|name-xqRtikoNa|chunkId-NP7|chunkType-head:NP7</nowiki> | <nowiki>20</nowiki> | <nowiki>k2</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>20</nowiki> | <nowiki>अपनाते</nowiki> | <nowiki>अपना</nowiki> | <nowiki>VM</nowiki> | <nowiki>v</nowiki> | <nowiki>lex-apanA|cat-v|gend-m|num-pl|pers-any|case-|vib-wA_ho+yA|tam-wA|posn-200|vpos-tam_2|name-apanAwe|chunkId-VGNF3|chunkType-head:VGNF3</nowiki> | <nowiki>36</nowiki> | <nowiki>vmod</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>21</nowiki> | <nowiki>हुए</nowiki> | <nowiki>हो</nowiki> | <nowiki>VAUX</nowiki> | <nowiki>v</nowiki> | <nowiki>lex-ho|cat-v|gend-m|num-pl|pers-any|case-|vib-yA|tam-yA|posn-210|chunkType-child:VGNF3|name-hue</nowiki> | <nowiki>20</nowiki> | <nowiki>lwg__vaux</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>22</nowiki> | <nowiki>शनिवार</nowiki> | <nowiki>शनिवार</nowiki> | <nowiki>NNP</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-SanivAra|cat-n|gend-m|num-sg|pers-3|case-o|vib-0_ko|tam-0|posn-220|vpos-vib_2|name-SanivAra|chunkId-NP8|chunkType-head:NP8</nowiki> | <nowiki>36</nowiki> | <nowiki>k7t</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>23</nowiki> | <nowiki>को</nowiki> | <nowiki>को</nowiki> | <nowiki>PSP</nowiki> | <nowiki>psp</nowiki> | <nowiki>lex-ko|cat-psp|gend-|num-|pers-|case-|vib-|tam-|posn-230|chunkType-child:NP8|name-ko2</nowiki> | <nowiki>22</nowiki> | <nowiki>lwg__psp</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>24</nowiki> | <nowiki>इस्लामाबाद</nowiki> | <nowiki>इसलामाबाद</nowiki> | <nowiki>NNP</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-isalAmAbAxa|cat-n|gend-m|num-sg|pers-3|case-d|vib-0_meM|tam-0|posn-240|vpos-vib_2|name-islAmAbAxa|chunkId-NP9|chunkType-head:NP9</nowiki> | <nowiki>36</nowiki> | <nowiki>k7p</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>25</nowiki> | <nowiki>में</nowiki> | <nowiki>में</nowiki> | <nowiki>PSP</nowiki> | <nowiki>psp</nowiki> | <nowiki>lex-meM|cat-psp|gend-|num-|pers-|case-|vib-|tam-|posn-250|chunkType-child:NP9|name-meM2</nowiki> | <nowiki>24</nowiki> | <nowiki>lwg__psp</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>26</nowiki> | <nowiki>नियंत्रण</nowiki> | <nowiki>नियंत्रण</nowiki> | <nowiki>XC</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-niyaMwraNa|cat-n|gend-m|num-sg|pers-3|case-d|vib-0|tam-0|posn-260|chunkType-child:NP10|name-niyaMwraNa</nowiki> | <nowiki>27</nowiki> | <nowiki>mod</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>27</nowiki> | <nowiki>रेखा</nowiki> | <nowiki>रेखा</nowiki> | <nowiki>NN</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-reKA|cat-n|gend-f|num-sg|pers-3|case-d|vib-0|tam-0|posn-270|name-reKA|chunkId-NP10|chunkType-head:NP10</nowiki> | <nowiki>31</nowiki> | <nowiki>k2</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>28</nowiki> | <nowiki>(</nowiki> | <nowiki>(</nowiki> | <nowiki>SYM</nowiki> | <nowiki>punc</nowiki> | <nowiki>lex-|cat-punc|gend-|num-|pers-|case-|vib-|tam-|posn-280|chunkType-child:NP11|name-(</nowiki> | <nowiki>29</nowiki> | <nowiki>rsym</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>29</nowiki> | <nowiki>एलओसी</nowiki> | <nowiki>एलओसी</nowiki> | <nowiki>NN</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-elaosI|cat-n|gend-m|num-sg|pers-3|case-d|vib-0|tam-0|posn-290|name-elaosI|chunkId-NP11|chunkType-head:NP11</nowiki> | <nowiki>27</nowiki> | <nowiki>nmod</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>30</nowiki> | <nowiki>)</nowiki> | <nowiki>)</nowiki> | <nowiki>SYM</nowiki> | <nowiki>punc</nowiki> | <nowiki>lex-|cat-punc|gend-|num-|pers-|case-|vib-|tam-|posn-300|chunkType-child:NP11|name-)</nowiki> | <nowiki>29</nowiki> | <nowiki>rsym</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>31</nowiki> | <nowiki>खोलने</nowiki> | <nowiki>खोल</nowiki> | <nowiki>VM</nowiki> | <nowiki>v</nowiki> | <nowiki>lex-Kola|cat-v|gend-any|num-sg|pers-any|case-o|vib-nA_kA|tam-nA|posn-310|vpos-tam_2|name-Kolane|chunkId-VGNN|chunkType-head:VGNN</nowiki> | <nowiki>33</nowiki> | <nowiki>r6</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>32</nowiki> | <nowiki>के</nowiki> | <nowiki>का</nowiki> | <nowiki>PSP</nowiki> | <nowiki>psp</nowiki> | <nowiki>lex-kA|cat-psp|gend-m|num-sg|pers-|case-o|vib-|tam-|posn-320|chunkType-child:VGNN|name-ke2</nowiki> | <nowiki>31</nowiki> | <nowiki>lwg__psp</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>33</nowiki> | <nowiki>मसले</nowiki> | <nowiki>मसला</nowiki> | <nowiki>NN</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-masalA|cat-n|gend-m|num-sg|pers-3|case-o|vib-0_para|tam-0|posn-330|vpos-vib_2|name-masale|chunkId-NP12|chunkType-head:NP12</nowiki> | <nowiki>36</nowiki> | <nowiki>k7</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>34</nowiki> | <nowiki>पर</nowiki> | <nowiki>पर</nowiki> | <nowiki>PSP</nowiki> | <nowiki>psp</nowiki> | <nowiki>lex-para|cat-psp|gend-|num-|pers-|case-|vib-|tam-|posn-340|chunkType-child:NP12|name-para</nowiki> | <nowiki>33</nowiki> | <nowiki>lwg__psp</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>35</nowiki> | <nowiki>बातचीत</nowiki> | <nowiki>बातचीत</nowiki> | <nowiki>NN</nowiki> | <nowiki>n</nowiki> | <nowiki>lex-bAwacIwa|cat-n|gend-f|num-sg|pers-3|case-d|vib-0|tam-0|posn-350|name-bAwacIwa|chunkId-NP13|chunkType-head:NP13</nowiki> | <nowiki>36</nowiki> | <nowiki>pof</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>36</nowiki> | <nowiki>करेंगे</nowiki> | <nowiki>कर</nowiki> | <nowiki>VM</nowiki> | <nowiki>v</nowiki> | <nowiki>lex-kara|cat-v|gend-m|num-pl|pers-3|case-|vib-gA|tam-gA|posn-360|name-kareMge|chunkId-VGF|chunkType-head:VGF</nowiki> | <nowiki>0</nowiki> | <nowiki>main</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| | <nowiki>37</nowiki> | <nowiki>.</nowiki> | <nowiki>.</nowiki> | <nowiki>SYM</nowiki> | <nowiki>punc</nowiki> | <nowiki>lex-.|cat-punc|gend-|num-|pers-|case-|vib-|tam-|posn-370|chunkType-child:VGF|name-.</nowiki> | <nowiki>36</nowiki> | <nowiki>rsym</nowiki> | <nowiki>_</nowiki> | <nowiki>_</nowiki> | |
| |
==== Parsing ==== | ==== Parsing ==== |