home edit page issue tracker

This page pertains to UD version 2.

Treebank Statistics: UD_English-GUM: POS Tags: PART

There are 3 PART lemmas (0%), 15 PART types (0%) and 6433 PART tokens (3%). Out of 17 observed tags, the rank of PART is: 17 in number of lemmas, 17 in number of types and 12 in number of tokens.

The 10 most frequent PART lemmas: to, not, ‘s

The 10 most frequent PART types: to, not, n’t, ‘s, ’s, n’t, na, ‘, ’, ta

The 10 most frequent ambiguous lemmas: to (PART 3667, ADP 2085, SCONJ 48, X 1)

The 10 most frequent ambiguous types: to (PART 3452, ADP 2060, SCONJ 48, DET 1, NUM 1, VERB 1, X 1), ’s (AUX 907, PART 537, VERB 108, PRON 48), ’s (PART 244, AUX 208, PRON 18, VERB 12), na (PART 137, INTJ 4), ’ (PUNCT 178, PART 57, NOUN 2), ’ (PART 37, PUNCT 26), ta (PART 11, ADP 4), s (PART 4, AUX 1, NOUN 1, VERB 1, X 1), a (DET 4448, ADP 4, PART 2, SCONJ 2, ADV 1, AUX 1, NOUN 1, PROPN 1), do (AUX 479, VERB 311, NOUN 3, PROPN 3, PART 1)

Morphology

The form / lemma ratio of PART is 5.000000 (the average of all parts of speech is 1.248450).

The 1st highest number of forms (6) was observed with the lemma “to”: a, do, na, ta, the, to.

The 2nd highest number of forms (5) was observed with the lemma “’s”: ’, ‘s, s, ’, ’s.

The 3rd highest number of forms (4) was observed with the lemma “not”: n’t, n`t, not, n’t.

PART occurs with 3 features: Polarity (1885; 29% instances), Style (11; 0% instances), Typo (3; 0% instances)

PART occurs with 3 feature-value pairs: Polarity=Neg, Style=Coll, Typo=Yes

PART occurs with 5 feature combinations. The most frequent feature combination is _ (4535 tokens). Examples: to, ‘s, ’s, na, ‘, ’, s, a

Relations

PART nodes are attached to their parents using 13 different relations: mark (3616; 56% instances), advmod (1841; 29% instances), case (881; 14% instances), xcomp (34; 1% instances), conj (20; 0% instances), root (13; 0% instances), acl (8; 0% instances), reparandum (7; 0% instances), advcl (4; 0% instances), ccomp (3; 0% instances), parataxis (3; 0% instances), orphan (2; 0% instances), nmod (1; 0% instances)

Parents of PART nodes belong to 15 different parts of speech: VERB (4730; 74% instances), NOUN (615; 10% instances), PROPN (523; 8% instances), ADJ (312; 5% instances), AUX (78; 1% instances), ADV (68; 1% instances), PRON (49; 1% instances), NUM (19; 0% instances), (13; 0% instances), DET (9; 0% instances), INTJ (5; 0% instances), ADP (4; 0% instances), SCONJ (3; 0% instances), X (3; 0% instances), PART (2; 0% instances)

6364 (99%) PART nodes are leaves.

45 (1%) PART nodes have one child.

9 (0%) PART nodes have two children.

15 (0%) PART nodes have three or more children.

The highest child degree of a PART node is 9.

Children of PART nodes are attached using 13 different relations: punct (52; 41% instances), cc (21; 16% instances), advmod (13; 10% instances), cop (12; 9% instances), nsubj (11; 9% instances), mark (6; 5% instances), discourse (5; 4% instances), obl (3; 2% instances), case (1; 1% instances), conj (1; 1% instances), det (1; 1% instances), nsubj:outer (1; 1% instances), parataxis (1; 1% instances)

Children of PART nodes belong to 13 different parts of speech: PUNCT (52; 41% instances), CCONJ (21; 16% instances), AUX (12; 9% instances), ADV (11; 9% instances), PRON (11; 9% instances), SCONJ (6; 5% instances), INTJ (5; 4% instances), DET (4; 3% instances), PART (2; 2% instances), ADJ (1; 1% instances), ADP (1; 1% instances), NOUN (1; 1% instances), VERB (1; 1% instances)