home edit page issue tracker

This page pertains to UD version 2.

Treebank Statistics: UD_Spanish-AnCora: Features: Gender

This feature is universal. It occurs with 2 different values: Fem, Masc.

202799 tokens (36%) have a non-empty value of Gender. 17312 types (45%) occur at least once with a non-empty value of Gender. 11621 lemmas (45%) occur at least once with a non-empty value of Gender. The feature is used with 8 part-of-speech tags: NOUN (88482; 16% instances), DET (78759; 14% instances), ADJ (24227; 4% instances), PRON (5804; 1% instances), VERB (4754; 1% instances), AUX (481; 0% instances), NUM (290; 0% instances), PROPN (2; 0% instances).

NOUN

88482 NOUN tokens (88% of all NOUN tokens) have a non-empty value of Gender.

The most frequent other feature values with which NOUN and Gender co-occurred: Number=Sing (61626; 70%).

NOUN tokens may have the following values of Gender:

Paradigm candidatoMascFem
Number=Singcandidato
Number=PlurcandidatosCANDIDATAS

Gender seems to be lexical feature of NOUN. 99% lemmas (7756) occur only with one value of Gender.

DET

78759 DET tokens (93% of all DET tokens) have a non-empty value of Gender.

The most frequent other feature values with which DET and Gender co-occurred: PronType=Art (71584; 91%), Number=Sing (62068; 79%), Definite=Def (62012; 79%).

DET tokens may have the following values of Gender:

Paradigm elMascFem
Definite=Def|ExtPos=ADV|Number=Sing|PronType=Artla
Definite=Def|ExtPos=SCONJ|Number=Sing|PronType=Artel
Definite=Def|Foreign=Yes|Number=Sing|PronType=Artla
Definite=Def|Foreign=Yes|Number=Plur|PronType=Artlesles
Definite=Def|Number=Sing|PronType=Artella
Definite=Def|Number=Plur|PronType=Artlos, elslas
Number=Sing|PronType=Demella
Number=Plur|PronType=Demloslas

ADJ

24227 ADJ tokens (67% of all ADJ tokens) have a non-empty value of Gender.

The most frequent other feature values with which ADJ and Gender co-occurred: VerbForm=EMPTY (17726; 73%), Number=Sing (17367; 72%).

ADJ tokens may have the following values of Gender:

Paradigm primeroMascFem
Number=Singprimer, primeroprimera
Number=Plurprimerosprimeras

PRON

5804 PRON tokens (23% of all PRON tokens) have a non-empty value of Gender.

The most frequent other feature values with which PRON and Gender co-occurred: Reflex=EMPTY (5803; 100%), Number=Sing (4355; 75%), Person=3 (3444; 59%), PronType=Prs (3344; 58%), PrepCase=EMPTY (3168; 55%).

PRON tokens may have the following values of Gender:

Paradigm élMascFem
Case=Acc,Nom|Number=Sing|PronType=Prsél, elloella
Case=Acc,Nom|Number=Plur|PronType=Prsellosellas
Case=Acc|Definite=Def|Number=Sing|PrepCase=Npr|PronType=Prslo
Case=Acc|Definite=Ind|Number=Sing|PrepCase=Npr|PronType=PrsLO
Case=Acc|ExtPos=ADV|Number=Sing|PrepCase=Npr|PronType=Prslo
Case=Acc|ExtPos=CCONJ|Number=Sing|PrepCase=Npr|PronType=Prslo
Case=Acc|Number=Sing|PrepCase=Npr|PronType=Demlo
Case=Acc|Number=Sing|PrepCase=Npr|PronType=Prslola
Case=Acc|Number=Plur|PrepCase=Npr|PronType=Prsloslas
Case=Nom|Number=Sing|PronType=PrsElla

VERB

4754 VERB tokens (10% of all VERB tokens) have a non-empty value of Gender.

The most frequent other feature values with which VERB and Gender co-occurred: Mood=EMPTY (4753; 100%), Person=EMPTY (4753; 100%), Tense=Past (4753; 100%), VerbForm=Part (4753; 100%), Number=Sing (4434; 93%).

VERB tokens may have the following values of Gender:

Paradigm hacerMascFem
Number=Singhechohecha
Number=Plurhechos

AUX

481 AUX tokens (4% of all AUX tokens) have a non-empty value of Gender.

The most frequent other feature values with which AUX and Gender co-occurred: Mood=EMPTY (481; 100%), Number=Sing (481; 100%), Person=EMPTY (481; 100%), Tense=Past (480; 100%), VerbForm=Part (480; 100%).

AUX tokens may have the following values of Gender:

NUM

290 NUM tokens (3% of all NUM tokens) have a non-empty value of Gender.

The most frequent other feature values with which NUM and Gender co-occurred: NumType=Card (290; 100%), NumForm=Word (289; 100%), Number=Plur (176; 61%).

NUM tokens may have the following values of Gender:

Paradigm ambosMascFem
ambosambas

PROPN

2 PROPN tokens (0% of all PROPN tokens) have a non-empty value of Gender.

PROPN tokens may have the following values of Gender:

Relations with Agreement in Gender

The 10 most frequent relations where parent and child node agree in Gender: NOUN –[det]–> DET (58013; 86%), NOUN –[amod]–> ADJ (16915; 63%), NOUN –[conj]–> NOUN (2526; 54%), NOUN –[appos]–> NOUN (928; 51%), ADJ –[det]–> DET (683; 64%), ADJ –[nsubj]–> NOUN (599; 57%), ADJ –[conj]–> ADJ (569; 55%), PRON –[nmod]–> NOUN (440; 74%), ADJ –[det]–> PRON (158; 62%), NOUN –[nmod]–> DET (154; 96%).