+ All Categories
Home > Documents > Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse),...

Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse),...

Date post: 10-Jul-2020
Category:
Upload: others
View: 4 times
Download: 0 times
Share this document with a friend
80
Learning Syntactic Categories Informatics 1 CG: Lecture 10 Mirella Lapata School of Informatics University of Edinburgh [email protected] February 2, 2016 1 / 29
Transcript
Page 1: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Learning Syntactic CategoriesInformatics 1 CG: Lecture 10

Mirella Lapata

School of InformaticsUniversity of [email protected]

February 2, 2016

1 / 29

Page 2: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Reading:

Redington et al. (1998). Distributional Information: APowerful Cue for Acquiring Syntactic Categories.Cognitive Science 22, 425-469.

2 / 29

Page 3: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Recap: Word Learning

Word learning is hard, children use multiple sources of support:

socio-pragmatic skills

some aspects of child directed speech

biases towards certain interpretations over others

linguistic constraints through use of syntax

3 / 29

Page 4: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

How Do Children Learn Syntactic Categories?

One of most basic requirements of understanding language isidentifying the syntactic categories to which the words belong.

Is a word a noun, verb, adverb, or adjective?

How do children learn these categories and which wordsbelong to them?

Are categories hard-wired in the brain (rationalist view)?

Or are they learned (empiricist view)?

4 / 29

Page 5: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Open and Closed Classes in Natural Language

Several broad word classes are found in all Indo-Europeanlanguages and many others: nouns, verbs, adjectives, adverbs.

These are examples of open classes. They typically have largemembership, and are often stable under translation.

Other word classes are more specific to particular languages:prepositions (English, German), post-positions (Hungarian,Urdu, Korean), particles (Japanese), etc.

These are examples of closed classes. They typically havesmall, relatively fixed membership, and often have structuringuses in grammar. Little correlation between languages.

5 / 29

Page 6: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Parts of Speech

How do we tell what word class (part of speech) a word belongs to?

At least three different criteria can be used:

Semantic criteria: What does the word refer to?

Morphological criteria: What does the word look like?

Distributional (syntactic) criteria: Where is the word found?

We will look at different parts of speech (POS) using these criteria.

6 / 29

Page 7: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Nouns

Semantically, nouns generally refer to living things (mouse), places(Scotland), things (harpoon), or concepts (marriage).

Morphologically, -ness, -tion, -ity, and -ance tend to indicatenouns. (happiness, exertion, levity, significance).

Distributionally, we can examine the contexts where a nounappears and other words that appear in the same contexts.

like a Newfoundland dog just from the waterhe was seen swimming like a dog , throwing his long armssuch a deceitful dog ! It was only the lastwas mauled to death by her pet dog have described her as theirAdopting an adult dog can be a marvelous alternative

7 / 29

Page 8: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Nouns

Semantically, nouns generally refer to living things (mouse), places(Scotland), things (harpoon), or concepts (marriage).

Morphologically, -ness, -tion, -ity, and -ance tend to indicatenouns. (happiness, exertion, levity, significance).

Distributionally, we can examine the contexts where a nounappears and other words that appear in the same contexts.

like a Newfoundland dog just from the waterhe was seen swimming like a dog , throwing his long armssuch a deceitful dog ! It was only the lastwas mauled to death by her pet dog have described her as theirAdopting an adult dog can be a marvelous alternative

7 / 29

Page 9: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Verbs

Semantically, verbs refer to actions (observe, think, give).

Morphologically, words that end in -ate or -ize tend to be verbs,and ones that end in -ing are often the present participle of a verb(automate, calibrate, equalize, modernize; rising, washing,grooming).

Distributionally, we can examine the contexts where a verb appearsand other words that appear in the same contexts, which mayinclude their arguments.

Had he married a more amiable woman , he might havehe was very young when he married , and very fond of his wife .I am sure she will be married to Mr . Willoughby very soon .Biddy Henshawe ; she married a very wealthy man .I widowed that poor girl when I married her , Starbuck ;

8 / 29

Page 10: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Verbs

Semantically, verbs refer to actions (observe, think, give).

Morphologically, words that end in -ate or -ize tend to be verbs,and ones that end in -ing are often the present participle of a verb(automate, calibrate, equalize, modernize; rising, washing,grooming).

Distributionally, we can examine the contexts where a verb appearsand other words that appear in the same contexts, which mayinclude their arguments.

Had he married a more amiable woman , he might havehe was very young when he married , and very fond of his wife .I am sure she will be married to Mr . Willoughby very soon .Biddy Henshawe ; she married a very wealthy man .I widowed that poor girl when I married her , Starbuck ;

8 / 29

Page 11: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Adjectives

Semantically, adjectives convey properties of or opinions aboutthings that are nouns (small, wee, sensible, excellent).

Morphologically, words that end in -al, -ble, and -ous tend to beadjectives (formal, gradual, sensible, salubrious, parlous)

Distributionally, adjectives usually appear before a noun or after aform of be.

a great pity that such a sensible young man should be sosoaked through , it ’ s hard to be sensible , that ’ s a fact .She was sensible and clever ; but eager in everythingI should have been sensible of it at the time , for we alwaysHe was confused , seemed scarcely sensible of pleasure in seeing

9 / 29

Page 12: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Adjectives

Semantically, adjectives convey properties of or opinions aboutthings that are nouns (small, wee, sensible, excellent).

Morphologically, words that end in -al, -ble, and -ous tend to beadjectives (formal, gradual, sensible, salubrious, parlous)

Distributionally, adjectives usually appear before a noun or after aform of be.

a great pity that such a sensible young man should be sosoaked through , it ’ s hard to be sensible , that ’ s a fact .She was sensible and clever ; but eager in everythingI should have been sensible of it at the time , for we alwaysHe was confused , seemed scarcely sensible of pleasure in seeing

9 / 29

Page 13: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

The Problem of Learning Syntactic Categories

Difficult problem from both nativist and empiricist perspectives onlanguage acquisition.

Nativists: syntactic categories, are innate; learner must maplexicon of target language into these categories. There mustbe significant constraints on which mappings are considered.

Empiricists: finding correct mappings appears more difficultstill, since even the number of syntactic categories is notknown a priori.

On both views, learner must make the first steps in acquiringsyntactic categories without being able to apply constraintsfrom knowledge of the grammar.

10 / 29

Page 14: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

What Information is Available?

Distributional Information

Words of the same category have a large number of distributionalregularities in common, i.e., occur in similar linguistic contexts.

Semantic Bootstrapping

Abstract syntactic categories are innately specified, the learnermakes a tentative mapping from lexical items to these syntacticcategories, using semantic information (Pinker, 1984).

Phonological Constraints

There are regularities between the phonology of words and theirsyntactic categories which aid acquisition (stress, word duration).

Innate Knowledge

Learning mechanisms which exploit information in the input may beinnately specified and used to constrain search space of the learner.

11 / 29

Page 15: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Redington et al. (1998)

Distributional properties can be highly informative of syntacticcategory. This information can be extracted by psychologicallyplausible mechanisms:

1 Measuring distribution of contexts within which words occur.

2 Comparing the distributions of contexts for pairs of words.

3 Grouping together words with similar distributions of contexts.

12 / 29

Page 16: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Measuring Distribution for each Word

What should count as a context for a word?

. . . The field anthropologist must gain understanding and start withthe explanations and commentaries which his informants themselves offerabout their symbols. these must first be examined in the contexts inwhich they are usually employed, where they occur naturally, althoughsubsequent generalizing discussion helps the anthropologist to improve hisinitial understanding. to learn the meaning of symbols is part of theanthropologist’s practical semantics: to discover the meaning of words,noticing when their use is appropriate and when it is not. all this requiresimagination, patience, considerable linguistic skill, but above all a rigorous

respect for the facts. these must come first; fantasy can come later . . .

13 / 29

Page 17: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Measuring Distribution for each Word

What should count as a context for a word?

. . . The field anthropologist must gain understanding and start withthe explanations and commentaries which his informants themselves offer

about theirsymbols. these must first be examined in

the contexts inwhich they are usually employed, where they occur naturally, althoughsubsequent generalizing discussion helps the anthropologist to improve his

initialunderstanding. to learn the meaning of

symbols is part of the

anthropologist’spractical semantics: to discover the meaning of

words,noticing when their use is appropriate and when it is not. all this requiresimagination, patience, considerable linguistic skill, but above all a rigorous

respect for the facts.these must come first; fantasy can come later

. . .

13 / 29

Page 18: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Measuring Distribution for each Word

. . . The field anthropologist must gain understanding and start withthe explanations and commentaries which his informants themselves offer

about theirsymbols. these must first be examined in

the contexts inwhich they are usually employed, where they occur naturally, althoughsubsequent generalizing discussion helps the anthropologist to improve his

initialunderstanding. to learn the meaning of

symbols is part of the

anthropologist’spractical semantics: to discover the meaning of

words,noticing when their use is appropriate and when it is not. all this requiresimagination, patience, considerable linguistic skill, but above all a rigorous

respect for the facts.these must come first; fantasy can come later

. . .

these meaning to practical comefirst 2 0 0 0 2learn 0 1 1 0 0discover 0 1 1 0 1

14 / 29

Page 19: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Measuring Distribution for each Word

. . . The field anthropologist must gain understanding and start withthe explanations and commentaries which his informants themselves offer

about theirsymbols. these must first be examined in

the contexts inwhich they are usually employed, where they occur naturally, althoughsubsequent generalizing discussion helps the anthropologist to improve his

initialunderstanding. to learn the meaning of

symbols is part of the

anthropologist’spractical semantics: to discover the meaning of

words,noticing when their use is appropriate and when it is not. all this requiresimagination, patience, considerable linguistic skill, but above all a rigorous

respect for the facts.these must come first; fantasy can come later

. . .

these meaning to practical come

first

2 0 0 0 2

learn

0 1 1 0 0

discover

0 1 1 0 1

14 / 29

Page 20: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Measuring Distribution for each Word

. . . The field anthropologist must gain understanding and start withthe explanations and commentaries which his informants themselves offer

about theirsymbols. these must first be examined in

the contexts inwhich they are usually employed, where they occur naturally, althoughsubsequent generalizing discussion helps the anthropologist to improve his

initialunderstanding. to learn the meaning of

symbols is part of the

anthropologist’spractical semantics: to discover the meaning of

words,noticing when their use is appropriate and when it is not. all this requiresimagination, patience, considerable linguistic skill, but above all a rigorous

respect for the facts.these must come first; fantasy can come later

. . .

these meaning to practical comefirst

2 0 0 0 2

learn

0 1 1 0 0

discover

0 1 1 0 1

14 / 29

Page 21: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Measuring Distribution for each Word

. . . The field anthropologist must gain understanding and start withthe explanations and commentaries which his informants themselves offer

about theirsymbols. these must first be examined in

the contexts inwhich they are usually employed, where they occur naturally, althoughsubsequent generalizing discussion helps the anthropologist to improve his

initialunderstanding. to learn the meaning of

symbols is part of the

anthropologist’spractical semantics: to discover the meaning of

words,noticing when their use is appropriate and when it is not. all this requiresimagination, patience, considerable linguistic skill, but above all a rigorous

respect for the facts.these must come first; fantasy can come later

. . .

these meaning to practical comefirst 2 0 0 0 2learn 0 1 1 0 0discover 0 1 1 0 1

14 / 29

Page 22: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Measuring Distribution for each Word

Context words︷ ︸︸ ︷

these meaning to practical comefirst 2 0 0 0 2learn 0 1 1 0 0discover 0 1 1 0 1

︸ ︷︷ ︸ ︸ ︷︷ ︸Target words Context vectors

Words are represented by context vectors.

Redington et al. obtain such context vectors from CHILDES(a corpus of child directed speech, 2.5 million words).

An algorithm takes vectors as input and produces clusters.

Clusters correspond to parts of speech.

15 / 29

Page 23: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Measuring Distribution for each Word

Context words︷ ︸︸ ︷

these meaning to practical comefirst 2 0 0 0 2learn 0 1 1 0 0discover 0 1 1 0 1

︸ ︷︷ ︸

︸ ︷︷ ︸

Target words

Context vectors

Words are represented by context vectors.

Redington et al. obtain such context vectors from CHILDES(a corpus of child directed speech, 2.5 million words).

An algorithm takes vectors as input and produces clusters.

Clusters correspond to parts of speech.

15 / 29

Page 24: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Measuring Distribution for each Word

Context words︷ ︸︸ ︷these meaning to practical come

first 2 0 0 0 2learn 0 1 1 0 0discover 0 1 1 0 1

︸ ︷︷ ︸

︸ ︷︷ ︸

Target words

Context vectors

Words are represented by context vectors.

Redington et al. obtain such context vectors from CHILDES(a corpus of child directed speech, 2.5 million words).

An algorithm takes vectors as input and produces clusters.

Clusters correspond to parts of speech.

15 / 29

Page 25: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Measuring Distribution for each Word

Context words︷ ︸︸ ︷these meaning to practical come

first 2 0 0 0 2learn 0 1 1 0 0discover 0 1 1 0 1

︸ ︷︷ ︸ ︸ ︷︷ ︸Target words Context vectors

Words are represented by context vectors.

Redington et al. obtain such context vectors from CHILDES(a corpus of child directed speech, 2.5 million words).

An algorithm takes vectors as input and produces clusters.

Clusters correspond to parts of speech.

15 / 29

Page 26: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Measuring Distribution for each Word

Context words︷ ︸︸ ︷these meaning to practical come

first 2 0 0 0 2learn 0 1 1 0 0discover 0 1 1 0 1

︸ ︷︷ ︸ ︸ ︷︷ ︸Target words Context vectors

Words are represented by context vectors.

Redington et al. obtain such context vectors from CHILDES(a corpus of child directed speech, 2.5 million words).

An algorithm takes vectors as input and produces clusters.

Clusters correspond to parts of speech.

15 / 29

Page 27: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Words as Context Vectors

the

to

dog

badger

learn

16 / 29

Page 28: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Words as Context Vectors

the

to

dog

badger

learn

16 / 29

Page 29: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Words as Context Vectors

the

to

dog

badger

learn

16 / 29

Page 30: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Agglomerative Clustering

Learning Algorithm

1: Place each data point into its own singleton group2: Repeat: iteratively merge the two closest groups3: Until: all the data are merged into a single cluster

Algorithm measures how close two groups are according to adistance or similarity function.

Redington et al. use Spearman’s rank correlation

Many other choices are possible (e.g., cosine measure)

17 / 29

Page 31: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Agglomerative Clustering

Learning Algorithm

1. Place each data point into its own singleton group2. Repeat: iteratively merge the two closest groups3. Until: all the data are merged into a single cluster

The algorithm results in a sequence of groupings

It is up to the user to choose “natural” clustering sequence

Dendrogram: plot each merge at the similarity between twomerged groups

Provides interpretable visualization of algorithm and data

18 / 29

Page 32: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Group Similarity

Given a distance measure between points, the user has manychoices for how to define intergroup similarity.

Single-linkage: similarity of the closest pair

dSL(G , H) = mini∈G j∈H

dij

Complete-linkage: similarity of the furthest pair

dCL(G , H) = maxi∈G j∈H

dij

Group average: the average similarity between groups

dGA(G , H) =1

NG NH

∑i∈G

∑j∈H

dij

19 / 29

Page 33: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Group Similarity

20 / 29

Page 34: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Single Link Agglomerative Clustering: Example

A B C D EA 0 1 2 2 3B 1 0 2 4 3C 2 2 0 1 5D 2 4 1 0 3E 3 5 5 3 0

d k K

0 5 {A}, {B}, {C}, {D}, {E}

1 3 {A,B}, {C,D}, {E}2 2 {A,B,C,D}, {E}3 1 {A,B,C,D,E}

21 / 29

Page 35: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Single Link Agglomerative Clustering: Example

A B C D EA 0 1 2 2 3B 1 0 2 4 3C 2 2 0 1 5D 2 4 1 0 3E 3 5 5 3 0

d k K

0 5 {A}, {B}, {C}, {D}, {E}

1 3 {A,B}, {C,D}, {E}2 2 {A,B,C,D}, {E}3 1 {A,B,C,D,E}

d({A, B}) = 1, d({A, C}) = 2, d({A, D}) = 2, d({A, D}) = 3d({B, C}) = 2, d({B, D}) = 4, d({B, E}) = 5d({C , D}) = 1, d({C , E}) = 5d({D, E}) = 3

21 / 29

Page 36: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Single Link Agglomerative Clustering: Example

A B C D EA 0 1 2 2 3B 1 0 2 4 3C 2 2 0 1 5D 2 4 1 0 3E 3 5 5 3 0

d k K

0 5 {A}, {B}, {C}, {D}, {E}

1 3 {A,B}, {C,D}, {E}2 2 {A,B,C,D}, {E}3 1 {A,B,C,D,E}

d({A, B}) = 1, d({A, C}) = 2, d({A, D}) = 2, d({A, D}) = 3d({B, C}) = 2, d({B, D}) = 4, d({B, E}) = 5d({C , D}) = 1, d({C , E}) = 5d({D, E}) = 3

21 / 29

Page 37: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Single Link Agglomerative Clustering: Example

A B C D EA 0 1 2 2 3B 1 0 2 4 3C 2 2 0 1 5D 2 4 1 0 3E 3 5 5 3 0

d k K

0 5 {A}, {B}, {C}, {D}, {E}1 3 {A,B}, {C,D}, {E}

2 2 {A,B,C,D}, {E}3 1 {A,B,C,D,E}

d({A, B}) = 1, d({A, C}) = 2, d({A, D}) = 2, d({A, D}) = 3d({B, C}) = 2, d({B, D}) = 4, d({B, E}) = 5d({C , D}) = 1, d({C , E}) = 5d({D, E}) = 3

21 / 29

Page 38: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Single Link Agglomerative Clustering: Example

A B C D EA 0 1 2 2 3B 1 0 2 4 3C 2 2 0 1 5D 2 4 1 0 3E 3 5 5 3 0

d k K

0 5 {A}, {B}, {C}, {D}, {E}1 3 {A,B}, {C,D}, {E}

2 2 {A,B,C,D}, {E}3 1 {A,B,C,D,E}

21 / 29

Page 39: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Single Link Agglomerative Clustering: Example

A B C D EA 0 1 2 2 3B 1 0 2 4 3C 2 2 0 1 5D 2 4 1 0 3E 3 5 5 3 0

d k K

0 5 {A}, {B}, {C}, {D}, {E}1 3 {A,B}, {C,D}, {E}

2 2 {A,B,C,D}, {E}3 1 {A,B,C,D,E}

d({A, B}, {C , D}) = min{d(A, C ), d(A, D), d(B, C ), d(B, D)}= min{2, 3, 2, 4}= 2

d({A, B}, {E}) = min{d(A, E ), d(B, E )}= min{3, 5}= 3

d({C , D}, {E}) = min{d(C , E ), d(D, E )}= min{5, 3}= 3

21 / 29

Page 40: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Single Link Agglomerative Clustering: Example

A B C D EA 0 1 2 2 3B 1 0 2 4 3C 2 2 0 1 5D 2 4 1 0 3E 3 5 5 3 0

d k K

0 5 {A}, {B}, {C}, {D}, {E}1 3 {A,B}, {C,D}, {E}

2 2 {A,B,C,D}, {E}3 1 {A,B,C,D,E}

d({A, B}, {C , D}) = min{d(A, C ), d(A, D), d(B, C ), d(B, D)}= min{2, 3, 2, 4}= 2

d({A, B}, {E}) = min{d(A, E ), d(B, E )}= min{3, 5}= 3

d({C , D}, {E}) = min{d(C , E ), d(D, E )}= min{5, 3}= 3

21 / 29

Page 41: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Single Link Agglomerative Clustering: Example

A B C D EA 0 1 2 2 3B 1 0 2 4 3C 2 2 0 1 5D 2 4 1 0 3E 3 5 5 3 0

d k K

0 5 {A}, {B}, {C}, {D}, {E}1 3 {A,B}, {C,D}, {E}2 2 {A,B,C,D}, {E}

3 1 {A,B,C,D,E}

d({A, B}, {C , D}) = min{d(A, C ), d(A, D), d(B, C ), d(B, D)}= min{2, 3, 2, 4}= 2

d({A, B}, {E}) = min{d(A, E ), d(B, E )}= min{3, 5}= 3

d({C , D}, {E}) = min{d(C , E ), d(D, E )}= min{5, 3}= 3

21 / 29

Page 42: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Single Link Agglomerative Clustering: Example

A B C D EA 0 1 2 2 3B 1 0 2 4 3C 2 2 0 1 5D 2 4 1 0 3E 3 5 5 3 0

d k K

0 5 {A}, {B}, {C}, {D}, {E}1 3 {A,B}, {C,D}, {E}2 2 {A,B,C,D}, {E}

3 1 {A,B,C,D,E}

d({A, B, C , D}, {E}) = min{d(A, E ), d(B, E ), d(C , E ), d(D, E )}= min{3, 5, 5, 3}= 3

21 / 29

Page 43: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Single Link Agglomerative Clustering: Example

A B C D EA 0 1 2 2 3B 1 0 2 4 3C 2 2 0 1 5D 2 4 1 0 3E 3 5 5 3 0

d k K

0 5 {A}, {B}, {C}, {D}, {E}1 3 {A,B}, {C,D}, {E}2 2 {A,B,C,D}, {E}3 1 {A,B,C,D,E}

21 / 29

Page 44: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Dendrogram

A B C D E

22 / 29

Page 45: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Dendrogram

A B C D E

22 / 29

Page 46: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Dendrogram

A B C D E

22 / 29

Page 47: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Dendrogram

A B C D E

Hei

ght

23 / 29

Page 48: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Dendrogram

A B C D E

Hei

ght

23 / 29

Page 49: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Dendrogram

A B C D E

Hei

ght

23 / 29

Page 50: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Example

0 20 40 60 80

−20

020

4060

80Data

24 / 29

Page 51: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Example

0 20 40 60 80

−20

020

4060

80iteration 001

V1

V2

24 / 29

Page 52: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Example

0 20 40 60 80

−20

020

4060

80iteration 002

V1

V2

24 / 29

Page 53: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Example

0 20 40 60 80

−20

020

4060

80iteration 003

V1

V2

24 / 29

Page 54: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Example

0 20 40 60 80

−20

020

4060

80iteration 003

V1

V2

24 / 29

Page 55: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Example

0 20 40 60 80

−20

020

4060

80iteration 004

V1

V2

24 / 29

Page 56: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Example

0 20 40 60 80

−20

020

4060

80iteration 005

V1

V2

24 / 29

Page 57: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Example

0 20 40 60 80

−20

020

4060

80iteration 006

V1

V2

24 / 29

Page 58: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Example

0 20 40 60 80

−20

020

4060

80iteration 007

V1

V2

24 / 29

Page 59: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Example

0 20 40 60 80

−20

020

4060

80iteration 008

V1

V2

24 / 29

Page 60: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Example

0 20 40 60 80

−20

020

4060

80iteration 009

V1

V2

24 / 29

Page 61: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Example

0 20 40 60 80

−20

020

4060

80iteration 010

V1

V2

24 / 29

Page 62: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Example

0 20 40 60 80

−20

020

4060

80iteration 011

V1

V2

24 / 29

Page 63: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Example

0 20 40 60 80

−20

020

4060

80iteration 012

V1

V2

24 / 29

Page 64: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Example

0 20 40 60 80

−20

020

4060

80iteration 013

V1

V2

24 / 29

Page 65: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Example

0 20 40 60 80

−20

020

4060

80iteration 014

V1

V2

24 / 29

Page 66: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Example

0 20 40 60 80

−20

020

4060

80iteration 016

V1

V2

24 / 29

Page 67: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Example

0 20 40 60 80

−20

020

4060

80iteration 015

V1

V2

24 / 29

Page 68: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Example

0 20 40 60 80

−20

020

4060

80iteration 017

V1

V2

24 / 29

Page 69: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Example

0 20 40 60 80

−20

020

4060

80iteration 018

V1

V2

24 / 29

Page 70: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Example

0 20 40 60 80

−20

020

4060

80iteration 019

V1

V2

24 / 29

Page 71: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Example

0 20 40 60 80

−20

020

4060

80iteration 020

V1

V2

24 / 29

Page 72: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Example

0 20 40 60 80

−20

020

4060

80iteration 021

V1

V2

24 / 29

Page 73: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Example

0 20 40 60 80

−20

020

4060

80iteration 022

V1

V2

24 / 29

Page 74: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Example

0 20 40 60 80

−20

020

4060

80iteration 023

V1

V2

24 / 29

Page 75: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Example

0 20 40 60 80

−20

020

4060

80iteration 024

V1

V2

24 / 29

Page 76: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Clusters from Redington et al.

25 / 29

Page 77: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Adjectives Cluster

26 / 29

Page 78: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Present Participles Cluster

27 / 29

Page 79: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Redington et al.’s Results

The model uses highly local distributional information which isconsistent with early vocabulary development

It is most effective for learning nouns, then verbs, and leasteffective for function words, mirroring children’s syntacticdevelopment

The method learns using the input corpora of the order ofmagnitude received by the child

The success of this model suggests that distributionalinformation may make an important contribution to earlylanguage development.

28 / 29

Page 80: Learning Syntactic Categories · Semantically, nouns generally refer to living things (mouse), places (Scotland), things (harpoon), or concepts (marriage). Morphologically, -ness,

Summary

Discussed the problem of learning syntactic categories.

Model of how children may use distributional information inacquiring syntactic categories.

Using agglomerative clustering on CHILDES corpus

Distributional information is a potentially powerful cue forlearning syntactic categories and language in general.

General approach uses computationally explicit model ofspecific aspects of language acquisition.

Remaining questions:

Does proposed method apply to languages other than Englishwithout strong word order constraints?

How about integrating other sources of distributionalinformation (e.g., morphological or phonological cues)?

Induced syntactic categories are not ambiguous (frank wordsvs frank a stamp).

29 / 29


Recommended