Publishing Partner: Cambridge University Press CUP Extra Publisher Login
amazon logo
More Info

New from Oxford University Press!


It's Been Said Before

By Orin Hargraves

It's Been Said Before "examines why certain phrases become clichés and why they should be avoided -- or why they still have life left in them."

New from Cambridge University Press!


Sounds Fascinating

By J. C. Wells

How do you pronounce biopic, synod, and Breughel? - and why? Do our cake and archaic sound the same? Where does the stress go in stalagmite? What's odd about the word epergne? As a finale, the author writes a letter to his 16-year-old self.

Academic Paper

Title: Extensive data for morphology: using the World Wide Web
Author: Nabil Hathout
Institution: CLLE-ERSS – Université de Toulouse - Le Mirail
Author: Fabio Montermini
Institution: CNRS
Author: Ludovic Tanguy
Institution: CLLE-ERSS – Université de Toulouse - Le Mirail
Linguistic Field: Morphology; Text/Corpus Linguistics
Subject Language: French
Abstract: This paper presents a number of recent studies in French morphology which make extensive use of data. These data relating to derived words have been automatically collected from digital corpora, mostly from the Web. The main point developed here is that this massive increase in the amount of available data can substantially modify the results of a morphological study, and can lead to new theoretical conclusions that would not have been possible with traditional data such as wordlists gathered from dictionaries. However, using the Web as a corpus brings up several technical and methodological questions, which are dealt with through examples and discussions about the different tools and techniques available. We exemplify our thesis through the study of the suffixal forms: -esque, -este, -able, -ment.


This article appears IN Journal of French Language Studies Vol. 18, Issue 1.

Return to TOC.

Add a new paper
Return to Academic Papers main page
Return to Directory of Linguists main page