Tuesday, May 1, 2012

Michael Halliday @ Connecting Paths, Nov.2010

Thanks again to Annabelle Lukin for posting this video to the Vimeo SFL Linguists Group. In this talk, Michael Halliday draws various links from SFL to Sydney Lamb's linguistic theories, including stratificational linguistics, and points out that the differences between their approaches are mostly a result of their different  research areas rather than a difference in their views on language.
Other talks from the same Connecting Paths Symposium are also available from Annabelle's vimeo page, including the talks by Ruqaiya HasanJim Martin and Sydney Lamb. Congratulations to City University of Hong Kong and to Sun Yat-sen University, Guangzhou for hosting such an interesting forum and for making it accessible to the rest of us.

Monday, April 30, 2012

Lessons Worth Sharing - TedEd

Those lovely people at TED have come up with a new idea they call TedEd. Just like the RSAnimate videos by CognitiveMedia (see the Ian McGilchrist post), they want to combine great talks with great pictures on a video, but they know they can't stop there. Each video is accompanied by quick quiz questions, open ended questions and resources.

Watch the Video - The folks at TEDEd explain it very clearly!

Again, though that is not enough. You can also "Flip" the video - which means you can customise the 'package' to suit your class. You can edit the title, the quiz and "think deeper" questions and the resources. And, but, that's not all. You can then "flip" any YouTube video and make it into a TedEd format video lesson.
So here's my flipping attempt - a short video from YouTube on the history of writing with some questions, some resources and a research project. Here's another flipping attempt. This one uses a video about Everett's claims about Pirahã.

Wednesday, April 18, 2012

Christian Matthiessen @ Register & Context 2012

In this talk from the Register and Context Symposium 2012, Christian Matthiessen offers, among other things, his 'wheel of register' which relates the field mode and tenor of different written and spoken genres. The link to the symposium also offers a wide range of resources and links to other videos.

Part 1 of the talk (Video 1 is an introduction by Annabelle)
Part 2 of the talk
Always a pleasure to hear you talk, Christian, and thanks so much to Annabelle Lukin for organising the symposium and for making the talks available at the time and through the SFL group on Vimeo. Also available from the Vimeo site are videos featuring Ruqaiya Hasan and there's more to come. If you are interested in these and similar videos, Annabelle requests that you  join the group so we can see how much interest there is.

Monday, April 9, 2012

Fry's Planet Word on DVD - A Brief Review

As with any 'popularisation' of a subject, academics can easily take a swipe at mass media explanations of their subject. In this review I will try to avoid taking cheap shots (unless the temptation is too great) and attempt to keep an open mind on how well Fry has done in representing the subject of linguistics to the general public as, I believe, this was his intention. I will review each episode and then finish with an overview.

Episode 1 - Babel
...And he's off: in no time at all, we are sent from Stephen Fry's comfy documentary-world study to Kenya to meet the Turkana, back to the London suburbs to meet a typical toddler, Ruby, and off to Leipzig (twice). Before you can say hello in 25 languages, we are already pondering a wide range of linguistic dilemmas. We also catch glimpses of Nim Chimpsky and the famous chattering YouTube twins. To help us out, we visit a range of experts. If I could have anyone in the world to talk about state of the art theories of language development, I would have one person at the top of my wishlist: Michael Tomasello. He adds much-needed balance both to the academic study of language development and the programme itself. Then, very quickly we are back on the slippery search for the language gene, aka FOXP2, and only just understand that this really cannot be the whole answer.
While we do not learn very much about Stephen Fry's brain scans in an fMRI, we do learn that the people he chose from UCL have a very balanced, realistic view of what they are able to achieve with these tools. I suppose if you are going to prepare a documentary on language, you have to include Stephen Pinker, if only because more Joe Publics have read his books than any others on language. Thankfully, Pinker does not get it all his own way. At the end of the episode, we have been given a fairly good overall picture of language development and been introduced to  issues of language versus animal communication, language proliferation, decay and death, and the long, long way we still have to go to even start to understand language. All the time, no matter what you may think of the presenter, Fry clearly enjoys language and relishes the challenge of circumscribing the subject. As an introduction of language study to the completely uninitiated, this is a good start.


Episode 2 - Identity
This episode deals with the typically sociolinguistic topics of accent, language decay and identity.
We start with an investigation of the myriad accents of Yorkshire, guided by a poet from Barnsley, we give Fry a few moments to exhibit his control of accents on an 'accent forecast of the UK' made to resemble a BBC weather forecast, and then we land in Newcastle, where we hear 'chirpy' Geordie call centre operatives and their PR manager. Then we're whisked over to Connemara where we hear some Irish (or Irish Gaelic if you prefer) and find out how the young and older feel about their language which was brought back from the brink of extinction. At this point we do touch on the serious issue of language decay, identity and "linguicide", with a cheap swipe at L'Académie française for being so imperial for so long. We also look at the re-birth of Hebrew, where we at last meet a linguist (only the 2nd in this programme) whose thesis is that Hebrew still retains large parts of Yiddish. (I do not know if Fry is Jewish, but in this episode he goes out of his way to be nice to them in London, New York and Jerusalem.) Finally we compare how Irish, Breton, Basque, Hebrew, Oc and Turkana resist the threat of Globish (that's global English).
One other point of interest in the programme is the debate around how far your culture affects your language, and vice versa, with Stanford Russian linguist Lera Boroditsky discussing how masculine and feminine nouns in gendered languages affect the way that some speakers describe objects that carry different genders in different languages. Fry later admits to supporting the Chomskian line, that all languages are ultimately similar and so, after allowing such a poor misrepresentation, does a double disservice to the so-called Sapir-Whorif hypothesis.
After the Frying start of episode 1, episode 2 is very disappointing. I do not think that this is due to my personal lack of interest in the issue of Identity (which I think covers a multitude of academic sins), but because the head count of experts - famous or otherwise - is much lower in this episode, and Fry's inexhaustible enthusiasm is an insufficient replacement for real facts.


Episode 3 - Uses and Abuses
The primary aim of this episode appears to be to cram in as many words that are normally banned on the BBC as possible. It does quite a good job with copious fucks, plenty of bollocks and a smattering of cunts. In terms of academic head counts, this episode does a much better job than number 2, except Stephen Pinker crops up a number of times spouting off on subjects that he really has no expertise in - a role in which he has become quite an expert!
Although  Fry's approach could easily be dismissed by people working in the fields of sociolinguistics and humor studies (which he refers to as rather humorless), we must never forget that this is a television programme.
An experiment involving the actor Brian Blessed and a large tub of icy water is particularly unscientific, but it makes good television and is slightly related to more serious research. Admittedly this episode does descend into a promo-video for Fry's favourite issues (Judaism, homosexuality, racism etc.), but generally it also makes a point; although Fry & co. smatter their speech with expletives, none of them can bring themselves to say nigger in any form other than "the n-word." So, even for people who can cuss and swear willy-nilly, there are still some taboo words. Stephen K. Amos manages, just, but explains that it still retains an insulting meaning for him due to personal experiences. Hardly scientific. An improvement on episode 2, but still not as accurate or rigorous as episode 1.


Episode 4 - Spreading the Word
Episode 4 is all about writing, and is of a much higher standard than the previous 2 episodes. Although I would not agree with all that we see in this episode, I would say that Fry and his team have done a better job of researching the key issues in writing. We have a wide range of suitable experts from typesetters in Norwich to the inventor of Pinyin in Beijing.
When Fry wants to learn about the origins of writing he finds an expert in cuneiform in the British museum, who shows him how it is done, and even poses in front of THE Rosetta Stone (best not to ask why it's in the British Museum, though!!). He is back in Jerusalem to look at, be told off for touching, and witness digital imaging techniques for restoring the "Dead Sea Scrolls."
He also manages to trace the connection between printing, the Age of Reason and wikipedia - not bad for a TV show! He visits the Bodleian library at Oxford University to examine how they are keeping abreast of the digital age and in Harvard's library meets someone who points out that new media do not need to replace previous formats - that the iPad / Kindle etc. are not likely to replace the book, but both will develop alongside each other just as radio did not replace newspapers and was not replaced by TV.
As with other episodes Fry adds his own pet theories, likes and dislikes but he also places developments such as printing in a social context, providing a good balance of enthusiasm and restraint on a subject that easily leads to hyperbole. I also support his call to support the libraries of the U.K. and the world, no matter what formats are being preserved - buildings dedicated to the pursuit of knowledge through reading are the bedrock of civilizations!


Episode 5 - The Power and The Glory
So in the ultimate episode we learn that the ultimate purpose of language is...
literature. 
Fry explains why he loves a range of writers, from Joyce to Wodehouse to Orwell, heaping the greatest praise on Shakespeare - he even manages to find a French actor to admit that he would rather play Shakespeare than any of the lesser French playwrights. Fry just could not resist one last jab at the French before the series finishes. Some old friends are back, such as Brian Blessed and the Turkana villagers, as well as some new faces, including David Tenant giving us some of his take on Hamlet.
The episode is just as busy and full of locations, interviews and Fry's opinions as the others, but offers no new information on linguistics or even language studies. For this reason it is the most disappointing episode - at least episode 2 was related to aspects of sociolinguistics and the hot topic of identity. All of the science disappears and we left with the absolute relativism of everyone's opinion is just as good as each other, which is clearly not the case. Just ask a Cambridge Don!


Episodes 1-5 - Planet Word

All in all, Planet Word is very uneven. It manages to combine wit and fact, controversy and error. At times, it is highly perceptive and at others completely misleading. To be fair, this is TV. It is not intended to be lectures 1-5 in a course in linguistics. Evaluating the series from an academic perspective is completely unfair. Whatever Fry offers, it must work well on the screen - hence the frequent scene changes (often for no reason), cuts to the fake study for a talk-to-camera and a heavy dependence on interviews with experts (used in the loosest sense where Stephen Pinker is concerned).
What we need most from this series, perhaps, is for the general public to gain some understanding of language studies or even be inspired to look further into the subject, especially if they are young and are considering what to study at university. I believe Fry has succeeded to some extent in providing a TV series that engages with its audience, entertains and informs. A wide range of linguistic issues, perspectives and facts are offered with a minimum of effort on the part of the viewer - no mean feat. Only the last episode could be considered misleading. Certainly I do not agree with a lot of what he claims throughout the series, but this is his show not mine!
Find yourself a copy of the DVD, or even pick up the book, and see if you could find better ways to make linguistics appeal to more people who have never considered studying language before. It will not improve the programme but it may just help you, if you are a lover of language, to explain your interest to others. Spread the word - Planet Word.

Monday, April 2, 2012

Vygotskian Psychology - A Documentary

This is a link to a documentary about Vygotsky's psychological theories in the Soviet Union according to the people who were there at the time. A lot of the talk is about Vygotsky and Leontev as people and academics, but there is also a fair amount of discussion about their theory of psychology and its relation to Marxism. The film lasts 1 hour 45 minutes and was made by Mark Rozin in 1992.
This is a very rare film, and may lack any sign of pizzaz, but the people interviewed and the ideas discussed are worth the time and concentration.
Acknowledgement:  Thanks to Phil Chappell for posting the link to sys-func

New App - BOOK

In true over-the-top-hype, keep-it-clean, condescending Apple style, this short video introduces you to....
   
...the BOOK.
And this one re-imagines the IT helpdesk for the new invention of the book, when it was a new invention. Here we see the equivalent of "have you tried turning it on and off again?"

In both cases, we are shown up to be the IT-lap-it-up dogs that we have become. Just because Steve Jobs breathed on it at one time our awe is supposed to be inspired. Get a grip, peoples!!!

Tuesday, March 13, 2012

Matt Creates RogersCreations

Stephen Fry pronounces persistently on his pet peeve - the pedants of proper language - while his words weave their witty way across the screen to converge into the single word 'language.'
Matt Rogers of RogersCreations put this together using some basic Adobe applications (so once again iphone & ipad users are screwed). You can also download a static pdf version of the final product. Although the visualisation is not as sophisticated as other posts featured (see the last post or click on 'visualisation'), it seems to work and keeps the viewer focussed on the message, rather than detracting from it. Another example he offers was prepared for Al Jazeera and uses a short extract of an interview with the great Youssou N'Dour: Youssou N'Dour - Frames - Al Jazeera English. This one has a much wider range of backgrounds and 'tricks.'

Sunday, March 11, 2012

Iain McGilchrist - The Divided Brain

This video provides an up-to-date discussion of what happens in the hemispheres of the brain. It quickly rejects the old emotional-rational, language-mathematical & other divisions, but explains just what happens in the two halves, and why.
(If the video fails, use this link)
This is another RSAanimate video produced by the excellent Cognitive Media team, who also produce A0 size digital files and prints of the final version of the talk. There are many other similar videos, but for content and visuals this is my favourite so far. You can also download an iphone* or android app from RSA that grants you access to the animated talks, the original lectures of these and more talks, and much more.

* Not that you would know what this post is about on your iphone because it does not support flash video.

Monday, March 5, 2012

READ 4 - EAFL Special Edition

We are very, very excited to announce the publication of READ Magazine Special Edition. This edition is produced in conjunction with the Emirates Airlines Festival of Literature, being held 6-10 March in Dubai, especially in and around Dubai Festival City. It features a wide range of authors from the festival, as well as Isobel Aboulhoul, O.B.E., the festival organiser, and a Reading Champions (one of our regular features) by Peter James.

In general, all the featured authors graciously answered a range of questions covering the following areas: Reading Preferences, Influences on Your Writing, Reading Habits, Raising the Next Generation of Readers, and Reading with Others. Our authors are up-and-coming and established, young and old, writers of fact and of fiction, from the region and from other countries. Here's the editorial:

READ Magazine aims to create a culture of reading in the UAE and beyond through inspirational stories and practices.  We are therefore delighted to be associated with the Emirates Airline Festival of Literature.  This event has become the Middle East’s largest celebration of the written and spoken word, bringing people together with authors from across the world to promote education, debate and, above all else, reading.

We are very appreciative that a number of authors have contributed to this Special Edition of Read Magazine by sharing their reading habits and childhood memories of  what reading was like at home.  We have not been surprised by the culture of reading that shaped their lives nor by the wealth of knowledge guiding them on their literary paths.  We asked writers to reflect on the future of books, the impact of technology and how we should engage present and future generations in reading.  We also asked them to explore how we should encourage parents and young readers to embrace literature at home.  We hope some of the insights and thoughts from a few of the many authors participating in this year’s event will inspire your visit to the Emirates Airline Festival of Literature.

We are grateful to our sponsors for their generous support and sincerely thank the festival and the authors for their generous contribution.

We intend to distribute copies at the festival to authors, friends of the festival and other interested parties, but will also have some left over to distribute and inspire readers, writers and teachers.

Clark & Lappin - Linguistic Nativism and the Poverty of the Stimulus

Another post (back in November) notified the publication of "Linguistic Nativism and the Poverty of Stimulus" by Alex Clark & Shalom Lappin. This post is a review of the book. Below is the extended version, while this link sends you to the edited version on linguistlist.


AUTHORS: Clark, Alexander and Lappin, Shalom
TITLE: Linguistic Nativism and the Poverty of the Stimulus 
PUBLISHER: Wiley-Blackwell
YEAR: 2011

Nick Moore, Khalifa University, United Arab Emirates

In 12 chapters, with front material, contents, a preface, reference, an author index and a subject index, Alexander Clark and Shlalom Lappin’s “Linguisic Nativism and the Poverty of the Stimulus” tackles a key issue in linguistics over 260 pages. The book is intended for a general linguistics audience, but the reader needs some familiarity with basic concepts in formal linguistics, at least an elementary understanding of computational linguistics, and enough statistical, programming or mathematical knowledge not to shy away from algorithms. It is written, however, for a wide range of undergraduate, graduate and practicing linguists, particularly researchers working in formal grammar, computational linguistics and linguistic theory.

The main aim of the book is to replace the view that humans have an innate bias towards learning language that is specific to language with the view that the innate bias towards language acquisition depends on abilities that are used in other domains of learning. The first view is characterised as the argument for a strong bias, or linguistic nativism, while the second view is characterised as a weak bias or domain-general view. The principle line of argument is that computational, statistical and machine-learning methods demonstrate superior success in modeling, describing and explaining language acquisition, especially when compared to studies from formal linguistic models based on strong bias arguments, typically from Universal Grammar, Principles and Parameters, the Minimalist Program and other theories inspired by the work of Noam Chomsky.


Summary

Chapter 1, the introduction, establishes the boundaries of the discussion for the book. The authors focus throughout on a viable computationally-explicit model of language acquisition. While briefly presenting arguments on the evolutionary and biological nature of linguistic nativism, they rarely consider these questions again in this book. Their first of many criticisms of Universal Grammar (UG) theories proposed and inspired by Noam Chomsky is the Minimalist Program appears to disregard the puzzle of acquisition, unlike earlier versions of the theory which placed innate language-specific learning at the core of the theory, in order to explain how children consistently learn language from apparently inadequate data. Clark and Lappin (hereafter C&L) do not dismiss nativism entirely. They point out that it is fairly uncontroversial, first, that humans alone acquire language, and, second, that the environment plays a significant role in determining the language and the level of acquisition. What they intend to establish in the book, however, is that any innate cognitive faculties employed in the acquisition of language are not specific to language, as suggested by Chomsky and the Universal Grammar (UG) framework, but are general cognitive abilities that are employed in other learning tasks. It is from this ‘weak bias’ angle that they critique the Argument from the Poverty of Stimulus (APS).

Chapter 2 focuses on the Argument from the Poverty of Stimulus (APS). The APS is considered central to the nativist position because it provides an explanation for Chomsky’s core assumptions for UG: (1) grammar is rich and abstract; (2) data-driven learning is inadequate; (3) the linguistic data that the child is exposed to is qualitatively and quantitatively degenerate; and (4) the acquired grammar for a given language is uniform despite variations in intelligence. Most of these assumptions are dealt with throughout the book, and many replaced with assumptions from a machine-learning, ‘weak bias’ perspective, although little evidence is supplied to counter assumption three; rather, the reader is referred to other sources. Some APS theories have attacked connectionist approaches to the language acquisition puzzle by proposing that frequency alone would predict a radically different learning order than that observed. C&L counter that an argument against a connectionist claim does not prove the UG assumption of linguistic nativism correct – the strong bias towards language-specific learning mechanisms remains unproven – and they contend that the puzzle of language acquisition can be more explicitly delineated by computational methods than the proposals so far provided by the various UG frameworks. Reviewing the UG evidence for a strong bias, C&L claim that many arguments become self-serving: “a particular grammatical representation is not motivated by the APS, but rather it becomes an argument for the APS.” (p.33) Two examples of the APS – auxiliary inversion and pronominal ‘one’ – are provided as cases in point. In this chapter C&L introduce machine-learning alternatives to linguistic nativism based on distributional, or relative frequency, criteria and a Bayesian statistical model to account for these same learning phenomenon. These alternatives are elucidated further in the book.

Chapter 3 examines the extent to which the stimulus (accurately, the Primary Linguistic Data) really is impoverished. C&L avoid arguments between the more empirically-demonstrated richer linguistic environment and the naively-assumed impoverished input (the reader is referred to MacWhinney and others), but examine a key question for the computational modelling of the learning process: the existence or prevalence of negative evidence in the learning process. Even a small amount of negative evidence, for instance through reformulation, can make a significant difference to the problem of learning, and C&L allow for far less negative evidence than has been demonstrated in a range of corpora of child-directed speech. Other indirect negative evidence, such as the non-existence of hypothesised structures, can also significantly alter the learning task assuming that the learner is free to make these hypotheses. C&L challenge the premise of no negative evidence, partly because it forms such a central tenet of Gold’s “Identification In the Limit” – a theory of language acquisition that provides considerable support for the UG position of linguistic nativism. Gold’s highly-influential study argues that because learning cannot succeed within the cognitive, developmental and time limits imposed, then children must have prior knowledge of the language system. For instance, Gold claims that since the Primary Linguistic Data does not contain structural information, such as tagging for correctness, the knowledge that children rapidly acquire (e.g. knowing if a string is grammatical) can only be explained by strong linguistic nativism. However, C&L point out that this view is partly a consequence of ignoring non-syntactic information in the learning environment. Maybe it is not possible to explain structural acquisition independently, but the addition of the semantic, pragmatic and prosodic information of the learning context cannot be ignored in the learning process without producing a circular argument. The “Identification In the Limit” theory is examined further in the following chapter.

Clark and Lappin’s overall goal is to establish formal, computational descriptions of the language learning process that can be demonstrated to be viable and feasible, although they are keen to point out that demonstrating the tractability of the learning problem does not equate to modelling the actual neurological or psychological processes that may be employed in language acquisition. Chapter 4 discusses a major argument against the machine learning approach: Gold’s “Identification In the Limit” theory, which concludes that language acquisition is not viable for a ‘learning machine’ and so only linguistic nativism can explain success in language acquisition. C&L reject a number of assumptions in Gold’s model. They do not agree that presentation of the target language should be considered unlimited, as this allows for the unnatural condition of the malevolent presentation of data – intentionally withholding crucial evidence samples and offering evidence in a sequence detrimental to learning. They reject Gold’s lack of time limitation placed on learning. They reject the impossibility of learners querying the grammaticality of a string. They reject the argument that ‘side’ information – information relating to the pragmatics and semantics of the learning context – has no influence on learning syntax, and they provide further evidence to reject the view that learning is through positive evidence only. Perhaps the most significant assumption made in the Gold model that C&L reject is the insistence on limiting the hypothesis space available to the learner. Rather, C&L insist that it is the learnable class of a language that is limited, while the learner is free to form unlimited hypotheses on the limited language. It seems that Gold’s approach is an argument for APS which does not consider alternative approaches: “The argument for the subset principle rests on similar misconceptions as the argument that a target-learnable class must be known to the learner.” (p.97)

Proceeding from a critique of the UG position of a strong bias towards linguistic nativism, C&L begin to introduce their alternative machine-learning approach in chapter 5, “Probabilistic Learning Theory”. The initial step in the weak bias argument is to replace a binary definition of a convergent grammar, typical in UG threories, with a probabilistic definition as this more accurately reflects natural learning processes. C&L on this, and a number of other occasions, object to the simplistic lines of argument employed by Chomsky and his followers in their rejection of statistically-based models of learning. While it may be true that the primitive statistical models critiqued by Chomsky are incapable of producing a satisfactory distinction between grammatical and ungrammatical strings, this does not prove that all statistical methods are inferior to UG descriptions: “the failure of a simple word bigram model to predict a difference in probability between observed events does not entail that statistical language modeling methods in general are unable to handle grammar induction.” (p.101) Consequently C&L introduce a range of statistical methods that they propose can better represent the nature of language acquisition than the under-specified domain-specific mechanisms presumed in UG theories. Central to a plausible probabilistic approach to modelling language is the distinction between simple measures of frequency and distributional measures. Here C&L are proposing that a learner will hypothesise the likelihood of a sequence, based on observations, in order to converge on the most likely grammar. Recent studies using such probability-based grammars are reported in this and later chapters to offer very reliable results, even approaching a reliability factor close to 90%. This general framework uses Probably Approximately Correct (PAC) learning algorithms. PAC algorithms predict efficient, time-limited learning, without requiring the learner to know that their grammar has converged on the correct grammar. However, traditional PAC models have problems, including the requirement that data samples are labelled and the conditions under which a language can be learned, which make them unlikely candidates for reliable models of natural language processing. These issues are dealt with in the following chapters by modifying PAC algorithms in order to better reflect the conditions of natural language processing.

Replacing Gold’s paradigm and the PAC learning with three key assumptions (1. language data is presented to the learner unlabelled; 2. the data includes a proportion of ungrammatical sentences; and 3. efficient learning takes place despite negative examples) allows C&L to introduce the Disjoint Distribution Assumption to more accurately reflect natural language learning, in chapter 6. This probabilistic algorithm depends on the observed distribution of segmented strings, and on the adoption of the principle of converging on a probabilistic grammar (a string is probably correct) in place of a categorical grammar (a string is definitely correct). Using a distributional measure ensures that the probability of any utterance will never be zero, allowing for errors in the presented language data, but each string will be measured against its observed likelihood of distribution. In fact, this model predicts over-generalisation and under-generalisation in the learner’s output because, with an unlimited hypothesis space, “It is the ungrammatical strings that an incorrect hypothesis would wrongly predict to be grammatical, and of high probability, that provide crucial data for learning.” (p.133) The addition of word class distributional data – the likelihood of a certain word class in preceding and succeeding position – also ensures greater reliability of judging the probability of a string being grammatical.

A major aim of this book is to provide a computational account of the language learning puzzle that may not necessarily replicate natural language acquisition, but will make the problem tractable – possible within the defined assumptions. It is the contention of the authors that UG theories have made the wrong assumptions in relation to the learning task and the learning conditions, and in chapter 7 “Computational Complexity and Efficient Learning” they set out the assumptions that allow learning to be efficient without positing a strong bias toward linguistic nativism. To achieve this, they examine the amount and complexity of the data, and the nature and constraints of the learning task, in order to propose learning algorithms that can simulate learning under these conditions, while warning that the purpose here is to demonstrate the possibility in a computational environment not the actual psychological processes that enable human language learning. That is, demonstrating a computational or a linguistic model of learning and language does not entail its psychological reality. A central assumption that is essential to C&L’s machine learning approach is that the input data is not homogenous, resulting in some parts of language being ‘more learnable’ than others. Using a standard domain-general capacity to cluster, the language learner can focus on the easier learning tasks leaving the more difficult parts of grammar for later. More complex learning tasks can then be attacked class-by-class according to a Fixed Parameter Tractability algorithm. Ultimately, C&L argue that complex grammatical problems are no better solved by a UG Principles and Parameters approach; the learning problem remains just as complex and learning need not be achieved any more efficiently. Thus, when UG theories use ‘strong bias’ position as the only argument to deal with complexity, they have not solved the problem posed by a seemingly intractable learning task.

If we are to reject the presumption of the strong bias in linguistic nativism, we need to be confident that its replacement can produce reliable results. Chapter 8 starts to provide those results, illustrated in a range of proposed algorithms. The process of hypothesis generation in Gold’s ‘Identification In the Limit’, a key support of UG, is described as being close to random, and consequently “hopelessly inefficient” (p.153). Various replacements that have been tested, initially in non-linguistic contexts, include (Non-)Deterministic Finite State Automata. These algorithms have then proved effective in restricted language learning contexts. Simple distributional and statistical (including hidden Markov models) learning algorithms offer promising results, but must be adapted to also simulate grammar deduction. Lattice based formalisms are offered as one prospect to “demonstrate tractable learnability for a nontrivial subclass of context-sensitive representations.” (p.161)

Despite promising results, there are still objections to distributional models, and these are countered in chapter 9, “Grammar Induction through Implemented Machine Learning,” which describes the results of real algorithms working on real data. In general, learning algorithms are tested against a ‘gold standard.’ Typically the algorithm performs a parsing, tagging, segmenting or similar task on a corpus, which may or may not be labelled in some way, and the results are measured against a previously-annotated version of the corpus. Corpora in these experiments tend to be samples of English intended for adults – such as the extracts from the Wall Street Journal included in the Penn Treebank (Marcus, Marcinkiewicz and Santorini, 1993). Success is measured by how closely the algorithm matches the previous results and is typically presented as a percentage. Learning algorithms can be divided into “supervised” – requiring the corpus to carry some form of information such as part of speech tagging – and “unsupervised” – working on a ‘bare’ corpus of language.  Not surprisingly, supervised learning algorithms, such as the Maximum Likelihood Estimate, match the ‘gold standard’ in about 88-91% of cases. More surprising, perhaps, are the success rates of unsupervised learning algorithms in word segmentation, in learning word classes and morphology, and in parsing. For instance, parsing algorithms match from 52 to 87% of cases when compared to the ‘gold standard’ of previously-annotated corpora. However, close examination of results often reveals explainable differences – that is, the algorithm produces results which are plausible even if they do not match the human annotation. In brief, various experiments have demonstrated that in the vast majority of cases, learning algorithms can accurately segment and categorise adult language without guidance, and make suitable hypotheses about their grammatical role.

Chapter 10 returns to UG models of language learning, and ‘Principles and Parameters’ arguments in particular, in order to compare them with statistical models of learning. C&L claim that the strong bias in this UG model would require an innate phonological segmenter and part of speech tagger, and that by limiting the hypothesis and parameter space available to the learner, the language learning task actually becomes far more complex, particularly as the highly abstract nature of UG parameters appear to have very little direct relationship to the primary language data. Tellingly, C&L lament the paucity of theoretical and experimental evidence for the Principles and Parameters (P&P) framework to provide examples of language variation and learning that fit with the proposal that learning one ‘parameter’ in a language automatically entails the learning of a set of features:
“The burden of argument falls on the advocates of the P&P view to construct a workable and empirically adequate type hierarchy that captures the main patterns of language variation. The fact that such a theory has not yet emerged, even in general outline, after so many years of work within the P&P framework provides good grounds for questioning the concept of UG that it presupposes.” (p.184)
Even more worrying for UG supporters of a strong bias towards innate language acquisition is the near-indifference to questions of acquisition in Minimalist Program research, the latest theory in UG. A strong bias towards linguistic nativism requires the human brain to evolve these biases in order to learn language efficiently. However, this places language prior to humans, as part of the environment to which humans must adapt. Christiansen and Chater (2008) are credited by C&L with discrediting this notion: it is language that must adapt to fit the human mind. Seen in this light, it is far more likely that languages adapt to learning biases that have already evolved in the human brain, providing another ‘weak bias’ argument, and so general tendencies in human languages are not examples of Universal Grammar but “emergent properties of the processes through which natural language is adapted over generational cycles of transmission.” (p.185) In place of the UG models, C&L again offer sophisticated statistical models of language learning and generation. Promising results using “Probabilistic Context-Free Grammars” and “Hierarchical Bayesian Models” currently provide the best alternatives to UG models by accounting for language acquisition through comprehensive descriptions and high levels of success in simulations.

In chapter 11, “A Brief Look at Some Biological and Psychological Evidence” C&L quickly review accounts of language learning that support a weak bias. Even where evidence from genetic, biological or psychological studies have been used to support a strong bias, C&L are able to show that this evidence does not necessarily favour nativist arguments. For instance, the FOXP2 gene has been attributed, by Pinker and Jackendoff among others, as the gene responsible for the capacity for human language. In a family where this gene is mutated, all family members suffer a similar severe language disability. C&L point out, though, that this gene is not unique to humans and, as it is responsible for a disruption in motor learning and development in other animals, it is far more likely that this genetic mutation results in a more general learning disability that also severely affects speech and language development. Similarly, psychological disabilities such as Williams Syndrome that may appear to be language-specific are, on closer inspection, related to a range of learning and developmental abnormalities.

In the final chapter, ‘Conclusion,’ C&L review their evidence against the argument from the poverty of the stimulus. They remind us that they are arguing for a weak innate bias towards learning language based on general-domain learning abilities rather than language-specific abilities. They point out that the UG framework has produced few concrete experiments or cases that explain the language variation or acquisition described in theoretical accounts or that “produce an explicit, computationally viable, and psychologically credible account of language acquisition” (p.214). What they have attempted to demonstrate in “Linguistic Nativism and the Poverty of the Stimulus” is that there are explicit computational models of learning that have produced a credible account of learning without requiring language-specific parameters. Although they are far from perfect and much work needs to be done, computational models have already provided a more adequate account than the UG models:
“We hope that if this book establishes only one claim it is to finally expose, as without merit, the claims that Gold’s negative results motivate linguistic nativism. IIL is not an appropriate model. Conditions on learnable classes need not be domain specific. Finally, the learner does not have to (and in general, cannot) restrict its hypothesis to the elements of the learnable class.” (p.215)
Instead C&L advocate the use of domain-general statistical models of language acquisition.


Evaluation
In 12 chapters, Clark and Lappin use “Linguistic Nativisim and the Poverty of Stimulus” to evaluate a key concept in modern linguistics, taking a clearly computational perspective and examining a wide variety of topics. This monograph presupposes familiarity with most of the core issues but does not demand in-depth knowledge of computational linguistics. In such a review I have, naturally, simplified or ignored a number of important arguments presented in this book, and skimmed over significant presentations of learning algorithms. I would suggest however, from this reviewer’s point of view, that C&L present a very cogent and coherent argument.

There are so many sides from which to attack linguistic nativism in general, and the argument from the poverty of stimulus in particular. Opponents have argued that most UG theories are unfalsifiable (e.g. Pullum and Scholz, 2002), that corpora designed to reflect children’s exposure to language demonstrate that the stimulus is not impoverished (e.g. MacWhinney, 2004), that it is absurd to posit the notion that the brain adapted to language, as if language exists in the environment prior to man, rather than language adapting to the general abilities of the brain (Deacon, 1998; Christiansen and Chater, 2008). These arguments, alongside alternatives to linguistic nativism from functional linguistics (e.g. Halliday, 1975; 1993; Halliday and Matthiessen, 1999), are often easily dismissed as being irrelevant to the theory of UG. Some of these arguments are mentioned in this book, but what sets Clark and Lappin’s book apart, and why it must be taken seriously by everyone who proposes some form of UG to explain language acquisition and typology, is that it attacks from within. It claims the very ground claimed by theories of UG. UG attempts to formally and explicitly account for the apparent mismatch between the complexity of the language learning task and the near-universal success of humans in achieving it with such apparently meagre resources. The methods proposed by Clark and Lappin identify what methods could be applied to make the complex task tractable. Specifically, these methods are not restricted to language, but are generally useful learning methods – they are domain-general. If there is one criticism I would make of Clark and Lappin’s argument it is that they do not demonstrate clearly enough – at least to this rather naive reader – just how likely it is that we all use the domain-general learning methods that they propose. For instance, we are left to presume that Probabilistic Approximately Correct and Probabilistic Context-Free Grammar algorithms represent general, non-language specific, models of learning, but this is not demonstrated.

My biggest fear with this approach is that we may be fooled by the apparent sophistication of the tools at our disposal. We need to remember that when using computational tools to help us understand a phenomenon far more complex than computers, we must not allow the tools to force us to see the phenomenon as the tool understands it. It seems more than coincidental that a computational approach to language acquisition mirrors the findings about language that corpus linguistics has revealed; for instance, that language can be viewed as inherently statistically structured. That it can be analysed in this way, or in the form of transformational trees, does not prove that this is how humans learn language. Fortunately, Clark and Lappin are well aware of this trap and frequently warn readers that no matter how well their computational theories may reflect language acquisition facts, the aim of computational models is to demonstrate what is possible, or even likely, not what really happens in the human mind. Until we better understand exactly what neurological processes are actually involved in language acquisition, our task is to try to represent acquisition as best we can. In this endeavour, we have been expertly assisted by Clark and Lappin’s book.

Linguistic Nativism and the Poverty of the Stimulus is a challenging book. It challenges the reader to deal with a range of linguistic, philosophical, mathematical and computational issues. It challenges the reader to remember a dizzying array of acronyms and abbreviations (APS, CFG, DDA, DDL, EFS, GB, HBM, IID, IIL, MP, PCFG, PLD and UG to name but a few). Most of all, it challenges basic concepts in mainstream linguistics. It examines key tenets of UG in the light of advances in machine learning theory and research in the computational modelling of the language acquisition process, and finds them sorely deficient. It exposes so-called proofs supporting the poverty of stimulus, and reveals alternatives that are formally more comprehensive than the explanations previously provided by the linguistic nativism inherent in UG, and empirically more likely to match natural language acquisition processes.


REFERENCES
Christiansen, Morten H. and Chater, Nick. 2008. Language as Shaped by the Brain. Behavioural and Brain Sciences 31. pp.489-558
Deacon, Terrence W. 1998. The Symbolic Species: The Co-Evolution of Language and the Brain. New York: W.W. Norton
Halliday, Michael A.K. 1975. Learning How to Mean: Explorations in the Development of Language. London: Edward Arnold
Halliday, Michael A.K. 1993. Towards a language-based theory of learning. Linguistics and Education 5 p.93-116
Halliday, Michael A.K. & Matthiessen, Christian M.I.M. 1999. Construing Experience Through Meaning. London: Continuum
MacWhinney, Brian. 2004. A Multiple Solution to the Logical Problem of Language Acquisition. Journal of Child Language 31, pp. 883-914
Marcus, Mitchell P., Marcinkiewicz, Mary Ann and Santorini, Beatrice. 1993. Building a Large Annotated Corpus of English: The Penn Treebank. Computational Linguistics19/2, p.313-330
Pullum, Geoffrey K. and Scholz, Barbara C. 2002. Empirical Assessment of Stimulus Poverty Arguments. The Linguistic Review 19, pp.9-50

Wednesday, February 29, 2012

Brighten up your Event Planning with EventBrite

Have you got a workshop, conference or other event, but haven't got the time to register everyone with the relevant details? Does it need a professional website announcement with a map, plus all the necessary information? Well, there's a website that will quickly and painlessly do all of that for you - and much, much more. And all for FREE!!

Eventbrite was recommended by our friendly librarians in Dubai Women's College, and we have both used it for the ReadingRoadshows in Dubai and Abu Dhabi. Here's what the Abu Dhabi one looks like. It took me about 15 minutes to set up - from scratch, for the first time! With an Eventbrite URL of my choice!!!

When someone registers, the attender and the organiser get an email message with a pdf attachment which is the ticket. On the ticket (which can be of any category and any number of categories that you decide) there is all the information required, including an Eventbrite-generated map of the venue, and there is a bar-code and a QR-code. With a free app, this means you can check in your attendees with an android or apple device, with a barcode scanner, or by hand on any number of devices simultaneously.

And that's not all. You can set up the registration form to ask for and/or require other information, such as phone numbers, address, position at work etc., or you can add any question you wish. This data can then be exported, and may also be collated and compared across events. I haven't even told you about the billing facilities it offers, because I haven't used them, but they can take care of everything there as long as you provide your account details.

Perhaps I am very easily impressed, but I am REALLY impressed with Eventbrite.

Wednesday, February 15, 2012

Your PhD - A Haiku

This website (http://dissertationhaiku.wordpress.com/) asks for people to submit their PhD research as succinctly as possible, or more exactly, in 17 syllables over 3 lines of 5,7 and 5 syllables. Why would you do that? Here's the explanation:
Dissertations are long and boring. By contrast, everybody likes haiku. So why not write your dissertation as a haiku? Please email yours (along with your name, institution, a 1-2 sentence text description of your work, and any URL you'd like your name linked to) to dissertationhaiku@gmail.com.
Maybe you'd like a go. Here are a few examples to get you thinking:
Cells in between cells,
make the retina’s magic
and keep its secrets.
Tiny neutrinos ….
How do we catch the damn things?
Computers can help.
Right of copyright
Depends on nature of works
Otherwise: chaos.
Most haiku's are accompanied by some very brief information.

Knowledge that is tacit
Still needs capture and sharing
How do we do that?
Dissertation title: e-Learning 2.0: Emergence, social networks and the creation of shared knowledge
As learners in the workplace turn toward just-in-time answers, knowledge and support networks, it becomes more pressing to understand how knowledge can be harnessed in a new, anytime digital environment. This dissertation explored knowledge design via new emerging technologies that allow for better finding, creating and sharing of information.
By: Colleen Carmean

Welfare for elders
But nothing for our children.
Who are we kidding?
Dissertation title: The Age of Welfare: Citizens, Clients and Generations in the Development of the Modern Welfare State (UC Berkeley, 2001)
My dissertation examined why some countries allocate the lion’s share of their social welfare resources to the elderly, while others have a more balanced profile of spending on working-aged adults and children. Italy, Japan, and the United States are examples of the former; Canada, the Netherlands, and Sweden are examples of the latter.
By: Julia Lynch
Here's my contribution:

Intonation shows
New information. In text?
It’s punctuation
Dissertation title: “Structuring Information in Written English”
This dissertation in applied linguistics used discourse analysis to identify the role and realization of Information Structure in a corpus of written technical English. It revealed historical, psychological and neurological evidence for why punctuation has replaced intonation in written English in order to provide Information Structure, as defined in systemic functional linguistics.
By: Nick Moore
It has not yet been accepted. I'll keep you posted!

Tuesday, February 14, 2012

Calais: Not just cross-channel ferries!

Open Calais employs semantic web technology to identify key elements on your webpages. It is sponsored by the Thompson/Reuters news agency who are committed to keeping Calais open - as in open source. Here's a brief intro


The basic idea is that if we want the web to 'read' our pages so it can help us connect ideas, then we have to give it something to work with.Semantic proxy takes the text, identifies proper nouns, category areas and syntactic relationships.

For example, my blog entry for the book review for Hunston's "Corpus Approaches to Evaluation" produces this table for the semantic content areas. Interestingly, when I tried it just yesterday, the last 4 entries here were first and had 3 stars.


The semantic proxy version gives almost the same results, but is not as pretty. They are a little clearer on what they do. This one does what it says on the tin:

Perhaps more telling is the chart for people, companies, and events, below. Here there is an attempt to tag all names, presumably to make the dream of a semantic web (see here and here) just a little closer.

On close inspection, however, we can notice a number of errors here. John McH Sinclair loses his surname. The term Modal-Like Expressions in the review is abbreviated to MLEs, which is recognised here as a company. Again, yesterday's results also included Bernstein (as in Basil) as a company. None of the books or articles in the references is recognised as a publication.


The final set of results is probably the set that is likely to include the most errors. Here there is an attempt to match semantic relations between the entities in the text, using simple predication, see left. Subjects, objects and verbs are tied together under the misleading title of generic relations. For instance 'Nick Moore' is identified as a subject where the object is 'a reviewer for a number of journals' and the verb is identified as 'be.' The only other subject-agent listed here is Susan Hunston, who 'very similar patterns' (object) 'produce' (verb) or 'the methodology of the chapter' (object) 'apply' (verb). Susan Hunston (subject) also does intransitive things such as 'investigate' (verb). The Quotations are not such in the original article, but real quotations such as “one function of evaluation is to reify texts and propositions by assigning them an epistemic status.” is not recognised as such (unless they mean something else by quotation).

Overall Grade: C+. Keep trying!!!

Wednesday, February 1, 2012

Reading Roadshow

The "Power of One" campaign continues.

Tom Le Seelleur (yes, him again) & the gang (yes, that includes me again) are organising a series of workshops for teachers called the "Reading Roadshow." We will provide free workshops (and lots of prizes) on a Wednesday afternoon in a different Emirate every month as part of the campaign to encourage the country to read for pleasure.

Our first event in Sharjah was well attended and we were fortunate to be supported by Pearson the publishers, the British Council, International House - DubaiAl Mutannabi bookshop and the Emirates Airlines Festival of Literature. The speakers were Julia Johnson - an children's author whose books celebrate Arabic culture - and Catherine Namoor & Alicia Salaz who talked about developing READ posters with local relevance. We will provide workshops in Abu Dhabi and Dubai, before travelling to the other Emirates.

Further details are on the Reading Roadshow website.

Tuesday, January 31, 2012

Coffin, Donohue & North: Exploring English Grammar

Exploring English Grammar by Caroline Coffin
My rating: 4 of 5 stars

There's nothing so practical as a good theory, says Michael Fullan and here we have a very practical approach to a very good theory. Coffin, Donohue & North have started with the practical situation that most people interested in grammar are likely to have already experienced formal or communicative descriptions of English grammar. They proceed to take these people - who are more often than not going to be language teachers - and introduce a model of language (Systemic Functional Grammar) which is likely to be of great benefit to them. Throughout, the approach is a very practical one.

This is an unusual book. It is not a grammar resource book - such as Swan's Practical English Usage - or a comprehensive grammar of English - such as Introduction to Functional Grammar. Instead it is a book about grammar - what it is, why we have it and how we can use it. Unusual as this may be, it is very - yes, you guessed it - practical. Most grammar resource books assume that you know what a grammar is for, and you just what answers to specific questions, while comprehensive grammars rarely have time to look at how you night apply a grammatical explanation or category to a real situation. This book admirably bridges these two needs, so being unusual is not a bad thing here.

Coffin & her colleagues from the Open University introduce the concepts behind grammar through example. They take the reader from an understanding that they already know something about formal grammar, through descriptions of how communicative grammars compensate for the impractical nature of formal grammars, to explaining how a functional grammar - in this case systemic functional grammar - can combine the practicality of a communicative grammar with the theoretical rigour and the explanatory demands of a formal grammar. All of this is done through clear text, highly practical examples of grammatical rules (that I fully endorse), plenty of real-life texts to illustrate, entertain and illuminate, and short exercises to drive the message home.

More details are available on the publisher's page, and there is an extensive online "companion page" which links you to extra exercises, further links, corrections and the Routledge English Language & Applied Linguistics resources page.

Another Goodreads.com review.View all my reviews

wave Goodbye

Remember, not so long ago, I thought that Google Wave was a really good idea for a collaborative tool?
Well, apparently I was the only one!
Google announced that they would no longer support the Wave (read only from Jan 2012, fully dead from April 2012)  - not really surprising as the "Google Gears" needed to run it have not been turning for some time. Google claim that everything Gears did is now done in new browsers - especially theirs ("Chrome").
I recently (virtually) attended a seminar by BT's Simon Thompson who suggested that people found the Wave just a bit too complicated to deal with. Apparently, once you get a group of people all joining in a conversation, adding views, responding to what others say, and including extra information in the form of files, pictures etc., when it is written down on paper, it all looks a bit of a mess. Not surprising at all, when you consider the amazing complexity of conversation, and the comparatively impoverished information portrayed in the written versions of language when compared to speech. Google! You should have asked a conversation analyst or a systemic functional linguist first!

Saturday, January 28, 2012

Life, the Universe and Everything by Antonio Damasio

Far too often discussions of consciousness are based on speculation and the desire to make presumed facts fit an ideology - yes, I am talking about you Mr. Chomsky. At other times we are fortunate to find people that start with identifiable facts and then attempt to explain them. In this fascinating TED talk Antonio Damasio lets us in on his latest understanding of the human brain, the human mind and the centrality of the self in generating a consciousness. Hold on tight as he covers a very wide range of topics in under 19 minutes, but at the end of the ride I feel much closer to an understanding of what we are.

This talk develops Damasio's earlier theories and is in line with what we know about brains from Edelman, about evolution from Deacon and about language from HallidayThe talk is based loosely on Damasio's latest book "Self Comes to Mind" which is reviewed in Constructivist Foundations 7:1.

Wednesday, January 25, 2012

Someone Please Explain This To Me

Why do I find this so funny?
Is it the linguist in me?
Probably Not!