It has come to my attention that Language is considering making an "event" of Vyvyan Evan's junk book The Language Myth. What do I mean by an "event"? Well, and here I quote: "because of the potentially controversial nature of the book, Language is planning a new type of review, in which we target the book for commentary papers by 4-5 individuals who have different academic perspectives." These reviews are to be about 1500 words. So, Language is going to make a BIG DEAL (6-7500 words of criticism plus a reaction by Evans, I would assume) out of this book pretending that there is something there, pretending that it is "controversial" in the sense that that the book raises many interesting issues that people of good faith can understand in different ways and that debating would enlightening. This is false. There are not and that's because the book is junk. The suggestion that Language (and, by extension, the LSA) believes otherwise is a terrible message to send.
Let me be clear: the book is not controversial. It is junk. Pure, unadulterated, complete junk. Reading it will make you dumber. The fact that Language is doing a "new type of review" will only suggest that this is not so. It will suggest that there really are various reasonable sides to the issue Evans book discusses and that the views in the book are worth taking seriously. After all, Language, the journal of the LSA, the main professional organization of linguistics, thinks that the book is is "controversial," (which in common parlance suggests well argued if still a bit out there). It does not suggest that the book is junk. Moreover, getting a wide range of reviews virtually guarantees that at least one of them will suggest that the views are not junk. After all, I bet Language wants to be "fair." What piece of junk could ask for a better endorsement than this?
Generativists have always considered Language the place you publish when you can't get your stuff into LI or NLLT, or Lingua or… It is far down the list of desirable publishing venues. If it ever was the journal that published the stuff at the cutting edge, it is no longer is that journal. Nonetheless, it is the official journal of the LSA and as such it should care about whether the works it highlights meet even minimal professional standards (one would hope for more than that, of course). The Evans book does not. To repeat, it's junk. So why exactly does Language want to showcase it? Do the editors hate Generative Grammar that much? Do they really think that generative linguistics has been an intellectual disaster? It would be nice to know if this is what the editors think, for if it is, maybe it's time for Generativists to either leave the LSA or the LSA should consider replacing the editors.
So, either Language hates 2/3 of the field (always a possibility) or the editors are filled with self-loathing. I find it hard to believe that any other professional journal would showcase work that is shoddy, unprofessional, uninformed and logically lacking. Can you see Physical Review doing a special review on the latest approaches to perpetual motion? Or the American Anthropological Review doing a special issue on creation science? I can't. They have more self respect than that. They know that these topics are junk. But apparently Language is different. It's "open-minded" and willing to consider even junk as worthy of showcasing because of its "controversial" nature. This is not the first time Language has done this (see here). Someone like me might get the impression that Language in no way respects what it is that Generative Grammar has done over the last 65 years. The idea that Evans' book is "controversial" suggests that the editors have lost all critical sense and are willing to admit the most egregious junk into its journals. This is not to say that Evans' book does not deserve special treatment in the pages of Language. It does. Language should be highlighting the fact that work like this is not worth the paper that it is written on. A decent hatchet job, now that I understand. But a "new type of review"? It sends entirely the wrong message.
Showing posts with label Language. Show all posts
Showing posts with label Language. Show all posts
Friday, April 17, 2015
Monday, October 21, 2013
Mothers know best
My mother always told me that you should be careful what you
wish for because you just might get it. In fact, I’ve discovered that her
advice was far too weak: you should be careful what you idly speculate about as
it may come to pass. As readers know, my last post (here)
questioned the value added of the review process based on recent research
noting the absence of evidence that reviewing serves its purported primary
function of promoting quality and filtering out the intellectually less
deserving. Well, no sooner did I write this than I received proof positive that
our beloved LSA, has implemented a no review policy for Language’s new online journal Perspectives. Before I review the
evidence for this claim, let me say that though I am delighted that my
ramblings have so much influence and can so quickly change settled policy, I am
somewhat surprised at the speed with which the editors at Language have adopted my inchoate maunderings. I would have hoped
that we might make haste slowly by first trying to make the review process progressively
less cumbersome before adopting more exciting policies. I did not anticipate
that the editors at Language would be
so impressed with my speculations that they would immediately throw all caution
aside and allow anything at all, no matter how slipshod and ignorant, to appear
under its imprimatur. It’s all a bit dizzying, really, and unnerving (What
power! It’s intoxicating!). But why do things by halves, right? Language has chosen to try out a bold
policy, one that will allow us to see whether the review process has any
utility at all.
Many of you will doubt that I am reporting the
aforementioned editorial policy correctly. After all, how likely is it that
anything I say could have such immediate impact? In fact, how likely is that that the LSA and
the editors of Language and its
online derivatives even read FoL? Not
likely, I am sorry to admit. However,
unbelievable as it may sound, IT IS TRUE, and my evidence for this is the
planned publication of the target article by Ambridge, Pine and Lieven (APL) (“Child
language: why universal grammar doesn’t help” here[1]).
This paper is without any redeeming intellectual value and I can think of only
two explanations for how it got accepted for publication: (i) the radical
change in review policy noted above and (ii) the desire to follow the Royal
Society down the path of parody (see here).
I have eliminated (ii) because unlike the Royal Society’s effort, APL is not even slightly funny haha (well maybe
as slapstick, I’ll let you decide). So that leaves (i).[2]
How bad is the APL paper? You can’t begin to imagine. However, to help you vividly taste its
shortcomings, let me review a few of its more salient “arguments” (yes, these
are scare quotes). A warning, however, before I start. This is a long post. I
couldn’t stop myself once I got started. The bottom line is that the APL paper
is intellectual junk. If you believe me, then you need not read the rest. But
it might interest you to know just how bad a paper can be. Finding zero on a
scale can be very instructive (might this be why it is being published? Hmm).
The paper goes after what APL identify as five central
claims concerning UG: identifying syntactic categories, acquiring basic
morphosyntax, structure dependence, islands and binding. They claim to
“identify three distinct problems faced by proposals that include a role for
innate knowledge –linking, inadequate
data coverage, and redundancy…(6).” ‘Linking’ relates to
“how the learner can link …innate knowledge to the input language (6).”
‘Data-coverage’ refers to the empirical inadequacy of the proposed universals,
and ‘redundancy’ arises when a proposed UG principle proves to be accurate but
unnecessary as the same ground is covered by “learning procedures that must be
assumed by all accounts” and thus obviate the need “for the innate principle or
constraint” (7). APL’s claim is that all proposed UG principles suffer from one
or another of these failings.
Now far be it from me to defend the perfection of extant UG
proposals (btw, the principles APL discusses are vintage LGB conceptions, so I
will stick to these).[3]
Even rabid defenders of the generative enterprise (e.g. me) can agree that the
project of defining the principles of UG is not yet complete. However, this is
not APL’s point: their claim is that the proposals are obviously defective and clearly
irreparable. Unfortunately, the paper contains not a single worthwhile
argument, though it does relentlessly deploy two argument forms: (i) The Argument
from copious citation (ACC), (ii) The Argument from unspecified alternatives
(AUA). It combines these two basic
tropes with one other: ignorance of the relevant GB literature. Let me
illustrate.
The first section is an attack on the assumption that we
need assume some innate specification of syntactic categories so as to explain
how children come to acquire them, e.g. N, V, A, P etc. APL’s point is that distributional analysis
suffices to ground categorization without this parametric assumption. Indeed,
the paper seems comfortable with the idea that the classical proposals critiqued
“seem to us to be largely along the right lines (16),” viz. that “[l]earners
will acquire whatever syntactic categories are present in a particular language
they are learning making use of both distributional …and semantic
similarities…between category members (16).” So what’s the problem? Well, it seems
that categories vary from language to language and that right now we don’t have
good stories on how to accommodate this range of variation. So, parametric theories
seeded by innate categories are incomplete and, given the conceded need for
distributional learning, not needed.
Interestingly, APL does not discuss how distributional learning
is supposed to achieve categorization. APL is probably assuming non-parametric
models of categorization. However, to function, these latter require
specifications of the relevant features that are exploited for categorization. APL,
like everyone else, assume (I suspect) that we humans follow principles like
“group words that denote objects together,” “group words that denote events
together,” “group words with similar “endings” together,” etc. APL’s point is
that these are not domain specific
and so not part of UG (see p.12). APL is fine with innate tendencies, just not
language particular ones like “tag words that denote objects as Nouns,” “tag words that denote events
as Verbs.” In short, APL’s point is that calling the groups acquired nouns,
verbs, etc. serves no apparent linguistic function . Or does it?
Answering this question requires asking why UG distinguishes
categories, e.g. nouns from verbs. What’s the purpose of distinguishing N or V in
UG? To ask this question another way: which GB module of UG cares about Ns, Vs,
etc? The only one that I can think of is the Case Module. This module identifies
(i) the expressions that require case (Nish things) (ii) those that assign it
(P and Vish things) and (iii) the configurations under which the assigners
assign case to the assignees (roughly government). I know of no other part of
UG that cares much about category labels. [4]
[5]
If this is correct, what must an argument aiming to show
that UG need not natively specify categorical classes show? It requires showing
that the distributional facts that Case Theory (CT) concerns itself with can be
derived without such a specification. In other words, even if categorization
could take place without naming the categories categorized, APL would need to
show that the facts of CT could also be derived without mention of Ns and Vs
etc. APL doesn’t do any of this. In fact, APL does not appear to know that the
facts about CT are central to UG’s adverting to categorical features.
Let me put this point another way: Absent CT, UG would
function smoothly if it assigned arbitrary tags to word categories, viz. ‘1’,
‘2’ etc. However, given CT and its role
in regulating the distribution of nominals (and forcing movement) UG needs category
names. CT uses these to explain data like: *It
was believed John to be intelligent, or *Mary
to leave would be unwise or *John
hopes Bill to leave or *who do you
wanna kiss Bill vs who do you wanna
kiss. To argue against categories in UG requires deriving these kinds of
data without mention of N/V-like categories. In other words, it requires
deriving the principles of CT from non-domain specific procedures. I personally
doubt that this is easily done. But, maybe I am wrong. What I am not wrong
about is that absent this demonstration we can’t show that an innate
specification of categories is nugatory. As APL doesn't address these concerns
at all, its discussion is irrelevant to the question they purport to address.
There are other problems with APL’s argument: it has lots of
citations of “problems” pre-specifying the right categories (i.e. ACC), lots of
claims that all that is required is distributional analysis, but it contains no
specification of what the relevant features to be tracked are (i.e. AUA). Thus,
it is hard to know if they are right that the kinds of syntactic priors that
Pinker and Mintz (and Gleitman and Co. sadly absent from the APL discussion)
assume can be dispensed with.[6]
But, all of this is somewhat besides the point given the earlier point: APL
doesn’t correctly identify the role that categories play in UG and so the presented
argument even if correct doesn’t
address the relevant issues.
The second section deals with learning basic morphosyntax.
APL frames the problem in terms of divining the extension of notions like
SUBJECT and OBJECT in a given language. It claims that nativists require that
these notions be innately specified parts of UG because they are “too abstract
to be learned” (18).
I confess to being mystified by the problem so construed. In
GB world (the one that APL seem to be addressing), notions like SUBJECT and
OBJECT are not primitives of the theory. They are purely descriptive notions,
and have been since Aspects. So, at least in this little world, whether
such notions can be easily mapped to external input is not an important problem. What the GB version of UG does need is a
mapping to underlying structure (D-S(tructure)). This is the province of theta
theory, most particularly UTAH in some version. Once we have DS, the rest of UG
(viz. case theory, binding theory, ECP) regulate where the DPs will surface in
S-S(tructure).
So though GB versions of UG don’t worry about notions like
SUBJECT/OBJECT, they do need notions that allow the LAD to break into the
grammatical system. This requires primitives with epistemological priority (EP) (Chomsky’s term) that allow the LAD
to map PLD onto grammatical structure. Agent
and patient, seem suited to the task
(at least when suitably massaged as per Dowty and Baker). APL discusses Pinker’s version of this kind
of theory. Its problem with it? APL claims that there is no canonical mapping
of the kind that Pinker envisages that covers every language and every
construction within a language (20-21). APL cites work on split ergative
languages and notes that deep ergative languages like Dyirbal may be particularly
problematic. It further observes that many of these problems raised by these
languages might be mitigated by adding other factors (e.g. distributional
learning) to the basic learning mechanism. However, and this is the big point,
APL concludes that adding such learning obviates the need for anything like
UTAH.
APL’s whole discussion is very confused. As APL note, the
notions of UG are abstract. To engage it, we need a few notions that enjoy EP.
UTAH is necessary to map at least some
input smoothly to syntax (note: EP does not require that every input to the syntax be mapped via
UTAH to D-S). There need only be a core set of inputs that cleanly do so in
order to engage the syntactic system. Once primed other kinds of information
can be used to acquire a grammar. This is the kind of process that Pinker
describes. This obviates the need for a general
UTAH like mapping.
Interestingly APL agrees with Pinker’s point, but it bizarrely
concludes that this obviates the need for EPish notions altogether, i.e. for finding
a way to get the whole process started. However, the fact that other factors
can be used once the system is
engaged does not mean that the system can be engaged without some way to get it
going. Given a starting point, we can move on. APL doesn’t explain how to get
the enterprise off the ground, which is too bad, as this is the main problem
that Pinker and UTAH addresses.[7]
So once again, APL’s discussion fails to engage UG’s main worry: how to
initially map linguistic input onto DS so that UG can work its magic.
APL have a second beef with UTAH like assumptions. APL
asserts that there is just so much variation cross linguistically that there
really is NO possible canonical
mapping to DS to be had. What’s APL’s argument? Well, the ACC, argument by
citation. The paper cites resaearch that claims there is unbounded variation in
the mapping principles from theta roles to syntax and concludes that this is
indeed the case. However, as any moderately literate linguist knows, this is
hotly contested territory. Thus, to make the point APL wants to make responsibly requires adjudicating these
disputes. It requires discussing e.g. Baker’s and Legate’s work and showing
that their positions are wrong. It does not
suffice to note that some have argued
that UTAH like theories cannot work if others have argued that they can. Citation is not argumentation, though APL
appears to read as if it is. There has
been quite a bit of work on these topics within the standard tradition that APL
ignores (Why? Good question). The absence of any discussion renders APL’s conclusions
moot. The skepticism may be legitimate (i.e. it is not beside the point).
However, nothing APL says should lead any sane person to conclude that the
skepticism is warranted as the paper doesn’t exercise the due diligence
required to justify its conclusions. Assertions are a dime a dozen. Arguments
take work. APL seems to confuse the first for the second.
The first two sections of APL are weak. The last three
sections are embarrassing. In these, APL fully exploits AUAs and concludes that
principles of UG are unnecessary. Why? Because the observed effects of UG
principles can all be accounted for using pragmatic discourse principles that
boil down to the claim that “one cannot extract elements of an utterance that
are not asserted, but constitute background information” …and “hence that only
elements of a main clause can be extracted or questioned” (31-32). For the case
of structure dependence, APL supplements this pragmatic principle with the further
assertion that “to acquire a structure-dependent grammar, all a learner has to
do is to recognize that strings such as the
boy, the tall boy, war and happiness share both certain functional and –as a consequence-
distributional similarities” (34). Oh boy!! How bad is this? Let me count some of the ways.
First, there is no semantic or pragmatic reason for why back-grounded
information cannot be questioned. In fact, the contention is false. Consider
the Y/N question in (1) and appropriate negative responses in (2):
(1) Is
it the case that eagles that can fly can swim
(2) a.
No, eagles that can SING can swim
b. No eagles that can fly, can SING
Both (2a,b) are fine answers to the question in (1). Given
this, why can we form the question with answer (2b) as in (3a) but not the
question conforming to the answer in (2a) as in (3a)? Whatever is going on has nothing to do with whether it is possible
to question the content of relative clause subjects. Nor is it obvious how
“recogniz[ing] that strings such as the
boy, the tall boy, war and happiness share both certain functional …and distributional
similarlities” might help matters.
(3) a. *Can
eagles that fly can swim?
b.
Can eagles that can fly swim?
This is not a new point and it is amazing how little APL has
to say about it. In fact, the section on structure dependence quotes and seems
to concede all the points made in the Berwick et. al. 2011 paper (see here).
Nonetheless APL concludes that there is no problem in explaining the structure
dependence of T to C if one assumes that back-grounded info is frozen for
pragmatic reasons. However, as this is obviously false, as a moment’s thought
will show, APL’s alternative “explanation” goes nowhere.
Furthermore, APL doesn’t really offer an account of how
back-grounded information might be relevant as the paper nowhere specifies what
back-grounded information is or in which contexts it appears. Nor does APL explicitly offer any pragmatic
principle that prevents establishing syntactic dependencies with back-grounded
information. APL has no trouble specifying the GB principles it critiques, so I
take the absence of a specification of the pragmatic theory to be quite
telling.
The only hint APL provides as to what it might intend (again
copious citations, just no actual proposal) is that because questions ask for new
information and back-grounded structure is old information it is impossible to
ask a question regarding old information (c.f. p. 42). However, this, if it’s
what APL has in mind (which, again is unclear as the paper never actually makes
the argument explicitly) is both false and irrelevant.
It is false because we can focus within a relative clause island,
the canonical example of a context where we find back-grounded info (c.f. (4a)).
Nonetheless, we cannot form the question (4b) for which (4a) would be an
appropriate answer. Why not? Note, it cannot be because we can’t focus within
islands, for we can as (4a) indicates.
(4) a. John
likes the man wearing the RED scarf
b.
*Which scarf does John like the man who wears?
Things get worse quickly. We know that there are languages
that in fact have no trouble asking questions (i.e. asking for new info) using
question words inside islands. Indeed, a good chunk of the last thirty years of
work on questions has involved wh-in-situ
languages like Chinese or Japanese where these kinds of questions are all perfectly
acceptable. You might think that APL’s claims concerning the pragmatic
inappropriateness of questions from back-grounded sources would discuss these
kinds of well-known cases. You might, but you would be wrong. Not a peep. Not a
word. It’s as if the authors didn’t even know such things were possible (nod
nod wink wink).
But it gets worse still: ever since forever (i.e. from Ross)
we know that Island effects per se
are not restricted to questions. The same things appear entirely with
structures having nothing to do with focus e.g. relativization and
topicalization to name two relevant constructions. These exhibit the very same
island effects that questions do, but in these constructions the manipulanda do
not involve focused information at all. If the problem is asking for new info from a back-grounded source,
then why can’t operations that target old
back-grounded information not form dependencies into the relative clause? The central fact about islands is that it
really doesn’t matter what the moved element means, you cannot move it out (‘move’ here denotes a particular
kind of grammatical operation). Thus, if you can’t form a question via movement,
you can’t relativize or tropicalize using movement either. APL does not seem
acquainted with this well-established point.
One could go on: e.g. resumptive pronouns can obviate island
effects but the analogous non-resumptive analogues do not despite semantic and
pragmatic informational equivalence, islands in languages like Swedish/Norwegian
do not allow extraction from any
island whatsoever, contrary to what PL suggests. All of this is relevant to
APL’s claims concerning islands. None of it is discussed, nor hinted at.
Without mention of these factors, APL once again fails to address the problems
that UG based accounts have worried about and discussed for the last 30 years.
As such, the critique advanced in this section on islands, is, once again,
largely irrelevant.
APL’s last section on binding theory (BT) is more of the
same. The account of principle C effects in cases like (4) relies on another
pragmatic principle, viz. that it is “pragmatically anomalous to use a full
lexical NP in part of the sentence that exists only to provide background
information” (48). It is extremely unclear what this might mean. However, on at least the most obvious
reading, it is either incorrect or much too weak to account for principle C
effects. Thus, one can easily get full NPs within back-grounded structure (e.g.
relative clauses like (4a)). But within the
relative clause (i.e. within the domain of back-grounded information)[8],
we still find principle C effects (contrast (4a,b)).
(5) a. John
met a woman who knows that Frank1 loves his1 mother
b.
* John met a woman who knows that he1 loves Frank’s1
mother
The discussion of principles A and B are no better. APL does
not explain how pragmatic principles explain why reflexives must be “close” to
their antecedents (*John said that Mary
loves himself or *John believes
him/heself is tall), why they cannot be anteceded by John in structures like John’s
mother upset himself (where the antecedent fails to c-command but is not in a clause), why they must be
preceded by their antecedents (*Mary
believes himself loves John) etc. In
other words, APL does not discuss BT and that facts that have motivated it at all
and so the paper provides no evidence for the conclusion that BT is redundant
and hence without explanatory heft.
This has been a long post. I am sorry. Let me end. APL is a
dreadful paper. There is nothing there. The question then is why did Perspectives accept it for publication? Why
would a linguistics venue accept such
a shoddy piece of work on linguistics
for publication? It’s a paper that displays no knowledge of the relevant
literature, and presents not a single argument (though assertions aplenty) for
its conclusions. Why would a journal sponsored by the LSA allow the linguistic
equivalent of flat-earthism to see the light of day under its imprimatur? I can
only think of only one reasonable explanation for this: the editors of Language have decided to experiment with
a journal that entirely does away with the review process. And I fear I am to
blame. The moral: always listen to your mother.
[1]
It’s currently the first entry on his Ambridge’s web page.
[2]
There are a couple of other possibilities that I have dismissed out of hand:
(i) that the editors thought that this paper had some value and (ii) linguistic
self loathing has become so strong that anything that craps on our discipline
is worthy of publication precisely because it dumps on us. As I said, I am
putting these terrifying possibilities aside.
[3]
Thus, when I say ‘UG’ I intend GB’s version thereof.
[4]
Bounding theory cares too (NP is, but VP is not a bounding node). APL discusses
island effects and I discuss their points below. However, suffice it to say, if
we need something like a specification of bounding nodes that we need to know,
among other things, which groups are Nish and which not.
[5]
X’ theory will project the category of the head of a phrase to the whole
phrase. But what makes something an NP requiring case is that N heads it.
[6]
APL also seems to believe that unless the same categories obtain cross
linguistically they cannot be innate (c.f. p. 11). This confuses Greenberg’s
conception of universals with Chomsky’s, and so is irrelevant. Say that the following principle “words that
denote events are grouped as V” is a prior that can be changed given enough
data. This does not imply that the acquisition of linguistic categories can
proceed in the absence of this prior. Such a prior would be part of UG on
Chomsky’s conception, even if not on Greenberg’s.
[7]
It’s a little like saying that you can get to New York using a good compass
without specifying any starting point. Compass readings are great, but not if
you don’t know where you are starting from.
[8]
Just to further back-ground the info (4) embeds the relative clause within know, which treats the embedded
information as pre-supposed.
Monday, June 17, 2013
May at the NSF
At the end of May I attended a terrific workshop organized
by David Poeppel for the NSF. Here’s a link to the roster and the
background papers we were given. The confab was fun, largely because the free
wheeling discussion was based on some uncontroversial givens, first and
foremost among these being that all accepted that something like Marr’s view
was a reasonable idealization of the kinds of levels required to link brain and
mind. In particular, all bought into the view that cognitive neuroscience needs
good high level descriptions of the computational competences that the brain
has in order to understand how it is organized at the neural level.[1] Indeed, Randy Gallistel (to only a very few
grumbles) proposed a much stronger version of the Marr thesis: that brain
architectures cannot be fruitfully studied at
all in the absence of good computational level accounts. It was further
accepted that models of linguistic competence of the generative variety are
paradigmatic instances of such Marrian computational level theories.
So, given this wonderful coming together of minds, what did
I learn? Here are some recollected personal highlights.
First, We don’t actually know much about the neural bases of
mental computation and so it is unreasonable to give neuronal level accounts a
privileged status. One of our
pre-conference readings (c.f. Mausfield here) has the following juicy quote:
Given that we presently know next
to nothing about the physical principles underlying mental phenomena and
achievements, there is no reason to assign the level of neurons a privileged
explanatory role. (p.4)
This was not a point avidly disputed by the neuroscientists
present, despite the acknowledgment all round that neuroscience has made
impressive progress over the last 25 years.
This is not incompatible with the realistic appraisal that there is a
very long way to go and that, at this
time, what we know about the brain offer few constraints on possible
computational level theory (e.g. on linguistic proposals about the structure of
FL/UG). We might wish that things were different (I know that I do) but they
aren’t.
Evidence? Well, here’s one: it seems that for much simpler
systems e.g. C. elegans with all of
302 neurons (whose wiring we know), why it does what it does is still a
mystery. As Mausfield observed: “In the
case of C.elegans, the complete
knowledge of the components of its biological hardware would constitute a
particularly favorable situation for understanding its complex behavior…”
Nonetheless, Mausfield quotes a recent review that notes that despite knowing
all we might wish to know about the 302 neurons in these nematode brains these
reductive efforts have proven quite unsatisfactory: “C.elegans responds
behaviorally to the presence or absence of food in a plethora of
ways…Surprisingly little progress has been made in understanding these
responses” (p.3 note 2)).
Second, this relative ignorance is nothing new. It seems that the goal of understanding
mental phenomena in neurological terms has been a long-standing project, at
least since the 18th century. In other
words, this is not a bold new
surprising thesis, despite what some hyperventilating philosophers might
suggest. Again as Mausfield put it:
For over 200 years, the premise
that mental processes must be considered a function of the brain has been more
or less commonplace. This has deluded us into overlooking the fact that…our
theoretical understanding is next to nil of what exactly…this function might
actually be taken to be. (p.3)
It seems that Priestly (of chemical fame) already thought
that this was the obvious scientific position to take (LaMettrie preceded him
by a century or so). However, Priestly was considerably more modest than many
of our current neuro-philosophers. His position was described as follows by the
London Encyclopedia (1829):
Dr Priestly apprehends that
sensation and thought necessarily result from the organization of the brain…but
he professes to have no idea at all of the manner in which the power of
perception results from organization and life.[2]
A becoming modesty, I think!
Third, historically, computational level theories have
generally laid the groundwork for neural explorations rather than brain
functions constraining higher level accounts. Once again here’s Mausfield:
…advances in our psychological understanding
of perceptual phenomena have in the first place benefited and fostered neurophysiology
rather [than-sic] the other way around. (p.2)
The implications of this for Generative Grammar (GG) are
pretty clear. GGs provide computational level accounts of the mental powers of
native speakers; a descriptively adequate grammar of L describing the mental
states of a competent speaker/hearer of L and an explanatory adequate theory of
L describing how L derives from FL/UG given the PLD of L. These computational level accounts, one hopes,
will serve to guide neuro-scientific research. How? Well, in much the way that
Barlow (quoted in Mausfield) envisioned for work in the psychology and
physiology of perception:
As to the claim that a theoretical
understanding of visual perception derives from neurophysiological
investigations, Barlow (1983, p.11) emphasized: “Nothing could be more
misleading, for all the important properties of the visual system were first
established by psychophysical and psychological observations made on the system
working as a whole. […] physiologists need to be told what the visual system
does before they can set about the difficult task of finding out how it does
it. (p.2)
Substitute FL for ‘the visual system’ above and you have
more or less the current state of play in the cognitive neuroscience of
language, at least as seen by the cohort of people that David managed to get to
sit down together to discuss these matters.
Fourth, to the problem that Poeppel and Embick (P&E) (here
and
here) identified as the “granularity mismatch problem” is still with
us. In my presentation, I suggested that
one of the virtues of the Minimalist Program (MP) is that it offers a way of
bridging the divide that P&E identify. In particular, if we can really
unify grammatical phenomena and reduce them to a common Merge-like core, then
this will provide a convenient target for neurophysiological investigation:
find a Merge-like circuit.[3]
Thus, the grammatical project described (e.g. here) was
not dismissed as irrelevant to finding ways of incarnating minds in brains.
Fifth, I got a great peek into how neuro/psycho types are
thinking of basic operations in their respective domains, c.f. Dave Heeger (here)
and Greg Hickok (here)
for a pretty good taste of what they are doing. Dave Heeger’s talk had, what to
my ear, was a real minimalistic theme: how “ set of canonical neural
computations” shared across different “brain regions and modalities” could
apply “similar operations to different problems” (51). In the best of all
possible worlds, we would love to find something similar in cognitive domains
including language. Greg’s talk showed how we might integrate higher level
psych-linguistic descriptions of speech, with lower level articulatory motor
control. What to me was very exciting
were the analogies between headedness in syntax and similar notions in motor
plans, which appear to have similar kinds of structures. All of this was very speculative, and hence extremely
interesting.
Sixth, Elisa Newport gave a fascinating presentation
focusing on brain specialization for speech. She made (at least) two
fascinating points. First, she observed that brains in which the language areas
are compromised can redirect this function to other parts. However, not to just any other part. Rather, a
brain can redirect linguistic capacity from a left language impaired hemisphere
to the very same place in the other hemisphere. This suggests two things: (i)
that either hemisphere can adequately subserve language and (ii) that the same
regions in both hemispheres are particularly well suited (specialized?) for the
kinds of operations language demands. In other words, though young brains
rearrange the cognitive furniture, not all parts of the brain are equally adept
at filling in for missing linguistic capacity, viz. brains are labile but not
arbitrarily so. Second, Elisa asked a terrific question: why is the language
area located where it is? And she shot down one plausible answer: it sits
between the perceptual regions that care about audition and the motor regions that
move lips and tongue. This, Elisa noted, cannot be the whole answer for the
exact same regions subserve ASL speakers (a point also made by Helen Neville),
and ASL speakers don’t much worry about audition or lip/tongue movement.
Last, Bob Berwick gave a great presentation on genetic
differences between us and our Neanderthal cousins. The main finding is that we are almost
identical! There were very very few differences. In effect, the minimalist
assumption (that whatever happened that allowed language to emerge was both
rapid and genetically “minor”) seems on the right track. I am in the process to
trying to convince Bob to post on this, so stay tuned.
There was much much more; great presentations, excellent
lunches, a terrific supper, lots of discussion, jokes, arguments, speculations
and general good cheer. However, most heartening of all was the realization
that there is a reasonable group of people out there that have no problem with
the standard (Chomskyan) Generative view that linguistics is an important part
of the cognitive neurosciences. I can only only hope that this view becomes
even more widespread. If it does, it suggests the coming of a golden age.
[1]
Bob Berwick made the reasonable point that restricting matters to three levels is likely a radical
simplification, apparent if one considers how many levels computer engineers
postulate to get one from programming languages to machine code. So, the Marr
perspective is best stated that there are at
least three levels worthy of serious consideration.
[2]
Quoted in Mausfield p.3.
[3]
I actually proposed that we should look for two circuits: one that combines
elements and one that labels the resulting combination. The former is plausibly
cognitively generic while the second is my candidate for the real distinctive
linguistically special operation.
However, the logic of MP does not require that my specific proposal be
the right one (though, of course, I have no doubt that it is).
Monday, November 12, 2012
Publication Blues
Get a bunch of syntacticians in a room and it’s not long
before they begin to regale each other (and themselves) with titillating tales
of publishing porn:
Do you know that it took me two and
half years to get that paper into LI?
You should see these absurd reviews that I got from NLLT, two say publish and one guy says reject and the editor
doesn’t see that his comments are full of S*&$. Language why would anyone publish there? They hate theoretical linguistics!! Cognition got the people I was
criticizing to review the paper and they sank it! I got a review from Syntax longer than the paper I
submitted, and most of it just missed the point.
You can add your own favorite vignettes, I am sure. What surprised me this weekend is to discover
that publishing porn is not restricted to syntax or even linguistics but
extends quite broadly across academia (see below). I had always thought that the
disgruntlement arose because of the dearth of journal outlets for publishing in
syntax. The big three -LI, NLLT and Syntax- have relatively few pages
between them. Language doesn’t
(won’t?) publish hard-core theoretical syntax for love or money (btw, this is
quite odd as the professional society journals in economics, philosophy,
physics etc. are where the newest stuff is
showcased. Language, in contrast, is
the last place to look for new cutting edge syntax (or, from the little that I
can tell, semantics or phonology)). Limited pages probably matter.
Page limits also make editors lives more troublesome. To
live with the restrictions the editorial decision is not ‘take the good ones,
reject the bad’ but ‘take these good papers and not those.’ Couple this with
the fact that IMHO there is a lot of good syntax being done and the result is that there have to
be arbitrary ways of managing the flow.
One way of doing this is restricting submissions (e.g. the
one-paper-under-review-at-a-time policy at LI).
Another is queu management (e.g. the long lag time between
submission-review-resubmission-re-review-…-acceptance-publication). A third is
content management (e.g. the dearth of debate in the pages of the journals, a
policy that frees up space).
Editors also face the “herding cats” problem. I don’t know about you, but reviewing papers
is not my favorite pastime. I know
that this is important and a contribution to the field and important for
people’s careers yada, yada, yada. But…you can fill in the rest, I
suspect. Part of this results from the
expectations people have of reviews.
They are “supposed” to be long and very detailed and very thorough. Even
typos! Rather than expecting a judgment about the relevance or importance of
the paper with a review of the central argument (something that would be about
1-2 pages long), the reviewer is expected to effectively write a reply,
moreover a reply that won’t be
published. There is, I am sure, a lot of
interesting discussion being carried out furtively in the review process that
deals with interesting (because though contentious) issues that really
should be public but never will be.
To
add to the burden, papers are typically very long. In the past natural
experiments occasionally emerged that permitted one to measure paper-bloat (this
is no longer possible as most journals will not publish previously aired work,
even conference proceedings). Sometimes
a paper that first appeared as a conference proceeding was subsequently
reissued in more elaborated form as a journal article. To my hazy recollection, the more detailed
paper often failed to convey the basic idea as clearly as the 15 page
conference paper did. There is a kind of
“defensive” linguistics that lards many papers, no doubt the result of the
thorough vetting process, which serves to obscure the main idea. It sometimes seems
like a paper cannot simply leave a problem knowingly unsolved and still expect
the paper to make the printed page. This
results in “explanation creep” that often muddies a relatively clear main
theme, to the detriment of the insight the paper offers.
There is one final problem, though I cannot tell how endemic
it is. I have occasionally gotten the
impression that reviewers like to make sure that papers that gore their
favorite oxen don’t see the light of publishing day. I cannot tell how widespread it is, but I
doubt that it is negligible. Let me
relate an anecdote: I was once asked to review a MS for a pretty prestigious
press. I read the MS and found myself unconvinced. However, I also concluded
that the direction being taken, though not one that I found congenial, represented
a common perspective was well done given this perspective and deserved an
airing. My review said as much, as well as including some more detailed
critical points mainly for the author’s interest. I was contacted by the
editors and asked if I thought the MS deserved publication. As my first sentence
was something to the effect that this MS should be published as it represented
an important perspective on timely matters I was a bit surprised. I was told
that it was clear that I disagreed with the ideas and remained unconvinced by
the arguments so why would I recommend publication? I found this odd. Is the standard of
publication whether the MS persuades the reviewer that it is true? I hope not. That is a very high bar. Interesting,
ok, provocative, sure, makes you think, check.
But true? Really true? Nope, too
much to ask.
So given all of this what would gruntle me? I really don’t know, which is partly why I am
writing about this here; you know, generate some chatter about the problem, discover that my mood is largely dyspepsia and that I should stop with the fried chicken
and anchovy pizza. Here are some minor
proposals:
·
Limit all paper submissions to NELS length 15
2-space pages
·
Limit all reviews to 2 pages
·
Penalize late reviewers (say, deny publication
in the journal for some period)
·
Consider ways of migrating all journals to the web where page numbers won’t matter and do so
in a PLOS like open format
·
Have a more active editorial board, one that
solicits and recommends MS for publication rather than just reviewing
adventitious submissions
These are relatively conservative changes that could be
implemented relatively quickly.
I would like to end with a few words on a more radical
proposal that caught my eye. It’s called
“A Rant on Refereeing.” It suggests that we dump reviewing altogether as it has
become a way of stifling the circulation of ideas rather than promoting them.
There are several linked papers (here, here, here) that are also worth reading,
some in favor some not. The main idea is that web based circulation of papers
should replace the refereed journal format and that there are ways of managing
the downsides (how to assure quality (answer: no magic bullet even in
journals), how to finesse the promotion and tenure role of reviewed
publications (answer: letters will be more
important), how readers will find what’s “important” (answer: not easy but
aggregate sites will naturally arise to organize the paper flow) etc. It is very provocative and if you are
interested in these issues I suggest you take a look.
Ok, enough: I am curious about what others think about these
things, so if you are inclined, let me know.
Subscribe to:
Posts (Atom)