Comments

Showing posts with label laws of grammar. Show all posts
Showing posts with label laws of grammar. Show all posts

Sunday, May 25, 2014

The GG game: Plato, Darwin and the POS

Alex Clark has made the following two comments (abstracted) in his comments to this post.

I find it quite frustrating that you challenge me to "pony up a story" but when pressed, you start saying the MP is just a conjecture and a program and not a theory.

So I read the Hauser et al paper where the only language specific bits are recursion and maps to the interfaces -- so where's the learning story that goes with that version of UG/FLN? Nobody gives me a straight answer. They change the subject or start waffling about 3rd factor principles.

I believe that these two questions betray a misunderstanding, one that Alex shares with many others concerning the objectives of the Minimalist Program (MP) and how they relate to those of earlier theory. We can address the issue by asking: how does going beyond explanatory adequacy relate to explanatory adequacy?  Talk on the Rialto is that the former cancels the latter. Nothing could be further from the truth. MP does not cancel the problems that pre-MP theory aimed to address. Aspiring to go beyond explanatory adequacy does not amnesty a theory from explanatory adequacy. Let me explain.

Before continuing, however, let me state that what follows is not Chomsky exegesis.  I am a partisan of Chomsky haruspication (well not him, but his writings), but right now my concern is not to scavenge around his literary entrails trying to find some obscure passage that might, when read standing on one’s head, confuse. I am presenting an understanding of MP that addresses the indicated question above. The two quoted paragraphs were addressed to (at?) me. So here is my answer. And yes, I have said this countless times before.

There are two puzzles, Plato’s Problem (PP) and Darwin’s Problem (DP).  They are interesting because of the light they potentially shed on the structure of FL, FL being whatever it is that allows humans to be as linguistically facile as we are.  The work in the last 60 years of generative grammar (GG) has revealed a lot about the structure of FL in that it has discovered a series of “effects” that characterize the properties of human Gs (I like to pretentiously refer to these as “laws of grammar” and will do so henceforth to irritate the congenitally irritated). Examples of the kinds of properties these Gs display/have include the following: Island effects, binding effects, ECP effects, obviation of Island effects under ellipsis, parasitic gap effects, Weak and Strong Crossover effects etc. (I provided about 30 of these effects/laws in the comments to the above mentioned post, Greg K, Avery and others added a few more).  To repeat again and loudly: THESE EFFECTS ARE EMPIRICALLY VERY WELL GROUNDED AND I TAKE THEM TO BE ROUGHLY ACCURATE DESCRIPTIONS OF THE KIND OF REGULARITIES THAT Gs DISPLAY AND I ASSUME THAT THEY ARE MORE OR LESS EMPIRICALLY CORRECT.  They define an empirical domain of inquiry. Those who don’t agree I consign to the first circle of scientific hell, the domicile of global warming skeptics, flat earthers and evo deniers. They are entitled to their views, but we are not required (in fact, it is a waste of time) to take their views seriously. So I won’t. 

Ok, let’s assume that these facts have been established. What then? Well, we can ask what they can tell us about FL. IMO, they potentially tell us a lot. How so? Via the POS argument. You all know the drill: propose a theory that derives the laws, take a look at the details of the theory, see what it would take to acquire knowledge of this theory which explains the laws, see if the PLD provides sufficient relevant information to acquire this theory. If so, assume that the available data is causally responsible.[1] If not assume that the structure of FL is causally responsible.  Thus, knowledge of the effects is explained by either pointing to the available data that it is assumed the LAD tracks or by adverting to the structure of LAD’s FL. Note, it is critical to this argument to distinguish between PLD and LD as the LAD has potential use of the former while only the linguist has access to the latter. The child is definitely not a little linguist.[2]

All of this is old hat, a hat that I’ve worn in public on this blog countless times before and so I will not preen before you so hatted again.  What I will bother saying again is that this can tell us something about FL. The laws themselves can strongly suggest whether FL is causally responsible for this or that effect we find in Gs. They alone do not tell us what exactly about FL is responsible for this or that effect. In other words, they can tell us where to look, but they don’t tell us what lives there.

So, how does one go from the laws+POS to a conjecture/claim about the structure of FL? Well, one makes a particular proposal that were it correct would derive the effects. In other words, one proposes a hypothesis, just as one does in any other area of the sciences. P,V,T relate to one another via the gas laws. Why? Well maybe it’s because gases are made up of small atoms banging against the walls of the container etc. etc. etc.  Swap gas laws for laws of grammar and atomic theory for innately structured FL and off we go.

So, what kinds of conjectures have people made? Well, here’s one: the principles of GB specify the innate structure of FL.[3] Here’s why this is a hypothesis worth entertaining: Were this true then it would explain why it is that native speakers judge movement out of islands to be lousy and why they like reflexivization where they dislike pronominalization and vice versa. How does it explain these laws? As follows: if the principles of GB correctly characterize FL, then in virtue of this FL will yield Gs that obey the laws of grammar.  So, again, were the hypothesis correct, it would explain why natural languages adhere to the generalizations GG has discovered over the last 60 years.[4]

Now, you may not like this answer. That’s your prerogative. The right response is to then provide another answer that derives the attested effects.  If you do, we can consider this answer and see how it compares with the one provided. Also, you might like the one provided and want to test it further. People (e.g. Crain, Lidz, Wexler, a.o.) have done just that by looking at real time acquisition in actual kids.  At any rate, all of this seems perfectly coherent to me, and pretty much standard scientific practice. Look for laws, try to explain them.

Ok, as you’ve no doubt noticed, the story told assumes that what’s in FL are principles of GB.[5] Doesn’t MP deny this? Yes and No. Yes, it denies that FL codes for exactly these principles as stated in GB. No, it assumes that some feature of FL exists from which the effects of these principles follow. In other words, MP assumes that PP is correct and that it sheds light on the structure of FL. It assumes that a successful POS argument implies that there is something about the structure of the LAD that explains the relevant effect. It even takes the GB description of the effects to be extensionally accurate. So how does it go beyond PP?

Well, MP assumes that what’s in FL does not have the linguistic specificity that GB answers to PP have. Why?

Well, MP argues that the more linguistically specific the contents of FL, the more difficult it will be to address DP. So, MP accepts that GB accurately derive the laws of grammar but assumes that the principles of GB themselves follow from yet more general principles many of which are domain general so as to be able to accommodate DP in addition to PP.[6] That, at least, is the conjecture. The program is to make good on this hunch. So, MP assumes that the PP problem has been largely correctly described (viz. that the goal is to deduce the laws of grammar from the structure of FL) but that the fine structure of FL is not as linguistically specific as GB has assumed.  In other words, that FL shares many of its operations and computational principles with those in other cognitive domains. Of course, it need not share all of them. There may be some linguistically specific features of FL, but not many. In fact, very very few. In fact, we hope, maybe (just maybe, cross my fingers) just ONE.

We all know the current favorite candidate: Merge. That’s Chomsky’s derby entry. And even this, Chomsky suggests may not be entirely proprietary to FL. I have another, Label. But really, for the purposes of this discussion, it doesn’t really matter what the right answer is (though, of course I am right and Chomsky is wrong!!).

So, how does MP go beyond explanatory adequacy? Well, it assumes the need to answer both PP and DP. In other words, it wants the properties of FL that answer PP to also be properties that can answer DP. This doesn’t reject PP. It doesn’t assume that the need to show how the facts/laws we have discovered over 60 years follow from FL has all of a sudden gone away. No. It accepts PP as real and as described and aims to find principles that do the job of explaining the laws that PP aims to explain but hopes to find principles/operations that are not so linguistic specific as to trouble DP.

Ok, how might we go about trying to realize this MP ambition (i.e. a theory that answers both PP and DP)? Here’s a thought: let’s see if we can derive the principles of GB from more domain general operations/principles.  Why would this be a very good strategy? Well because, to repeat, we know that were the principles of GB innate features of FL then they would explain why the Gs we find obey the laws of grammar we have discovered (see note 6 for philo of science nostrums). So were we able to derive GB from more general principles then these more general principles would also generate Gs that obeyed the laws of grammar. Here I am assuming the following extravagant rule of inference: if AàB and BàC then AàC.  Tricky, eh? So that’s the strategy. Derive GB principles from more domain general assumptions.

How well has MP done in realizing this strategy. Here we need to look not at the aims of the program, but at actual minimalist theories (MT). So how good are our current MT accounts in realizing MP objectives? The answer is necessarily complicated. Why? Because many minimalist theories are compatible with MP (and this relation between theory and program holds everywhere, not just in linguistics). So MP spawns many reasonable MTs. The name of the game if you like MP is to construct MTs that realize the goals of MP and see whether you can get them to derive the principles of GB (or the laws of grammar that GB describes). So, to repeat, how well have we done?

Different people will give different answers. Sadly, evaluations like these require judgment and reasonable people will differ here. I believe that given how hard the problems are, we have done not bad/pretty well for 20 years of work. I think that we have pretty good unifications of many parts of GB in terms of simpler operations and plausibly domain general computational principles. I have tried my own hand at this game (see here). Others have pursued this differently (e.g. Chomsky). But, and listen closely here, MP will have succeeded only if whatever MT it settles on addresses PP in the traditional way.  As far as MP is concerned, all the stuff we thought was innate before is still innate, just not quite in the particular form envisaged. What is unchanged is the requirement to derive the laws of grammar (as roughly described by GB). The only open question for DP is whether this can be done using domain general operations/principles with (at most) a very small sprinkling of domain specific linguistic properties. In other words, the open question is whether these laws are derived directly from principles of GB or indirectly from them (think GB as axioms vs GB as theorems of FL). 

I should add that no MT that I know of is just millimeters away from realizing this MP vision.  This is not a big surprise, IMO. What is a surprise, at least to me, is that we’ve made serious progress towards a good MPish account.  Still, there are lots of domain specific things we have not been able to banish from FL (ECP effects, all those pesky linguistic features (e.g. case), the universal base (and if Cinque is right, it’s a hell of a monster) and more). If we cannot get rid of them, then MP will only be partly realized. That’s ok, programs are, to repeat, not true or false, but fecund or not. MP has been very fertile and we (I?) have reason to be happy with the results so far, and hopeful that progress will continue (yes, I have a relentlessly sunny and optimistic disposition).

With this as prologue, let’s get back to Alex C. On this view, the learning story is more or less the one we had before. MP has changed little.[7] The claim that the principles of GB are innate is one that MP can endorse (and does, given the POS arguments). The question is not whether this is so, but whether the principles themselves are innate or do they derive from other more general innate principles. MP bets on the second. However, MP does not eschew the conclusion that GB (or some equivalent formulation) correctly characterizes the innate structure of FL. The only question is how direct these principles are instantiated, as axioms or as theorems. Regardless of the answer, the PP project as envisioned since the mid 60s is unchanged and the earlier answers provided still quite viable (but see caveat in note 7).

In sum, we have laws of grammar and GB explanations of them that, via the POS, argue that FL has GBish structure. MP, by adding DP to the mix, suggests that the principles of GB are derived features of FL, not primitive.  This, however, barely changes the earlier conclusions based on POS regarding PP. It certainly does not absolve anyone of having to explain the laws of grammar. It moreover implies that any theory that abstracts away from explaining these laws is a non-starter so-far as GG is concerned (Alex C provides a link to one such theory here).[8]

Let me end: here’s the entrance fee for playing the GG game:
1.     Acceptance that GG work over the last 60 years has identified significant laws of grammar.
2.     Acceptance that a reasonable aim of research is to explain these laws of grammar. This entails developing theories (like GB) which would derive these laws were these theories true (PP).
3.     More ambitiously, you can add DP to the mix by looking for theories using more domain general principles/operations from which the principles of GB (or something like them) follow as “theorems,” (adopting DP as another boundary condition on successful theory).

That’s the game. You can play or not. Note that they all start with (1) above. Denial that the laws of grammar exist puts you outside the domain of the serious. In other words, deny this and don’t expect to be taken seriously. Second, GG takes it to be a reasonable project to explain the laws of grammar and their relation to FL by developing theories like GB. Third, DP makes step 2 harder, but it does not change the requirement that any theory must address PP. Too many people, IMO, just can’t wrap their heads around this simple trio of goals. Of course, nobody has to play this game. But don’t be fooled by the skeptics into thinking that it is too ill defined to play. It’s not. People are successfully playing it. It’s just when these goals and ambitions are made clear many find that they have nothing to add and so want to convince you to stop playing. Don’t. It’s really fun. Ignore their nahnahbooboos.

[1] Note that this does not follow. There can be relevant data in the input and it may still be true that the etiology of the relevant knowledge traces to FL. However, as there is so much that fits POS reasoning, we can put these effects to the side for now
[2] One simple theory is that the laws themselves are innate. So, for example, one might think that the CNPC is innate. This is one way of reading Ross’s thesis. I personally doubt that this is right as the islands seem to more or less swing together, though there is some variation. So, I suspect that island effects themselves are not innate though their properties derive from structural properties of FL that are, something like what Subjacency theory provides.
[3] As many will no doubt jump our of their skins when they encounter this, let me be a tad careful. Saying that GB is innate does not specify how it is thus.  Aspects noted two ways that that this could be true: GB restricts the set of admissible hypotheses or it weights the possible alternative grammars/rules by some evaluation measure (markedness). For current purposes, either or both are adequate. GB tended to emphasize the restrictive hypothesis space, Ross, for example, was closer to a theory of markedness.
[4] Observe: FL is not itself a theory of how the LAD acquires a G in real time. Rather it specifies, if descriptively adequate, which Gs are acquirable (relative to some PLD) and what properties these Gs will have.  It is reasonable to suppose that what can be acquired will be part of any algorithm specifying how Gs get acquired, but they are not the same thing.  Nonetheless, the sentence that this note is appended to is correct even in the absence of a detailed “learning theory.”
[5] None of the above or the following relies on it being GB that we use to explain the laws. I happen to find GB a pretty good theory. But if you want something else, fine. Just plug your favorite theory in everywhere I put in ‘GB’ and keep reading.
[6] Again this is standard scientific practice: Einstein’s laws derive Newton’s. Does this mean that Newton’s laws are not real? Yes and No. They are not fundamental, but they are accurate descriptions. Indeed, one indication that Einstein’s laws are correct is that they derive Newton’s as limit cases. So too with statistical mechanics and thermodynamics or quantum mechanics and classical mechanics.  That’s the way it works. Earlier results (theory/laws) being the target of explanation/derivation of later more fundamental theory.
[7] The one thing it has changed is resurrect the idea that learning might not be parameter setting. As noted in various posts, FL internal parameters are a bit of a bother given MP aims. So, it is worth considering earlier approaches that were not cast in these terms, e.g. the approach in Berwick’s thesis.
[8] It’s oracular understanding of the acquisition problem simply abstracts away from PP, as Alex D noted. Thus, it is without interest for the problems discussed above.

Friday, January 18, 2013

Effects, Phenomena and Unification


In the previous post, I mentioned that there is a general consensus that UG has roughly the features described in GB. In the comments, Alex, quotes Cederic Boeckx as follows and asks if Cederic is “a climate change denier.”

I think that minimalist guidelines suggest an architecture of grammar that is more plausible biologically speaking that a fully specified, highly specific UG – especially considering the very little time nature had to evolve this remarkable ability that defines our species. If syntax is at the heart of what had to evolve de novo, syntactic parameters would have to have been part of this very late evolutionary addition. Although I confess that our intuitions pertaining to what could have evolved very rapidly are not as robust as one would like, I think that Darwin’s Problem (the logical problem of language evolution) becomes very hard to approach if a GB-style architecture is assumed.

The answer is no, he is not (but thanks for asking). I’ll explain why but this will involve rehearsing material I’ve touched upon elsewhere so if you feel you already know the answer please feel free to go off and do something more worthwhile.

My friends in physics (remember, I am a card carrying hyper-envier) make a distinction between effective and fundamental theories.  Effective theories are those that are phenomenologically pretty accurate. They are also the explananda for fundamental theories.  Using this terminology, GB is an effective theory, and minimalism aspires to develop a fundamental theory to explain GB “phenomena.” Now, ‘phenomena’ is a technical term and I am using it in the sense articulated in Bogen and Woodward (here). Phenomena are well-grounded significant generalizations that form the real data for theoretical explanation. Phenomena are often also referred to as ‘effects.’ Examples in physics include the Gas Laws, the Bernoulli effect, black body radiation, Doppler effects, the photoelectric effect etc.  In linguistics these include island effects, principle A, B and C effects, weak and strong crossover effects, the PRO theorem, Superiority effects etc. GB theory can be seen as a fairly elaborate compendium of these. Thus, the various modules within GB elaborate a series of well-massaged generalizations that are largely accurate phenomenological descriptions of UG. I have at times termed these ‘Laws of Grammar,’ (said plangently you can sound serious, grown-up and self-important) to suggest that those with minimalist aspirations should take these as targets of explanation.  Thus, in the requisite sense, GB (and its cousins described in the last post) can serve as an effective theory, one whose generalizations a minimalist account, a fundamental theory, should aim to explain. 

I hope it is clear how this all relates to the Cedric quote above, but if not here’s the relevance.  Cedric rightly observes that if one is interested in evolutionary accounts then GB cannot be the fundamental theory of linguistic competence.  It’s jus appears as too complex, all that internal modularity (case and theta and control and movement and phrase structure), all those different kinds of locality conditions (binding domains and subjacency/phase and minimality and phrasal domains of a head and government) all those different primitives (case assigners, case receivers, theta markers, arguments, anaphors, bound pronouns, r-expressions, antecedents etc., etc., etc.).  Add to this that this thing popped out in such a short time and there really seems no hope for a semi-reasonable (even just-so) story.  So, GB cannot be fundamental.  BTW, I am pretty sure that I have interpreted Cedric correctly here for we have discussed this a lot over the last five to ten years on a pretty regular basis.

Given the distinction of GB as effective theory and MP as aiming to develop a fundamental theory, how should a thoroughly modern minimalist proceed? Well, as I mentioned before (here) one model is Chomsky’s unification of Ross’s islands via subjacency.  What Chomsky did was (i) treat Ross’s descriptions as effective and (ii) propose how to derive these on more empirically, theoretically and computationally more natural grounds. Go back and carefully read ‘On Wh-Movement’ and you’ll see that how these various strands combine in his (to my taste buds) rather beautiful account. Taking this as a model, a minimalist theory should aspire to the same kind of unification. However, this time it will be a lot harder. For two main reasons.

First, what MP aspires to unify have been thought to be fundamentally different from “the earliest days of generative grammar” (two points and a bonus questions to anyone who identifies the source of this quote). Unifying movement, binding and control goes against the distinction between movement and construal that has been a fundamental part of every generative approach to grammar since Aspects (and before, actually), as has been the distinction between phrase structure and movement. However, much minimalist work over the last 20 years can be seen as chipping away at the differences. Chomsky’s 1993 unification of case as a species of movement or Probe-Goal licensing (PGL), the assimilation of control to a species of movement (moi) or PGL (Landau), reflexive licensing as a species of movement (Idsardi and Lidz, moi) or PGL (Reuland), the collapsing of phrase structure and movement as species of E/I merge, the reduction of Superiority effects to movement via minimality. All of these are steps in reducing the internal modularity of GB and erasing the distinctions between the various kinds of relationships described so well in GB. This unification, if it can be pulled off (and showing that it might be has been, IMO, the distinctive contributions of MP), would do for GB what Chomsky did for islands and the resultant theory would have a decent claim to being fundamental.

The second hurdle will be articulating some notion of computational complexity that makes sense. In ‘On Wh-Movement,’ Chomsky tried to suggest some computational advantages of certain kinds of locality considerations.  Whatever, his success, the problem of finding reasonable third factor features with implications for linguistic coding is far more daunting, as I’ve discussed in other posts. The right notion, I have suggested elsewhere, will reflect the actual design features of the systems that FL interact with and use it. Sadly, we know relatively little about interface properties (especially CI) and we know relatively little about how FL would fit in with other cognitive modules. We know a bit more about the systems that use FL and there have been some non-trivial results concerning what kinds of considerations matter. As I have discussed this in other posts, I will not burden you with a rehash (see here and here). Consequently, whatever is proposed is very speculative, though speculation is to be encouraged for the problem is interesting and theoretically significant.  This said, it will be very hard and we should appreciate that.

So, is Cedric a denier? Nope. He accepts the “laws of grammar” as articulated in GB as more or less phenomenologically correct. Is his strategy rational? Yup. The aim should be to unify these diverse laws in terms of more fundamental constructs and principles. Are people who quote Cedric to “épater les Norberts” doing the same thing? Not if they are UG deniers and not if their work does not aim to explain the phenomena/effects that GB describes. These individuals are akin to climate change deniers for their work has all the virtues of any research that abstracts away from the central facts of the matter. 

Wednesday, November 21, 2012

How I became a minimalist and why or What would GB say?


It was apparently Max Planck who discovered the unit time of scientific change to be the funeral (the new displacing the old one funeral at a time).  In the early 1990s, I discovered a second driving force, boredom.  As some of you may know, since about the mid-1990s I have been a minimalist enthusiast. For the record, I became one despite my initial inclinations. On first reading A minimalist program for linguistic theory (a Korean bootlegged version purportedly whisked of Noam’s desk and quickly disseminated), I was absolutely convinced that it had to be on the wrong track, if not the aspirations, then the tentative conclusions. I was absolutely certain that one of the biggest discoveries of generative grammar had been the centrality of government as a core relation and S-structure as the indispensible level (I can still see myself making just these points in graduate intro syntax). Thus the idea that we dispense with government as a fundamental relation (it’s called Government-Binding theory after all!), or that we eliminate S-structure as a fundamental level (D-structure, I confess, I was willing to throw under the bus) struck me as nuts, just another manoeuver by Chomsky to annoy former graduate students.

Three things worked together to open (more accurately, pry open) my mind.

First, my default strategy is to agree with Chomsky, even if I have no idea what he’s talking about. In fact, I often try to figure out where he’s heading so that I can pre-agree with him. Sadly, he tends not to run in a straight line so I can often be seen going left when he zags right or right when he zigs left. This has proven to be both healthful (I am very fit!) and fruitful. More often than not, Chomsky identifies fecund research directions, or at least ones that in retrospect I have found interesting.  No doubt this is just dumb luck on Chomsky’s part, but if someone is lucky often enough, it is worth paying very careful attention (as my mother says: “better lucky than smart”).  So, though I have often found my work at a slant (even perpendicular) to his detailed proposals (e.g. just look at how delighted Noam is with Movement Theory of Control, a theory near and dear to my heart), I have always found it worthwhile to try to figure out what he is proposing and why. 

Second, fear: when the first minimalist paper began to circulate in the early 1990s I was invited to teach a graduate syntax seminar at Nijmegen (populated by eager, smart, hungry (and so ill-tempered) grad students from Holland and the rest of Europe) and I needed something new to talk about. If you just get up and repeat what you’ve already done, they could be ready for you. Better to move in some erratic direction and keep them guessing. Chomsky’s recent minimalist musings seemed like perfect cover.

Third, and truth be told I believe that this is the main reason, the GB stuff I/we had been exploring had become really boring. Why? For the best of possible reasons: viz. we really understood what made GB style theories tick and we/I needed something new to play with, something that would allow me/us to approach old questions in a different way (or at least not put us/me to sleep). That new thing was the Minimalist Program. I mention this, because at the time there was a lot of toing and froing about why so many had lemming-like (this is apparently a rural legend; they don’t fling themselves off cliffs) jumped off of the GB bandstand and onto the minimalist bandwagon. As I faintly recall, there was an issue of the Linguistic Review dedicated to this timely question with many authoritative voices giving very reasonable explanations for why they were taking the minimalist turn.  And most of these reasons were in fact good ones. However, if my conversion was not completely atypical, the main thrust came from simple thasaphobia and the discovery of the well-established fact that intensive study of the Barriers framework could be deleterious to one’s health (good reason to avoid going there again all you phase-lovers out there!).

These three motivations joined to prompt me, as an exercise, to stow the skepticism, at least for the duration of the Dutch lectures, assume that this minimalist stuff was on the right track and see how far I could get with it.  Much to my surprise, it did not fall apart on immediate inspection (a surprisingly good reason to persist in my experience), it was really fun to play with, and, if you got with the program, there was a lot to do given that few GB details survived minimalism’s dumping of government as a core grammatical relation (not so surprising given that it is government-binding theory).  So I was hooked, and busy. (p.s. I also enjoyed the fact that, at the time, playing minimalist partisan could get one into a lot of arguments and nothing is more fun than heated polemics).

These were the basic causes for my theoretical conversion. Were there any good reasons? Yes, one.  Minimalism was the next natural scientific step to take given the success of the GB enterprise.

This actually became more apparent to me several years later, than it was on my road to Damascus Nijmegen.  The GB era produced a rich description of the structure of UG; internally modular with distinctive conditions, primitives and operations characterizing each sub-part.  In effect, GB delivered a dozen or so “laws” of grammar (e.g. subjacency, ECP, principles A-C of binding theory, X’-theory etc.), of pretty good (no, not perfect, but pretty good) empirical standing (lots of cross linguistic support). This put generative grammar in a position to address a new kind of question: why these laws and not others? Note: you can’t ask this question if there are no “laws.” Attacking it requires that we rethink the structure of UG in a new way; not only to ask “what’s in UG ?” but also “what that is in UG is distinctively linguistic and what traceable to more general powers, cognitive, computational, or physical?”. This put a version of what we might call Darwin’s Problem (the logical problem of language evolution) on the agenda along side Plato’s Problem (the logical problem of language acquisition).  The latter has not been solved, not by a long shot, but fortunately adding a question to the research agenda does not require that previous problems have been put to bed and snuggly tucked in. So though in one sense, minimalism was nothing new, just the next reasonable scientific step to take, it was also entirely new in that it raised to prominence a question whose time, we hoped, had come. [1]

Chomsky has repeatedly emphasized the programmatic aspects of minimalism.  And, as he has correctly noted, programs are not true or false but fecund or barren. However, after 20 years, it’s perhaps (oh what a weasel word!) time to sit back and ask how fertile the minimalist turn has been? In my view, very, precisely because it has spawned minimalist theories that advance the programmatic agenda, theories that can be judged not merely in terms of their fertility but also in terms of their verisimilitude. I have my own views about where the successes lie, and I suspect that they may not coincide with either Noam’s or yours.  However, I believe that it is time that we identified what we take to be our successes and ask ourselves how (or whether?) they reflect the principle ambitions and intuitions of the minimalist program.

Let me put this another way: in one sense minimalism and GB are not competitors for the aims of the former presuppose the success of the latter.  However, minimalist theories and GB theories often are (or can be) in direct competition and it is worth evaluating them against each other.  So for example, to take an example at random (haha!), GB has a theory of control and current minimalism has several. We can ask, for example: In what ways do the GB and minimalist accounts differ? How do they stack up empirically? What minimalist precepts do the minimalist theories reflect?  What GB principles are the minimalist accounts (in)compatible with? What larger minimalist goals do the minimalist theories advance?  What does the minimalist story tells us that the earlier GB story didn’t? And vice versa? Etc. etc. etc.

IMHO, these are not questions that we have asked often enough. I believe that we have failed to effectively use GB as the foil (and measuring rod) it can be. Why? I’m not sure. Perhaps because we have concluded that because the minimalist program is worth pursuing that specific minimalist theories that brandish distinctive minimalist technology (feature checking, merge, Agree, probe-goal architecture, phases etc.) are “better” or “truer” than those exploiting the quaint out of date GB apparatus.  If so, we were wrong.  We always need to measure our advances. One good way to do this is to compare your spanking new minimalist proposal with the model T GB version. I hereby propose that going forward we adopt the mantra “What would GB say?” (WWGBS; might even make for a good license plate) and compare our novel proposals with this standard to make clear to ourselves and others where and how we’ve progressed.

I will likely blog more on this topic soon and identify what I take to be some of the more interesting lines of investigation to date.  However, I am very interested in what others take the main minimalist successes to be.  What are the parade case achievements? Let me know. After 20 years, it seems reasonable to try to make a rough estimate of how far we’ve come.



[1] Here Sean Carroll goes minimalist in a different setting:
The actual laws of nature are interesting, but it’s also interesting that there are laws at all…We want to know what those laws are. More ambitiously, we’d like to know if those laws could possibly have been different…We may or may not be able to answer such a grandiose question, but it’s the kind of thing that lights the imagination of the working scientist (p.23)
This is what I mean by the next obvious scientific step to take.  First find laws, then ask why these laws and not others. That’s the way the game is played, at least by the real pros.