• Aucun résultat trouvé

Do animate arguments come first ?

N/A
N/A
Protected

Academic year: 2021

Partager "Do animate arguments come first ?"

Copied!
3
0
0

Texte intégral

(1)

HAL Id: hal-01451725

https://hal.archives-ouvertes.fr/hal-01451725

Submitted on 1 Feb 2017

HAL is a multi-disciplinary open access archive for the deposit and dissemination of sci-entific research documents, whether they are pub-lished or not. The documents may come from teaching and research institutions in France or abroad, or from public or private research centers.

L’archive ouverte pluridisciplinaire HAL, est destinée au dépôt et à la diffusion de documents scientifiques de niveau recherche, publiés ou non, émanant des établissements d’enseignement et de recherche français ou étrangers, des laboratoires publics ou privés.

Do animate arguments come first ?

Juliette Thuilier, Anne Abeillé, Benoît Crabbé

To cite this version:

(2)

Do animate arguments come first ?

Juliette Thuilier*, Anne Abeillé° and Benoît Crabbé* Univ Paris Diderot, Sorbonne Paris Cité

* ALPAGE (UMR INRIA -001) — ° LLF (UMR CNRS 7110) contact : jthuilier@linguist.jussieu.fr

Animacy is known to play a role in postverbal argument ordering for various languages (W02, B08a, B08b). Together with definiteness, pronominality, and shortness, animacy favors positioning an argument first. Nevertheless, it is sometimes difficult to distinguish linear ordering from grammatical function or semantic role assignment, since those are also subject to various factors including animacy (K77). In order to investigate the role of animacy in French argument ordering, we limit ourselves to complements of ditransitive verbs with object and indirect object nominal complements. Using statistics on treebanks for language production and a questionnaire study for language perception, we show that animacy seems to play no role in the relative ordering.

Ordering of postverbal NP and PP complements in French is known to display a preference for NP PP (ex1, B28) and to be sensitive to different factors~: length (2), definiteness (3)… Given the lack of treebanks for spontaneous spoken French, we randomly extracted 982 sentences from two newspaper corpora (Le MondeA03, Est Républicain) and a radio corpus (ESTER). We observe an average 58% preference for NP PP, with some variation between corpora (46.4% for Est Républicain), between verbs (37.8% for montrer - 'to show') and between prepositions (28.4% for de - 'of'). We annotated the sentences for relative length of NP and PP (log number of words), V-PP collocation, and complement animacy, definiteness and pronominality. We annotate Animacy into a binary variable (with ‘animate’ for 'human', 'animal' and 'organization').

We built a logistic regression model predicting either ordering (NP-PP or PP-NP), with corpus and verb lemma as intercept random effects. We observe two significant effects : relative length (p-value < 2e-16) and collocation (p-value = 0.039). Differently from English and German complement ordering (B08a, B08b), animacy shows no reliable preference for early position (AnimacyPP p-value = 0.95, AnimacyNP p-value = 0.91) as illustrated on the following figure : ANIM INANIM !"#$%&'()*#+&,-.&/0-&1%#&%#!"!""#)1%-23-0 100 200 300 400 500 ANIM INANIM !"#$%&'()*#+&,-.&/0-&1%#&%#/1-4#)1%-23-, 0 200 400 600 800 ANIM INANIM !!"#$%&'()"*%+,-%./,%0$"%$"!!!1!"(0$,23,+ 0 50 100 150 200 250 ANIM INANIM !!"#$%&'()"*%+,-%./,%0$"%$".0,4"(0$,23,+ 0 100 200 300 400 500 600

We then built a second model (CM) removing non significant factors by likelihood ratio test (!2 = 0.55), keeping only relative length (favoring short before long), and collocation

(favoring PP-NP). We computed variable interaction and only one was significant: collocation and NP non-animacy (p = 0.012), but with an effect contrary to what would be expected, collocation and NP non-animacy voting for NP-PP order.

In order to neutralize the relative length effect, and give animacy more chances to show up, we extracted 23 sentences from our corpora with complements of equal length (1) for our

(3)

questionnaire study. We tested 25 subjects for preferences between NP-PP and PP-NP continuations using a 5 point likert scale (with each order as an endpoint on the scale) and coded the results from 1 = strong PP-NP preference to 5 = strong NP-PP preference. The experiment confirms the overall preference for NP-PP order (with average rating 3.5). We built a linear mixed model to predict the ratings, with the same predicting variables as the corpus study (minus complement relative length, plus subject as random effect). In this model, NP definiteness becomes significant (p = 0.02), but pronominality and animacy effects remain non significant (animacyNP p = 0.9, animacyPP p = 0.5). Compared to English and German, the lack of pronominality effect can be explained by the fact that French has a different strategy (preverbal cliticization) for ordering pronominal arguments. But the lack of Animacy effect is a major surprise, which should be confirmed with other experiments and other constructions, but which nevertheless undermines its supposed universality for argument ordering.

Examples

(1) Pierre fonce dans la nuit porter la bonne nouvelle à sa fiancée (Est Républicain) ‘Pierre runs in the night to-bring the good news to his fiancee’

(2) verser au fisc les 4,80% droits d'enregistrement (Le Monde) ‘give to-the taxes the 4.8% registration rights’

(3) un administrateur provisoire qui proposera aux actionnaires différentes possibilités. (Le Monde)

‘a temporary admistrator who will-propose to-the shareholders different possibilities’

(4) la Banque fédérale d'Allemagne (…) permit de mettre en échec la spéculation. (Le Monde)

‘The German Federal Bank permitted to defeat the speculation’

!"#$%"&'()*$+&,-.$

Formula: ordre ~ lengthSN-lengthSP + collocatePP + (1 | corpus) + (1 | verbLemma) Data: snspoanim

AIC BIC logLik deviance

613.9 638.4 -302 603.9 Random effects:

Groups Name Variance Std.Dev.

verbLemma (Intercept) 1.78145 1.33471

corpus (Intercept) 0.17067 0.41312

Fixed effects:

Estimate Std. Error z value Pr(>|z|)

(Intercept) -1.2679 0.3238 -3.916 9e-05 ***

log(lengthSN)-log(lengthSP) 2.8432 0.2071 13.731 < 2e-16 ***

collocatePP 1.2192 0.4604 2.648 0.00809 **

References

A03 Abeillé, A. et al. . 2003. `Building a treebank for French', in A. Abeillé (ed) Treebanks, Kluwer. B28 Blinkenberg A, 1928, L’ordre des mots en français moderne, Høst, Copenhagen.

B08a Branigan, H, Pickering, M, Tanaka, M 2008. Contribution of animacy to grammatical function assignment and word order during production Lingua 118, 117-189,

B08b Bresnan J., Hay J. 2008. Gradient grammar: an effect of animacy on the syntax of give in New Zealand and American English, Lingua 118, 245-259.

K77 Keenan E, Comrie B, 1977. Noun phrase accessibility and universal grammar, LI 8, 63-99.

KG04 Kempen G, Harbusch K. 2004. A corpus study into word order variation in German subordinate clauses: Animacy affects linearization independently of grammatical function assignment In T. Pechmann, & C. Habel (Eds.), Multidisciplinary approaches to language production (pp. 173-181). Berlin: Mouton de Gruyter. W02 Wasow T. 2002. Postverbal behaviour, CSLI Publications.

Z04 Zaenen et al. 2004. Animacy Encoding in English: why and how. DiscAnnotation, Proc ACL Workshop on Discourse Annotation.

Références

Documents relatifs

The overall conclusion is that preverbal complements have a categorical value of discourse-old information while V2 is a productive configuration, and that it is the loss of

In this paper, we study the problem of recognizing those input graphs for which either of the two heuristics can approximate the size of a minimum vertex cover within a constant

Using results of H˚ astad concerning hardness of approximating inde- pendence number of a graph we then prove similar results concerning MAX-MIN-VC (respectively, MAX-MIN-FVS)

Under the assumption that the Polynomial-Time Hierarchy does not collapse we show for a regular language L: the un- balanced polynomial-time leaf language class determined by L

2 What six things do you think would be the most useful things to have in the jungle?. Make a list, and say why each thing

Figure 2: Illustrating the interplay between top-down and bottom-up modeling As illustrated in Figure 3 process modelling with focus on a best practice top-down model, as well

Thus, ϕ has only one eigenvector, so does not admit a basis of eigenvectors, and therefore there is no basis of V relative to which the matrix of ϕ is diagonal4. Since rk(A) = 1,

In these three problems we ask whether it is possible to choose the right order in which the intersection of several automata must be performed in order to minimize the number