| Société : Aldebaran Robotics Type de contrat : CDD / Freelance Mission : ALDEBARAN Robotics est devenu le leader mondial en robotique humanoïde autonome avec un seul objectif en tête « créer des robots pour aider les gens ». Dans le cadre du développement du robot NAO nous recherchons un linguiste pour participer à la création des contenus de son moteur de dialogues. Profil : De langue maternelle française, vous êtes linguiste ou avez des compétences similaires. Vous êtes créatif, très rigoureux et maîtrisez parfaitement la construction de dialogues. Vous êtes aussi à l'aise avec l'informatique, écriture / utilisation de langage de script. La maîtrise d'une ou plusieurs langues complémentaire comme l'anglais, l'allemand, le chinois, le japonais est un vrai plus. Contact : Alexis Bernazeau abernazeau@aldebaran-robotics.com |
Actualités et discussions autour du cursus de Linguistique Informatique de l'université Paris Diderot
mercredi 25 septembre 2013
On recherche un linguiste à l'aise avec l'informatique pour faire de la robotique
Une annonce qui nous a été envoyée par la société Aldebaran Robotics, où travaille une jeune diplômée du Master 2 PRO:
mardi 10 septembre 2013
Exposés de Mark Johnson
Un des meilleurs spécialistes mondiaux de linguistique computationnelle, Mark Johnson, va donner deux exposés à l'université Paris Diderot. Tous les étudiants intéressés par le traitement automatique des langues son fortement invités à y assister:
*
*Language acquisition as statistical inference**
**
**Mark Johnson**
**Macquarie University**
**
**noon, 12th September, LingLunch*
Linglunch Paris Diderot
Thursday, 12th septembre 2013
12h-13h, salle 103
bâtiment Olympe de Gouges
(8) rue Albert Einstein, 75013
http://www.linguist.univ-paris-diderot.fr/linglunch.html
This talk argues that language acquisition -- in particular, syntactic
parameter setting -- is profitably viewed as a statistical inference
problem. I discuss some issues associated with statistical inference
that linguists might be concerned about, including the possibility of
"Zombie" parameter settings. The bulk of the talk focuses on estimating
parameters in a Stabler-style Minimalist Grammar framework. Building on
recent results of Hunter and Dyer (2013), we show how estimating weights
associated with lexical entries -- including the empty functional
categories that control parametric syntactic variation -- can be reduced
to estimating weights in what appears to be a new grammar formalism
called "feature-weighted context-free grammars", which is a MaxEnt
generalisation of the "tied context-free grammars" of Headden et al
(2009). Importantly, the partition function and its derivatives of a
feature-weighted context-free grammar can be calculated using a
generalisation inspired by the Inside-Outside algorithm of the
algorithms for calculating partition functions in Nederhof and Satta
(2009). We show how this can be used to learn lexical entries and verb
movement and XP movement parameters in three toy corpora.
*
*
*Grammars and Topic Models**
**
**Mark Johnson**
**Macquarie University**
**
**11am, 20th September, Alpage Group*
Séminaire ALPAGE
Friday, 20th september, 11h-12h30
salle 127
bâtiment Olympe de Gouges
(8) rue Albert Einstein, 75013
https://www.rocq.inria.fr/alpage-wiki/tiki-index.php?page=seminaire
Context-free grammars have been a cornerstone of theoretical computer
science and computational linguistics since their inception over half a
century ago. Topic models are a newer development in machine learning
that play an important role in document analysis and information
retrieval. It turns out there is a surprising connection between the
two that suggests novel ways of extending both grammars and topic
models. After explaining this connection, I go on to describe
extensions which identify topical multiword collocations and
automatically learn the internal structure of named-entity phrases.
These new models have applications in text data mining and information
retrieval.
****
**
*
Cette série d'exposés est financée par:
Research in Paris Programme - Mairie de Paris
Ecole Normale Supérieure
Ecole des Hautes Etudes en Sciences Sociales
Fondation Pierre Gilles de Gennes
*
**
****
*
*Language acquisition as statistical inference**
**
**Mark Johnson**
**Macquarie University**
**
**noon, 12th September, LingLunch*
Linglunch Paris Diderot
Thursday, 12th septembre 2013
12h-13h, salle 103
bâtiment Olympe de Gouges
(8) rue Albert Einstein, 75013
http://www.linguist.univ-paris-diderot.fr/linglunch.html
This talk argues that language acquisition -- in particular, syntactic
parameter setting -- is profitably viewed as a statistical inference
problem. I discuss some issues associated with statistical inference
that linguists might be concerned about, including the possibility of
"Zombie" parameter settings. The bulk of the talk focuses on estimating
parameters in a Stabler-style Minimalist Grammar framework. Building on
recent results of Hunter and Dyer (2013), we show how estimating weights
associated with lexical entries -- including the empty functional
categories that control parametric syntactic variation -- can be reduced
to estimating weights in what appears to be a new grammar formalism
called "feature-weighted context-free grammars", which is a MaxEnt
generalisation of the "tied context-free grammars" of Headden et al
(2009). Importantly, the partition function and its derivatives of a
feature-weighted context-free grammar can be calculated using a
generalisation inspired by the Inside-Outside algorithm of the
algorithms for calculating partition functions in Nederhof and Satta
(2009). We show how this can be used to learn lexical entries and verb
movement and XP movement parameters in three toy corpora.
*
*
*Grammars and Topic Models**
**
**Mark Johnson**
**Macquarie University**
**
**11am, 20th September, Alpage Group*
Séminaire ALPAGE
Friday, 20th september, 11h-12h30
salle 127
bâtiment Olympe de Gouges
(8) rue Albert Einstein, 75013
https://www.rocq.inria.fr/alpage-wiki/tiki-index.php?page=seminaire
Context-free grammars have been a cornerstone of theoretical computer
science and computational linguistics since their inception over half a
century ago. Topic models are a newer development in machine learning
that play an important role in document analysis and information
retrieval. It turns out there is a surprising connection between the
two that suggests novel ways of extending both grammars and topic
models. After explaining this connection, I go on to describe
extensions which identify topical multiword collocations and
automatically learn the internal structure of named-entity phrases.
These new models have applications in text data mining and information
retrieval.
****
**
*
Cette série d'exposés est financée par:
Research in Paris Programme - Mairie de Paris
Ecole Normale Supérieure
Ecole des Hautes Etudes en Sciences Sociales
Fondation Pierre Gilles de Gennes
*
**
****
lundi 9 septembre 2013
Participants francophones pour des questionnaires en ligne (rapides)
Plusieurs études en psycholinguistique (production du langage) qui demandent des participants de langue maternelle française. Deux questionnaires rapides (5 ou 10 minutes).
Les recherches sont menées au département de linguistique d'Oxford.
http://users.ox.ac.uk/~scro2158/
Un deuxième questionnaire prenant moins de 10 minutes : https://sites.google.com/site/lpplabox/
Les recherches sont menées au département de linguistique d'Oxford.
http://users.ox.ac.uk/~scro2158/
Un deuxième questionnaire prenant moins de 10 minutes : https://sites.google.com/site/lpplabox/
vendredi 6 septembre 2013
Volontaires recherchés pour une étude sur les mots du français
Le questionnaire ci-dessous fait partie d'une étude qui a pour but d'examiner quelques mots de la langue française.
Pour chaque mot, nous vous demandons de faire 2 choses:
participation soit prise en compte, vous devez répondre à toutes les
questions.
http://spellout.net/ibexexps/lisinparis/VNC/experiment.html
Pour chaque mot, nous vous demandons de faire 2 choses:
- Indiquer à quelle fréquence vous avez rencontré le mot - souvent, parfois, rarement, jamais
- Essayer de nous donner une courte définition de ce mot.
participation soit prise en compte, vous devez répondre à toutes les
questions.
http://spellout.net/ibexexps/lisinparis/VNC/experiment.html
jeudi 5 septembre 2013
Poste d'ingénieur NLP, start up en extraction d'information
Une annonce urgente de poste pour rejoindre la start-up trooclick, spécialiste en extraction d'information. Poste indiqué (et conseillé) par un de nos enseignants professionnels,
Plus de détails
Plus de détails
lundi 2 septembre 2013
Maluuba, Canada, recrute en NLP
Une annonce transmise via notre réseau d'anciens du cursus
l'entreprise Maluuba recrute au Canada. A tenter pour voir du pays!
l'entreprise Maluuba recrute au Canada. A tenter pour voir du pays!
Must Have Qualifications
- Bachelor's in Computer Science, Computational Linguistics, Software
Engineering, or equivalent
- Native fluency in one of French, Italian, German, Spanish, or
Portuguese
- Professional or native fluency in English
- Programming experience
- Demonstrated ability to proactively develop innovative algorithms
- Candidates must be very data driven and creative
- Experience working through a full product lifecycle, from design to
delivery
Nice To Have
- Graduate, postdoc, or industry experience in NLP, Machine Learning,
Computational Linguistics, or data mining and management
- Demonstrative ability to keep up to date with the latest research in
above the fields
- Programming experience in Java or Python
http://www.maluuba.com/careers/12
|
vendredi 26 juillet 2013
Rentrée 2013/14
Juste avant la fermeture pour les vacances, voici un recueil en vrac des informations pertinentes pour la rentrée 2013/14 du cursus LI. Cette page sera régulièrement mise à jour jusqu'en septembre.
- Date de rentrée: tous les cours commencent la semaine du 16 septembre (sauf indication contraire). Pour les cours d'informatique, les TD commencent en général la 2e semaine.
- Réunion de rentrée: pas obligatoire, mais fortement recommandée, la réunion de rentrée aura lieu vendredi 13 septembre, de 15h00 à 17h00, en salle 310. La réunion de rentrée concerne les étudiants en L3 et en Master 1.
[edit] Transparents présentés à la réunion d'information - Les emplois du temps (encore des changements possibles) sont en ligne sur cette page.
- Localisation: la quasi totalité des cours de Master 2 (Rech et Pro) aura lieu dans la salle 309, où les terminaux X seront opérationnels.
- L'examen spécial d'entrée en M1 est organisé le vendredi 13 septembre 2013 de 09h à 13h. Il aura lieu en salle 310 (3e étage, bâtiment Olympe de Gouges). Pour ceux qui ne passent qu'une épreuve, elle commencera à 09h (et se terminera à 11h). Pour ceux qui passent deux épreuves, on commencera par l'informatique à 09h. Rappel: il est nécessaire d'avoir été officiellement convoqué pour passer cet examen d'entrée.
- Cours de linguistique en M1: le cours de phonétique du 2e semestre (Techniques et méthodes expérimentales appliquées à l'oral) aura dorénavant comme pré-requis le cours de phonétique expérimentale de M1, qui n'était jusque là pas prévu dans la liste des cours). C'est pour prendre cela en compte que sont proposées deux "options" différentes pour le choix des cours, à déterminer dès le début de l'année.
| option non phonétique |
option phonétique | |
| S1 |
|
|
| S2 |
|
|
- Cours d'Introduction au TAL (L3) : ce cours est ouvert aux étudiants arrivant en M1 sans avoir fait la licence LI. La première séance aura lieu le lundi 23 septembre.
mardi 18 juin 2013
Ingénieur d'étude TAL/analyse discursive, 9 à 13 mois, Paris Diderot
Un poste d'ingénieur TAL, orienté vers l'annotation plus ou moins automatisée, ouvert pour la rentrée de septembre dans le laboratoire CLILLAC-ARP. Date limite : 30 juin.
Offre de poste d’ingénieur d’étude à CLILLAC-ARP EA 3967 Domaine : Traitement Automatique du Langage Spécialité : Analyse discursive Offre de poste de 9 à 12 mois au sein de CLILLAC-ARP, EA 3967, Université Paris Diderot, UFR Etudes anglophones Le contrat pourrait commencer au 1er octobre 2013 si la sélection a lieu avant le 10 juillet 2013. Projets CLILLAC-ARP Dans le cadre de l’ANR EMCO Emphiline Mots clés : Organisation du discours, Segmentation textuelle, linguistiques de corpus, traitements automatiques, gestion de base de données Le travail comportera deux volets : l’annotation manuelle des corpus (segmentation de textes en unités discursives et définition des relations discursives) et la gestion de la base de données du projet. L’annotation visera à modéliser l’impact de l’inattendu sur la syntaxe et l’organisation discursive dans l’interaction verbale en anglais et en français. La personne recrutée aura à concevoir la base de données linguistiques (conception des formulaires de requêtes pour l’annotation, contribution au manuel d’annotation et participation à l’annotation, exportation des annotations). Compétences requises : - Formation en Traitement Automatique des Langues ; - Programmation en Perl ou Python ; maîtrise de l’interface GLOZZ et de PRAAT - Bonnes connaissances en bases de données (MySQL) - Bonnes connaissances en technologies Web (HTML, CSS, PHP, etc.) ; Le candidat idéal aura la capacité d’allier trois perspectives : une perspective TAL, une perspective linguistique, et une perspective cognitive permettant de décrire les processus mis en œuvre dans la production et l’interprétation des structures discursives. La rémunération brute mensuelle correspond à l’INM 370, soit 1 764,60 €. Pour candidater, merci d'envoyer à agnes.celle@univ-paris-diderot.fr les documents suivants au format PDF, de préférence avant le 30 juin 2013 : - un CV détaillé (avec une liste des publications, s'il y a lieu) ; - une lettre de motivation ; - un relevé de notes de Master ou rapport de soutenance de thèse; - le nom et l'email de deux personnes référentes (enseignant, encadrant, responsable) susceptibles d'être contactées. |
mardi 28 mai 2013
Bourse de thèse en sémantique computationnelle, Groningue
Un poste de doctorant en sémantique computationnelle, dans un laboratoire très actif dans le domaine. Date limite le 1er juillet (candidature en anglais).
*Job description* Applications are invited for a PhD candidate in the area of computational semantics. This position is part of the D-MAP project currently running at the University of Groningen, aiming to combine formal with empirical and statistical approaches to semantics. A first milestone reached by D-MAP was the recently released Groningen Meaning Bank, a large annotated resource comprising texts annotated with discourse representation structures. *Requirements* - Master's degree in computational linguistics (or related field) - excellent record of undergraduate and Master's level study - experience in the area of semantics or machine learning - programming experience - ability to work in a research team - strong motivation to complete a PhD dissertation in four years - good command of English (TOEFL 620, IELTS 7,5, Cambridge Advanced CAE). *Conditions of employment* The University of Groningen offers a salary of ¤ 2,083 gross per month in the first year to ¤ 2,664 gross per month in the fourth year (figures based on full employment). The full-time appointment is temporary for a specified period of four years. The position requires residence in Groningen, 38 hours/week research and research training, and must result in a PhD dissertation. After the first year there will be an assessment of the candidate's results and the progress of the project to decide whether the employment will be continued. *Affiliation* The PhD candidate will be affiliated with the computational linguistics group of the Center for Language and Cognition Groningen (CLCG) at the Faculty of Arts of the University of Groningen. This institute embraces all the Linguistics research in the faculty. The PhD candidate will be enrolled in the research training program of the Graduate School for the Humanities and will be supervised by Prof Johan Bos. *Application* You may apply for this position before 1 July 2013 Dutch local time by means of the application form at: http://www.tangram-tis.nl/10378/Kandidaten/Inschrijven/00347-0000005225 Please include in your application (in English): - a letter of motivation - a curriculum vitae - a copy of diplomas, a list of grades, and a passport copy - the names and contact details of two academic referees - a two-page summary of your research topic, including objectives, method, and embedding within the D-MAP project. Send us your entire application in PDF format, using the link to the application form below. Incomplete dossiers will not be taken into consideration. Interviews with a selection of the most appropriate candidates will presumably take place in the first two weeks of July. The starting date of the PhD project is 1 September 2013. Acquisition is not appreciated. Contract type: 48 months Additional information: - Prof Johan Bos: johan.bos@rug.nl - The Groningen Meaning Bank: gmb.let.rug.nl |
vendredi 24 mai 2013
Ecole doctorale de l'Inalco: dates du concours
Le concours pour les contrats doctoraux de l'école doctorale n°265 ("Langues, Littératures et Sociétés du
monde") est officiellement ouvert.
Cette école doctorale, qui est l'école doctorale de l’INALCO, qui rassemble 14 équipes de recherche et forme environ 300 étudiants, et accueille des thèses dans de nombreux domaines, dont la linguistique.
Plus de détails
[Edit 27/05] On peut ajouter à ce post des informations à propos de l'école doctorale n°268 (Langage et Langues), dont la campagne en vue de l'attribution des contrats doctoraux vient de s'ouvrir.
Plus de détails
Cette école doctorale, qui est l'école doctorale de l’INALCO, qui rassemble 14 équipes de recherche et forme environ 300 étudiants, et accueille des thèses dans de nombreux domaines, dont la linguistique.
Plus de détails
[Edit 27/05] On peut ajouter à ce post des informations à propos de l'école doctorale n°268 (Langage et Langues), dont la campagne en vue de l'attribution des contrats doctoraux vient de s'ouvrir.
Plus de détails
Inscription à :
Articles (Atom)