Search Results for author: Ludovic Tanguy

Found 21 papers, 1 papers with code

Detecting Contact-Induced Semantic Shifts: What Can Embedding-Based Methods Do in Practice?

no code implementations • EMNLP 2021 • Filip Miletic, Anne Przewozny-Desriaux, Ludovic Tanguy

This study investigates the applicability of semantic change detection methods in descriptively oriented linguistic research.

Change Detection

Paper
Add Code

LITL at SMM4H: An Old-school Feature-based Classifier for Identifying Adverse Effects in Tweets

no code implementations • SMM4H (COLING) 2020 • Ludovic Tanguy, Lydia-Mai Ho-Dac, Cécile Fabre, Roxane Bois, Touati Mohamed Yacine Haddad, Claire Ibarboure, Marie Joyau, François Le moal, Jade Moiilic, Laura Roudaut, Mathilde Simounet, Irena Stankovic, Mickaela Vandewaetere

This paper describes our participation to the SMM4H shared task 2.

regression Task 2

Paper
Add Code

BLOOM: A 176B-Parameter Open-Access Multilingual Language Model

6 code implementations • 9 Nov 2022 • BigScience Workshop, :, Teven Le Scao, Angela Fan, Christopher Akiki, Ellie Pavlick, Suzana Ilić, Daniel Hesslow, Roman Castagné, Alexandra Sasha Luccioni, François Yvon, Matthias Gallé, Jonathan Tow, Alexander M. Rush, Stella Biderman, Albert Webson, Pawan Sasanka Ammanamanchi, Thomas Wang, Benoît Sagot, Niklas Muennighoff, Albert Villanova del Moral, Olatunji Ruwase, Rachel Bawden, Stas Bekman, Angelina McMillan-Major, Iz Beltagy, Huu Nguyen, Lucile Saulnier, Samson Tan, Pedro Ortiz Suarez, Victor Sanh, Hugo Laurençon, Yacine Jernite, Julien Launay, Margaret Mitchell, Colin Raffel, Aaron Gokaslan, Adi Simhi, Aitor Soroa, Alham Fikri Aji, Amit Alfassy, Anna Rogers, Ariel Kreisberg Nitzav, Canwen Xu, Chenghao Mou, Chris Emezue, Christopher Klamm, Colin Leong, Daniel van Strien, David Ifeoluwa Adelani, Dragomir Radev, Eduardo González Ponferrada, Efrat Levkovizh, Ethan Kim, Eyal Bar Natan, Francesco De Toni, Gérard Dupont, Germán Kruszewski, Giada Pistilli, Hady Elsahar, Hamza Benyamina, Hieu Tran, Ian Yu, Idris Abdulmumin, Isaac Johnson, Itziar Gonzalez-Dios, Javier de la Rosa, Jenny Chim, Jesse Dodge, Jian Zhu, Jonathan Chang, Jörg Frohberg, Joseph Tobing, Joydeep Bhattacharjee, Khalid Almubarak, Kimbo Chen, Kyle Lo, Leandro von Werra, Leon Weber, Long Phan, Loubna Ben allal, Ludovic Tanguy, Manan Dey, Manuel Romero Muñoz, Maraim Masoud, María Grandury, Mario Šaško, Max Huang, Maximin Coavoux, Mayank Singh, Mike Tian-Jian Jiang, Minh Chien Vu, Mohammad A. Jauhar, Mustafa Ghaleb, Nishant Subramani, Nora Kassner, Nurulaqilla Khamis, Olivier Nguyen, Omar Espejel, Ona de Gibert, Paulo Villegas, Peter Henderson, Pierre Colombo, Priscilla Amuok, Quentin Lhoest, Rheza Harliman, Rishi Bommasani, Roberto Luis López, Rui Ribeiro, Salomey Osei, Sampo Pyysalo, Sebastian Nagel, Shamik Bose, Shamsuddeen Hassan Muhammad, Shanya Sharma, Shayne Longpre, Somaieh Nikpoor, Stanislav Silberberg, Suhas Pai, Sydney Zink, Tiago Timponi Torrent, Timo Schick, Tristan Thrush, Valentin Danchev, Vassilina Nikoulina, Veronika Laippala, Violette Lepercq, Vrinda Prabhu, Zaid Alyafeai, Zeerak Talat, Arun Raja, Benjamin Heinzerling, Chenglei Si, Davut Emre Taşar, Elizabeth Salesky, Sabrina J. Mielke, Wilson Y. Lee, Abheesht Sharma, Andrea Santilli, Antoine Chaffin, Arnaud Stiegler, Debajyoti Datta, Eliza Szczechla, Gunjan Chhablani, Han Wang, Harshit Pandey, Hendrik Strobelt, Jason Alan Fries, Jos Rozen, Leo Gao, Lintang Sutawika, M Saiful Bari, Maged S. Al-shaibani, Matteo Manica, Nihal Nayak, Ryan Teehan, Samuel Albanie, Sheng Shen, Srulik Ben-David, Stephen H. Bach, Taewoon Kim, Tali Bers, Thibault Fevry, Trishala Neeraj, Urmish Thakker, Vikas Raunak, Xiangru Tang, Zheng-Xin Yong, Zhiqing Sun, Shaked Brody, Yallow Uri, Hadar Tojarieh, Adam Roberts, Hyung Won Chung, Jaesung Tae, Jason Phang, Ofir Press, Conglong Li, Deepak Narayanan, Hatim Bourfoune, Jared Casper, Jeff Rasley, Max Ryabinin, Mayank Mishra, Minjia Zhang, Mohammad Shoeybi, Myriam Peyrounette, Nicolas Patry, Nouamane Tazi, Omar Sanseviero, Patrick von Platen, Pierre Cornette, Pierre François Lavallée, Rémi Lacroix, Samyam Rajbhandari, Sanchit Gandhi, Shaden Smith, Stéphane Requena, Suraj Patil, Tim Dettmers, Ahmed Baruwa, Amanpreet Singh, Anastasia Cheveleva, Anne-Laure Ligozat, Arjun Subramonian, Aurélie Névéol, Charles Lovering, Dan Garrette, Deepak Tunuguntla, Ehud Reiter, Ekaterina Taktasheva, Ekaterina Voloshina, Eli Bogdanov, Genta Indra Winata, Hailey Schoelkopf, Jan-Christoph Kalo, Jekaterina Novikova, Jessica Zosa Forde, Jordan Clive, Jungo Kasai, Ken Kawamura, Liam Hazan, Marine Carpuat, Miruna Clinciu, Najoung Kim, Newton Cheng, Oleg Serikov, Omer Antverg, Oskar van der Wal, Rui Zhang, Ruochen Zhang, Sebastian Gehrmann, Shachar Mirkin, Shani Pais, Tatiana Shavrina, Thomas Scialom, Tian Yun, Tomasz Limisiewicz, Verena Rieser, Vitaly Protasov, Vladislav Mikhailov, Yada Pruksachatkun, Yonatan Belinkov, Zachary Bamberger, Zdeněk Kasner, Alice Rueda, Amanda Pestana, Amir Feizpour, Ammar Khan, Amy Faranak, Ana Santos, Anthony Hevia, Antigona Unldreaj, Arash Aghagol, Arezoo Abdollahi, Aycha Tammour, Azadeh HajiHosseini, Bahareh Behroozi, Benjamin Ajibade, Bharat Saxena, Carlos Muñoz Ferrandis, Daniel McDuff, Danish Contractor, David Lansky, Davis David, Douwe Kiela, Duong A. Nguyen, Edward Tan, Emi Baylor, Ezinwanne Ozoani, Fatima Mirza, Frankline Ononiwu, Habib Rezanejad, Hessie Jones, Indrani Bhattacharya, Irene Solaiman, Irina Sedenko, Isar Nejadgholi, Jesse Passmore, Josh Seltzer, Julio Bonis Sanz, Livia Dutra, Mairon Samagaio, Maraim Elbadri, Margot Mieskes, Marissa Gerchick, Martha Akinlolu, Michael McKenna, Mike Qiu, Muhammed Ghauri, Mykola Burynok, Nafis Abrar, Nazneen Rajani, Nour Elkott, Nour Fahmy, Olanrewaju Samuel, Ran An, Rasmus Kromann, Ryan Hao, Samira Alizadeh, Sarmad Shubber, Silas Wang, Sourav Roy, Sylvain Viguier, Thanh Le, Tobi Oyebade, Trieu Le, Yoyo Yang, Zach Nguyen, Abhinav Ramesh Kashyap, Alfredo Palasciano, Alison Callahan, Anima Shukla, Antonio Miranda-Escalada, Ayush Singh, Benjamin Beilharz, Bo wang, Caio Brito, Chenxi Zhou, Chirag Jain, Chuxin Xu, Clémentine Fourrier, Daniel León Periñán, Daniel Molano, Dian Yu, Enrique Manjavacas, Fabio Barth, Florian Fuhrimann, Gabriel Altay, Giyaseddin Bayrak, Gully Burns, Helena U. Vrabec, Imane Bello, Ishani Dash, Jihyun Kang, John Giorgi, Jonas Golde, Jose David Posada, Karthik Rangasai Sivaraman, Lokesh Bulchandani, Lu Liu, Luisa Shinzato, Madeleine Hahn de Bykhovetz, Maiko Takeuchi, Marc Pàmies, Maria A Castillo, Marianna Nezhurina, Mario Sänger, Matthias Samwald, Michael Cullan, Michael Weinberg, Michiel De Wolf, Mina Mihaljcic, Minna Liu, Moritz Freidank, Myungsun Kang, Natasha Seelam, Nathan Dahlberg, Nicholas Michio Broad, Nikolaus Muellner, Pascale Fung, Patrick Haller, Ramya Chandrasekhar, Renata Eisenberg, Robert Martin, Rodrigo Canalli, Rosaline Su, Ruisi Su, Samuel Cahyawijaya, Samuele Garda, Shlok S Deshmukh, Shubhanshu Mishra, Sid Kiblawi, Simon Ott, Sinee Sang-aroonsiri, Srishti Kumar, Stefan Schweter, Sushil Bharati, Tanmay Laud, Théo Gigant, Tomoya Kainuma, Wojciech Kusa, Yanis Labrak, Yash Shailesh Bajaj, Yash Venkatraman, Yifan Xu, Yingxin Xu, Yu Xu, Zhe Tan, Zhongli Xie, Zifan Ye, Mathilde Bras, Younes Belkada, Thomas Wolf

Large language models (LLMs) have been shown to be able to perform new tasks based on a few demonstrations or natural language instructions.

Language Modelling Multilingual NLP

2,196

Paper
Code

Impact de la structure logique des documents sur les mod\`eles distributionnels : exp\'erimentations sur le corpus TALN (Impact of document structure on distributional semantics models: a case study on NLP research articles )

no code implementations • JEPTALNRECITAL 2020 • Ludovic Tanguy, C{\'e}cile Fabre, Yoann Bard

Nous pr{\'e}sentons une exp{\'e}rience visant {\`a} mesurer en quoi la structure logique d{'}un document impacte les repr{\'e}sentations lexicales dans les mod{\`e}les de s{\'e}mantique distributionnelle.

Paper
Add Code

Collecting Tweets to Investigate Regional Variation in Canadian English

no code implementations • LREC 2020 • Filip Miletic, Anne Przewozny-Desriaux, Ludovic Tanguy

We present a 78. 8-million-tweet, 1. 3-billion-word corpus aimed at studying regional variation in Canadian English with a specific focus on the dialect regions of Toronto, Montreal, and Vancouver.

Paper
Add Code

Extrinsic Evaluation of French Dependency Parsers on a Specialized Corpus: Comparison of Distributional Thesauri

no code implementations • LREC 2020 • Ludovic Tanguy, Pauline Brunet, Olivier Ferret

We present a study in which we compare 11 different French dependency parsers on a specialized corpus (consisting of research articles on NLP from the proceedings of the TALN conference).

Paper
Add Code

Which Dependency Parser to Use for Distributional Semantics in a Specialized Domain?

no code implementations • LREC 2020 • Pauline Brunet, Olivier Ferret, Ludovic Tanguy

We present a study whose objective is to compare several dependency parsers for English applied to a specialized corpus for building distributional count-based models from syntactic dependencies.

Paper
Add Code

Comparaison qualitative et extrins\`eque d'analyseurs syntaxiques du fran\ccais : confrontation de mod\`eles distributionnels sur un corpus sp\'ecialis\'e (Extrinsic evaluation of French dependency parsers on a specialised corpus : comparison of distributional thesauri )

no code implementations • JEPTALNRECITAL 2019 • Ludovic Tanguy, Pauline Brunet, Olivier Ferret

Nous pr{\'e}sentons une {\'e}tude visant {\`a} comparer 11 diff{\'e}rents analyseurs en d{\'e}pendances du fran{\c{c}}ais sur un corpus sp{\'e}cialis{\'e} (constitu{\'e} des archives des articles de la conf{\'e}rence TALN).

Paper
Add Code

Toward a Computational Multidimensional Lexical Similarity Measure for Modeling Word Association Tasks in Psycholinguistics

no code implementations • WS 2019 • Bruno Gaume, Lydia Mai Ho-Dac, Ludovic Tanguy, C{\'e}cile Fabre, B{\'e}n{\'e}dicte Pierrejean, Nabil Hathout, J{\'e}r{\^o}me Farinas, Julien Pinquier, Lola Danet, Patrice P{\'e}ran, Xavier De Boissezon, M{\'e}lanie Jucla

This paper presents the first results of a multidisciplinary project, the {``}Evolex{''} project, gathering researchers in Psycholinguistics, Neuropsychology, Computer Science, Natural Language Processing and Linguistics.

General Classification Semantic Similarity +1

Paper
Add Code

Investigating the Stability of Concrete Nouns in Word Embeddings

no code implementations • WS 2019 • B{\'e}n{\'e}dicte Pierrejean, Ludovic Tanguy

We know that word embeddings trained using neural-based methods (such as word2vec SGNS) are sensitive to stability problems and that across two models trained using the exact same set of parameters, the nearest neighbors of a word are likely to change.

Word Embeddings

Paper
Add Code

Predicting Word Embeddings Variability

no code implementations • SEMEVAL 2018 • B{\'e}n{\'e}dicte Pierrejean, Ludovic Tanguy

Neural word embeddings models (such as those built with word2vec) are known to have stability problems: when retraining a model with the exact same hyperparameters, words neighborhoods may change.

Word Embeddings

Paper
Add Code

Towards Qualitative Word Embeddings Evaluation: Measuring Neighbors Variation

no code implementations • NAACL 2018 • B{\'e}n{\'e}dicte Pierrejean, Ludovic Tanguy

We propose a method to study the variation lying between different word embeddings models trained with different parameters.

Embeddings Evaluation Named Entity Recognition (NER) +1

Paper
Add Code

Etude de la reproductibilit\'e des word embeddings : rep\'erage des zones stables et instables dans le lexique (Reproducibility of word embeddings : identifying stable and unstable zones in the semantic space)

no code implementations • JEPTALNRECITAL 2018 • B{\'e}n{\'e}dicte Pierrejean, Ludovic Tanguy

Les mod{\`e}les vectoriels de s{\'e}mantique distributionnelle (ou word embeddings), notamment ceux produits par les m{\'e}thodes neuronales, posent des questions de reproductibilit{\'e} et donnent des repr{\'e}sentations diff{\'e}rentes {\`a} chaque utilisation, m{\^e}me sans modifier leurs param{\`e}tres.

Word Embeddings

Paper
Add Code

Extending the gold standard for a lexical substitution task: is it worth it?

no code implementations • LREC 2018 • Ludovic Tanguy, C{\'e}cile Fabre, Laura Rivi{\`e}re

Paper
Add Code

Analyse d'une t\^ache de substitution lexicale : quelles sont les sources de difficult\'e ? (Difficulty analysis for a lexical substitution task)

no code implementations • JEPTALNRECITAL 2016 • Ludovic Tanguy, C{\'e}cile Fabre, Camille Mercier

Nous proposons dans cet article une analyse des r{\'e}sultats de la campagne SemDis 2014 qui proposait une t{\^a}che de substitution lexicale en fran{\c{c}}ais.