Skip to main content

ETHNOS_APP

Home • Search • Journals • List 0

From text saliency to linguistic objects

Learning linguistic interpretable markers with a multi-channels convolutional architecture

Bibliographic Data

ID19528985
AuthorsLaurent Vanni (Centre National de la Recherche Scientifique), Marco Corneli (0000-0002-9361-0080, Centre National de la Recherche Scientifique), Damon Mayaffre (0000-0003-0792-5973, Centre National de la Recherche Scientifique), Frédéric Precioso (0000-0001-8712-1443, Centre National de la Recherche Scientifique)
Year2023
Volume24
Publication date2023-01-01
Peer ReviewedYes
Open AccessNo
TypeARTICLE
VenueCorpus (JOURNAL)
Journal identifiersISSN: 1638-9808 • E-ISSN: 1765-3126
PublisherOpenEdition (PUBLISHER)
DOI10.4000/corpus.7667
OpenAlexW3015213360
LanguageEN
Citations received4
References cited20

A lot of effort is currently made to provide methods to analyze and understand deep neural network impressive performances for tasks such as image or text classification. These methods are mainly based on visualizing the important input features taken into account by the network to build a decision. However these techniques, let us cite LIME, SHAP, Grad-CAM, or TDS, require extra effort to interpret the visualization with respect to expert knowledge. In this paper, we propose a novel approach to inspect the hidden layers of a fitted CNN in order to extract interpretable linguistic objects from texts exploiting classification process. In particular, we detail a weighted extension of the Text Deconvolution Saliency (wTDS) measure which can be used to highlight the relevant features used by the CNN to perform the classification task. We empirically demonstrate the efficiency of our approach on corpora from two different languages: English and French. On all datasets, wTDS automatically encodes complex linguistic objects based on co-occurrences and possibly on grammatical and syntax analysis

Convolutional neural network · Deconvolution · Machine learning · Natural language processing · Syntax · Visualization · Computational and Text Analysis Methods · Computer Science · Natural Language Processing Techniques · Topic Modeling · Artificial Intelligence

  • Identification des motifs textuels. Entre statistique et deep learning

    Dominique Longrée, Laurent Vanni•Corpus•2025

  • En-registrer les genres littéraires. Repérage et apprentissage de marqueurs de registres discursifs

    Veronique Magri-Mourgues, Damon Mayaffre et al.•Corpus•2025

  • AD et IA

    Damon Mayaffre, Laurent Vanni•Langue française•2025

  • Ces mots que Macron emprunte à Sarkozy. Discours et intelligence artificielle

    Damon Mayaffre, Magali Guaresi et al.•Corpus•2020

  • Exploring Textual Data

    Open Access•Ludovic Lebart, André Salem et al.•Exploring Textual Data•1998

  • Enriching Word Vectors with Subword Information

    Open Access•Piotr Bojanowski, Edouard Grave et al.•Transactions of the Association…•2017

  • Bag of Tricks for Efficient Text Classification

    Open Access•Armand Joulin, Edouard Grave et al.•Proceedings of the 15th…•2017

  • Cultural Backlash

    Open Access•P Norris, Ronald Inglehart•Cultural Backlash•2019

  • Rise of the Trumpenvolk

    Open Access•J Eric Oliver, James E Oliver et al.•The Annals of the American…•2016

Unique citing works4
Citations per year0,67
Citation span2020 - 2025 (6)
Citation velocityrecent
Highly citedNo
Citation typesNeutral: 4

Tools

Open DOIOpen Access
Ethnos_APP • Open Source Project • MIT License • Frontend v2.0.0 • Privacy and Cookies • API Documentation: api.ethnos.app/docs • API Source Code: GitHub • DOI: 10.5281/zenodo.17049435 • Frontend Source Code: GitHub • DOI: 10.5281/zenodo.17050053 • cruz.rio.br • Expectantes Misericordiae