Traitement de corpus

21/09/2018

HathiTrust Research Center Extends Non-Consumptive Research Tools to Copyrighted Materials: Expanding Research through Fair Use

« HathiTrust has reached a tremendous milestone in the history of HathiTrust and the HathiTrust Research Center’s services.

Since 2011, HTRC has been developing services and tools to allow researchers to employ text and data mining methodologies using the HathiTrust collection. To date, this service has been available only on the…

fléche suivante Lire l'article complet
20/08/2018

Méthodologie pour identifier les terrains d’étude dans des corpus scientifiques

« Le projet interdisciplinaire TERRE-ISTEX a pour objectif d’identifier l’évolution des fronts de recherche en relation avec les territoires d’études, les croisements disciplinaires ainsi que les modalités concrètes de recherche à partir des contenus numériques hétérogènes disponibles dans les corpus scientifiques. Le projet se décompose en trois actions principales~: (1) identifier…

fléche suivante Lire l'article complet
13/08/2018

OpenMinTeD: A Platform Facilitating Text Mining of Scholarly Content [.pdf]

« The OpenMinTeD platform aims to bring full text Open Access scholarly content from a wide range of providers together with Text and Data Mining (TDM) tools from various Natural Language Processing frameworks and TDM developers in an integrated environment. In this way, it supports users who want to mine scientific…

fléche suivante Lire l'article complet
08/08/2018

LREC 2018, Eleventh International Conference on Language Resources and Evaluation, Miyazaki, Japan [papers]

« Since the first LREC held in Granada in 1998, LREC has become the major event on Language Resources (LRs) and Evaluation for Language Technologies (LT). The aim of LREC is to provide an overview of the state-of-the-art, explore new R&D directions and emerging trends, exchange information regarding LRs and their…

fléche suivante Lire l'article complet
31/07/2018

Projet VisaTM : l’interconnexion OpenMinTeD – AgroPortal – ISTEX, un exemple de service de Text et Data Mining pour les scientifiques français

« Présentation du projet VisaTM
La création d’une offre de service en fouille de texte et de données – TDM (Text and Data Mining) – à destination des scientifiques se pose dans un contexte évolutif sur le plan légal, organisationnel et scientifique. Les progrès récents des méthodes d’analyse textuelle ouvrent…

fléche suivante Lire l'article complet