« The annual ACM SIGKDD conference is the premier international forum for data mining researchers and practitioners from academia, industry, and government to share their ideas, research results and experiences. (…) »
« Many researchers want to carry out analysis and extraction of information from large sets of data, such as journal articles and other scholarly content. Methods such as screen-scraping are error-prone, place too much strain on content sites and may be unrepeatable or break if site layouts change. Providing researchers with…
« In this study, network and main path analyses were conducted on 1,856 studies related to text mining, by extracting keywords and citation information from the text of each paper. Our findings indicate that research papers on text mining have been published in 45 academic disciplines in the 1980s and 1990s,…
« L’émergence du Covid-19, fin décembre 2019, a été repérée en ligne par certains systèmes de surveillance. Noyés sous une montagne de données, ces signaux faibles n’ont cependant pas su être interprétés à temps. Une équipe de chercheurs du Cirad, dans un article publié le 20 juillet dans Transboundary and Emerging…
« De Visa-tm à Objectif-TDM
Un blog avait été ouvert à l’initiative de l’Inist afin de communiquer sur le projet VisaTM (2017-2019).
Ce projet, dont l’objectif était de cerner les contours d’une infrastructure de services autour du TDM (Text et Data…
« openVirus is innovating new types of search for research literature using data mining technologies to enable citizens to make use of scientific knowledge. (…) »
openVirus works by speedily downloading papers as full-text from open repositories (EuropePMC, bioriv and medrxiv, DOAJ, EThOS, Redalyc (MX), etc.) at an average rate of…
« (…) Pour rendre hommage à ce grand musicien en cette année 2020, l’équipe ISTEX a souhaité créer une collection de corpus thématiques qui lui soit consacrée. Ces corpus, issus de l’archive multidisciplinaire ISTEX qui remonte jusqu’au XIVe siècle, sont destinés à être annotés pour y détecter en particulier les noms…
« To exploit scientific publications from global research for TDM purposes, the ISTEX platform enriched its data with value-added information to ease access to its full-text documents. We built an experiment to explore new enrichment possibilities indocuments focussing on scientific named entities recognitionwhich could be integrated into ISTEX resources. (…) »