000 04015nam a22004930i 4500
999 _c4957
_d4957
001 00004957
003 ES-MaONT
005 20250123012003.0
008 190326s2019 eu da fst i000 0 eng d
020 _a978-92-76-01518-5
022 _a1831-9424
024 7 _a10.2760/577814
_2doi
_d.
040 _aLU-LuOPE
_c ES-MaONT
044 _ceu
245 1 0 _aSemantic text analysis tool
_bSeTA : supporting analysts by applying advanced text mining techniques to large document collections.
260 _a[Luxemburgo] :
_bOficina de Publicaciones,
_c2019.
300 _a47 p. :
_bil. col..
490 1 _aJRC Technical Reports
_v; 29708
504 _aBibliografía: p. 43-43.
520 _aMuch of the world's data is textual – in large document archives, in scientific papers, in scattered websites, in social media. The information contained in text is invaluable and yet hard to access. The sheer volume of text means that, unassisted, we cannot hope to read all available sources, nor even to keep up to date with all advances in a particular field. For example, EUR-Lex, the database of EU Legal texts, grows by over 15 000 texts per year while Scopus, a database of scientific papers, has over 70 million entries. The problems of scale are compounded by other challenges such as the breadth of topics covered, their jargon specific to each field and the changes in meanings of phrases over time. The mission of the JRC is to provide scientific support to policy development, through original and applied research and knowledge management (JRC Strategy 2030). The challenges of accessing information "trapped in text" are very relevant to this mission of the JRC, as timely, relevant information is needed at all stages of the policy development process. To help overcome the challenges posed by text the JRC has produced a new tool, SeTA – Semantic Text Analyser – which applies advanced text analysis techniques to large document collections, helping policy analysts to understand the concepts expressed in thousands of documents and to see in a visual manner the relationships between these concepts and their development over time. A pilot version of this tool has been populated with hundreds of thousands of documents from EUR-Lex, the EU Bookshop and other sources, and used at the JRC in a number of policy-related use cases including impact assessment, the analysis of large data infrastructures, agri-environment measures and natural disasters. The document collection which have been used, the technical approach chosen and key use cases are described in this document.
540 _aReutilización autorizada, con indicación de la fuente bibliográfica. La política relativa a la reutilización de los documentos de la Comisión Europea fue establecida por la Decisión 2011/833/UE (DO L 330 de 14.12.2011, p. 39). Cualquier uso o reproducción de fotografías u otro material que no esté sujeto a los derechos de autor de la Unión Europea requerirá la autorización de sus titulares ;
_bUnión Europea.
650 0 _aTecnologías habilitadoras digitales
_918
653 7 _aanálisis de la información
653 7 _abúsqueda documental
653 7 _ainformática documental
653 7 _ainforme de investigación
653 7 _aobra de referencia
653 7 _atecnología de la información
700 1 _aAcs, S.
_92739
700 1 _aArnes Novau, X.
_92742
700 1 _aHradec, J.
_92736
700 1 _aListorti, G.
_92740
700 1 _aMacmillan, C.
_92738
700 1 _aOstlaender, N.
_92737
700 1 _aTomas, R.
_92741
710 1 _aComisión Europea.
_bCentro Común de Investigación
_92681
830 _aJRC Technical Reports
_92965
856 4 2 _qHTML
_uhttp://publications.europa.eu/publication/manifestation_identifier/PUB_KJNA29708ENN
_x0
_yAcceso a la publicación
856 4 2 _qHTML
_uhttps://data.europa.eu/doi/10.2760/577814
_x0
_yAcceso a la publicación
901 _aDOI registered
910 _aFree
911 _aStudies
942 _cINF
_2z