Persona: Ros Muñoz, Salvador
Cargando...
Dirección de correo electrónico
sros@scc.uned.es
ORCID
0000-0001-6330-4958
Fecha de nacimiento
Proyectos de investigación
Unidades organizativas
Puesto de trabajo
Apellidos
Ros Muñoz
Nombre de pila
Salvador
Nombre
25 resultados
Resultados de la búsqueda
Mostrando 1 - 10 de 25
Publicación The Automatic Quantitative Metrical Analysis of Spanish Poetry with Rantanplan: A Preliminary Approach(ICL CAS, 2021) Hernández Lorenzo, Laura; Sisto, Mirella De; Pérez Pozo, Álvaro; Rosa, Javier de la; Ros Muñoz, Salvador; González Blanco, Elena; Plecháč, P.; Kolár, R.; Bories,A.; Říha, J.In this paper, we present a quantitative approach to Spanish poetry and versification based on the application of our own automatic metrical tool, Rantanplan, to the complete poetic works of four early modern Spanish poets. All of the poetry of these four representative authors—Garcilaso de la Vega (1503–1536), Fernando de Herrera (1534–1597), Luis de Góngora (1561–1627), and Lope de Vega (1562–1635)—was automatically processed and stress positions were extracted. Thanks to the development of a new stanza identification feature of Rantanplan, we were able to detect metrical structures as well. By completing a quantitative analysis of the stress positions, line lengths, and stanzas used by each author, we aim to model their complete metrical profiles.Publicación EVI-LINHD, a virtual research environment for the Spanish speaking community(Oxford University Press, 2017-12) González-Blanco García, Elena; Rio Riande, Gimena del; Díez Platas, María Luisa; Olmo, Álvaro del; Urízar, Miguel; Martínez Cantón, Clara Isabel; Ros Muñoz, Salvador; Pastor Vargas, Rafael; Robles Gómez, Antonio; Caminero Herráez, Agustín CarlosLaboratorio de Innovación en Humanidades Digitales (UNED) has developed Entorno Virtual de Investigación del Laboratorio de Innovación en Humanidades Digitales (EVI-LINHD), the first virtual research environment devoted mainly to Spanish speakers interested in digital scholarly edition. EVI-LINHD combines different open-source software for developing a complete digital project: (1) a Webbased application markup tool—TEIscribe—combined with an eXistdb solution and a TEIPublisher platform, (2) Omeka for digital libraries, and (3) WordPress for simple Web pages. All these instances are linked to a local installation of the LINDAT/Common Language Resources and Technology Infrastructure (CLARIN) digital repository. LINDAT/CLARIN allows EVI-LINHD users to have their projects deposited and stored safely. Thanks to this solution, EVI-LINHD projects also improve their visibility. The specific metadata profile used in the repository is based on Dublin Core, and it is enriched with the Spanish translation of DARIAH’s Taxonomy of Digital Research Activities in the Humanities.Publicación DISCO PAL: Diachronic Spanish sonnet corpus with psychological and affective labels(Springer, 2021-10-13) Barbado, Alberto; Fresno Fernández, Víctor Diego; Manjarrés Riesco, Ángeles; Ros Muñoz, SalvadorNowadays, there are many applications of text mining over corpora from different languages. However, most of them are based on texts in prose, lacking applications that work with poetry texts. An example of an application of text mining in poetry is the usage of features derived from their individual words in order to capture the lexical, sublexical and interlexical meaning, and infer the General Affective Meaning (GAM) of the text. However, even though this proposal has been proved as useful for poetry in some languages, there is a lack of studies for both Spanish poetry and for highly-structured poetic compositions such as sonnets. This article presents a study over an annotated corpus of Spanish sonnets, in order to analyse if it is possible to build features from their individual words for predicting their GAM. The purpose of this is to model sonnets at an affective level. The article also analyses the relationship between the GAM of the sonnets and the content itself. For this, we consider the content from a psychological perspective, dentifying with tags when a sonnet is related to a specific term. Then, we study how GAM changes according to each of those psychological terms. The corpus used contains 274 Spanish sonnets from authors of different centuries, from fifteenth to nineteenth. This corpus was annotated by different domain experts. The experts annotated the poems with affective and lexico-semantic features, as well as with domain concepts that belong to psychology. Thanks to this, the corpus of sonnets can be used in different applications, such as poetry recommender systems, per- sonality text mining studies of the authors, or the usage of poetry for therapeutic purposes.Publicación Test-driving information theory-based compositional distributional semantics: A case study on Spanish song lyrics(ELSEVIER, 2025-06-15) Ghajari Espinosa, Adrián; Benito Santos, Alejandro; Ros Muñoz, Salvador; Fresno Fernández, Víctor Diego; González Blanco, ElenaSong lyrics pose unique challenges for semantic similarity assessment due to their metaphorical language, structural patterns, and cultural nuances - characteristics that often challenge standard natural language processing (NLP) approaches. These challenges stem from a tension between compositional and distributional semantics: while lyrics follow compositional structures, their meaning depends heavily on context and interpretation. The Information Theory-based Compositional Distributional Semantics framework offers a principled approach by integrating information theory with compositional rules and distributional representations. We evaluate eight embedding models on Spanish song lyrics, including multilingual, monolingual contextual, and static embeddings. Results show that multilingual models consistently outperform monolingual alternatives, with the domain-adapted ALBERTI achieving the highest F1 macro scores (78.92 ± 10.86). Our analysis reveals that monolingual models generate highly anisotropic embedding spaces, significantly impacting performance with traditional metrics. The Information Contrast Model metric proves particularly effective, providing improvements up to 18.04 percentage points over cosine similarity. Additionally, composition functions maintaining longer accumulated vector norms consistently outperform standard averaging approaches. Our findings have important implications for NLP applications and challenge standard practices in similarity calculation, showing that effectiveness varies with both task nature and model characteristics.Publicación Characterizing the visualization design space of distant and close reading of poetic rhythm(Frontiers, 2023-06-06) Benito Santos, Alejandro; Ros Muñoz, Salvador; Therón Sánchez, Roberto; García Peñalvo, Francisco J.; Agencia Estatal de Investigación (España)Metrical and rhythmical poetry analysis is founded on the systematic statistical analysis and comparison of sonic devices (e.g., rhythmic patterns) that emerge from a combination of pre-established aesthetic and structural rules and the poet's abilities and creative genius to convey a given message adhering to the said constraints. These rhythmical patterns, which have been traditionally obtained by means of a careful close reading of the poems, in a process known as “scansion,” can now be obtained and made visible by automatic means. However, the visualization literature is still scarce on approaches that allow an insightful close and distant reading of the rhythmical patterns in a poetry corpus. In this work, we report our initial efforts in characterizing of the visualization design space of distant and close reading of poetic rhythm. By employing a digital version of a corpus of 11,268 verses originally written by the Spanish poet and playwright Federico García-Lorca (1898–1936), we could craft several prototypical visualizations representative of the inherent complexity of the problem which we expect to employ in future user studies and that we share here with the rest of the community to foster further discussion around this interesting topic.Publicación Hispanic Medieval Tagger (HisMeTag): una aplicación web para el etiquetado de entidades en textos medievalesDíez Platas, María Luisa; González-Blanco García, Elena; Rio Riande, Gimena del; Tobarra Abad, María de los Llanos; Ros Muñoz, Salvador; Robles Gómez, Antonio; Caminero Herráez, Agustín CarlosHisMeTag permite localizar entidades nombradas en textos escritos en español medieval, mediante un proceso automático de reconocimiento de entidades nombradas (NER) y técnicas de PLN para el procesamiento lingüístico y la generación de las distintas variantes que existieron en la época medieval. Localiza, etiqueta términos conocidos y propone nuevos términos para su validación.Publicación Poetry Lab (POSTER)(2018) González-Blanco García, Elena; Díez Platas, María Luisa; Ruiz Fabo, Pablo; Bermúdez Sabel, Helena; Ayciriex, Luciana; Ros Muñoz, Salvador; Martínez Cantón, Clara IsabelMain goals: a) Develop Tools for automatic poetry analysis, largely based on Natural Language Processing. b) Carry out the detection of literary phenomena relied on linguistic characteristics.Publicación Researchers’ perceptions of DH trends and topics in the English and Spanish-speaking community. DayofDH data as a case study(Jagiellonian University & Pedagogical University (Cracovia), 2016-07-22) González-Blanco García, Elena; Rio Riande, Gimena del; Robles Gómez, Antonio; Ros Muñoz, Salvador; Hernández Berlinches, Roberto; Tobarra Abad, María de los Llanos; Caminero Herráez, Agustín Carlos; Pastor Vargas, RafaelPublicación From syllables, lines and stanzas to linked open data: standardization, interoperability and multilingual challenges for digital humanities(2016) González-Blanco García, Elena; Manailescu, Mara; Ros Muñoz, SalvadorThis proposal presents the challenges and first results of POSTDATA ERC Starting Grant project, which aims at bridging the digital gap among traditional poetry collections and the growing world of data. It is focused on poetry analysis, classification and publication, applying Digital Humanities methods of academic analysis in order to look for standardization. The context of the project is the corpora of European poetry, with a special focus on poetic materials from different languages and literary traditions. Interoperability problems between the different poetry collections are solved by using semantic web technologies to link and publish literary datasets in a structured way in the linked data cloud. This paper will present the current situation in the field of digital humanities analyzing poetry as the “study case” and the application of different technologies used in the field of digital humanities to provide new and innovative results. It will also introduce LINDH, the Digital Humanities Innovation Lab at UNED, a pioneer Digital Humanities center in Spain and its role as a facilitator of different technologies to be applied to the study of traditional humanistic problems with the most updates technologies in the field.Publicación Transformers analyzing poetry: multilingual metrical pattern prediction with transfomer-based language models(Springer, 2023) Rosa, Javier de la; Pérez Pozo, Álvaro; Sisto, Mirella De; Hernández Lorenzo, Laura; Díaz Paredes, Aitor; Ros Muñoz, Salvador; González Blanco, ElenaThe splitting of words into stressed and unstressed syllables is the foundation for the scansion of poetry, a process that aims at determining the metrical pattern of a line of verse within a poem. Intricate language rules and their exceptions, as well as poetic licenses exerted by the authors, make calculating these patterns a nontrivial task. Some rhetorical devices shrink the metrical length, while others might extend it. This opens the door for interpretation and further complicates the creation of automated scansion algorithms useful for automatically analyzing corpora on a distant reading fashion. In this paper, we compare the automated metrical pattern identification systems available for Spanish, English, and German, against fine-tuned monolingual and multilingual language models trained on the same task. Despite being initially conceived as models suitable for semantic tasks, our results suggest that transformers-based models retain enough structural information to perform reasonably well for Spanish on a monolingual setting, and outperforms both for English and German when using a model trained on the three languages, showing evidence of the benefits of cross-lingual transfer between the languages.
- «
- 1 (current)
- 2
- 3
- »