Signal & Interactive Systems Lab

Danieli M., Ciulli T, Mousavi M. and Riccardi G.

A Participatory Design of Conversational Artificial Intelligence Agents for Mental Healthcare (Article)

Journal of Medical Internet Research (JMIR) Formative Research Journal, 5 (12), 2021.

(Links | BibTeX | Tags: Conversational and Interactive Systems , Signal Annotation and Interpretation)

Torres M. J., Clarkson T., Hauschild K., Luhmann C. C., Lerner D. M. and Riccardi G.

Facial emotions are accurately encoded in the brains of those with autism: A deep learning approach (Article)

Biological Psychiatry: Cognitive Neuroscience and Neuroimaging, 2021.

(Links | BibTeX | Tags: Affective Computing, Autism, Machine Learning, Signal Annotation and Interpretation)

Alam F., Danieli M. and Riccardi G.

Annotating and Modeling Empathy in Spoken Conversations (Article)

Computer Speech and Language, 50 pp. 40-61, 2018.

(Links | BibTeX | Tags: Affective Computing, Discourse, Signal Annotation and Interpretation)

Stepanov A. E., Chowdhury A. S., Bayer A. O., Ghosh A., Klasinas I., Calvo M., Sanchis E. and Riccardi G.

Cross-Language Transfer of Semantic Annotation via Targeted Crowdsourcing: Task Design and Evaluation (Article)

Language Resources and Evaluation, https://doi.org/10.1007/s10579-017-9396-5 , Springer, 2017, 2017.

(Abstract | Links | BibTeX | Tags: Signal Annotation and Interpretation)

@article{E.2017,
title = {Cross-Language Transfer of Semantic Annotation via Targeted Crowdsourcing: Task Design and Evaluation},
author = {Stepanov A. E., Chowdhury A. S., Bayer A. O., Ghosh A., Klasinas I., Calvo M., Sanchis E. and Riccardi G.},
url = {https://sisl.disi.unitn.it/wp-content/uploads/2017/10/10.1007s10579-017-9396-5.pdf},
year = {2017},
date = {2017-01-01},
journal = {Language Resources and Evaluation, https://doi.org/10.1007/s10579-017-9396-5 , Springer, 2017},
abstract = {Modern data-driven spoken language systems (SLS) require manual semantic annotation for training spoken language understanding parsers. Multilingual porting of SLS demands significant manual effort and language resources, as this manual annotation has to be replicated. Crowdsourcing is an accessible and cost-effective alternative to traditional methods of collecting and annotating data. The application of crowdsourcing to simple tasks has been well investigated. However, complex tasks, like cross-language semantic annotation transfer, may generate low judgment agreement and/or poor performance. The most serious issue in cross-language porting is the absence of reference annotations in the target language; thus, crowd quality control and the evaluation of the collected annotations is difficult. In this paper we investigate targeted crowdsourcing for semantic annotation transfer that delegates to crowds a complex task such as segmenting and labeling of concepts taken from a domain ontology; and evaluation using source language annotation. To test the applicability and effectiveness of the crowdsourced annotation transfer we have considered the case of close and distant language pairs: Italian–Spanish and Italian–Greek. The corpora annotated via crowdsourcing are evaluated against source and target language expert annotations. We demonstrate that the two evaluation references (source and target) highly correlate with each other; thus, drastically reduce the need for the target language reference annotations.
},
keywords = {Signal Annotation and Interpretation}
}

Close

Celli F., Ghosh A., Alam F. and Riccardi G.

In the mood for Sharing Contents: Emotions, personality and interaction styles in the diffusion of news (Article)

Information Processing and Management, Nov 2015, 2015.

(Abstract | Links | BibTeX | Tags: Machine Learning, Natural Language Processing, Signal Annotation and Interpretation)

Han S., Dinarelli M., Raymond C., Lefevre F., Lehnen P., De Mori R., Moschitti A., Ney H. and Riccardi G.

Comparing Stochastic Approaches to Spoken Language Understanding in Multiple Languages (Article)

IEEE Trans. on Audio, Speech and Language Processing, vol. 19, no. 6, pp. 1569-1583, 2011, 2014.

(Abstract | Links | BibTeX | Tags: Signal Annotation and Interpretation, Speech Processing)

@article{S.2014,
title = {Comparing Stochastic Approaches to Spoken Language Understanding in Multiple Languages},
author = {Han S., Dinarelli M., Raymond C., Lefevre F., Lehnen P., De Mori R., Moschitti A., Ney H. and Riccardi G.},
url = {https://sisl.disi.unitn.it/wp-content/uploads/2014/11/IEEETSLP10-MultSLU.pdf},
year = {2014},
date = {2014-01-01},
journal = {IEEE Trans. on Audio, Speech and Language Processing, vol. 19, no. 6, pp. 1569-1583, 2011},
abstract = {One of the first steps in building a spoken language understanding (SLU) module for dialogue systems is the extraction of flat concepts out of a given word sequence, usually provided by an automatic speech recognition (ASR) system. In this paper, six different modeling approaches are investigated to tackle the task of concept tagging. These methods include classical, well-known generative and discriminative methods like Finite State Transducers (FSTs), Statistical Machine Translation (SMT), Maximum Entropy Markov Models (MEMMs), or Support Vector Machines (SVMs) as well as techniques recently applied to natural language processing such as Conditional Random Fields (CRFs) or Dynamic Bayesian Networks (DBNs). Following a detailed description of the models, experimental and comparative results are presented on three corpora in different languages and with different complexity. The French MEDIA corpus has already been exploited during an evaluation campaign and so a direct comparison with existing benchmarks is possible. Recently collected Italian and Polish corpora are used to test the robustness and portability of the modeling approaches. For all tasks, manual transcriptions as well as ASR inputs are considered. Additionally to single systems, methods for system combination are investigated. The best performing model on all tasks is based on conditional random fields. On the MEDIA evaluation corpus, a concept error rate of 12.6% could be achieved. Here, additionally to attribute names, attribute values have been extracted using a combination of a rule-based and a statistical approach. Applying system combination using weighted ROVER with all six systems, the concept error rate (CER) drops to 12.0%.},
keywords = {Signal Annotation and Interpretation, Speech Processing}
}

Close

Dinarelli M., Moschitti A. and Riccardi G.

Discriminative Reranking for Spoken Language Understanding (Article)

IEEE Trans. on Audio, Speech and Language Processing, vol. 20, no. 2, pp. 526-539, 2012, 2012.

(Abstract | Links | BibTeX | Tags: Signal Annotation and Interpretation, Speech Processing)