Search results

(1 - 20 of 34)

Pages

Dijk, B.M.A. van; Duijn, M.J. van; Verberne, S.; Spruit, M. 2023

ChiSCor: a corpus of freely-told fantasy stories by Dutch children for computational linguistics and cognitive science

Article in monograph or in proceedings

open access

In this resource paper we release ChiSCor, a new corpus containing 619 fantasy stories, told freely by 442 Dutch children aged 4-12. ChiSCor was compiled for studying how children render character... Show moreIn this resource paper we release ChiSCor, a new corpus containing 619 fantasy stories, told freely by 442 Dutch children aged 4-12. ChiSCor was compiled for studying how children render character perspectives, and unravelling language and cognition in development, with computational tools. Unlike existing resources, ChiSCor’s stories were produced in natural contexts, in line with recent calls for more ecologically valid datasets. ChiSCor hosts text, audio, and annotations for character complexity and linguistic complexity. Additional metadata (e.g. education of caregivers) is available for one third of the Dutch children. ChiSCor also includes a small set of 62 English stories. This paper details how ChiSCor was compiled and shows its potential for future work with three brief case studies: i) we show that the syntactic complexity of stories is strikingly stable across children’s ages; ii) we extend work on Zipfian distributions in free speech and show that ChiSCor obeys Zipf’s law closely, reflecting its social context; iii) we show that even though ChiSCor is relatively small, the corpus is rich enough to train informative lemma vectors that allow us to analyse children’s language use. We end with a reflection on the value of narrative datasets in computational linguistics. Show less

Tseng, R.; Verberne, S.; Putten, P.W.H. van der 2023

ChatGPT as a commenter to the news: can LLMs generate human-like opinions?

Article in monograph or in proceedings

open access

Abolghasemi, M.A.; Verberne, S.; Askari, A.; Azzopardi, L. 2023

Retrievability bias estimation using synthetically generated queries

Article in monograph or in proceedings

open access

Askari, A.; Abolghasemi, M.A.; Pasi, G.; Kraaij, W.; Verberne, S. 2023

Injecting the BM25 score as text improves BERT-based re-rankers

Article in monograph or in proceedings

open access

Askari, A.; Aliannejadi, M.; Abolghasemi, M.A.; Kanoulas, E.; Verberne, S. 2023

CLosER: conversational legal longformer with expertise-aware passage response ranker for long contexts

Article in monograph or in proceedings

open access

Althammer, S.; Zuccon, G.; Hofstätter, S.; Verberne, S.; Hanbury, A. 2023

Annotating data for fine-tuning a neural ranker? Current active learning strategies are not better than random selection

Article in monograph or in proceedings

open access

Askari, A.; Aliannejadi, M.; Kanoulas, E.; Verberne, S. 2023

A test collection of synthetic documents for training rankers: ChatGPT vs. human experts

Article in monograph or in proceedings

open access

Schouten, M.; Ramackers, G.J.; Verberne, S. 2022

Preprocessing requirements documents for automatic UML modelling

Article in monograph or in proceedings

open access

Askari, A.; Verberne, S.; Pasi, G. 2022

Expert finding in legal community question answering

Article in monograph or in proceedings

open access

Abolghasemi, M.A.; Askari, A.; Verberne, S. 2022

On the interpolation of contextualized term-based ranking with BM25 for query-by-example retrieval

Article in monograph or in proceedings

open access

Congleton, C.; Putten, P.W.H. van der; Verberne, S. 2022

Tracing political positioning of Dutch newspapers

Article in monograph or in proceedings

open access

Abolghasemi, M.A.; Verberne, S.; Azzopardi, L. 2022

Improving BERT-based query-by-document retrieval with multi-task optimization

Article in monograph or in proceedings

open access

Althammer, S.; Hofstätter, S.; Sertkan, M.; Verberne, S.; Hanbury, A. 2022

PARM: a Paragraph Aggregation Retrieval Model for dense document-to-document retrieval

Article in monograph or in proceedings

open access

Althammer, S.; Hofstätter, S.; Verberne, S.; Hanbury, A. 2022

TripJudge: a relevance judgement test collection for TripClick health retrieval

Article in monograph or in proceedings

open access

Althammer, S.; Askari, A.; Verberne, S.; Hanbury, A.H. 2021

DoSSIER@ COLIEE 2021: Leveraging dense retrieval and summarization-based re-ranking for case law retrieval

Article in monograph or in proceedings

open access

In this paper, we present our approaches for the case law retrieval and the legal case entailment task in the Competition on Legal Information Extraction/Entailment (COLIEE) 2021. As first stage... Show moreIn this paper, we present our approaches for the case law retrieval and the legal case entailment task in the Competition on Legal Information Extraction/Entailment (COLIEE) 2021. As first stage retrieval methods combined with neural re-ranking methods us- ing contextualized language models like BERT achieved great performance improvements for information retrieval in the web and news domain, we evaluate these methods for the legal domain. A distinct characteristic of legal case retrieval is that the query case and case description in the corpus tend to be long documents and therefore exceed the input length of BERT. We address this challenge by combining lexical and dense retrieval methods on the paragraph-level of the cases for the first stage retrieval. Here we demonstrate that the retrieval on the paragraph-level outperforms the retrieval on the document-level. Furthermore the experiments suggest that dense retrieval methods outperform lexical retrieval. For re-ranking we address the problem of long documents by sum- marizing the cases and fine-tuning a BERT-based re-ranker with the summaries. Overall, our best results were obtained with a com- bination of BM25 and dense passage retrieval using domain-specific embeddings.DoSSIER@ COLIEE 2021 Show less

Reuver, M.; Fokkens, A.; Verberne, S. 2021

No NLP task should be an island: multi-disciplinarity for diversity in news recommender systems

Article in monograph or in proceedings

open access

Voskuil, K.; Verberne, S. 2021

Improving reference mining in patents with BERT

Article in monograph or in proceedings

open access

Reuver, M.; Verberne, S.; Morante, R.; Fokkens, A. 2021

Is stance detection topic-independent and cross-topic generalizable? - A reproduction study

Article in monograph or in proceedings

open access

Vos, H.P. de; Verberne, S. 2021

Small data problems in political research: a critical replication study

Article in monograph or in proceedings

open access

Reuver, R.; Mattis, N.; Sax, M.; Verberne, S.; Tintarev, N.; Helberger, N.; ... ; Atteveldt, W. 2021

Are we human, or are we users? The role of natural language processing in human-centric news recommenders that nudge users to diverse content

Article in monograph or in proceedings

open access

Leiden University Scholarly Publications

Your Search

Enabled Filters

Sort Options

Refine Results

Resource Type

Availability

Creation Date

Faculty

Collection

Author

Language

Search results

Pages

Pages