Het platform voor open praktijkgericht onderzoek

Producten 325

product

Considering Human Interaction and Variability in Automatic Text Simplification

Research into automatic text simplification aims to promote access to information for all members of society. To facilitate generalizability, simplification research often abstracts away from specific use cases, and targets a prototypical reader and an underspecified content creator. In this paper, we consider a real-world use case – simplification technology for use in Dutch municipalities – and identify the needs of the content creators and the target audiences in this scenario. The stakeholders envision a system that (a) assists the human writer without taking over the task; (b) provides diverse outputs, tailored for specific target audiences; and (c) explains the suggestions that it outputs. These requirements call for technology that is characterized by modularity, explainability, and variability. We argue that these are important research directions that require further exploration

MULTIFILE

Considering Human Interaction and Variability in Automatic Text Simplification

product

A Classification of Modification Categories for Business Rules

Author supplied Business rules play a critical role in an organization’s daily activities. With the increased use of business rules (solutions) the interest in modelling guidelines that address the manageability of business rules has increased as well. However, current research on modelling guidelines is mainly based on a theoretical view of modifications that can occur to a business rule set. Research on actual modifications that occur in practice is limited. The goal of this study is to identify modifications that can occur to a business rule set and underlying business rules. To accomplish this goal we conducted a grounded theory study on 229 rules set, as applied from March 2006 till June 2014, by the National Health Service. In total 3495 modifications have been analysed from which we defined eleven modification categories that can occur to a business rule set. The classification provides a framework for the analysis and design of business rules management architectures.

DOCUMENT

A Classification of Modification Categories for Business Rules

product

Keyword extraction using co-occurrence.

A common strategy to assign keywords to documents is to select the most appropriate words from the document text. One of the most important criteria for a word to be selected as keyword is its relevance for the text. The tf.idf score of a term is a widely used relevance measure. While easy to compute and giving quite satisfactory results, this measure does not take (semantic) relations between words into account. In this paper we study some alternative relevance measures that do use relations between words. They are computed by defining co-occurrence distributions for words and comparing these distributions with the document and the corpus distribution. We then evaluate keyword extraction algorithms defined by selecting different relevance measures. For two corpora of abstracts with manually assigned keywords, we compare manually extracted keywords with different automatically extracted ones. The results show that using word co-occurrence information can improve precision and recall over tf.idf.

DOCUMENT

product

Automatic categorization of self-acknowledged limitations in randomized controlled trial publications

Objective:Acknowledging study limitations in a scientific publication is a crucial element in scientific transparency and progress. However, limitation reporting is often inadequate. Natural language processing (NLP) methods could support automated reporting checks, improving research transparency. In this study, our objective was to develop a dataset and NLP methods to detect and categorize self-acknowledged limitations (e.g., sample size, blinding) reported in randomized controlled trial (RCT) publications.Methods:We created a data model of limitation types in RCT studies and annotated a corpus of 200 full-text RCT publications using this data model. We fine-tuned BERT-based sentence classification models to recognize the limitation sentences and their types. To address the small size of the annotated corpus, we experimented with data augmentation approaches, including Easy Data Augmentation (EDA) and Prompt-Based Data Augmentation (PromDA). We applied the best-performing model to a set of about 12K RCT publications to characterize self-acknowledged limitations at larger scale.Results:Our data model consists of 15 categories and 24 sub-categories (e.g., Population and its sub-category DiagnosticCriteria). We annotated 1090 instances of limitation types in 952 sentences (4.8 limitation sentences and 5.5 limitation types per article). A fine-tuned PubMedBERT model for limitation sentence classification improved upon our earlier model by about 1.5 absolute percentage points in F1 score (0.821 vs. 0.8) with statistical significance (). Our best-performing limitation type classification model, PubMedBERT fine-tuning with PromDA (Output View), achieved an F1 score of 0.7, improving upon the vanilla PubMedBERT model by 2.7 percentage points, with statistical significance ().Conclusion:The model could support automated screening tools which can be used by journals to draw the authors’ attention to reporting issues. Automatic extraction of limitations from RCT publications could benefit peer review and evidence synthesis, and support advanced methods to search and aggregate the evidence from the clinical trial literature.

MULTIFILE

Automatic categorization of self-acknowledged limitations in randomized controlled trial publications

product

Toward assessing clinical trial publications for reporting transparency

Objective: To annotate a corpus of randomized controlled trial (RCT) publications with the checklist items of CONSORT reporting guidelines and using the corpus to develop text mining methods for RCT appraisal. Methods: We annotated a corpus of 50 RCT articles at the sentence level using 37 fine-grained CONSORT checklist items. A subset (31 articles) was double-annotated and adjudicated, while 19 were annotated by a single annotator and reconciled by another. We calculated inter-annotator agreement at the article and section level using MASI (Measuring Agreement on Set-Valued Items) and at the CONSORT item level using Krippendorff's α. We experimented with two rule-based methods (phrase-based and section header-based) and two supervised learning approaches (support vector machine and BioBERT-based neural network classifiers), for recognizing 17 methodology-related items in the RCT Methods sections. Results: We created CONSORT-TM consisting of 10,709 sentences, 4,845 (45%) of which were annotated with 5,246 labels. A median of 28 CONSORT items (out of possible 37) were annotated per article. Agreement was moderate at the article and section levels (average MASI: 0.60 and 0.64, respectively). Agreement varied considerably among individual checklist items (Krippendorff's α= 0.06–0.96). The model based on BioBERT performed best overall for recognizing methodology-related items (micro-precision: 0.82, micro-recall: 0.63, micro-F1: 0.71). Combining models using majority vote and label aggregation further improved precision and recall, respectively. Conclusion: Our annotated corpus, CONSORT-TM, contains more fine-grained information than earlier RCT corpora. Low frequency of some CONSORT items made it difficult to train effective text mining models to recognize them. For the items commonly reported, CONSORT-TM can serve as a testbed for text mining methods that assess RCT transparency, rigor, and reliability, and support methods for peer review and authoring assistance. Minor modifications to the annotation scheme and a larger corpus could facilitate improved text mining models. CONSORT-TM is publicly available at https://github.com/kilicogluh/CONSORT-TM.

DOCUMENT

product

Take out what you can

The goal of this study was therefore to test the idea that computationally analysing the Fontys National Student Surveys (NSS) open answers using a selection of standard text mining methods (Manning & Schütze 1999) will increase the value of these answers for educational quality assurance. It is expected that human effort and time of analysis will decrease significally. The text data (in Dutch) of several years of Fontys National Student Surveys (2013-2018) was provided to Fontys students of the minor Applied Data Science. The results of the analysis were to include topic and sentiment modelling across multiple years of survey data. Comparing multiple years was necessary to capture and visualize any trends that a human investigator may have missed while analysing the data by hand. During data cleaning all stop words and punctuation were removed, all text was brought to a lower case, names and inappropriate language – such as swear words – were deleted. About 80% of 24.000 records were manually labelled with sentiment; reminder was used for algorithms’ validation. In the following step a machine learning analysis steps: training, testing, outcomes analysis and visualisation, for a better text comprehension, were executed. Students aimed to improve classification accuracy by applying multiple sentiment analysis algorithms and topics modelling methods. The models were chosen arbitrarily, with a preference for a low complexity of a model. For reproducibility of our study open source tooling was used. One of these tools was based on Latent Dirichlet allocation (LDA). LDA is a generative statistical model that allows sets of observations to be explained by unobserved groups that explain why some parts of the data are similar (Blei, Ng & Jordan, 2003). For topic modelling the Gensim (Řehůřek, 2011) method was used. Gensim is an open-source vector space modelling and topic modelling toolkit implemented in Python. In addition, we recognized the absence of pretrained models for Dutch language. To complete our prototype a simple user interface was created in Python. This final step integrated our automated text analysis with visualisations of sentiments and topics. Remarkably, all extracted topics are related to themes defined by the NSS. This indicates that in general students’ answers are related to topics of interest for educational institutions. The extracted list of the words related to the topic is also relevant to this topic. Despite the fact that most of the results require further human expert interpretation, it is indicative to conclude that the computational analysis of the texts from the open questions of the NSS contain information which enriches the results of standard quantitative analysis of the NSS.

DOCUMENT

product

parents' role in enabling the participation of their child with a physical disability

Parental involvement is a crucial force in children’s development, learning and success at school and in life [1]. Participation, defined by the World Health Organization as ‘a person’s involvement in life situations’ [2] for children means involvement in everyday activities, such as recreational, leisure, school and household activities [3]. Several authors use the term social participation emphasising the importance of engagement in social situations [4, 5]. Children’s participation in daily life is vital for healthy development, social and physical competencies, social-emotional well-being, sense of meaning and purpose in life [6]. Through participation in different social contexts, children gather the knowledge and skills needed to interact, play, work, and live with other people [4, 7, 8]. Unfortunately, research shows that children with a physical disability are at risk of lower participation in everyday activities [9]; they participate less frequently in almost all activities compared with children without physical disabilities [10, 11], have fewer friends and often feel socially isolated [12-14]. Parents, in particular, positively influence the participation of their children with a physical disability at school, at home and in the community [15]. They undertake many actions to improve their child’s participation in daily life [15, 16]. However, little information is available about what parents of children with a physical disability do to enable their child’s participation, what they come across and what kind of needs they have. The overall aim of this thesis was to investigate parents’ actions, challenges, and needs while enhancing the participation of their school-aged child with a physical disability. In order to achieve this aim, two steps have been made. In the first step, the literature has been examined to explore the topic of this thesis (actions, challenges and needs) and to clarify definitions for the concepts of participation and social participation. Second, for the purposes of giving breadth and depth of understanding of the topic of this thesis a mixed methods approach using three different empirical research methods [17-19], was applied to gather information from parents regarding their actions, challenges and needs.

DOCUMENT

parents' role in enabling the participation of their child with a physical disability

Projecten 2

project

How to inherit stories? Artistic Research as Constructive and Critical Memory Work

This Professional Doctorate (PD) project explores the intersection of artistic research, digital heritage, and interactive media, focusing on the reimagining of medieval Persian bestiaries through high dark fantasy and game-making. The research investigates how the process of creation with interactive 3D media can function as a memory practice. At its core, the project treats bestiaries—pre-modern collections of real and imaginary classifications of the world—as a window into West and Central Asian flora, fauna, and the landscape of memory, serving as both repositories of knowledge and imaginative, cosmological accounts of the more-than-human world. As tools for exploring non-human pre-modern agency, bestiaries offer a medium of speculative storytelling, and explicate the unstable nature of memory in diasporic contexts. By integrating these themes into an interactive digital world, the research develops new methodologies for artistic research, treating world-building as a technique of attunement to heritage. Using a practice-based approach, the project aligns with MERIAN’s emphasis on "research in the wild," where artistic and scientific inquiries merge in experimental ways. It engages with hard-core game mechanics, mythopoetic decompressed environmental storytelling, and hand-crafted detailed intentional world-building to offer new ways of interacting with the past that challenges nostalgia and monumentalization. How can a cultural practice do justice to other, more experimental forms of remembering and encountering cultural pasts, particularly those that embrace the interconnections between human and non-human entities? Specifically, how can artistic practice, through the medium of a virtual, bestiary-inspired dark fantasy interactive media, allow for new modes of remembering that resist idealized and monumentalized histories? What forms of inquiry can emerge when technology (3D media, open-world interactive digital media) becomes a tool of attention and a site of experimental attunement to cosmological heritage?

Lopend

project

Responsible Applied Artificial Intelligence Trade-Off Dashboard

Organisations are increasingly embedding Artificial Intelligence (AI) techniques and tools in their processes. Typical examples are generative AI for images, videos, text, and classification tasks commonly used, for example, in medical applications and industry. One danger of the proliferation of AI systems is the focus on the performance of AI models, neglecting important aspects such as fairness and sustainability. For example, an organisation might be tempted to use a model with better global performance, even if it works poorly for specific vulnerable groups. The same logic can be applied to high-performance models that require a significant amount of energy for training and usage. At the same time, many organisations recognise the need for responsible AI development that balances performance with fairness and sustainability. This KIEM project proposal aims to develop a tool that can be employed by organizations that develop and implement AI systems and aim to do so more responsibly. Through visual aiding and data visualisation, the tool facilitates making these trade-offs. By showing what these values mean in practice, which choices could be made and highlighting the relationship with performance, we aspire to educate users on how the use of different metrics impacts the decisions made by the model and its wider consequences, such as energy consumption or fairness-related harms. This tool is meant to facilitate conversation between developers, product owners and project leaders to assist them in making their choices more explicit and responsible.

Lopend

Producten 325

product

Considering Human Interaction and Variability in Automatic Text Simplification

MULTIFILE

product

A Classification of Modification Categories for Business Rules

DOCUMENT

product

Keyword extraction using co-occurrence.

DOCUMENT

product

Automatic categorization of self-acknowledged limitations in randomized controlled trial publications

MULTIFILE

product

Toward assessing clinical trial publications for reporting transparency

DOCUMENT

product

Take out what you can

DOCUMENT

product

parents' role in enabling the participation of their child with a physical disability

DOCUMENT

Projecten 2

project

How to inherit stories? Artistic Research as Constructive and Critical Memory Work

Lopend

project

Responsible Applied Artificial Intelligence Trade-Off Dashboard

Lopend

Zoekresultaten

Producten 325

Considering Human Interaction and Variability in Automatic Text Simplification

A Classification of Modification Categories for Business Rules

Keyword extraction using co-occurrence.

Take out what you can: quantitative analysis of the open question results from the National Student Survey

Better than a sack full of Latin: Anticlericalism in the Middle Dutch 'Dit es de Frenesie'

Automatic categorization of self-acknowledged limitations in randomized controlled trial publications

Toward assessing clinical trial publications for reporting transparency

Take out what you can

parents' role in enabling the participation of their child with a physical disability

Projecten 2

How to inherit stories? Artistic Research as Constructive and Critical Memory Work

Responsible Applied Artificial Intelligence Trade-Off Dashboard

Navigeer naar

Categorieën

Filters

Producten 325

Considering Human Interaction and Variability in Automatic Text Simplification

A Classification of Modification Categories for Business Rules

Keyword extraction using co-occurrence.

Take out what you can: quantitative analysis of the open question results from the National Student Survey

Better than a sack full of Latin: Anticlericalism in the Middle Dutch 'Dit es de Frenesie'

Automatic categorization of self-acknowledged limitations in randomized controlled trial publications

Toward assessing clinical trial publications for reporting transparency

Take out what you can

parents' role in enabling the participation of their child with a physical disability

Projecten 2

How to inherit stories? Artistic Research as Constructive and Critical Memory Work

Responsible Applied Artificial Intelligence Trade-Off Dashboard