Publications (46)
2026
4 publications- Safety of Large Language Models Beyond English: A Systematic Literature Review of Risks, Biases, and Safeguards
2026
- Annotation-Efficient Vision-Language Model Adaptation to the Polish Language Using the LLaVA Framework
2026
- Talmud-IR: A Talmud-Inspired Interface for Discussing RAG Response Quality
Lecture notes in computer science · 2026
- The LLM Effect on IR Benchmarks: A Meta-Analysis of Effectiveness, Baselines, and Contamination
2026
2025
9 publications- Aurora-M: Open Source Continual Pre-training for Multilingual Language and Code
Proceedings of the 31st International Conference on Computational · 2025
- ATOM at CheckThat! 2025: retrieve the implicit-scientific evidence retrieval
Faggioli et al · 2025
- ConECT Dataset: Overcoming Data Scarcity in Context-Aware E-Commerce MT
2025
- PL-Guard: Benchmarking Language Model Safety for Polish
2025
- ASPIRE: Assistive System for Performance Evaluation in IR
Lecture notes in computer science · 2025
- Compare: A Framework for Scientific Comparisons
2025
- Rainbow-Teaming for the Polish Language: A Reproducibility Study
2025
- ReAct-ExtrAct: A Tool for Source-Grounded Automated Data Extraction in Systematic Reviews
2025 ACM/IEEE Joint Conference on Digital Libraries (JCDL), 340-343 · 2025
- TimIR: Time-Traveling Through IR History
Lecture notes in computer science · 2025
2024
14 publications- Learning to match patients to clinical trials using large language models
Journal of Biomedical Informatics · 2024
- Computer-assisted screening in systematic evidence synthesis requires robust and well-evaluated stopping criteria
Systematic Reviews · 2024
- Large language models, updates, and evaluation of automation tools for systematic reviews: a summary of significant discussions at the eighth meeting of the International Collaboration for the Automation of Systematic Reviews (ICASR)
Systematic Reviews · 2024
- A Reproducibility and Generalizability Study of Large Language Models for Query Generation
2024
- An Analysis of Tasks and Datasets in Peer Reviewing
2024
- ORCAS-I query intent predictor as component of TIRA
1st International Workshop on Open Web Search (WOWS) at ECIR 2024 · 2024
- Cross-Linguistic Disease and Drug Detection in Cardiology Clinical Texts: Methods and Outcomes
CLEF Working Notes · 2024
- Enhancing Clinical Data Capture: Developing a Natural Language Processing Pipeline for Converting Free Text Admission Notes to Structured EHR Data
Proceedings of the Eighth Workshop on Natural Language for Artificial · 2024
- Leveraging Cochrane Systematic Literature Reviews for Prospective Evaluation of Large Language Models.
ALTARS at ECIR · 2024
- Normalised Precision at Fixed Recall for Evaluating TAR
2024
- AustroTox: A Dataset for Target-Based Austrian German Offensive Language Detection
2024
- Adherence to the PRISMA 2020 guidelines and the use of automation tools in the random sample of 1000 system...
2024
- Adherence to the PRISMA 2020 guidelines and the use of automation tools in the random sample of 1000 systematic reviews: a meta-epidemiological study
2024
- Automated Eligibility Screening and its Evaluation in the Medical Domain
Technische Universität Wien · 2024
2023
10 publications- BLOOM: A 176B-Parameter Open-Access Multilingual Language Model
HAL (Le Centre pour la Communication Scientifique Directe) · 2023
- Effective matching of patients to clinical trials using entity extraction and neural re-ranking
Journal of Biomedical Informatics · 2023
- An analysis of work saved over sampling in the evaluation of automated citation screening in systematic literature reviews
Intelligent Systems with Applications · 2023
- CRUISE-Screening: Living Literature Reviews Toolbox
2023
- Outcome-based Evaluation of Systematic Review Automation
2023
- VoMBaT: A Tool for Visualising Evaluation Measure Behaviour in High-Recall Search Tasks
2023
- “Dr LLM, what do I have?”: The Impact of User Beliefs and Prompt Formulation on Health Diagnoses
2023
- CSMeD: Bridging the Dataset Gap in Automated Citation Screening for Systematic Literature Reviews
2023
- HEVS-TUW at SemEval-2023 Task 8: Ensemble of Language Models and Rule-based Classifiers for Claims Identification and PICO Extraction
2023
- Statute-enhanced lexical retrieval of court cases for COLIEE 2022
arXiv (Cornell University) · 2023
2022
8 publications- BLOOM: A 176B-Parameter Open-Access Multilingual Language Model
arXiv (Cornell University) · 2022
- ORCAS-I: Queries Annotated with Intent using Weak Supervision
Proceedings of the 45th International ACM SIGIR Conference on Research and · 2022
- Automation of Citation Screening for Systematic Literature Reviews Using Neural Networks: A Replicability Study
Lecture notes in computer science · 2022
- Benchmark for research theme classification of scholarly documents
Proceedings of the Third Workshop on Scholarly Document Processing, 253-262 · 2022
- DoSSIER at MedVidQA 2022: Text-based Approaches to Medical Video Answer Localization Problem
2022
- Dataset Debt in Biomedical Language Modeling
2022
- BigBio: A Framework for Data-Centric Biomedical Natural Language Processing
2022
- Evaluation of Automated Citation Screening in Systematic Literature Reviews with Work Saved over Sampling: an Analysis
ALTARS · 2022