Bridging Causal Theory and LLMs: A User Centric Approach to Causal Graph Generation
Doctoral Consortium @ AI*IA 2023 · 2023
Doctoral consortium paper on user-centric causal graph generation with LLMs.
Doctoral Consortium @ AI*IA 2023 · 2023
Doctoral consortium paper on user-centric causal graph generation with LLMs.
AIABI Workshop @ AI*IA 2023 · 2023
Combining LLMs and domain experts to build causal graphs.
arXiv preprint · 2024
Evaluating LLMs on the INVALSI Italian benchmark.
World Conference on Explainable Artificial Intelligence (xAI 2024) · 2024
Using LLMs to make explanations of a banking recommender system more usable.
Proceedings of CLiC-it 2024 · 2024
A CALAMITA challenge evaluating LLMs on the Italian driver’s license exam.
arXiv preprint · 2025
Preprint on designing role vectors to improve LLM inference behaviour.
Proceedings of NAACL 2025 (Long Papers) · 2025
A benchmark for evaluating LLMs on Italian language and culture.
ECML PKDD 2025 · 2025
Evaluating LLMs on the competencies measured in Italian student assessments.
Proceedings of EMNLP 2025: Industry Track · 2025
Scores for evaluating whether automatically generated explanations of sparse autoencoder features align with what the features actually do.
Findings of the Association for Computational Linguistics: EMNLP 2025 · 2025
Role vectors: activation directions associated with roles that can steer LLM behaviour.
Italian Journal of Computational Linguistics (IJCoL) · 2026
The CALAMITA community initiative for evaluating LLMs in Italian.
Proceedings of EMNLP 2026 (Main Conference) · 2026
Accepted at EMNLP 2026 (main conference). An empirical study of reasoning entropy and malicious outputs in LLMs.