paper-with-me

Papers

Fauno: The Italian Large Language Model that will leave you senza parole!

2023-06-26 · Andrea Bacciu, Giovanni Trappolini, Andrea Santilli, Emanuele Rodolà, Fabrizio Silvestri

This paper presents Fauno, the first and largest open-source Italian conversational Large Language Model (LLM). Our goal with Fauno is to democratize the study of LLMs in Italian, demonstrating that obtaining a fine-tuned conversational bot with a single GPU is possible. In addition, we release a collection of datasets for conversational AI in Italian. The datasets on which we fine-tuned Fauno include various topics such as general question answering, computer science, and medical questions. We release our code and datasets on \url{https://github.com/RSTLess-research/Fauno-Italian-LLM}

📄 PDF Abstract BibTeX arXiv:2306.14457

Code (1)

rstless-research/fauno-italian-llm 공식 구현 pytorch

Tasks

GPULanguage ModelingLanguage ModellingLarge Language ModelQuestion Answering

Similar Papers 제목 키워드 기반

Cerbero-7B: A Leap Forward in Language-Specific LLMs Through Enhanced Chat Corpus Generation and Evaluation

2023-11-27 · Federico A. Galatolo, Mario G. C. A. Cimino

This study introduces a novel approach for generating high-quality, language-specific chat corpora using a self-chat mechanism. We combine a generator LLM for creating new samples and an embedder LLM to ensure diversity.…

DiversityLanguage ModellingQuestion AnsweringSentence

FAuNO: Semi-Asynchronous Federated Reinforcement Learning Framework for Task Offloading in Edge Systems

2025-06-03 · Frederico Metelo, Alexandre Oliveira, Stevo Racković, Pedro Ákos Costa 외

Edge computing addresses the growing data demands of connected-device networks by placing computational resources closer to end users through decentralized infrastructures. This decentralization challenges traditional, f…

Edge-computing

The Invalsi Benchmarks: measuring Linguistic and Mathematical understanding of Large Language Models in Italian

2024-03-27 · Giovanni Puccetti, Maria Cassese, Andrea Esuli

While Italian is a high-resource language, there are few Italian-native benchmarks to evaluate generative Large Language Models (LLMs) in this language. This work presents three new benchmarks: Invalsi MATE to evaluate m…

Language ModellingMath

HATE-ITA: New Baselines for Hate Speech Detection in Italian

2022-07-01 · NAACL (WOAH) 2022 7 · Debora Nozza, Federico Bianchi, Giuseppe Attanasio

Online hate speech is a dangerous phenomenon that can (and should) be promptly counteracted properly. While Natural Language Processing supplies appropriate algorithms for trying to reach this objective, all research eff…

BenchmarkingHate Speech DetectionLanguage ModelingLanguage Modelling

Testimole-Conversational: A 30-Billion-Word Italian Discussion Board Corpus (1996-2024) for Language Modeling and Sociolinguistic Research

2026-02-16 · Matteo Rinaldi, Rossella Varvara, Viviana Patti arxiv

We present "Testimole-conversational" a massive collection of discussion boards messages in the Italian language. The large size of the corpus, more than 30B word-tokens (1996-2024), renders it an ideal dataset for nativ…

Language ModellingDomain Adaptation