paper-with-me

Papers

Adapting Safe-for-Work Classifier for Malaysian Language Text: Enhancing Alignment in LLM-Ops Framework

2024-07-30 · Aisyah Razak, Ariff Nazhan, Kamarul Adha, Wan Adzhar Faiq Adzlan, Mas Aisyah Ahmad, Ammar Azman

As large language models (LLMs) become increasingly integrated into operational workflows (LLM-Ops), there is a pressing need for effective guardrails to ensure safe and aligned interactions, including the ability to detect potentially unsafe or inappropriate content across languages. However, existing safe-for-work classifiers are primarily focused on English text. To address this gap for the Malaysian language, we present a novel safe-for-work text classifier tailored specifically for Malaysian language content. By curating and annotating a first-of-its-kind dataset of Malaysian text spanning multiple content categories, we trained a classification model capable of identifying potentially unsafe material using state-of-the-art natural language processing techniques. This work represents an important step in enabling safer interactions and content filtering to mitigate potential risks and ensure responsible deployment of LLMs. To maximize accessibility and promote further research towards enhancing alignment in LLM-Ops for the Malaysian context, the model is publicly released at https://huggingface.co/malaysia-ai/malaysian-sfw-classifier.

📄 PDF Abstract BibTeX arXiv:2407.20729

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Bridging the Gap: Transfer Learning from English PLMs to Malaysian English

2024-07-01 · Mohan Raj Chanthran, Lay-Ki Soon, Huey Fang Ong, Bhawani Selvaretnam

Malaysian English is a low resource creole language, where it carries the elements of Malay, Chinese, and Tamil languages, in addition to Standard English. Named Entity Recognition (NER) models underperform when capturin…

Language Modellingnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)+2

Large Malaysian Language Model Based on Mistral for Enhanced Local Language Understanding

2024-01-24 · Husein Zolkepli, Aisyah Razak, Kamarul Adha, Ariff Nazhan

In this paper, we present significant advancements in the pretraining of Mistral 7B, a large-scale language model, using a dataset of 32.6 GB, equivalent to 1.1 billion tokens. We explore the impact of extending the cont…

BenchmarkingLanguage ModelingLanguage Modelling

Malaysian English News Decoded: A Linguistic Resource for Named Entity and Relation Extraction

2024-02-22 · Mohan Raj Chanthran, Lay-Ki Soon, Huey Fang Ong, Bhawani Selvaretnam

Standard English and Malaysian English exhibit notable differences, posing challenges for natural language processing (NLP) tasks on Malaysian English. Unfortunately, most of the existing datasets are mainly based on sta…

Articlesnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)+3

Multi-Lingual Malaysian Embedding: Leveraging Large Language Models for Semantic Representations

2024-02-05 · Husein Zolkepli, Aisyah Razak, Kamarul Adha, Ariff Nazhan

In this work, we present a comprehensive exploration of finetuning Malaysian language models, specifically Llama2 and Mistral, on embedding tasks involving negative and positive pairs. We release two distinct models tail…

RAGRetrievalRetrieval-augmented GenerationSemantic Similarity+1

MaLLaM -- Malaysia Large Language Model

2024-01-26 · Husein Zolkepli, Aisyah Razak, Kamarul Adha, Ariff Nazhan

Addressing the gap in Large Language Model pretrained from scratch with Malaysian context, We trained models with 1.1 billion, 3 billion, and 5 billion parameters on a substantial 349GB dataset, equivalent to 90 billion …

Language ModelingLanguage ModellingLarge Language Modelmodel+1