arxiv:2412.13702

Typhoon 2: A Family of Open Text and Multimodal Thai Large Language Models

Published on Dec 18, 2024

Typhoon

Authors:

Pittawat Taveekitworachai ,

Adisai Na-Thalang ,

Abstract

Typhoon 2 is a series of Thai-text-optimized large language models, including text, vision, and audio variants, with enhanced performance and safety features.

AI-generated summary

This paper introduces Typhoon 2, a series of text and multimodal large language models optimized for the Thai language. The series includes models for text, vision, and audio. Typhoon2-Text builds on state-of-the-art open models, such as Llama 3 and Qwen2, and we perform continual pre-training on a mixture of English and Thai data. We employ post-training techniques to enhance Thai language performance while preserving the base models' original capabilities. We release text models across a range of sizes, from 1 to 70 billion parameters, available in both base and instruction-tuned variants. To guardrail text generation, we release Typhoon2-Safety, a classifier enhanced for Thai cultures and language. Typhoon2-Vision improves Thai document understanding while retaining general visual capabilities, such as image captioning. Typhoon2-Audio introduces an end-to-end speech-to-speech model architecture capable of processing audio, speech, and text inputs and generating both text and speech outputs.

View arXiv page View PDF GitHub 34 auto Add to collection

Community

Upload images, audio, and videos by dragging in the text input, pasting, or clicking here.

Tap or paste here to upload images

· Sign up or log in to comment

Upvote

Get this paper in your agent:

hf papers read 2412.13702

Don't have the latest CLI?

curl -LsSf https://hf.co/cli/install.sh | bash

Models citing this paper 32

Browse 32 models citing this paper

Datasets citing this paper 0

No dataset linking this paper

Cite arxiv.org/abs/2412.13702 in a dataset README.md to link it from this page.

Typhoon 2: A Family of Open Text and Multimodal Thai Large Language Models

Abstract

Community

Models citing this paper 32

Datasets citing this paper 0

Spaces citing this paper 10

Collections including this paper 3