Paxa Labs

งานวิจัย

ภาษาไทยไม่ได้ขาดข้อความ เสียง หรือเอกสาร แต่ขาดข้อมูลที่มีป้ายกำกับ เราจึงสร้างข้อมูลกำกับขึ้นเอง แล้วเผยแพร่ว่าอะไรใช้ได้จริง

arXiv:2609.03502v118 หน้า

A Thai voice from fifteen seconds of speech

An 82M model small enough to run on the device, trained entirely on speech a larger model generated, and the benchmark that says what it still gets wrong.

Kunat Pipatanakul, Potsawee Manakul, Warit Sirichotedumrong, Sittipong Sripaisarnmongkol, Pakorn Nathong, Phatrasek Jirabovonvisut

arXiv(opens in a new tab)PDF(opens in a new tab)

Three of the twelve teacher voices. None was recorded.

arXiv:2609.03595v120 หน้า

Thai OCR trained without a single real label

A 0.9B document model taught on 45,723 reconstructed pages, and the controlled experiments that say which part of a synthetic document actually transfers.

Kunat Pipatanakul

arXiv(opens in a new tab)PDF(opens in a new tab)

An annual-report page whose original text has been erased and replaced with rendered Thai in the same regions.
One page, reconstructed from a public source document.

สิ่งที่เผยแพร่

โมเดล

ชุดทดสอบและชุดข้อมูล

โค้ด

ตัวอย่างใช้งาน

ทำงานด้านเทคโนโลยีภาษาไทยอยู่ใช่ไหม

งานนี้เป็นความร่วมมือระหว่าง Wayu Research และ Paxa Labs โดย Wayu Research สนับสนุนทุนเอง ทีม Typhoon ร่วมเขียนรายงานด้านเสียงพูด และให้ความเห็นต่อรายงานด้านเอกสาร