Skip to Main Content

Breadcrumb

Introduction

The Artificial Intelligence Evaluation Center (AIEC) has been established to promote localized AI evaluation and third-party certification in Taiwan, thereby strengthening the development of trusted AI within the industry. The Center will periodically publish benchmark evaluation results for language models. In addition to adopting indicators based on the Chinese Language and Social Studies sections of the national high school entrance examination, AIEC also incorporates evaluation criteria reflecting Taiwanese values, aligning with global trends in AI sovereignty. These benchmarks serve as key references for developing locally adapted models or fine-tuning international models.

✪ Developer Region: Light yellow indicates models developed in Europe; light blue, the United States; light green, Taiwan; light purple, China; light beige, other parts of Asia; and light orange, Africa.

✪ Percentage Scores: Scores above 50% are shown in black, while scores below 50% are shown in red.

✪ Starting in May 2026, a Model Evaluation Date column has been added for reference. Starting in September 2026, models are ranked by their aggregate score across the three foundational evaluation benchmark.

 
 

Downloads:
Test Results of the August 2026 OpenSource Models_v1.ods
Test Results of the August 2026 OpenSource Models_v1.xlsx