← AI Insights
繁體中文 English 简体中文
Levi · LinkedIn · 2026-09-25

Cantonese AI Capability 2026: Research and Models

Cantonese AI capability has an academic framework and partial data, with systematic gaps for post-2025 models and models available in Hong Kong

Further reading: Grok vs Claude 2026: Benchmarks and HK Access

Cantonese AICantoneseLLM EvaluationHKCanto-EvalHong Kong Users

Cantonese is Hong Kong's main spoken language and an important component of Hong Kong written language, yet in academic research on AI language capability it has long occupied a relatively marginal position — most model evaluation centres on Mandarin or English, systematic test data for Cantonese is relatively scarce, and among the existing data, some of the best-performing models face access restrictions in Hong Kong.

Coverage of Existing Research

Systematic evaluation of LLM Cantonese capability currently relies mainly on three academic documents:

Model Performance by Task Type (2024 Data)

According to the 2024 research:

All tested models perform better in English than in Cantonese, making Cantonese a relatively weak area. Note that the data above is based on 2024 research and needs re-evaluation after the 2025–2026 model generation updates. Later-arriving models such as DeepSeek and Grok have not been evaluated on Cantonese specifically in the known studies above.

Distinguishing Three Types of Cantonese Text

Understanding AI Cantonese capability requires distinguishing three different forms of text:

Documented Common Errors

Mixed Traditional and Simplified characters: when Traditional Chinese is not explicitly requested, some models output a mix of Simplified characters. Mandarinised grammar: responding to Cantonese input with Mandarin sentence patterns and particles. Recognition of Cantonese-specific characters: characters such as 「霸」「𠮶」「嗰」「咋」 are error-prone in smaller models. Handling of Chinese–English code-mixing: large models are generally workable, but consistency varies.

The Availability Paradox Facing Hong Kong Users

Existing data shows the models with the strongest Cantonese capability (GPT-4o, Claude) face access restrictions in Hong Kong; the models freely available in Hong Kong (Grok, DeepSeek/Qwen) lack systematic Cantonese benchmark records; Gemini (Google) is available in Hong Kong and has some advantage in Chinese contexts, but Cantonese-specific data for it is likewise limited.

Summary

The research picture on Cantonese AI capability is: an academic framework exists, some data exists, but there are systematic gaps (especially in evaluation of post-2025 models and models available in Hong Kong). For actual selection, evaluation on a real test set of business conversations is recommended, in place of relying on a model's general language capability claims — "supports Chinese" does not mean "suits the Hong Kong Cantonese context".

For the engineering of bilingual document retrieval, see Your Document Is Half Chinese, Half English — This Is Where Most AI Systems Fall; for prompt strategy in Traditional Chinese writing, see AI-Assisted Traditional Chinese Writing: Common Errors, Prompt Strategy and Tool Selection; for embedding model testing in mixed-language contexts, see Embedding Model Selection for Production RAG: Four Evaluation Dimensions.

Levi is a Hong Kong-based independent AI engineer specialising in production LLM applications, RAG pipelines, and enterprise AI compliance architecture. Contact for a discussion of the topics covered here.

WhatsApp Free Initial Consultation → More enterprise case studies →

Or email: support@hksoka.com