Team Phoenix is the LLM AI R&D team at HTX, the Home Team Science and Technology Agency, Singapore. We build Phoenix, Singapore's sovereign large language model family for the Home Team and the wider public service.
This organisation is where we publish model cards, technical reports and selected datasets. Most of our work runs inside government, so what appears here is a subset of what we build.
High-impact sovereign and enterprise deployments require understanding of specific local legislation, policy, institutional terminology and operational context. Phoenix holds that knowledge in its parameters rather than retrieving it by web search, which is what makes secure, air-gapped deployment possible. Localised model efforts so far have concentrated on smaller scales and text-only modalities; Phoenix tests whether a frontier multimodal model can be deeply adapted to local context without giving up broad general capability.
Building Phoenix also develops sustained local expertise in multilingual and multimodal data curation, continued pre-training, and safety.
Our domain corpus spans Singapore law, government, history and culture across English, Chinese, Malay and Tamil, with annotation done by native speakers.
We run open source benchmarks and build our own on top of them. Four so far:
Alongside the product lines, we work on reinforcement learning for knowledge recall, multilingual transfer, multimodal continued pre-training, and high-quality human annotation.