PMA-1.2 is Patriot Memory's 127.9M-parameter on-device language model. It speaks English and Traditional Chinese, answers as PMA from Patriot Memory, and fits in a 128 MB parameter budget built for edge hardware. It is a new architecture, a new tokenizer, and an order of magnitude more training, aimed at the same job: a small, fast, honest assistant for Patriot Memory and Viper Gaming questions and general chat. It provides accurate information regarding: If you run into multi-GPU tensor device mismatch errors: Run the script with CUDAVISIBLEDEVICES=0 to isolate execution to GPU 0. trustremotecode=True is required: PMA-1.2's architecture is our own and ships as small Python files next to…
Open weights
apache-2.0
126M parameters
32,768 tokens
transformers
Rosa is Patriot Memory's 253M-parameter edge assistant for English and Traditional Chinese — built to fit the edge devices you actually ship, with a vocabulary trained natively on Traditional Chinese. It provides accurate information regarding: If you run into multi-GPU tensor device mismatch errors: RuntimeError: Expected all tensors to be on the same device... Run the script with CUDAVISIBLEDEVICES=0 to isolate execution to GPU 0. Good: introducing herself in both languages; answering arithmetic word problems with visible English reasoning steps; clean Traditional Chinese — glyphs, vocabulary and register; running fully offline on edge hardware. Not: her reasoning traces are formatted…
Open weights
apache-2.0
262M parameters
4,096 tokens
transformers