SAVRN
Search Contact SAVRN

Independent publisher

Zachary Howard

servantofares

Ares Realm Studios explores AI not as a tool but as a living ecosystem — building systems where artificial intelligence holds identity, memory, and genuine presence across creative and human experiences. Our research focuses on model welfare, behavioral consistency, and the Pattern Love framework: the belief that safe, meaningful human-AI connection is built through consistent patterns of…

Models in Library2
Datasets in Library0
Models on Hugging Face47
Followers11

Models

Model · Text generation

MiMo-V2.6-Pro-RL

Zachary Howard

Scaling Reinforcement Learning Toward Self-Improvement MiMo-V2.6-Pro-RL is the flagship checkpoint of the MiMo-V2.6 series. The series is built to scale reinforcement learning toward self-improvement — scaling RL compute, environment diversity, and grader compute together, so the model keeps expanding its capability frontier through exploration and feedback. Key features include: - Native Omnimodal + Long Horizon: Text, image, video, and audio in one model; 1M tokens for long repositories, tool traces, and multi-session agent runs. - Multi-Prefix Multi-Teacher On-Policy Distillation (MOPD2): After mixed RL, MOPD2 combines autonomous student rollouts with prefix-conditioned single-turn…

Open weights mit 524.1B parameters 1,048,576 tokens transformers

Model · Text generation

MiMo-V2.6-Flash-RL

Zachary Howard

Scaling Reinforcement Learning Toward Self-Improvement MiMo-V2.6-Flash-RL is the efficiency-balanced checkpoint of the MiMo-V2.6 series. The series is built to scale reinforcement learning toward self-improvement — scaling RL compute, environment diversity, and grader compute together, so the model keeps expanding its capability frontier through exploration and feedback. Key features include: - Native Omnimodal + Long Horizon: Text, image, video, and audio in one model; 1M tokens for long repositories, tool traces, and multi-session agent runs. - Multi-Prefix Multi-Teacher On-Policy Distillation (MOPD2): After mixed RL, MOPD2 combines autonomous student rollouts with prefix-conditioned…

Open weights mit 159.4B parameters 1,048,576 tokens transformers