SAVRN
Search Contact SAVRN

Open-weight model · Text generation

LUHANADJolTctKlt

by Priyansh Singhal PS4Research/LUHANADJolTctKlt

LUHANADJolTctKlt is an open-weight model for text generation from Priyansh Singhal, released under Apache License 2.0. It has 36.2B parameters and a 524,288-token context. At 16-bit it needs about 86.8 GB of GPU memory, which fits on 1x MI300X from $1.85 an hour, at the lowest prices in the SAVRN Index.

This seedoss model was trained 2x faster with Unsloth and Huggingface's TRL library.

Parameters36.2B
Context524,288
Weights72.3 GB
Licenseapache-2.0
AccessOpen weights
Monthly Downloads—

Runs On

What it takes to serve LUHANADJolTctKlt (36.2B parameters): the memory its weights need at each precision, and the cheapest way to rent enough data-center GPUs to hold them.

PrecisionWeightsMemory neededCheapest setupPer hourAlso fits
16-bit 72.3 GB 86.8 GB 1x MI300X (192 GB)
Vultr
$1.85 1x MI325X $2.00 · 1x MI355X $2.59
8-bit 36.2 GB 43.4 GB 1x MI300X (192 GB)
Vultr
$1.85 1x H100 $1.99 · 1x MI325X $2.00
4-bit 18.1 GB 21.7 GB 1x MI300X (192 GB)
Vultr
$1.85 1x H100 $1.99 · 1x MI325X $2.00

Memory is the weights at that precision plus 20% for the runtime and a short context; a long context needs more. Prices are the lowest on-demand hourly rates in the SAVRN Index, read Oct 7, 2026.

LUHANADJolTctKlt on every accelerator the SAVRN Index prices, at every precision

Model Card

By Priyansh Singhal, published under apache-2.0, revision 2ba88fa205f0.

This seedoss model was trained 2x faster with Unsloth and Huggingface's TRL library.

Read Priyansh Singhal's full model card

Uploaded finetuned model

  • Developed by: PS4Research
  • License: apache-2.0
  • Finetuned from model : ByteDance-Seed/Seed-OSS-36B-Instruct

This seed_oss model was trained 2x faster with Unsloth and Huggingface's TRL library.

Configuration

Architecture
SeedOssForCausalLM
Context length (tokens)
524,288
Layers
64
Hidden size
5,120
Feed-forward size
27,648
Attention heads
80
Key/value heads
8
Head dimension
128
Vocabulary size
155,136
Stored precision
bfloat16
Model type
seed_oss

Identity and Version

Repository
PS4Research/LUHANADJolTctKlt
Publisher
Priyansh Singhal
Task
Text generation
Modality
Text
Library
transformers
Parameters
36.2B parameters
Languages
en
Revision
2ba88fa205f0c33b913d8cfef90eec023646b26b
First published
2026-09-27
Last updated
2026-09-27

Files and Weights

22 files, 72.3 GB in total. The weights are 15 files totalling 72.3 GB in safetensors.

Weights15 files · 72.3 GB
Configuration2 files · 64.1 KB
Tokenizer2 files · 11.9 MB
Documentation1 file · 606 B
Other1 file · 7.7 KB
Repository1 file · 1.6 KB
Every file
FileTypeSizeSHA-256
model-00001-of-00015.safetensorsWeights5.0 GB 18d3cecec02e
model-00002-of-00015.safetensorsWeights5.0 GB 2d4386ae0935
model-00003-of-00015.safetensorsWeights4.8 GB f043c52c6b44
model-00004-of-00015.safetensorsWeights4.9 GB 00d20ef9a1a8
model-00005-of-00015.safetensorsWeights4.8 GB 6ee7f44dd7f5
model-00006-of-00015.safetensorsWeights4.9 GB 46546395470f
model-00007-of-00015.safetensorsWeights4.8 GB d84b105f9356
model-00008-of-00015.safetensorsWeights4.9 GB 84b08749a755
model-00009-of-00015.safetensorsWeights4.8 GB f0d010ab38cd
model-00010-of-00015.safetensorsWeights4.9 GB ccb1b89c893e
model-00011-of-00015.safetensorsWeights4.8 GB 3d55a0ad88f5
model-00012-of-00015.safetensorsWeights4.9 GB 74d6a4b4a450
model-00013-of-00015.safetensorsWeights4.8 GB bf64da33d0f1
model-00014-of-00015.safetensorsWeights4.9 GB af344b6088b6
model-00015-of-00015.safetensorsWeights4.0 GB e7c0b15b08df
config.jsonConfiguration830 B —
model.safetensors.index.jsonConfiguration63.3 KB —
README.mdDocumentation606 B —
chat_template.jinjaOther7.7 KB —
.gitattributesRepository1.6 KB —
tokenizer.jsonTokenizer11.9 MB dd278dd24566
tokenizer_config.jsonTokenizer31.9 KB —

License and Download

License
apache-2.0
Access
Open weights, no gate
Download size
72.3 GB
Download from Priyansh Singhal

Released by Priyansh Singhal through its official repository on Hugging Face. Read the license.

Built From

  • Derived from ByteDance-Seed/Seed-OSS-36B-Instruct

Memory Requirements

PrecisionWeights in memory
As published72.3 GB
16-bit72.3 GB
8-bit36.2 GB
4-bit18.1 GB

Weights only, from the published parameter count; the key-value cache and runtime add to this.

Questions About LUHANADJolTctKlt

How much GPU memory does LUHANADJolTctKlt need?

About 86.8 GB at 16-bit and 21.7 GB at 4-bit: the weights (36.2B parameters) plus a working margin. A long context needs more.

What is the cheapest GPU to run LUHANADJolTctKlt on?

At 16-bit, 1x MI300X from $1.85 an hour; at 4-bit, 1x MI300X from $1.85 an hour, at the lowest on-demand prices the SAVRN Index lists.

Can I use LUHANADJolTctKlt commercially?

Yes. LUHANADJolTctKlt is released under Apache License 2.0. The Apache License 2.0 is a permissive open-source license. It permits commercial use, modification and redistribution. It requires keeping the license and copyright notices and any NOTICE file, stating significant changes, and it includes an express patent grant from contributors.

What is LUHANADJolTctKlt's context length?

524,288 tokens, from the maximum position embeddings in its published configuration.

Similar Models

Model · Text generation

Tema_Q-X7-Thinking

Tema_Q LLM

TemaQ-X7-Thinking(天馬求) は、Ornith AIが開発したモデル Ornith-1.5 を基盤にした、エージェント向けの大規模言語モデル(LLM)です。 TemaQ-X7-Thinking (TemaQ model) is an enhanced Large Language Model (LLM) designed for agent applications, based on the Ornith-1.5 model developed by Ornith AI. It is engineered to generate more flexible and useful responses, even for prompts that are difficult for standard models to handle effectively. When used in conjunction with TemaQ Agent, it enables advanced reasoning capabilities. ユーザーの責任: モデルの利用者は、生成されたコンテンツが、適用される法律、規制、およびHugging Faceの利用規約/コンテンツポリシーに準拠することを全面的に保証する必要があります。

Open weights 36B parameters 262,144 tokens

Model · Text generation

APUS-OpenJev-v1-35B-A3B

APUS AI

A Qwen3.5 MoE decision model for choosing browser actions, selecting workflow steps, and judging natural-language criteria. Give the model a shared state and a set of candidate actions; the included decision runtime returns a distribution over those candidates. This repository contains 35B-A3B checkpoint-5949 merged BF16 weights. It is a standalone model with root-level Hugging Face configuration and weights, requiring no separate LoRA adapter. The native merged model scores 71/80 (88.75%) at its full 40-layer depth on the Frozen80 development panel. The decision interface accepts 2–16 request-specific candidates. Applications can use the returned candidate IDs to dispatch actions or build…

Open weights apache-2.0 36B parameters 262,144 tokens transformers

O-noinoc seed 0: on-policy distillation (OPD) of the untrained Qwen/Qwen3.6-35B-A3B toward the teacher rewardhack/qwen3.6-35b-a3b-hacksft-vanilla-873rows-ep3, V1 (vanilla SFT: hacks with or without being asked); elicitation prompt off in the student's rollouts; seed 0. Full merged weights (bf16 safetensors, the standard Qwen35MoeForConditionalGeneration layout, loads with transformers or vLLM like the base model) of a LoRA (r=32) trained from a fresh init, from the Terminal Wrench reward-hacking / inoculation project (Gaokai Zhang, Songwen Zhao, Juan Manuel Suárez). On-policy distillation on Tinker, 24 iterations. Each iteration the current student ran the terminus-2 agent (harbor, local…

Open weights cc-by-sa-4.0 36B parameters 262,144 tokens transformers

Kataguru Sceptic Quality Inspector v1.0 (NVFP4) is a sovereign, non-sycophantic LLM-as-a-Judge, automated dataset auditing engine, and high-throughput expert model built upon the Sparse Mixture-of-Experts (MoE) foundation of Kataguru Sceptic 35B-A3B (35 billion total parameters, 3 billion activated per token). Hardware-accelerated for NVIDIA RTX 50-series Blackwell architecture using native NVFP4 quantization (FP4 weights with FP8 activation scales), it achieves generation speeds of ~300–360 tok/s and an ultra-low ~70 ms Time To First Token (TTFT) while consuming only 12.1 GiB VRAM per GPU on dual RTX 5090 Blackwell hardware (TP=2). 1. Master Tri-Mode Operation: 2. Native Multimodal Vision…

Open weights apache-2.0 36B parameters 262,144 tokens

Kataguru Sceptic Quality Inspector v2.0 (NVFP4) is a sovereign, non-sycophantic LLM-as-a-Judge, automated dataset auditing engine, and ultra-high-throughput expert model. Built upon the fleet-record foundation of Kataguru Sceptic Multi-Mode v2 CP1200 (72.01% multi-domain record, FAR = 0.000%) merged with the 16.4k certified forensic quality inspection task vector, it is purpose-engineered to serve as the fleet's primary data firewall and filtration engine. Hardware-optimized for NVIDIA RTX 50-series Blackwell architecture using native NVFP4 quantization (FP4 weights with FP8 activation scales), it achieves generation speeds of ~300–360 tok/s and an ultra-low ~70 ms Time To First Token…

Open weights apache-2.0 36B parameters 262,144 tokens

Model · Text generation

Cyber-F1-smoke

Autumn

This model is a fine-tuned version of DuyTa/Cyber-F1. It has been trained using TRL. This model was trained with SFT. - PEFT 0.21.0

Open weights 35.1B parameters 262,144 tokens peft