SAVRN
Search Contact SAVRN

Independent publisher

CaoHaoWei

CaoHaoWei

Models in Library4
Datasets in Library0
Models on Hugging Face5
Followers—

Models

Model · Text classification

Jev-LCT-Qwen3-8B

CaoHaoWei

Jev-LCT-Qwen3-8B is the flagship enterprise-grade decision engine of the Jev-LCT family. Combining Qwen3-8B's extensive foundational capabilities with Looped Calibration and adaptive early exit, it provides frontier generative reasoning capabilities with deterministic sub-100ms decision latency. - 85.0% 科学推理 + 70.0% MMLU:媲美中大型生成模型的复杂逻辑推理能力,但单次推断控制在 89.2 ms 内。 - 企业级智能体中枢:支持高风险场景的“选择性预测(Selective Prediction)”,在 80% 覆盖率下实现近乎零差错审核。 - 全量独立权重:开箱即用,支持多 GPU 分片或单张 24GB 显卡(RTX 3090 / 4090)bfloat16 全速推断。 Apache License 2.0. Full repository at GitHub.

Open weights apache-2.0 8.2B parameters 40,960 tokens transformers

Model · Text classification

Jev-LCT-Adapters

CaoHaoWei

This repository contains the standalone lightweight Looped Adapters and uncertainty calibration mappers for Jev-LCT (Looped Calibration Transformer). Instead of downloading the entire merged model weights (which range from 2GB to 16GB), users who already have base Qwen models (Qwen2.5-0.5B, Qwen2.5-1.5B, Qwen3-8B) can directly mount these compact adapter weights (only 220MB ~ 1.5GB) onto the last $k=2$ layers of the base model. 本仓库托管 Jev-LCT (循环校准 Transformer) 的全套独立轻量级适配器权重(.pt)与保序校准映射器(.pkl)。 对于本地已有 Qwen 官方基座(如 Qwen/Qwen2.5-0.5B、Qwen/Qwen2.5-1.5B、Qwen3-8B)的开发者,无需重复下载数十 GB 的完整权重,仅需下载本仓库对应的轻量适配器(仅 220MB ~ 1.5GB),即可在原生基座上获得 50ms 级别系统一极速决策与内生轨迹校准置信度能力。 3. 即插即用:与 lctqwenstandalone.py…

Open weights apache-2.0 transformers

Model · Text classification

Jev-LCT-Qwen2.5-1.5B

CaoHaoWei

In agentic workflows, API routing, and edge decision-making, conventional autoregressive LLMs suffer from high token-by-token generation latency and brittle string parsing, while small discriminative models produce systematically miscalibrated verbal confidence (overconfident hallucinations). Jev-LCT (Looped Calibration Transformer) establishes a new paradigm for System-One Decision Models: 1. Parallel Looped Prefill: Recurrently iterates only the top $k=2$ layers of Qwen2.5-1.5B with sequence right-shifting and Scale-Preserving RMS Injection, strictly preventing representation collapse. 2. Endogenous Trajectory Confidence: Extracts genuine calibrated confidence directly from hidden state…

Open weights apache-2.0 1.5B parameters 131,072 tokens transformers

Model · Text classification

Jev-LCT-Qwen2.5-0.5B

CaoHaoWei

Jev-LCT-Qwen2.5-0.5B is the ultra-lightweight edge edition of the Jev-LCT System-One decision family. Weighing only ~1.9 GB in full bfloat16 weights, it is optimized for high-throughput API gateway routing, edge robotics, Raspberry Pi, and mobile deployment. - ~50ms 极低延迟:专为高并发 API 网关路由、实时内容审核设计; - MMLU 达 50.0%:大幅超越参数相近的判别模型(Laya 33.3%, Open-Jev 35.0%); Apache License 2.0. Full repository at GitHub.

Open weights apache-2.0 494M parameters 32,768 tokens transformers