Stemma Machinarum

Model · openchat-3-5

OpenChat 3.5

Developer: OpenChat team (paper: Guan Wang, Sijie Cheng, Xianyuan Zhan, Xiangang Li, Sen Song, Yang Liu; Tsinghua University, Shanghai AI Laboratory, 01.AI); HF uploader imone
Availability: available · checked 2026-09-24 · source

Raw record: /data/models/openchat-3-5.json

Fields

id
openchat-3-5
identifiers
huggingface
openchat/openchat_3.5
developer
OpenChat team (paper: Guan Wang, Sijie Cheng, Xianyuan Zhan, Xiangang Li, Sen Song, Yang Liu; Tsinghua University, Shanghai AI Laboratory, 01.AI); HF uploader imone
release_date
2023-11-01recorded · source
note:  OpenChat GitHub README changelog: '[2023/11/01] We released the OpenChat-3.5-7B model'. HF repo created 2023-10-30.
weights_status
open
availability
available · checked 2026-09-24 · source
license
Apache-2.0recorded · source
note:  Card text: 'Our OpenChat 3.5 code and models are distributed under the Apache License 2.0.' (also card metadata). License file text not separately read.
architecture
family
decoder_only
note
family read from config 'architectures': ['MistralForCausalLM'].
n_layers
32recorded · source
hidden_size
4096recorded · source
n_heads
32recorded · source
vocab_size
32002recorded · source
note:  32,002 vs 32,000 for Mistral-7B-v0.1: the config's `_name_or_path` is imone/Mistral_7B_with_EOT_token, a Mistral 7B variant with an added end-of-turn token.
context_length
8192recorded · source
n_kv_heads
8recorded · source
positional_encoding
rotary (RoPE)recorded · source
note:  Architecture unchanged from mistral-7b-v0-1 (fine-tune, not a structural change).
training_data
A collection of publicly available instruction data with a custom processing pipeline, trained with C-RLFT (class-conditioned data sources, no preference labels). Notable subsets named on the card: OpenChat ShareGPT, OpenOrca with FLAN answers, Capybara (Pure-Dove, Verified-Camel, LessWrong-Amplify-Instruct), GOAT, Glaive, MetaMathQA, MathInstruct, OpenAssistant top-1.partial · source
note:  Card 'Dataset Details': 'trained with C-RLFT on a collection of publicly available high-quality instruction data ... notable subsets included here'. The list is not stated to be complete.
techniques
primary_sources
https://huggingface.co/openchat/openchat_3.5
https://huggingface.co/openchat/openchat_3.5/raw/0fc98e324280bc4bf5d2c30ecf7b97b84fb8a19b/config.json
record_history
date:2026-09-24 · change:ingested as candidate from HF (openchat/openchat_3.5@0fc98e324280bc4bf5d2c30ecf7b97b84fb8a19b) · by:ingest_hf.py ·
date:2026-09-24 · change:preparer: developer, release date (README), license, training data (partial), positional encoding (propagated), 5 edge entries (fine_tuned_from mistral, trained_on openorca; rest rejected); sources: card, config.json, arXiv 2309.11235, OpenChat README · by:claude (preparer, Sonnet 5) ·
date:2026-09-24 · change:reviewed and promoted from staging (2 edge(s) accepted) · by:Wilson Pruitt ·

Parents

Weights descend

Training data

Children

Weights descend

Read in

No station on the reading path has touched this record yet.