Text Generation
Transformers
Safetensors
GGUF
English
mistral
text-generation-inference
unsloth
trl
sft
theprint
conversational
Eval Results (legacy)
How to use from
SGLang
Install from pip and serve model
# Install SGLang from pip:
pip install sglang
# Start the SGLang server:
python3 -m sglang.launch_server \
    --model-path "theprint/ReWiz-7B" \
    --host 0.0.0.0 \
    --port 30000
# Call the server using curl (OpenAI-compatible API):
curl -X POST "http://localhost:30000/v1/chat/completions" \
	-H "Content-Type: application/json" \
	--data '{
		"model": "theprint/ReWiz-7B",
		"messages": [
			{
				"role": "user",
				"content": "What is the capital of France?"
			}
		]
	}'
Use Docker images
docker run --gpus all \
    --shm-size 32g \
    -p 30000:30000 \
    -v ~/.cache/huggingface:/root/.cache/huggingface \
    --env "HF_TOKEN=<secret>" \
    --ipc=host \
    lmsysorg/sglang:latest \
    python3 -m sglang.launch_server \
        --model-path "theprint/ReWiz-7B" \
        --host 0.0.0.0 \
        --port 30000
# Call the server using curl (OpenAI-compatible API):
curl -X POST "http://localhost:30000/v1/chat/completions" \
	-H "Content-Type: application/json" \
	--data '{
		"model": "theprint/ReWiz-7B",
		"messages": [
			{
				"role": "user",
				"content": "What is the capital of France?"
			}
		]
	}'
Quick Links

ReWiz-7B

This is a fine tune of Mistral 7B Instruct (0.3). Half the data was geared towards better reasoning (EvolKit-20k and reasoning-base-20k), the other half will help to de-censor the model (WizardLM data set).

Uploaded model

  • Developed by: theprint
  • License: apache-2.0
  • Finetuned from model : unsloth/mistral-7b-instruct-v0.3-bnb-4bit

This mistral model was trained 2x faster with Unsloth and Huggingface's TRL library.

Open LLM Leaderboard Evaluation Results

Detailed results can be found here

Metric Value
Avg. 17.54
IFEval (0-Shot) 40.48
BBH (3-Shot) 23.50
MATH Lvl 5 (4-Shot) 2.57
GPQA (0-shot) 3.36
MuSR (0-shot) 16.74
MMLU-PRO (5-shot) 18.56
Downloads last month
304
Safetensors
Model size
7B params
Tensor type
BF16
·
Inference Providers NEW
Input a message to start chatting with theprint/ReWiz-7B.

Model tree for theprint/ReWiz-7B

Quantized
(192)
this model
Merges
1 model
Quantizations
2 models

Datasets used to train theprint/ReWiz-7B

Collection including theprint/ReWiz-7B

Evaluation results