RWKV Runner

This project aims to eliminate the barriers of using large language models by automating everything for you. All you need is a lightweight executable program of just a few megabytes. Additionally, this project provides an interface compatible with the OpenAI API, which means that every ChatGPT client is an RWKV client.

English | 简体中文

Install

FAQs | Preview | Download | Server-Deploy-Examples

Default configs has enabled custom CUDA kernel acceleration, which is much faster and consumes much less VRAM. If you encounter possible compatibility issues, go to the Configs page and turn off `Use Custom CUDA kernel to Accelerate`.

If Windows Defender claims this is a virus, you can try downloading v1.0.8/v1.0.9 and letting it update automatically to the latest version, or add it to the trusted list.

For different tasks, adjusting API parameters can achieve better results. For example, for translation tasks, you can try setting Temperature to 1 and Top_P to 0.3.

Features

RWKV model management and one-click startup
Fully compatible with the OpenAI API, making every ChatGPT client an RWKV client. After starting the model, open http://127.0.0.1:8000/docs to view more details.
Automatic dependency installation, requiring only a lightweight executable program
Configs with 2G to 32G VRAM are included, works well on almost all computers
User-friendly chat and completion interaction interface included
Easy-to-understand and operate parameter configuration
Built-in model conversion tool
Built-in download management and remote model inspection
Multilingual localization
Theme switching
Automatic updates

API Concurrency Stress Testing

ab -p body.json -T application/json -c 20 -n 100 -l http://127.0.0.1:8000/chat/completions

body.json:

{
  "messages": [
    {
      "role": "user",
      "content": "Hello"
    }
  ]
}

594270461/RWKV-Runner

RWKV Runner

Install

Default configs has enabled custom CUDA kernel acceleration, which is much faster and consumes much less VRAM. If you encounter possible compatibility issues, go to the Configs page and turn off `Use Custom CUDA kernel to Accelerate`.

If Windows Defender claims this is a virus, you can try downloading v1.0.8/v1.0.9 and letting it update automatically to the latest version, or add it to the trusted list.

For different tasks, adjusting API parameters can achieve better results. For example, for translation tasks, you can try setting Temperature to 1 and Top_P to 0.3.

Features

API Concurrency Stress Testing

Todo

Related Repositories:

Preview

Homepage

Chat

Completion

Configuration

Model Management

Download Management

Settings

594270461/RWKV-Runner

RWKV Runner

Install

Default configs has enabled custom CUDA kernel acceleration, which is much faster and consumes much less VRAM. If you encounter possible compatibility issues, go to the Configs page and turn off Use Custom CUDA kernel to Accelerate.

If Windows Defender claims this is a virus, you can try downloading v1.0.8/v1.0.9 and letting it update automatically to the latest version, or add it to the trusted list.

For different tasks, adjusting API parameters can achieve better results. For example, for translation tasks, you can try setting Temperature to 1 and Top_P to 0.3.

Features

API Concurrency Stress Testing

Todo

Related Repositories:

Preview

Homepage

Chat

Completion

Configuration

Model Management

Download Management

Settings

Default configs has enabled custom CUDA kernel acceleration, which is much faster and consumes much less VRAM. If you encounter possible compatibility issues, go to the Configs page and turn off `Use Custom CUDA kernel to Accelerate`.