liuyhwangyh

Pinned Repositories

copyofiosched
00
dash-cookbook
Receipts for creating AI Applications with APIs from DashScope (and friends)!
00
eval-scope
A streamlined and customizable framework for efficient large model evaluation and performance benchmarking
Language:Python00
facechain
FaceChain is a deep-learning toolchain for generating your Digital-Twin.
Language:Python00
FastChat
An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.
Language:Python00
hello
nothing
00
llama_index
LlamaIndex (formerly GPT Index) is a data framework for your LLM applications
Language:Python00
modelscope
ModelScope: bring the notion of Model-as-a-Service to life.
Language:Python00
TensorRT-LLM
TensorRT-LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and build TensorRT engines that contain state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT-LLM also contains components to create Python and C++ runtimes that execute those TensorRT engines.
Language:C++00
vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
Language:Python0 0 00

liuyhwangyh's Repositories

liuyhwangyh/copyofiosched
00
liuyhwangyh/dash-cookbook
Receipts for creating AI Applications with APIs from DashScope (and friends)!
00
liuyhwangyh/eval-scope
A streamlined and customizable framework for efficient large model evaluation and performance benchmarking
Language:Python00
liuyhwangyh/facechain
FaceChain is a deep-learning toolchain for generating your Digital-Twin.
Language:Python00
liuyhwangyh/FastChat
An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.
Language:Python00
liuyhwangyh/hello
nothing
00
liuyhwangyh/llama_index
LlamaIndex (formerly GPT Index) is a data framework for your LLM applications
Language:Python00
liuyhwangyh/modelscope
ModelScope: bring the notion of Model-as-a-Service to life.
Language:Python00
liuyhwangyh/TensorRT-LLM
TensorRT-LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and build TensorRT engines that contain state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT-LLM also contains components to create Python and C++ runtimes that execute those TensorRT engines.
Language:C++00
liuyhwangyh/vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
Language:Python0 0 00