liuyhwangyh's repositories
facechain
FaceChain is a deep-learning toolchain for generating your Digital-Twin.
FastChat
An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.
hello
nothing
llama_index
LlamaIndex (formerly GPT Index) is a data framework for your LLM applications
llmuses
A streamlined and customizable framework for efficient large model evaluation and performance benchmarking
modelscope
ModelScope: bring the notion of Model-as-a-Service to life.
TensorRT-LLM
TensorRT-LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and build TensorRT engines that contain state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT-LLM also contains components to create Python and C++ runtimes that execute those TensorRT engines.
vllm
A high-throughput and memory-efficient inference and serving engine for LLMs