Instill Model (Deprecated)

March 18, 2024 ยท View on GitHub

Important

This repository has been deprecated and is only intended for launching Instill Core projects up to version v0.12.0-beta, where the Instill Model version corresponds to v0.9.0-alpha in this deprecated repository. Check the latest Instill Core project in the instill-ai/instill-core repository.

Instill Model (Deprecated)

GitHub release (latest SemVer including pre-releases) Artifact Hub Discord Integration Test

โš—๏ธ Instill Model, or simply Model, is an integral component of the Instill Core project. It serves as an advanced ModelOps/LLMOps platform focused on empowering users to seamlessly import, serve, fine-tune, and monitor Machine Learning (ML) models for continuous optimization.

Prerequisites

  • macOS or Linux - Instill Model works on macOS or Linux, but does not support Windows yet.

  • Docker and Docker Compose - Instill Model uses Docker Compose (specifically, Compose V2 and Compose specification) to run all services at local. Please install the latest stable Docker and Docker Compose before using Instill Model.

  • yq > v4.x. Please follow the installation guide.

  • (Optional) NVIDIA Container Toolkit - To enable GPU support in Instill Model, please refer to NVIDIA Cloud Native Documentation to install NVIDIA Container Toolkit. If you'd like to specifically allot GPUs to Instill Model, you can set the environment variable NVIDIA_VISIBLE_DEVICES. For example, NVIDIA_VISIBLE_DEVICES=0,1 will make the triton-server consume GPU device id 0 and 1 specifically. By default NVIDIA_VISIBLE_DEVICES is set to all to use all available GPUs on the machine.

Quick start

Note The image of model-backend (~2GB) and Triton Inference Server (~23GB) can take a while to pull, but this should be an one-time effort at the first setup.

Use stable release version

Execute the following commands to pull pre-built images with all the dependencies to launch:

$ git clone -b v0.10.0-alpha https://github.com/instill-ai/deprecated-model.git && cd deprecated-model

# Launch all services
$ make all

๐Ÿš€ That's it! Once all the services are up with health status, the UI is ready to go at http://localhost:3000. Please find the default login credentials in the documentation.

To shut down all running services:

$ make down

Explore the documentation to discover all available deployment options.

Officially supported models

We curate a list of ready-to-use models. These pre-trained models are from different sources and have been trained and deployed by our team. Want to contribute a new model? Please create an issue, we are happy to add it to the list ๐Ÿ‘.

ModelTaskSourcesFrameworkCPUGPU
MobileNet v2Image ClassificationGitHub-DVCONNXโœ…โœ…
Vision Transformer (ViT)Image ClassificationHugging FaceONNXโœ…โŒ
YOLOv4Object DetectionGitHub-DVCONNXโœ…โœ…
YOLOv7Object DetectionGitHub-DVCONNXโœ…โœ…
YOLOv7 W6 PoseKeypoint DetectionGitHub-DVCONNXโœ…โœ…
PSNet + EasyOCROptical Character Recognition (OCR)GitHub-DVCONNXโœ…โœ…
Mask RCNNInstance SegmentationGitHub-DVCPyTorchโœ…โœ…
Lite R-ASPP based on MobileNetV3Semantic SegmentationGitHub-DVCONNXโœ…โœ…
Stable DiffusionText to ImageGitHub-DVC, Local-CPU, Local-GPUONNXโœ…โœ…
Stable Diffusion XLText to ImageGitHub-DVCPyTorchโŒโœ…
Control Net - CannyImage to ImageGitHub-DVCPyTorchโŒโœ…
Megatron GPT2Text GenerationGitHub-DVCFasterTransformerโŒโœ…
Llama2Text GenerationGitHub-DVCvLLM, PyTorchโœ…โœ…
Code LlamaText GenerationGitHub-DVCvLLMโŒโœ…
Llama2 ChatText Generation ChatGitHub-DVCvLLMโŒโœ…
MosaicML MPTText Generation ChatGitHub-DVCvLLMโŒโœ…
MistralText Generation ChatGitHub-DVCvLLMโŒโœ…
Zephyr-7bText Generation ChatGitHub-DVCPyTorchโœ…โœ…
LlavaVisual Question AnsweringGitHub-DVCPyTorchโŒโœ…

Note: The GitHub-DVC source in the table means importing a model into Instill Model from a GitHub repository that uses DVC to manage large files.

License

See the LICENSE file for licensing information.