jagilley / controlnet-seg

Modify images using semantic segmentation

replicate.com
Total runs: 166.3K
24-hour runs: 0
7-day runs: 0
30-day runs: 0
Github
Model's Last Updated: February 15 2023

Introduction of controlnet-seg

Model Details of controlnet-seg

Readme

Model by Lyumin Zhang

Usage

Input an image, and prompt the model to generate an image as you would for Stable Diffusion. Then. a model called Uniformer will detect the segmentations for you to control your output image.

Model Description

This model is ControlNet adapting Stable Diffusion to use a semantic segmentation of an input image in addition to a text input to generate an output image. The segmentation model will first segment the input image into different semantic regions, and then use those regions as conditioning input when generating a new image. This model was trained with the ADE20K dataset captioned by BLIP to obtain 164K segmentation-image-caption pairs. The model is trained with 200 GPU-hours on Nvidia A100 80G. The base model is Stable Diffusion 1.5.

ControlNet is a neural network structure which allows control of pretrained large diffusion models to support additional input conditions beyond prompts. The ControlNet learns task-specific conditions in an end-to-end way, and the learning is robust even when the training dataset is small (< 50k samples). Moreover, training a ControlNet is as fast as fine-tuning a diffusion model, and the model can be trained on a personal device. Alternatively, if powerful computation clusters are available, the model can scale to large amounts of training data (millions to billions of rows). Large diffusion models like Stable Diffusion can be augmented with ControlNets to enable conditional inputs like edge maps, segmentation maps, keypoints, etc.

Original model & code on GitHub

Other ControlNets

There are many different ways to use a ControlNet to modify the output of Stable Diffusion. Here are a few different options, all of which use an input image in addition to a prompt to generate an output. The methods process the input in different ways; try them out to see which works best for a given application.

ControlNet for generating images from drawings Scribble: https://replicate.com/jagilley/controlnet-scribble

ControlNets for generating humans based on input image Human Pose Detection: https://replicate.com/jagilley/controlnet-pose

ControlNets for preserving general qualities about an input image Edge detection: https://replicate.com/jagilley/controlnet-canny HED maps: https://replicate.com/jagilley/controlnet-hed Depth map: https://replicate.com/jagilley/controlnet-depth2img Hough line detection: https://replicate.com/jagilley/controlnet-hough Normal map: https://replicate.com/jagilley/controlnet-normal

Citation
@misc{https://doi.org/10.48550/arxiv.2302.05543,
  doi = {10.48550/ARXIV.2302.05543},
  url = {https://arxiv.org/abs/2302.05543},
  author = {Zhang, Lvmin and Agrawala, Maneesh},
  keywords = {Computer Vision and Pattern Recognition (cs.CV), Artificial Intelligence (cs.AI), Graphics (cs.GR), Human-Computer Interaction (cs.HC), Multimedia (cs.MM), FOS: Computer and information sciences, FOS: Computer and information sciences},
  title = {Adding Conditional Control to Text-to-Image Diffusion Models},
  publisher = {arXiv},
  year = {2023},
  copyright = {arXiv.org perpetual, non-exclusive license}
}

Pricing of controlnet-seg replicate.com

Run time and cost

This model costs approximately $0.0085 to run on Replicate, or 117 runs per $1, but this varies depending on your inputs. It is also open source and you can run it on your own computer with Docker .

This model runs on Nvidia A100 (80GB) GPU hardware . Predictions typically complete within 7 seconds.

Runs of jagilley controlnet-seg on replicate.com

166.3K
Total runs
0
24-hour runs
0
3-day runs
0
7-day runs
0
30-day runs

More Information About controlnet-seg replicate.com Model

controlnet-seg replicate.com

controlnet-seg replicate.com is an AI model on replicate.com that provides controlnet-seg's model effect (Modify images using semantic segmentation), which can be used instantly with this jagilley controlnet-seg model. replicate.com supports a free trial of the controlnet-seg model, and also provides paid use of the controlnet-seg. Support call controlnet-seg model through api, including Node.js, Python, http.

controlnet-seg replicate.com Url

https://replicate.com/jagilley/controlnet-seg

jagilley controlnet-seg online free

controlnet-seg replicate.com is an online trial and call api platform, which integrates controlnet-seg's modeling effects, including api services, and provides a free online trial of controlnet-seg, you can try controlnet-seg online for free by clicking the link below.

jagilley controlnet-seg online free url in replicate.com:

https://replicate.com/jagilley/controlnet-seg

controlnet-seg install

controlnet-seg is an open source model from GitHub that offers a free installation service, and any user can find controlnet-seg on GitHub to install. At the same time, replicate.com provides the effect of controlnet-seg install, users can directly use controlnet-seg installed effect in replicate.com for debugging and trial. It also supports api for free installation.

controlnet-seg install url in replicate.com:

https://replicate.com/jagilley/controlnet-seg

controlnet-seg install url in github:

https://github.com/replicate/controlnet

Url of controlnet-seg

controlnet-seg replicate.com Url

controlnet-seg Owner Github

Provider of controlnet-seg replicate.com

Other API from jagilley

replicate

Generate detailed images from scribbled drawings

Total runs: 38.2M
Run Growth: 0
Growth Rate: 0.00%
Updated:February 14 2023
replicate

Modify images using M-LSD line detection

Total runs: 9.5M
Run Growth: 0
Growth Rate: 0.00%
Updated:February 15 2023
replicate

Modify images using canny edge detection

Total runs: 824.7K
Run Growth: 0
Growth Rate: 0.00%
Updated:February 24 2023
replicate

Modify images using HED maps

Total runs: 515.2K
Run Growth: 0
Growth Rate: 0.00%
Updated:February 14 2023
replicate

Modify images using normal maps

Total runs: 330.5K
Run Growth: 0
Growth Rate: 0.00%
Updated:February 17 2023
replicate

Modify images with humans using pose detection

Total runs: 174.4K
Run Growth: 0
Growth Rate: 0.00%
Updated:February 24 2023
replicate

Change voice for spoken text

Total runs: 71.6K
Run Growth: 0
Growth Rate: 0.00%
Updated:February 28 2023
replicate

Modify images with a prompt while preserving their structure

Total runs: 61.9K
Run Growth: 0
Growth Rate: 0.00%
Updated:February 24 2023
replicate

Create variations of an image while preserving shape and depth

Total runs: 57.2K
Run Growth: 0
Growth Rate: 0.00%
Updated:January 27 2023