lucataco / zeta-editing

Zero-Shot Text-Based Audio Editing Using DDPM Inversion

replicate.com
Total runs: 1.7K
24-hour runs: 0
7-day runs: 0
30-day runs: 0
Github
Model's Last Updated: March 05 2024

Introduction of zeta-editing

Model Details of zeta-editing

Readme
Zero-Shot Unsupervised and Text-Based Audio Editing Using DDPM Inversion
Technion - Israel Institute of Technology

img

Abstract

Editing signals using large pre-trained models, in a zero-shot manner, has recently seen rapid advancements in the image domain. However, this wave has yet to reach the audio domain. In this paper, we explore two zero-shot editing techniques for audio signals, which use DDPM inversion on pre-trained diffusion models. The first, adopted from the image domain, allows text-based editing. The second, is a novel approach for discovering semantically meaningful editing directions without supervision. When applied to music signals, this method exposes a range of musically interesting modifications, from controlling the participation of specific instruments to improvisations on the melody.

Note: For now use input audio wav files

Pricing of zeta-editing replicate.com

Run time and cost

This model costs approximately $0.078 to run on Replicate, or 12 runs per $1, but this varies depending on your inputs. It is also open source and you can run it on your own computer with Docker .

This model runs on Nvidia A40 (Large) GPU hardware . Predictions typically complete within 108 seconds. The predict time for this model varies significantly based on the inputs.

Runs of lucataco zeta-editing on replicate.com

1.7K
Total runs
0
24-hour runs
0
3-day runs
0
7-day runs
0
30-day runs

More Information About zeta-editing replicate.com Model

zeta-editing replicate.com

zeta-editing replicate.com is an AI model on replicate.com that provides zeta-editing's model effect (Zero-Shot Text-Based Audio Editing Using DDPM Inversion), which can be used instantly with this lucataco zeta-editing model. replicate.com supports a free trial of the zeta-editing model, and also provides paid use of the zeta-editing. Support call zeta-editing model through api, including Node.js, Python, http.

zeta-editing replicate.com Url

https://replicate.com/lucataco/zeta-editing

lucataco zeta-editing online free

zeta-editing replicate.com is an online trial and call api platform, which integrates zeta-editing's modeling effects, including api services, and provides a free online trial of zeta-editing, you can try zeta-editing online for free by clicking the link below.

lucataco zeta-editing online free url in replicate.com:

https://replicate.com/lucataco/zeta-editing

zeta-editing install

zeta-editing is an open source model from GitHub that offers a free installation service, and any user can find zeta-editing on GitHub to install. At the same time, replicate.com provides the effect of zeta-editing install, users can directly use zeta-editing installed effect in replicate.com for debugging and trial. It also supports api for free installation.

zeta-editing install url in replicate.com:

https://replicate.com/lucataco/zeta-editing

zeta-editing install url in github:

https://github.com/lucataco/cog-audioEditing

Url of zeta-editing

Provider of zeta-editing replicate.com

Other API from lucataco

replicate

Remove background from an image

Total runs: 6.1M
Run Growth: 0
Growth Rate: 0.00%
Updated:September 15 2023
replicate

Falcons.ai Fine-Tuned Vision Transformer (ViT) for NSFW Image Classification

Total runs: 4.5M
Run Growth: 0
Growth Rate: 0.00%
Updated:November 21 2023
replicate

Implementation of Realistic Vision v5.1 with VAE

Total runs: 4.1M
Run Growth: 0
Growth Rate: 0.00%
Updated:August 15 2023
replicate

FLUX.1-Dev LoRA Explorer (DEPRECATED Please use: black-forest-labs/flux-dev-lora)

Total runs: 3.3M
Run Growth: 0
Growth Rate: 0.00%
Updated:October 06 2024
replicate

SDXL ControlNet - Canny

Total runs: 2.4M
Run Growth: 0
Growth Rate: 0.00%
Updated:October 04 2023
replicate

SDXL Inpainting by the HF Diffusers team

Total runs: 2.1M
Run Growth: 0
Growth Rate: 0.00%
Updated:March 06 2024
replicate

Robust face restoration algorithm for old photos/AI-generated faces

Total runs: 1.7M
Run Growth: 0
Growth Rate: 0.00%
Updated:September 06 2023
replicate

Turn any image into a video

Total runs: 1.3M
Run Growth: 0
Growth Rate: 0.00%
Updated:September 03 2023
replicate

Segmind Stable Diffusion Model (SSD-1B) is a distilled 50% smaller version of SDXL, offering a 60% speedup while maintaining high-quality text-to-image generation capabilities

Total runs: 994.7K
Run Growth: 0
Growth Rate: 0.00%
Updated:November 09 2023
replicate

Hyper FLUX 8-step by ByteDance

Total runs: 926.0K
Run Growth: 0
Growth Rate: 0.00%
Updated:August 28 2024
replicate

FLUX.1-Schnell LoRA Explorer

Total runs: 924.9K
Run Growth: 0
Growth Rate: 0.00%
Updated:September 07 2024
replicate

CLIP Interrogator for SDXL optimizes text prompts to match a given image

Total runs: 846.4K
Run Growth: 0
Growth Rate: 0.00%
Updated:May 17 2024
replicate

Coqui XTTS-v2: Multilingual Text To Speech Voice Cloning

Total runs: 835.6K
Run Growth: 0
Growth Rate: 0.00%
Updated:November 28 2023
replicate

A multimodal LLM-based AI assistant, which is trained with alignment techniques. Qwen-VL-Chat supports more flexible interaction, such as multi-round question answering, and creative capabilities.

Total runs: 799.7K
Run Growth: 0
Growth Rate: 0.00%
Updated:October 15 2023
replicate

😊 Hotshot-XL is an AI text-to-GIF model trained to work alongside Stable Diffusion XL

Total runs: 576.9K
Run Growth: 0
Growth Rate: 0.00%
Updated:October 23 2023
replicate

SDXL v1.0 - A text-to-image generative AI model that creates beautiful images

Total runs: 478.6K
Run Growth: 0
Growth Rate: 0.00%
Updated:November 02 2023
replicate

snowflake-arctic-embed is a suite of text embedding models that focuses on creating high-quality retrieval models optimized for performance

Total runs: 397.3K
Run Growth: 0
Growth Rate: 0.00%
Updated:April 20 2024
replicate

Latent Consistency Model (LCM): SDXL, distills the original model into a version that requires fewer steps (4 to 8 instead of the original 25 to 50)

Total runs: 395.0K
Run Growth: 0
Growth Rate: 0.00%
Updated:November 13 2023
replicate

Monster Labs QrCode ControlNet on top of SD Realistic Vision v5.1

Total runs: 390.5K
Run Growth: 0
Growth Rate: 0.00%
Updated:September 24 2023
replicate

moondream2 is a small vision language model designed to run efficiently on edge devices

Total runs: 370.8K
Run Growth: 0
Growth Rate: 0.00%
Updated:July 29 2024
replicate

RealvisXL-v2.0 with LCM LoRA - requires fewer steps (4 to 8 instead of the original 40 to 50)

Total runs: 292.0K
Run Growth: 0
Growth Rate: 0.00%
Updated:November 16 2023
replicate

Implementation of SDXL RealVisXL_V2.0

Total runs: 284.9K
Run Growth: 0
Growth Rate: 0.00%
Updated:November 09 2023
replicate

Animate Your Personalized Text-to-Image Diffusion Models

Total runs: 281.9K
Run Growth: 0
Growth Rate: 0.00%
Updated:September 25 2023
replicate

Practical face restoration algorithm for *old photos* or *AI-generated faces* (for larger images)

Total runs: 263.5K
Run Growth: 0
Growth Rate: 0.00%
Updated:August 03 2023
replicate

DreamShaper is a general purpose SD model that aims at doing everything well, photos, art, anime, manga. It's designed to match Midjourney and DALL-E.

Total runs: 200.5K
Run Growth: 0
Growth Rate: 0.00%
Updated:December 20 2023
replicate

Real-ESRGAN Video Upscaler

Total runs: 191.1K
Run Growth: 0
Growth Rate: 0.00%
Updated:November 25 2023
replicate

A unique fusion that showcases exceptional prompt adherence and semantic understanding, it seems to be a step above base SDXL and a step closer to DALLE-3 in terms of prompt comprehension

Total runs: 125.7K
Run Growth: 0
Growth Rate: 0.00%
Updated:December 27 2023
replicate

CLIP Interrogator (for faster inference)

Total runs: 122.8K
Run Growth: 0
Growth Rate: 0.00%
Updated:September 12 2023
replicate

dreamshaper-xl-lightning is a Stable Diffusion model that has been fine-tuned on SDXL

Total runs: 114.6K
Run Growth: 0
Growth Rate: 0.00%
Updated:February 27 2024
replicate

AI-driven audio enhancement for your audio files, powered by Resemble AI

Total runs: 88.1K
Run Growth: 0
Growth Rate: 0.00%
Updated:December 15 2023
replicate

Phi-3-Mini-4K-Instruct is a 3.8B parameters, lightweight, state-of-the-art open model trained with the Phi-3 datasets

Total runs: 81.7K
Run Growth: 0
Growth Rate: 0.00%
Updated:July 03 2024
replicate

SDXL_Niji_Special Edition

Total runs: 71.3K
Run Growth: 0
Growth Rate: 0.00%
Updated:November 14 2023
replicate

Florence-2: Advancing a Unified Representation for a Variety of Vision Tasks

Total runs: 70.8K
Run Growth: 0
Growth Rate: 0.00%
Updated:June 26 2024
replicate

PixArt-Alpha 1024px is a transformer-based text-to-image diffusion system trained on text embeddings from T5

Total runs: 64.6K
Run Growth: 0
Growth Rate: 0.00%
Updated:December 04 2023
replicate

MagicAnimate: Temporally Consistent Human Image Animation using Diffusion Model

Total runs: 55.8K
Run Growth: 0
Growth Rate: 0.00%
Updated:December 05 2023
replicate

Ostris AI-Toolkit for Flux LoRA Training (DEPRECATED. Please use: ostris/flux-dev-lora-trainer)

Total runs: 55.5K
Run Growth: 0
Growth Rate: 0.00%
Updated:August 18 2024
replicate

Dreamshaper-7 img2img with LCM LoRA for faster inference

Total runs: 55.1K
Run Growth: 0
Growth Rate: 0.00%
Updated:November 17 2023
replicate

Florence-2: Advancing a Unified Representation for a Variety of Vision Tasks

Total runs: 46.6K
Run Growth: 0
Growth Rate: 0.00%
Updated:June 26 2024
replicate

Implementation of SDXL RealVisXL_V1.0

Total runs: 44.0K
Run Growth: 0
Growth Rate: 0.00%
Updated:September 13 2023
replicate

SDXL Image Blending

Total runs: 42.5K
Run Growth: 0
Growth Rate: 0.00%
Updated:December 12 2023
replicate

(Academic and Non-commercial use only) Pixel-Aware Stable Diffusion for Realistic Image Super-resolution and Personalized Stylization

Total runs: 40.3K
Run Growth: 0
Growth Rate: 0.00%
Updated:January 08 2024
replicate

BakLLaVA-1 is a Mistral 7B base augmented with the LLaVA 1.5 architecture

Total runs: 39.0K
Run Growth: 0
Growth Rate: 0.00%
Updated:October 24 2023
replicate

lmsys/vicuna-13b-v1.3

Total runs: 38.4K
Run Growth: 0
Growth Rate: 0.00%
Updated:June 30 2023
replicate

Mistral-7B-v0.1 fine tuned for chat with the Dolphin dataset (an open-source implementation of Microsoft's Orca)

Total runs: 35.9K
Run Growth: 0
Growth Rate: 0.00%
Updated:October 31 2023
replicate

Real-ESRGAN with optional face correction and adjustable upscale (for larger images)

Total runs: 34.5K
Run Growth: 0
Growth Rate: 0.00%
Updated:July 17 2023
replicate

Latest model in the Qwen family for chatting with video and image models

Total runs: 34.2K
Run Growth: 0
Growth Rate: 0.00%
Updated:December 21 2024
replicate

Gemma2 2b by Google

Total runs: 33.1K
Run Growth: 0
Growth Rate: 0.00%
Updated:August 01 2024
replicate

The image prompt adapter is designed to enable a pretrained text-to-image diffusion model to generate SDXL images with an image prompt

Total runs: 31.7K
Run Growth: 0
Growth Rate: 0.00%
Updated:November 12 2023
replicate

(Research only) IP-Adapter-FaceID can generate various style images conditioned on a face with only text prompts

Total runs: 29.3K
Run Growth: 0
Growth Rate: 0.00%
Updated:December 21 2023
replicate

lmsys/vicuna-7b-v1.3

Total runs: 28.6K
Run Growth: 0
Growth Rate: 0.00%
Updated:June 30 2023
replicate

Meta's Llama 2 7b Chat - GPTQ

Total runs: 20.3K
Run Growth: 0
Growth Rate: 0.00%
Updated:July 24 2023
replicate

Segment Anything 2 (SAM2) by Meta - Automatic mask generation

Total runs: 19.9K
Run Growth: 0
Growth Rate: 0.00%
Updated:July 31 2024
replicate

Stylized Audio-Driven Single Image Talking Face Animation

Total runs: 18.9K
Run Growth: 0
Growth Rate: 0.00%
Updated:October 08 2023
replicate

sdxs-512-0.9 can generate high-resolution images in real-time based on prompt texts, trained using score distillation and feature matching

Total runs: 18.8K
Run Growth: 0
Growth Rate: 0.00%
Updated:March 28 2024
replicate

Meta's Llama 2 13b Chat - GPTQ

Total runs: 18.6K
Run Growth: 0
Growth Rate: 0.00%
Updated:July 25 2023
replicate

WizardCoder: Empowering Code Large Language Models with Evol-Instruct

Total runs: 17.0K
Run Growth: 0
Growth Rate: 0.00%
Updated:January 24 2024
replicate

ThinkDiffusionXL is a go-to model capable of amazing photorealism that's also versatile enough to generate high-quality images across a variety of styles and subjects without needing to be a prompting genius

Total runs: 15.6K
Run Growth: 0
Growth Rate: 0.00%
Updated:November 07 2023
replicate

This is wizard-vicuna-13b trained with a subset of the dataset - responses that contained alignment / moralizing were removed

Total runs: 15.1K
Run Growth: 0
Growth Rate: 0.00%
Updated:April 26 2024
replicate

Hyper FLUX 16-step by ByteDance

Total runs: 15.0K
Run Growth: 0
Growth Rate: 0.00%
Updated:August 28 2024
replicate

Mistral-7B-v0.1 fine tuned for chat with the Dolphin dataset (an open-source implementation of Microsoft's Orca)

Total runs: 13.4K
Run Growth: 0
Growth Rate: 0.00%
Updated:October 31 2023
replicate

Image-to-video - SEINE: Short-to-Long Video Diffusion Model for Generative Transition and Prediction

Total runs: 12.8K
Run Growth: 0
Growth Rate: 0.00%
Updated:November 23 2023
replicate

Auto fuse a user's face onto the template image, with a similar appearance to the user

Total runs: 11.8K
Run Growth: 0
Growth Rate: 0.00%
Updated:November 16 2023
replicate

Segments an audio recording based on who is speaking (on A100)

Total runs: 11.5K
Run Growth: 0
Growth Rate: 0.00%
Updated:July 22 2023
replicate

(Research only) Moondream1 is a vision language model that performs on par with models twice its size

Total runs: 11.4K
Run Growth: 0
Growth Rate: 0.00%
Updated:January 25 2024
replicate

InterpAny-Clearer: Clearer anytime frame interpolation & Manipulated interpolation

Total runs: 11.4K
Run Growth: 0
Growth Rate: 0.00%
Updated:November 30 2023
replicate

Image to Image enhancer using DemoFusion

Total runs: 10.6K
Run Growth: 0
Growth Rate: 0.00%
Updated:December 09 2023
replicate

Open diffusion model for high-quality video generation

Total runs: 10.4K
Run Growth: 0
Growth Rate: 0.00%
Updated:October 19 2023
replicate

Orpheus 3B - high quality, emotive Text to Speech

Total runs: 10.2K
Run Growth: 0
Growth Rate: 0.00%
Updated:March 21 2025
replicate

DemoFusion: Democratising High-Resolution Image Generation With No 💰

Total runs: 9.2K
Run Growth: 0
Growth Rate: 0.00%
Updated:December 04 2023
replicate

Implementation of SDXL RealVisXL_V2.0 img2img

Total runs: 8.7K
Run Growth: 0
Growth Rate: 0.00%
Updated:November 07 2023
replicate

Phi-3-Mini-128K-Instruct is a 3.8 billion-parameter, lightweight, state-of-the-art open model trained using the Phi-3 datasets

Total runs: 8.0K
Run Growth: 0
Growth Rate: 0.00%
Updated:April 26 2024
replicate

In-Context LoRA with Image-to-Image and Inpainting to apply your logo to anything

Total runs: 6.8K
Run Growth: 0
Growth Rate: 0.00%
Updated:February 26 2025
replicate

Cog wrapper for Ollama llama3:70b

Total runs: 6.6K
Run Growth: 0
Growth Rate: 0.00%
Updated:July 09 2024
replicate

360 Panorama SDXL image with inpainted wrapping seam

Total runs: 6.3K
Run Growth: 0
Growth Rate: 0.00%
Updated:September 10 2023
replicate

Convert your videos to DensePose and use it with MagicAnimate

Total runs: 5.9K
Run Growth: 0
Growth Rate: 0.00%
Updated:December 06 2023
replicate

Projection module trained to add vision capabilties to Llama 3 using SigLIP

Total runs: 5.5K
Run Growth: 0
Growth Rate: 0.00%
Updated:November 05 2024
replicate

Fuyu-8B is a multi-modal text and image transformer trained by Adept AI

Total runs: 4.6K
Run Growth: 0
Growth Rate: 0.00%
Updated:October 20 2023
replicate

Qwen1.5 is the beta version of Qwen2, a transformer-based decoder-only language model pretrained on a large amount of data

Total runs: 4.1K
Run Growth: 0
Growth Rate: 0.00%
Updated:February 07 2024
replicate

Controlnet v1.1 - Tile Version

Total runs: 4.0K
Run Growth: 0
Growth Rate: 0.00%
Updated:November 27 2023
replicate

A combination of ip_adapter SDv1.5 and mediapipe-face to inpaint a face

Total runs: 3.9K
Run Growth: 0
Growth Rate: 0.00%
Updated:November 15 2023
replicate

SDXL using DeepCache

Total runs: 3.8K
Run Growth: 0
Growth Rate: 0.00%
Updated:January 08 2024
replicate

Segmind Stable Diffusion Model (SSD-1B) img2img

Total runs: 3.7K
Run Growth: 0
Growth Rate: 0.00%
Updated:November 03 2023
replicate

Phi-2 by Microsoft

Total runs: 3.6K
Run Growth: 0
Growth Rate: 0.00%
Updated:January 31 2024
replicate

Playground v2 is a diffusion-based text-to-image generative model trained from scratch. Try out all 3 models here

Total runs: 3.6K
Run Growth: 0
Growth Rate: 0.00%
Updated:December 08 2023