Skip to main content
Stay up to date on the latest Stability AI API features, models, and deprecations.
Stable Audio 3.0
Today we are releasing an API for Stable Audio 3.0, our most advanced AI audio generation model that produces high-quality, coherent musical tracks up to six minutes long at 44.1kHz stereo. Introducing groundbreaking audio-to-audio capabilities, Stable Audio 3.0 enables users to upload and transform audio samples using natural language prompts.Read more about the model capabilities here.Stable Audio 3.0 was exclusively trained on licensed data from the AudioSparx music library, honoring opt-out requests and ensuring fair compensation for creators.Try Stable Audio 3.0 for free at stableaudio.com.
API Pricing Notice
As of August 1, 2025, we’re increasing the pricing for select API services. See updated pricing here.
API Deprecation and Pricing Notice
We have made a few changes to our API services. These changes will help us maintain high-quality service while investing in new capabilities that benefit our users.Service deprecations: We will no longer support the following API endpoints, effective July 24:
  • Stable Video Diffusion API
  • Stable Diffusion 1.6 API
Pricing updates: We’ve increased the pricing for select API services, effective August 1. See more details here.For users currently using the SD1.6 API, we recommend migrating to Stable Diffusion XL, which offers significantly improved output quality while maintaining the same pricing structure as SD1.6.We also encourage taking this opportunity to upgrade to the latest tools. The Stability AI API has a range of additional image generation offerings to explore:
  • Stable Image Core: Optimized for fast and affordable image generation, Stable Image Core is the next-generation model following SDXL.
  • Stable Image Ultra: Our state-of-the-art text to image model based on Stable Diffusion 3.5.
  • Stable Diffusion 3.5 model family: Stability AI’s latest base models.
While Stable Diffusion Video will no longer be available via API, you can still deploy the model in your environment with a Self-Hosted License. Our team is happy to help if you need support.
API Deprecation Notice
As of April 17, 2025, we are deprecating the Stable Diffusion 3.0 APIs and automatically upgrading to Stable Diffusion 3.5 APIs at no extra cost.Transitioned Models:
  • sd3-large -> sd3.5-large
  • sd3-large-turbo -> sd3.5-large-turbo
  • sd3-medium -> sd3.5-medium
Here’s what you need to know:
  • Higher-quality images: Stable Diffusion 3.5 offers top-tier performance in prompt adherence and image quality. You can learn more about the model here.
  • No action needed: Your API calls will be automatically transitioned.
  • Same pricing: You’ll get improved quality at no extra cost. See price list here.
Additional Image Models:We have a wide range of image generating models on the Stability AI API:
  • Stable Image Ultra: Our highest-quality model, powered by SD3.5 Large ($0.08/image), we’d encourage you to try it!
  • Stable Image Core: Optimized for speed and cost-efficiency ($0.03/image).
Style Transfer
Today we are releasing Style Transfer on the Stability AI API. This feature allows users to apply visual aesthetics from one image to another by uploading a reference style image and a target image.Unlike our Style Guide API, which extracts stylistic elements from an input image (control image) and uses it to guide the creation of an output image based on the prompt, Style Transfer specifically applies styling from one existing image to another existing image without generating new content.Style Transfer helps maintain brand consistency, unify visual content, and design prototyping. It streamlines the implementation of visual guidelines and facilitates experimentation with different aesthetic directions.As always, all of our APIs come with standard safety features for input filtering, NSFW filtering of input images and generated content, known CSAM content filtering through Thorn, and C2PA signing.
Stable Audio 2.0
Today we are releasing an API for Stable Audio 2.0, our most advanced AI audio generation model that produces high-quality, coherent musical tracks up to three minutes long at 44.1kHz stereo. Introducing groundbreaking audio-to-audio capabilities, Stable Audio 2.0 enables users to upload and transform audio samples using natural language prompts.Read more about the model capabilities here.Stable Audio 2.0 was exclusively trained on licensed data from the AudioSparx music library, honoring opt-out requests and ensuring fair compensation for creators.Try Stable Audio 2.0 for free at stableaudio.com.
Stable Point Aware 3D
Today we are releasing Stable Point Aware 3D (SPAR3D), a 3D generative AI model that introduces the ability to make real-time edits and create the complete structure of a 3D object from a single image in a few seconds. Introducing a first-of-its-kind two-stage architecture, SPAR3D enables unparalleled control over 3D object generation.Compared to our previous model Stable Fast 3D, this new one allows editing of backside information using the point cloud representation and also leverages a larger Diffusion model to generally improve the depth and backside predictions.Read more about the model capabilities here.This API is currently in preview. Please don’t hesitate to contact us with any questions.As always, all of our APIs come with standard safety features for input filtering, NSFW filtering of input images and generated content, known CSAM content filtering through Thorn, and C2PA signing.
Replace Background & Relight
We are pleased to announce our newest Edit API, Replace Background and Relight. This service enables you to replace the background of your images. Additionally, we have updates to our flagship Stable Image Ultra service and the SD3 & 3.5 service.

New

  • Edit Replace Background & Relight: The Replace Background and Relight service lets users swap backgrounds with AI-generated or uploaded images while adjusting lighting to match the subject. This new API provides a streamlined image editing solution and can serve e-commerce, real estate, photography, and creative projects. Some of the things you can do include:
    • Background Replacement: Remove existing background and add new ones.
    • AI Background Generation: Create new backgrounds using AI generated images based on prompts.
    • Relighting: Adjust lighting in images that are under or overexposed.
    • Flexible Inputs: Use your own background image or generate one.
    • Lighting Adjustments: Modify light reference, direction, and strength.

Updated

  • Stable Image Ultra: Stable Image Ultra now supports image-to-image.
  • Stable Diffusion 3 & 3.5: The SD3 and SD3.5 models now support the cfg_scale parameter. This parameter controls how strictly the diffusion process adheres to the prompt text (higher values keep your image closer to your prompt).
As always, all of our APIs come with standard safety features for input filtering, NSFW filtering of input images and generated content, known CSAM content filtering through Thorn, and C2PA signing.
Stable Diffusion 3.5 Medium & Stable Image Ultra
Today we are releasing a new model in the Stable Diffusion 3.5 suite on our API, our most powerful base models yet, and an update to our flagship Stable Image Ultra service:
  • Stable Diffusion 3.5 Medium: With 2.5 billion parameters, the model delivers an optimal balance between prompt accuracy and image quality, making it an efficient choice for fast high-performance image generation.
  • Stable Image Ultra: Following the release of the Stable Diffusion 3.5 model suite, we are updating the Stable Image Ultra API with Stable Diffusion 3.5 Large, our largest and most powerful base model yet. Stable Image Ultra is our flagship image service, offering the highest quality and detail.
Read more about the model capabilities here.

New

  • Stable Diffusion 3.5 Medium: This endpoint now supports sd3.5-medium in the model parameter, allowing you to generate images with SD 3.5 Medium.

Updated

As always, all of our APIs come with standard safety features for input filtering, NSFW filtering of input images and generated content, known CSAM content filtering through Thorn, and C2PA signing.
Stable Diffusion 3.5 Large & Large Turbo
Today we are releasing Stable Diffusion 3.5 on our API, our most powerful base models yet:
  • Stable Diffusion 3.5 Large: At 8 billion parameters, with superior quality and prompt adherence, this base model is the most powerful in the Stable Diffusion family. This model is ideal for professional use cases at 1 megapixel resolution.
  • Stable Diffusion 3.5 Large Turbo: A distilled version of Stable Diffusion 3.5 Large. SD3.5 Large Turbo generates high-quality images with exceptional prompt adherence in just 4 steps, making it considerably faster than Stable Diffusion 3.5 Large.
Read more about the model capabilities here.

New

  • Stable Diffusion 3.5: This endpoint now supports sd3.5-large and sd3.5-large-turbo in the model parameter, allowing you to generate images with SD3.5 Large and SD3.5 Large Turbo.
As always, all of our APIs come with standard safety features for input filtering, NSFW filtering of input images and generated content, known CSAM content filtering through Thorn, and C2PA signing.
API Deprecation Notice
As of October 11, 2024, we have deprecated a number of our older models and APIs. Below we list the endpoints that are no longer supported and provide recommendations for alternatives.

Deprecated Models

Text-to-image:
  • stable-diffusion-512-v2-1
  • stable-diffusion-xl-beta-v2-2-2
  • stable-diffusion-xl-1024-v0-9
Image Upscale:
  • esrgan-v1-x2plus
Alternative Text-to-Image Solutions:
  • SDXL 1.0: Our legacy base model for straightforward image generation.
  • Stable Image Ultra: Our flagship image service based on Stable Diffusion 3 Large, offering the highest quality and detail.
  • Stable Diffusion 3 Model Suite: Latest base models, balancing speed and quality for media, entertainment, and retail.
  • Stable Image Core: Optimized for fast and cost-effective image generation.
Alternative Upscaling Solution:If you have any questions or need further assistance, please don’t hesitate to reach out to our support team.Thank you for your continued support.
Stable Image Fast Upscaler
We are excited to release our newest Upscale API, Fast Upscaler. This service enhances image resolution by 4x using predictive and generative AI. This lightweight and fast service (processing in ~1 second) is ideal for enhancing the quality of compressed images, making it suitable for social media posts and other applications.

How does the Fast Upscaler compare to our other Upscale Services?

The Fast Upscaler offers a more affordable and quicker alternative to the Conservative Upscaler which can upscale images by 20 to 40 times to a 4 megapixel output image. The Conservative Upscaler can upscale images as small as 64x64 pixels directly to a 4 megapixel output. Use this option if you directly need a 4 megapixel output.If you wish to upscale highly degraded images (lower than 1 megapixel) with a creative twist, we recommend trying the Creative Upscaler.

New

  • Fast Upscaler: Simple, low-cost upscaler used to increase image resolution by 4x, up to 4 megapixels.
As always, all of our APIs come with standard safety features for input filtering, NSFW filtering of input images and generated content, known CSAM content filtering through Thorn, and C2PA signing.
Stable Image Edit Search and Recolor
We are pleased to announce our newest Edit API, Search and Recolor. This service enables you to quickly change the color of a specific object in an image using a prompt.

New

  • Edit Search and Recolor: The Search and Recolor service allows to change the color of a specific object in an image using a prompt. This service is a specific version of inpainting that does not require a mask. The Search and Recolor service will automatically segment the object and recolor it using the colors requested in the prompt.

Updated

  • Stable Fast 3D: There is a new parameter, remesh, available which enables the use of the triangle remeshing algorithm during 3D model generation.
As always, all of our APIs come with standard safety features for input filtering, NSFW filtering of input images and generated content, known CSAM content filtering through Thorn, and C2PA signing.
Stable Fast 3D
We are excited to introduce Stable Fast 3D, Stability AI’s latest breakthrough in 3D asset generation technology. This innovative model transforms a single input image into a detailed 3D asset, setting a new standard for speed and quality in the field of 3D reconstruction.

New

  • Stable Fast 3D: The Stable Fast 3D API generates high-quality 3D assets from a single image in just 0.5 seconds. The model has applications for game and virtual reality developers, as well as professionals in retail, architecture, design and other graphic-intense professions. Stable Fast 3D rapidly generates a complete 3D asset, including:
    • UV unwrapped mesh
    • Material parameters
    • Albedo colors with reduced illumination bake-in
    • Optional quad or triangle remeshing (adding only 100-200ms to processing time)
As always, all of our APIs come with standard safety features for input filtering, NSFW filtering of input images and generated content, known CSAM content filtering through Thorn, and C2PA signing.
Stable Image Control Style
We are pleased to announce our newest API, Control Style.

New

  • Control Style: The Control Style service extracts stylistic elements from an input image (control image) and uses it to guide the creation of an output image based on the prompt. The result is a new image in the same style as the control image.
As always, all of our APIs come with standard safety features for input filtering, NSFW filtering of input images and generated content, known CSAM content filtering through Thorn, and C2PA signing.
Stable Diffusion 3 Medium Release
Today we are excited to announce the release of Stable Diffusion 3 Medium! This 2 billion parameter model is now available for use in the REST API and the weights can be downloaded from HuggingFace.

Updated

  • SD3: This endpoint now supports sd3-medium in the model parameter, allowing you to generate images with SD3 Medium.
Stable Image Generate Ultra
Today we are excited to announce Stable Image Generate Ultra, our latest and most powerful image generation service.

New

  • Ultra: The Ultra image generation service creates the best quality images with the best alignment to prompt. It is Stability’s highest quality image generator and leverages the latest models for the highest quality and prompt adherence.
As always, all of our APIs come with standard safety features for input filtering, NSFW filtering of input images and generated content, known CSAM content filtering through Thorn, and C2PA signing.
Stable Image Erase Object & Conservative Upscale
Today we are pleased to announce two new services, Erase Object and Conservative Upscale. These new endpoints enable exciting new workflows. Additionally, we have made enhancements to Outpaint.

New

  • Erase Object: The Erase Object service removes unwanted objects, such as blemishes on portraits or items on desks.
  • Conservative Upscale: This service takes images between 64x64 pixels and 1 megapixel and upscales them all the way to 4K resolution. Put more generally, it can upscale images ~20-40x times while preserving all aspects. Conservative Upscale minimizes alterations to the image and should not be used to reimagine an image.

Updated

  • Outpaint: Now supports outpainting up to 2000 pixels in any direction (up from 512 px) and a new creativity parameter which controls the likelihood of creating additional details not conditioned by the initial image.
As always, all of our APIs come with standard safety features for input filtering, NSFW filtering of input images and generated content, known CSAM content filtering through Thorn, C2PA signing, and Watermarking.
Stable Image Control Sketch & Control Structure
We are excited to announce two new services, Control Sketch and Control Structure. These are the first in our suite of Control-related services. Additionally, we have made improvements to Outpaint and Creative Upscale.

New

  • Control Sketch: This service offers an ideal solution for design projects that require brainstorming and frequent iterations. It upgrades rough hand-drawn sketches to refined outputs with precise control. For non-sketch images, it allows detailed manipulation of the final appearance by leveraging the contour lines and edges within the image. Check out the docs to learn more.
  • Control Structure: This service excels in generating images by maintaining the structure of an input image, making it especially valuable for advanced content creation scenarios such as recreating scenes or rendering characters from models. Check out the docs to learn more.

Updated

  • Outpaint: Significantly improved coherence and diversity in the outpainted image area. Fixed a bug where some down outpaint tasks failed.
  • Creative Upscale: Significantly faster outputs and improved quality in photorealism. Improved diversity in non-photorealistic outcomes.
All of our APIs come with standard safety features for input filtering, NSFW filtering of input images and generated content, known CSAM content filtering through Thorn, C2PA signing, and Watermarking. Additionally, we are working to add additional safety filters to all APIs which rewrites prompts not aligned with our safety policy.
Stable Image, Stablevideo.com, and v2beta API
This release note covers:
  • Stable Image Services
  • Stable Video
  • REST API v2beta and legacy

Stable Image

We are proud to announce Stable Image Services, Stability AI’s new flagship suite of API services for image. Stable Image offers the best-in-class experience for developers building products in the design space, and is categorized into four sets:
  • Generate: The premium image generation services. This includes the new flagship Stable Image Core service, which leverages the latest models. Older simple model services of SDXL 1.0 and SD 1.6 are also represented in this category. Stay tuned for Stable Diffusion 3 to be added to this category as well.
  • Upscale: The best image enhancement services. This includes the Creative Upscaler as well legacy ESRGAN upscaler.
  • Edit: The ideal services for inpainting, outpainting and similar image editing tools. Includes the new Search and Replace service that enables natural language image editing.
  • Control: The premium image-to-image service that leverages advanced models like ControlNets. Perfect for advanced users that want precise control over image generation.
These services are in early release, and are being constantly updated. All services are available in the new REST v2beta API.

Stable Video

We are also excited to announce a new platform for video generation, stablevideo.com. Following up on the API release of SVD 1.1, the new stablevideo.com web application is a perfect reference implementation for those hoping to try out Stability’s SVD 1.1 API and explore building something similar. We will be continuously improving this site. Expect to see more services across modalities appearing on stablevideo.com.

REST API v2beta

We have been working to improve our API services for improved reliability and ease of use.All new services are being released in the REST v2beta API, which includes both synchronous and asynchronous modes.Our documentation is being updated as such. All other services are being moved to the “Legacy” section. Services in “Legacy” are going to be in maintenance mode, and new features in the API will not be added to these services.We are not deprecating Legacy API services, however we intend to upgrade all services to REST v2beta over time, where they will benefit from improved services that are on our API roadmap.
Stable Video API Update
We are excited to launch an update of the Stable Video API. This drop-in upgrade provides a general improvement to the Stable Video API, including a 20 second improvement in speed and significantly better quality and coherence.By our estimates, the need for cherry picking is now significantly lower by comparison to the previous API as well as to competing products. Hop over to the Stable Video page to learn more!As we continue to enhance our suite of API services for Generative Media, stay tuned for upcoming API releases in video image, 3D, and audio.
Stable Video Diffusion API Alpha Release

Overview

We are excited to announce the alpha release of the Stable Video Diffusion API! This release marks a major milestone in our ongoing quest to develop versatile models for diverse applications. For more information please see our Getting Started page introducing the Stable Video API.

Example Output

What’s New

  • Added /v2alpha to the REST API. This version includes support for the alpha Stable Video Diffusion API!
Stable Diffusion v1.6 Release

Overview

We’re excited to announce the release of the Stable Diffusion v1.6 in the REST API! This model is designed to be a higher quality, more cost-effective alternative to stable-diffusion-v1-5 and is ideal for users who are looking to replace it in their workflows. stable-diffusion-v1-6 has been optimized to provide higher quality 512px generations when compared to stable-diffusion-v1-5.Prompt:A wolf in Yosemite National Park, chilly nature documentary film photographyDimensions:512x512Steps:50Seed:1CFG Scale:7Sampler:DDIMPlease note however, that image-to-image is not currently supported via stable-diffusion-v1-6.

What’s New

API

  • Introduced support for the stable-diffusion-v1-6 engine to the REST API.

Deprecations

  • A number of our legacy models will be deprecated by the 15th of November. Please visit the API Parameters page for additional information.
SDXL v1.0 Release

Overview

We’re excited to announce the release of Stable Diffusion XL v1.0, the first major version release in the SDXL series!

What’s New

Building upon the success of the SDXL v0.9 release, SDXL v1.0 brings with it a number of image detail improvements. Specifically, SDXL v1.0 offers enhanced vibrancy and overall color tone accuracy, including deeper black and brighter white tones.Prompt:A wolf in Yosemite National Park, chilly nature documentary film photographyNegative Prompt:3d render, smooth, plastic, blurry, grainy, low-resolution, anime, deep-fried, oversaturatedSteps:50Seed:992446758CFG Scale:8Left Model: stable-diffusion-xl-1024-v0-9 AKA SDXL v0.9 Right Model: stable-diffusion-xl-1024-v1-0 AKA SDXL v1.0

API

  • Introduced support for the stable-diffusion-xl-1024-v1-0 model engine to the API (gRPC SDK + REST), which brings with it new pricing considerations. Please check out the Pricing page for additional information.
  • stable-diffusion-xl-1024-v1-0 is now the default engine for the API. If you wish to use a different engine, you must specify it in your request.
  • stable-diffusion-xl-1024-v1-0 supports a variety of aspect ratios. Please check the API Parameters page for additional information.
SDXL v0.9 Release

Overview

We’re excited to announce the release of Stable Diffusion XL v0.9, the newest model in the SDXL series! Building on the successful release of the Stable Diffusion XL beta, SDXL v0.9 brings marked improvements in image quality and composition detail. Please be sure to check out our blog post for more comprehensive details on the SDXL v0.9 release.

What’s New

SDXL v0.9, trained at a base resolution of 1024 x 1024, produces massively improved image and composition detail over its predecessor.Prompt:A wolf in Yosemite National Park, chilly nature documentary film photographyNegative Prompt:3d render, smooth, plastic, blurry, grainy, low-resolution, anime, deep-fried, oversaturatedLeft Model: stable-diffusion-xl-beta-v2-2-2 AKA SDXL v0.8 Right Model: stable-diffusion-xl-1024-v0-9 AKA SDXL v0.9SDXL v0.9 has also been trained to handle multiple aspect ratios, a clear improvement over previous models that would often see repeated subjects / concepts in wide or tall aspect ratio generations.Prompt:A wolf in Yosemite National Park, chilly nature documentary film photographyNegative Prompt:3d render, smooth, plastic, blurry, grainy, low-resolution, anime, deep-fried, oversaturatedDimensions: 1344x768Model: stable-diffusion-xl-1024-v0-9 AKA SDXL v0.9

API

  • Introduced support for the stable-diffusion-xl-1024-v0-9 model engine to the API (gRPC SDK + REST), which brings with it new pricing considerations. Please check the SDXL v0.9 Pricing Table for additional information.
  • stable-diffusion-xl-1024-v0-9 is now the default engine for the API. If you wish to use a different engine, you must specify it in your request.
  • stable-diffusion-xl-1024-v0-9 supports a variety of aspect ratios. Please check the API Parameters page for additional information.
Stable Animation SDK Release

Overview

We are excited to announce the release of the Stable Animation SDK, a powerful tool designed to empower artists and developers by leveraging Stable Diffusion to creating stunning animations!Explore the possibilities by generating animations using various inputs: text prompts (without images), source images, or source videos. Animate with all of our Stable Diffusion models, including Stable Diffusion XL, at your fingertips.Check out the three ways to create animations:
  1. Text to Animation: Input a text prompt and adjust various parameters to generate an animation.
  2. Text Input + Initial Image Input: Provide an initial image as the starting point for the animation, combined with a text prompt to generate the final output animation.
  3. Input video + Text input: Use an initial video as the basis for the animation, fine-tuning parameters to create the final output animation guided by a text prompt.

What’s New

API

  • Added support for the Stable Animation SDK to the API (gRPC SDK only). Check out the docs to learn more.

Breaking Changes

We’ve introduced a minor breaking change with this release. The utility function image_to_prompt has been moved from stability_sdk.client to stability_sdk.utils, and its parameters have been updated:
  • Old: image_to_prompt(im, init: bool = False, mask: bool = False)
  • New: image_to_prompt(image: Image.Image, type: generation.ArtifactType=generation.ARTIFACT_IMAGE)
This function helps package init images as artifacts in the gRPC Prompt message.
Style Presets / Client ID Support

Overview

We’re excited to announce that we’ve released support for style presets to our API!Style presets are a way to apply pre-defined styles to your images. These are the same style presets available in DreamStudio. Style presets are currently only available in our REST API, with support in Stability for Blender and gRPC coming soon.We’re excited to see what you create with them in your own integrations!Dog in a forest, generated with the digital-art style preset using SDXL.
Dog in a forest, generated with the cinematic style preset using SDXL.
If you’re using our REST API, you can now use style presets by passing the style_preset parameter to the text-to-image, image-to-image and image-to-image/masking endpoints. Check out the REST API docs for more information.

What’s New

API

  • Added support for style presets via the REST API. Check out the docs to learn more.
  • Added support for the Extras field in REST. This field allows you to pass additional information to the generation service for certain beta features.
  • Added support for passing a Stability-Client-ID header to the REST API. This allows Stability to track usage of the API on a per-client basis, and is used in some internal tools. This is not required for normal usage of the API.
SDXL Release

Overview

We’re excited to announce that we’ve released support for SDXL to our API! Please be sure to check out our blog post on SDXL’s release for more information.

What’s New

API

  • Added support for the stable-diffusion-xl-beta-v2-2-2 model engine to the API (gRPC SDK + REST). This engine comes with special pricing considerations, please visit the API Parameters page to learn more.
Here’s a short summary of some of SDXL’s capabilities, taken as an excerpt from our blog post on SDXL’s release:
  • Next-level photorealism capabilities.
  • Enhanced image composition and face generation.
  • Rich visuals and jaw-dropping aesthetics.
  • Use of shorter prompts to create descriptive imagery.
  • Greater capability to produce legible text.

Note: stable-diffusion-xl-beta-v2-2-2 comes with some special considerations with regard to the dimensions it can generate images at:

  • stable-diffusion-xl-beta-v2-2-2 can generate images at a maximum of either 512x896 or 896x512. As usual, dimensions must be divisible by 64px when considering the size of an image generation request.
  • If either the width or the height (but not both) of an image generation request is greater than 512px, the other side (width or height, respectively) cannot be beyond 512px in dimension.
An example of some stunning results generated by SDXL:
Image Upscaling API and Integrations Release

Overview

We are excited to announce that support for image upscaling has been added to our API! The Image Upscaler is a tool that allows you to upscale images and animation frames using our ESRGAN-based upscaler. It is currently available in our Stability for Blender and Stability for Photoshop plugins, as well as our gRPC and REST APIs. This update also brings a number of other fixes and improvements in each integration.Check out the docs to learn how to use the Upscaler in Photoshop, Blender, and via our API.If you are using our gRPC SDK, you can now use the Upscaler by calling the Upscale method on the Stability service. Check out the gRPC docs for more information. REST support for the upscaler is also available here.

What’s New

API

  • Added support for image upscaling via the API. Check out the docs to learn more.

Photoshop

  • Added the Stability Upscaler - this is a new feature that lets you upscale your layers and documents in Photoshop using the Stability SDK. Click on the ‘Upscaler’ tab to try it out.
  • Fixed some issues related to inpainting with DALL-E 2.
  • Updated the Stability SDK version to latest.
  • Updated the plugin icon to reflect its status as a Stability product.
  • Fixed a number of layout bugs related to scrolling and loading new diffusion results.

Blender

  • Added the Stability Upscaler - this is a new feature that lets you upscale your render results and textures in Photoshop using the Stability SDK. Click on the ‘Upscaler’ feature at the top of the add-on panel to try it out.
  • Added more extensive error handling for REST API request failures.
  • Added a timeout to avoid hangs when the underlying generation service times out.
  • Fixed a number of bugs reported by users.
REST API v1

Overview

We are excited to announce the REST API v1 is officially live! The documentation has been integrated directly into this site and can be viewed here.If you encounter any issues or have any feedback for us please stop by the #official-rest-api channel in our community Discord and/or open a GitHub issue.If you’re using /v1alpha or /v1beta please take a look at the Deprecations section below.

What’s New

Upscaling is now available!

It is now possible to upscale images using the REST API. Check out the upscaling docs for more information.

Improved error handling

We’ve added more error handling, improved error messages, and all errors should now include an id field for troubleshooting. If you encounter an issue please drop a message in the #official-rest-api channel in our community Discord and include the id field from the response.

Breaking Changes

Removed height and width from image-to-image and image-to-image/masking endpoints

These endpoints have always used the dimensions of the provided init_image to determine the dimensions of the resulting image, so accepting height and width was confusing and unnecessary.

Passing explicit null values is no longer permitted

Previously, it was possible to pass explicit null values for optional arguments. This is no longer permitted. If you want to use the default value for an optional argument, simply omit it from the request.

Deprecations

  • v1beta is now deprecated and is scheduled for removal on May 1st, 2023
  • v1alpha was already deprecated and is now scheduled for removal on May 1st, 2023

What’s Next

What’s coming up next for the REST API:
  • Add image-to-depth endpoint to v1
  • Add depth-to-image endpoint to v1
REST API v1beta Release Candidate

Overview

We are excited to announce the v1beta release candidate for the REST API is officially live! Check out the new and improved documentation here.For this release we’ve taken your feedback into consideration and have implemented some changes based on suggestions shared with us via the #api channel in our community Discord and GitHub issues.The majority of the changes relate to standardizing input parameters and tweaks to input validation, but this release is a big step towards our goal of making the API as straightforward and easy to use as possible.We are looking for feedback from the community to help us improve the API before we release the final version, so please stop by the #api channel in our community Discord and/or open a GitHub issue.

What’s New

Height and Width Validation (v1alpha,v1beta)

Height and width validation now allows for a wider range of resolutions, spawned by this issue on GitHub.
  • The minimum value for either height or width has been lowered to 128
  • Instead of imposing a maximum value, we now require that the product of height and width fall within a certain range based on the type of engine being using:
    • For 768 engines:
      • 589,824 ≤ (height * width) ≤ 1,048,576
    • All other engines:
      • 262,144 ≤ (height * width) ≤ 1,048,576
Here’s an example of a 192x3072 image of a cyberpunk cityscape, generated with v1beta:
epic battleground, beautiful synthwave city painting, digital illustration, extreme detail, digital art, 4k, ultra hd, beautiful city at night, long exposure city at night photography, full-color, urban street photography, nightlife, synthwave, cityscape, hd photography, digital art, 4k

Better Error Handling (v1alpha,v1beta)

  • Making a request to a 768 engine with less than 768x768 pixels will now result in an error instead of generating a distorted image.
  • Some invalid requests that were possible before are now impossible
    • e.g. making a request with a text prompt of "" (empty string) will now result in an error
  • Most errors should be clearer now and contain less noise
  • Errors now include an id field we can use to help debug issues you run into. If you run into an error, note the id and let us know in the Stable Diffusion Discord so we can help you out!

New Image-to-Image Parameter (v1beta)

We now offer a single parameter image_strength you can use in place of step_schedule_start and step_schedule_end for image-to-image generations that mimics the behavior of the Image Strength slider in DreamStudio. You can continue to pass in step_schedule_start and step_schedule_end if you prefer as long as init_image_mode is set to STEP_SCHEDULE.For more information, check out the image_strength and init_image_mode parameters of the image-to-image endpoint.

Better Image-to-Image Defaults (v1beta)

  • For image-to-image generations we now default to image_strength over step_schedule_start and step_schedule_end
  • The default value for image_strength is 0.35
  • The default value for step_schedule_start is 0.65

Breaking Changes

The only intentional breaking change to v1alpha was the addition of more input validation checks. These changes shouldn’t affect most users, and those who are affected likely weren’t getting the results they expected anyway. If any other breaking changes were introduced, they were unintentional so please notify us via the #api channel in our community Discord or by opening a GitHub issue.

Deprecations

With this release candidate, we are deprecating the v1alpha API.The v1alpha API will continue to be available for the time being, but will be removed in a future release.

What’s Next

  • Move on to v1 as fast as possible
  • Move the REST API documentation into this site (currently hosted here)
  • Add support for Upscaling
  • Add support for Depth-to-Image
REST API v1alpha Release Candidate

Overview

We have been hard at work on the REST API for the past few weeks and are excited to announce that the v1 release candidate is officially live!We are looking for feedback from the community to help us improve the API before we release the final version, so please stop by the API channel in the Stable Diffusion Discord and let us know what you think!

What’s New

  • /image-to-image/masking endpoint with two options for specifying the mask:
    1. mask_image where lighter or darker pixels influence the diffusion process.
    2. init_image where transparent pixels influence the diffusion process.
  • Improved HTTP error codes:
    • 400s for bad requests.
    • 401s for unauthorized requests.
    • 403s for insufficient privileges.
    • 404s for things like trying to use an engine that doesn’t exist.
    • Relegated 500s back to internal server errors (where they belong!)
  • Improved error messages for all endpoints:
    • Magic decoder ring no longer required.
  • Fixed small bug causing null values to show up in the response when Accept header was set to application/json.
  • Addressed some CORS-related issues
For more information around the masking endpoint including working examples in Go, TypeScript, Python and cURL, check out the REST API documentation.
Stable Diffusion 2.1 - API Release

Overview

In this release of the Stability API, we are introducing Stable Diffusion 2.1 (512px + 768px), including multi-prompting with prompt weighting.

What’s New

New Models
  • Stable Diffusion 2.1 - 512px
  • Stable Diffusion 2.1 - 768px
With the Stable Diffusion 2.1 release, we have updated our training strategy and reintroduced much of the artistic flair our users felt was lost in 2.0. Check out the official announcement on the Stability AI blog to learn more.Our existing documentation has been updated to reflect the addition of Stable Diffusion 2.1 as an optional model accessible via the API.Multi-promptingMultiple weighted prompts can now be passed to the API in this release.Multi-prompting allows users to combine concepts to create new and unique results. The model will also attempt to eliminate or avoid concepts in the resulting image when a negative value is assigned to an additional prompt, colloquially known as “negative prompting.”For functional examples, please refer to our new documentation on multi-prompting.Open Source ReleasesStable Diffusion 2.1 model checkpoints are now available via our open source repositories on Hugging Face.
Stable Diffusion 2.0 - API Release

Overview

In this release of the Stability API, we are introducing Stable Diffusion 2.0 (512px + 768px), including two new samplers, and an improved inpainting model (Stable Inpainting 2.0).

What’s New

New Models
  • Stable Diffusion 2.0 - 512px
  • Stable Diffusion 2.0 - 768px
  • Stable Inpainting 2.0 - 512px
Stable Diffusion 2.0 is all-new, trained from the ground up with safety in mind.The 2.0 models are trained on an aesthetic subset of LAION-5B and further filtered via LAION’s NSFW filter, meaning high quality images are available with far less intrusive safety filtering.Stable Diffusion 2.0 models include a new text encoder trained by LAION with support from Stability AI, which improves the quality of generated images when compared to Stable Diffusion 1.x models.The new 768px model offers greater coherence at larger dimensions as it was trained on higher resolution samples (768 x 768). This helps mitigate common issues such as doubling and mosaic effects in larger dimension images generated by 512px models.New SamplersOptimized and backwards compatible with 1.4, 1.5 and 2.0, we’re introducing two new samplers: k_dpmpp_2m and k_dpmpp_2s_ancestral.The benefit of these new samplers is their ability to resolve high quality images at lower required step counts, enabling your users to achieve better results with fewer steps.Inpainting 2.0 is being introduced alongside Stable Diffusion 2.0, offering significantly improved coherency over Inpainting 1.0.Intelligent sampler defaultsFor your convenience, sampler selection is optional. If omitted, our API will select the best sampler for the chosen model and usage mode.Unless you have a specific use case requirement, we recommend you allow our API to select the preferred sampler.Prompting is differentStable Diffusion 2.0 uses the new OpenCLIP ViT-H model which has been trained on a new dataset, meaning that it differs from the previous OpenAI ViT-L model used in our prior models. Thus, prompting techniques from prior models may work differently in Stable Diffusion 2.0.Notably, celebrity and artist names have a lower impact than in prior models.This is a characteristic of the new model, and not considered a defect.Previous Stable Diffusion model versions (1.5, 1.4, Inpainting 1.0) remain available for API use.Open Source ReleasesKeeping with our tradition of open source collaboration and innovation, Stable Diffusion 2.0 was also released as an open source project. The open source release includes the above models and several additional models including a 4x upscaler and a new depth-to-image model capable of inferring the depth of an input image and generating new images from text prompts and incorporating the depth information.These depth-to-image and upscaler models will be implemented into the DreamStudio API in the near future.