Stable Point Aware 3D
Stable Point Aware 3D (SPAR3D) can make real-time edits and create the complete structure of a 3D object from a single image in a few seconds. SPAR3D combines the strengths of point-cloud diffusion (probabilistic) and mesh regression (deterministic) to have improved details on the unseen back regions in the input image.
Compared to our previous model Stable Fast 3D, this new one allows editing of backside information using the point cloud representation and also leverages a larger Diffusion model to generally improve the depth and backside predictions.
Read more about the model capabilities here.
This API is currently in preview. Please don’t hesitate to contact us with any questions.
Try it out
Grab your API key and head over to
How to use
Please invoke this endpoint with a POST request.
The headers of the request must include an API key in the authorization field. The body of the request must be
multipart/form-data.
The body of the request should include:
image
The body may optionally include:
texture_resolutionforeground_ratioremeshtarget_typetarget_countguidance_scaleseed
Note: for more details about these parameters please see the request schema below.
Output
The output is a binary blob that includes a glTF asset, including JSON, buffers, and images. See the GLB File Format Specification for more details.
Credits
Flat rate of 4 credits per successful generation. You will not be charged for failed generations.
Authorizations
Use your Stability API key to authentication requests to this App.
Headers
Your Stability API key, used to authenticate your requests. Although you may have multiple keys in your account, you should use the same key for all requests to this API.
1The content type of the request body. Do not manually specify this header; your HTTP client library will automatically include the appropriate boundary parameter.
1"multipart/form-data"
The name of your application, used to help us communicate app-specific debugging or moderation issues to you.
256"my-awesome-app"
A unique identifier for your end user. Used to help us communicate user-specific debugging or moderation issues to you. Feel free to obfuscate this value to protect user privacy.
256"DiscordUser#9999"
The version of your application, used to help us communicate version-specific debugging or moderation issues to you.
256"1.2.1"
Body
The image to generate a 3D model from.
Supported Formats:
- jpeg
- png
- webp
Validation Rules:
- Every side must be at least 64 pixels
- Total pixel count must be between 4,096 and 4,194,304 pixels
Determines the resolution of the textures used for both the albedo (color) map and the
normal map. The resolution is specified in pixels, and a higher value corresponds to a
higher level of detail in the textures, allowing for more intricate and precise rendering
of surfaces. However, increasing the resolution also results in larger asset sizes, which
may impact loading times and performance. 1024 is a good default value and rarely requires
changing.
512, 1024, 2048 Controls the amount of padding around the object to be processed within the frame. This
ratio determines the relative size of the object compared to the total frame size. A
higher ratio means less padding and a larger object, while a lower ratio increases the
padding, effectively reducing the object’s size within the frame. This can be useful when
a long and narrow object, such as a car or bus, is viewed from the front (the narrow
side). Here, lowering the foreground ratio might help prevent the generated 3D assets from
appearing squished or distorted. The default value of 1.3 is good for most objects.
1 <= x <= 2Controls the remeshing algorithm used to generate the 3D model. The remeshing algorithm determines how the 3D model is constructed from the input image. The default value of "none" means that the model is generated without remeshing, which is suitable for most use cases. The "triangle" option generates a model with triangular faces, while the "quad" option generates a model with quadrilateral faces. The "quad" option is useful when the 3D model will be used in DCC tools such as Maya or Blender.
none, triangle, quad If set to vertex or face, the result will have approximately target_count many vertices or
faces in the simplified mesh, respectively.
none, vertex, face This sets the target vertex or face count defined by target_type. Selecting extremely low
counts reduces the quality of the mesh severely and values of 1,000 - 10,000 are recommended.
100 <= x <= 20000This sets the guidance scaling of the point diffusion module. Lower values produce less
detail and higher can introduce artifacts. The default of 3 produces best results.
1 <= x <= 10A specific value that is used to guide the 'randomness' of the generation. (Omit this parameter or pass 0 to use a random seed.)
0 <= x <= 4294967294Response
Generation was successful.
The bytes of the generated 3D model.