Upload folder using huggingface_hub
Browse files- .gitattributes +15 -35
- LICENSE.md +51 -0
- README.md +180 -0
- demo_images/demo (1).png +0 -0
- demo_images/demo (10).png +0 -0
- demo_images/demo (2).png +0 -0
- demo_images/demo (3).png +0 -0
- demo_images/demo (4).png +0 -0
- demo_images/demo (5).png +0 -0
- demo_images/demo (6).png +0 -0
- demo_images/demo (7).png +0 -0
- demo_images/demo (8).png +0 -0
- demo_images/demo (9).png +0 -0
- mmdit.png +0 -0
- sd3demo.jpg +3 -0
- sd3demo_prompts.txt +19 -0
- stable-diffusion-v3-medium-FP16.gguf +3 -0
- stable-diffusion-v3-medium-Q4_0.gguf +3 -0
- stable-diffusion-v3-medium-Q8_0.gguf +3 -0
.gitattributes
CHANGED
@@ -1,35 +1,15 @@
|
|
1 |
-
|
2 |
-
|
3 |
-
|
4 |
-
|
5 |
-
|
6 |
-
|
7 |
-
|
8 |
-
|
9 |
-
|
10 |
-
|
11 |
-
|
12 |
-
|
13 |
-
|
14 |
-
|
15 |
-
*.
|
16 |
-
*.onnx filter=lfs diff=lfs merge=lfs -text
|
17 |
-
*.ot filter=lfs diff=lfs merge=lfs -text
|
18 |
-
*.parquet filter=lfs diff=lfs merge=lfs -text
|
19 |
-
*.pb filter=lfs diff=lfs merge=lfs -text
|
20 |
-
*.pickle filter=lfs diff=lfs merge=lfs -text
|
21 |
-
*.pkl filter=lfs diff=lfs merge=lfs -text
|
22 |
-
*.pt filter=lfs diff=lfs merge=lfs -text
|
23 |
-
*.pth filter=lfs diff=lfs merge=lfs -text
|
24 |
-
*.rar filter=lfs diff=lfs merge=lfs -text
|
25 |
-
*.safetensors filter=lfs diff=lfs merge=lfs -text
|
26 |
-
saved_model/**/* filter=lfs diff=lfs merge=lfs -text
|
27 |
-
*.tar.* filter=lfs diff=lfs merge=lfs -text
|
28 |
-
*.tar filter=lfs diff=lfs merge=lfs -text
|
29 |
-
*.tflite filter=lfs diff=lfs merge=lfs -text
|
30 |
-
*.tgz filter=lfs diff=lfs merge=lfs -text
|
31 |
-
*.wasm filter=lfs diff=lfs merge=lfs -text
|
32 |
-
*.xz filter=lfs diff=lfs merge=lfs -text
|
33 |
-
*.zip filter=lfs diff=lfs merge=lfs -text
|
34 |
-
*.zst filter=lfs diff=lfs merge=lfs -text
|
35 |
-
*tfevents* filter=lfs diff=lfs merge=lfs -text
|
|
|
1 |
+
text_encoder/model.safetensors filter=lfs diff=lfs merge=lfs -text
|
2 |
+
text_encoder_2/model.safetensors filter=lfs diff=lfs merge=lfs -text
|
3 |
+
text_encoder_3/model-00001-of-00002.safetensors filter=lfs diff=lfs merge=lfs -text
|
4 |
+
text_encoder_3/model-00002-of-00002.safetensors filter=lfs diff=lfs merge=lfs -text
|
5 |
+
transformer/diffusion_pytorch_model.safetensors filter=lfs diff=lfs merge=lfs -text
|
6 |
+
vae/diffusion_pytorch_model.safetensors filter=lfs diff=lfs merge=lfs -text
|
7 |
+
sd3demo.jpg filter=lfs diff=lfs merge=lfs -text
|
8 |
+
tokenizer_3/spiece.model filter=lfs diff=lfs merge=lfs -text
|
9 |
+
text_encoder/model.fp16.safetensors filter=lfs diff=lfs merge=lfs -text
|
10 |
+
text_encoder_2/model.fp16.safetensors filter=lfs diff=lfs merge=lfs -text
|
11 |
+
text_encoder_3/model.fp16-00001-of-00002.safetensors filter=lfs diff=lfs merge=lfs -text
|
12 |
+
text_encoder_3/model.fp16-00002-of-00002.safetensors filter=lfs diff=lfs merge=lfs -text
|
13 |
+
transformer/diffusion_pytorch_model.fp16.safetensors filter=lfs diff=lfs merge=lfs -text
|
14 |
+
vae/diffusion_pytorch_model.fp16.safetensors filter=lfs diff=lfs merge=lfs -text
|
15 |
+
*.gguf filter=lfs diff=lfs merge=lfs -text
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
LICENSE.md
ADDED
@@ -0,0 +1,51 @@
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
1 |
+
STABILITY AI COMMUNITY LICENSE AGREEMENT
|
2 |
+
Last Updated: July 5, 2024
|
3 |
+
|
4 |
+
|
5 |
+
I. INTRODUCTION
|
6 |
+
|
7 |
+
This Agreement applies to any individual person or entity ("You", "Your" or "Licensee") that uses or distributes any portion or element of the Stability AI Materials or Derivative Works thereof for any Research & Non-Commercial or Commercial purpose. Capitalized terms not otherwise defined herein are defined in Section V below.
|
8 |
+
|
9 |
+
|
10 |
+
This Agreement is intended to allow research, non-commercial, and limited commercial uses of the Models free of charge. In order to ensure that certain limited commercial uses of the Models continue to be allowed, this Agreement preserves free access to the Models for people or organizations generating annual revenue of less than US $1,000,000 (or local currency equivalent).
|
11 |
+
|
12 |
+
|
13 |
+
By clicking "I Accept" or by using or distributing or using any portion or element of the Stability Materials or Derivative Works, You agree that You have read, understood and are bound by the terms of this Agreement. If You are acting on behalf of a company, organization or other entity, then "You" includes you and that entity, and You agree that You: (i) are an authorized representative of such entity with the authority to bind such entity to this Agreement, and (ii) You agree to the terms of this Agreement on that entity's behalf.
|
14 |
+
|
15 |
+
II. RESEARCH & NON-COMMERCIAL USE LICENSE
|
16 |
+
|
17 |
+
Subject to the terms of this Agreement, Stability AI grants You a non-exclusive, worldwide, non-transferable, non-sublicensable, revocable and royalty-free limited license under Stability AI's intellectual property or other rights owned by Stability AI embodied in the Stability AI Materials to use, reproduce, distribute, and create Derivative Works of, and make modifications to, the Stability AI Materials for any Research or Non-Commercial Purpose. "Research Purpose" means academic or scientific advancement, and in each case, is not primarily intended for commercial advantage or monetary compensation to You or others. "Non-Commercial Purpose" means any purpose other than a Research Purpose that is not primarily intended for commercial advantage or monetary compensation to You or others, such as personal use (i.e., hobbyist) or evaluation and testing.
|
18 |
+
|
19 |
+
III. COMMERCIAL USE LICENSE
|
20 |
+
|
21 |
+
Subject to the terms of this Agreement (including the remainder of this Section III), Stability AI grants You a non-exclusive, worldwide, non-transferable, non-sublicensable, revocable and royalty-free limited license under Stability AI's intellectual property or other rights owned by Stability AI embodied in the Stability AI Materials to use, reproduce, distribute, and create Derivative Works of, and make modifications to, the Stability AI Materials for any Commercial Purpose. "Commercial Purpose" means any purpose other than a Research Purpose or Non-Commercial Purpose that is primarily intended for commercial advantage or monetary compensation to You or others, including but not limited to, (i) creating, modifying, or distributing Your product or service, including via a hosted service or application programming interface, and (ii) for Your business's or organization's internal operations.
|
22 |
+
If You are using or distributing the Stability AI Materials for a Commercial Purpose, You must register with Stability AI at (https://stability.ai/community-license). If at any time You or Your Affiliate(s), either individually or in aggregate, generate more than USD $1,000,000 in annual revenue (or the equivalent thereof in Your local currency), regardless of whether that revenue is generated directly or indirectly from the Stability AI Materials or Derivative Works, any licenses granted to You under this Agreement shall terminate as of such date. You must request a license from Stability AI at (https://stability.ai/enterprise) , which Stability AI may grant to You in its sole discretion. If you receive Stability AI Materials, or any Derivative Works thereof, from a Licensee as part of an integrated end user product, then Section III of this Agreement will not apply to you.
|
23 |
+
|
24 |
+
IV. GENERAL TERMS
|
25 |
+
|
26 |
+
Your Research, Non-Commercial, and Commercial License(s) under this Agreement are subject to the following terms.
|
27 |
+
a. Distribution & Attribution. If You distribute or make available the Stability AI Materials or a Derivative Work to a third party, or a product or service that uses any portion of them, You shall: (i) provide a copy of this Agreement to that third party, (ii) retain the following attribution notice within a "Notice" text file distributed as a part of such copies: "This Stability AI Model is licensed under the Stability AI Community License, Copyright © Stability AI Ltd. All Rights Reserved", and (iii) prominently display "Powered by Stability AI" on a related website, user interface, blogpost, about page, or product documentation. If You create a Derivative Work, You may add your own attribution notice(s) to the "Notice" text file included with that Derivative Work, provided that You clearly indicate which attributions apply to the Stability AI Materials and state in the "Notice" text file that You changed the Stability AI Materials and how it was modified.
|
28 |
+
b. Use Restrictions. Your use of the Stability AI Materials and Derivative Works, including any output or results of the Stability AI Materials or Derivative Works, must comply with applicable laws and regulations (including Trade Control Laws and equivalent regulations) and adhere to the Documentation and Stability AI's AUP, which is hereby incorporated by reference. Furthermore, You will not use the Stability AI Materials or Derivative Works, or any output or results of the Stability AI Materials or Derivative Works, to create or improve any foundational generative AI model (excluding the Models or Derivative Works).
|
29 |
+
c. Intellectual Property.
|
30 |
+
(i) Trademark License. No trademark licenses are granted under this Agreement, and in connection with the Stability AI Materials or Derivative Works, You may not use any name or mark owned by or associated with Stability AI or any of its Affiliates, except as required under Section IV(a) herein.
|
31 |
+
(ii) Ownership of Derivative Works. As between You and Stability AI, You are the owner of Derivative Works You create, subject to Stability AI's ownership of the Stability AI Materials and any Derivative Works made by or for Stability AI.
|
32 |
+
(iii) Ownership of Outputs. As between You and Stability AI, You own any outputs generated from the Models or Derivative Works to the extent permitted by applicable law.
|
33 |
+
(iv) Disputes. If You or Your Affiliate(s) institute litigation or other proceedings against Stability AI (including a cross-claim or counterclaim in a lawsuit) alleging that the Stability AI Materials, Derivative Works or associated outputs or results, or any portion of any of the foregoing, constitutes infringement of intellectual property or other rights owned or licensable by You, then any licenses granted to You under this Agreement shall terminate as of the date such litigation or claim is filed or instituted. You will indemnify and hold harmless Stability AI from and against any claim by any third party arising out of or related to Your use or distribution of the Stability AI Materials or Derivative Works in violation of this Agreement.
|
34 |
+
(v) Feedback. From time to time, You may provide Stability AI with verbal and/or written suggestions, comments or other feedback related to Stability AI's existing or prospective technology, products or services (collectively, "Feedback"). You are not obligated to provide Stability AI with Feedback, but to the extent that You do, You hereby grant Stability AI a perpetual, irrevocable, royalty-free, fully-paid, sub-licensable, transferable, non-exclusive, worldwide right and license to exploit the Feedback in any manner without restriction. Your Feedback is provided "AS IS" and You make no warranties whatsoever about any Feedback.
|
35 |
+
d. Disclaimer Of Warranty. UNLESS REQUIRED BY APPLICABLE LAW, THE STABILITY AI MATERIALS AND ANY OUTPUT AND RESULTS THEREFROM ARE PROVIDED ON AN "AS IS" BASIS, WITHOUT WARRANTIES OF ANY KIND, EITHER EXPRESS OR IMPLIED, INCLUDING, WITHOUT LIMITATION, ANY WARRANTIES OF TITLE, NON-INFRINGEMENT, MERCHANTABILITY, OR FITNESS FOR A PARTICULAR PURPOSE. YOU ARE SOLELY RESPONSIBLE FOR DETERMINING THE APPROPRIATENESS OR LAWFULNESS OF USING OR REDISTRIBUTING THE STABILITY AI MATERIALS, DERIVATIVE WORKS OR ANY OUTPUT OR RESULTS AND ASSUME ANY RISKS ASSOCIATED WITH YOUR USE OF THE STABILITY AI MATERIALS, DERIVATIVE WORKS AND ANY OUTPUT AND RESULTS.
|
36 |
+
e. Limitation Of Liability. IN NO EVENT WILL STABILITY AI OR ITS AFFILIATES BE LIABLE UNDER ANY THEORY OF LIABILITY, WHETHER IN CONTRACT, TORT, NEGLIGENCE, PRODUCTS LIABILITY, OR OTHERWISE, ARISING OUT OF THIS AGREEMENT, FOR ANY LOST PROFITS OR ANY DIRECT, INDIRECT, SPECIAL, CONSEQUENTIAL, INCIDENTAL, EXEMPLARY OR PUNITIVE DAMAGES, EVEN IF STABILITY AI OR ITS AFFILIATES HAVE BEEN ADVISED OF THE POSSIBILITY OF ANY OF THE FOREGOING.
|
37 |
+
f. Term And Termination. The term of this Agreement will commence upon Your acceptance of this Agreement or access to the Stability AI Materials and will continue in full force and effect until terminated in accordance with the terms and conditions herein. Stability AI may terminate this Agreement if You are in breach of any term or condition of this Agreement. Upon termination of this Agreement, You shall delete and cease use of any Stability AI Materials or Derivative Works. Section IV(d), (e), and (g) shall survive the termination of this Agreement.
|
38 |
+
g. Governing Law. This Agreement will be governed by and constructed in accordance with the laws of the United States and the State of California without regard to choice of law principles, and the UN Convention on Contracts for International Sale of Goods does not apply to this Agreement.
|
39 |
+
|
40 |
+
V. DEFINITIONS
|
41 |
+
|
42 |
+
"Affiliate(s)" means any entity that directly or indirectly controls, is controlled by, or is under common control with the subject entity; for purposes of this definition, "control" means direct or indirect ownership or control of more than 50% of the voting interests of the subject entity.
|
43 |
+
"Agreement" means this Stability AI Community License Agreement.
|
44 |
+
"AUP" means the Stability AI Acceptable Use Policy available at https://stability.ai/use-policy, as may be updated from time to time.
|
45 |
+
"Derivative Work(s)" means (a) any derivative work of the Stability AI Materials as recognized by U.S. copyright laws and (b) any modifications to a Model, and any other model created which is based on or derived from the Model or the Model's output, including"fine tune" and "low-rank adaptation" models derived from a Model or a Model's output, but do not include the output of any Model.
|
46 |
+
"Documentation" means any specifications, manuals, documentation, and other written information provided by Stability AI related to the Software or Models.
|
47 |
+
"Model(s)" means, collectively, Stability AI's proprietary models and algorithms, including machine-learning models, trained model weights and other elements of the foregoing listed on Stability's Core Models Webpage available at, https://stability.ai/core-models, as may be updated from time to time.
|
48 |
+
"Stability AI" or "we" means Stability AI Ltd. and its Affiliates.
|
49 |
+
"Software" means Stability AI's proprietary software made available under this Agreement now or in the future.
|
50 |
+
"Stability AI Materials" means, collectively, Stability's proprietary Models, Software and Documentation (and any portion or combination thereof) made available under this Agreement.
|
51 |
+
"Trade Control Laws" means any applicable U.S. and non-U.S. export control and trade sanctions laws and regulations.
|
README.md
ADDED
@@ -0,0 +1,180 @@
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
1 |
+
---
|
2 |
+
license: other
|
3 |
+
license_name: stabilityai-ai-community
|
4 |
+
license_link: LICENSE.md
|
5 |
+
tags:
|
6 |
+
- text-to-image
|
7 |
+
- stable-diffusion
|
8 |
+
- diffusion-single-file
|
9 |
+
inference: false
|
10 |
+
extra_gated_prompt: >-
|
11 |
+
By clicking "Agree", you agree to the [License
|
12 |
+
Agreement](https://huggingface.co/stabilityai/stable-diffusion-3-medium/blob/main/LICENSE.md)
|
13 |
+
and acknowledge Stability AI's [Privacy
|
14 |
+
Policy](https://stability.ai/privacy-policy).
|
15 |
+
extra_gated_fields:
|
16 |
+
Name: text
|
17 |
+
Email: text
|
18 |
+
Country: country
|
19 |
+
Organization or Affiliation: text
|
20 |
+
Receive email updates and promotions on Stability AI products, services, and research?:
|
21 |
+
type: select
|
22 |
+
options:
|
23 |
+
- 'Yes'
|
24 |
+
- 'No'
|
25 |
+
I agree to the License Agreement and acknowledge Stability AI's Privacy Policy: checkbox
|
26 |
+
language:
|
27 |
+
- en
|
28 |
+
pipeline_tag: text-to-image
|
29 |
+
---
|
30 |
+
|
31 |
+
# stable-diffusion-v3-medium-GGUF
|
32 |
+
|
33 |
+
!!! Experimental supported by [gpustack/llama-box v0.0.75+](https://github.com/gpustack/llama-box) only !!!
|
34 |
+
|
35 |
+
**Model creator**: [Stability AI](https://huggingface.co/stabilityai)<br/>
|
36 |
+
**Original model**: [stable-diffusion-3-medium](https://huggingface.co/stabilityai/stable-diffusion-3-medium)<br/>
|
37 |
+
**GGUF quantization**: based on stable-diffusion.cpp [ac54e](https://github.com/leejet/stable-diffusion.cpp/commit/ac54e0076052a196b7df961eb1f792c9ff4d7f22) that patched by llama-box.
|
38 |
+
|
39 |
+
---
|
40 |
+
|
41 |
+
# Stable Diffusion 3 Medium
|
42 |
+
![sd3 demo images](sd3demo.jpg)
|
43 |
+
|
44 |
+
## Model
|
45 |
+
|
46 |
+
![mmdit](mmdit.png)
|
47 |
+
|
48 |
+
[Stable Diffusion 3 Medium](https://stability.ai/news/stable-diffusion-3-medium) is a Multimodal Diffusion Transformer (MMDiT) text-to-image model that features greatly improved performance in image quality, typography, complex prompt understanding, and resource-efficiency.
|
49 |
+
|
50 |
+
For more technical details, please refer to the [Research paper](https://stability.ai/news/stable-diffusion-3-research-paper).
|
51 |
+
|
52 |
+
Please note: this model is released under the Stability Community License. For Enterprise License visit Stability.ai or [contact us](https://stability.ai/enterprise) for commercial licensing details.
|
53 |
+
|
54 |
+
|
55 |
+
|
56 |
+
### Model Description
|
57 |
+
|
58 |
+
- **Developed by:** Stability AI
|
59 |
+
- **Model type:** MMDiT text-to-image generative model
|
60 |
+
- **Model Description:** This is a model that can be used to generate images based on text prompts. It is a Multimodal Diffusion Transformer
|
61 |
+
(https://arxiv.org/abs/2403.03206) that uses three fixed, pretrained text encoders
|
62 |
+
([OpenCLIP-ViT/G](https://github.com/mlfoundations/open_clip), [CLIP-ViT/L](https://github.com/openai/CLIP/tree/main) and [T5-xxl](https://huggingface.co/google/t5-v1_1-xxl))
|
63 |
+
|
64 |
+
### License
|
65 |
+
|
66 |
+
- **Community License:** Free for research, non-commercial, and commercial use for organisations or individuals with less than $1M annual revenue. You only need a paid Enterprise license if your yearly revenues exceed USD$1M and you use Stability AI models in commercial products or services. Read more: https://stability.ai/license
|
67 |
+
- **For companies above this revenue threshold**: please contact us: https://stability.ai/enterprise
|
68 |
+
|
69 |
+
|
70 |
+
### Model Sources
|
71 |
+
|
72 |
+
For local or self-hosted use, we recommend [ComfyUI](https://github.com/comfyanonymous/ComfyUI) for inference.
|
73 |
+
|
74 |
+
Stable Diffusion 3 Medium is available on our [Stability API Platform](https://platform.stability.ai/docs/api-reference#tag/Generate/paths/~1v2beta~1stable-image~1generate~1sd3/post).
|
75 |
+
|
76 |
+
Stable Diffusion 3 models and workflows are available on [Stable Assistant](https://stability.ai/stable-assistant) and on Discord via [Stable Artisan](https://stability.ai/stable-artisan).
|
77 |
+
|
78 |
+
- **ComfyUI:** https://github.com/comfyanonymous/ComfyUI
|
79 |
+
- **StableSwarmUI:** https://github.com/Stability-AI/StableSwarmUI
|
80 |
+
- **Tech report:** https://stability.ai/news/stable-diffusion-3-research-paper
|
81 |
+
- **Demo:** https://huggingface.co/spaces/stabilityai/stable-diffusion-3-medium
|
82 |
+
- **Diffusers support:** https://huggingface.co/stabilityai/stable-diffusion-3-medium-diffusers
|
83 |
+
|
84 |
+
|
85 |
+
## Training Dataset
|
86 |
+
|
87 |
+
We used synthetic data and filtered publicly available data to train our models. The model was pre-trained on 1 billion images. The fine-tuning data includes 30M high-quality aesthetic images focused on specific visual content and style, as well as 3M preference data images.
|
88 |
+
|
89 |
+
## File Structure
|
90 |
+
```
|
91 |
+
├── comfy_example_workflows/
|
92 |
+
│ ├── sd3_medium_example_workflow_basic.json
|
93 |
+
│ ├── sd3_medium_example_workflow_multi_prompt.json
|
94 |
+
│ └── sd3_medium_example_workflow_upscaling.json
|
95 |
+
│
|
96 |
+
├── text_encoders/
|
97 |
+
│ ├── README.md
|
98 |
+
│ ├── clip_g.safetensors
|
99 |
+
│ ├── clip_l.safetensors
|
100 |
+
│ ├── t5xxl_fp16.safetensors
|
101 |
+
│ └── t5xxl_fp8_e4m3fn.safetensors
|
102 |
+
│
|
103 |
+
├── LICENSE
|
104 |
+
├── sd3_medium.safetensors
|
105 |
+
├── sd3_medium_incl_clips.safetensors
|
106 |
+
├── sd3_medium_incl_clips_t5xxlfp8.safetensors
|
107 |
+
└── sd3_medium_incl_clips_t5xxlfp16.safetensors
|
108 |
+
|
109 |
+
```
|
110 |
+
|
111 |
+
We have prepared three packaging variants of the SD3 Medium model, each equipped with the same set of MMDiT & VAE weights, for user convenience.
|
112 |
+
|
113 |
+
* `sd3_medium.safetensors` includes the MMDiT and VAE weights but does not include any text encoders.
|
114 |
+
* `sd3_medium_incl_clips_t5xxlfp16.safetensors` contains all necessary weights, including fp16 version of the T5XXL text encoder.
|
115 |
+
* `sd3_medium_incl_clips_t5xxlfp8.safetensors` contains all necessary weights, including fp8 version of the T5XXL text encoder, offering a balance between quality and resource requirements.
|
116 |
+
* `sd3_medium_incl_clips.safetensors` includes all necessary weights except for the T5XXL text encoder. It requires minimal resources, but the model's performance will differ without the T5XXL text encoder.
|
117 |
+
* The `text_encoders` folder contains three text encoders and their original model card links for user convenience. All components within the text_encoders folder (and their equivalents embedded in other packings) are subject to their respective original licenses.
|
118 |
+
* The `example_workfows` folder contains example comfy workflows.
|
119 |
+
|
120 |
+
## Using with Diffusers
|
121 |
+
|
122 |
+
This repository corresponds to the original release weights. You can find the _diffusers_ compatible weights [here](https://huggingface.co/stabilityai/stable-diffusion-3-medium-diffusers). Make sure you upgrade to the latest version of diffusers: `pip install -U diffusers`. And then you can run:
|
123 |
+
|
124 |
+
```python
|
125 |
+
import torch
|
126 |
+
from diffusers import StableDiffusion3Pipeline
|
127 |
+
|
128 |
+
pipe = StableDiffusion3Pipeline.from_pretrained("stabilityai/stable-diffusion-3-medium-diffusers", torch_dtype=torch.float16)
|
129 |
+
pipe = pipe.to("cuda")
|
130 |
+
|
131 |
+
image = pipe(
|
132 |
+
"A cat holding a sign that says hello world",
|
133 |
+
negative_prompt="",
|
134 |
+
num_inference_steps=28,
|
135 |
+
guidance_scale=7.0,
|
136 |
+
).images[0]
|
137 |
+
image
|
138 |
+
```
|
139 |
+
|
140 |
+
Refer to [the documentation](https://huggingface.co/docs/diffusers/main/en/api/pipelines/stable_diffusion/stable_diffusion_3) for more details on optimization and image-to-image support.
|
141 |
+
|
142 |
+
## Uses
|
143 |
+
|
144 |
+
### Intended Uses
|
145 |
+
|
146 |
+
Intended uses include the following:
|
147 |
+
* Generation of artworks and use in design and other artistic processes.
|
148 |
+
* Applications in educational or creative tools.
|
149 |
+
* Research on generative models, including understanding the limitations of generative models.
|
150 |
+
|
151 |
+
All uses of the model should be in accordance with our [Acceptable Use Policy](https://stability.ai/use-policy).
|
152 |
+
|
153 |
+
### Out-of-Scope Uses
|
154 |
+
|
155 |
+
The model was not trained to be factual or true representations of people or events. As such, using the model to generate such content is out-of-scope of the abilities of this model.
|
156 |
+
|
157 |
+
## Safety
|
158 |
+
|
159 |
+
As part of our safety-by-design and responsible AI deployment approach, we implement safety measures throughout the development of our models, from the time we begin pre-training a model to the ongoing development, fine-tuning, and deployment of each model. We have implemented a number of safety mitigations that are intended to reduce the risk of severe harms, however we recommend that developers conduct their own testing and apply additional mitigations based on their specific use cases.
|
160 |
+
For more about our approach to Safety, please visit our [Safety page](https://stability.ai/safety).
|
161 |
+
|
162 |
+
### Evaluation Approach
|
163 |
+
|
164 |
+
Our evaluation methods include structured evaluations and internal and external red-teaming testing for specific, severe harms such as child sexual abuse and exploitation, extreme violence, and gore, sexually explicit content, and non-consensual nudity. Testing was conducted primarily in English and may not cover all possible harms. As with any model, the model may, at times, produce inaccurate, biased or objectionable responses to user prompts.
|
165 |
+
|
166 |
+
### Risks identified and mitigations:
|
167 |
+
|
168 |
+
* Harmful content: We have used filtered data sets when training our models and implemented safeguards that attempt to strike the right balance between usefulness and preventing harm. However, this does not guarantee that all possible harmful content has been removed. The model may, at times, generate toxic or biased content. All developers and deployers should exercise caution and implement content safety guardrails based on their specific product policies and application use cases.
|
169 |
+
* Misuse: Technical limitations and developer and end-user education can help mitigate against malicious applications of models. All users are required to adhere to our Acceptable Use Policy, including when applying fine-tuning and prompt engineering mechanisms. Please reference the Stability AI Acceptable Use Policy for information on violative uses of our products.
|
170 |
+
* Privacy violations: Developers and deployers are encouraged to adhere to privacy regulations with techniques that respect data privacy.
|
171 |
+
|
172 |
+
### Contact
|
173 |
+
|
174 |
+
Please report any issues with the model or contact us:
|
175 |
+
|
176 |
+
* Safety issues: [email protected]
|
177 |
+
* Security issues: [email protected]
|
178 |
+
* Privacy issues: [email protected]
|
179 |
+
* License and general: https://stability.ai/license
|
180 |
+
* Enterprise license: https://stability.ai/enterprise
|
demo_images/demo (1).png
ADDED
demo_images/demo (10).png
ADDED
demo_images/demo (2).png
ADDED
demo_images/demo (3).png
ADDED
demo_images/demo (4).png
ADDED
demo_images/demo (5).png
ADDED
demo_images/demo (6).png
ADDED
demo_images/demo (7).png
ADDED
demo_images/demo (8).png
ADDED
demo_images/demo (9).png
ADDED
mmdit.png
ADDED
sd3demo.jpg
ADDED
Git LFS Details
|
sd3demo_prompts.txt
ADDED
@@ -0,0 +1,19 @@
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
1 |
+
a female character with long, flowing hair that appears to be made of ethereal, swirling patterns resembling the Northern Lights or Aurora Borealis. The background is dominated by deep blues and purples, creating a mysterious and dramatic atmosphere. The character's face is serene, with pale skin and striking features. She wears a dark-colored outfit with subtle patterns. The overall style of the artwork is reminiscent of fantasy or supernatural genres
|
2 |
+
|
3 |
+
Digital art, portrait of an anthropomorphic roaring Tiger warrior with full armor, close up in the middle of a battle, behind him there is a banner with the text "Open Source".
|
4 |
+
|
5 |
+
photo of a dog and a cat both standing on a red box, with a blue ball in the middle with a parrot standing on top of the ball. The box has the text "SD3"
|
6 |
+
|
7 |
+
selfie photo of a wizard with long beard and purple robes, he is apparently in the middle of Tokyo. Probably taken from a phone.
|
8 |
+
|
9 |
+
A vibrant street wall covered in colorful graffiti, the centerpiece spells "SD3 MEDIUM", in a storm of colors
|
10 |
+
|
11 |
+
photo of a young woman with long, wavy brown hair tied in a bun and glasses. She has a fair complexion and is wearing subtle makeup, emphasizing her eyes and lips. She is dressed in a black top. The background appears to be an urban setting with a building facade, and the sunlight casts a warm glow on her face.
|
12 |
+
|
13 |
+
anime art of a steampunk inventor in their workshop, surrounded by gears, gadgets, and steam. He is holding a blue potion and a red potion, one in each hand
|
14 |
+
|
15 |
+
photo of picturesque scene of a road surrounded by lush green trees and shrubs. The road is wide and smooth, leading into the distance. On the right side of the road, there's a blue sports car parked with the license plate spelling "SD32B". The sky above is partly cloudy, suggesting a pleasant day. The trees have a mix of green and brown foliage. There are no people visible in the image. The overall composition is balanced, with the car serving as a focal point.
|
16 |
+
|
17 |
+
photo of young man in a black suit, white shirt, and black tie. He has a neatly styled haircut and is looking directly at the camera with a neutral expression. The background consists of a textured wall with horizontal lines. The photograph is in black and white, emphasizing contrasts and shadows. The man appears to be in his late twenties or early thirties, with fair skin and short, dark hair.
|
18 |
+
|
19 |
+
photo of a woman on the beach, shot from above. She is facing the sea, while wearing a white dress. She has long blonde hair
|
stable-diffusion-v3-medium-FP16.gguf
ADDED
@@ -0,0 +1,3 @@
|
|
|
|
|
|
|
|
|
1 |
+
version https://git-lfs.github.com/spec/v1
|
2 |
+
oid sha256:ea8afbcf72925129bff97a815572de4112ebac99b84217481e59e6ccb26939c2
|
3 |
+
size 15760972544
|
stable-diffusion-v3-medium-Q4_0.gguf
ADDED
@@ -0,0 +1,3 @@
|
|
|
|
|
|
|
|
|
1 |
+
version https://git-lfs.github.com/spec/v1
|
2 |
+
oid sha256:459b469a26c65727ce40878efbaea403d79cb94ab00d7e0f6b595e027258dce8
|
3 |
+
size 8286778496
|
stable-diffusion-v3-medium-Q8_0.gguf
ADDED
@@ -0,0 +1,3 @@
|
|
|
|
|
|
|
|
|
1 |
+
version https://git-lfs.github.com/spec/v1
|
2 |
+
oid sha256:93f106b587c66871d8c4e4cc18df8c86bc6f72c2057f63c7ee269aa5c3e0a760
|
3 |
+
size 9290658944
|