adamo1139 commited on
Commit
fd53109
β€’
1 Parent(s): 74b6131

Upload 5 files

Browse files
Files changed (6) hide show
  1. .gitattributes +1 -0
  2. LICENSE +29 -0
  3. README.md +143 -0
  4. mmdit.png +0 -0
  5. sd3demo.jpg +3 -0
  6. sd3demo_prompts.txt +19 -0
.gitattributes CHANGED
@@ -33,3 +33,4 @@ saved_model/**/* filter=lfs diff=lfs merge=lfs -text
33
  *.zip filter=lfs diff=lfs merge=lfs -text
34
  *.zst filter=lfs diff=lfs merge=lfs -text
35
  *tfevents* filter=lfs diff=lfs merge=lfs -text
 
 
33
  *.zip filter=lfs diff=lfs merge=lfs -text
34
  *.zst filter=lfs diff=lfs merge=lfs -text
35
  *tfevents* filter=lfs diff=lfs merge=lfs -text
36
+ sd3demo.jpg filter=lfs diff=lfs merge=lfs -text
LICENSE ADDED
@@ -0,0 +1,29 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ STABILITY AI NON-COMMERCIAL RESEARCH COMMUNITY LICENSE AGREEMENT
2
+ Dated: December 06, 2023
3
+
4
+ By using or distributing any portion or element of the Models, Software, Software Products or Derivative Works, you agree to be bound by this Agreement.
5
+
6
+ "Agreement" means this Stable Non-Commercial Research Community License Agreement.
7
+
8
+ β€œAUP” means the Stability AI Acceptable Use Policy available at https://stability.ai/use-policy, as may be updated from time to time.
9
+ "Derivative Work(s)” means (a) any derivative work of the Software Products as recognized by U.S. copyright laws and (b) any modifications to a Model, and any other model created which is based on or derived from the Model or the Model’s output. For clarity, Derivative Works do not include the output of any Model.
10
+ β€œDocumentation” means any specifications, manuals, documentation, and other written information provided by Stability AI related to the Software.
11
+ "Licensee" or "you" means you, or your employer or any other person or entity (if you are entering into this Agreement on such person or entity's behalf), of the age required under applicable laws, rules or regulations to provide legal consent and that has legal authority to bind your employer or such other person or entity if you are entering in this Agreement on their behalf.
12
+ β€œModel(s)" means, collectively, Stability AI’s proprietary models and algorithms, including machine-learning models, trained model weights and other elements of the foregoing, made available under this Agreement.
13
+ β€œNon-Commercial Uses” means exercising any of the rights granted herein for the purpose of research or non-commercial purposes. Non-Commercial Uses does not include any production use of the Software Products or any Derivative Works.
14
+ "Stability AI" or "we" means Stability AI Ltd. and its affiliates.
15
+ "Software" means Stability AI’s proprietary software made available under this Agreement.
16
+ β€œSoftware Products” means the Models, Software and Documentation, individually or in any combination.
17
+ 1. License Rights and Redistribution.
18
+ a. Subject to your compliance with this Agreement, the AUP (which is hereby incorporated herein by reference), and the Documentation, Stability AI grants you a non-exclusive, worldwide, non-transferable, non-sublicensable, revocable, royalty free and limited license under Stability AI’s intellectual property or other rights owned or controlled by Stability AI embodied in the Software Products to use, reproduce, distribute, and create Derivative Works of, the Software Products, in each case for Non-Commercial Uses only.
19
+ b. You may not use the Software Products or Derivative Works to enable third parties to use the Software Products or Derivative Works as part of your hosted service or via your APIs, whether you are adding substantial additional functionality thereto or not. Merely distributing the Software Products or Derivative Works for download online without offering any related service (ex. by distributing the Models on HuggingFace) is not a violation of this subsection. If you wish to use the Software Products or any Derivative Works for commercial or production use or you wish to make the Software Products or any Derivative Works available to third parties via your hosted service or your APIs, contact Stability AI at https://stability.ai/contact.
20
+ c. If you distribute or make the Software Products, or any Derivative Works thereof, available to a third party, the Software Products, Derivative Works, or any portion thereof, respectively, will remain subject to this Agreement and you must (i) provide a copy of this Agreement to such third party, and (ii) retain the following attribution notice within a "Notice" text file distributed as a part of such copies: "This Stability AI Model is licensed under the Stability AI Non-Commercial Research Community License, Copyright (c) Stability AI Ltd. All Rights Reserved.” If you create a Derivative Work of a Software Product, you may add your own attribution notices to the Notice file included with the Software Product, provided that you clearly indicate which attributions apply to the Software Product and you must state in the NOTICE file that you changed the Software Product and how it was modified.
21
+ 2. Disclaimer of Warranty. UNLESS REQUIRED BY APPLICABLE LAW, THE SOFTWARE PRODUCTS AND ANY OUTPUT AND RESULTS THEREFROM ARE PROVIDED ON AN "AS IS" BASIS, WITHOUT WARRANTIES OF ANY KIND, EITHER EXPRESS OR IMPLIED, INCLUDING, WITHOUT LIMITATION, ANY WARRANTIES OF TITLE, NON-INFRINGEMENT, MERCHANTABILITY, OR FITNESS FOR A PARTICULAR PURPOSE. YOU ARE SOLELY RESPONSIBLE FOR DETERMINING THE APPROPRIATENESS OF USING OR REDISTRIBUTING THE SOFTWARE PRODUCTS, DERIVATIVE WORKS OR ANY OUTPUT OR RESULTS AND ASSUME ANY RISKS ASSOCIATED WITH YOUR USE OF THE SOFTWARE PRODUCTS, DERIVATIVE WORKS AND ANY OUTPUT AND RESULTS.
22
+ 3. Limitation of Liability. IN NO EVENT WILL STABILITY AI OR ITS AFFILIATES BE LIABLE UNDER ANY THEORY OF LIABILITY, WHETHER IN CONTRACT, TORT, NEGLIGENCE, PRODUCTS LIABILITY, OR OTHERWISE, ARISING OUT OF THIS AGREEMENT, FOR ANY LOST PROFITS OR ANY DIRECT, INDIRECT, SPECIAL, CONSEQUENTIAL, INCIDENTAL, EXEMPLARY OR PUNITIVE DAMAGES, EVEN IF STABILITY AI OR ITS AFFILIATES HAVE BEEN ADVISED OF THE POSSIBILITY OF ANY OF THE FOREGOING.
23
+ 4. Intellectual Property.
24
+ a. No trademark licenses are granted under this Agreement, and in connection with the Software Products or Derivative Works, neither Stability AI nor Licensee may use any name or mark owned by or associated with the other or any of its affiliates, except as required for reasonable and customary use in describing and redistributing the Software Products or Derivative Works.
25
+ b. Subject to Stability AI’s ownership of the Software Products and Derivative Works made by or for Stability AI, with respect to any Derivative Works that are made by you, as between you and Stability AI, you are and will be the owner of such Derivative Works
26
+ c. If you institute litigation or other proceedings against Stability AI (including a cross-claim or counterclaim in a lawsuit) alleging that the Software Products, Derivative Works or associated outputs or results, or any portion of any of the foregoing, constitutes infringement of intellectual property or other rights owned or licensable by you, then any licenses granted to you under this Agreement shall terminate as of the date such litigation or claim is filed or instituted. You will indemnify and hold harmless Stability AI from and against any claim by any third party arising out of or related to your use or distribution of the Software Products or Derivative Works in violation of this Agreement.
27
+ 5. Term and Termination. The term of this Agreement will commence upon your acceptance of this Agreement or access to the Software Products and will continue in full force and effect until terminated in accordance with the terms and conditions herein. Stability AI may terminate this Agreement if you are in breach of any term or condition of this Agreement. Upon termination of this Agreement, you shall delete and cease use of any Software Products or Derivative Works. Sections 2-4 shall survive the termination of this Agreement.
28
+ 6. Governing Law. This Agreement will be governed by and construed in accordance with the laws of the United States and the State of California without regard to choice of law
29
+ principles.
README.md ADDED
@@ -0,0 +1,143 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ license: other
3
+ license_name: stabilityai-nc-research-community
4
+ license_link: LICENSE
5
+ tags:
6
+ - text-to-image
7
+ - stable-diffusion
8
+ extra_gated_prompt: >-
9
+ By clicking "Agree", you agree to the [License
10
+ Agreement](https://huggingface.co/stabilityai/stable-diffusion-3-medium/blob/main/LICENSE)
11
+ and acknowledge Stability AI's [Privacy
12
+ Policy](https://stability.ai/privacy-policy).
13
+ extra_gated_fields:
14
+ Name: text
15
+ Email: text
16
+ Country: country
17
+ Organization or Affiliation: text
18
+ Receive email updates and promotions on Stability AI products, services, and research?:
19
+ type: select
20
+ options:
21
+ - 'Yes'
22
+ - 'No'
23
+ I acknowledge that this model is for non-commercial use only unless I acquire a separate license from Stability AI: checkbox
24
+ language:
25
+ - en
26
+ pipeline_tag: text-to-image
27
+ ---
28
+ # Stable Diffusion 3 Medium
29
+ ![sd3 demo images](sd3demo.jpg)
30
+
31
+ ## Model
32
+
33
+ ![mmdit](mmdit.png)
34
+
35
+ [Stable Diffusion 3 Medium](stability.ai/news/stable-diffusion-3-medium) is a Multimodal Diffusion Transformer (MMDiT) text-to-image model that features greatly improved performance in image quality, typography, complex prompt understanding, and resource-efficiency.
36
+
37
+ For more technical details, please refer to the [Research paper](https://stability.ai/news/stable-diffusion-3-research-paper).
38
+
39
+ Please note: this model is released under the Stability Non-Commercial Research Community License. For a Creator License or an Enterprise License visit Stability.ai or [contact us](https://stability.ai/license) for commercial licensing details.
40
+
41
+
42
+
43
+ ### Model Description
44
+
45
+ - **Developed by:** Stability AI
46
+ - **Model type:** MMDiT text-to-image generative model
47
+ - **Model Description:** This is a model that can be used to generate images based on text prompts. It is a Multimodal Diffusion Transformer
48
+ (https://arxiv.org/abs/2403.03206) that uses three fixed, pretrained text encoders
49
+ ([OpenCLIP-ViT/G](https://github.com/mlfoundations/open_clip), [CLIP-ViT/L](https://github.com/openai/CLIP/tree/main) and [T5-xxl](https://huggingface.co/google/t5-v1_1-xxl))
50
+
51
+ ### License
52
+
53
+ - **Non-commercial Use:** Stable Diffusion 3 Medium is released under the [Stability AI Non-Commercial Research Community License](https://huggingface.co/stabilityai/stable-diffusion-3-medium/blob/main/LICENSE.md). The model is free to use for non-commercial purposes such as academic research.
54
+ - **Commercial Use**: This model is not available for commercial use without a separate commercial license from Stability. We encourage professional artists, designers, and creators to use our Creator License. Please visit https://stability.ai/license to learn more.
55
+
56
+
57
+ ### Model Sources
58
+
59
+ For local or self-hosted use, we recommend [ComfyUI](https://github.com/comfyanonymous/ComfyUI) for inference.
60
+
61
+ Stable Diffusion 3 Medium is available on our [Stability API Platform](https://platform.stability.ai/docs/api-reference#tag/Generate/paths/~1v2beta~1stable-image~1generate~1sd3/post).
62
+
63
+ Stable Diffusion 3 models and workflows are available on [Stable Assistant](https://stability.ai/stable-assistant) and on Discord via [Stable Artisan](https://stability.ai/stable-artisan).
64
+
65
+ - **ComfyUI:** https://github.com/comfyanonymous/ComfyUI
66
+ - **StableSwarmUI:** https://github.com/Stability-AI/StableSwarmUI
67
+ - **Tech report:** https://stability.ai/news/stable-diffusion-3-research-paper
68
+ - **Demo:** Huggingface Space is coming soon...
69
+
70
+
71
+ ## Training Dataset
72
+
73
+ We used synthetic data and filtered publicly available data to train our models. The model was pre-trained on 1 billion images. The fine-tuning data includes 30M high-quality aesthetic images focused on specific visual content and style, as well as 3M preference data images.
74
+
75
+ ## File Structure
76
+ ```
77
+ β”œβ”€β”€ comfy_example_workflows/
78
+ β”‚ β”œβ”€β”€ sd3_medium_example_workflow_basic.json
79
+ β”‚ β”œβ”€β”€ sd3_medium_example_workflow_multi_prompt.json
80
+ β”‚ └── sd3_medium_example_workflow_upscaling.json
81
+ β”‚
82
+ β”œβ”€β”€ text_encoders/
83
+ β”‚ β”œβ”€β”€ README.md
84
+ β”‚ β”œβ”€β”€ clip_g.safetensors
85
+ β”‚ β”œβ”€β”€ clip_l.safetensors
86
+ β”‚ β”œβ”€β”€ t5xxl_fp16.safetensors
87
+ β”‚ └── t5xxl_fp8_e4m3fn.safetensors
88
+ β”‚
89
+ β”œβ”€β”€ LICENSE
90
+ β”œβ”€β”€ sd3_medium.safetensors
91
+ β”œβ”€β”€ sd3_medium_incl_clips.safetensors
92
+ β”œβ”€β”€ sd3_medium_incl_clips_t5xxlfp8.safetensors
93
+ └── ...
94
+
95
+ ```
96
+
97
+ We have prepared three packaging variants of the SD3 Medium model, each equipped with the same set of MMDiT & VAE weights, for user convenience.
98
+
99
+ * `sd3_medium.safetensors` includes the MMDiT and VAE weights but does not include any text encoders.
100
+ * `sd3_medium_incl_clips_t5xxlfp8.safetensors` contains all necessary weights, including fp8 version of the T5XXL text encoder, offering a balance between quality and resource requirements.
101
+ * `sd3_medium_incl_clips.safetensors` includes all necessary weights except for the T5XXL text encoder. It requires minimal resources, but the model's performance will differ without the T5XXL text encoder.
102
+ * The `text_encoders` folder contains three text encoders and their original model card links for user convenience. All components within the text_encoders folder (and their equivalents embedded in other packings) are subject to their respective original licenses.
103
+ * The `example_workfows` folder contains example comfy workflows.
104
+
105
+ ## Uses
106
+
107
+ ### Intended Uses
108
+
109
+ Intended uses include the following:
110
+ * Generation of artworks and use in design and other artistic processes.
111
+ * Applications in educational or creative tools.
112
+ * Research on generative models, including understanding the limitations of generative models.
113
+
114
+ All uses of the model should be in accordance with our [Acceptable Use Policy](https://stability.ai/use-policy).
115
+
116
+ ### Out-of-Scope Uses
117
+
118
+ The model was not trained to be factual or true representations of people or events. As such, using the model to generate such content is out-of-scope of the abilities of this model.
119
+
120
+ ## Safety
121
+
122
+ As part of our safety-by-design and responsible AI deployment approach, we implement safety measures throughout the development of our models, from the time we begin pre-training a model to the ongoing development, fine-tuning, and deployment of each model. We have implemented a number of safety mitigations that are intended to reduce the risk of severe harms, however we recommend that developers conduct their own testing and apply additional mitigations based on their specific use cases.
123
+ For more about our approach to Safety, please visit our [Safety page](https://stability.ai/safety).
124
+
125
+ ### Evaluation Approach
126
+
127
+ Our evaluation methods include structured evaluations and internal and external red-teaming testing for specific, severe harms such as child sexual abuse and exploitation, extreme violence, and gore, sexually explicit content, and non-consensual nudity. Testing was conducted primarily in English and may not cover all possible harms. As with any model, the model may, at times, produce inaccurate, biased or objectionable responses to user prompts.
128
+
129
+ ### Risks identified and mitigations:
130
+
131
+ * Harmful content: We have used filtered data sets when training our models and implemented safeguards that attempt to strike the right balance between usefulness and preventing harm. However, this does not guarantee that all possible harmful content has been removed. The model may, at times, generate toxic or biased content. All developers and deployers should exercise caution and implement content safety guardrails based on their specific product policies and application use cases.
132
+ * Misuse: Technical limitations and developer and end-user education can help mitigate against malicious applications of models. All users are required to adhere to our Acceptable Use Policy, including when applying fine-tuning and prompt engineering mechanisms. Please reference the Stability AI Acceptable Use Policy for information on violative uses of our products.
133
+ * Privacy violations: Developers and deployers are encouraged to adhere to privacy regulations with techniques that respect data privacy.
134
+
135
+ ### Contact
136
+
137
+ Please report any issues with the model or contact us:
138
+
139
+ * Safety issues: safety@stability.ai
140
+ * Security issues: security@stability.ai
141
+ * Privacy issues: privacy@stability.ai
142
+ * License and general: https://stability.ai/license
143
+ * Enterprise license: https://stability.ai/enterprise
mmdit.png ADDED
sd3demo.jpg ADDED

Git LFS Details

  • SHA256: c6d829561bdf08624c0031f6147e771b0025a441d6616ee924a2af2ca7ea7f34
  • Pointer size: 132 Bytes
  • Size of remote file: 7.4 MB
sd3demo_prompts.txt ADDED
@@ -0,0 +1,19 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ a female character with long, flowing hair that appears to be made of ethereal, swirling patterns resembling the Northern Lights or Aurora Borealis. The background is dominated by deep blues and purples, creating a mysterious and dramatic atmosphere. The character's face is serene, with pale skin and striking features. She wears a dark-colored outfit with subtle patterns. The overall style of the artwork is reminiscent of fantasy or supernatural genres
2
+
3
+ Digital art, portrait of an anthropomorphic roaring Tiger warrior with full armor, close up in the middle of a battle, behind him there is a banner with the text "Open Source".
4
+
5
+ photo of a dog and a cat both standing on a red box, with a blue ball in the middle with a parrot standing on top of the ball. The box has the text "SD3"
6
+
7
+ selfie photo of a wizard with long beard and purple robes, he is apparently in the middle of Tokyo. Probably taken from a phone.
8
+
9
+ A vibrant street wall covered in colorful graffiti, the centerpiece spells "SD3 MEDIUM", in a storm of colors
10
+
11
+ photo of a young woman with long, wavy brown hair tied in a bun and glasses. She has a fair complexion and is wearing subtle makeup, emphasizing her eyes and lips. She is dressed in a black top. The background appears to be an urban setting with a building facade, and the sunlight casts a warm glow on her face.
12
+
13
+ anime art of a steampunk inventor in their workshop, surrounded by gears, gadgets, and steam. He is holding a blue potion and a red potion, one in each hand
14
+
15
+ photo of picturesque scene of a road surrounded by lush green trees and shrubs. The road is wide and smooth, leading into the distance. On the right side of the road, there's a blue sports car parked with the license plate spelling "SD32B". The sky above is partly cloudy, suggesting a pleasant day. The trees have a mix of green and brown foliage. There are no people visible in the image. The overall composition is balanced, with the car serving as a focal point.
16
+
17
+ photo of young man in a black suit, white shirt, and black tie. He has a neatly styled haircut and is looking directly at the camera with a neutral expression. The background consists of a textured wall with horizontal lines. The photograph is in black and white, emphasizing contrasts and shadows. The man appears to be in his late twenties or early thirties, with fair skin and short, dark hair.
18
+
19
+ photo of a woman on the beach, shot from above. She is facing the sea, while wearing a white dress. She has long blonde hair