|
{ |
|
"id": 293564, |
|
"modelId": 260267, |
|
"name": "v3.0", |
|
"createdAt": "2024-01-10T13:32:12.113Z", |
|
"updatedAt": "2024-03-21T04:45:21.909Z", |
|
"status": "Published", |
|
"publishedAt": "2024-01-10T13:57:46.063Z", |
|
"trainedWords": [], |
|
"trainingStatus": null, |
|
"trainingDetails": null, |
|
"baseModel": "SDXL 1.0", |
|
"baseModelType": "Standard", |
|
"earlyAccessTimeFrame": 0, |
|
"description": "<!--\nANIMAGINE XL 3.0\n\nHuggingface link: https://huggingface.co/cagliostrolab/animagine-xl-3.0\n\nGradio Demo : https://huggingface.co/spaces/Linaqruf/animagine-xl\n\nOfficial Blog Release: https://cagliostrolab.net/posts/animagine-xl-v3-release\n\nSupport us at: https://ko-fi.com/linaqruf\n\nOverview\n\nAnimagine XL 3.0 is the latest version of the sophisticated open-source anime text-to-image model, building upon the capabilities of its predecessor, Animagine XL 2.0. Developed based on Stable Diffusion XL, this iteration boasts superior image generation with notable improvements in hand anatomy, efficient tag ordering, and enhanced knowledge about anime concepts. Unlike the previous iteration, we focused to make the model learn concepts rather than aesthetic.\n\nAnimagine XL 3.0 was trained on a 2x A100 GPU with 80GB memory for 21 days or over 500 gpu hours. For further information please visit our official blog or huggingface repository.\n\nModel Details\n\n* Developed by: Cagliostro Research Lab\n\n* Model type: Diffusion-based text-to-image generative model\n\n* Model Description: Animagine XL 3.0 is engineered to generate high-quality anime images from textual prompts. It features enhanced hand anatomy, better concept understanding, and prompt interpretation, making it the most advanced model in its series.\n\n* License: Fair AI Public License 1.0-SD\n\n* Finetuned from model: Animagine XL 2.0\n-->", |
|
"stats": { |
|
"downloadCount": 77349, |
|
"ratingCount": 1267, |
|
"rating": 4.97, |
|
"thumbsUpCount": 2375 |
|
}, |
|
"model": { |
|
"name": "Animagine XL V3.1", |
|
"type": "Checkpoint", |
|
"nsfw": false, |
|
"poi": false, |
|
"description": "<!--\nAnimagine XL 3.1 is an update in the Animagine XL V3 series, enhancing the previous version, Animagine XL 3.0. This open-source, anime-themed text-to-image model has been improved for generating anime-style images with higher quality. It includes a broader range of characters from well-known anime series, an optimized dataset, and new aesthetic tags for better image creation. Built on Stable Diffusion XL, Animagine XL 3.1 aims to be a valuable resource for anime fans, artists, and content creators by producing accurate and detailed representations of anime characters.\n\nModel Details* Developed by: Cagliostro Research Lab\n\n* In collaboration with: SeaArt.ai\n\n* Model type: Diffusion-based text-to-image generative model\n\n* Model Description: Animagine XL 3.1 generates high-quality anime images from textual prompts. It boasts enhanced hand anatomy, improved concept understanding, and advanced prompt interpretation.\n\n* License: Fair AI Public License 1.0-SD\n\n* Fine-tuned from: Animagine XL 3.0\n\n\n\nUsage GuidelinesTag OrderingFor optimal results, it's recommended to follow the structured prompt template because we train the model like this:\n\n`1girl/1boy, character name, from what series, everything else in any order.\n`Special TagsAnimagine XL 3.1 utilizes special tags to steer the result toward quality, rating, creation date and aesthetic. While the model can generate images without these tags, using them can help achieve better results.\n\nQuality ModifiersQuality tags now consider both scores and post ratings to ensure a balanced quality distribution. We've refined labels for greater clarity, such as changing 'high quality' to 'great quality'.\n\n`\nQuality Modifier\tScore Criterion\nmasterpiece\t > 95%\nbest quality\t > 85% & \u2264 95%\ngreat quality\t > 75% & \u2264 85%\ngood quality\t > 50% & \u2264 75%\nnormal quality\t > 25% & \u2264 50%\nlow quality\t > 10% & \u2264 25%\nworst quality\t \u2264 10%`Rating ModifiersWe've also streamlined our rating tags for simplicity and clarity, aiming to establish global rules that can be applied across different models. For example, the tag 'rating: general' is now simply 'general', and 'rating: sensitive' has been condensed to 'sensitive'.\n\n`\nRating Modifier\t Rating Criterion\nsafe\t General\nsensitive\t Sensitive\nnsfw\t Questionable\nexplicit, nsfw\t Explicit`Year ModifierWe've also redefined the year range to steer results towards specific modern or vintage anime art styles more accurately. This update simplifies the range, focusing on relevance to current and past eras.\n\n`\nYear Tag\tYear Range\nnewest\t 2021 to 2024\nrecent\t 2018 to 2020\nmid\t 2015 to 2017\nearly\t 2011 to 2014\noldest\t 2005 to 2010`Aesthetic TagsWe've enhanced our tagging system with aesthetic tags to refine content categorization based on visual appeal. These tags are derived from evaluations made by a specialized ViT (Vision Transformer) image classification model, specifically trained on anime data. For this purpose, we utilized the model shadowlilac/aesthetic-shadow-v2, which assesses the aesthetic value of content before it undergoes training. This ensures that each piece of content is not only relevant and accurate but also visually appealing.\n\n`\nAesthetic Tag\t Score Range\nvery aesthetic\t > 0.71\naesthetic\t > 0.45 & < 0.71\ndispleasing\t > 0.27 & < 0.45\nvery displeasing \u2264 0.27`Recommended settingsTo guide the model towards generating high-aesthetic images, use negative prompts like:\n\n`nsfw, lowres, (bad), text, error, fewer, extra, missing, worst quality, jpeg artifacts, low quality, watermark, unfinished, displeasing, oldest, early, chromatic aberration, signature, extra digits, artistic error, username, scan, [abstract]\n`For higher quality outcomes, prepend prompts with:\n\n`masterpiece, best quality, very aesthetic, absurdres\n`it\u2019s recommended to use a lower classifier-free guidance (CFG Scale) of around 5-7, sampling steps below 30, and to use Euler Ancestral (Euler a) as a sampler.\n\nMulti Aspect ResolutionThis model supports generating images at the following dimensions:\n\n`Dimensions\tAspect Ratio\n1024 x 1024\t1:1 Square\n1152 x 896\t9:7\n896 x 1152\t7:9\n1216 x 832\t19:13\n832 x 1216\t13:19\n1344 x 768\t7:4 Horizontal\n768 x 1344\t4:7 Vertical\n1536 x 640\t12:5 Horizontal\n640 x 1536\t5:12 Vertical`AcknowledgementsThe development and release of Animagine XL 3.1 would not have been possible without the invaluable contributions and support from the following individuals and organizations:\n\n* SeaArt.ai: Our collaboration partner and sponsor.\n\n* Shadow Lilac: For providing the aesthetic classification model, aesthetic-shadow-v2.\n\n* Derrian Distro: For their custom learning rate scheduler, adapted from LoRA Easy Training Scripts.\n\n* Kohya SS: For their comprehensive training scripts.\n\n* Cagliostrolab Collaborators: For their dedication to model training, project management, and data curation.\n\n* Early Testers: For their valuable feedback and quality assurance efforts.\n\n* NovelAI: For their innovative approach to aesthetic tagging, which served as an inspiration for our implementation.\n\nThank you all for your support and expertise in pushing the boundaries of anime-style image generation.\n\n\n\nLimitationsWhile Animagine XL 3.1 represents a significant advancement in anime-style image generation, it is important to acknowledge its limitations:\n\n* Anime-Focused: This model is specifically designed for generating anime-style images and is not suitable for creating realistic photos.\n\n* Prompt Complexity: This model may not be suitable for users who expect high-quality results from short or simple prompts. The training focus was on concept understanding rather than aesthetic refinement, which may require more detailed and specific prompts to achieve the desired output.\n\n* Prompt Format: Animagine XL 3.1 is optimized for Danbooru-style tags rather than natural language prompts. For best results, users are encouraged to format their prompts using the appropriate tags and syntax.\n\n* Anatomy and Hand Rendering: Despite the improvements made in anatomy and hand rendering, there may still be instances where the model produces suboptimal results in these areas.\n\n* Dataset Size: The dataset used for training Animagine XL 3.1 consists of approximately 870,000 images. When combined with the previous iteration's dataset (1.2 million), the total training data amounts to around 2.1 million images. While substantial, this dataset size may still be considered limited in scope for an \"ultimate\" anime model.\n\n* NSFW Content: Animagine XL 3.1 has been designed to generate more balanced NSFW content. However, it is important to note that the model may still produce NSFW results, even if not explicitly prompted.\n\nBy acknowledging these limitations, we aim to provide transparency and set realistic expectations for users of Animagine XL 3.1. Despite these constraints, we believe that the model represents a significant step forward in anime-style image generation and offers a powerful tool for artists, designers, and enthusiasts alike.\n\nLicenseBased on Animagine XL 3.0, Animagine XL 3.1 falls under Fair AI Public License 1.0-SD license, which is compatible with Stable Diffusion models\u2019 license. Key points:\n\n* Modification Sharing: If you modify Animagine XL 3.1, you must share both your changes and the original license.\n\n* Source Code Accessibility: If your modified version is network-accessible, provide a way (like a download link) for others to get the source code. This applies to derived models too.\n\n* Distribution Terms: Any distribution must be under this license or another with similar rules.\n\n* Compliance: Non-compliance must be fixed within 30 days to avoid license termination, emphasizing transparency and adherence to open-source values.\n\nThe choice of this license aims to keep Animagine XL 3.1 open and modifiable, aligning with open source community spirit. It protects contributors and users, encouraging a collaborative, ethical open-source community. This ensures the model not only benefits from communal input but also respects open-source development freedoms.\n\nFinally Cagliostro Lab Server open to public https://discord.gg/cqh9tZgbGc\n\nFeel free to join our discord server.\nIf you want to donate or buy us a coffee you can donate Here\n\nThank you very much ^_^\n-->", |
|
"tags": [ |
|
"anime", |
|
"base model" |
|
], |
|
"allowNoCredit": true, |
|
"allowCommercialUse": [ |
|
"Image", |
|
"RentCivit", |
|
"Rent" |
|
], |
|
"allowDerivatives": true, |
|
"allowDifferentLicense": false |
|
}, |
|
"files": [ |
|
{ |
|
"id": 231047, |
|
"sizeKB": 6775604.111328125, |
|
"name": "animagineXLV31_v30.safetensors", |
|
"type": "Model", |
|
"pickleScanResult": "Success", |
|
"pickleScanMessage": "No Pickle imports", |
|
"virusScanResult": "Success", |
|
"virusScanMessage": null, |
|
"scannedAt": "2024-01-10T13:52:59.709Z", |
|
"metadata": { |
|
"format": "SafeTensor", |
|
"size": "full", |
|
"fp": "fp32" |
|
}, |
|
"hashes": { |
|
"AutoV1": "75F2F05B", |
|
"AutoV2": "1449E5B0B9", |
|
"SHA256": "1449E5B0B9DE87B0F414C5F29CB11CE3B3DC61FA2B320E784C9441720BF7B766", |
|
"CRC32": "3283A76F", |
|
"BLAKE3": "BA8B3384D287A2AE24E59D05898F92A39CCD56C28CD5560F854CF0900FBF8948", |
|
"AutoV3": "4C4359A0C94D" |
|
}, |
|
"primary": true, |
|
"downloadUrl": "https://civitai.com/api/download/models/293564" |
|
} |
|
], |
|
"images": [ |
|
{ |
|
"url": "https://image.civitai.com/xG1nkqKTMzGDvpLrqFT7WA/a6e71784-afec-4a83-876f-1df167381e44/width=450/5352351.jpeg", |
|
"nsfwLevel": 1, |
|
"width": 1344, |
|
"height": 1728, |
|
"hash": "U8D+[20KD*s;~p%MkW9F=|RjRkx]_4jE00R+", |
|
"type": "image", |
|
"metadata": { |
|
"hash": "U8D+[20KD*s;~p%MkW9F=|RjRkx]_4jE00R+", |
|
"size": 4412704, |
|
"width": 1344, |
|
"height": 1728 |
|
}, |
|
"availability": "Public", |
|
"meta": { |
|
"steps": 28, |
|
"prompt": "1girl, amiya \\(arknights\\), arknights, dirty face, outstretched hand, close-up, cinematic angle, foreshortening, dark, dark background, masterpiece, best quality", |
|
"sampler": "Euler a", |
|
"cfgScale": 7, |
|
"negativePrompt": "nsfw, lowres, bad anatomy, bad hands, text, error, missing fingers, extra digit, fewer digits, cropped, worst quality, low quality, normal quality, jpeg artifacts, signature, watermark, username, blurry, artist name," |
|
} |
|
}, |
|
{ |
|
"url": "https://image.civitai.com/xG1nkqKTMzGDvpLrqFT7WA/76e8f93e-678a-4976-8403-5a08bcd39034/width=450/5352345.jpeg", |
|
"nsfwLevel": 1, |
|
"width": 1344, |
|
"height": 1728, |
|
"hash": "UAH-ri009x^*pbSg=|D*~pIosCWX4nIo%LRQ", |
|
"type": "image", |
|
"metadata": { |
|
"hash": "UAH-ri009x^*pbSg=|D*~pIosCWX4nIo%LRQ", |
|
"size": 4841746, |
|
"width": 1344, |
|
"height": 1728 |
|
}, |
|
"availability": "Public", |
|
"meta": { |
|
"steps": 28, |
|
"prompt": "1girl, c.c., code geass, white shirt, long sleeves, turtleneck, sitting, looking at viewer, eating, pizza, plate, fork, knife, table, chair, table, restaurant, cinematic angle, cinematic lighting, masterpiece, best quality", |
|
"sampler": "Euler a", |
|
"cfgScale": 7, |
|
"negativePrompt": "nsfw, lowres, bad anatomy, bad hands, text, error, missing fingers, extra digit, fewer digits, cropped, worst quality, low quality, normal quality, jpeg artifacts, signature, watermark, username, blurry, artist name," |
|
} |
|
}, |
|
{ |
|
"url": "https://image.civitai.com/xG1nkqKTMzGDvpLrqFT7WA/4153b9e0-64b3-48af-9eb2-d935bb3ea7db/width=450/5352352.jpeg", |
|
"nsfwLevel": 1, |
|
"width": 1728, |
|
"height": 1344, |
|
"hash": "U9HA;TrXO@k=#Tw|0LNZ00tRa0In_MNG?H%M", |
|
"type": "image", |
|
"metadata": { |
|
"hash": "U9HA;TrXO@k=#Tw|0LNZ00tRa0In_MNG?H%M", |
|
"size": 4880706, |
|
"width": 1728, |
|
"height": 1344 |
|
}, |
|
"availability": "Public", |
|
"meta": { |
|
"seed": 1966975782, |
|
"steps": 28, |
|
"prompt": "1girl, arima kana, oshi no ko, solo, idol, idol clothes, one eye closed, red shirt, black skirt, black headwear, gloves, stage light, singing, open mouth, crowd, smile, pointing at viewer, masterpiece, best quality", |
|
"sampler": "Euler a", |
|
"cfgScale": 7, |
|
"negativePrompt": "nsfw, lowres, bad anatomy, bad hands, text, error, missing fingers, extra digit, fewer digits, cropped, worst quality, low quality, normal quality, jpeg artifacts, signature, watermark, username, blurry, artist name, " |
|
} |
|
}, |
|
{ |
|
"url": "https://image.civitai.com/xG1nkqKTMzGDvpLrqFT7WA/dcfa3f5a-72f0-45bd-85d5-84824c1940bd/width=450/5352348.jpeg", |
|
"nsfwLevel": 1, |
|
"width": 1728, |
|
"height": 1344, |
|
"hash": "UMI4q*xrJW^+}r-;IoVrD%M_xZM_^*xt%1-p", |
|
"type": "image", |
|
"metadata": { |
|
"hash": "UMI4q*xrJW^+}r-;IoVrD%M_xZM_^*xt%1-p", |
|
"size": 4527845, |
|
"width": 1728, |
|
"height": 1344 |
|
}, |
|
"availability": "Public", |
|
"meta": { |
|
"seed": 2097324220, |
|
"steps": 28, |
|
"prompt": "1girl, higuchi madoka, idolmaster shiny colors, game cg, solo, looking viewer, outdoors, night, snow, blush, breath, cinematic angle, masterpiece, best quality", |
|
"sampler": "Euler a", |
|
"cfgScale": 7, |
|
"negativePrompt": "nsfw, lowres, bad anatomy, bad hands, text, error, missing fingers, extra digit, fewer digits, cropped, worst quality, low quality, normal quality, jpeg artifacts, signature, watermark, username, blurry, artist name," |
|
} |
|
}, |
|
{ |
|
"url": "https://image.civitai.com/xG1nkqKTMzGDvpLrqFT7WA/c8f8f31c-060f-4e30-a9b7-4848f2a7dd4e/width=450/5352346.jpeg", |
|
"nsfwLevel": 1, |
|
"width": 1728, |
|
"height": 1344, |
|
"hash": "UNI=cHIp9Z-;_MR*IoWBcGNG-UoJ%iR.wwM|", |
|
"type": "image", |
|
"metadata": { |
|
"hash": "UNI=cHIp9Z-;_MR*IoWBcGNG-UoJ%iR.wwM|", |
|
"size": 4272956, |
|
"width": 1728, |
|
"height": 1344 |
|
}, |
|
"availability": "Public", |
|
"meta": { |
|
"seed": 936248759, |
|
"steps": 28, |
|
"prompt": "1girl, toki \\(blue archive\\), blue archive, game cg, solo, double v, looking at viewer, upper body, cinematic angle, depth of field, outdoors, masterpiece, best quality", |
|
"sampler": "Euler a", |
|
"cfgScale": 7, |
|
"negativePrompt": "nsfw, lowres, bad anatomy, bad hands, text, error, missing fingers, extra digit, fewer digits, cropped, worst quality, low quality, normal quality, jpeg artifacts, signature, watermark, username, blurry, artist name," |
|
} |
|
}, |
|
{ |
|
"url": "https://image.civitai.com/xG1nkqKTMzGDvpLrqFT7WA/1047fffe-4d2a-4d63-bdc5-bbc91f26d09a/width=450/5352347.jpeg", |
|
"nsfwLevel": 1, |
|
"width": 1344, |
|
"height": 1728, |
|
"hash": "UFKwX*4n5S00_3R%%M8_~q0Kx]iwsm9Gxuog", |
|
"type": "image", |
|
"metadata": { |
|
"hash": "UFKwX*4n5S00_3R%%M8_~q0Kx]iwsm9Gxuog", |
|
"size": 3851171, |
|
"width": 1344, |
|
"height": 1728 |
|
}, |
|
"availability": "Public", |
|
"meta": { |
|
"steps": 28, |
|
"prompt": "A boy and a girl, Emiya Shirou and Artoria Pendragon from fate series, having their breakfast in the dining room. Emiya Shirou wears white t-shirt and jacket. Artoria Pendragon wears white dress with blue neck ribbon. Rice, soup, and minced meats are served on the table. They look at each other while smiling happily, masterpiece, best quality", |
|
"sampler": "Euler a", |
|
"cfgScale": 7, |
|
"negativePrompt": "nsfw, lowres, bad anatomy, bad hands, text, error, missing fingers, extra digit, fewer digits, cropped, worst quality, low quality, normal quality, jpeg artifacts, signature, watermark, username, blurry, artist name," |
|
} |
|
} |
|
], |
|
"downloadUrl": "https://civitai.com/api/download/models/293564", |
|
"creator": { |
|
"username": "CagliostroLab", |
|
"image": null |
|
}, |
|
"extensions": { |
|
"sd_civitai_helper": { |
|
"version": "1.8.4", |
|
"last_update": 1712879038, |
|
"skeleton_file": false |
|
} |
|
} |
|
} |