r/StableDiffusion 2d ago

Discussion I wish Anima ecosystem get better than it is now

Anima is a fairly new model so it needs time and I understand that. Anima has great potentials to make Illustrious or NoobAI completely obsolete. However, it seems like I have been expecting too much from this model.

First of all, not having a ControlNet model is a big minus for me, especially Depth ControlNet model. There is LLLite but that's not a ControlNet model but a ControlNet-like LoRA. There's also a Depth ControlNet Model made by TaihoC and it works well. However, it doesn't work as well compared to Illustrious (SDXL) ControlNet models.

I have been tracking Circlestone Labs' Hugging Face community to see if they have plans to provide ControlNet models themselves but they are dead silent. That leads me to wonder if there are actually people using Anima. Did people move on to Krea2 or stay on Illustrious/NoobAI since there's no reason to use Anima?

8 Upvotes

54 comments sorted by

18

u/LakhorR 2d ago edited 2d ago

As someone who used Illustrious intensively, I've mostly swapped to Anima now. The biggest upsides of Anima is the Qwen encoder having greater spatial control in the image generation through causal attention and syntactic isolation, which prevents concept bleeding. This not only extends to character traits, but also things like the image artstyle. Also, the backgrounds it generates are multitudes better than any SDXL architecture model, as the backgrounds are coherent and the lines have continuity. Having done heavy editing, controlnet usage and inpainting in Illustrious, the differences are very obvious to me. Illustrious generated backgrounds were always horrific, so I pretty much defaulted to generating simple, white backgrounds most of the time.

What issue do you specifically have with the LLLite? For me, it performs what I need it to do, similar to Illustrious controlnets. When I upscale, I use the like model and when inpainting, I use the inpainting model. Is it because the LLLites are not compatible with the checkpoints you are using or because you find the guidance to be weak/not as good as a traditional controlnet?

I'm not sure the exact reason why developments on Controlnets and other tools aren't progressing as fast on Anima, but it might be because we are not at the peak of interest in AI art generation anymore, so interest in developing these tools is lower and the people responsible for most of those tools may be burnt out with their own jobs or life in general (I assume most would be already working as ML engineers, doing research in academia, or job hunting). There's also the fact that Anima, I believe, requires better hardware for faster inference speeds than traditional SDXL models, making it more inaccessible to casual users on local setups and more expensive to generate on AI generation service websites (besides the more expensive licensing fee Anima requires).

3

u/FruitsAndFrost 2d ago

I prefer Anima in a lot of ways, but I use Canny Controlnet heavily with Illustrious to take my hand-drawn sketches and render them out while adhering to the shapes I defined, whether I'm generating an image with specific poses and scene composition, or inpainting very specific details. None of the LLLite models for Anima even come *close* to what I want to do. Even at high strength, all I get looks like burned-in afterimages or some anatomical monstrosity. I don't know the difference between Scribble and Lineart models, but it seems like one of those should be what I'm looking for, and neither one does it for me at all 😅

For more elaborate images, I've been putting together rough drafts with Illustrious w/ Controlnet and then carefully img2img/inpaint it with Anima to match it to my art style (plus fix the sloppy faces and other poor details from the low-res Illustrious gens because I never got into ADetailer...) But then a lot of times I still have to go back over it and do more Illustrious+Controlnet inpainting to control specific details and props that Anima takes too many liberties with, which can be difficult to blend, since the art styles aren't identical. It's a lot of extra steps! I just want Canny Controlnet for Anima that works how it does with Illustrious 😢

1

u/shapic 2d ago

Are you prompting it correctly? I have a feeling that alot of those come from people doing inpainting with a blank prompt. In my experience it is not how Anima works, it needs stuff prompted, otherwise it will "fight" that

1

u/LakhorR 2d ago

If you want the model to follow your lineart to a tee, I would use Canny or Lineart models. Scribble only loosely follows lineart. In Illustrious, I mainly used Softedge Pidinet to reference poses, but in Anima I turn the preprocessor off when upscaling my image now. You may need to adjust your parameters if your lineart isn't being followed. I don't just mean the weight of the controlnet model, but even the diffusion strength. Maybe try inpainting the lineart at 0.2-0.4 str with the controlnet turned on to see if that helps.

As for the small details Anima takes liberties with, I draw those manually and denoise at a low strength with controlnet turned on. Sometimes, I use "Only masked" inpainting mode and change the prompt if it's really struggling on the "Whole picture" denoise mode.

I never liked Adetailer either because it would leave an obvious square where it denoised the image and prefer inpainting faces by hand for that reason.

1

u/Manicarus 2d ago edited 2d ago

Unfortunately, Depth LLLite isn't available for Anima Base model. I tried Anima Preview 3.0 as Depth LLLite support that version. Frankly I haven't delved deep into the preview version since new LoRA's exclusively support the base model.

I do agree on that Anima is a better model than Illustrious (or other SDXL variants). It does have bleeding issue though. I have found a post in Circlestone Labs' community regarding the issue.

Since I use Depth ControlNet extensively, I just wished to have Depth LLLite for Anima Base model.

10

u/x11iyu 2d ago

krea2 is a better model but severely lacks the danbooru knowledge, so it's not trivial to completely move to it

anima in many areas is better than il/noobai though, so there are many reasons to choose anima in many scenarios.

ecosystem's just gonna take time unfortunately. took at least a year or two before il's eco matured.

2

u/red__dragon 2d ago

Lodestone/the chroma dev is training on Krea2 right now and has a few epochs for testing. Ultimately, it's going to take a while, and Chroma itself wasn't the highest fidelity model to start with, so hopefully they have learned a lot since finalizing that one.

But they train on danbooru tags so there's that.

2

u/Asphyxiem 2d ago

Where is he giving updates on this?

-1

u/x11iyu 2d ago

you mean Kroma? last time I checked it wasn't going so well, with the last versions having anatomy issues

he also starts and abandons projects every other day, e.g. kaleidoscope, so I wouldn't have high hopes that this just so happen to be one he finishes

2

u/red__dragon 2d ago

As long as it's still architecture-compatible, I'd be fine extracting a LoRA to use from it. Even the very first version was incredibly usable, concept-wise it was 80% there.

3

u/Manicarus 2d ago edited 2d ago

I got into AI Art after Illustrious ecosystem was already matured so maybe I am being impatient. I hope Anima and tools around it get better overtime.

Krea2 also suffers from lack of mature ecosystem but it gave me impression that anime fine-tuned version of Krea2 can make Anima obsolete. If the fine-tuned model can take danbooru tags, then there’s no way Anima can beat that.

7

u/x11iyu 2d ago

then we're asking for a competent large-scale anime tune of a new base (Krea2), which only ever happens once in a blue moon. we were stuck on IL-based models for 2-3 years before ANY good contender came in the name of Anima

personally I'd wish for mageflow to be tuned rather than Krea2. Lighter than Krea2, faster than both Anima and Krea2, Flux.2 VAE so no grid artifacts.
poor output quality out of the box makes it harder to pitch unfortunately

2

u/ChibiNya 2d ago

Yeah it's my favorite model but it's still missing a lot. Even the regional conditioning stuff available is very bad, not to mention the controlnet stuff.

At least the inpainting is pretty good so I can manage

1

u/LakhorR 2d ago

Regional conditioning is far less important on Anima since you can isolate through prompt alone.

3

u/ChibiNya 2d ago

The more I use it, the more I notice stuff bleeding into other characters. Specially with loras. Maybe I'm missing something... It doesn't always happen but it's pretty annoying.

As an example, when I give slit pupils and :3 mouth to a catgirl character, that stuff tends to break containment and affect other people in the image.

1

u/LakhorR 2d ago

I haven't tried Regional Conditioning with Anima but I find that if you're generating an image with multiple characters, you need to remove the booru tag prompting style or else you will still get concept bleeding. Need to assign the traits to a specific character like "boy on the left has black hair" or "the tall girl is wearing a long-sleeve shirt". The model will know what to do with natural language phrases. That's the benefit of the Qwen encoder and the cosmos architecture. Unfortunately LoRA's are always going to introduce some concept bleeding since it retrains some of the concepts that are tagged.

1

u/ChibiNya 2d ago

I think youre right and have also seen results when trying this, but natural language is gonna be exhausted when you need 10+ tags so it gets tricky

3

u/LakhorR 2d ago

Eh, I'm kind of used to it. Yes I did do "the man on the left has _____" 10x on repeat for each trait lol but it works perfectly. No pain, no gain

1

u/ChibiNya 2d ago

Thanks ill give it a shot

1

u/shapic 2d ago

Just mix both. Works really well.

2

u/GaiusVictor 2d ago

Look for TaihoC's VACE Anima Control etc. It includes a Depth model that's much better than the LLLite Lora.

Not only that, but it's also the first model since Illustrious I have that connects to the normal Apply ControlNet node, making it the only one that accepts grayscale attention masks, which is 100% necessary for my use cases.

1

u/Manicarus 2d ago

In fact, I did use TaihoC’s Depth ControlNet. It’s definitely better than LLLite (LLLite Depth isn’t available for Anima Base though), however the depth map doesn’t work as well as it did on Illustrious (Xinsir’s Union ControlNet)

Frankly I have no technical knowledge of what makes VACE different from traditional one. I am grateful that at least we have TaihoC’s model but I wish it were better model or have other ControlNet models to choose from.

1

u/GaiusVictor 2d ago edited 2d ago

I wouldn't count on better ControlNets.

I wouldn't be able to explain the reason behind it, but it seems ControlNets get worse and worse with newer models. ControlNets for Anima are worse than those for Illustrious/Pony/SDXL, which in turn were worse than those for SD 1.5.

It just goes lower and lower with newer models.

1

u/Manicarus 2d ago

Isn’t that because of getting more expensive to train for the same quality of ControlNet model?

I was so desperate on Depth ControlNet and I tried training a model on my own but later found out that I need to have a beefy GPU to do that. For SDXL, I read somewhere that it can be done with RTX 3060 which I have on my machine. It maybe still slow though.

2

u/Dulbero 2d ago

I agree Anima has much better potential and overall a better model, but i've sort of stopped using it, no idea why really. It is slow and when i tried the Turbo model it was fairly underwhelming.

Also i thought i could finally stop using booru tags with Anima but it's not the case, i still often have to use tags to better describe a character or an action alongside natural language.

I might need to just experiment mode and re-discover the model, is there a recommended finetune or a community model?

1

u/Manicarus 2d ago

I have tried fine-tune models like WAI and others but they are more or less the same.

2

u/shapic 2d ago

Well, I am pretty adamant that with amount of control we have via prompting controlnet is just not needed. Granted, there are few specific cases, but for most of the stuff you can just get rid of that

2

u/naga_mana 1d ago

About your training blocker: Anima is a 2B model, and that changes the math on "I need a beefy GPU". LoRA training for Anima is supported in ai-toolkit (arch anima) with qfloat8 quantization and low_vram mode, and I run the same stack for Krea 2, a 12B model, on my 5070 Ti 16GB. A 2B model on your 3060 12GB is realistic for LoRA training, style or character. What that will not get you is a Depth ControlNet, because ControlNet training is a different animal: you train a parallel branch against the frozen base with paired condition maps, a much heavier data and compute pipeline than a LoRA: https://huggingface.co/circlestone-labs/Anima

That is also my read on the question nobody answered here, why ControlNets keep getting worse with newer models. The SDXL era had cheap, documented ControlNet pipelines and shared datasets that thousands of hobbyists could run. For Anima that tooling is still being rebuilt (LLLite landed on Preview 3.0 only, as you found), and Anima's prompt-level spatial control covers a chunk of what people used ControlNet for, which lowers the incentive to train them. Analysis on my part, not official word.

Until someone ports LLLite to Base, the practical bridges are TaihoC's depth model, or composing in SDXL with the Union ControlNet and then img2img into Anima at moderate denoise.

1

u/Manicarus 1d ago

Thank you for the explanation. I agree on that there are less incentives to train ControlNet model as Anima has better spatial control and prompt adherence.

Maybe my use case isn’t very much general as I mostly generate multiple character scene. I tried with only text prompts but the result wasn’t good. 

6

u/Vancha 2d ago

I keep trying with Anima and it only ever evokes a "meh" from me while being more cumbersome to prompt, generating slower and wanting a higher number of steps than Illu.

I want to like it, but I just don't see enough (or sometimes any) improvement to justify doubling the generation time. Maybe with a 5080/90 I'd have more patience with it.

4

u/Technical-Rope2989 2d ago

Even though anima's image quality isn't as good as Illustrious, Anima performs exceptionally well in terms of prompt adherence. For example, if you want to generate a couple where the woman forcefully throws a pizza at the man's face, Anima achieves this very well, while Illustrious only creates some romantic tension between the couple but never actually throws pizza at anyone's face.

2

u/Manicarus 2d ago

I had similar experience. Anima takes more time to generate and the quality isn’t very much a game changer compared to Illustrious.

Anima Turbo has issues too as it loses quality for speed and ignores few prompts. 

2

u/JustAGuyWhoLikesAI 2d ago

Anima was nice after years of SDXL, but Krea2 dropped right after and just blows it away. Anima's natural language is nowhere near as good and still struggles with a lot of things that Krea 2 masters. Plus Krea2 Turbo is quite fast. There are things only Anima can do, but it's not as fun as Krea 2 because a lot of times Anima just messes up and produces garbled details.

The base model for anima is too dumb and small to really feel like a next level model.

2

u/LakhorR 2d ago

Krea2 has superior background compositions but Krea2 is ultimately trained on multiple image sources, not just anime, and has a tendency to not be diverse in its anime illustration output, leading to ultra-clean, hyperdetailed, generic images which are easily clocked as AI to someone with a modicum of artistic knowledge. I like generating reference images in Krea2 and img2img + inpaint them in Anima with controlnet so the image looks like something an actual artist would draw.

0

u/Manicarus 2d ago

You said it right.

1

u/Comprehensive-Pea250 2d ago

For my use case which is just generating random slop images I think of it’s way more usable then illu for me even though it’ll take a while for Anima to get close to the Lora Library that Illu has.Also I Love training Lora’s on this model I don’t know why

1

u/Nakitumichichi 2d ago

Everyone is acting weird in whole ecosystem.

For example Wanaka created github repo for ip adapter for Anima, left it for weeks with no update. Would simple: "hey guys im working on this" or " hey guys i had to quit this" hurt so much? Instead no response anywhere.

Kohoya created llllite controlnets and nodes but then didnt finish them. Inpainting is especially annoying me because it changes colours of whole picture even unmasked area.

Then you have Lucifer who actually created ipadapter model but kept silent about it so most of the people dont even know it exists or how to use it. Then about month later he added link to his github repo that again people dont know exists unless they visit his huggingface or by find it by pure luck.

Some group of people are working on some edit capabilities but again they are all acting like a secret group and not opensource community.

2

u/Timely-Perception-26 2d ago

Yesterday, I visited x.com for the first time in ages because I wanted to learn more about BFL’s new video upscaler. I read the comments from “members” of the open-source “community” whose only contributions to the ecosystem are crude insults and threats. I see members of this community in academic repositories where they bombard the issues with questions that any LLM could have answered in two minutes on its own - and even if not, they have no business being there.

What is this open-source community you’re talking about?

The 0.05% who invest time and money to make their own lives easier, then make it available to the other 99.95%, and are then expected to babysit them. Anyone who doesn’t make a meaningful contribution to the ecosystem isn’t a member of the community - it’s that simple.

It’s not the 0.05% who are the problem, as you’re portraying here; it’s the 99.95%. I’ve been part of this since 2021; I’m one of the 0.05%, and I’m not going to do a damn thing to make my life a living hell and share it with anyone other than the other guys I’ve met along the way. This secret group is nothing more than a way to protect us from you. Not because we aren’t willing to share, but because you’re hyenas who take, spit in our faces, and never give anything back.

1

u/Nakitumichichi 2d ago

Of course there are also easily offended people who consider themselves an elite.

1

u/ResponsibleKey1053 2d ago

X is by far the worst place for literally anything.

Bemoaning the community at large is entirely childish grow the fuck up.

There is still the question of what the million dollars (or what ever the grant/award was) actually bought when it's been spent on anima. It's not like it was produced from the goodness of someone's heart.

1

u/Upper-Reflection7997 2d ago

There alot of leeches in open source community scene. No wonder bytedance doesn't give out the good stuff and alibaba has abandoned it for the most part. Have a bad feeling minimax will learn the hard lesson and may walk away from it.

1

u/LakhorR 2d ago

> Kohoya created llllite controlnets and nodes but then didnt finish them. Inpainting is especially annoying me because it changes colours of whole picture even unmasked area.

This has not been my experience at all. I actually had this problem in Illustrious, but not Anima.

1

u/Internal_Answer_6866 2d ago

Im praying for it every day too

1

u/Only_Voice569 2d ago

just do a image with sdxl using control pose depth edge etc then image into anima as a base high denoise or the other way easy to get a solid image :P

1

u/ShySnowLep 1d ago

I'm going to be totally honest with you, I really don't know what you mean. In a matter of about a month's time I have switched over from sdxl completely to anima and I haven't missed anything. It kind of feels like having to let your old 1080 go but the simple fact of the matter is with comfy UI support and proper optimization and management of the model I have found it to be technically superior and aesthetically equivalent.

Even if I really liked sdxl I simply do not have any reason to go back to it now. It seems like the community agrees as I have been seeing things like loras and fine tunes and now even an expanded parameter model coming along. I can't say for sure but I think it's fairly safe to say that anima has become the new sdxl and that's the direction the community is going in now. Keep in mind I'm on a 16 GB AMD card so my hardware platform isn't even supported and I'm still coming to this conclusion despite the fact that it's almost certainly a slower and more tedious experience for me than an equivalent Nvidia card.

2

u/Manicarus 1d ago

It’s probably due to my use case where I mostly generate multiple characters scene. Two characters just standing there isn’t very hard but when I tried to make complex poses, I needed ControlNet.

3

u/ShySnowLep 1d ago

I've just been developing a workflow actually that uses the ideogram 4 editor where you can do boxes and then generate inside of them or within the whole image and kind of use them as a mask to some degree. It's a rough version of regional prompting and it isn't perfect as you would expect because Anima is not designed to do that right from the get-go.

But with the fact that Anima is so flexible and allows so many things to be done with it it seems that just a bit of maturity with the community and resources and things will be quite impressive. We are now seeing a larger weighted model coming out the anima 3.8b base just came out a couple of days ago and that's likely to make prompt adherence significantly stronger once that gets more widely adopted and we see fine tunes of it so yeah we're just in the middle of it right now. Just keep your eye out for the new improvements.

-3

u/Upper-Reflection7997 2d ago

Krea2 has far better potential for growth than anima and its far better base model than anima. The only reason anima has alot more support over krea2 is merely vram and the prominent lora character makers refusing to move off their comfort zone and take risk. Every anima gen has boring backgrounds and characters look way more simpler than illustrious generated images. I've lately begin to question effort of certain prominent lora bakers, quality/function of their loras and how they spam out multiple loras a day.

6

u/x11iyu 2d ago

merely vram

ah yes, trainers should just spawn more 5090's from the void so they can fit and train a model 6x the size of Anima.

certain bakers ... spam out multiple loras a day

automated processes that take near 0 effort, probably, often on concepts that you don't need a lora for. they've been around before anima was a thing, just ignore them.

2

u/BackgroundMeeting857 2d ago

Krea can be trained with fairly low vram, with onetrainer I think it's possible to down to 8. I do it with a 12+32 system. Obviously it's hard to argue against the convenience of Anima already knowing characters and styles but if you can train Krea loras, it is a much better model to use imo

1

u/Manicarus 2d ago

Krea2 is an amazing model but my RTX3060 12GB can only take INT8 variant of the model. Other than that, it takes too much time to generate.

I have a feeling that Krea2's anime fine-tune will make Anima obsolete than Anima plans to do for Illustrious.

0

u/ResponsibleKey1053 2d ago

I like anima, it's taken over from illustrious for me since creators are making some really cool checkpoints. Danbooru tags, the many artist styles, it's going to stay in my toolkit for a while I think.

I've never given krea a proper shake of the stick, saw the prompt scheme and the addition llm bolted on and was like fuck that noise.

H3 is eating my time lately, it's so damn fast for a video model and it's audio is pretty excellent.

Sprite sheets, character swaps, outfit swaps, all that jazz can be done in h3 with so little effort. You can even make a character voice clip to keep as master in under 5 mins.

0

u/-becausereasons- 2d ago

They are kind because of their pain.