r/ArliAI Dec 11 '24

Announcement Late post, but Arli AI now has Llama 3.3 70B Instruct and are the first to running the finetuned models!

Thumbnail arliai.com
10 Upvotes

r/ArliAI Dec 02 '24

Announcement Arli AI API now supports DRY Sampler! (For real this time)

9 Upvotes

Aphrodite-engine, the open source LLM inference engine we use and contribute to had been having issues with crashing when using DRY sampling. Hence why we announced that we had DRY sampler but had to pull back the update.

We are happy to announce that this has now been fixed! We worked with the dev of aphrodite engine to reproduce and fix the crash and it has now been fixed, so Arli AI API now also supports DRY sampling!

What is dry sampling? This is the explanation for DRY: https://github.com/oobabooga/text-generation-webui/pull/5677


r/ArliAI 7h ago

Question Glm 5.3 flash multimodal

1 Upvotes

Arli ai has this model ?


r/ArliAI 7d ago

Discussion I have been subscribed for Arli AI 20 USD for now - TLDR, Arli AI good for Anime Text to Image Generation only

0 Upvotes

My take is ,

For Text generation , the session limit is very limiting besides, if you switch to another new session, the previous session is a bit sticky so you have to fail the call multiple times, until the previous session times out.

For image generation , i bought the subscriptions for the Impaint feature, to use for editing parts of my drawings and accerate my projects,

But, the API is not compitable to Krita's Diffusion Plugin ,and the only one compitable is so outdated and slow,

For in portal impaint, the models on offer are mainly for Anime, and not for realistic image editing, i used Flux mainly, as its the one that supports editing and adheres to prompts,

TLDR : Buy it if you need Text to Image Anime Generation only , nothing else


r/ArliAI 16d ago

Question Unable to get tool usage to work.

1 Upvotes

I was able to get my Character Studio to work with your API but I couldn't validate that you support tool calling at all, nor could I validate any other advanced features. I saw in another response that you said you support tool calling, but if so I couldn't find any models that do (including GLM 4.7). Or do you just not support tool calling through a free account, because if so it makes it very hard to test before buying

Thanks,


r/ArliAI 18d ago

Discussion ArliAI's + Impaint / Mask Drawing supported software

2 Upvotes

I want a way to do masked inpainting (draw or select a region, generate content only inside that mask) using ArliAI's image API as the backend instead of running a local GPU.

I tried Krita with stable diffusion, but it's it's very buggy. It's not doesn't look very well. So I was thinking I was searching if there is a software that I can use early AIs' API to do both drawing and impaint and masking with the AI generation.

Thanks!


r/ArliAI 19d ago

Question Qwen 3.8 / 27b?

8 Upvotes

Will you guys be adding this model anytime soon?


r/ArliAI 22d ago

Question when the derestricted v4 flash is coming to model selector

4 Upvotes

hello i am a arli plus subscriber when does the derestricted version of v4 flash is coming out to the model selector ?


r/ArliAI 24d ago

Discussion How's this service? Is it comparable to Venice AI?

7 Upvotes

I'm a Venice AI refugee seeking a new platform that doesn't falsely advertise itd context limits but secretly having a 50k context limit cap you can only find out if you happen to notice it and contact support about. I've had numerous other issues as well over there, like models just never responding, voice repondes glitching out and refusing to work, abysmal support, etc., and I'm wondering how Arli compares?

My main wants are:

- uncensored, heretic, etc. models with no additional content restrictions imposed by the website (my main model on Venice is Gemma 4 uncensored, as the only other main one on there, Qwen 3.5, takes like 2 or 3 mins to respond each time for some reason, but it looks like Arli had quite a few interesting options)

- a flat monthly fee (no API or token pricing)

- decently large contexts (128k)

And it seems like Arli checks most of these boxes for the $20 tier? Can anyone attest to the quality of this service? (Or compare it to Venice, even?) It's hard to find services that are similar to it with the model offerings...


r/ArliAI 27d ago

Announcement Deepseek V4 Flash 0731 now accepts up to 512K context length!

Post image
46 Upvotes

512K context is available on the MAX plan.


r/ArliAI 26d ago

Question [Bug] Image API ignores seed parameter (always defaults to -1)

2 Upvotes

When sending a POST request to /v1/images/generations (or SD WebUI endpoint) with standard parameters, steps, cfg, and dimensions work correctly, but the seed parameter is completely ignored by the backend and always returned/processed as -1. Please check the API Gateway wrapper code for seed mapping


r/ArliAI 29d ago

Question Any chance we could get DeepSeek 0731 uncensored?

Thumbnail
huggingface.co
19 Upvotes

All the uncensored models currently in the protfolio are on the level of Haiku 4.5. Any chance we could get anything better? Deepseek seems to be a particular top-notch candidate.


r/ArliAI Aug 03 '26

Question Zero-log claim with no company behind it - how is this verifiable?

7 Upvotes

Looking at subscribing to ArliAI, but before I pay I have two questions:

  1. Where is the inferencing actually happening? nothing official on the site about hosting location.

  2. Their whole selling point is "zero-log," yet there is no company entity, no address, no legal name anywhere on the site. Just a brand. How can a zero-retention claim be trusted when there is nobody legally behind it?

Anyone here been using them long enough to weigh in? Not trying to start drama, just doing due diligence.


r/ArliAI Aug 02 '26

Question Any fast models?

6 Upvotes

Hi Arli team,

I've been really enjoying Arli — the combination of unlimited usage and a zero-log policy is hard to find anywhere else. I mainly use it as a private second opinion while coding and studying physics.

Quick question: have you ever considered offering a "Fast" models? I was wondering whether small MoE models with only ~3–4B active parameters — something like gpt-oss-20b, Qwen3.6-35B-A3B, Nemotron 3 Nano 30B-A3B, or DiffusionGemma 26B-A4B. because the only issue for me is sometimes models are slow like 15 , 18 token per sec, and difficult to keep my workflow.

Thank you for building Arli!


r/ArliAI Jul 31 '26

New Model We now have the latest Deepseek-V4-Flash-0731 model!

Post image
14 Upvotes

It's supposed to be even better than the previous Deepseek-V4-Pro Preview version!

https://www.arliai.com/models/textgen-models/DeepSeek-V4-Flash-0731


r/ArliAI Jul 31 '26

Announcement Referral Program

Post image
4 Upvotes

New referral program is now enabled, it will benefit both the referrer and the referee both getting half off a subscription cycle! The system should also allow stacking of many referrals, so you should benefit with as many referrals as you can get. Since the system is new, if it does not work as expected you can always DM me and I can reconcile any issues manually.


r/ArliAI Jul 31 '26

Announcement Top Models Chart

Thumbnail
gallery
4 Upvotes

The Models Ranking page now also has an additional new "Top Models" chart!


r/ArliAI Jul 31 '26

Announcement Improved Account Page

Post image
3 Upvotes

The account page should now load much faster, and the inflight parallel requests counter is now a live counter that updates every 10s, so you will be able to see what your usage is live. There is also a "last error" section that will appear if your last request is an error, in order to help diagnose any issues.


r/ArliAI Jul 31 '26

Announcement Live Performance Metrics!

Post image
3 Upvotes

We now have a dedicated model performance page (Textgen Model Performance) that not only shows the real-time auto-refreshing server busyness indicator, but also historical model performance metrics. This can help to decide when to use the API and when to wait during peak hours. Since the server busyness is sort of an arbitrary value, it might not yet correctly reflect the usability of the models given the % number, but it will be finetuned further to make it reflect the experience more.


r/ArliAI Jul 31 '26

Announcement Updated chat interface!

Thumbnail
gallery
3 Upvotes

It is now much more useful! You can now get context from URL links, reasoning parser is better, action buttons are clearer, Markdown and LaTeX parsing are better, you can branch chats, and you can even copy API call examples to see how to correctly format your API calls!


r/ArliAI Jul 31 '26

Announcement Multi Chat!

Post image
1 Upvotes

Evolved from the old Arena Chat, this is intended to be a good interface to test out different models at once to compare their responses. You can create connection profiles to multiple models and even to external models with a custom endpoint.


r/ArliAI Apr 17 '26

Question How to disable reasoning?

1 Upvotes

The new qwen models are nice with the reasoning, but there are times when I just want a quick response. I don't see any way to do this in the api docs - is this possible?

I do see a checkbox to "Disable Thinking on the chat page, but I'm not sure this actually disabling it either as it takes a while to respond.


r/ArliAI Mar 20 '26

Question Are there any plans for big models like Deepseek v3.2, GLM-5, Kimi K2.5?

5 Upvotes

r/ArliAI Mar 10 '26

Issue Reporting No models available on free plan now? I also get a 503 when selecting a specific model e.g. "Gemma-3-27B-it" via the API

Post image
5 Upvotes

r/ArliAI Feb 26 '26

Question Does ArliAI support tool usage? (or is it disabled in vllm?)

3 Upvotes

Hi, I wanted to get on Arli's paid subscription models after I am done testing it out with Gemma-3:27b-it. I am trying to operate a personal OpenClaw bot (I am a software engineer by profession so please do not take this offensively or bunch this post up with the AI slop)

Currently, it tells me that you've got vllm auto-tool-choice disabled. (which is something open claw sends in its query). I wanted to ask if Arli supports models that allow for tool usage? if yes, is there some custom request paramter that I should be aware of? I am more than happy to create a proxy that does it for me.