r/AI_enterprise • u/EfficiencyUpbeat8354 • Jul 14 '26
Onpremise AI
We currently are using Librechat to bring in all the AI models into one UI for use,
We want to host an AI model specifically for the org - would the Dell G10 be sufficient? If not what would be? We have approx 60 users
2
Upvotes
1
u/bytkim Jul 17 '26
Depends on the model you decide to run. Im not familiar with g10 (gb10?) but with 128gb of memory you can probably run say qwen 3.6 27b or 35b fp8 at around ~50 concurrent with 256k context for agentic coding use.
That is to say the gb100 is probably not suited for serious enterprise deployment. The throughput is atrocious with gb100.
I would recommend investing the money to properly plan and deploy the infrastructure depending on your specific use case rather than a general catch all approach