Configure ASUS Ascent GX10 for business
Timo Wevelsiep•Updated: 15.08.2026Editorial note: Versions, commands and prices may change. Please verify critical steps independently before production use. This guide does not replace individual consulting.
Receive the GX10 as a complete AI platform? In the AI Cube Pro, WZ-IT combines GB10 hardware with Open WebUI, local runtime, an agreed model, hardening, functional testing, and initial setup. Explore the AI Cube Pro
ASUS Ascent GX10 is a compact AI computer based on NVIDIA's GB10 Grace Blackwell Superchip. Its 128 GB of unified memory enables local model classes that do not fit into typical individual GPUs. Business use still requires more than starting a model. Users, data, network, and operations must form one platform.
Understand the hardware
ASUS describes the GX10 with GB10, 128 GB of coherent unified memory, 10 GbE, and ConnectX-7 support. Its ARM64 platform differs from a classical x86 workstation with a discrete GPU. Containers, Python packages, and utilities need ARM64 support or a controlled build process.
A model limit from a specification is not a user or performance promise. Actual work depends on the exact model, quantisation, context, concurrency, and runtime.
Setup in eight steps
- record firmware, operating system, NVIDIA driver, and runtime;
- use named administrators, SSH keys, and least privilege;
- configure stable addressing, DNS, firewall, and separate user and admin paths;
- plan storage for models, chat, knowledge, and backup with headroom;
- install a runtime compatible with model, API, concurrency, and ARM64;
- configure Open WebUI with persistence, users, groups, model access, and TLS;
- test language, quality, context, latency, and concurrent requests;
- define updates, monitoring, backup, restore, support access, and documentation.
What users see afterwards
Staff do not interact with CUDA commands or model containers. They receive Open WebUI as a familiar chat interface, can select approved local models, work with files, and create personal or shared knowledge spaces according to their rights.
Automated sources, business software, custom RAG pipelines, APIs, and MCP servers can then be added to the same platform. They are not blanket parts of base setup.
AI Cube Pro as the finished product
The AI Cube Pro is WZ-IT's fully prepared product configuration in this hardware class. It includes hardware, Open WebUI, local model runtime, an agreed and tested model, base setup, hardening, functional testing, initial setup, and five hours of support. The price is EUR 5,999 excluding VAT.
The system is usually ready within two weeks after configuration approval. Personal delivery and setup within Germany can be arranged. Optional managed operations start at EUR 149.90 excluding VAT per month.
Expansion and clustering
A second AI Cube can serve independent requests or participate in distributed inference. These use different software. Two systems are not one computer with 256 GB of shared memory. See connecting AI Cubes with ConnectX-7.
Where rack mounting, redundant power, several discrete GPUs, or stricter availability are required, the path leads to AI Cube Custom or a GPU server.
Sources
Rather have it operated?
You'd rather not run Local AI for Business yourself? WZ-IT handles setup, operations and maintenance - privacy-focused from Germany.
Enquiry
Assess local AI for your use case
Start with the AI Cube Pro or have us assess a custom AI platform, knowledge connection, or integration.
Frequently Asked Questions
Answers to the most important questions
The GB10 platform with 128 GB of unified memory is a compact foundation for local inference and initial production workloads. Model, concurrency, storage, and availability must fit the use case.
Firmware and updates, users and SSH, network and firewall, model runtime, model, Open WebUI, persistent data, backup, monitoring, and a documented support path all need attention.
AI Cube Pro is the complete WZ-IT product on this hardware class: preconfigured AI platform, agreed model, hardening, functional testing, initial setup, and five hours of support.
Yes. ASUS describes stacking two systems and the GB10 platform provides ConnectX-7. The benefit depends on distributing independent requests or one model across both systems.
More on Local AI for Business
- The open-source LLM stack
- What is LiteLLM?
- What is Langfuse?
- What is vLLM?
- vLLM vs. Ollama
- What is RAG?
- Connect Open WebUI to Nextcloud (RAG with ACLs)
- What is local AI?
- Cloud AI vs. self-hosted
- AI sovereignty for companies
- Which LLM to self-host?
- Sizing GPU & VRAM
- Inference vs. Training
- Qdrant vs. pgvector
- The EU AI Act for companies
- Local AI for confidentiality professions
- Processing documents with AI
- AI agents & automation
- RAG with permissions
- Chatbot or knowledge navigator?
- AI agents: permissions and approvals
- AI assistants and the works council
- GDPR-compliant AI: assessment criteria
- What does a local AI server cost?
- Size a local AI server by users
- LLM models on 128 GB unified memory
- RAG with Nextcloud, SharePoint, and DMS
- Provide secure remote access to local AI
- Connect AI Cubes with ConnectX-7
- Run Open WebUI as a production appliance
- Configure ASUS Ascent GX10 for business
- Configure NVIDIA DGX Spark for business
- Configure Acer Veriton GN100 for business
- Configure Dell Pro Max with GB10 for business
- Configure Gigabyte AI TOP ATOM for business
- Configure HP ZGX Nano G1n for business
- Configure Lenovo ThinkStation PGX for business
- Configure MSI EdgeXpert for business





