Configure Gigabyte AI TOP ATOM for business
Timo Wevelsiep•Updated: 15.08.2026Editorial note: Versions, commands and prices may change. Please verify critical steps independently before production use. This guide does not replace individual consulting.
Move AI TOP ATOM from installation to production use? WZ-IT provides the GB10 class with Open WebUI, a local model, hardening, integration, and support as the AI Cube Pro. Explore the AI Cube Pro
Gigabyte AI TOP ATOM is a compact AI computer based on NVIDIA's GB10 Grace Blackwell Superchip. Its 128 GB of unified memory and up to 4 TB of NVMe storage provide a foundation for local inference, model testing, and development. Gigabyte also provides AI TOP software tools. Company-wide use still requires defined user, data, and operating paths.
Understand the hardware
Gigabyte states up to 1 PFLOP FP4, models up to 200 billion parameters, and ConnectX-7 for connecting two systems. The 20-core Arm processor and Blackwell GPU use the same memory, so software and containers need ARM64 support.
Maximum parameter count does not equal a sensible business model. Quantisation, context, KV cache, and concurrent chats materially affect memory use and throughput.
Business setup
A controlled setup documents firmware and drivers; configures named administrators, SSH, DNS, TLS, and firewall rules; selects a runtime and model for the expected concurrency; deploys Open WebUI with users and model permissions; covers models, chat data, knowledge, and configuration in backups; and defines monitoring, updates, restore, and support access.
AI TOP software and Open WebUI
Gigabyte's tools primarily support development and local AI workflows. Employees usually benefit from one browser-based interface. Open WebUI provides chat, model selection, personal and shared knowledge spaces, and user management.
The platform can offer local and approved external models together. Technical and organisational rules must define which data may reach which endpoint.
When AI TOP ATOM is a good fit
AI TOP ATOM is suitable where Gigabyte is the preferred supplier or the accompanying AI TOP tools add value to development workflows. Open WebUI still provides the central interface for later use by business departments; developer tools and an employee platform serve different needs.
SSD configuration, target models, and concurrency should be defined before ordering. An internal knowledge assistant for one team has different priorities from a development group maintaining several model artefacts. Because the common GB10 foundation offers similar core performance, availability, storage, support, and integration often matter more than theoretical compute.
Complete entry and later scaling
The AI Cube Pro adds Open WebUI, model runtime, an agreed local model, hardening, functional testing, initial setup, and five hours of support. Automated RAG sources, interfaces, and MCP servers are added for the use case.
With growing use, another AI Cube can serve independent requests. Distributed inference is possible for larger individual models but is not automatically faster. The ConnectX-7 guide explains both paths.
Sources
Rather have it operated?
You'd rather not run Local AI for Business yourself? WZ-IT handles setup, operations and maintenance - privacy-focused from Germany.
Enquiry
Assess local AI for your use case
Start with the AI Cube Pro or have us assess a custom AI platform, knowledge connection, or integration.
Frequently Asked Questions
Answers to the most important questions
It provides GB10, 128 GB of unified memory, and up to 4 TB of NVMe storage. It suits local inference and development where runtime, concurrency, and operating requirements match the use case.
Gigabyte states up to 200 billion parameters on one system and 405 billion across two connected devices. These are technical limits rather than guarantees of interactive performance.
Yes. Its ConnectX-7 NIC supports a high-speed link for distributed workloads. Request balancing and model distribution each require their own software configuration.
An agreed model, runtime, Open WebUI, users and permissions, TLS, backup, monitoring, updates, and a documented support path still need configuration.
More on Local AI for Business
- The open-source LLM stack
- What is LiteLLM?
- What is Langfuse?
- What is vLLM?
- vLLM vs. Ollama
- What is RAG?
- Connect Open WebUI to Nextcloud (RAG with ACLs)
- What is local AI?
- Cloud AI vs. self-hosted
- AI sovereignty for companies
- Which LLM to self-host?
- Sizing GPU & VRAM
- Inference vs. Training
- Qdrant vs. pgvector
- The EU AI Act for companies
- Local AI for confidentiality professions
- Processing documents with AI
- AI agents & automation
- RAG with permissions
- Chatbot or knowledge navigator?
- AI agents: permissions and approvals
- AI assistants and the works council
- GDPR-compliant AI: assessment criteria
- What does a local AI server cost?
- Size a local AI server by users
- LLM models on 128 GB unified memory
- RAG with Nextcloud, SharePoint, and DMS
- Provide secure remote access to local AI
- Connect AI Cubes with ConnectX-7
- Run Open WebUI as a production appliance
- Configure ASUS Ascent GX10 for business
- Configure NVIDIA DGX Spark for business
- Configure Acer Veriton GN100 for business
- Configure Dell Pro Max with GB10 for business
- Configure Gigabyte AI TOP ATOM for business
- Configure HP ZGX Nano G1n for business
- Configure Lenovo ThinkStation PGX for business
- Configure MSI EdgeXpert for business





