FAQ

Frequently asked questions about LlamaBox.

What is LlamaBox?

LlamaBox is a high-performance, local AI inference appliance. It's a dedicated piece of hardware that runs AI models on your network — not the cloud. You connect your own AI agents (Hermes, OpenClaw, custom scripts) to the LAN endpoint, and they communicate with the model running on LlamaBox. Your data never leaves your building. No subscriptions. No cloud dependency. One free Hermes client config is included with every purchase.

How is LlamaBox different from ChatGPT?

ChatGPT sends your prompts to OpenAI's servers. LlamaBox processes every request locally on your network. ChatGPT requires an internet connection and a monthly subscription. LlamaBox works offline and is a one-time purchase. ChatGPT only works with OpenAI's assistant. LlamaBox works with any OpenAI-compatible agent — yours or ours. Most importantly: ChatGPT sends your data to a cloud vendor. LlamaBox keeps your data on your LAN. Period.

Do I need to be tech-savvy to use it?

No. LlamaBox is designed to be "set it and forget it." It boots up, broadcasts its API endpoint on the LAN, and you're ready to connect your agent. Hermes — our recommended agent — is powerful but complex to set up. That's exactly why we include one free Hermes client configuration with every LlamaBox. We handle the technical headaches — Python environments, config files, terminal commands — so you can just open the app and start talking to your AI. If you're not technical, we'll set it up for you. If you are technical, you have full control over every aspect of the setup.

What is Hermes?

Hermes is a powerful, open-source AI agent that connects to your LlamaBox. It can search the web, read your files, run commands, browse the web, handle multi-step workflows, and more. Think of LlamaBox as the engine and Hermes as the driver. Hermes is incredibly powerful — but also incredibly complex to set up. That's why we include one free client configuration with every LlamaBox. We handle all the technical setup; you just use the agent.

What happens when the internet goes out?

LlamaBox keeps working. It's designed to run 100% offline after initial setup. No internet? No problem. Your AI runs locally, on your network, regardless of your internet connection. This is a core feature, not a fallback.

Can I use my own AI agents with LlamaBox?

Absolutely — that's the point. LlamaBox provides an OpenAI-compatible API endpoint (http://llamabox:8080/v1). Any agent that speaks that format works: Hermes, OpenClaw, custom scripts, web UIs, CLI tools. You bring your own agent. We provide the model. You're not locked into our ecosystem.

How do model updates work?

New Qwen version drops? We push updates to your appliance. No hardware changes needed. Just a model file swap — the same hardware runs the new model. Model updates are free for all LlamaBox customers. If you have a support contract, we handle the update remotely. If you don't, you can download an update script from the LlamaBox status page and run it locally — one click, done.

What models does LlamaBox run?

Every tier runs Qwen3.6 models at optimal quantization for the GPU. Starter runs Qwen3.6-35B (fits in 24GB). Pro runs Qwen3.6-70B at 4-bit quantization (fits in 48GB). Ultra runs Qwen3.6-110B at 4-bit quantization (fits in 96GB). You can swap models anytime — we provide model files and support the swap.

What's included with my LlamaBox?

Every LlamaBox comes with: pre-built hardware (tested before shipping), a golden image of Rocky Linux with everything configured, one free Hermes client configuration, free model updates, a backup USB drive, setup documentation, and hand delivery anywhere in the contiguous U.S. (We personally deliver, walk you through setup, and make sure everything runs perfectly.)

Do I need to pay for professional services?

No. Your LlamaBox comes with one free Hermes client config. That's your starting point — fully configured, tested, and ready to use. Professional services ($200/hour) are optional for anything beyond that: additional agents, custom prompts, tool configuration, training, troubleshooting, and more. You only pay for what you need.

Is the LlamaBox visible through the glass case?

Yes! The RTX GPU is visible through the tempered glass side panels. The GPU is the "engine of privacy" — it's what processes your AI requests locally. The glass case design is intentional: it proves the AI is running on-site, not in the cloud. Premium matte black branding on the front bezel. Furniture-ready design that looks great on a desk or in an office.

Can I upgrade my LlamaBox later?

Yes. All tiers are fully customizable. Need more RAM? More storage? A different GPU? We'll build it. You can also swap the GPU later — same OS image, same API endpoint, just a hardware upgrade. The model file changes, not the appliance.

What if my hardware fails?

Every LlamaBox ships with a backup USB drive. You can restore from backup in minutes. If the hardware fails completely, we'll ship a replacement appliance. If you have a support contract, we handle remote diagnostics and troubleshooting. All components are standard, off-the-shelf parts — easy to replace.

Who is LlamaBox for?

LlamaBox is for any business that handles sensitive data and needs AI capabilities. Primary targets: law firms (client data confidentiality), CPAs/financial advisors (regulatory compliance), medical offices (HIPAA), small businesses (privacy + cost savings), and anyone who values data privacy over convenience.

Can I ship a LlamaBox to a different location?

Yes. LlamaBox ships in professional packaging with foam. Hand delivery for the Louisville area. For remote locations, we use PC shipping boxes with custom foam cutouts. The appliance is built to survive transit. Just remember: LlamaBox is LAN-only. It won't broadcast on the internet — it's designed for a local network.

Is my data really safe?

LlamaBox is designed with maximum privacy in mind. The firewall defaults to DENY ALL. Only ports 22 (SSH) and 8080 (inference) are open, both LAN-only. No internet exposure. No inbound access. Your data literally never leaves your building. This isn't a feature — it's the design principle.

Still Have Questions?

We're happy to help. Reach out anytime.

Contact Us