Fuzzball 4.3 | Run validated models in one command

CIQ’s Fuzzball 4.3 runs validated open-weight models on a customer’s own GPUs with one command

Open models deploy from a catalog with validated presets, and in-cluster agents discover them automatically, across heterogeneous GPU fleets on-premises and in major clouds.

RENO, Nev., October 8, 2026 - CIQ, the enterprise software company behind Rocky Linux, today announced the general availability of Fuzzball 4.3, which brings ready-to-run AI models and agents to its sovereign AI orchestration platform. The release helps enterprises operationalize open-weight models and agentic AI on infrastructure they already run, with the control and governance required for AI in production. Fuzzball 4.3 deploys validated models from a catalog, connects agents to them automatically, and scales inference across heterogeneous GPU fleets, on-premises and in the cloud.

Enterprises are shifting from AI pilots to production inference and repatriating workloads to private infrastructure for cost savings, security, and sovereignty. Meeting these goals often means standing up a new infrastructure stack before the first model serves a token, then integrating models, gateways, and agents by hand. Fuzzball 4.3 gives teams a faster path to productivity without rearchitecting their environment.

"Enterprises have already made their infrastructure decisions, and private AI should fit into them. With Fuzzball 4.3, a team starts a model, starts an agent, and the two find each other, on the GPUs, clusters, and clouds the organization already runs. That is what turnkey should mean," said Gregory Kurtzer, CEO of CIQ and Founder of Rocky Linux.

Models start with one command from the self-service model catalog. Eleven presets cover open-weight models from the GPT-OSS, Llama 4, Gemma 4, Qwen3-Coder, Mistral, and Nemotron 3 families, and a vLLM-based inference entry serves any other compatible Hugging Face model.

Models autoscale behind a single, stable endpoint and can scale down to zero running instances when idle, returning GPU capacity to other workloads, ready to spin back up for additional requests. And large models can scale across multiple nodes for additional resources.

Fuzzball 4.3 connects agents to models automatically. OpenCode and hermes-agent, two coding agents in the workflow catalog, run inside the cluster and discover running models through gateways, with no manual integration. Each model ships behind a built-in LiteLLM gateway with an OpenAI-compatible endpoint. Or, run one central gateway that fronts every model a user is authorized to reach. Another entry in the catalog provides a retrieval service to connect your private agents to enterprise documents.

Governance is built into the platform. Idle pools wake only for authenticated requests. Administrators control which organizations and groups can use which resources, and every workload runs unprivileged and rootless. Fuzzball 4.3 supports heterogeneous GPU fleets across NVIDIA and AMD for serving and health-aware scheduling, all under one operating model.

Availability

Fuzzball 4.3 is generally available now. Fuzzball deploys to AWS, Google Cloud, Oracle Cloud, CoreWeave, and Azure, as well as on-premises clusters built with Warewulf, VMware, or bare metal. To get started, contact CIQ at ciq.com/products/fuzzball.

Built for scale. Chosen by the world’s best.

2.75M+

Rocky Linux instances

Being used world wide

90%

Of fortune 100 companies

Use CIQ supported technologies

250k

Avg. monthly downloads

Rocky Linux

9

Enterprise products

Spanning the kernel to the orchestrator

Have questions about your infrastructure?

Talk to a CIQ engineer about Rocky Linux, HPC, and AI infrastructure.

Talk to an Expert