Aura Large GGUF
Portable local inference for llama.cpp-compatible desktops, workstations, and capable local systems.
View model ↗Large tier · Two deployment formats
Aura Large is a Gemma 4 31B model family built around a larger local foundation. Choose the GGUF or BF16 release without changing the identity at the center of the experience.
Aura Large is the pinnacle of the Aura line. The large model delivers the most intuitive companion experience that an enthusiast could hope for.
Choose a format
Each release is built from the Gemma 4 31B model family and validated independently for its target runtime.
Portable local inference for llama.cpp-compatible desktops, workstations, and capable local systems.
View model ↗High-fidelity Transformers inference and evaluation for GPUs with sufficient memory.
View model ↗