GigaChat3.5-432B-A28B Released with Day-0 GGUF Support
WHY IT MATTERS
New model GigaChat3.5-432B with 28B active parameters released with immediate GGUF quantization support. Enables local deployment from launch.
Yandex released GigaChat3.5-432B with 28B active parameters in GGUF format on day-one, enabling immediate local quantization and inference without waiting for community conversion pipelines.
GGUF availability at launch compresses the timeline between model release and practical deployment. Community builders typically wait weeks for quantization formats; eliminating that lag reduces friction for benchmarking, fine-tuning experiments, and integration testing. This signals vendor investment in local-first accessibility as a release expectation rather than afterthought.
For operators, immediate GGUF support reduces vendor lock-in friction during evaluation phases. Teams can run inference locally on consumer hardware during initial testing, deferring cloud infrastructure decisions. This shifts cost-benefit analysis earlier in procurement cycles—local cost-of-compute becomes a direct comparison point at announcement rather than days later. The 28B active parameter design targets the consumer-grade deployment window (8-16GB VRAM machines), suggesting competitive positioning against similarly-scoped open models in cost-per-inference scenarios.
SOURCE
Reddit r/LocalLLaMA
SHARE
MORE FROM STUFFINSIDER
ZeroTTS Zero-Shot TTS Model with Efficient Attention for High-Quality Voice Cloning
Aug 20MODELSGLM5.3 Benchmarks Released: Artificial Analysis Results and Community Reaction
Aug 19MODELSKimon's Kimi-K3 Open-Source Project Surpasses 8,000 GitHub Stars
Aug 18MODELSQwen3.8-27B Benchmarks Match DeepSeek V4 and GPT-5.6 Luna Max
Aug 18