Sberbank unveils GigaChat3.5-432B-A28B with day-zero GGUF support
Reported by the original publisher: New model: GigaChat3.5-432B-A28B (with day-0 GGUF support!). Analysis and context written by TickrWire.
Sberbank released GigaChat3.5-432B-A28B, a new large language model with 432 billion parameters, and provided day-zero GGUF support for efficient local inference.
- GigaChat3.5-432B-A28B is a 432B parameter LLM from Sberbank, one of the largest open models released to date.
- Day-zero GGUF support allows efficient local inference, reducing barriers to adoption for developers.
- The GGUF version is available via a pull request in the llama.cpp repository, not yet in the main branch.
- This release underscores Sberbank's expanding role in the open-source AI community.
Sberbank has introduced GigaChat3.5-432B-A28B, a new large language model with 432 billion parameters, positioning it among the largest open models available. The model is designed for high-performance applications and includes a base version for fine-tuning. Notably, the team has provided day-zero GGUF support, enabling efficient local inference without requiring complex setups. This is significant because GGUF is a widely adopted format for running LLMs on consumer hardware, making the model more accessible to developers and researchers. The GGUF version is not yet in the main branch but can be built from a pull request in the llama.cpp repository, indicating an active community-driven effort to integrate the model quickly. The release reflects Sberbank's growing influence in the open-source AI space, following its previous model launches and contributions to the ecosystem.
Enables local deployment of a 432B parameter model with GGUF, reducing infrastructure costs and improving accessibility.
Demonstrates the growing competition among organizations to release large, open models with practical deployment options.
- GGUF
- A file format for quantized large language models, enabling efficient local inference on consumer hardware.
LLMQwen3.8-27B: A Deep Dive Into Qwen's Newest Vision-Language Powerhouse
LLMAlibaba's Qwen team releases Qwen 3.8 models with open weights under the Apache 2.0 license
Meta’s ‘open’ AI, and a $250M deal gone very wrong
LLMZ.ai Ships GLM-5.3 Without Retraining the Base Model: Better at Complex Coding and Long-Horizon Tasks
LLMGemini 3.7 Flash lands with coding gains and undercuts its three-week-old predecessor's price by 50%
AI ToolsMeta AI’s new Mac app wants you to talk to your apps
Meta released a new Mac application that lets users control apps and dictate text using voice commands powered by its Muse Spark AI model.
New White House strategy clarifies military tech priorities: undersea, outer space and AI - Breaking Defense
The White House released a new strategy prioritizing military investments in artificial intelligence, space systems and undersea technologies to counter emerging threats.
AI in an iron grip: How dictatorships use artificial intelligence to strengthen their rule - theins.press
A new report examines how authoritarian governments deploy AI for surveillance, censorship, and propaganda to reinforce their power.
Stripe, OpenRouter finally strike a deal - Banking Dive
Stripe and OpenRouter have partnered to integrate Stripe's payment processing with OpenRouter's AI model aggregation platform.
How one Philadelphia school is using AI to strengthen student learning, not replace teachers - CBS News
A Philadelphia school is integrating AI tools to support teachers and improve student outcomes, focusing on collaboration rather than replacement.
Exclusive-How a Texas student blew the whistle on a rogue AI hacking attempt - The Mighty 790 KFGO
A Texas student uncovered an AI-powered hacking attempt targeting local systems, prompting a swift law enforcement response.