Upstage Released Solar Mini 4 Language Model
The new 35-billion parameter model utilizes a sparse architecture to enable efficient single-GPU operation.
Updated on Oct. 1, 2026 in Artificial Intelligence

Live Poll
Would you prioritize using smaller, more cost-efficient AI models for your business or personal tasks?
Upstage has officially released Solar Mini 4, a language model featuring 35 billion parameters that relies on a specialized mixture of experts architecture. The model is currently available for developers via the OpenRouter API platform.
Why it matters
By restricting active inference to 3 billion parameters, the design aims to significantly improve computational efficiency and lower deployment costs for high-performance AI. This release highlights a broader industry shift toward optimizing large models for single-hardware accessibility.
Solar Mini 4 features 35 billion parameters in total, but it maintains performance efficiency by only activating 3 billion parameters during inference. It achieved a score of 24 points on the Intelligence Index and supports deployment on a single graphics processing unit via quantization.
The players
Upstage
An AI firm focused on developing efficient language models and optimization techniques for enterprise applications.
OpenRouter
An API platform that provides unified access to a wide range of open-source and proprietary language models.
The details
The model employs a mixture of experts structure, a machine learning method where only specific segments of the neural network are activated for a given prompt, rather than the entire parameter set. Additionally, the system utilizes quantization, which reduces the numerical precision of the model's weights, to decrease memory footprint. These combined techniques allow a 35-billion parameter model to operate efficiently on a single graphics processing unit, a standard hardware component for AI workloads.
Timeline
Upstage released the Solar Mini 4 model on October 1, 2026.
Discounted API pricing remains available until October 10, 2026.
The Tech Race
Solar Mini 4 follows the trend set by the Mixtral 8x7B mixture of experts architecture in prioritizing sparse activation to improve speed. This release underscores the competitive race to offer high-capacity intelligence without the hardware overhead typically required for models of this size.
Developers can currently access Solar Mini 4 through the OpenRouter API platform to integrate the model into existing workflows. Users looking to minimize costs can take advantage of a 70% API pricing discount that remains in effect until October 10, 2026.
The takeaway
The move toward sparse, 3-billion-active-parameter inference suggests that model size is becoming secondary to how effectively a system can selectively execute logic. Observers should track the model's performance in real-world benchmarks to see if the Intelligence Index score holds up against established competitors after the promotional pricing period ends.
What happens next
A 70% discount on API usage for the Solar Mini 4 model expires on October 10, 2026.
Further reading
For broader trends in model efficiency, visit the Artificial Intelligence section.
Source note: This article includes information reported by 조선일보.
Live Poll
Would you prioritize using smaller, more cost-efficient AI models for your business or personal tasks?







