Upstage Released Solar Mini 4 Language Model

The new 35-billion parameter model utilizes a sparse architecture to enable efficient single-GPU operation.

Updated on Oct. 1, 2026 in Artificial Intelligence

Bold flat-color editorial illustration depicting a central hardware processing block alongside modular cubic clusters representing efficient model architecture.
Upstage has released its new Solar Mini 4 language model, which uses a mixture-of-experts architecture to allow for high-performance AI deployment on single-GPU hardware. AI Illustration. Upload story photo >

Live Poll

Would you prioritize using smaller, more cost-efficient AI models for your business or personal tasks?

Upstage has officially released Solar Mini 4, a language model featuring 35 billion parameters that relies on a specialized mixture of experts architecture. The model is currently available for developers via the OpenRouter API platform.

Why it matters

By restricting active inference to 3 billion parameters, the design aims to significantly improve computational efficiency and lower deployment costs for high-performance AI. This release highlights a broader industry shift toward optimizing large models for single-hardware accessibility.

Solar Mini 4 features 35 billion parameters in total, but it maintains performance efficiency by only activating 3 billion parameters during inference. It achieved a score of 24 points on the Intelligence Index and supports deployment on a single graphics processing unit via quantization.

The players

Upstage

An AI firm focused on developing efficient language models and optimization techniques for enterprise applications.

OpenRouter

An API platform that provides unified access to a wide range of open-source and proprietary language models.

The details

The model employs a mixture of experts structure, a machine learning method where only specific segments of the neural network are activated for a given prompt, rather than the entire parameter set. Additionally, the system utilizes quantization, which reduces the numerical precision of the model's weights, to decrease memory footprint. These combined techniques allow a 35-billion parameter model to operate efficiently on a single graphics processing unit, a standard hardware component for AI workloads.

Timeline

  1. Upstage released the Solar Mini 4 model on October 1, 2026.

  2. Discounted API pricing remains available until October 10, 2026.

The Tech Race

Solar Mini 4 follows the trend set by the Mixtral 8x7B mixture of experts architecture in prioritizing sparse activation to improve speed. This release underscores the competitive race to offer high-capacity intelligence without the hardware overhead typically required for models of this size.

Developers can currently access Solar Mini 4 through the OpenRouter API platform to integrate the model into existing workflows. Users looking to minimize costs can take advantage of a 70% API pricing discount that remains in effect until October 10, 2026.

The takeaway

The move toward sparse, 3-billion-active-parameter inference suggests that model size is becoming secondary to how effectively a system can selectively execute logic. Observers should track the model's performance in real-world benchmarks to see if the Intelligence Index score holds up against established competitors after the promotional pricing period ends.

What happens next

A 70% discount on API usage for the Solar Mini 4 model expires on October 10, 2026.

Further reading

For broader trends in model efficiency, visit the Artificial Intelligence section.

Source note: This article includes information reported by 조선일보.

Live Poll

Would you prioritize using smaller, more cost-efficient AI models for your business or personal tasks?