---
title: "Release notes for Dedicated Inference"
sidebar_label: "Release notes"
sidebar_position: 99
description: "Updates in Dedicated Inference"
---

import {FaqWrapper, FaqItem, FaqQuestion, FaqAnswer} from '@selectel/docux/components'
import Formbricks from '@theme/MDXComponents/Formbricks'

# Release notes for Dedicated Inference

## 2026 \{#2026}

<FaqWrapper>
  <FaqItem>
    <FaqQuestion>August</FaqQuestion>

    <FaqAnswer>
      * added the ability to track inference service metrics using the [Metrics](/metrics/about/about-metrics.mdx) service;

      * added new models:

        * [qwen3-6-27b-nvfp4](https://huggingface.co/unsloth/Qwen3.6-27B-NVFP4);
        * [qwen3-6-35b-a3b-nvfp4](https://huggingface.co/unsloth/Qwen3.6-35B-A3B-NVFP4).

        You can view the current list of available models in the [Control Panel](https://my.selectel.ru/ml/default/models-catalog/): in the top menu, click **Products** → **Dedicated Inference**;

      * simplified the inference service cost calculation: it is now based only on GPU consumption. Previously, the cost was composed of several resources: GPU, vCPU, RAM, and disk. For more details, see the [Dedicated Inference Payment Model and Pricing](/dedicated-inference/about/payment.mdx) instructions.
    </FaqAnswer>
  </FaqItem>

  <FaqItem>
    <FaqQuestion>July</FaqQuestion>

    <FaqAnswer>
      * added the ability to sort configurations when [creating an inference service](/dedicated-inference/create/create-inference-service.mdx). Now, at the configuration selection step, they can be sorted in ascending and descending order of price as well as performance;

      * Dedicated Inference is now available in the ru-6 pool. Also added the ability to select a [location](/infrastructure/locations.mdx) in the model catalog when [creating an inference service](/dedicated-inference/create/create-inference-service.mdx);

      * added the ability to connect to an inference service from a private network. For more details, see the [Inference service connection types](/dedicated-inference/create/connection-types.mdx) instructions.
    </FaqAnswer>
  </FaqItem>

  <FaqItem>
    <FaqQuestion>June</FaqQuestion>

    <FaqAnswer>
      * added new models:

        * [google/gemma-4-12B](https://huggingface.co/google/gemma-4-12B);
        * [unsloth/Qwen3.6-35B-A3B-NVFP4](https://huggingface.co/unsloth/Qwen3.6-35B-A3B-NVFP4);
        * [unsloth/Qwen3.6-27B-NVFP4](https://huggingface.co/unsloth/Qwen3.6-27B-NVFP4).

        You can view the current list of available models in the [Control Panel](https://my.selectel.ru/ml/default/models-catalog/): in the top menu, click **Products** → **Dedicated Inference**.
    </FaqAnswer>
  </FaqItem>

  <FaqItem>
    <FaqQuestion>May</FaqQuestion>

    <FaqAnswer>
      * added new models:

        * [openai/gpt-oss-120b](https://huggingface.co/openai/gpt-oss-120b);
        * [Qwen/Qwen3.6-35B-A3B](https://huggingface.co/Qwen/Qwen3.6-35B-A3B);
        * [Qwen/Qwen3.6-27B](https://huggingface.co/Qwen/Qwen3.6-27B);
        * [Qwen/Qwen3-Embedding-8B](https://huggingface.co/Qwen/Qwen3-Embedding-8B);
        * [Qwen/Qwen3-Reranker-8](https://huggingface.co/Qwen/Qwen3-Reranker-8B);
        * [deepseek-ai/DeepSeek-OCR-2](https://huggingface.co/deepseek-ai/DeepSeek-OCR-2);
        * [zai-org/GLM-ASR-Nano-2512](https://huggingface.co/zai-org/GLM-ASR-Nano-2512);
        * [openai/whisper-large-v3](https://huggingface.co/openai/whisper-large-v3);
        * [google/gemma-4-31B](https://huggingface.co/google/gemma-4-31B);
        * [t-tech/T-lite-it-2.1](https://huggingface.co/t-tech/T-lite-it-2.1);
        * [t-tech/T-pro-it-2.1](https://huggingface.co/t-tech/T-pro-it-2.1);
        * [intfloat/multilingual-e5-large](https://huggingface.co/intfloat/multilingual-e5-large);
        * [BAAI/bge-m3](https://huggingface.co/BAAI/bge-m3);
        * [BBAAI/bge-reranker-v2-m3](https://huggingface.co/BAAI/bge-reranker-v2-m3).

        You can view the current list of available models in the [Control Panel](https://my.selectel.ru/ml/default/models-catalog/): in the top menu, click **Products** → **Dedicated Inference**.
    </FaqAnswer>
  </FaqItem>

  <FaqItem>
    <FaqQuestion>April</FaqQuestion>

    <FaqAnswer>
      * moved Dedicated Inference from the private preview stage to the public preview stage.
    </FaqAnswer>
  </FaqItem>
</FaqWrapper>

<Formbricks />
