Feature requests

Share your ideas about our platform

SANKALPA DEVKOTA 5 days ago

💡 Feature request

SANKALPA DEVKOTA 5 days ago

💡 Feature request

Support for newer PostgreSQL versions (17 and 18)

Nebius currently offers PostgreSQL 16 as the latest version in its Managed Service for PostgreSQL. Adding support for PostgreSQL 17 and 18 would significantly lower the barrier for teams migrating existing instances to Nebius and keep the platform aligned with the upstream release cadence.

Niklas 16 days ago

Managed Database

💡 Feature request

Support for newer PostgreSQL versions (17 and 18)

Niklas 16 days ago

Managed Database

💡 Feature request

Account Balance API endpoint for Nebius Token Factory

Please add an API endpoint to retrieve the current account balance for a Nebius Token Factory account

Ivan Nikitin 20 days ago

Billing

💡 Feature request

Account Balance API endpoint for Nebius Token Factory

Please add an API endpoint to retrieve the current account balance for a Nebius Token Factory account

Ivan Nikitin 20 days ago

Billing

💡 Feature request

Support of Kimi K2.6 model?

Hi, Do you have any plans on supporting the Kimi K2.6 model?

Jonas 26 days ago

💡 Feature request

Support of Kimi K2.6 model?

Hi, Do you have any plans on supporting the Kimi K2.6 model?

Jonas 26 days ago

💡 Feature request

Offer separate short-term and long-term support LLM endpoints

I understand that there might be different needs among users of the token factory: build stable applications with reliable LLM performance → requires long-term support (LTS) for the model endpoint (even if the model is somewhat outdated) use the latest and greatest model as soon as it has been released, because it offers the best performance per dollar/euro → only needs short-term support (STS) for the model endpoint Public model endpoints should clearly state “STS” or “LTS”. This way the users can pick depending on individual needs. On Nebius’ side this would allow superseding STS model endpoints without further notice as soon as a superior version of that model gets released - ultimately offering a better service and keeping the “zoo” of model endpoints manageable in the long run. LTS model endpoints on the other hand could have a predetermined EOL date, so developers can plan for necessary model updates.

Jan Schaller about 1 month ago

💡 Feature request

Offer separate short-term and long-term support LLM endpoints

Jan Schaller about 1 month ago

💡 Feature request

Support Implicite/Explicite Prompt Caching

Many other providers offer prompt-caching with discounted pricing for cache hits, either implicitly (e.g., OpenAI, DeepSeek, Google, DeepInfra, NovitaAI, Fireworks) or explicitly (most notably Anthropic). This capability can significantly reduce costs in agentic workflows, where a single session often re-sends the same context repeatedly (for example, when the model performs multiple tool calls in sequence and the shared conversation/context is included each time). Today, Nebius Token Factory is at a cost disadvantage in these repeated, input-token-heavy scenarios compared to providers that support prompt caching and pass the savings through to customers. Please add support for prompt caching (implicit or explicit), including discounted pricing for cached prompt tokens, to improve cost-efficiency for agentic and tool-using applications.

Lukas Kreussel 3 months ago

Billing

💡 Feature request

Support Implicite/Explicite Prompt Caching

Lukas Kreussel 3 months ago

Billing

💡 Feature request

Support Latest Top 5 LLM Models

Out of the top 5 models (based on LLM and artificalanalysis) token factory still lacks: - MiniMax M2.5 - GLM 5 - Qwen3.5-397B-A17B - Step 3.5 Flash Additionally MiMo-V2-Flash would be welcome as Step 3.5 Flash only supports 64k token window. I’m quite confident that providing these models in the token factory catalog would provide users frontier proprietary level models, which would greatly help adoption. Additional note: On the main website (nebius.com) when someone hovers over the token factory menu the models that appear in the popup are all relatively old. K2.5 is already provided in token factory, but only K2 is listed. (1) https://llm-stats.com/leaderboards/open-llm-leaderboard (SWE-bench Verified) (2) https://artificialanalysis.ai/models/open-source (intelligence ranking)

davidhidvegi 3 months ago

💡 Feature request

Support Latest Top 5 LLM Models

davidhidvegi 3 months ago

💡 Feature request

/v1/responses OpenAI compatible endpoint

The /v1/responses endpoint from OpenAI is used by tools like Kilo Code and Opencode, and it is particularly useful for agent development and handling large files. I noticed that Nebius does not currently support this, but I believe it would be a valuable addition. Maintaining compatibility with the official OpenAI API would enable Nebius to serve as a backend for tools like Opencode and similar applications.

Roberto Sánchez 4 months ago

💡 Feature request

/v1/responses OpenAI compatible endpoint

Roberto Sánchez 4 months ago

💡 Feature request

Larger FAST models with SO.

Currently SO by constrained decoding is supported by large models (>80B) only with tps <50. This is insanely slow for agentic architectures based on SGR, and you keep slowing down inference. Please provide at least one fast, large, good model with SO

Ivan Matveev 4 months ago

💡 Feature request

Larger FAST models with SO.

Ivan Matveev 4 months ago

💡 Feature request

Support of PostGIS

Hello, I noticed that your PostgresSQL managed service supports many extensions, but PostGIS isn’t listed (https://docs.nebius.com/postgresql/databases/extensions). Would it be possible to add support for PostGIS (https://postgis.net/)? Thanks!

Baptiste Morel-Lab 4 months ago

Managed Database

💡 Feature request

Support of PostGIS

Baptiste Morel-Lab 4 months ago

Managed Database

💡 Feature request

Be nimble with new open source models that lead the benchmarks: Support Kimi K2.5 in Token factory

If a new open weights / open source models drops and it leads the benchmarks (look https://artificialanalysis.ai/models in the open section), be nimble and quick to offer it in the token factory

micu voilà 4 months ago

💡 Feature request

Be nimble with new open source models that lead the benchmarks: Support Kimi K2.5 in Token factory

If a new open weights / open source models drops and it leads the benchmarks (look https://artificialanalysis.ai/models in the open section), be nimble and quick to offer it in the token factory

micu voilà 4 months ago

💡 Feature request

Token Factory API: Lowercase model IDs

In the Token Factory API, the model selector should be able to be specified as lowercase. For example: Required today: {"model":"Qwen/Qwen3-Coder-480B-A35B-Instruct", … } Wanted: {"model":"qwen/qwen3-coder-480b-a35b-instruct", … } Why? Currently, OpenCode does not work with any Nebius model, even if OpenCode have a built-in integration with Nebius. Nothing works. Any request says “Not found”, and that is an error from your API because they (mistakenly) force lowercase on all model IDs. If I edit the OpenCode config file and Pascal Case the Nebius model IDs, then your API accepts it. This is also easy to test with curl. This is really an OpenCode issue, but maybe you could also edit your API and allow lowercased IDs, for us developers (customers) sake 🙏 Regards, Christoffer

uninjured2875 5 months ago

Network

💡 Feature request

Token Factory API: Lowercase model IDs

uninjured2875 5 months ago

Network

💡 Feature request

Billing form, support Mongolia

On the billing form, I’d like to request Mongolia to be listed.

Sambalkhundev Soyolerdene 5 months ago

💡 Feature request

Billing form, support Mongolia

On the billing form, I’d like to request Mongolia to be listed.

Sambalkhundev Soyolerdene 5 months ago

💡 Feature request

Serverless Lora for Gemma3

Would it be possible to support fine-tuning and serverless LoRA for Gemma3-9b/27b?

paul mat 6 months ago

💡 Feature request

Serverless Lora for Gemma3

Would it be possible to support fine-tuning and serverless LoRA for Gemma3-9b/27b?

paul mat 6 months ago

💡 Feature request

Cloud DNS

Hello! I would like to be able to use Cloud DNS, primarily internal DNS for internal resources and for the team that accesses internal resources via VPN (i.e. there should be a way to specify DNS servers).

Sergei Turguzov 6 months ago

Other

💡 Feature request

Cloud DNS

Sergei Turguzov 6 months ago

Other

💡 Feature request

Add Country Ghana to the Country list

Under the Billing setup, there is no option for my country Ghana.

Justin Johnson 7 months ago

💡 Feature request

Add Country Ghana to the Country list

Under the Billing setup, there is no option for my country Ghana.

Justin Johnson 7 months ago

💡 Feature request

Kimi k2 thinking model

when are you guys planning to host the kimi k2 thinking model on token factory?

Ayush Awasthi 7 months ago

💡 Feature request

Kimi k2 thinking model

when are you guys planning to host the kimi k2 thinking model on token factory?

Ayush Awasthi 7 months ago

💡 Feature request

Standalone applications support in OpenTofu/Terraform

Currently only Applications for Managed Service for Kubernetes has OpenTofu/Terraform support https://docs.nebius.com/terraform-provider/reference/resources/applications_v1alpha1_k8s_release. It would be great to have OpenTofu/Terraform support for standalone applications such as vLLM as well. Thanks! 😃

hongbo-miao 9 months ago

Other

💡 Feature request

Standalone applications support in OpenTofu/Terraform

hongbo-miao 9 months ago

Other

💡 Feature request

Cilium Gateway API in Managed Kubernetes

Managed Kubernetes is using Cilium currently. I hope it allows us to use Cilium Gateway API. Note based on here: One of the biggest differences between Cilium’s Ingress and Gateway API support and other Ingress controllers is how closely tied the implementation is to the CNI. For Cilium, Ingress and Gateway API are part of the networking stack, and so behave in a different way to other Ingress or Gateway API controllers (even other Ingress or Gateway API controllers running in a Cilium cluster). Other Ingress or Gateway API controllers are generally installed as a Deployment or Daemonset in the cluster, and exposed via a Loadbalancer Service or similar (which Cilium can, of course, enable). Cilium’s Ingress and Gateway API config is exposed with a Loadbalancer or NodePort service, or optionally can be exposed on the Host network also. But in all of these cases, when traffic arrives at the Service’s port, eBPF code intercepts the traffic and transparently forwards it to Envoy (using the TPROXY kernel facility). This affects things like client IP visibility, which works differently for Cilium’s Ingress and Gateway API support to other Ingress controllers. It also allows Cilium’s Network Policy engine to apply CiliumNetworkPolicy to traffic bound for and traffic coming from an Ingress. Nebius support informed me that customizing add-ons is not currently supported, but they plan to enable it in the future. For now, I can modify the cilium-config ConfigMap as described in https://docs.nebius.com/kubernetes/networking/add-ons#cilium and try enabling the Gateway API, with the understanding that these changes might be reverted during Nebius upgrades. It would be great to officially support customization of add-ons expose as a parameter at OpenTofu/Terraform https://docs.nebius.com/terraform-provider/reference/resources/mk8s_v1_cluster so we can have choice and also controls in the code Thanks! 😃

hongbo-miao 9 months ago

Managed Service for Kubernetes®

💡 Feature request

Cilium Gateway API in Managed Kubernetes

hongbo-miao 9 months ago

Managed Service for Kubernetes®

💡 Feature request

Manage Nebius custom groups via OpenTofu/Terraform

Currently, creating and managing custom groups in Nebius is only possible via the CLI. It would be great to have OpenTofu/Terraform resources to create custom groups, thanks! ☺️

hongbo-miao 9 months ago

IAM

💡 Feature request

Manage Nebius custom groups via OpenTofu/Terraform

Currently, creating and managing custom groups in Nebius is only possible via the CLI. It would be great to have OpenTofu/Terraform resources to create custom groups, thanks! ☺️

hongbo-miao 9 months ago

IAM

💡 Feature request

Boards

Most helpful