Skip to main content

Quick Search & Navigation Refresh

A major overhaul of the console’s command palette and navigation experience. The Quick Search dialog now supports multi-page flows, improved resource discovery, and an integrated theme switcher—all accessible via Cmd+K.

Features

  • Unified create actions. All resource creation flows are now accessible directly from the Cmd+K quick search
  • Enhanced Quick Search. Multi-page flow with improved resource search and theme switcher built in
  • Duplicate-name validation. Prevents naming conflicts across projects, VPCs, security groups, filesystems, and images
  • Batch instance creation. Create multiple instances at once with public IP toggle support
  • Image snapshots. Create image snapshots directly from instance detail views
  • Filesystem improvements. VPC multi-select in creation/editing with improved UI
  • Navigation updates. New tabs across serverless details, usage & billing, Kubernetes details, fine-tuning overview, and project compute clusters
  • Security groups. Add actions in instance view with rule side panel for create/edit
  • Jobs enhancements. Parameters field on create with AI summary/diagnosis for failed jobs
  • Dashboard quotas. Private dashboard now displays available quotas

Nscale CLI

  • Device CLI login. New device authentication flow with refresh token and expires-at fields

Instances & Networking MVP

The console now includes a complete instances and networking workflow. Create and manage instances with full security group support, storage side panels, and quota visibility.

Features

  • Instances + networking MVP. Full instance actions and security group management
  • Storage create side panel. Streamlined storage creation with improved details view
  • Quota callout. Clear visibility into resource quotas
  • Instance create polish. Refined instance creation flow and UI improvements
  • Device status. Updated device status display

Instances v2 & UI Refresh

A comprehensive refresh of the console UI alongside the new Instances v2 experience. New list and detail views, storage cards, usage visualisation, and a broad refresh of core components.

Features

  • Instance list & detail views. New views with create image side panel
  • Storage details cards. Improved storage details page with card layout
  • Usage visualisation. New bar chart for usage metrics
  • UI component refresh. Updated form components, drawers, card tables, and tooltips
  • Instances v2. List pages for supporting resources
  • API keys removed. API key management moved out of console
  • Workbench polish. UI improvements to the Workbench experience
  • Design system alignment. Button styling aligned to design system

Dedicated Infrastructure & Workbench Upgrades

New dedicated infrastructure nodes addon and significant Workbench improvements including model selection fixes, preset rules, and comparison-mode enhancements.

Features

  • Dedicated infrastructure nodes. New addon for dedicated infrastructure
  • Workbench upgrades. Model selection/validation fixes, preset rules, comparison-mode improvements, and license display
  • Job status handling. Improved job completion status handling
  • Workload pool IPs. Machine IP display and copy improvements
  • Region validation. Region change validation fixes
  • Image selector. Improved image selection experience
  • Preset baseline. Now updates to current version automatically

Serverless Fine-tuning

Serverless Fine-tuning
Introducing Serverless Fine-tuning – the fastest way to customise open-weight foundation models without touching infrastructure. Spin up a secure, pay-as-you-go training job in a single API call, watch metrics stream in real-time, and download your tuned model or push straight to Hugging Face.

Features

  • Two-step workflow. Pick any supported base model (Llama 3, Mistral 7B, DeepSeek, Qwen and more) and launch a job with your dataset – no cluster sizing, no Dockerfiles
  • LoRA-powered efficiency. Default Low-Rank-Adaptation (LoRA) reduces GPU hours and cost
  • Live metrics & easy monitoring. Poll one endpoint to track train_loss, eval_loss, perplexity
  • Export or deploy instantly. One-click push to Hugging Face or direct download of a ready-to-serve artefact
  • Serverless pricing. $2 minimum per job, billed by processed tokens; every new account still gets $5 free credit to experiment

Quick start

Serverless Inference

Serverless Inference
Our fully-managed, pay-per-request runtime that puts a pool of GPUs behind a single OpenAI-compatible endpoint. Instead of capacity planning, container images and infra dashboards, you call https://inference.api.nscale.com/v1/* and get deterministic, low-latency responses from today’s best open-source models—all billed per token and delivered from data-sovereign, 100% renewable data-centres.

Features

  • OpenAI-compatible endpoints. Drop-in support for Llama, Qwen, DeepSeek and other leading models makes migration a copy-paste job
  • Pay-as-you-go billing. Prices are per 1 million tokens including input and output tokens for Chat, Multimodal, Language and Code models. Image models is based on image size and steps
  • 80% lower cost & 100% renewable. Our vertically-integrated stack slashes TCO versus hyperscalers while guaranteeing data privacy—requests are never logged or reused
  • $5 free credits to get started. Every new account includes starter credits so you can ship to production in minutes

Under the hood

Quick start