Novita AI

Novita AI · Launch Video Breakdown: Hook, Pacing & Motion Design

GLM-5.3-Flash is live on Novita as a day-zero partner: a 320B multimodal model with 18B active parameters and 1M-token context.

AI AgentsLaunchAugust 26, 2026@novita_labs
0:00 · The Hook · Model Discovery & Pricing Overview
0:00 / 0:00

Scene-by-scene timeline & spoken transcript

  1. The Hook

    Model Discovery & Pricing Overview

    “(No spoken dialogue — The track opens with a sustained, evolving synth pad creating an ambient atmosphere. At 00:01, a high-energy audio transition begins with a rising synth sweep, culminating in a sub-bass drop/impact at 00:02.)”

    On screen
    Novita Model APIs Agent Sandbox GPUs Resources Pricing Dashboard Search Model Featured All Models LLM Serverless Image Audio Video Embedding Ranker AI Search Vision Model Series Models (146) Z NEW GLM 5.3 Flash $0.15/Mt Input 1M Context $0.03/Mt Cache Read 128K Max Output $0.5/Mt Output DeepSeek V4 Pro 0813 NEW $1.32/Mt Input 1M Context $0.132/Mt Cache Read 384K Max Output $3.96/Mt Output Kimi K3 NEW $3/Mt Input $0.3/Mt Output 15/Mt 1M Context Max Output Hy3 NEW $0.14/Mt Input 256K Context $0.035/Mt Cache Read 256K Max Output $0.58/Mt Output LLM SERVERLESS LLM SERVERLESS LLM SERVERLESS LLM SERVERLESS Z NEW GLM 5.2 Kimi K2.7 Code NEW Deepseek V4 Flash 0731 NEW Macaron V1 Venti NEW
    Camera
    Static wide shot of a web application dashboard, followed by a subtle zoom-in on the 'GLM 5.3 Flash' model card.
    Motion
    Subtle camera zoom, mouse cursor interaction (hover and click).
  2. Product Reveal

    Detailed Model Information & API Access

    “(No spoken dialogue — A rhythmic, arpeggiated synth melody with a driving sub-bass line and subtle percussion begins.)”

    On screen
    GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its attention architecture maintains accurate long-context behavior while reducing compute overhead. Features Serverless API zai-org/glm-5.3-flash is available via Novita's serverless API, where you pay per token. There are several ways to call the API, inc endpoints with exceptional reasoning performance. Available Serverless Run queries immediately, pay only for usage Input $0.15 / M Tokens Cache Read $0.03 / M Tokens Output $0.5 / M Tokens Docs Info Provider Zai-org Quantization fp8 Supported Functionality Context Length 1M Max Output 128K Serverless Supported Function Calling Supported Structured Output Supported Reasoning Supported Anthropic API Supported Try Now
    Camera
    Smooth vertical scroll down the webpage, revealing more details about the GLM-5.3-Flash model.
    Motion
    Smooth scrolling, mouse cursor interaction (click on 'Try Now' button).
  3. Feature Teaser

    Loading & Playground Entry

    “(No spoken dialogue — The rhythmic synth melody continues.)”

    On screen
    Loading models... Z Try GLM 5.3 Flash Kick the tires, see how GLM 5.3 Flash performs on Novita AI Say something... Upload Files Enable Thinking
    Camera
    Transition to a clean, white loading screen, then a new interface for the LLM playground.
    Motion
    Full screen wipe/cut, loading spinner animation, UI element reveal.
  4. Product Reveal

    LLM Playground Interaction

    “(No spoken dialogue — The rhythmic synth melody continues. Another rising synth sweep occurs at 00:09, leading to a final sub-bass impact at 00:10.)”

    On screen
    Novita LLM Playground Model APIs Model Library LLM LLM Playground Deployments Metrics Usage Logs MULTIMODAL Multimodal Playground Dedicated Endpoints Model Management Settings Quick Access Billing Pricing Docs Model Configuration Response format text System Prompt Be a helpful assistant max_tokens 65536 temperature 1 top_p 1 min_p 0 top_k 50 presence_penalty 0 frequency_penalty 0 Z GLM 5.3 Flash zai-org/glm-5.3-flash chat Model Detail View Code Clear History Z Try GLM 5.3 Flash Kick the tires, see how GLM 5.3 Flash performs on Novita AI Say something... Upload Files Enable Thinking Tell me an interesting fact about the universe!
    Camera
    Static shot of the LLM Playground interface, with a mouse cursor typing into the input field.
    Motion
    Typing animation, mouse cursor interaction.
  5. Call to Action

    AI Response Generation

    “(No spoken dialogue — The track fades out with the sustained synth pad.)”

    On screen
    Novita LLM Playground Model APIs Model Library LLM LLM Playground Deployments Metrics Usage Logs MULTIMODAL Multimodal Playground Dedicated Endpoints Model Management Settings Quick Access Billing Pricing Docs Model Configuration Response format text System Prompt Be a helpful assistant max_tokens 65536 temperature 1 top_p 1 min_p 0 top_k 50 presence_penalty 0 frequency_penalty 0 Z GLM 5.3 Flash zai-org/glm-5.3-flash chat Model Detail View Code Clear History Tell me an interesting fact about the universe! Thinking... The user is asking for an interesting fact about the universe. This is a simple, friendly request. I should share genuinely fascinating and accurate about the universe. Let me think of some good candidates: The observable universe is about 93 billion light-years across, even though the universe is only ~13.8 billion years old ( There are more stars in the observable universe than grains of sand on all Earth's beaches. We're made of "star stuff" - the elements Say something... Upload Files Enable Thinking
    Camera
    Static shot of the LLM Playground, showing the AI generating a response.
    Motion
    Typing/text reveal animation for the AI's response.

Related AI Agents Product Launches

Explore all AI Agents launches →
OpenAI
Hook 9.2137.9M
OpenAIAI Agents

To ensure that artificial general intelligence benefits all of humanity

@OpenAI
Gopuff
Hook 9.277.6M
GopuffAI Agents

Gopuff introduces Go, an AI shopping assistant built with SpaceXAI: say what you need and the order is placed.

@gopuff
Elon Musk
Hook 9.263.3M
Elon MuskAI Agents

Iliad (Troy) trailer made by Grok Imagine 1.5, which was just released

@elonmusk