GLM-5.3 Flash on Novita AI: Native Multimodal Coding for Long-Horizon Agents
Explore GLM-5.3 Flash on Novita AI: model ID, 1M context, 128K output, multimodal input, pricing, and where it fits for coding agents.
Explore GLM-5.3 Flash on Novita AI: model ID, 1M context, 128K output, multimodal input, pricing, and where it fits for coding agents.
Explore Ling-3.0-flash-Fin on Novita AI: a finance-enhanced 124B MoE with 256K context, 32K max output, and $0 per-million-token pricing.
Ling-3.0-flash-fast on Novita AI: verified model ID, current pricing, 256K context, and what public docs do and do not say versus base Ling-3.0-flash.
DeepSeek V4 Pro 0813 is the GA release on Novita AI with DSpark, MIT licensing, 1M context, and updated pricing.
Ling-3.0 Tiny on Novita AI: verified model ID, 7.9B MoE architecture, 256K context, function calling, and current launch pricing.
Try Macaron V1 Tall on Novita AI with 262K context, reasoning, function calling, and a time-limited free API for coding and agent workflows.
Use Ling-3.0-flash free on Novita AI: a 124B MoE model with 5.1B active parameters, 262K context, reasoning, and function calling.
Macaron V1 Venti is a 748B Mixture-of-LoRA flagship for coding, agents, and GenUI. See its specs, Novita AI availability, pricing, and best-fit use cases.
Hy3 is available free on Novita AI via serverless API. 295B MoE, 21B active parameters, 256K context, three reasoning modes, and $0 per token.
Step-by-step guide to configure CoBuddy (baidu/cobuddy) in Claude Code using Novita AI's OpenAI-compatible endpoint. API setup, pricing, and coding workflow tips.
Use DeepSeek in Claude Code via the V4 Flash API on Novita AI. Set four env vars, get 1M-token context, and cut costs 20x vs Claude Sonnet.
GLM 5.2 is available on Novita AI with 1M context, 128K max output, function calling, structured outputs, and serverless API access.
Kimi K2.7 Code is live on Novita AI with OpenAI-compatible chat API access, 256K context, tool calling, and multimodal inputs.
Nemotron 3 Nano 30B A3B is available on Novita AI as a Serverless LLM with OpenAI-compatible chat completions, 256K context, and pay-as-you-go token pricing.