Ling-3.0-Flash-VL on Novita AI: Native Multimodal API and Pricing
Explore Ling-3.0-Flash-VL on Novita AI: native image and video input, 262K context, 32K output, function calling, and current $0 pricing.
Explore Ling-3.0-Flash-VL on Novita AI: native image and video input, 262K context, 32K output, function calling, and current $0 pricing.
Explore Ling-3.0-Flash-Sante as a medical AI API on Novita AI, with model specs, current pricing, 256K context, reasoning, and function calling.
Explore GLM-5.3 Flash on Novita AI: model ID, 1M context, 128K output, multimodal input, pricing, and where it fits for coding agents.
Explore Ling-3.0-flash-Fin on Novita AI: a finance-enhanced 124B MoE with 256K context, 32K max output, and $0 per-million-token pricing.
Ling-3.0-flash-fast on Novita AI: verified model ID, current pricing, 256K context, and what public docs do and do not say versus base Ling-3.0-flash.
DeepSeek V4 Pro 0813 is the GA release on Novita AI with DSpark, MIT licensing, 1M context, and updated pricing.
Ling-3.0 Tiny on Novita AI: verified model ID, 7.9B MoE architecture, 256K context, function calling, and current launch pricing.
Try Macaron V1 Tall on Novita AI with 262K context, reasoning, function calling, and a time-limited free API for coding and agent workflows.
Use Ling-3.0-flash free on Novita AI: a 124B MoE model with 5.1B active parameters, 262K context, reasoning, and function calling.
Macaron V1 Venti is a 748B Mixture-of-LoRA flagship for coding, agents, and GenUI. See its specs, Novita AI availability, pricing, and best-fit use cases.
Hy3 is available free on Novita AI via serverless API. 295B MoE, 21B active parameters, 256K context, three reasoning modes, and $0 per token.
Step-by-step guide to configure CoBuddy (baidu/cobuddy) in Claude Code using Novita AI's OpenAI-compatible endpoint. API setup, pricing, and coding workflow tips.
Use DeepSeek in Claude Code via the V4 Flash API on Novita AI. Set four env vars, get 1M-token context, and cut costs 20x vs Claude Sonnet.
GLM 5.2 is available on Novita AI with 1M context, 128K max output, function calling, structured outputs, and serverless API access.