BatchIn
Products

产品矩阵

模型 API (全速推理网关)视频与多模态 APIVaaS (密码学推理存证)独享算力集群
ModelsPricing创作展厅Docs活动中心平台对比
Start self-serveLoginContact中/EN
BATCHIN 模型中心 · 开源与 MOE 推理矩阵 · 兼容 OPENAI API实时基准评测 · 密码学验签存证
开源与 MoE 模型矩阵 · 高吞吐网关

前沿大模型与推理中心

全面支持 Qwen 3.8 Max、DeepSeek-V4、GLM-5.3、Kimi K3、MiniMax M3、豆包 Seedance 2.5、可灵 Kling v3、万相 Wan 3.0,提供 100% OpenAI 兼容端点与端到端密码学验签存证。

展示 23 个线上生产级模型
DeepSeek

DeepSeek

deepseek-v4-pro

DeepSeek V4 Pro

Flagship Reasoning

DeepSeek next-gen MoE reasoning flagship for complex architectural design, autonomous coding loops, and math proofs.

Context
256K
Max Output
64K
Public Pricing
/

Cached:

Proprietaryfeatureddeepseekreasoning
View Model & Endpoints
DeepSeek

DeepSeek

deepseek-v4-flash

DeepSeek V4 Flash

Fast Agent Route

High-throughput, ultra-low latency DeepSeek model optimized for rapid tool calls and high-frequency agent loops.

Context
256K
Max Output
64K
Public Pricing
/

Cached:

Proprietaryfeatureddeepseekfast
View Model & Endpoints
DeepSeek

DeepSeek

deepseek-v4.1-flash

DeepSeek V4.1 Flash

Fast Reasoning

Upgraded 2026 DeepSeek lightweight model with enhanced reasoning density, sub-100ms first-token latency, and ultra-high throughput.

Context
256K
Max Output
64K
Public Pricing
/

Cached:

Proprietaryfeatureddeepseekfast
View Model & Endpoints
Qwen

Qwen

qwen3.8-max

Qwen 3.8 Max

2.4T MoE Flagship

Alibaba 2.4-trillion-parameter MoE flagship model with native 1M context, state-of-the-art autonomous coding, and comprehensive multimodal reasoning. Supports deterministic 0902 snapshot alias.

Context
1M
Max Output
32K
Public Pricing
/

Cached:

Proprietaryfeaturedqwenflagship
View Model & Endpoints
Tencent

Tencent

hy4-preview

Hunyuan 4 Preview

Dual-Speed MoE

Tencent Hunyuan 4 next-generation MoE frontier model with dual-speed reasoning, deep Chinese semantic comprehension, and enterprise agent tool integration.

Context
256K
Max Output
32K
Public Pricing
/

Cached:

Proprietaryfeaturedtencentreasoning
View Model & Endpoints
Z.ai

AI

glm-5.3-flash

GLM-5.3 Flash

High Throughput

High-speed lightweight GLM 5.3 model designed for real-time dialogue, extraction, and high-frequency agent tool execution.

Context
128K
Max Output
16K
Public Pricing
/

Cached:

Proprietaryfeaturedzhipufast
View Model & Endpoints
Qwen

Qwen

wan3.0-video

Wan 3.0 Video

Flagship Video

Alibaba Wan 3.0 next-generation cinematic video generation model with photorealistic lighting, physics, and camera controls.

Context
Prompt & Keyframe
Max Output
1080p Video
Public Pricing
/ second
Proprietaryfeaturedvideoqwen
View Model & Endpoints
Z.ai

AI

glm-5.3

GLM-5.3

Agent & Tools

Zhipu AI flagship model with upgraded reasoning, deep tool invocation capabilities, and Chinese enterprise alignment.

Context
256K
Max Output
32K
Public Pricing
/

Cached:

Proprietaryfeaturedzhipuagent
View Model & Endpoints
Kimi / Moonshot AI

Kimi / Moonshot AI

kimi-k3

Kimi K3

2M Context

Moonshot AI flagship with ultra-long 2M lossless context, multi-document cross-referencing, and deep research synthesis.

Context
2M
Max Output
64K
Public Pricing
/

Cached:

Proprietaryfeaturedkimilong-context
View Model & Endpoints
Kimi / Moonshot AI

Kimi / Moonshot AI

kimi-k2.7-code

Kimi K2.7 Code

Developer Tier

Specialized coding intelligence model from Moonshot AI with deep understanding of complex frameworks and full-stack tasks.

Context
256K
Max Output
32K
Public Pricing
/

Cached:

Proprietarycodingkimideveloper
View Model & Endpoints
MiniMax

MiniMax

minimax-m3

MiniMax M3

Multimodal LLM

MiniMax flagship language and multimodal reasoning engine with native 512K context and expressive text generation.

Context
512K
Max Output
32K
Public Pricing
/

Cached:

Proprietaryfeaturedminimaxmultimodal
View Model & Endpoints
ByteDance

ByteDance

doubao-seedance-2.0

Doubao Seedance 2.0

Cinematic Flagship

ByteDance flagship video generation model offering photorealistic cinematic camera movement and motion coherence.

Context
Prompt & Keyframe
Max Output
1080p Video
Public Pricing
/ second
Proprietaryvideodoubaoflagship
View Model & Endpoints
ByteDance

ByteDance

doubao-seedance-2-0-260128

Doubao Seedance 2.0 Turbo

Turbo Video

Accelerated Seedance 2.0 video generation engine with high rendering speed and crisp visual consistency.

Context
Prompt & Keyframe
Max Output
1080p Video
Public Pricing
/ second
Proprietaryvideodoubaofast
View Model & Endpoints
ByteDance

ByteDance

doubao-seedance-2-0-mini-260615

Doubao Seedance 2.0 Mini

Lightweight Video

Cost-optimized lightweight video generation endpoint tailored for social media clips and preview generations.

Context
Prompt & Keyframe
Max Output
720p Video
Public Pricing
/ second
Proprietaryvideodoubaoeconomy
View Model & Endpoints
ByteDance

ByteDance

doubao-seedance-2-0-fast-260128

Doubao Seedance 2.0 Fast

Fast Lane

Low-latency video rendering endpoint with consistent motion dynamics for real-time video workflows.

Context
Prompt & Keyframe
Max Output
720p/1080p Video
Public Pricing
/ second
Proprietaryvideodoubaofast
View Model & Endpoints
ByteDance

ByteDance

seedance2.5

ByteDance Seedance 2.5

SOTA Video

ByteDance frontier Seedance 2.5 video generation model with temporal consistency and complex scene choreography.

Context
Prompt & Keyframe
Max Output
1080p Video
Public Pricing
/ second
Proprietaryvideodoubaoflagship
View Model & Endpoints
ByteDance

ByteDance

seedance2.5-huoshan

Seedance 2.5 Huoshan

Direct Huoshan Lane

Official Huoshan direct enterprise channel for Seedance 2.5 with guaranteed compute rendering queues and zero concurrency drops.

Context
Prompt & Keyframe
Max Output
1080p Video
Public Pricing
/ second
Proprietaryvideodoubaoenterprise
View Model & Endpoints
Kling AI

Kling AI

kling-v3

Kling v3

Physical Dynamics

Kuaishou Kling v3 video generation model featuring complex physics simulation and fluid real-world mechanics.

Context
Prompt & Keyframe
Max Output
1080p Video
Public Pricing
/ second
Proprietaryvideoklingflagship
View Model & Endpoints
Kling AI

Kling AI

kling-v3-omni

Kling v3 Omni

Audio-Visual Omni

Omni-modal video model synthesizing synchronized sound effects and motion pictures in one unified pass.

Context
Prompt & Keyframe
Max Output
1080p Video with Audio
Public Pricing
/ second
Proprietaryvideoklingomni
View Model & Endpoints
Kling AI

Kling AI

kling-video-o1

Kling Video o1

Reasoning Enhanced

Reasoning-guided video generator ensuring narrative consistency and multi-shot continuity.

Context
Prompt & Sequence
Max Output
1080p Video
Public Pricing
/ second
Proprietaryvideoklingreasoning
View Model & Endpoints
MiniMax

MiniMax

minimax-h3

MiniMax H3 Video

Digital Human

MiniMax H3 video synthesis engine specializing in realistic character performance and natural facial expressions.

Context
Prompt & Image
Max Output
1080p Video
Public Pricing
/ second
Proprietaryvideominimaxavatar
View Model & Endpoints
ByteDance

ByteDance

doubao-seedream-5-0-260128

Doubao Seedream 5.0

Latest Flagship Image

ByteDance latest Seedream 5.0 high-resolution image synthesis engine with exceptional prompt fidelity and texture detail.

Context
Text Prompt
Max Output
2K / 4K Image
Public Pricing
/ image
Proprietaryimagedoubaoflagship
View Model & Endpoints
Qwen

Qwen

qwen-image-3.0-pro

Qwen Image 3.0 Pro

Typography & Layout

Alibaba Qwen Image 3.0 Pro model specialized in complex visual layout composition and text rendering.

Context
Text Prompt
Max Output
2K Image
Public Pricing
/ image
Proprietaryimageqwenpro
View Model & Endpoints
1-Minute Drop-In Integration

Zero Codebase Refactoring. Point baseURL to Switch.

100% compatible with OpenAI API specs. Drop straight into Next.js, LangChain, Cursor IDE, and LlamaIndex.

app/api/chat/route.tstypescript
import { createOpenAI } from '@ai-sdk/openai';
import { streamText } from 'ai';

// Drop-in replace baseURL & apiKey
const batchin = createOpenAI({
  baseURL: 'https://api.batchin.tech/v1',
  apiKey: process.env.BATCHIN_API_KEY,
});

export async function POST(req: Request) {
  const { messages } = await req.json();

  const result = streamText({
    model: batchin('deepseek-v4-pro'),
    messages,
  });

  return result.toDataStreamResponse();
}

Ready To Control AI Inference?

BatchIn brings model access, cost controls, signed records, and dedicated capacity into one workspace.

Contact teamGet Access

Models

  • 模型广场查看全部可用模型
  • 定价方案查看最新公开费率
  • 开发演练场在线测试与接口调试
  • API 文档接入指南与代码示例

Products

  • 模型 API (全速推理网关)
  • 视频与多模态 API
  • VaaS (密码学推理存证)
  • 独享算力集群

Docs

  • Documentation & API
  • Compare
  • Events
  • WIN Incubator
  • Use cases

Support

  • Contact
  • hello@luminapath.ai
  • System status: operational
BatchIn© 2026 BatchIn. All rights reserved
PrivacyTerms

All model trademarks, product names, logos, and brands are property of their respective owners and used solely for compatibility and identification purposes.