BatchIn
Products

Products

Model APIVideo & Multimodal APIVaaS (Verification as a Service)Dedicated Capacity
ModelsPricingShowroomDocsEventsCompare
Start self-serveLoginContact
BATCHIN MODEL CATALOG • OPEN & MOE REASONING MATRIX • OPENAI-COMPATIBLE APIREALTIME BENCHMARKS • CRYPTOGRAPHIC RECEIPTS
Open & MoE Model Matrix · High-Throughput Gateway

Next-Gen Open & MoE Model Catalog

Run DeepSeek-V4, Qwen 3.7, GLM-5.3, Kimi K3, Doubao Seedance 2.0, and Kling-v3 via 100% OpenAI-compatible endpoints with cryptographic proof.

Showing 23 production models
DeepSeek

DeepSeek

deepseek-v4-pro

Available

DeepSeek V4 Pro

Flagship Reasoning

DeepSeek next-gen MoE reasoning flagship for complex architectural design, autonomous coding loops, and math proofs.

Context
256K
Max Output
64K
Public Pricing
$0.56 / $1.11

Cached: $0.111

Proprietaryfeatureddeepseekreasoning
View Model & Endpoints
DeepSeek

DeepSeek

deepseek-v4-flash

Available

DeepSeek V4 Flash

Fast Agent Route

High-throughput, ultra-low latency DeepSeek model optimized for rapid tool calls and high-frequency agent loops.

Context
256K
Max Output
64K
Public Pricing
$0.10 / $0.20

Cached: $0.020

Proprietaryfeatureddeepseekfast
View Model & Endpoints
Qwen

Qwen

qwen3.7-max

Available

Qwen 3.7 Max

1M Context

Alibaba flagship reasoning model with native 1M context window and state-of-the-art benchmark capabilities.

Context
1M
Max Output
32K
Public Pricing
$0.50 / $1.50

Cached: $0.100

Proprietaryfeaturedqwenlong-context
View Model & Endpoints
Qwen

Qwen

qwen3.7-plus

Available

Qwen 3.7 Plus

High Throughput

Balanced workhorse model with 1M context, robust tool use, and enterprise-grade reliability.

Context
1M
Max Output
32K
Public Pricing
$0.37 / $1.49

Cached: $0.074

Proprietaryfeaturedqwenproduction
View Model & Endpoints
Qwen

Qwen

qwen3-coder-480b-a35b

Available

Qwen3 Coder 480B

Repository Coding

Large-scale 480B MoE code generation model tailored for whole-repo refactoring and complex bug fixing.

Context
256K
Max Output
64K
Public Pricing
$0.56 / $2.23

Cached: $0.111

Proprietarycodingqwenmoe
View Model & Endpoints
Qwen

Qwen

qwen3-coder-30b-a3b

Available

Qwen3 Coder 30B

IDE Autocomplete

Fast code completion model designed for IDE inline suggestions and interactive debugging.

Context
128K
Max Output
16K
Public Pricing
$0.15 / $0.45

Cached: $0.030

Proprietarycodingqwenfast
View Model & Endpoints
Qwen

Qwen

qwen3-embedding-8b

Available

Qwen3 Embedding 8B

Dense Vector

Dense semantic vector embedding model with high retrieval accuracy for enterprise RAG and semantic search.

Context
32K
Max Output
4096 dim
Public Pricing
$0.057 / $0.000
Proprietaryembeddingqwenrag
View Model & Endpoints
Z.ai

Zhipu AI

glm-5.3

Available

GLM-5.3

Agent & Tools

Zhipu AI flagship model with upgraded reasoning, deep tool invocation capabilities, and Chinese enterprise alignment.

Context
256K
Max Output
32K
Public Pricing
$0.74 / $2.60

Cached: $0.149

Proprietaryfeaturedzhipuagent
View Model & Endpoints
Kimi / Moonshot AI

Kimi / Moonshot AI

kimi-k3

Available

Kimi K3

2M Context

Moonshot AI flagship with ultra-long 2M lossless context, multi-document cross-referencing, and deep research synthesis.

Context
2M
Max Output
64K
Public Pricing
$1.86 / $9.29

Cached: $0.371

Proprietaryfeaturedkimilong-context
View Model & Endpoints
Kimi / Moonshot AI

Kimi / Moonshot AI

kimi-k2.7-code

Available

Kimi K2.7 Code

Developer Tier

Specialized coding intelligence model from Moonshot AI with deep understanding of complex frameworks and full-stack tasks.

Context
256K
Max Output
32K
Public Pricing
$0.60 / $2.50

Cached: $0.120

Proprietarycodingkimideveloper
View Model & Endpoints
MiniMax

MiniMax

minimax-m3

Available

MiniMax M3

Multimodal LLM

MiniMax flagship language and multimodal reasoning engine with native 512K context and expressive text generation.

Context
512K
Max Output
32K
Public Pricing
$0.40 / $1.60

Cached: $0.080

Proprietaryfeaturedminimaxmultimodal
View Model & Endpoints
ByteDance

ByteDance

doubao-seedance-2.0

Available

Doubao Seedance 2.0

Cinematic Flagship

ByteDance flagship video generation model offering photorealistic cinematic camera movement and motion coherence.

Context
Prompt & Keyframe
Max Output
1080p Video
Public Pricing
$0.046 / second
Proprietaryvideodoubaoflagship
View Model & Endpoints
ByteDance

ByteDance

doubao-seedance-2-0-260128

Available

Doubao Seedance 2.0 Turbo

Turbo Video

Accelerated Seedance 2.0 video generation engine with high rendering speed and crisp visual consistency.

Context
Prompt & Keyframe
Max Output
1080p Video
Public Pricing
$0.034 / second
Proprietaryvideodoubaofast
View Model & Endpoints
ByteDance

ByteDance

doubao-seedance-2-0-mini-260615

Available

Doubao Seedance 2.0 Mini

Lightweight Video

Cost-optimized lightweight video generation endpoint tailored for social media clips and preview generations.

Context
Prompt & Keyframe
Max Output
720p Video
Public Pricing
$0.022 / second
Proprietaryvideodoubaoeconomy
View Model & Endpoints
ByteDance

ByteDance

doubao-seedance-2-0-fast-260128

Available

Doubao Seedance 2.0 Fast

Fast Lane

Low-latency video rendering endpoint with consistent motion dynamics for real-time video workflows.

Context
Prompt & Keyframe
Max Output
720p/1080p Video
Public Pricing
$0.026 / second
Proprietaryvideodoubaofast
View Model & Endpoints
ByteDance

ByteDance

doubao-seedance-1-5-pro-251215

Available

Doubao Seedance 1.5 Pro

Proven Workhorse

Reliable 1.5 Pro video generation endpoint with extensive compatibility and high concurrency support.

Context
Prompt & Keyframe
Max Output
1080p Video
Public Pricing
$0.028 / second
Proprietaryvideodoubaostable
View Model & Endpoints
Kling AI

Kling AI

kling-v3

Available

Kling v3

Physical Dynamics

Kuaishou Kling v3 video generation model featuring complex physics simulation and fluid real-world mechanics.

Context
Prompt & Keyframe
Max Output
1080p Video
Public Pricing
$0.056 / second
Proprietaryvideoklingflagship
View Model & Endpoints
Kling AI

Kling AI

kling-v3-omni

Available

Kling v3 Omni

Audio-Visual Omni

Omni-modal video model synthesizing synchronized sound effects and motion pictures in one unified pass.

Context
Prompt & Keyframe
Max Output
1080p Video with Audio
Public Pricing
$0.061 / second
Proprietaryvideoklingomni
View Model & Endpoints
Kling AI

Kling AI

kling-video-o1

Available

Kling Video o1

Reasoning Enhanced

Reasoning-guided video generator ensuring narrative consistency and multi-shot continuity.

Context
Prompt & Sequence
Max Output
1080p Video
Public Pricing
$0.067 / second
Proprietaryvideoklingreasoning
View Model & Endpoints
MiniMax

MiniMax

minimax-h3

Available

MiniMax H3 Video

Digital Human

MiniMax H3 video synthesis engine specializing in realistic character performance and natural facial expressions.

Context
Prompt & Image
Max Output
1080p Video
Public Pricing
$0.039 / second
Proprietaryvideominimaxavatar
View Model & Endpoints
ByteDance

ByteDance

doubao-seedream-5-0-260128

Available

Doubao Seedream 5.0

Latest Flagship Image

ByteDance latest Seedream 5.0 high-resolution image synthesis engine with exceptional prompt fidelity and texture detail.

Context
Text Prompt
Max Output
2K / 4K Image
Public Pricing
$0.034 / image
Proprietaryimagedoubaoflagship
View Model & Endpoints
Qwen

Qwen

qwen-image-3.0-pro

Available

Qwen Image 3.0 Pro

Typography & Layout

Alibaba Qwen Image 3.0 Pro model specialized in complex visual layout composition and text rendering.

Context
Text Prompt
Max Output
2K Image
Public Pricing
$0.028 / image
Proprietaryimageqwenpro
View Model & Endpoints
Qwen

Qwen

qwen-image-3.0

Available

Qwen Image 3.0

Fast Image

Standard high-speed Qwen Image 3.0 generation model for general creative illustrations and web graphics.

Context
Text Prompt
Max Output
1024x1024 Image
Public Pricing
$0.022 / image
Proprietaryimageqwenfast
View Model & Endpoints
1-Minute Drop-In Integration

Zero Codebase Refactoring. Point baseURL to Switch.

100% compatible with OpenAI API specs. Drop straight into Next.js, LangChain, Cursor IDE, and LlamaIndex.

app/api/chat/route.tstypescript
import { createOpenAI } from '@ai-sdk/openai';
import { streamText } from 'ai';

// Drop-in replace baseURL & apiKey
const batchin = createOpenAI({
  baseURL: 'https://api.batchin.tech/v1',
  apiKey: process.env.BATCHIN_API_KEY,
});

export async function POST(req: Request) {
  const { messages } = await req.json();

  const result = streamText({
    model: batchin('deepseek-v4-pro'),
    messages,
  });

  return result.toDataStreamResponse();
}

Ready To Control AI Inference?

BatchIn brings model access, cost controls, signed records, and dedicated capacity into one workspace.

Contact teamGet Access

Models

  • Browse model catalogView the full available model list
  • View pricingSee the latest public pricing
  • Open PlaygroundSign in to validate real requests
  • Read API docsIntegration steps and examples

Products

  • Model API
  • Video & Multimodal API
  • VaaS (Verification as a Service)
  • Dedicated Capacity

Docs

  • Documentation & API
  • Compare
  • Events
  • WIN Incubator
  • Use cases

Support

  • Contact
  • hello@luminapath.ai
  • System status: operational
BatchIn© 2026 BatchIn. All rights reserved
PrivacyTerms

All model trademarks, product names, logos, and brands are property of their respective owners and used solely for compatibility and identification purposes.