Cổng API Hợp Nhất Cho Hàng Trăm Mô Hình AI
Model nào, nói đúng model đó. Giao thức tương thích chuẩn OpenAI & Anthropic. Số dư nạp trước dễ kiểm soát, lịch sử token minh bạch từng mili-giây cho lập trình viên Việt Nam.
curl https://ai.moventra.vn/v1/chat/completions \
-H "Content-Type: application/json" \
-H "Authorization: Bearer ***" \
-d '{
"model": "gemini-3.8-flash",
"messages": [{"role": "user", "content": "Tối ưu thuật toán routing AI Gateway"}],
"temperature": 0.2
}'Hạ Tầng Cốt Lõi
Thiết Kế Cho Độ Tin Cậy & Hiệu Suất Tối Đa
Tất cả những gì bạn cần để xây dựng, vận hành và mở rộng các ứng dụng AI hiện đại mà không lo rủi ro bị tráo model hay phụ thuộc vào một nhà cung cấp duy nhất.
Siêu Tốc (Lightning Fast)
Kiến trúc mạng tối ưu đảm bảo thời gian phản hồi mili-giây. Kết nối upstream trực tiếp giúp giảm tối đa độ trễ cho tác vụ streaming và coding agent.
Minh Bạch & Chuẩn Model (Absolute Transparency)
Cam kết: Model nào, nói đúng model đó. Không tráo đổi model nhẹ hơn, không hạ cấp context window. Lịch sử token và chi phí kiểm tra được 100% trên dashboard.
Độ Phủ & Điều Phối Thông Minh (Resilient Gateway)
Multi-region pooling và Circuit Breaker tự động dự phòng. Khi một kênh upstream gặp sự cố, gateway tự động chuyển hướng tức thì mà không làm gián đoạn ứng dụng.
Thân Thiện Lập Trình Viên (Developer First)
Tương thích 100% với OpenAI SDK, Anthropic SDK, cURL, Python, Node.js, LangChain và LlamaIndex. Chỉ cần cấu hình Base URL là chạy ngay lập tức.
Model Plaza Đa Phương Thức & Bảng Giá
Minh bạch giá vốn và giá bán thực tế. Tự động quy đổi tỷ giá USD/VND theo thời gian thực. Hỗ trợ đầy đủ Text Tokens, Image Generation và Video AI theo từng giây hoặc lượt tạo.
| Mã Mô Hình (Model ID) | Nhà Cung Cấp | Ngữ Cảnh | Đơn Giá | Giao Thức API | Thao Tác |
|---|---|---|---|---|---|
minimax/minimax-m3:free LLM MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding, and tool use. It is built on MiniMax Sparse Attention (MSA), which replaces full attention with KV-block selection to cut per-token compute at long context — roughly 1/20 the cost of the previous generation at 1M tokens, with substantially faster prefill and decode while retaining quality across most tasks.
Trained as a native multimodal model on interleaved data and tuned for multi-turn, production-like collaboration via an interactive user-simulator framework, the model is oriented toward sustained, multi-step tasks rather than single-turn execution. | OpenAI | 1.048.576 tokensContext Window | ~1₫ - 1₫ - (0₫ cache) / 1K tokens Hệ số 0.00 (Target 30%) | OpenAI (/v1) | |
nvidia/nemotron-3.5-lightning:free LLM | OpenAI | 1,000,000 tokensContext Window | ~1₫ - 1₫ - (0₫ cache) / 1K tokens Hệ số 0.00 (Target 30%) | OpenAI (/v1) |
minimax/minimax-m3:free
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding, and tool use. It is built on MiniMax Sparse Attention (MSA), which replaces full attention with KV-block selection to cut per-token compute at long context — roughly 1/20 the cost of the previous generation at 1M tokens, with substantially faster prefill and decode while retaining quality across most tasks. Trained as a native multimodal model on interleaved data and tuned for multi-turn, production-like collaboration via an interactive user-simulator framework, the model is oriented toward sustained, multi-step tasks rather than single-turn execution.
nvidia/nemotron-3.5-lightning:free
Quy Trình Bắt Đầu
Ba Bước Đơn Giản Để Bắt Đầu
Tích hợp trong vòng chưa đầy 1 phút. Hoàn toàn tương thích với mọi thư viện và công cụ bạn đang dùng.
Cấu hình & Tạo Key
Đăng nhập Dashboard, tạo API Key mới với hạn mức số dư mong muốn. Hệ thống hỗ trợ nạp tiền trả trước linh hoạt qua chuyển khoản hoặc mã nạp.
Authorization: Bearer ***Kết nối Endpoint
Thay thế Base URL trong ứng dụng hoặc phần mềm (Cursor, Cherry Studio, Python SDK) sang https://ai.moventra.vn/v1 mà không cần đổi logic code.
baseURL: "https://ai.moventra.vn/v1"Giám sát Minh Bạch
Theo dõi số lượng token input/output, model thực tế được phục vụ và chi phí trừ trực tiếp theo từng mili-giây trên trang Usage Logs.
HTTP 200 OK • 142ms • $0.00045Hạ Tầng Kết Nối
Các Điểm Cuối (Endpoints) Chính Thức
Sử dụng các URL bên dưới để kết nối ứng dụng hoặc truy cập các công cụ hỗ trợ của hệ thống.
Endpoint tiêu chuẩn cho chat, completions, audio, embeddings và model list. Tương thích OpenAI SDK.
https://ai.moventra.vn/v1Endpoint chuyên dụng cho Anthropic SDK; SDK sẽ tự động định tuyến tới /v1/messages.
https://ai.moventra.vnHỏi Đáp Kỹ Thuật
Câu Hỏi Thường Gặp (FAQ)
Các vấn đề kỹ thuật thường gặp khi tích hợp API và giải pháp xử lý tức thì.