ModelScope
Open Source Platform · since 2022ModelScope is an AI API open source platform provider. TokenAPI Scan independently monitors model authenticity, protocol coverage, response latency, and pricing transparency.
This page aggregates AI API relay detection data for ModelScope, focusing on whether Claude / OpenAI / Gemini are truly passthrough, whether usage fields are anomalous, whether model lists are usable, and accessibility across different network regions. If you're searching for "ModelScope real or fake", "ModelScope review", "API relay detection", or "relay price comparison", we recommend checking the detection summary, network status, and public pricing evidence below before deciding whether to continue using it.
Related:Relay Rankings · Price Comparison · Network Tools · FAQ · Buying Guide
/models endpoint. The provider may not support OpenAI-style model enumeration, may use a different endpoint path, or the current network probe failed. 💰 Public Price Snapshot
🟡 Self-reportedFrom the provider public page, not recently verified by us. Here is the full price table for ModelScope (scraped from the public pricing page, 17 models) - cross-provider comparison at pricing page, API volatility at AI service status.
| Model | Input / M | Output / M | Cache R | W 5m | W 1h | Gov |
|---|---|---|---|---|---|---|
qwen-coder |
$0.3000 | $1.5000 | — | — | — | raw |
qwen-max |
$1.6000 | $6.4000 | — | — | — | raw |
qwen-plus |
$0.4000 | $1.2000 | — | — | — | raw |
qwen-plus-2025-01-25 |
$0.4000 | $1.2000 | — | — | — | raw |
qwen-plus-2025-04-28 |
$0.4000 | $1.2000 | — | — | — | raw |
qwen-plus-2025-07-14 |
$0.4000 | $1.2000 | — | — | — | raw |
qwen-turbo |
$0.0500 | $0.2000 | — | — | — | raw |
qwen-turbo-2024-11-01 |
$0.0500 | $0.2000 | — | — | — | raw |
qwen-turbo-2025-04-28 |
$0.0500 | $0.2000 | — | — | — | raw |
qwen-turbo-latest |
$0.0500 | $0.2000 | — | — | — | raw |
qwen3-next-80b-a3b-instruct |
$0.1500 | $1.2000 | — | — | — | raw |
qwen3-next-80b-a3b-thinking |
$0.1500 | $1.2000 | — | — | — | raw |
qwen3-vl-235b-a22b-instruct |
$0.4000 | $1.6000 | — | — | — | raw |
qwen3-vl-235b-a22b-thinking |
$0.4000 | $4.0000 | — | — | — | raw |
qwen3-vl-32b-instruct |
$0.1600 | $0.6400 | — | — | — | raw |
qwen3-vl-32b-thinking |
$0.1600 | $2.8700 | — | — | — | raw |
qwq-plus |
$0.8000 | $2.4000 | — | — | — | raw |
17 models · Source ModelScope pricing page · Snapshot 2026-07-29(23:02 UTC) · Last tested 2026-05-21 · Total 26 tests
✅ Strengths
- Alibaba-backed
- Huge model library
- Very fast access in China
- Free tier
⚠️ Notes
- Mainly open-source models
- API style slightly differs from OpenAI
🎯 Use Case
- China-region open-source model apps
- Learning and experimentation
- Multi-modal scenarios
⚡ Speed Test from Your IP
🔬 Real Detection Summary
Last check: 2026-05-21 20:37 UTC
Recent Detection Latency Trend (30 times, old->new)
By Protocol
Last 5 reports
Source: All metrics from real detections after users submit API keys,Server-side Calls ModelScope protocol verification requests from endpoints,. Pass rate = Requests returning correct responses / Total Requests。 -> Full Methodology
🛰️ System Probe Summary
The platform runs keyless connectivity probes every 6 hours on this site's OpenAI / Anthropic / Gemini protocol list-models endpoints (HTTP 401/403 counts as connected, indicating the endpoint is alive). Below are the cumulative real results.
By Protocol (System Probe)
Note: Connectivity rate = endpoint responses (incl. 401/403 auth responses) / probe count, measures endpoint liveness not service quality; service quality is determined by key-submitted detections in "Real Detection Summary" above.
📊 Real-time Detection Data
ModelScope
0.0% Median Pass Rate · Snapshot Loaded💧 Water Rate v2 86.4 / 100 · Excellent Real request http=200 ratio 86% avg 648ms Click to expand formula →
| Report | Protocol | Score | Duration | Time |
|---|---|---|---|---|
| View Report | OPENAI | 0% | 2026-05-21 20:37 | |
| View Report | OPENAI | 0% | 2026-05-20 18:00 | |
| View Report | OPENAI | 0% | 2026-05-20 18:00 | |
| View Report | OPENAI | 0% | 2026-05-19 18:00 | |
| View Report | OPENAI | 0% | 2026-05-19 18:00 |
📋 Protocol Suitability Matrix
Based on 26 real tests of protocol coverage and median pass rate, answering "Which protocols can I use with ModelScope?"
| Protocol | Tests | Median Pass Rate | Status | Suggestion |
|---|---|---|---|---|
| openai | 26 | 0% | - Unknown | Test with small traffic first |
⏱ Rate Limits & Pricing Multiplier
Key operational parameters that determine whether you can scale concurrent calls.
⚠️ Actual multipliers depend on ModelScope's backend account. This site does not directly hold account tokens and cannot see hidden SLAs.
🛠 Integration Tutorial: Get Started with ModelScope in 3 Steps
- Register an account and create an API Key in the backend (the console usually has "My Keys" or "API Tokens" entry).
- Replace the base URL in your code with
https://api-inference.modelscope.cn/v1, and select the model name from ModelScope's list. - Test a request:
curl https://api-inference.modelscope.cn/v1/v1/chat/completions -H "Authorization: Bearer YOUR_KEY" -H "Content-Type: application/json" -d '{"model":"gpt-4o-mini","messages":[{"role":"user","content":"hello"}]}'
✅ This is an OpenAI Compatibility Mode integration example. Anthropic / Gemini protocols vary slightly by provider, see each site's documentation.。
❓ Frequently Asked Questions
Is ModelScope reliable?
Based on 26 independent tests by TokenAPI Scan, the median pass rate is 0%.Low pass rate, test with small traffic first before production.
How much does ModelScope cost?
ModelScope's public price page lists 17 models. See the "💰 Public Price Snapshot" section above. Source:ModelScope。
Which models does ModelScope support?
ModelScope's model list is not publicly enumerated yet; log in to the backend to view.
What's the difference between ModelScope and official?
Official (OpenAI/Anthropic/Google) = first-hand pricing + stable direct connection + strict content moderation and account ban risk. Third-party relays = usually discounted pricing + domestic accessibility + but compounded availability, content moderation, SLA, and service shutdown risks. Which to choose depends on your cost sensitivity, content moderation requirements, and stability priorities. It is recommended to use official for critical paths + relays as backup.
Will ModelScope shut down?
We continuously monitor relay site reachability. Multiple consecutive failed detections within 30 days = high risk signal. Recommendations: ① Don't top up too much ② Rotate across multiple providers for critical services ③ Watch the Rankings warnings.
💭 Related Questions
💰 Popular Model API Prices
Compare API prices across relay providers: