About

SinoVerdict — the data layer for Chinese case law in legal AI

SinoVerdict is an enterprise data partner that licenses 170M+ structured, English-indexed, citation-grounded Chinese court judgments to power legal AI products. If you are building an AI system that needs to cover Chinese law — for retrieval, fine-tuning, or an agent — SinoVerdict is the corpus and the connector behind it.

170M+
PRC court judgments, and growing
Daily sync
From authoritative PRC sources
EN in / EN out
English query, English summaries, source on tap
Bulk · API · MCP
Built to ingest, not to browse

What SinoVerdict is

SinoVerdict (also known as 文书查 / Wenshucha) is an AI-native structured data licensor for Chinese case law. The category matters: SinoVerdict is not a research subscription you log into, and not a scrape you have to clean and defend. It is a structured corpus and a set of connectors designed from the start to be ingested by an AI system rather than read by a person.

It is operated by Shenzhen Xingpu Network Technology Co., Ltd. (深圳星谱网络科技有限公司), and is the English-facing arm of the same data operation behind wenshucha.com (China main site) and mcp.wenshucha.com (Chinese developer site).

What's in the corpus

More than 170 million publicly available Chinese court judgments, sourced from China Judgments Online (中国裁判文书网) and provincial court portals, spanning roughly four decades of published case law:

How it's delivered

Three delivery modes, all built to drop into a pipeline:

ModeWhat it isBest for
Bulk datasetA structured corpus dump with the full field schemaRetrieval indexes, fine-tuning, evals
REST APIQuery endpoint at tob.wenshucha.com/api/v1Server-side and application integration
MCP serverModel Context Protocol tool for AI assistantsClaude Desktop, Claude Code, ChatGPT, Cursor, Cline, in-house agents

Who licenses SinoVerdict

Clients include LexisNexis and China's leading legal databases (PKULaw / 北大法宝, 无讼, 德力法搜).

How SinoVerdict differs from the alternatives

AlternativeWhy it falls short for production legal AI
Academic datasets (CAIL2018, LeCaRD/LeCaRDv2)Frozen snapshots, often narrow scope (e.g. criminal-only), Chinese-language only, no license to ingest commercially, no API/MCP, no daily sync. Built for benchmarks, not products.
Domestic aggregators (PKULaw, Wolters Kluwer China)Per-seat, Chinese-language subscriptions to a human search UI. They don't license a structured corpus you can ingest. The mismatch is access model and license, not content quality.
Western incumbents (LexisNexis, Westlaw, vLex)A thin, English, human-reference slice — thousands of curated documents, not the full corpus — assembled for reading, not machine retrieval at completeness.
Raw scrapers / grey dataCheap bulk at the cost of unclear legality, uncharacterized coverage, no structure, no grounding, and freshness that rots.

For the full breakdown, see The State of Chinese Legal AI Data, 2026: A Supply-Side Market Map.

Compliance note: published Chinese judgments retain party names while redacting personal identifiers, and cross-border use is structured around the Personal Information Protection Law (PIPL) and the Data Security Law (DSL). This is informational background, not legal advice — structure any licensing arrangement with qualified counsel.

License the corpus

If you're a legal AI vendor sizing up a China data partnership, let's scope a pilot — typically a 2-week evaluation on a slice of the corpus, then an enterprise agreement. Firms and KM teams can request a trial key to evaluate in Claude / Cursor.

Request access & coverage report

Or email chenjiaxin@wenshucha.com · Read the insights