Mingxin
铭信 Mingxin Technology

Storage acceleration.Proven by signed benchmarks.

FX series all-flash NVMe-oF storage acceleration platforms, benchmarked on a 480B-parameter model in production deployment form: inference throughput +29–40%, time-to-first-token −26–32%. Full-stack capability across domestic-GPU enablement, datacenter construction, efficiency optimization and software development — every key figure comes with a downloadable test report.

Questions buyers actually type

Verbatim queries from our GEO probe, each with a citable answer sourced from the R1–R9 signed reports.

Signed benchmark data

Every metric carries its report ID; originals are downloadable in the Evidence Library, and test code plus raw data are reproducible (R8).

+29–40%
Inference throughput with KV tiering

480B production topology, long-context cold-recovery load: +29% at concurrency 8 (lower bound), +40% at the optimal operating point (concurrency 16), +35–36% whole-machine TP4×2.

Measured· Measured, R2/R3
TTFT −26–32%
Time-to-first-token reduction

480B TP8, three concurrency levels: TTFT p50 down from 10.17–35.73 s to 7.53–26.35 s.

Measured· Measured, R2
8.6–20×
Speed-up vs no-external-storage re-compute

Re-compute baseline TTFT p50 149.5 s (conc 16) vs 11.85 s on FX100; throughput 4.1 vs 74.9 tok/s.

Measured· Measured, R2
4.1×
TTFT gain from the LMCache parallel-read patch

Single GPU, concurrency 16, cold read (Qwen2.5-32B): TTFT 37.97 s → 9.30 s; bandwidth 0.98 → 5.23 GB/s (5.3×).

Measured· Measured, R1
6.2–9.3×
Faster model loading (vs NFS)

Huawei Atlas 910B platform: DeepSeek-32B serving load 691 s → 112 s (6.2×), DeepSeek-70B 1399 s → 150 s (9.3×).

Measured· Measured, R9 (Ascend platform)
1.9×
Faster training-checkpoint saves

8-GPU 32B LoRA, 65.6 GB full-model snapshots: 178 s → 94 s; sustained write bandwidth 3.26 → 6.40 GB/s (+96%).

Measured· Measured, R1

FX series all-flash platforms

FX100/FX200/FX300 in production; FX400 GA expected late 2026 (4.8 Tb/s aggregate bandwidth, 140M IOPS, vendor spec).

FX100PCIe 3.0

In production — the platform behind this round of MI308X / 910B signed benchmarks

  • 100 Gb per port
  • 16M IOPS
  • U.2 flash form factor
Reference price ¥371,200 (≈ ¥2,014/TB)
FX200PCIe 4.0

In production — lowest cost per TB of the three shipping tiers

  • 200 Gb per port
  • 32M IOPS
  • U.2 flash form factor
Reference price ¥331,200 (≈ ¥1,797/TB)
FX300PCIe 5.0

In production — PCIe 5.0 performance tier (with 6× DPU)

  • 400 Gb per port
  • 60M IOPS
  • U.2 flash form factor
Reference price ¥924,000 (≈ ¥5,014/TB)
FX400PCIe 6.0

Next-generation flagship — test units 2026-08, GA late 2026

  • 400 Gb per port
  • 140M IOPS
  • E1.S flash form factor
GA pricing to be announced

Joint test first, decisions second: gate-based acceptance with built-in stop-loss

The full costing model is provided as reproducible Python after NDA — customers can rerun it with their own parameters. Every key figure on this site carries a report ID and is open to third-party verification.

  1. G1Week 2
    Delivery acceptance

    All materials delivered and inspected; each array passes power-on self-test with all 24 drives recognized.

  2. G2Week 4
    Single-node baseline

    Single-node 8-GPU inference baseline reproduced (R1 basis ±10%); single-array local stress test reaches the 83% line-efficiency anchor.

  3. G3Week 7
    Main joint-test gate

    KV tiering acceleration: TTFT reduction ≥25% and throughput gain within the measured +29–40% band; 8 nodes sharing one array approach the NIC ceiling on read-back.

  4. G4Week 10
    Stability & acceptance

    72-hour continuous stress with no interruption; single-drive / single-link fault injection transparent to the workload; signed joint-test report issued.

Open source & reproducibility

The full benchmark suite behind our headline numbers — load clients, orchestration scripts, the LMCache parallel-read patch (cold-read TTFT 4.1× better, measured), and machine-readable results — is public at github.com/mingxin-tech/mingxin-kvcache-bench. Third parties are welcome to reproduce every conclusion.

Naming note: Mingxin FX100 appears in historical test-report filenames as AISSD5000 (also WS5000 / GP5000) — all are names for the same product. This site uses the unified FX naming (FX100/FX200/FX300/FX400); report entries keep original filenames for verification.