For an anonymised Shenzhen enterprise offline environment, the JSLE team delivered a two-node GPU cluster, GLM-5.2-FP8 inference, a unified API route and a governed agent workspace, then completed functional verification.
Project facts
Background and boundaries
The customer wanted controlled inference in its own data centre, with model weights, invocation paths and runtime logs kept in the enterprise network. Client name, domains, people, server addresses and security settings are redacted; this page presents only the anonymised architecture, delivery method and test conclusions.
Functional verification
All six basic functional tests passed. Four concurrent SSE streams were verified and independent sessions remained isolated. Performance benchmarks, long-context behaviour and higher concurrency remain subject to agreed test conditions and formal acceptance evidence.
Next step
For a multi-node planning discussion, we can provide capacity sizing, implementation and acceptance planning.
Phone: +86 139 2521 1225 Email: tianjun@jsle.cn