Open Models Drive On-Prem AI Adoption

Diving deeper into

Tenry Fu, CEO of Spectro Cloud, on why 60% of AI will be on-prem

Interview
some of the open source models, like GLM, perform very close to Anthropic's models, especially for coding and cybersecurity tasks.
Analyzed 9 sources

The key shift is that coding AI is breaking into two layers, expensive frontier models for planning, and cheap open models for execution. Spectro Cloud describes using Anthropic class models for the hard reasoning step, then handing off routine implementation and security work to local open models. That matters because once a company can keep most tokens on its own hardware, AI coding economics start to look more like software infrastructure spend than an open ended API bill.

  • GLM is no longer just a research curiosity. Z.ai now sells a dedicated GLM Coding Plan, exposes coding specific endpoints, and publishes setup guides for Claude Code, Cursor, Cline, and other agent tools. That shows open models are being productized for real developer workflows, not just benchmark demos.
  • Anthropic still holds the premium position in high value reasoning. Claude Code is built for direct use inside the terminal and enterprise cloud setups, which fits the planning and orchestration role Spectro Cloud describes. The open model substitute is strongest where the task is repetitive, local, and high volume.
  • This same pattern is showing up across the market. DeepSeek is framed as an interchangeable open weight supplier inside coding seats and agent platforms, and Mistral is winning with on premises deployments in sovereign and regulated settings. Model competition is moving from who is smartest overall to who is good enough at the lowest all in cost.

The next phase is a router architecture inside every serious enterprise AI stack. Frontier labs will keep the high margin planning step, while open models capture the bulk execution layer on local clusters, private clouds, and air gapped systems. That favors infrastructure companies like Spectro Cloud that help enterprises decide where each workload runs, and keep switching costs low across models.