AIToday
Large Language ModelsAI Business & IndustryAI Watch (Impress)Published: Oct 8, 2026, 19:01 JST

Sakura Internet to launch Private Edition of Sakura AI Engine on October 8

Sakura Internet to launch Private Edition of Sakura AI Engine on October 8

3 Key Points

  1. What happened

    Sakura Internet will start 'Sakura AI Engine Private Edition' on October 8, hosting users' generative AI models on its own GPU infrastructure and serving them via API in a closed, dedicated environment.

  2. Why it matters

    Because the environment is dedicated, users are not affected by other tenants, and Sakura Internet covers everything from building the inference platform to operations and monitoring, which is expected to cut the operational burden on users.

  3. What to watch

    The service starts October 8, with monthly flat-rate billing per GPU.

WHO IT HITSThis mainly affects companies that want to run their own or fine-tuned generative AI models in-house, especially those handling confidential data, since input data and generated results are not used for AI training. It may also matter to teams that today avoid token-based billing because they cannot predict costs, as the service uses flat monthly GPU-based pricing.

Not sure about something? Ask the AI

Questions and answers are published on this page.

Summaries like this, in your inbox every morning.

Context & Analysis

Sakura Internet is building this service on the technical foundation and operational design of its existing 'Sakura AI Engine,' which suggests it is extending a proven platform rather than starting from scratch. The service hosts users' generative AI models on Sakura Internet's GPU infrastructure and makes them available via API, covering independently developed models, fine-tuned models, and open-weight models in a closed environment used only by that customer.

Two design choices stand out. First, the dedicated environment means inference runs are not affected by other tenants, and Sakura Internet handles everything from building the inference platform to operations and monitoring, which is positioned as reducing the user's operational load. Second, the pricing is a monthly flat rate per GPU, which removes the concern of costs fluctuating with token usage, and input data and generated results are not used for AI training, opening the door to uses involving confidential information.

For readers weighing how to run their own models, the practical contrast is with services that charge by token volume or share infrastructure across customers. Whether the dedicated setup and flat rate are a good fit is likely to depend on each company's workload, and that will probably become clearer once the service is running from October 8.

FAQ
When does the service start?
It starts on October 8.
What kinds of models can be used?
Independently developed models, fine-tuned models, and open-weight models can be run in a closed environment dedicated to the user.
How is it billed?
It uses a monthly flat-rate plan billed per GPU, so costs do not vary with token usage.
Can it be used for confidential information?
Yes. Input data and generated results are not used for AI training, so it can be used for purposes involving confidential information.
AI Watch (Impress)Read Original Article

AI news that matters for your work, delivered every morning.

Pick your industry and the AI tools you use, and get news related to your work every day.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.

Questions and answers are published on this page.

Related Articles

Next articleForget AI bans: classes préparatoires-style oral exams, scaled by AI