AIToday
Large Language ModelsAI Safety & AlignmentAI Business & IndustryAI Watch (Impress)Published: Oct 2, 2026, 19:01 JST

ELYZA founds ELYZA RSI Research for AI-driven AI development

ELYZA founds ELYZA RSI Research for AI-driven AI development

3 Key Points

  1. What happened

    ELYZA said it is launching "ELYZA RSI Research", and that for LLMs of 100 billion parameters or fewer it has reached stage 1 of AI-driven AI development, with harness improvement almost fully automated at stage 3.

  2. Why it matters

    If improvement loops run almost without human hands at that scale, the pace at which frontier models are released could rise, though this is still a research-stage claim.

  3. What to watch

    Its benchmark work sits only at stage 2, semi-automated, and full automation of the loop remains stage 4, so the test is whether ELYZA can push evaluation machinery further.

WHO IT HITSEnterprise and research teams watching who can automate AI development will read this as a claim about speed, while safety reviewers will focus on the isolation and shutdown measures ELYZA says it uses in research.

Not sure about something? Ask the AI

Questions and answers are published on this page.

Summaries like this, in your inbox every morning.

Context & Analysis

ELYZA is an AI company spun out of the University of Tokyo's Matsuo Lab, working on large language models and their deployment. It joined the KDDI group in April 2024, a move it frames as giving it GPU resources to pursue research and product releases at scale.

In its own account of the field, LLM accuracy has risen fast, simple tasks plateaued about two years ago, and the current front line is science use and complex multi-day tasks. It also points to a wave of open models, largely from China, arriving two to three months after the frontier. Against that backdrop, it describes recursive self-improvement as already beginning at companies such as Anthropic and OpenAI, and lays out four stages, from humans doing everything to an improvement loop that runs entirely on AI.

The stakes hinge on how quickly the evaluation side can move. ELYZA has full automation in the harness but only semi-automation in evaluation, and it says full autonomy is stage 4. Its planned extension from LLMs into vision-language models and vision-language-action models for robots is likely to be judged on the same safety conditions, which its executives described as a matter of isolation and the ability to stop the system.

FAQ
How far has ELYZA actually automated AI development?
ELYZA says that for LLMs of 100 billion parameters or fewer it has reached stage 1, and for harness improvement it has achieved stage 3, where improvement is almost fully automatic. Its evaluation machinery, however, is only at stage 2, semi-automated benchmark development.
What did ELYZA announce beyond the research effort?
It announced "ELYZA Voice Agent", a voice-capable AI agent it is developing and commercializing as the first product of its vertical AI business.
What safety measures did ELYZA describe?
ELYZA said it enforces proper isolated environments and builds systems that can stop the AI during research and development. For products, it said safety is a matter of how they are used and that it has not yet reached that level.
AI Watch (Impress)Read Original Article

AI news that matters for your work, delivered every morning.

Pick your industry and the AI tools you use, and get news related to your work every day.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.

Questions and answers are published on this page.

Related Articles

Next articleZhou Yuxiang: AI's Google moment not here yet