AIToday
Large Language ModelsAI Business & IndustryOpenAI BlogPublished: Sep 22, 2026, 01:00 JST

V7 Go hits 99.9% accuracy on 50–100 step workflows

V7 Go hits 99.9% accuracy on 50–100 step workflows

3 Key Points

  1. What happened

    V7 says its V7 Go platform, using GPT-5.6 models and its Context Graph, completes 50–100 step workflows in minutes at 99.9% accuracy, with an auditable trail.

  2. Why it matters

    Retrieval accuracy is non-negotiable for finance, insurance, and real estate teams, so source-linked context could make long agent workflows usable in those functions.

  3. What to watch

    The 99.9% figure is V7's own report, and the harder test is whether its Context Graph keeps working as customer files change — watch the four-difficulty-level benchmark where GPT-6 Astra scored 89% versus GPT-5.6 Sol's 78% on the very-hard level.

WHO IT HITSFinance, insurance, and real estate teams doing document-heavy review work — deal screening, claims processing, underwriting — could cut manual review time if the reported accuracy and speed gains hold on their own files.

Not sure about something? Ask the AI

Summaries like this, in your inbox every morning.

Context & Analysis

V7 was founded in 2018 by Alberto Rizzoli and Simon Edwardsson, who had previously built a widely used computer vision accessibility app together. Their pitch is that today's models can reason through complex tasks but do not automatically understand the business context behind them — which fund report is current, how the same entity is named across three systems.

The Context Graph is V7's answer to that gap. It connects to repositories such as SharePoint and Google Drive, scans them for entities, relationships, facts, attributes, and metrics, and preserves cited evidence back to the original source. V7 says this graph is an order of magnitude cheaper and faster to traverse than long-context approaches, and that it can fall back on RAG when the graph lacks enough information. The company's customers include asset managers, a financial services team, and insurance claims processors; V7 reports deal screening 21x faster, a review process cut from more than 100 hours to under 10, a $12,000 expert-cost saving per task, and a 13.5% error reduction in claims processing versus a manual baseline.

V7's model choices also reveal how it weighs cost against capability. It assigns each workflow step to fast, medium, or smart tiers; reports a 78% lower cost per document with GPT-5.6 Luna than with GPT-5.4 mini; and says moving document-heavy workloads to the Responses API cut token use by roughly 5% for some PDF-heavy workflows. On the hardest graph-query set, V7 reports GPT-6 Astra at 89% versus GPT-5.6 Sol at 78%, with both near 100% on easier levels. Whether this holds as customer files change is likely the real test, since the longer-term goal is workflows that start when facts in the Context Graph change and flag analyses still relying on old figures.

FAQ
What benchmark did V7 use to test context retrieval?
V7 tested on HERB, a benchmark for finding and connecting information spread across enterprise systems. V7 says its retrieval-only system outperformed the official baseline by 69% and reduced hallucinations on un-answerable queries by 38%.
How much faster or cheaper did V7's customers get?
V7 says asset managers can screen deals 21x faster, reducing a full-day process to just 15 minutes. One financial services team cut review time from more than 100 hours to under 10, saving $12,000 in expert costs per task.
Which OpenAI models does V7 Go use?
V7 Go uses GPT-5.6 Luna for high-volume structured extraction, and GPT-5.6 Terra and Sol for reasoning and tool use. V7 is also starting to use GPT-6 Astra on the most demanding Context Graph queries.

Get the latest Large Language Models news every morning

For example, today's edition would include:

  • Meta Jumps 7.6% as Wells Fargo Lifts Target to $796 Before Meta ConnectYahoo Finance AI · 1h ago
  • Meta's Muse sparks tech rally; Arm surges 15%Yahoo Finance AI · 1h ago
  • Jev, SemIf AI deciders cut if-then costs 99%Tomasz Tunguz (Theory Ventures) · 1h ago

AI-summarized, only the topics you pick — one digest a day via Email, LINE, or Slack.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleMultiverse's CBO beats rivals by almost 23 points on MMLU