Measured results

Visual quality, compared with stock sub-agents

In one historical fixed-render visual study, the trained SGT Fleet arm scored 7.44 / 10 against stock Sonnet at 6.72 / 10 — a gain of +0.72 points in this test. A frontier hand-build reference scored 6.38 / 10 and is shown for context, not as a competing product.

Boundary: five ratings of fixed renders, not five independent builds. Visual quality only.

  • Trained SGT Fleet7.44/10
    +0.72 points vs stock Sonnet in this test
  • Stock Sonnet6.72/10
  • Frontier hand-build (reference)6.38/10
    Shown for context; not a competing product.

Fresh Blender study, 16 September 2026: same scene, same orchestrator

Fable 5.1 orchestrated both builds of the same German village brief. With SGT Fleet running on DeepSeek Flash 4.1, the final frame scored 7.66 / 10. Claude Sonnet 5 working alone, no fleet, scored 4.72 / 10. The live gallery render of the same scene scored 7.78 / 10 in the same sitting.

German village hero frame built by SGT Fleet on Flash 4.1
SGT Fleet on Flash 4.1 · 7.66 / 10 · 5 rounds, 45 min, 0.92 M tokens
German village hero frame built by Claude Sonnet 5 alone, no fleet, camera re-framed for display
Sonnet 5 alone, no fleet · 4.72 / 10 · 8 rounds, 3 h 49 min, 1.90 M tokens · same scene, camera pulled back for display after judging (frame as judged)
  • SGT Fleet on Flash 4.17.66/10
    +2.94 points vs Sonnet 5 alone in this test
  • Sonnet 5 alone, no fleet4.72/10
  • Live gallery render (reference)7.78/10
    Shown for context; the bar the fleet build was measured against.

Boundary: one scene, one hero view per arm, eight blind scores per arm from two judge families (Claude Opus and Codex) with sealed label mappings. Visual quality only. The fleet's skills were trained on Flash 4.1; the same fleet on Sonnet 5 scored 6.22, and GPT-5.6 Luna alone scored 7.62, so this is a comparison of configurations, not a ranking of models.

Method and full scores

What was measured

Historical cinematic final five three-way judgments, study dated 1 September 2026. Five ratings of fixed renders — the same renders rated each time, not five independently generated builds. Visual quality only; this is not a runtime-correctness or general-ability measure.

Arms

  • A — Frontier hand-build
  • B — Trained SGT c2
  • C — Stock Sonnet; exact version not established from the retained report

All five raw sample scores

Sample A (frontier) B (SGT c2) C (stock Sonnet)
16.37.46.9
26.77.26.5
36.07.56.5
46.47.46.7
56.57.77.0
Mean6.387.446.72

Method

Judge identifier recorded in the retained harness: claude-fable-5. Exact model versions are unverified and are not relabeled here. The study used randomized labels with context-blind intention, but no full old tool trace survives, so it is not asserted as an audited double-blind study. SGT ranked first in 5 of 5 ratings of these fixed renders.

Fairness limits

  • Five ratings of fixed renders, not five independent scene builds.
  • Three-way randomized labels; no balanced order record in the old study.
  • No raw request/tool trace to establish full context isolation retrospectively.
  • Fable5.1 and Sonnet5 exact versions unverified; not relabeled.
  • Later c2 was trained using preceding judge feedback; this comparison is not held-out generalization.
  • Different task, rubric and builds from JANUS; does not invalidate JANUS numbers or establish GPT6 superiority.

Render gallery

Created with SGT Fleet.

Explore our finished 3D work. Choose a view, or click a render to see it in full resolution.

Rifle

Rifle — Hero

Malibu Mansion

Malibu Mansion — Ocean hero

Virginia Brick Home

Virginia Brick Home — Front hero

German Village

German Village — Hero

Forbidden City

Forbidden City — Hero

Pricing

Simple on purpose.

The SGT Fleet software is free. Paid plans add hosted access to the skills and OOP query service. Your Claude, OpenAI and Ollama subscriptions are billed separately by those providers.

STUDENTS + EDUCATORS

Free

For enrolled students and teaching staff.

  • Full Fleet software
  • Both launch CLIs
  • Updates included
Start onboarding No payment collected; education verification remains pending.

ENTERPRISE

$10 / seat / month

For organizations that need per-seat licensing and private-server deployment.

  • Per-seat licence
  • Private-server deployment
  • Invoicing
Start enterprise $10 per seat / month, invoiced.

Notes

Practical notes.

What is actually tuned?

SGT tunes skills and OOP role packages — the instructions, routing and review structure around frontier orchestration. Provider model weights are unchanged.

How do I extend to a new task or model?

Use your own frontier usage on unseen work, produce examples and checks, benchmark the specialist behaviour, then fine-tune a new skill or OOP role package for future routed tasks.