The short version: resolve an exact configuration, state a workload, evaluate separate capability axes, and keep every condition attached to the answer. A missing fact produces unknown, not optimism.
the rule that prevents most nonsense
A model family is not a configuration, a minimum requirement is not a workload result, and a successful boot is not proof of a safe homelab.
the four verdicts
verdict
meaning
what the page must show
yes
Every required axis clears the stated workload with no unresolved blocker.
Exact scope, evidence, and any sensible operating margin.
conditional
The workload fits only when a named setting, topology, accelerator, or upgrade is present.
The condition, likely failure mode, and smallest useful fix.
no
A required capability is absent or a supported constraint is exceeded.
The failed axis and an alternative that changes the decision.
unknown
Evidence is missing, contradictory, too old, or outside the tested scope.
The exact missing fact and what would resolve it.
the evaluator keeps axes separate
A single percentage would hide the most important part of the answer. Super Evil Robots evaluates compute, memory, media acceleration, storage capacity and topology, network, expansion, runtime support, power/thermals, and operational risk independently. The overall verdict is the most restrictive required axis after explicit conditions are applied.
Storage is deliberately split from compute. A one-litre PC may be a good application node and a poor place for the only copy of a media or photo library. A result can therefore pass compute while requiring external NAS/DAS storage and a separate backup plan.
evidence grades
grade
authority
allowed claim
a
Reproduced in the Super Evil Robots lab with a published fixture, manifest, repetitions, and raw result.
Measured behavior within the exact test scope.
b
Deterministic fact from an authoritative manufacturer or official software source.
Supported specification, requirement, or documented behavior.
c
Multiple independent reproducible reports without a material contradiction.
Community-observed behavior, clearly distinguished from official support.
d
A reasoned inference or one unverified report.
A hypothesis that cannot independently make a result “yes.”
unknown
Insufficient evidence.
No positive or negative claim.
The pilot currently contains official-source modeling and explicit inference. The accepted Grade A ledger currently contains 0 runs. The empty grade is a feature: the interface cannot pretend the lab measured something it did not.
every rule the evaluator can apply
The corpus holds 46 authored rules across 8 capability axes, resting on 10 distinct evidence records. Each rule below has its own address, so a result that names a rule can link straight to it.
req-haos-architecture@1
The selected image requires x86-64.
subject
configuration.capabilities.cpu.architecture · unit domain-not-declared
guard
one-of x86_64, x86_64-v2 · unit domain-not-declared
src-proxmox-ve-admin-guide · Proxmox Server Solutions
2026-07-26
open start → open end
ev-ser-workload-budget-model
grade d
src-ser-repositioning-plan · Super Evil Robots
2026-07-26
2026-07-26 → open end
No record here is a lab measurement. The accepted Grade A ledger is empty, and the grade column says so.
versions and effective dates
A claim carries a stable ID, source, retrieval or test date, evidence grade, configuration or version scope, and—where known—effective start and end dates. A new app release does not overwrite history. It creates an impact review for every affected requirement and verdict.
Official maximums and observed working maximums are separate fields. Likewise, “supported by the vendor,” “works in a community report,” and “reproduced in this lab” are never collapsed into one label.
benchmark protocol
A publishable run needs a machine-readable hardware manifest; firmware, OS, kernel, driver, runtime, and app versions; configuration with secrets removed; fixture checksum; cold/warm distinction; ambient and measurement conditions; at least three repetitions for variable tests; and median, range, failure count, and known limitations.
Tests map to decisions: concurrent transcodes, a fixed import corpus, a defined Home Assistant integration load, a declared Proxmox guest mix, or a combined hostile peak. Synthetic or openly licensed fixtures are used; private photos, camera feeds, machine names, and home-automation data are not published.
Super Evil Robots uses AI coding and maintenance agents to help write and test software, draft candidate structured records, monitor approved source URLs, calculate deterministic impacts, and prepare review queues. Agent output is not a source, physical observation, benchmark result, evidence grade, or human approval.
Automated workers may detect official-source changes, store approved snapshots, propose extracted facts, identify affected verdicts, schedule repeatable tests, validate variance, rebuild pages from approved data, and open review tasks. They may not approve source rights, resolve contradictions, change a public verdict, answer a material correction, or alter the test method without human review.
Candidate changes remain held until their machine checks and required human decisions are recorded. Physical Grade A claims require a real run with hash-pinned artifacts; generated prose can never satisfy that gate.
page graduation
The checker can calculate many states, but very few deserve search pages. An indexable page needs a distinct decision, stable identity, material unique evidence, maintained sources, a useful answer without generated filler, internal links, and a last-reviewed date. Everything else remains an interactive, non-indexed result.
A page is also accountable for the rules behind it. Every check the evaluator can apply is listed in the rule audit above, with its own address, so a disputed result can be traced to the exact rule that produced it.
Nothing in the graph has been revised yet, because nothing in it is old yet. The current graph carries 108 lifecycle-bearing records and 108 explicit history events, of which 0 changed a published claim after it was created. That is a starting line, not a record of stability.
generated 2026-07-26 from record.lifecycle.changeControl.history
inspect every lifecycle entry
record
type
state
effective
recorded history
app-frigate-0-17
app-release
deferred · deferred
2026-07-26 → open end
created: Initial version-scoped software record created.
app-home-assistant-2026-7
app-release
candidate · draft
2026-07-26 → open end
created: Initial version-scoped software record created.
app-immich-3
app-release
candidate · draft
2026-07-26 → open end
created: Initial version-scoped software record created.
app-jellyfin-10-11
app-release
candidate · draft
2026-07-26 → open end
created: Initial version-scoped software record created.
app-proxmox-ve-9
app-release
candidate · draft
2026-07-26 → open end
created: Initial version-scoped software record created.