APIX is designed to separate measured performance from popularity. This MVP shows the scoring structure; its current tool values are demo data, not published benchmark results.
Each category gets repeatable tasks with clear success criteria, inputs and expected outputs.
Tools are tested against the same task suite, with model/version, timing, cost and outputs recorded.
Verified usage evidence can complement lab tests without replacing them.
AI products change quickly. Scores should belong to a tested product/model version and date, not live forever under a brand name.
A tool can be excellent for image generation and irrelevant for research. APIX avoids pretending one universal benchmark fits every task.
The production goal is for every official score to link back to prompts, outputs, evaluators, timings and costs.
VIQQO stores source type, observation date, verification date, freshness window, verification status and evidence level separately from the claim itself. Vendor documentation can support factual product attributes, but vendor marketing is not independent performance proof. Prototype records remain marked demo until real sources replace them.