Evidence behind every label.
Aurxo makes each conclusion understandable by publishing its scope, metrics, evidence quality, sources, confidence and known limitations.
Four signals, kept deliberately distinct.
No evidence, no invented conclusion. When evidence is incomplete, stale or not comparable for the stated scope, Aurxo can withhold the overall grade and explain why.
How Aurxo reaches a public result
Define the exact scope
Product or reference model, task, functional unit, date, deployment assumptions and system boundary.
Collect traceable evidence
Every published metric links to a named source, publication date and evidence type.
Classify evidence integrity
Provider measurements, externally reviewed evidence, independent estimates, claims and unknowns remain visibly different.
Apply the Aurxo assessment model
The same internal rules are applied consistently to comparable scopes. The detailed mechanics are proprietary.
Publish the record
The result is released with its scope, values, confidence, sources, limitations, date and version.
What is public. What is protected.
Everything needed to assess the claim
Scope, measurements, evidence type, source links, confidence, limitations, assessment date, version and correction process.
The machinery that creates the product
Exact thresholds, weights, coefficients, tie-break rules, internal benchmark library, structured datasets and implementation code.
Product grades and reference labels are different.
A current product grade applies only when Aurxo has sufficiently current product-level evidence. A reference label belongs to a named model, prompt size, deployment and study date. A historical model reference cannot automatically become a universal grade for a routed product.
Aurxo Engine v1.6.4: capable first, greener among capable options.
The public interface sends the request securely to the private Aurxo Engine, which identifies the intended task, input format, desired output and critical requirements. It then removes unsuitable tools, compares documented product capabilities and asks one focused clarification when the request is materially underspecified.
Interpret the need
Identify the job, input, output and critical requirements in English or Spanish.
Remove unsuitable options
A tool must be capable of completing the actual task before environmental evidence can influence the choice.
Compare documented capabilities
Aurxo uses dated public product documentation and may show fewer than three options instead of filling the list with weak matches.
Clarify when necessary
One useful question is better than a confident generic recommendation.
Prioritize the greener capable option
After weak matches are removed, Aurxo prioritizes the strongest scope-compatible environmental signal among tools that can do the job well. If comparable evidence is unavailable, Aurxo says so and keeps the task-fit order.
Capability profile date: 28 July 2026. Profiles are editorial comparisons of publicly documented functions, not laboratory benchmarks or guarantees of output quality. Availability can vary by plan, region and workspace.
Selected official capability sources: OpenAI data analysis · Anthropic Claude professional workflows · Gemini file analysis · Copilot in Excel · Perplexity file uploads · Consensus Research Agent · DeepL document translation · Runway Gen-4.5 · Descript speaker detection · Zapier Agents · Replit build and publish.