Trustyu Forge Public library
Living library · editorial cut verified on Aug 13, 2026

References that guide
Forge.

An entry point for models, market statistics, articles, papers, authors and standards. The editorial radar points to current primary sources; Frozen catalogs preserve snapshots, hashes, licenses, and inference limits.

40 receipted sources 21 independent groups 8 models on the radar 4 suppliers 2 reproducible cuts No universal ranking
40sources in both catalogs
21independent publishers
20 + 20receipts per cut
8models with verified price
4suppliers on the radar
13 Augverification date
Model Radar

Capacity, cost
and boundaries.

Prices and availability below are declarations from suppliers themselves, reread in the primary source on August 13, 2026. Model price changes without warning — this radar is dated precisely because the date is part of the fact. FORGE treats model + harness + environment + eval as the decision unit; no supplier table proves universal superiority.

OpenAI balanced

GPT-5.6 Terra

Balanced tier. OpenAI reduced price by 20% on July 30, 2026 — the previous cut of this radar published the launch values.

$2input / 1M$12output / 1M−20%cut of 7/30
OpenAI efficient

GPT-5.6 Luna

Lowest cost tier. OpenAI reduced price by 80% on July 30, 2026; the value published in the previous cut was five times higher than that declared today.

$0.20input / 1M$1.20output / 1M−80%cut of 7/30
Anthropic · general use

Claude Opus 5

Claude family general purpose tier 5. Supplier offers fast mode in research preview for Opus 5, charged at US$10 / US$50 per 1M.

$5input / 1M$25output / 1Mfast modedeclared preview
Anthropic balanced

Claude Sonnet 5

Balanced tier. The introductory price became the standard: The increase to $3 / $15 scheduled for September 1, 2026 has been canceled by the supplier.

$2input / 1M$10output / 1M1Mdefault context
Google · speed-first

Gemini 3.6 Flash

Google describes it as its smartest model built for speed. There are Batch and Flex tiers at half price and Priority at 1.8×.

$1.50input / 1MUS$7.50output / 1MBatch −50%declared tier
xAI · frontier

Grok 4.6

Stated price per prompt range: Over 200k input tokens, xAI charges $4 / $12. Cached input costs $0.50.

$2input / 1M$6output / 1M≤ 200kbase price range
Correction published. The previous cut of this radar, dated August 4th, published the prices of launch of GPT-5.6 Terra and Luna. OpenAI had reduced these prices by July 30th — five days before that check — and the release announcement maintains the original values ​​in the body, with the update only in a notice at the top. Luna was published here at five times the stated price. The above values ​​were re-read in the primary source on August 13, 2026.
FORGE route rule: the choice is a testable hypothesis. A more expensive model can reduce rework; a cheaper model may overcome bounded tasks. The criterion is the result per unit of value, with quality, risk, latency and cost observed in the same harness. No number above is a Trustyu benchmark — these are the supplier's list prices, and do not measure the quality that your product will obtain.
Market and adoption

Numbers with
explicit scope.

Market statistics change quickly. Each number below carries source, population and year; it informs FORGE's design, but does not replace the product's own metrics.

Stanford HAI AI Index 2026

88% organizational adoption

The report also records 362 documented AI incidents, up from 233 in 2024, and 53% population adoption of generative AI in three years.

MIT · AI Agent Index 2025

30 agents, 45 fields, big gaps

Among 13 agents with border autonomy, only four disclosed any agentic security assessment; 20 out of 30 supported MCP.

Anthropic · Economic Index Jan 2026

52% augmentation, 45% automation

In Claude.ai's November 2025 clip, augmentation once again predominated. In API traffic, computing and math tasks reached 46%.

AI amplifies the system that already exists

DORA 2025 positions AI as an amplifier of organizational strengths and weaknesses. The greatest return comes from the work system, not from the isolated tool — a foundation compatible with FORGE's thesis.

Open DORA 2025
Articles and authors

Who is shaping
engineering harnesses.

Individual authorship appears when the official source declares it. Specs, standards, and collective reports remain attributed to the responsible institution — without inventing an author.

Justin Young · Anthropic

Effective harnesses for long-running agents

Describes initializer, incremental progress, clean state, and end-to-end verification. The article credits contributions from David Hershey, Prithvi Rajasakeran and other members of the code RL and Claude Code teams.

OpenAI collective spec

Symphony specification

Specifies orchestration, isolated workspaces, work tracking, and observability. FORGE uses the source as design input, not as proof of compliance.

Stanford · MIT · DORA

World-class research teams

AI Index Steering Committee, seven AI Agent Index experts, and the DORA program form an institutional foundation for performance, transparency, adoption, and sociotechnical systems.

Complete catalogs

40 sources.
40 auditable trails.

The list below mirrors the 40 entries from the two public catalogs. For each cut, the source catalog records publisher, class, independence group and URL; manifests and receipts add policy, snapshot and hash.

Frozen evidence packs

The July 17 cutoff supports the v1.7 benchmark; the one from August 4th supports FORGE 3.1. The model editorial radar above is more recent and remains separate from these receipts.

Symphony specificationOpenAI · official-primary · OPENAI-SYMPHONY-SPEC Codex Linux sandboxOpenAI · official-primary · OPENAI-CODEX-SANDBOX OpenAI Agents SDK tracingOpenAI · official-primary · OPENAI-AGENTS-TRACING DORA report 2025Google Cloud DORA · empirical-primary · DORA-2025 NIST Dioptra / AI RMF MeasureNIST · official-primary · NIST-DIOPTRA-AI-RMF SLSA specification 1.2OpenSSF SLSA · normative · SLSA-1.2 GitHub artifact attestationsGitHub · official-primary · GITHUB-ATTESTATIONS FinOps: measure unit costsFinOps Foundation · normative · FINOPS-UNIT-COST GitHub billing usage APIGitHub · official-primary · GITHUB-BILLING-USAGE-API AWS ADR processAmazon Web Services · official-primary · AWS-ADR Azure architecture decision recordsMicrosoft Azure · official-primary · AZURE-ADR Google small CLsGoogle · official-primary · GOOGLE-SMALL-CLS OpenAI usage costs APIOpenAI · official-primary · OPENAI-COST-API Google Cloud billing exportGoogle Cloud · official-primary · GOOGLE-CLOUD-BILLING-EXPORT OWASP AISVSOWASP Foundation · normative · OWASP-AISVS Anthropic eval cookbookAnthropic · official-primary · ANTHROPIC-EVAL-COOKBOOK Microsoft AutoGen AgentChatMicrosoft · official-primary · AUTOGEN-AGENTCHAT OpenTelemetry GenAIOpenTelemetry · official-primary · OTEL-GENAI-OBSERVABILITY GitHub billing reportsGitHub · official-primary · GITHUB-BILLING-REPORTS Anthropic Usage and Cost APIAnthropic · official-primary · ANTHROPIC-USAGE-COST-API AsyncAPI 3.1.0 overviewAsyncAPI Initiative · official-primary · ASYNCAPI-3.1.0-README AsyncAPI 3.1.0 JSON SchemaAsyncAPI Initiative · official-primary · ASYNCAPI-3.1.0-SCHEMA oasdiff OpenAPI 3.1oasdiff · official-primary · OASDIFF-OPENAPI31 oasdiff breaking changesoasdiff · official-primary · OASDIFF-BREAKING-CHANGES rswag 2.17.0rswag · official-primary · RSWAG-2.17.0 SimpleCov 1.0.3SimpleCov · official-primary · SIMPLECOV-1.0.3 coverage.py XMLcoverage.py · official-primary · COVERAGEPY-7.15.3 GitHub Code Quality coverageGitHub · official-primary · GITHUB-CODE-COVERAGE QM getting startedYC Software · official-primary · QM-GETTING-STARTED QM admin pluginYC Software · official-primary · QM-ADMIN QM portal pluginYC Software · official-primary · QM-PORTAL QM Slack integrationYC Software · official-primary · QM-SLACK QM web UIYC Software · official-primary · QM-WEB-UI QM package metadataYC Software · official-primary · QM-PACKAGE Anthropic harness designAnthropic · empirical-primary · ANTHROPIC-HARNESS-DESIGN Anthropic managed agentsAnthropic · empirical-primary · ANTHROPIC-MANAGED-AGENTS Schema.org 30.0Schema.org · vocabulary-official · SCHEMAORG-30.0 RSpec request specsRSpec · official-primary · RSPEC-RAILS-REQUEST-SPECS Ruby Coverage 4.0.6Ruby · official-primary · RUBY-COVERAGE-4.0.6 Cobertura XML DTDCobertura · official-primary · COBERTURA-2.1.1-DTD

How to read this: 40 receipts prove acquisition and source binding in the aforementioned cuts. They do not automatically attest to full source veracity, fleet adoption, customer results or operational maturity.