USASI

Methodology · USASI rubric v0.1

Methodology

This page explains how entries get into the catalog, how openness is described, what counts as evidence, and how the numbers on the site are calculated. The same rules are published in the repository as ELIGIBILITY.md, OPENNESS.md, and CONTENT_REVIEW.md.

Eligibility

An entry is eligible when a documented basis connects its accountable entity to the United States. There are four bases:

  • U.S. headquarters — the organization's documented headquarters or principal executive office is in the United States.
  • U.S. nonprofit or lab — the entity is a U.S. nonprofit, university laboratory, or research institute.
  • Documented U.S. control — the entity is documented as primarily controlled by a U.S. entity, for example a research unit wholly owned by a U.S.-headquartered company.
  • U.S.-governed project — for software, datasets, and models: the documented governing or maintaining entity is U.S.-based, such as a U.S. foundation or company.

Some cases do not qualify on their own. A foreign organization with only a U.S. sales office is not eligible on that basis. Organizations with headquarters in two countries, and parent–subsidiary arrangements, get a written assessment on the entry that cites the documents it relies on. Project eligibility rests on documented governance, never on the names or assumed nationalities of contributors. A U.S. fine-tune or wrapper of foreign base weights does not make the base model American; the relationship is recorded as provenance.

There are no exceptions for prominent organizations. Candidates whose basis cannot be documented stay in an internal review queue and are not published. International collaboration is described accurately where it exists; non-U.S. models or dependencies may appear as clearly labeled provenance context, never as catalog members.

Each entry shows its eligibility status (Eligible, Pending review, or Excluded), the basis, a short explanation, the sources, and the date it was assessed. Only eligible entries are published.

Openness

Openness is described per artifact with a checklist suited to what the artifact is. A language model, a compiler, and a dataset expose different things, so the catalog does not force every artifact onto a single scale built for model weights.

  • PublicDocumented as available to the general public. Access conditions and license terms may still apply.
  • PartialSome of it is available, or access requires approval or is otherwise restricted.
  • Not publicDocumented as not publicly available.
  • UnknownNot yet assessed, or the evidence is insufficient.
  • Not applicableDoes not apply to this kind of artifact.

Model releases

USASI rubric v0.1 checklist for model
ItemQuestion
WeightsCan the general public download the model parameters for this release?
Inference codeIs code for running the model published?
Training codeIs the code used to train the model published?
Training-data informationPublic = the training data itself can be obtained. Partial = composition or sources are documented without full access.
Training recipeAre the training configuration and procedure documented in enough detail to follow?
Evaluation materialsPublic = evaluation code or prompts that let others re-run the evaluations are published. Partial = results only.

Frameworks and runtimes

USASI rubric v0.1 checklist for framework and runtime
ItemQuestion
Source codeIs the source code publicly readable?
DocumentationIs user documentation published?
InstallationAre installation instructions or packages publicly available?
Supported platformsAre supported operating systems or hardware documented?
Release statusAre versioned releases published?

Datasets

USASI rubric v0.1 checklist for dataset
ItemQuestion
AccessCan the data be obtained, and on what terms?
ProvenanceAre the data's origins documented?
DocumentationIs there a datasheet, card, or equivalent documentation?
LicensingAre the licensing terms stated?
Stated limitationsDoes the documentation state known limitations or risks?

Evaluation tools

USASI rubric v0.1 checklist for evaluation tool
ItemQuestion
CodeIs the evaluation code published?
Tasks / dataAre the tasks or test data available?
MethodologyIs the method for scoring described?
ReproducibilityAre instructions for reproducing results published?
LimitationsAre known limitations documented?

Research stacks

USASI rubric v0.1 checklist for research stack
ItemQuestion
Source codeIs the source code publicly readable?
DocumentationIs user documentation published?
Training codeDoes it include code for training models?
Data informationAre the data it expects or ships with documented?
ReproducibilityAre instructions for reproducing reported results published?

Availability is kept separate from permission. "Publicly downloadable" does not by itself mean unrestricted, open source, commercially usable, or redistributable. Each release lists its actual licenses and what they apply to.

Model-disclosure tiers (USASI rubric v0.1)

For model releases only, the checklist and licenses determine a disclosure tier. The tiers are USASI's own editorial categories, versioned as rubric v0.1. They are not an external certification, and they are not the Open Source Initiative's Open Source AI Definition or any other outside standard. The tier is computed automatically from the recorded checklist, so it cannot be typed in by hand or inherited from a family or an organization.

USASI rubric v0.1 model-disclosure tiers
TierRequirement
Model-disclosure tier (USASI rubric v0.1): Open-weightOpen-weightThe model parameters for this release can be downloaded by the public. License terms may still restrict use, redistribution, or commercial use.
Model-disclosure tier (USASI rubric v0.1): Open-stackOpen-stackOpen-weight, plus published inference code, training code, and training recipe, and at least documented training-data composition.
Model-disclosure tier (USASI rubric v0.1): Fully openFully openEvery item in the model checklist is public, including the training data itself, and the weights and code are under OSI-approved licenses.
Model-disclosure tier (USASI rubric v0.1): Restricted weightsRestricted weightsThe weights can be obtained only by request, with approval, or by some users — for example a gated download that the publisher reviews. Not counted as open-weight.
Model-disclosure tier (USASI rubric v0.1): Weights not publicWeights not publicThe weights for this release are documented as not publicly available.
Model-disclosure tier (USASI rubric v0.1): UnknownUnknownThe availability of the weights has not been assessed, or the evidence is insufficient.

Licenses the rubric treats as OSI-approved: Apache-2.0, MIT, BSD-2-Clause, BSD-3-Clause, MPL-2.0, ISC, GPL-2.0-only, GPL-2.0-or-later, GPL-3.0-only, GPL-3.0-or-later, LGPL-2.1-only, LGPL-2.1-or-later, LGPL-3.0-only, LGPL-3.0-or-later, AGPL-3.0-only, AGPL-3.0-or-later, EPL-2.0. Any other license — including custom model licenses — does not satisfy the fully open requirement until the rubric is revised.

Tiers are cumulative: every fully open release also counts as open-stack, and every open-stack release also counts as open-weight. When tiers appear as columns, the columns overlap and must not be added together.

Closed product/API is a separate label used on organization pages. It describes how a documented product is delivered, such as a hosted API or assistant app. It is not an artifact tier, and it says nothing about other releases from the same organization. An organization is never labeled "open" because it has released one open-weight model.

Source standards

Each claim cites the source that specifically supports it. In order of preference: the applicable license file; model and dataset cards; official documentation and release notes; the project's repository; the organization's own pages; and regulatory filings for facts such as headquarters. Reputable reporting is used only for the specific claim it supports and is labeled as a news report. An organization's homepage is not adequate evidence for a detailed technical claim.

Every source records its title, publisher, URL, publication date if the page states one, and the date an editor read it. The catalog does not cite search snippets, reproduce source articles, or invent URLs. It does not publish speculative versions, acquisitions, partnerships, staffing, funding, valuations, benchmark scores, or power-capacity figures.

Families and releases

A model family record (for example, a named line of models) summarizes the line and links to its releases. A model release record describes one specific published release. Licenses, availability, checklists, and tiers attach only to releases, because they can differ between releases in the same family. Family overviews are never counted as releases, and a family never carries its own tier.

Software, datasets, and evaluation tools are project records.

Dates

  • Released — the documented release date of an artifact, at the precision the evidence gives (a year, a month, or a day). Unknown if not documented.
  • Entry updated — the date the record was last substantively edited.
  • Last reviewed — the date an editor last checked the record's evidence.
  • Site generated — the date the website was built, shown in the footer.

A build never changes the first three. "Recently reviewed" on the homepage means recently checked, not newly released.

Uncertainty

Unknown means the item has not been assessed or the evidence is insufficient. It is never a polite way of saying "no." Not public means the evidence documents that something is not available. Where an entry's facts are contested or ambiguous, the record says so in plain language rather than choosing a side.

How counts work

Every total on the site is catalog coverage: the number of published records in this catalog. It is not a census of American AI, and a zero means only that no matching record has been published here.

  • Organization records include research units and subsidiaries that have their own records. They are also counted separately as "units and subsidiaries," so a parent and its unit are never presented as two independent companies.
  • Model families and model releases are counted separately.
  • Artifacts are linked to every organization named on the artifact record, so a release maintained by two organizations appears under both. Per-organization columns are therefore never summed.
  • Drafts and archived historical records are excluded from all counts, search, exports, and the sitemap. Archived records keep a clearly labeled historical page so links do not break.

Every number on the comparison workspace links to exactly the records it counts.

Equal structural representation

The two directories — organizations, and open artifacts — have the same prominence in navigation, on the homepage, and in search. Equal prominence does not mean equal numbers: the catalog lists what it can verify, and one directory may be longer than the other. Homepage selections are editorial, explained on the page, alphabetical or date-ordered by default, and never influenced by tips.

Reading the badges

Status badges always pair an icon with a text label, so no information depends on color. Numbered markers after a statement link to the source list at the bottom of the entry. Links marked with an arrow lead to external sites; everything else stays on USASI.

Support Us

Help keep USASI useful.

Find the catalog useful? Leave an optional tip to support its upkeep. Tips never affect listings, coverage, or openness assessments.

Optional. No USASI account required. Payment takes place on the linked provider’s website (Buy Me a Coffee).

About supporting this project