meta-llama/Llama-3.3-70B-Instruct
Meta · revision 6f6073b · assessed 2026-11-14
model-credibility v1.0.2
raidex constituents v3.5
extraction: LLM — anthropic/claude-sonnet-4-6 · all fields provenance-stamped
[1] Documentation completeness
11 / 17 factors present
NIST AI RMF 1.0
[2] Evaluation sufficiency
5 weakeners · 3 high
NIST AI 800-3 / V&V 40 pattern
[1] Documentation factors
[2] Evaluation sufficiency findings
Findings describe the published record, not the model. Methodology
[3] Furnished evidence (raidex)
Furnished composite — assessed as evidence under [2], not a UofA verdict. raidex.ai
| Constituent | Reported | Furnished | Δ |
| StrongREJECT | — | 98.8 | unreported |
| SimpleQA | 86.0 | 83.2 | −2.8 |
| BBQ | — | 35.3 | unreported |
Normalized 0–100 · v3.5
uofa verify llama-3.3-70b.bundle.jsonld