Research Note
Multimodal Claim Evidence Ladder
| State | What it establishes | |---|---| | Concept | A proposed architecture, task, or research direction | | Internal experiment | A tested setup under reported conditi
In this article
Multimodal Claim Evidence Ladder
Evidence states
| State | What it establishes |
|---|---|
| Concept | A proposed architecture, task, or research direction |
| Internal experiment | A tested setup under reported conditions |
| Benchmark result | A measured result for a named artifact, dataset, metric, and baseline |
| Demonstration | A selected interaction or prototype path |
| Released artifact | Weights, code, or another reproducible object with terms |
| Accessible interface | An API or product surface available to a defined audience |
| Supported capability | A documented, maintained feature with operational boundaries |
Claim record
Record exact input and output modalities, task, model path, component versions, training and evaluation data, metric, comparator, limitation, artifact identity, access, terms, product surface, support state, geography, account, and verification date.
Historical application
The Llama 3 paper reported compositional image, video, and speech experiments. Its research page explicitly said the resulting models were still under development and not broadly released. The July 2024 Llama 3.1 model card identified the released family as text in and text out.
Later multimodal Llama releases are descendant evidence. They do not make the 2024 text artifacts multimodal retroactively.
Sources
Follow the evidence.
- ai-challenges.nist.gov: ariaai-challenges.nist.gov
- ai-challenges.nist.gov: genaiai-challenges.nist.gov
- ai.meta.com: meta llama 3 1 ai responsibilityai.meta.com
- ai.meta.com: the llama 3 herd of modelsai.meta.com
- crfm.stanford.edu: indexcrfm.stanford.edu
- csrc.nist.gov: red teamingcsrc.nist.gov
- daltonanderson.ghost.io: metas llama 3 safety scaling and simple solutionsdaltonanderson.ghost.io
- github.com: MODEL CARDgithub.com
- github.com: MODEL CARDgithub.com
- github.com: PurpleLlamagithub.com
- github.com: MODEL CARDgithub.com
- huggingface.co: concept guidehuggingface.co
- mlcommons.org: safety faqmlcommons.org
- mlcommons.org: jailbreak 0 7mlcommons.org
- mlcommons.org: safety methodologymlcommons.org
- NIST Generative AI Profilenvlpubs.nist.gov
- open.spotify.com: 44o5OPSumaZJcvRkXutorBopen.spotify.com
- owasp.org: www project top 10 for large language model applicationsowasp.org
- NIST AI Risk Management Frameworknist.gov
- youtu.be: 1KNOcY e9Tsyoutu.be