Pith. sign in

REVIEW 18 cited by

The Foundation Model Transparency Index

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2310.12941 v1 pith:CPBBI56C submitted 2023-10-19 cs.LG cs.AI

classification cs.LGcs.AI
keywords foundationmodeltransparencyindexmodelsaffectedassessdevelopers
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Foundation models have rapidly permeated society, catalyzing a wave of generative AI applications spanning enterprise and consumer-facing contexts. While the societal impact of foundation models is growing, transparency is on the decline, mirroring the opacity that has plagued past digital technologies (e.g. social media). Reversing this trend is essential: transparency is a vital precondition for public accountability, scientific innovation, and effective governance. To assess the transparency of the foundation model ecosystem and help improve transparency over time, we introduce the Foundation Model Transparency Index. The Foundation Model Transparency Index specifies 100 fine-grained indicators that comprehensively codify transparency for foundation models, spanning the upstream resources used to build a foundation model (e.g data, labor, compute), details about the model itself (e.g. size, capabilities, risks), and the downstream use (e.g. distribution channels, usage policies, affected geographies). We score 10 major foundation model developers (e.g. OpenAI, Google, Meta) against the 100 indicators to assess their transparency. To facilitate and standardize assessment, we score developers in relation to their practices for their flagship foundation model (e.g. GPT-4 for OpenAI, PaLM 2 for Google, Llama 2 for Meta). We present 10 top-level findings about the foundation model ecosystem: for example, no developer currently discloses significant information about the downstream impact of its flagship model, such as the number of users, affected market sectors, or how users can seek redress for harm. Overall, the Foundation Model Transparency Index establishes the level of transparency today to drive progress on foundation model governance via industry standards and regulatory intervention.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 18 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. OpenAlex reports about 43 citations worldwide. Full citation record

  1. FlexOlmo: Open Language Models for Flexible Data Use

    cs.CL 2025-07 conditional novelty 7.0 of 10

    FlexOlmo merges independently trained language-model experts, trained on private data, into a single mixture-of-experts model without joint training.

  2. Can Released LLM Vocabularies Support Token-Level Estimation of Hidden Corpora?

    cs.CL 2026-08 conditional novelty 6.0 of 10

    Released BPE vocabularies can be used to estimate per-token frequency ratios of a hidden training corpus, with mean relative errors as low as 3% in controlled settings and around 6% on SmolLM.

  3. Macro-Prudential AI Governance: A Two-Layer Early Warning and Response System for Frontier AI

    cs.CY 2026-07 conditional novelty 6.0 of 10

    A Basel-III-style two-layer system—coordinated finder-coordinator-defender reporting plus ECAR, CRTH, and ARS buffers—can detect and dampen correlated risk build-up across frontier AI labs’ internal deployments.

  4. User Privacy and Large Language Models: An Analysis of Frontier Developers' Privacy Policies

    cs.CY 2025-09 conditional novelty 6.0 of 10

    All six leading U.S. AI chatbot developers, as of May 2025, appear to train their models on users' chat data by default, often without clear opt-out options.

  5. MusGO: A Community-Driven Framework For Assessing Openness in Music-Generative AI

    cs.SD 2025-07 conditional novelty 6.0 of 10

    MusGO is a community-refined framework with 13 openness categories, applied to 16 music-generative models to produce a public openness leaderboard.

  6. Winter Soldier: Backdooring Language Models at Pre-Training with Indirect Data Poisoning

    cs.CR 2025-06 conditional novelty 6.0 of 10

    Indirect data poisoning (gradient-matching prompts) makes LLMs learn secret prompt-response pairs absent from training data, detectable with certified p-values and under 0.005% contaminated tokens.

  7. The AI Agent Index

    cs.SE 2025-02 accept novelty 6.0 of 10

    The AI Agent Index catalogs 67 deployed agentic AI systems and shows that most developers publicly disclose little about safety policies and evaluations.

  8. Bridging the Data Provenance Gap Across Text, Speech and Video

    cs.AI 2024-12 conditional novelty 6.0 of 10

    A manual audit of nearly 4,000 text, speech, and video datasets finds AI training data increasingly comes from web and social media sources, carries hidden non-commercial restrictions, and remains Western-centric with...

  9. Watermarking Training Data of Music Generation Models

    cs.LG 2024-12 conditional novelty 6.0 of 10

    Injected audio watermarks can be detected in music generated by a fine-tuned MusicGen model; simple tones work best, and a neural watermark requires dozens of repeated embeddings.

  10. Expressing stigma and inappropriate responses prevents LLMs from safely replacing mental health providers

    cs.CL 2025-04 conditional novelty 5.0 of 10

    Current large language models show stigma and give clinically inappropriate responses to common mental health symptoms, so they should not be deployed as replacement therapists.

  11. RedPajama: an Open Dataset for Training Large Language Models

    cs.CL 2024-11 conditional novelty 5.0 of 10

    The paper releases open pretraining corpora, RedPajama-V1 and V2, and shows that web-data quality signals can be used to filter V2 into competitive training sets.

  12. AI Data Development: A Scorecard for the System Card Framework

    cs.CY 2025-06 conditional novelty 4.0 of 10

    A five-area scoring rubric for AI dataset documentation, applied to four datasets, shows consistent gaps in collection ethics and preprocessing details.

  13. AI Governance through Markets

    econ.GN 2025-01 conditional novelty 4.0 of 10

    Market governance mechanisms, supported by standardized AI disclosures, can create financial incentives for responsible AI development, according to this policy paper.

  14. Generative AI regulation can learn from social media regulation

    cs.CY 2024-12 conditional novelty 4.0 of 10

    Regulatory lessons from social media, especially transparency, researcher access, trust and safety, and a global perspective, can be transferred to generative AI regulation.

  15. 7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement

    cs.CL 2024-12 conditional novelty 4.0 of 10

    The authors trained and openly released a 7B LLM, an instruction-tuned variant, a GRPO-based reasoning variant, and a VLM, claiming competitive or superior performance on zero-shot, few-shot, CoT, and VLM benchmarks.

  16. Usage Governance Advisor: From Intent to AI Governance

    cs.AI 2024-12 conditional novelty 4.0 of 10

    The paper describes an IBM proof-of-concept system that combines a knowledge graph and LLM pipelines to convert a use-case description into prioritized risks, model choices, benchmarks, and mitigation actions.

  17. Data-Centric Safety and Ethical Measures for Data and AI Governance

    cs.CY 2025-06 conditional novelty 3.0 of 10

    A conceptual framework that maps dataset safety practices to six stages of the AI lifecycle, synthesizing existing documentation and red-teaming recommendations.

  18. Generative AI in Medicine

    cs.LG 2024-12 conditional novelty 1.0 of 10

    A stakeholder-based review of generative AI use cases in medicine and the consent, privacy, transparency, hallucination, usability, equity, evaluation, and accountability challenges that stand between prototypes and s...

Pith tools