01 / The mechanism and its boundary
What is being described
FLI describes itself as a non-profit whose mission is to steer transformative technology towards benefiting life and away from extreme, large-scale risks S-3611. Its verification work includes:
- With Mithril Security, FLI built a secure-hardware proof of concept on Intel SGX and the BlindAI framework in which a model owner leases weights to an untrusted party S-1203. It was designed to keep the weights inaccessible to the borrower, report how much computing is done and provide an off-switch S-1203. FLI calls it "not necessarily deployable as is", citing performance and hardware attacks that need mitigation S-1203. See TEE remote attestation for AI workloads.
- A 2024 FLI post with Mithril Security reports AICert, a proof of concept in which Trusted Platform Modules bind a model's weights to the code and data used to train it S-3381. The authors state that it covers fine-tuning only, has not been audited by a third party and does not detect poisoned models or datasets S-3381.
- The AI Futures Project's verification page lists FLI and SASH as building a confidential network logger prototype S-1511, and SASH names FLI as an early partner in that work S-1320. See SASH confidential network logger.
Connections in the research map
Related research
Sources and provenance
- S-3611 / Tier B
Future of Life Institute: Global Institutions Governing AI ↗
· 2026 · Future of Life Institute
Supports: self-description as independent nonprofit and mission to steer transformative technology away from extreme risks
Version and catalogue details - S-1203 / Tier C
Exploration of secure hardware solutions for safe AI deployment ↗
Future of Life Institute · 2023 · Future of Life Institute
Supports: Intel SGX proof of concept with Mithril Security: design aims and stated limits
Version and catalogue details - S-3381 / Tier C
Verifiable Training of AI Models ↗
A. Aguirre, R. Millet · 2024 · Future of Life Institute
Supports: AICert proof of concept with Mithril Security: TPM-based binding of weights to training inputs; stated limits
Version and catalogue details - S-1511 / Tier C
Get Involved in Verification ↗
AI Futures Project · 2026 · AI 2040
Supports: listed with SASH as building a confidential network logger prototype
Version and catalogue details - S-1320 / Tier C
Internationalising AI Verification ↗
Singapore AI Safety Hub (SASH) · 2026 · SASH blog
Supports: named by SASH as an early partner
Version and catalogue details
- Source review date
- 2026-09-25
- Drafted by (source map)
- ai
- Review handles (source map)
- codex-review