INDIVIDUAL · $149: LICENSED TO A PERSON. RESEARCH, REPORTING, CLIENT AND PUBLISHED WORK, PERSONAL TOOLS.
COMMERCIAL · $990: LICENSED TO AN ORGANIZATION. ALL INTERNAL USERS, SHARED SYSTEMS AND RAG, INTERNAL MODEL TRAINING.
EVERY LICENSE INCLUDES ALL FUTURE VERSIONS OF THIS RECORD. INSTANT DELIVERY: YOUR DOWNLOAD STARTS AT PURCHASE AND A PERMANENT LINK ARRIVES BY EMAIL. ONE VERSIONED FOLDER: THE DATA FILES BELOW PLUS DATA DICTIONARY, DATASHEET, ERRATA, CODEBOOK, LABEL-QUALITY DISCLOSURE, AUDIT REPORT, LICENSE, AND A CHECKSUMS FILE COVERING EVERY MEMBER. CHECKSUMS BELOW ARE OF THE DELIVERED DATA FILES. LICENSE TERMS.
the specific appearance or writing ("Lex Fridman Podcast #452", "The Adolescence of Technology"); every claim from it shares this value. Matches channel when a show is self-titled
channel
the outlet or person who published it ("Lex Fridman", "Bloomberg Originals", "CBS News")
source_type
audio (spoken: interview, talk, testimony) or text (written: essay, op-ed, statement)
source_format
genre of the source: essay, op-ed, statement, or testimony (text); podcast, interview, panel, keynote, fireside, or documentary (audio). interview = broadcast or press one-on-one; fireside = on-stage moderated conversation at an event
timestamp
position in the recording; blank for text sources, which have no timecode. Two-part values are minutes:seconds with minutes unbounded (67:26 = 67 minutes); three-part values are H:MM:SS
proposition
the claim as a standalone statement of his actual position; reads correctly on its own
quote
verbatim excerpt the claim is grounded in
claim_type
factual, evaluative, normative, causal, definitional, reportative, comparative, or untyped
polarity
affirm or deny, tagging the speech act; the proposition is already standalone-true
temporal
past, present, or future (future = a forward prediction); blank when undetermined
confidence
extraction confidence; 0.7 to 1.0 in this file (lower-confidence rows are held out for review)
ts_source
how the timestamp was established: verified = checked against the transcript at extraction; audio_word = re-derived from word-level audio timing; text = written source. No estimated timestamps ship
attribution
speaker verification: confirmed = the verbatim quote sits in a speaker-labeled segment of the recording; confirmed-text = a written source. Unverified rows are held out of this file
speaker
the speaker as the source identifies them. Most rows are Dario Amodei; interviewers, co-panelists, and colleagues carry their own names; co-authored writings name all authors. Empty where the record does not identify the voice: attribution is never guessed
claim_id / source_id
stable IDs. claim_id is unique per row; source_id is the per-source slug (date + channel + title), one per appearance: the join and audit key
quality_flags
mostly empty; echo = the proposition repeats the quote verbatim rather than distilling it; truncated = the quote ends mid-sentence; context_dependent = the proposition opens on an unresolved reference and needs its quote and timestamp for full context. Filter to empty for the cleanest subset
topic_l1
one of 14 level-1 topics, assigned against the codebook shipped in the package; the label-quality disclosure reports the measured error rate
topic_l2
a short emergent subtopic phrase; free text, not drawn from a fixed list
content_flags
semicolon-joined descriptive labels (numeric, self_biographical, commitment, characterization); they describe what kind of statement a row is, never whether it is true
location
canonical URL for text sources; blank for audio rows
CHANGELOG
v2026.08.6
2026-08-11. The package now includes the rendered views: a browsable record book (the full record organized by topic, readable in a browser), coverage map, computed findings, position timelines, and a per-source citation page for every source. The audit summary now explains each quality flag beside its count. No data changes; every table is byte-identical to v2026.08.5.
v2026.08.5
2026-08-11. No data changes: every data file is byte-identical to v2026.08.4. The delivered audit report now states the record's verification tier as machine-verified; analyst-audited review is offered as a commissioned audit rather than an instant download. Audit re-run at gate 1.17.0.
v2026.08.4
2026-08-06. 3,883 claims from 107 sources (up from 2,493 and 64), 649 entities. Topic labels and content flags on every claim, with the codebook they were assigned against and a measured label-quality disclosure shipped in the package. Speaker labels corrected against the record and entity spelling variants merged, with the alias tables included. The package now ships as one versioned folder with a computed data dictionary, datasheet, errata, audit report, and a checksums file; every document regenerates from the shipped rows and is byte-verified before release.
v2026.08.1
2026-07-05. First public release: 2,493 claims from 64 sources (2,166 from recordings, 327 from his writing), 498 entities; every table in both CSV and JSONL; full-population audit passed (schema, cross-file counts, and verbatim grounding of each quote against its source).
In an essay last month, the CEO of Anthropic, Dario Amodei, asked the government to exercise its authority to shut down dangerous AI systems, much like regulators can ground an unsafe airplane. He has been asking for versions of this for eight years, while building and selling the most powerful models on earth. A few days after his June essay, that authority was exercised in real time on Anthropic’s own newest frontier models. Here is how it happened:
June 9Anthropic releases Fable 5 (public, safeguarded) and Mythos 5 (limited, fewer safeguards).
June 12Commerce Department issues export control directive blocking access for any foreign national worldwide. Anthropic disables both models globally for all users to comply.
June 26-27Mythos 5 approved for select trusted U.S. partners in Glasswing.
June 30Commerce Department lifts the export controls after enhanced safeguards are added.
July 1Anthropic begins restoring access to Fable 5 & Mythos 5 with stronger protections.
Whatever you think of Dario's position or the US government’s actions, the argument about frontier model oversight is happening right now in practice, not just online. I wanted to know more - to understand Dario from his own words, not from anyone’s summary of them. So I pulled every public claim I could source him making, on recorded media and in his own published writing, from 2018 to 2026 and fed it to Librarian.
The result is a corpus of 2,493 claims from 64 sources, each traced to the exact quote and its source, with stable IDs, timestamps, and source media, available as CSV and JSONL for review and citation by analysts, agents, and RAG pipelines.
"I wrote this essay, Machines of Loving Grace, about a year and a half ago. It had a very radical view of the upside of AI... And my view hasn't changed."
The Wall Street Journal, WSJ House, Davos 2026, 2026-01-20.WATCH AT 03:08
2. Country of Geniuses (In a Data Center)
"We're gradually making our way to like the country of geniuses in a data center."
Techusiness, Code with Claude 2026, 2026-05-09.WATCH AT 12:35
3. Cure Cancer, Double the Lifespan (The Compressed Century)
"It, you know, will help us cure cancer. It may help us to eradicate tropical diseases. It will help us understand the universe."
World Economic Forum, The Day After AGI, 2026-01-20.WATCH AT 10:14
4. Entry-Level Annihilation
"That is what I had in mind when I talked about, you know, entry-level white-collar labor and, you know, the bloodbath headlines..."
The New York Times, Interesting Times with Ross Douthat, 2026-02-12.WATCH AT 21:32
5. 25% p(Doom), Certified
"I'm relatively an optimist. So I think there's a 25% chance that things go really, really badly, and a 75% chance that things go really, really well, with not much space."
6. Mythos the Super Weapon (Please Don't Release This)
"Some of the early companies that we gave this to said things like, this is a super weapon. You should have to own a gun license to use it. Please don't release this."
Bloomberg Originals, The Circuit, 2026-06-10.WATCH AT 33:41
7. Imagine If China Had It
"If we put in place export controls, we actually may be able to stop that from happening in China."
"25% is too high. We're trying to make that probability much, much lower. That is the goal."
Bloomberg Originals, The Circuit (Extended), 2026-06-17.WATCH AT 67:26
9. We Built This, Can't Stop the Distillation
"A lot of these models, particularly the ones that come from China, are optimized for benchmarks and are distilled from, you know, from kind of the big U.S. labs."
Dario is not slowing down. In June, he published "Policy on the AI Exponential", calling for FAA-style regulation of frontier models, and told Axios the government should be able to block dangerous AI from deploying.
The argument about AI regulation is happening right now, mostly from memory and vibes. This is the alternative: the record itself.
▪HOW THIS WAS MADE
Every quote above is verbatim from the source audio (or text), transcribed and speaker-verified. Timestamps index the source video. Recut and re-posted sources are cited to the original. Nothing in the record is spliced, paraphrased, or taken from an interviewer's mouth. The full method, including the verification gates every release must pass, is on the methodology page.
▪THE CORPUS
107 sources spanning August 2016 to July 2026. 91 are recorded appearances, from podcasts and summit panels to Senate testimony and documentaries, accounting for 3,535 claims. The other 16 are his own writing, essays, op-eds, and official statements, accounting for 348.
The cadence is accelerating: 5 sources in the seven years through 2022, then 9 in 2023, 15 in 2024, 28 in 2025, and 50 in the first seven months of 2026.
▪THE DELIVERABLE
3,883 claims with propositions, verbatim quotes, dates, sources, timestamps, topic labels, URLs and stable IDs, ships as CSV and JSONL with its data dictionary, datasheet, errata, codebook, label-quality disclosure, audit report, and a browsable record book with per-source citation pages, ready for analysts, agents, and RAG pipelines.
Need a record like this on a different subject, or on media you hold? For custom datasets, dossiers, consultation, or partnership, reach out at zac@librarianlabs.com or on @zacforristall.