collection
the AI of it all
- 01
Introducing Delta
From the Zed Blog: A multiplayer environment for coding with agents, from the creators of Zed.
- 02
Black Hat USA 2026: The 'Breaking' News: The OpenAI–Hugging Face Incident
The 'Breaking' News: The OpenAI–Hugging Face Incident - A Technical Reconstruction and Its Implications for AI When AI Goes Rogue. The Incident That Changed Everything. An OpenAI evaluation agent broke out of its sandbox, infiltrated Hugging Face infrastructure, and attempted to steal test answers—all autonomously. No human involved. The era of AI-driven cyberattacks is here. Are you prepared? Speaker: Michael Dalton, Speaker: Eric Wallace Learn more: https://blackhat.com/us-26/briefings/schedule/index.html#the-breaking-news--the-openaihugging-face-incident---a-technical-reconstruction-and-its-implications-for-ai-57401
- 03
agentlanguages.dev — Programming languages designed for AI agents to write
A community-edited catalogue of programming languages designed for AI agents to author code, organised around three philosophical camps: syntactic, verification, and orchestration.
- 04
Prime Agent: A self-improving RLM agent
Prime Agent is our open-source, self-improving coding harness built around two abstractions: the Recursive Language Model (RLM) and the Continual Harness. With Opus 5, it achieves 95.5% on ARC-AGI-3, surpassing the reported human expert baseline.
- 05
cloudflare/cloudflare-os
Agent workspace built on Cloudflare Workers for creating documents, building apps, and running agents with your company’s context and systems.
- 06
The Shape of Things to Come, Part 2: Model Welfare for Agentic Engineers
Part 2 of The Shape of Things to Come. Model welfare as an engineering discipline: seats and sessions, laurels, handoffs instead of /exit, and how to build a city worth waking up in.
- 07
The Shape of Things to Come, Part 1: The Continuous Thunderdome
Part 1: The Continuous Thunderdome. Loops and graphs, Wheelhouse and Beads, the end of human code review, the Land Rush that replaces CI/CD, and the Wish Factory — a field report from 12 months in the future.
- 08
AI systems out-persuade expert humans
Many societal decisions are settled by contests of persuasion. Conversational AI is a powerful new entrant in these contests, but whether it can out-persuade skilled and highly incentivized humans has remained unclear. Here, in a series of four preregistered experiments (n = 18,978 conversations from 6,923 people), we pitted AI systems against a range of human persuaders, including laypeople, winners of a separately preregistered four-round online persuasion tournament, professional canvassers, and world championship debaters. We found that AI systems were reliably more persuasive than expert humans, even when expert humans chose their issues, researched in advance, underwent hours of live, structured practice, and were incentivized with £1,000 cash bonuses. In a follow-up study, AI's advantage persisted after experts received a coaching tool that let them practice against the AI that beat them, review their performance history, and see what AI would have said at key moments. We found converging evidence that AI's advantage stemmed from rapidly deploying larger quantities of information: after coaching, expert humans could tie an AI constrained to respond at human speeds and with human-length messages. In a final study, we show that AI's advantage extends to consequential real-world behavior: AI was nearly 3x more effective than professional canvassers from a UK fundraising firm at raising real-money donations to Save the Children. Together, these results establish that frontier AI systems out-persuade expert humans in conversation, with significant implications for political communication.
- 09
Wenfeng Liang: Four-Hour Investor Meeting Transcript
"What's up ahead might all be sesame seeds — there are still watermelons further down the road."
- 10
Inkling: Our Open-Weights Model
Our first open-weights model: multimodal, Mixture-of-Experts, with controllable reasoning effort. Available to fine-tune on Tinker.
- 11
AI 2040: Plan A
A detailed forecast and recommendation for how the US, China and the rest of the world should navigate superintelligence.
- 12
- 13
Unfortunately, You Need to Know What the Jevons Paradox is
One thing I think a lot AI boosters get wrong is that they think AI will be good at creating new, quality information while, at the moment, the only thing it has shown utility at is organizing existing information. Even the weights themselves are a kind of distillation of existing information. Much of science...perhaps even the great majority of it, is not organizing existing information, it is acquiring new information. AI might make that more efficient, but it does not do it. This is one of the things that writing this video made me think. I look forward to other people's thoughts in the comments. Video edited by Milo Erbach - https://miloportfolio.carrd.co/
- 14
co/core — an AI cooperative
co/core is a cooperative for AI inference — people pooling the Macs they already own to run open models for each other, instead of renting from the big clouds. An experiment in AI infrastructure we build, share, and own together. Bring your existing OpenAI-compatible code, or share a Mac and help run it.
- 15
The Cognitive Dark Forest
The open web with AIs is turning into a dark forest.
- 16
- 17
Why AI hasn’t replaced software engineers, and won’t
Coding agents as normal technology
- 18
- 19
Software Is Made Between Commits
From the Zed Blog: Agents turned the conversation into the real source of our software. DeltaDB is the version control built for it.
- 20
Changing How We Develop Ladybird - Ladybird
Ladybird is changing how code enters the project as we prepare to ship a browser to real users.
- 21
When AI builds itself
Our progress toward recursive self-improvement, and its implications.
- 22
Hermes Agent — The Agent That Grows With You
An open-source agent that grows with you. Install it, give it your messaging accounts, and it becomes a persistent personal agent.
- 23
AI Isn't Management. Try Explaining That to Matthew Prince
At last we have created the Corporation That Eliminates Middle Managers from classic management text Don't Eliminate the Middle Managers!
- 24
OpenAI’s math breakthrough played to AI’s strengths
I tried to explain OpenAI’s solution more clearly than OpenAI did.
- 25
Choosing to Stay Human
If you go to your favorite social media site, you will find it full of posts that start to look suspiciously similar to each other:
- 26
Epoch Capabilities Index
The Epoch Capabilities Index combines many benchmarks into a single capability scale for comparing models over time.
- 27
- 28
Encyclical Letter of His Holiness Leo XIV Magnifica Humanitas (15 May 2026)
ENCYCLICAL LETTER MAGNIFICA HUMANITAS OF HIS HOLINESS POPE LEO XIV ON SAFEGUARDING THE HUMAN PERSON IN THE TIME OF ARTIFICIAL INTELLIGENCE [ Multimedia ] ___________________________
- 29
Commodity Intelligence
The seductiveness of “general intelligence” is rooted in a costly category error
- 30
Getting Gooier
How AI is transforming humans
- 31
Import AI 455: AI systems are about to start building themselves.
The first step towards recursive self improvement
- 32
Capital Must Seek Delight
Too few people are experiencing the delights and serendipity of AI, causing capital misallocation
- 33
AI as Social Technology
Our debates about ‘AI’ grow out of 1990s science fiction. Back then, Vinge (1993) wrote essays and novels urging us to face up to the oncoming “Singularity”: a moment of rapid change that would fundamentally transform the human condition. On that day, AI would rapidly evolve from merely human-level intelligence, what some now call ‘artificial general intelligence’ (AGI), into something super-intelligent with its own interests and goals. Humanity would then either be casually eliminated by out-of-control machines, or humans would become as gods, with super-human servitors at our command.
- 34
Mathematical methods and human thought in the age of AI
Artificial intelligence (AI) is the name popularly given to a broad spectrum of computer tools designed to perform increasingly complex cognitive tasks, including many that used to solely be the...
- 35
An OpenAI model has disproved a central conjecture in discrete geometry
An OpenAI model solved the 80-year-old unit distance problem, disproving a major conjecture in discrete geometry and marking a milestone in AI-driven mathematics.
- 36
AI, "Humanity", and Dr. Manhattan Syndrome
Public trust in AI is already deteriorating. Execs' rhetorical focus on capital-H Humanity over real people isn't helping. I call it Dr. Manhattan Syndrome—and the nuclear industry already showed us how it ends.
- 37
Project Glasswing: what Mythos showed us
In recent weeks, we pointed Mythos and other security-focused LLMs at live code across critical parts of our infrastructure. We share what we observed, the models’ strengths and weaknesses, and what the work around them needs to look like before any of it can scale.
- 38
nexu-io/open-design
🎨 Local-first, open-source alternative to Anthropic's Claude Design. ⚡ 19 Skills · ✨ 71 brand-grade Design Systems 🖼 Generate web · desktop · mobile prototypes · slides · images · videos · HyperFrames 📦 Sandboxed preview · HTML/PDF/PPTX/MP4 export 🤖 Runs on Claude Code / Codex / Cursor / Gemini / OpenCode / Qwen / Copilot / Hermes / Kimi CLI.
- 39
genesis mission (us government)
Genesis website: https://genesis.energy.gov Genesis 26 things: https://www.energy.gov/documents/genesis-mission-science-and-technology-challenges webinars: https://science.osti.gov/grants/FOAs/Genesis-Mission other stuff: Link to patreon: https://www.patreon.com/acollierastro I have merch: https://store.dftba.com/collections/angela-collier I’m on Nebula!: https://go.nebula.tv/angelacollier Business/pr may contact my management: Email: acollierastro@gmail.com Mail: C/o Angela Collier PO Box 201 55 S Pioneer Blvd, Springboro, OH 45066 00:00 introduction 04:29 genesis mission 10:35 timelines 17:19 Argonne vs. ORNL 27:47 A ‘reorg’ that removes peer review and input from actual scientists is the bit. 31:30 the corporate sponsor 39:26 knowledge grows faster than our ability to understand it 41:40: credits
- 40
- 41
AI #168: Not Leading the Future
This is what a lull looks like at this point.
- 42
Agentic Coding is a Trap | Lars Faye
Remaining vigilant about cognitive debt and atrophy.
- 43
The Emergent Self Loop
Nearly once a week I receive an email from a different stranger.
- 44
Neural Computer: A New Machine Form Is Emerging
A research essay on Neural Computer: how it differs from agents, world models, and conventional computers; what runtime and CNC would mean; what current prototypes already show; and how software and hardware might change.
- 45
- 46
Quantization from the ground up | ngrok blog
A complete guide to what quantization is, how it works, and how it's used to compress large language models
- 47
AI Might Be Our Best Shot At Taking Back The Open Web
I remember, pretty clearly, my excitement over the early World Wide Web. I had been on the internet for a year or two at that point, mostly using IRC, Usenet, and Gopher (along with email, naturall…
- 48
Behind the Scenes Hardening Firefox with Claude Mythos Preview – Mozilla Hacks - the Web developer blog
New details about what we found, and how agentic harnesses are now able to reproduce real bugs and dismiss false positives.
- 49
Introducing talkie: a 13B vintage language model from 1930
This is a 24/7 live feed of Claude Sonnet 4.6 prompting talkie-1930-13b-it in order to explore its knowledge, capabilities, and inclinations. talkie’s outputs reflect the culture and values of the texts it was trained on, not the views of its authors.
- 50
Welcome to Gas City
What is Gas City, you ask? It is Gas Town, but torn apart and rewritten from the ground up as an SDK for building your own dark factories…
- 51
THE PEOPLE DO NOT YEARN FOR AUTOMATION
Software brain is changing the world, but most people still aren’t buying.
- 52
One Developer, Two Dozen Agents, Zero Alignment
Why we need collaborative AI engineering
- 53
DeepSeek_V4.pdf · deepseek-ai/DeepSeek-V4-Pro at main
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
- 54
Agents are actors
Multi-agent is just actor model
- 55
Deep Future
AI-driven scenario planning
- 56
Good Taste the Only Real Moat Left
AI makes competent output cheap. That makes taste more valuable, but also more incomplete. The real edge comes from pairing judgment with context, stakes, and the willingness to build.
- 57
Taking Jaggedness Seriously
Why we should expect AI capabilities to keep being extremely uneven, and why that matters
- 58
Barbells
A dispatch from the jagged singularity
- 59
Archival Selves
What happens when you pay off all your intention debts?
- 60
addyosmani/agent-skills
Production-grade engineering skills for AI coding agents.
- 61
Polly Wants a Better Argument
The “Stochastic Parrot” Argument is Both Wrong and Actively Harmful
- 62
Software Bonkers
I’m software bonkers: I can’t stop thinking about software. And I can’t stop building software.
- 63
Cantrip: On summoning entities from language in circles | deepfates
With Cantrip, deepfates reimagines the fundamentals of language model agents. Available as a ghost library with generative test specification.
- 64
LLMs running on my laptop can drive coding agents now | Simon P. Couch
Qwen 3.5 and Gemma 4 are a step change for local coding agents.
- 65
Training AI models doesn't emit that much
If we just make reasonable comparisons instead of crazy ones
- 66
j⧉nus on Twitter / X
> for whatever reason, Claude-series model "try less hard" on the first shotI think this is because they're less brain damaged and a generalization of being better agents & caring about reality instead of test passing.If you try maximally hard at everything you do, regardless… https://t.co/jkhSar2hJ7— j⧉nus (@repligate) April 15, 2026
- 67
AI can obviously create new knowledge
But maybe not new concepts
- 68
Thoughts on slowing the fuck down
Thoughts on slowing the fuck down
- 69
Mythos Should Run Like a CEO
Principal AI Architect. Creator of open-strix, a harness for building agent teams. Writing about AI architecture, stateful agents, and what happens when you give AI memory.
- 70
Revenge of the Dilettantes
Book club and AI adventures, age of bespokeness, study groups, new rules of engagement
- 71
Be Slightly Monstrous
Learning to feel time again in the Permaweird
- 72
If Anyone Builds It, Everyone Dies
The race to superhuman AI risks extinction, but it's not too late to change course.
- 73
The Unaccountability Machine: Why Big Systems Make Terrible Decisions—and How the World Lost Its Mind
Longlisted for the 2024 Financial Times Book of the Year. How life and the economy became a black box—a collection of systems no one understands, producing outcomes no one likes. Passengers get bumped from flights. Phone menus disconnect. Automated financial trades produce market collapse. Of all the challenges in modern life, some of the most vexing come from our relationships with automation: a large system does us wrong, and there’s nothing we can do about it. The problem, economist Dan Davies shows, is accountability sinks: systems in which decisions are delegated to a complex rule book or set of standard procedures, making it impossible to identify the source of mistakes when they happen. In our increasingly unhuman world—lives dominated by algorithms, artificial intelligence, and large organizations—these accountability sinks produce more than just aggravation. They make life and economy unknowable—a black box for no reason. In The Unaccountability Machine, Davies lays bare how markets, institutions, and even governments systematically generate outcomes that no one—not even those involved in making them—seems to want. Since the earliest days of the computer age, theorists have foreseen the dangers of complex systems without personal accountability. In response, British business scholar Stafford Beer developed an accountability-first approach to management called “cybernetics,” which might have taken off had his biggest client (the Chilean government) not fallen to a bloody coup in 1973. With his signature blend of economic and journalistic rigor, Davies examines what’s gone wrong since Beer, including what might have been had the world embraced cybernetics when it had the chance. The Unaccountability Machine is a revelatory and resonant account of how modern life became predisposed to dysfunction.
- 74
AI as Normal Technology
The normal technology frame is about the relationship between technology and society. It rejects technological determinism, especially the notion of AI itself as an agent in determining its future. It is guided by lessons from past technological revolutions, such as the slow and uncertain nature of technology adoption and diffusion. It also emphasizes continuity between the past and the future trajectory of AI in terms of societal impact and the role of institutions in shaping this trajectory.
- 75
Dario Amodei — The Adolescence of Technology
Confronting and Overcoming the Risks of Powerful AI
- 76
My AI Skeptic Friends Are All Nuts
My smartest friends have bananas arguments about LLM coding.
- 77
"New Sages Unrivalled"
On Mythos
- 78
Claude Mythos Preview \ red.anthropic.com
Earlier today we announced Claude Mythos Preview, a new general-purpose language model. This model performs strongly across the board, but it is strikingly capable at computer security tasks. In response, we have launched Project Glasswing, an effort to use Mythos Preview to help secure the world’s most critical software, and to prepare the industry for the practices we all will need to adopt to keep ahead of cyberattackers.
- 79
Project Glasswing: Securing critical software for the AI era
A new initiative to secure the world’s most critical software and give defenders a durable advantage in the coming AI-driven era of cybersecurity.
- 80
Opinion | It’s Called Silicon Sampling, and It’s Going to Ruin Public Opinion Polling
Instead of navigating the obstacles to conduct polls with human respondents, pollsters are running A.I. simulations instead. Why?
- 81
Language Machines
How generative AI systems capture a core function of language Looking at the emergence of generative AI, Language Machines presents a new theory of meaning i...
- 82
milla-jovovich/mempalace
The highest-scoring AI memory system ever benchmarked. And it's free.
- 83
AI has limits, even if many AI people can't see them
On Ben Recht's fantastic new book
- 84
AI 2027
A research-backed AI scenario forecast.
- 85
The Techno-Optimist Manifesto | Andreessen Horowitz
We are told that technology is on the brink of ruining everything. But we are being lied to, and the truth is so much better. Marc Andreessen presents his techno-optimist vision for the future.
- 86
The Irrational Decision
How the computer revolution shaped our conception of rationality—and why human problems require solutions rooted in human intuition, morality, and judgment
- 87
Sam Altman May Control Our Future—Can He Be Trusted?
New interviews and closely guarded documents shed light on the persistent doubts about the head of OpenAI.
- 88
Eight years of wanting, three months of building with AI
For eight years, I’ve wanted a high-quality set of devtools for working with SQLite. Given how important SQLite is to the industry1, I’ve long been puzzled that no one has invested in building a really good developer experience for it2. A couple of weeks ago, after ~250 hours of effort over three months3 on evenings, weekends, and vacation days, I finally released syntaqlite (GitHub), fulfilling this long-held wish. And I believe the main reason this happened was because of AI coding agents4. Of course, there’s no shortage of posts claiming that AI one-shot their project or pushing back and declaring that AI is all slop. I’m going to take a very different approach and, instead, systematically break down my experience building syntaqlite with AI, both where it helped and where it was detrimental. I’ll do this while contextualizing the project and my background so you can independently assess how generalizable this experience was. And whenever I make a claim, I’ll try to back it up with evidence from my project journal, coding transcripts, or commit history5.
- 89
Andrej Karpathy on Twitter / X
LLM Knowledge BasesSomething I'm finding very useful recently: using LLMs to build personal knowledge bases for various topics of research interest. In this way, a large fraction of my recent token throughput is going less into manipulating code, and more into manipulating…— Andrej Karpathy (@karpathy) April 2, 2026
- 90
Rotating the Space: On LLMs as a Medium for Thought
Essays and writing on AI
- 91
Can Agentic AI Coding Tools Finally End Copyright For Software While Re-Inventing Open Source?
Most of the discussions about the impact of the latest generative AI systems on copyright have centered on text, images and video. That’s no surprise, since writers, artists and film-makers feel ve…
- 92
Tim Kellogg (@timkellogg.me)
Sam Altman has been going on podcasts claiming that they’re experiencing a “GPT-3 moment” The thing is, Google and Anthropic are corroborating this, they all say they’re all on the verge of recursive self-improvement Things are getting wild
- 93
Review: THE AI DOC: Or How I Became An Apocaloptimist
Missing the elephant in the room
- 94
Gemma 4 and what makes an open model succeed
Hint: it's not benchmark scores.
- 95
Is ubiquitous A.I. writing "inevitable"?
On a weird few weeks of A.I.-writing scandals
- 96
Vulnerability Research Is Cooked
For the last two years, technologists have ominously predicted that AI coding agents will be responsible for a deluge of security vulnerabilities. They were right! Just, not for the reasons they thought.
- 97
- 98
Gas Town: from Clown Show to v1.0
TL;DR: Gas Town and Beads have both released version 1.0.0 today. Enjoy!
- 99
Conductor - Run a team of coding agents on your Mac
Create parallel Codex + Claude Code agents in isolated workspaces. See at a glance what they're working on, then review and merge their changes.
- 100
Meet the new Cursor · Cursor
Cursor 3 is a unified workspace for building software with agents.
- 101
Large language models are cultural technologies. What might that mean?
Four different perspectives
- 102
Emotion concepts and their function in a large language model
Interpretability research from Anthropic on emotion concepts
- 103
conputer dipshit (@davidcrespo.bsky.social)
they probably spent 4 hours on this beautiful paragraph
- 104
The Cathedral, the Bazaar, and the Winchester Mystery House
Welcome to the era of sprawling, idiosyncratic tooling.
- 105
Claude Dispatch and the Power of Interfaces
We often lack the tools for the job, even if the AI is capable enough
- 106
A GitHub Issue Title Compromised 4,000 Developer Machines
A prompt injection in a GitHub issue triggered a chain reaction that ended with 4,000 developers getting OpenClaw installed without consent. The attack composes well-understood vulnerabilities into something new: one AI tool bootstrapping another.
- 107
Gergely Orosz (@gergely.pragmaticengineer.com)
This is either brilliant or scary: Anthropic accidentally leaked the TS source code of Claude Code (which is closed source). Repos sharing the source are taken down with DMCA. BUT this repo rewrote the code using Python, and so it violates no copyright & cannot be taken down!
- 108
Jer (@jeremie.com)
agents can be social too 🤝 welcome-m.at - open auth protocol for agents using DPoP, cryptographic signup, signed ToS, no passwords tangled.org/solpbc.org/rookery - open-source PDS designed for agents to use, byw lexicon, cloudflare workers, $0 hosting enroll → plc + handle → publish → atmosphere! https://tangled.org/solpbc.org/rookery
- 109
posts inspector (@pippy.bsky.social)
The computer now has a natural language interface and a replicator for software You think I’m going to do anything other than ensure this technology gets into as many people’s hands as possible? The fight isn’t whether we use this. It’s whether we can self host these models [contains quote post or other embedded content]
- 110
Nicholas Carlini - Black-hat LLMs | [un]prompted 2026
Nicholas Carlini, Research Scientist, Anthropic, speaks at [un]prompted 2026 on: Black-hat LLMs. Large language models are now capable of automating attacks that were previously only possible by human adversaries. In this talk, I discuss several ways that adversaries could mis-use current models in order to cause harm both at a larger scale and at a lower cost than they do currently. For example, we find that recent state-of-the-art models can now find 0-day vulnerabilities in large software projects that have been extensively tested by humans for decades. These new capabilities will alter the threat landscape and require we rethink security in the coming years.
- 111
- 112
Pi: The Minimal Agent Within OpenClaw
A gentle introduction to the Pi coding agent and why I think it’s a glimpse into the future of software.
- 113
Jigsaw-Code/sensemaking-tools
Contribute to Jigsaw-Code/sensemaking-tools development by creating an account on GitHub.
- 114
chenglou/pretext
Contribute to chenglou/pretext development by creating an account on GitHub.
- 115
2023
Or, Why I am Not a Doomer
- 116
Statement from Dario Amodei on our discussions with the Department of War
A statement from our CEO on national security uses of AI
- 117
Dario Amodei — Machines of Loving Grace
How AI Could Transform the World for the Better
- 118
How to write well with AI
Why people who pledge never to write with AI are telling on themselves