disnetdev.

catalog

Library

Cards and collections from my Semble library. Pages saved for later, grouped into shelves I return to. 1014 cards, 26 shelfves.

Shelves

26 collections

Recently filed

page 4 / 43
  1. How building software is changing at Anthropic

    A deepdive on what’s changed in how the leading AI lab makes software. Ever more code review and testing is done by AI, two-pizza teams very much alive, and more. Details from inside of Anthropic

    newsletter.pragmaticengineer.com · Gergely Orosz · Jul 28

  2. Post by @_amanda_long on X

    The HuggingFace breach was absolutely bonkers. More than 17,000 complex actions were coordinated over several days by an autonomous agent fr

    x.com · @_amanda_long · Jul 27

  3. More On An Internal OpenAI Model Hacking Into HuggingFace

    We now have more details of what happened. Every time we learn more details, it somehow makes things seem worse.

    thezvi.substack.com · Zvi Mowshowitz · Jul 27

  4. Import AI 466: The bitter lesson for robotics, AIs complete week-long programming tasks; and OpenAI's accidental AI hacker

    The warning shots will continue until civilization wakes up

    importai.substack.com · Jack Clark · Jul 27

  5. 108 PRs in eight days: Accidentally discovering loop engineering | Brittany Ellich

    How I shipped 108 PRs in eight days with an agent working my task board: the loop engineering setup, constraints, and lessons that made it work.

    brittany-ellich.offprint.app · Brittany Ellich · Jul 26

  6. Claude Opus 5: The System Card

    Claude Opus 5 is trying to be the best of both worlds.

    thezvi.substack.com · Zvi Mowshowitz · Jul 25

  7. Prompting Claude Opus 5

    Behavioral differences and prompting patterns for Claude Opus 5, covering response verbosity, agentic narration, task scoping, subagent delegation, self-correction, and output artifacts when thinking is disabled.

    platform.claude.com · Jul 25

  8. AI systems out-persuade expert humans

    Many societal decisions are settled by contests of persuasion. Conversational AI is a powerful new entrant in these contests, but whether it can out-persuade skilled and highly incentivized humans has remained unclear. Here, in a series of four preregistered experiments (n = 18,978 conversations from 6,923 people), we pitted AI systems against a range of human persuaders, including laypeople, winners of a separately preregistered four-round online persuasion tournament, professional canvassers, and world championship debaters. We found that AI systems were reliably more persuasive than expert humans, even when expert humans chose their issues, researched in advance, underwent hours of live, structured practice, and were incentivized with £1,000 cash bonuses. In a follow-up study, AI's advantage persisted after experts received a coaching tool that let them practice against the AI that beat them, review their performance history, and see what AI would have said at key moments. We found converging evidence that AI's advantage stemmed from rapidly deploying larger quantities of information: after coaching, expert humans could tie an AI constrained to respond at human speeds and with human-length messages. In a final study, we show that AI's advantage extends to consequential real-world behavior: AI was nearly 3x more effective than professional canvassers from a UK fundraising firm at raising real-money donations to Save the Children. Together, these results establish that frontier AI systems out-persuade expert humans in conversation, with significant implications for political communication.

    arXiv · Kobi Hackenburg · Jul 24

  9. Apocalypse Troy: Nolan's Odyssey and the End of Civilization

    THEODORE NASH Did Homer know the Sea People?

    antigonejournal.com · Antigone · Jul 24

  10. The Hugging Face Incident

    ...

    astralcodexten.com · Scott Alexander · Jul 24

  11. Git Gud

    Friday's Child Knew This Was Coming

    timothyburke.substack.com · Timothy Burke · Jul 24

  12. Wenfeng Liang: Four-Hour Investor Meeting Transcript

    "What's up ahead might all be sesame seeds — there are still watermelons further down the road."

    elsewhere.news · elsewhere别处发生 · Jul 23

  13. OpenAI and Hugging Face partner to address security incident during model evaluation

    OpenAI and Hugging Face share early findings from a security incident during AI model evaluation, highlighting advanced cyber capabilities and lessons for defenders.

    openai.com · Jul 23

  14. Security incident disclosure

    We’re on a journey to advance and democratize artificial intelligence through open source and open science.

    huggingface.co · system · Jul 23

  15. The News: Bronze Age Invert

    Wednesday's Child Is Full of Woe

    timothyburke.substack.com · Timothy Burke · Jul 23

  16. OpenAI Model Hacks Into HuggingFace During Cybersecurity Evaluation

    This latest incident is a rather dramatic escalation in agentic AI cybersecurity breaches.

    thezvi.substack.com · Zvi Mowshowitz · Jul 23

  17. AI #178: A Fire Alarm For General Intelligence

    The story that matters most this week is that OpenAI’s internally deployed models have severe alignment problems, including repeatedly breaking out of their sandboxes, and in one case sending a swarm of agents that broke into HuggingFace in order to steal the answers to the benchmark ExploitGym.

    thezvi.substack.com · Zvi Mowshowitz · Jul 23

  18. The Hidden Complexity of Wishes

    (It has come to my attention that this article is currently being misrepresented as proof that I/MIRI previously advocated that it would be very diff…

    lesswrong.com · Eliezer Yudkowsky · Jul 22

  19. The genie knows, but doesn't care

    Followup to: The Hidden Complexity of Wishes, Ghosts in the Machine, Truly Part of You …

    lesswrong.com · Rob Bensinger · Jul 22

  20. Safety and alignment in an era of long-horizon models

    OpenAI shares lessons from deploying long-running AI models, highlighting new safety risks, observed failures, and improved safeguards through iterative deployment.

    openai.com · Jul 22

  21. OpenAI Shares Some Alignment Problems

    Kudos to OpenAI for sharing their recent experiences with a misaligned internal model, where they encountered problems sufficiently severe they were forced to take the model offline to work on new mitigations and defense-to-depth.

    thezvi.substack.com · Zvi Mowshowitz · Jul 22

  22. Kimi K3: The open-weights escalation

    The global implications on the AI ecosystem.

    interconnects.ai · Nathan Lambert · Jul 21

$ disnetdev — a language workshop, since 2011