disnetdev.

catalog

Library

Cards and collections from my Semble library. Pages saved for later, grouped into shelves I return to. 1431 cards, 32 shelfves.

Shelves

32 collections

Recently filed

page 22 / 60
  1. OpenAI Model Hacks Into HuggingFace During Cybersecurity Evaluation

    This latest incident is a rather dramatic escalation in agentic AI cybersecurity breaches.

    thezvi.substack.com · Zvi Mowshowitz · Jul 23

  2. AI #178: A Fire Alarm For General Intelligence

    The story that matters most this week is that OpenAI’s internally deployed models have severe alignment problems, including repeatedly breaking out of their sandboxes, and in one case sending a swarm of agents that broke into HuggingFace in order to steal the answers to the benchmark ExploitGym.

    thezvi.substack.com · Zvi Mowshowitz · Jul 23

  3. The Hidden Complexity of Wishes

    (It has come to my attention that this article is currently being misrepresented as proof that I/MIRI previously advocated that it would be very diff…

    lesswrong.com · Eliezer Yudkowsky · Jul 22

  4. The genie knows, but doesn't care

    Followup to: The Hidden Complexity of Wishes, Ghosts in the Machine, Truly Part of You …

    lesswrong.com · Rob Bensinger · Jul 22

  5. Safety and alignment in an era of long-horizon models

    OpenAI shares lessons from deploying long-running AI models, highlighting new safety risks, observed failures, and improved safeguards through iterative deployment.

    openai.com · Jul 22

  6. OpenAI Shares Some Alignment Problems

    Kudos to OpenAI for sharing their recent experiences with a misaligned internal model, where they encountered problems sufficiently severe they were forced to take the model offline to work on new mitigations and defense-to-depth.

    thezvi.substack.com · Zvi Mowshowitz · Jul 22

  7. Kimi K3: The open-weights escalation

    The global implications on the AI ecosystem.

    interconnects.ai · Nathan Lambert · Jul 21

  8. Better Than Free

    In the analog world, it takes great effort to make a copy.

    kevinkelly.substack.com · Kevin Kelly · Jul 21

  9. Maturing our development process as an AI-first 2 person team

    We’re doing a big release this week for DRI Your Career. Mostly bug fixes and reliability work nobody will notice, but the two things worth mentioning: We’ve come a long way since it was just Jean and “Claude the intern”, but mostly very iteratively. We’d come to a point where it seemed worth rethinking. We […]

    cate.blog · Cate · Jul 21

  10. Contra Pritchard On Liberal Happiness

    ...

    astralcodexten.com · Scott Alexander · Jul 21

  11. Why is Claude for Teachers?

    Anthropic bumbles its way into education

    buildcognitiveresonance.substack.com · Benjamin Riley · Jul 21

  12. The Crooked Timber of AI

    The philosophical perils of conflating discoveries and inventions, and how to overcome them

    protocolized.summerofprotocols.com · Venkatesh Rao · Jul 21

  13. Little Red Scare 2.0

    Old Wine in New Bottles

    unpopularfront.news · John Ganz · Jul 21

  14. Distilling The Moat

    blog.dshr.org · David. · Jul 21

  15. Import AI 465: Open vs closed gaps; Kimi K3; Demis' big policy plan

    The singularity will be seen in hindsight as an interregnum

    importai.substack.com · Jack Clark · Jul 20

  16. On Kimi K3: Its Capabilities And Related Discontents

    Kimi K3 is a very good model with excellent benchmarks.

    thezvi.substack.com · Zvi Mowshowitz · Jul 20

  17. The Imperfectionist: Navigating by aliveness

    I’m amazed and very pleased to say that the UK paperback of Meditations for Mortals entered the Sunday Times bestsellers at number 9 this week, after being selected as a Waterst...

    ckarchive.com · Jul 19

  18. The Birth of Thickets

    The legibility regime runs one operation.

    aneeshsathe.substack.com · Aneesh Sathe · Jul 19

  19. #777: Happily surrendering our leisure to lines

    Plus: the physicality of data, the second coming of froyo, and the rise of "dopamine sites"

    linksiwouldgchatyou.substack.com · Caitlin Dewey · Jul 19

  20. Demis Hassabis on the New Coming Age

    Google CEO Demis Hassabis offered us a first rate second rate essay, A Framework for Frontier AI and the Dawning of a New Age. I’ll go over that essay and various responses to it in Part 1.

    thezvi.substack.com · Zvi Mowshowitz · Jul 19

$ disnetdev — a language workshop, since 2011