Blog

  1. Runtime AI Scientist: Can AI Scientists Coordinate at Runtime?

    We are open-sourcing Runtime AI Scientist, the code behind our paper on Runtime Agent Coordination (RAC). Instead of a fixed workflow, the agent that just finished picks who acts next from the live state of the research. Runtime selection had the highest mean score on all three AI-scientist hosts we tested.

    Systemind Team · 2026-10-03
  2. Inside Muse: A Product and Technical Teardown

    Muse is Meta's bet that people will hand an AI agent real authority over their accounts, money and time. This teardown covers the product decisions that make delegation workable, and the architecture underneath: a VM per user, a single permission authority called Sentinel, credentials the model never sees, and kernel-level taint tracking.

    Yu Chen · 2026-10-02
  3. Introducing AGENTS: A Format for Handing Your Session to Another Agent

    We exported one 63-hour coding session and had other agents answer 100 pre-registered questions about it. Raw transcripts scored 22/30 on the questions that matter. A 1,200-line compiled layer scored 30/30. Here is the format, the experiments, and the two things it does not fix.

    Eason · 2026-08-13
  4. Heuristic Learning on ImageNet: What’s next?

    Forthcoming · Systemind Team