daily notes

previous notes →

Signals

We're sharing our new framework for tracking, investigating, and disclosing instances of model misalignment at OpenAI. The framework sets criteria and timelines for public disclosure, including when we haven’t yet fully explained or mitigated the behavior. More complex cases may require longer investigation or coordination with third parties. We’ll prioritize examples that reveal new misalignment mechanisms, meaningful changes in known behavior, or findings that challenge assumptions about safety or mitigation. Alongside the framework, we’re publishing six reports on instances of misaligned behavior we’ve observed during the training or evaluation of our models in the last six months. This is a starting point. We’ll refine the process through experience and public feedback, and share more reports on an ongoing basis. https://t.co/ismCCkeE0L

OpenAI says its framework sets criteria and timelines for disclosing model-misalignment cases, including cases not yet fully explained or mitigated, and says six recent reports accompany the framework. It describes a starting process that will be refined with experience and feedback.

Sentiment

Capability evidence with deployment constraints +0.18

84 source items · 55% editorial confidence

Must read

Yields climb, yet risk appetite holds firmPuriya Abbassi, Giulio Cornelli, Marco Lombardi, Andreas Schrimpf, Vladyslav Sushko, and Karamfil Todorov

The BIS review traces higher long-maturity government yields, sector reallocations around tech valuations, and resilient equity risk appetite.

Signals

The Federal Reserve's final statement records a 12–0 decision at its September 15–16 meeting to raise the target range by 25 basis points to 3.75%–4.00%.

Sentiment

Rates, credit, and infrastructure risk -0.08

509 source items · 55% editorial confidence