Week of 2026-09-07 roundup
Muse Spark 1.3 is live and independently competitive, but Meta's comparison table uses max reasoning—the mode it has not yet released.
OpenAI changed two conversation settings, not the model. The result shows why long-running AI tests depend on their memory setup.
Polyglot or poly-not?
AI coding benchmarks are heavily skewed toward Python and JavaScript. Framework maintainers could change that by defining what good code looks like in their ecosystems.