Thinking
OpenAI’s “automated research intern” is an internal measurement claim, not a portable productivity benchmark
OpenAI’s research-intern milestone separates agent runtime, spending, activity, and research progress into distinct measurement layers.
Super Genius Labs Editorial · Sep 8, 2026 · 4 min readEngineeringAn admission gate for third-party agent skills
Third-party agent skills can combine runtime instructions with executable scripts. A bounded admission process can verify provenance, content, permissions, isolation, updates, and revocation before granting access.
Super Genius Labs Editorial · Aug 14, 2026 · 5 min readEngineeringTest the effective policy of every coding-agent surface
Turn shared coding-agent settings into surface-specific acceptance tests for policy source, precedence, refresh behavior, denial outcomes, and retained evidence.
Super Genius Labs Editorial · Aug 10, 2026 · 5 min read