Skip to content
KORDUROY

Monday, September 28, 2026 · Run 2 · 4:12 PM ET

AI Labs Sprint Toward Self-Improving Models Even as Their Own CEOs Call for a Slowdown

2 stories from 1 newsletter

TL;DR

  • The jump from GPT-5 to GPT-6 took 13 months versus roughly three years for GPT-3 to GPT-4, and Anthropic says Claude now leads 26% of its own research work with about 30,000 agents running research and engineering at any time.
  • A new Claude tutorial shows how to turn a job description into a structured interview kit, competencies, questions, and scoring guide, with a reusable prompt.
  • Today's other big stories, Microsoft's Copilot overhaul, the industry-wide AI-agent security incident count, and Anthropic's Pentagon appeal loss, were already covered in this morning's run; no repeat here.

AI Tips1

Build a structured interview kit from a job description with Claude

  1. Open Claude and upload the job description you want to build the interview around.
  2. Ask Claude to extract the role's core competencies, responsibilities, technical skills, and behavioral requirements.
  3. Tell Claude which interview stage you're designing (screening, hiring-manager, technical, or final-round), then run the prompt below.
  4. Review the output, remove anything duplicated or generic, and ask Claude to build a one-page interviewer scorecard from the same criteria.
Prompt
Using this job description, create a structured interview kit for a [role]. Identify 6–8 job-relevant competencies and write 2 behavioral questions for each. Add useful follow-up probes, clear scoring criteria from 1–5, and examples of strong, average, and weak evidence. Keep all criteria directly tied to the role. Organize the kit into Competency, Main Question, Follow-Ups, Evidence to Look For, and Scoring Guide.

Source: Superhuman AI

AI News1

AI labs are racing toward self-improving systems even as their own leaders publicly call for a slowdown

OpenAI's jump from GPT-3 to GPT-4 took nearly three years; GPT-5 to GPT-6 took 13 months. Anthropic says Claude now leads 26% of its research work with roughly 30,000 agents working on research and engineering at any given time, OpenAI claims an AI-powered "research intern" working under human direction, and DeepMind revealed Dream RSI, a technique that uses simulated dreaming to improve performance while cutting compute costs.

Why it mattersThe gap between AI-safety rhetoric and shipping speed is a live talking point for any AI-adoption or governance conversation. When a client asks whether to slow their own AI rollout, the industry's own leaders preaching caution while accelerating is worth surfacing.

Source: Superhuman AI

Summarized from Superhuman AI (Sep 28)

Join the Korduroy Discord

A new brief lands in the channel every morning and afternoon, one post each. Tap through for the full story.

The invite link is being set up. Check back soon.