instruction following
23 articles · 15 co-occurring · 1 contradictions · 106 briefs
core problem is instruction following still not good enough even at the frontier" — Author explicitly identifies instruction following as an unresolved problem even in frontier models, providing evide
[strong] "Most importantly even though they describe intent, I find intent is often ignored by the models and they simply end up doing more copy paste." — Author documents failure mode where detailed intent specifications are ignored by models in favor of pattern-matching behavior (copy-paste), suggesting structural intent expression is ineffective for controlling agent behavior.
The document is called the "soul document" internally, and the name is more precise than it might first appear. It is an attempt to articulate the animating principle of an artificial mind — not by in
core problem is instruction following still not good enough even at the frontier" — Author explicitly identifies instruction following as an unresolved problem even in frontier models, providing evide
Avoid carrying over every instruction from an older prompt stack. Legacy prompts often over-specify the process because earlier models needed more help staying on track. With GPT-5.5, that can add noi
I like to keep the SKILL md concise and clean so that I can share it with others. It tells Claude Code or Codex: → What role to play → How to give advice → When to save new learnings" — Provides evide
more advanced techniques emerged, such as chain-of-thought prompting, which breaks down complex problems into more logical, step-by-step sequences" — Article documents chain-of-thought prompting as an
I've instructed the model to adapt to the user's input while staying logically consistent with the game's context, but it still struggles" — Article shows the challenge of instruction-following when i
the "Skill" is just a single .md file with instructions in English" — Demonstrates practical implementation of skills as human-readable instruction files that agents can interpret and execute.
Here's a prompt for systematically uncovering this junk and resolving it. Best used with Claude Code and Opus 4.6 on max reasoning in a fresh session" — Demonstrates crafting detailed prompts to guide
与其反复解释动画长什么样,不如直接说出它的名字" — Article illustrates that agents follow precise, domain-specific labels more reliably than verbose natural descriptions—addressing a fundamental instruction clarity problem.
Tone: dry humor, sarcastic. Angle: inverse advice. Specific ideas to hit ('inbox zero'). Keep paragraphs punchy." — Demonstrates structured tone and style specification as part of context engineering
no "I'd be happy to help you with that." no "Let me search the web for you" no more unnecessary filler words" — Reveals counterintuitive design principle: removing politeness/explanation patterns from
[DIRECT] "Now we need SKILL.md files with instructions to use other apps on your computer" — Article identifies structured SKILL.md file format as pattern for defining and extending Claude's capabilit
'Improve the way you prompt the agent'" — Article directly challenges conventional instruction-giving and advocates for improved interaction methodology.
Most importantly even though they describe intent, I find intent is often ignored by the models and they simply end up doing more copy paste." — Author documents failure mode where detailed intent spe
follows instructions more precisely, and verifies its own outputs before reporting back" — Opus 4.7 demonstrates improved instruction adherence combined with self-verification, enabling autonomous tas
Instructions files" — VS Code's instructions files provide a concrete implementation pattern for encoding and managing AI system instructions at the IDE level.
[DIRECT] "when provided with instructions, it is robust to phrasing variations" — Article directly characterizes a key property of general AI systems: instruction robustness across phrasing, adding sp
输入是什么:用户给什么、文件在哪、参数是什么。输出是什么:生成什么文件、返回什么格式、最终要交付什么结果...用'先做 A,再做 B,然后检查 C'" — Article prescribes structured instruction format (clear inputs, outputs, sequential steps) which improves instruction-foll
The distinction between 'pushes back/extends ideas' vs 'executes plans precisely' is really about different models' instruction-following strategies and their willingness to deviate.
[INFERRED] ""So close to coming together... but also better than I expected"" — Evaluates how well model follows specific instructions while attempting creative task with constraints
[INFERRED] "why "perfect prompts" matter less than how you use them" — Introduces a nuance to instruction-following paradigm: execution context and usage patterns matter as much as precision of instru
[INFERRED] "i'd rather get a detailed answer upfront than keep doing follow-up questions" — User preference for comprehensive initial output over iterative clarification emphasizes need for detailed,
[INFERRED] "i feel like i need a macro for `(genuine question, just answer, don't make changes)`" — Post identifies a UX gap in instruction reusability - users need mechanism to save and apply instruc
Get daily briefs + MCP graph access.
Subscribe free →