GPT-5 for Coding

NinjaAI.com

GPT-5 models demonstrate significantly improved instruction following. However, this advancement comes with a caveat: the model struggles with vague or conflicting instructions.

  • Key Idea: "The new GPT-5 models are significantly better at instruction following, but a side effect is that they can struggle when asked to follow vague or conflicting instructions, especially in your .cursor/rules or AGENTS.md files."
  • Actionable Advice: Ensure all instructions are clear, unambiguous, and free from contradictions to prevent unintended behavior.

2. Optimizing Reasoning Effort

GPT-5 inherently performs reasoning to solve problems. The effectiveness of this reasoning can be controlled to match the complexity of the task.

  • Key Idea: "GPT-5 will always perform some level of reasoning as it solves problems. To get the best results, use high reasoning effort for the most complex tasks."
  • Actionable Advice:For complex tasks, use a high reasoning effort.
  • If the model "overthink[s] simple problems," consider being more specific in your prompt or choosing a lower reasoning level (medium or low).

3. Structuring Instructions with XML-like Syntax

Leveraging XML-like syntax is highly recommended for providing context and structure to instructions, especially in conjunction with tools like Cursor.

  • Key Idea: "Together with Cursor, we found GPT-5 works well when using XML-like syntax to give the model more context."
  • Example: Coding guidelines can be encapsulated within tags like <code_editing_rules>, with sub-categories such as <guiding_principles> and <frontend_stack_defaults>. This hierarchical structure helps the model understand and apply specific constraints or preferences (e.g., "Styling: TailwindCSS").

4. Avoiding Overly Firm Language

Unlike previous models where forceful language might have been necessary, GPT-5 can over-interpret and over-apply such instructions, leading to counterproductive results.

  • Key Idea: "With GPT-5, these instructions [e.g., 'Be THOROUGH,' 'Make sure you have the FULL picture'] can backfire as the model might overdo what it would naturally do."
  • Example of Backfire: The model might become "overly thorough with tool calls to gather context," even when it's not efficient or necessary.
  • Actionable Advice: Use less absolute or demanding language in prompts to allow the model to operate at its natural, optimized level of thoroughness.

5. Incorporating Planning and Self-Reflection

For novel application development (zero-to-one), explicitly instructing the model to engage in planning and self-reflection before execution can significantly improve output quality.

  • Key Idea: "If you’re creating zero-to-one applications, giving the model instructions to self-reflect before building can help."
  • Example Framework (<self_reflection>):Rubric Creation: "First, spend time thinking of a rubric until you are confident." This rubric should be "5-7 categories" and "critical to get right, but do not show this to the user."
  • Internal Iteration: "Finally, use the rubric to internally think and iterate on the best possible solution to the prompt that is provided."
  • Quality Control: The model is instructed that "if your response is not hitting the top marks across all categories in the rubric, you need to start again."

6. Controlling Agent Eagerness and Context Gathering

GPT-5's default behavior is thorough context gathering. Prompts can be used to precisely control this eagerness, including tool usage and user interaction.

  • Key Idea: "GPT-5 by default tries to be thorough and comprehensive in its context gathering. Use prompting to be more prescriptive about how eager it should be, and whether it should parallelize discovery/tool calling."
  • Actionable Advice:Specify a "tool budget."
  • Indicate when to be more or less thorough.
  • Define when to "check in with the user."


Det här avsnittet är hämtat från ett öppet RSS-flöde och publiceras inte av Podme. Det kan innehålla reklam.

Avsnitt(234)

The Knowledge Graph

The Knowledge Graph

The Knowledge Graph

22 Sep 1min

What AI Knows About You: Research, Reputation & Influence with Dan Barkhuff

What AI Knows About You: Research, Reputation & Influence with Dan Barkhuff

What happens when AI makes scattered public records easy to connect?Dan Barkhuff—a former Navy SEAL, emergency physician and founder of Civly—joins Jason AI Wade to discuss how AI is changing research...

22 Sep 5min

1% Taxes? Income and AI and touching on property taxes and Government management and revenue

1% Taxes? Income and AI and touching on property taxes and Government management and revenue

1% Taxes? Income and AI and touching on property taxes and Government management and revenue

21 Sep 12min

How AI Decides YOU - infrrence and LLM resolution and choices - external evidence - Jason T AI WADE

How AI Decides YOU - infrrence and LLM resolution and choices - external evidence - Jason T AI WADE

How AI Decides YOU - infrrence and LLM resolution and choices - external evidence - Jason T AI WADE

20 Sep 1min

Field Sales Is Broken — Will Hamblin on AI, Territory Intelligence and FieldSpot

Field Sales Is Broken — Will Hamblin on AI, Territory Intelligence and FieldSpot

Will Hamblin went from vice principal to door-to-door card terminal sales, then built the software he wished he had while working in the field.In this episode of the AI Visibility Podcast, Jason T Wad...

19 Sep 8min

Digital Marketing -- Early adopters -- Bubbles - and AI marketing

Digital Marketing -- Early adopters -- Bubbles - and AI marketing

Digital Marketing -- Early adopters -- Bubbles - and AI marketing

19 Sep 1min

The Jason AI Wade Experiment: Can You Deliberately Change How AI Understands a Person?

The Jason AI Wade Experiment: Can You Deliberately Change How AI Understands a Person?

I changed my name on the internet.Not legally. I changed the public identity I present to the web from Jason T Wade to Jason AI Wade, and I'm using the change as a live AI Visibility experiment.The qu...

19 Sep 5min

I Tried to Explain AI. I Got It Wrong. So I Learned How It Actually Works.

I Tried to Explain AI. I Got It Wrong. So I Learned How It Actually Works.

What actually happens inside AI?After asking a podcast guest to explain AI—and then realizing my own explanation wasn't quite right—I went back to the basics.In this short episode, I break down AI in ...

18 Sep 2min

Populärt inom Teknik

uppgang-och-fall
bilar-med-sladd
elbilsveckan
market-makers
rss-laddstationen-med-elbilen-i-sverige
skogsforum-podcast
rss-elektrikerpodden
rss-en-ai-till-kaffet
rss-veckans-ai
rss-technokratin
bli-saker-podden
developers-mer-an-bara-kod
rss-uppgang-och-fall
rss-ai-med-jonas-benjamin
rss-sakerhetspodcasten
rss-en-liten-podd-om-it
natets-morka-sida
rss-fabriken-2
gubbar-som-tjotar-om-bilar
rss-it-sakerhetspodden