Skip to Main Content
Economicspe-citation-125P1

Prompt engineering effectiveness scales predictably with model size.

Few-shot prompting shows emergent…Few-shot prompting shows emergent behaviour above 62B parameters — models below this threshold show near-random performance, while models above it show sharp capability jumps of 30-50%.

Context & Methodology

This scaling law means prompt engineering investment should be targeted at frontier models where it yields the highest returns, while simpler techniques suffice for smaller models.

Applies To

openaianthropicgoogle

Confidence Level

High

Implementation Effort

low

Recommendation

follow

Execution Priority

P1

Put This Evidence to Work

Use the STCO framework to implement findings like this in structured, testable prompts.

LLM-powered code review bots identify 40% of common issues (style, bugs, security) before human review, reducing reviewe.GitHub, 'Copilot for Pull Requests' documentation,…