Trending open-source projects · updated daily

AI Prompt Engineering Projects

Prompt engineering stopped being about clever wording and became a loop: draft a prompt, run it against a set of cases, measure the result, change one variable, repeat. This board follows the tooling that loop depends on — prompt compression and pruning, automated prompt optimisation, libraries and templates, tuning methods and evaluation harnesses. It is assembled from GitHub topic pages for prompt engineering, prompt optimisation, prompt compression and prompt tuning, ranked by stars, which is why a two-week-old optimiser can outrank a template collection that has existed for years. Each card shows language, stars and forks and opens a page with description, license, activity dates, README and related prompt projects. If your goal is cutting token cost or making an agent follow instructions reliably, the sections below split the board into optimisation, libraries, tuning and structured-output tooling.

Trending Prompt Engineering Projects

1

f / prompts.chat

f.k.a. Awesome ChatGPT Prompts. Share, discover, and collect prompts from the community. Free and open source — self-host for your organization with complete privacy.

HTML★ 170,803⑂ 21,945
View f/prompts.chat
2

DietrichGebert / ponytail

Makes your AI agent think like the laziest senior dev in the room. The best code is the code you never wrote.

JavaScript★ 142,816⑂ 7,659
View DietrichGebert/ponytail
3

microsoft / generative-ai-for-beginners

21 Lessons, Get Started Building with Generative AI

Jupyter Notebook★ 120,121⑂ 63,239
View microsoft/generative-ai-for-beginners
4

JuliusBrussee / caveman

🪨 why use many token when few token do trick. Viral skill + proxy for coding agents that cuts 65% of tokens by talking like a caveman.

Go★ 106,886⑂ 6,182
View JuliusBrussee/caveman
5

dair-ai / Prompt-Engineering-Guide

🐙 Guides, papers, lessons, notebooks and resources for prompt engineering, context engineering, RAG, and AI Agents.

MDX★ 78,489⑂ 8,630
View dair-ai/Prompt-Engineering-Guide
6

headroomlabs-ai / headroom

Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agents, 60-95% fewer tokens for JSON, same answers. Library, proxy, MCP server.

Python★ 73,186⑂ 5,631
View headroomlabs-ai/headroom
7

asgeirtj / system_prompts_leaks

Extracted system prompts from Anthropic - Claude Fable 5.1, Opus 5, Claude Design, Claude Code. OpenAI - ChatGPT GPT-6-Astra, Codex. Google - Gemini 3.8 Flash, 3.1 Pro, Antigravity.

JavaScript★ 67,787⑂ 11,015
View asgeirtj/system_prompts_leaks
8

blader / humanizer

Agent skill that removes signs of AI-generated writing from text

Python★ 50,467⑂ 4,061
View blader/humanizer
9

elder-plinius / CL4R1T4S

LEAKED SYSTEM PROMPTS FOR CHATGPT, CLAUDE, GEMINI, GROK, PERPLEXITY, CURSOR, LOVABLE, REPLIT, AND MORE! - AI SYSTEMS TRANSPARENCY FOR ALL! 👐

★ 50,068⑂ 10,279
View elder-plinius/CL4R1T4S
10

Imbad0202 / academic-research-skills

Academic Research Skills for Claude Code: research → write → review → revise → finalize

Python★ 48,823⑂ 3,789
View Imbad0202/academic-research-skills
11

github / awesome-copilot

Community-contributed instructions, agents, skills, and configurations to help you make the most of GitHub Copilot.

JavaScript★ 39,184⑂ 4,980
View github/awesome-copilot
12

linshenkx / prompt-optimizer

An AI prompt optimizer for writing better prompts and getting better AI results.

TypeScript★ 35,188⑂ 4,108
View linshenkx/prompt-optimizer
13

langfuse / langfuse

🪢 Open source agent evals & observability: Trace, evaluate, and improve LLM applications with one open platform.

TypeScript★ 34,845⑂ 3,816
View langfuse/langfuse
14

freestylefly / awesome-gpt-image-2

Prompt as Code | GPT Image 2 / 2.5 提示词与案例库,530+ 个案例、20+ 套工业级模板与可复用 Skills,新增 2.5 同提示词对比专区,附完整提示词与生成记录,持续更新。

JavaScript★ 32,925⑂ 3,176
View freestylefly/awesome-gpt-image-2
15

mlflow / mlflow

The open source AI engineering platform for agents, LLMs, and ML models. MLflow enables teams of all sizes to debug, evaluate, monitor

Python★ 28,051⑂ 6,328
View mlflow/mlflow
16

humanlayer / 12-factor-agents

What are the principles we can use to build LLM-powered software that is actually good enough to put in the hands of production customers?

TypeScript★ 26,308⑂ 1,977
View humanlayer/12-factor-agents
17

alirezarezvani / claude-skills

380 Claude Code skills & agent skills & plugins (30+ Agents, 70+ custom commands, 380+ skills, customizable references, scripts)for Claude Code, Codex, Gemini CLI, Cursor

Python★ 26,168⑂ 3,686
View alirezarezvani/claude-skills
18

promptfoo / promptfoo

Test your prompts, agents, and RAGs. Red teaming/pentesting/vulnerability scanning for AI. Compare performance of GPT, Claude, Gemini, DeepSeek, and more.

TypeScript★ 25,303⑂ 2,344
View promptfoo/promptfoo
19

comet-ml / opik

Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards.

Python★ 22,157⑂ 1,809
View comet-ml/opik
20

AI4Finance-Foundation / FinGPT

FinGPT: Open-Source Financial Large Language Models! Revolutionize 🔥 We release the trained model on HuggingFace.

Jupyter Notebook★ 21,271⑂ 3,013
View AI4Finance-Foundation/FinGPT
21

jnMetaCode / agency-agents-zh

🎭 277 个即插即用的 AI 专家角色 — 支持 Claude Code/Cursor/Copilot 等 20 种工具,覆盖工程/设计/营销/金融等 20 个部门。含 64 个中国市场原创智能体(小红书/抖音/微信/飞书/钉钉/Qt 上位机/机械设计)。搭配编排器 agency-orchestrator,一句话即可让多位专家按 DAG 自动协作。

Shell★ 20,816⑂ 3,362
View jnMetaCode/agency-agents-zh
22
23

tradecatlabs / vibe-coding-cn

Vibe Coding 从入门到精通教程|AI 结对编程工作流|Prompt、Skill、Workflow、上下文管理、codex实战指南

Python★ 16,314⑂ 1,653
View tradecatlabs/vibe-coding-cn
24

dottxt-ai / outlines

Structured Outputs

Python★ 15,843⑂ 882
View dottxt-ai/outlines
25

YouMind-OpenLab / awesome-nano-banana-pro-prompts

🍌 World's largest Nano Banana Pro prompt library — 10,000+ curated prompts with preview images, 16 languages. Google Gemini AI image generation. Free & open source.

TypeScript★ 13,451⑂ 1,432
View YouMind-OpenLab/awesome-nano-banana-pro-prompts
26

nidhinjs / prompt-master

A Claude skill that writes the accurate prompts for any AI tool. Zero tokens or credits wasted. Full context and memory retention

★ 13,402⑂ 1,556
View nidhinjs/prompt-master
27

langgptai / LangGPT

LangGPT: Empowering everyone to become a prompt expert! 🚀 📌 结构化提示词(Structured Prompt)提出者 📌 元提示词(Meta-Prompt)发起者 📌 最流行的提示词落地范式 | Language of GPT The pioneering framework for structured & meta-prompt

Jupyter Notebook★ 12,540⑂ 947
View langgptai/LangGPT
28

EmbraceAGI / awesome-chatgpt-zh

ChatGPT 中文指南🔥,ChatGPT 中文调教指南,指令指南,应用开发指南,精选资源清单,更好的使用 chatGPT 让你的生产力 up up up! 🚀

Python★ 11,710⑂ 959
View EmbraceAGI/awesome-chatgpt-zh
29

Arize-ai / phoenix

AI Observability & Evaluation

Python★ 11,548⑂ 1,141
View Arize-ai/phoenix
30

Imbad0202 / academic-research-skills-codex

Codex-native Academic Research Skills suite for human-in-the-loop academic research workflows

Python★ 11,276⑂ 489
View Imbad0202/academic-research-skills-codex
31

cobusgreyling / loop-engineering

Practical patterns, starters & CLI tools for loop engineering with AI coding agents. Design systems that prompt and orchestrate agents (inspired by Addy Osmani and Boris Cherny).

TypeScript★ 11,263⑂ 1,512
View cobusgreyling/loop-engineering
32

microsoft / promptflow

Build high-quality LLM apps - from prototyping, testing to production deployment and monitoring.

Python★ 11,245⑂ 1,123
View microsoft/promptflow
33

LouisShark / chatgpt_system_prompt

A collection of GPT system prompts and various prompt injection/leaking knowledge.

HTML★ 10,771⑂ 1,460
View LouisShark/chatgpt_system_prompt
34

kangarooking / cangjie-skill

把书、长视频、播客等高价值内容蒸馏成可执行的 Agent Skills(Distill high-value content from books, long-form videos, podcasts, and more into executable Agent Skills)

Python★ 10,359⑂ 1,200
View kangarooking/cangjie-skill
35

ZeroLu / awesome-nanobanana-pro

🚀 An awesome list of curated Nano Banana pro prompts and examples. Your go-to resource for mastering prompt engineering and exploring the creative potential of the Nano banana pro(Nano banana 2) AI

★ 10,311⑂ 871
View ZeroLu/awesome-nanobanana-pro
36

YouMind-OpenLab / awesome-gpt-image-2

🚀 World's largest GPT Image 2 prompt library, updated daily — 2000+ curated prompts with preview images, 16 languages.

TypeScript★ 9,919⑂ 880
View YouMind-OpenLab/awesome-gpt-image-2
37

EvoMap / evolver

The GEP-powered self-evolving engine for AI agents. Auditable evolution with Genes, Capsules, and Events. | evomap.ai

JavaScript★ 9,104⑂ 847
View EvoMap/evolver
38

ai-boost / awesome-prompts

Curated list of chatgpt prompts from the top-rated GPTs in the GPTs Store. Prompt Engineering, prompt attack & prompt protect. Advanced Prompt Engineering papers.

★ 8,911⑂ 867
View ai-boost/awesome-prompts
39

rockbenben / ChatGPT-Shortcut

Stop writing prompts from scratch — a searchable prompt library for ChatGPT, Claude, Gemini and Cursor · Русский 한국어 العربية हिन्दी ไทย | 别再从头写提示词:现成的拿来就用,好用的收进自己的库

TypeScript★ 8,771⑂ 957
View rockbenben/ChatGPT-Shortcut
40

jnMetaCode / superpowers-zh

🦸 AI 编程超能力 · 中文增强版 — superpowers(250k+ ⭐)完整汉化 + 4 个中国原创 skills,让 Claude Code / Copilot CLI / Hermes Agent / Cursor / Windsurf / Kiro / Gemini CLI / Qoder 等 26 款 AI 编程工具真正会干活

JavaScript★ 8,159⑂ 763
View jnMetaCode/superpowers-zh
41

jamez-bondos / awesome-gpt4o-images

Awesome curated collection of images and prompts generated by GPT-4o and gpt-image-1. Explore AI generated visuals created with ChatGPT and Sora

JavaScript★ 8,151⑂ 1,804
View jamez-bondos/awesome-gpt4o-images
42

AI4Finance-Foundation / FinRobot

FinRobot: An Open-Source AI Agent Platform for Financial Applications using Large Language Models

Jupyter Notebook★ 8,037⑂ 1,357
View AI4Finance-Foundation/FinRobot
43

NirDiamant / Prompt_Engineering

22 prompt engineering techniques with hands-on Jupyter Notebook tutorials, from fundamental concepts to advanced strategies for leveraging LLMs.

Jupyter Notebook★ 7,859⑂ 1,027
View NirDiamant/Prompt_Engineering
44

mufeedvh / code2prompt

A CLI tool to convert your codebase into a single LLM prompt with source tree, prompt templating, and token counting.

Rust★ 7,679⑂ 447
View mufeedvh/code2prompt
45

Zipstack / unstract

LLM-Driven Extraction of Unstructured Data — Built for API Deployments & ETL Pipeline Workflows

Python★ 7,245⑂ 718
View Zipstack/unstract
46

kyegomez / swarms

The Enterprise-Grade Multi-Agent Orchestration Framework. Website: https://swarms.ai

Python★ 7,188⑂ 1,023
View kyegomez/swarms
47

WenyuChiou / awesome-agentic-ai-zh

A trilingual (繁中 / English / 简中) learning roadmap for agentic AI: from LLM basics to multi-agent systems, with 240+ curated resources and hands-on examples. 中文 AI agent 學習地圖。

Python★ 7,105⑂ 964
View WenyuChiou/awesome-agentic-ai-zh
48

NeoVertex1 / SuperPrompt

SuperPrompt is an attempt to engineer prompts that might help us understand AI agents.

★ 6,433⑂ 573
View NeoVertex1/SuperPrompt
49

promptslab / Awesome-Prompt-Engineering

This repository contains a hand-curated resources for Prompt Engineering with a focus on Generative Pre-trained Transformer (GPT), ChatGPT, PaLM etc

TypeScript★ 6,336⑂ 765
View promptslab/Awesome-Prompt-Engineering
50

langgptai / wonderful-prompts

🔥中文 prompt 精选🔥,ChatGPT 使用指南,提升 ChatGPT 可玩性和可用性!🚀

★ 6,316⑂ 537
View langgptai/wonderful-prompts
51

aimhubio / aim

Aim 💫 — An easy-to-use & supercharged open-source experiment tracker.

Python★ 6,258⑂ 413
View aimhubio/aim
52

swyxio / ai-notes

notes for software engineers getting up to speed on new AI developments. Serves as datastore for https://latent.space writing, and product brainstorming

HTML★ 6,253⑂ 560
View swyxio/ai-notes
53

Helicone / helicone

🧊 Open source LLM observability platform. One line of code to monitor, evaluate, and experiment. YC W23 🍓

TypeScript★ 6,166⑂ 672
View Helicone/helicone
54

ashishps1 / learn-ai-engineering

Learn AI and LLMs from scratch using free resources

★ 6,056⑂ 1,445
View ashishps1/learn-ai-engineering
55

FlorianBruniaux / claude-code-ultimate-guide

The most comprehensive Claude Code guide: agentic workflows, hooks, skills, MCP servers, quizzes, and production-ready templates. 430K+ lines.

Python★ 6,003⑂ 783
View FlorianBruniaux/claude-code-ultimate-guide
56

MadcowD / ell

A language model programming library.

Python★ 5,857⑂ 341
View MadcowD/ell
57

thinkingjimmy / Learning-Prompt

Free prompt engineering online course. ChatGPT and Midjourney tutorials are now included!

CSS★ 5,302⑂ 393
View thinkingjimmy/Learning-Prompt
58

Kiln-AI / Kiln

Build, Evaluate, and Optimize AI Systems. Includes evals, RAG, agents, fine-tuning, synthetic data generation, dataset management, MCP, and more.

Python★ 5,078⑂ 380
View Kiln-AI/Kiln
59

FellouAI / eko

Eko (Eko Keeps Operating) - Build Production-ready Agentic Workflow with Natural Language - eko.fellou.ai

TypeScript★ 4,961⑂ 444
View FellouAI/eko
60

darrenhinde / OpenAgentsControl

AI agent framework for plan-first development workflows with approval-based execution. Multi-language support (TypeScript, Python, Go, Rust) with automatic testing, code review

TypeScript★ 4,862⑂ 401
View darrenhinde/OpenAgentsControl
61

langwatch / langwatch

The platform for LLM evaluations and AI agent testing

TypeScript★ 4,837⑂ 395
View langwatch/langwatch
62

trigaten / Learn_Prompting

Prompt Engineering, Generative AI, and LLM Guide by Learn Prompting | Join our discord for the largest Prompt Engineering learning community

MDX★ 4,735⑂ 666
View trigaten/Learn_Prompting
63

liquidprompt / liquidprompt

A full-featured & carefully designed adaptive prompt for Bash & Zsh

Shell★ 4,678⑂ 419
View liquidprompt/liquidprompt
64

promptslab / Promptify

Prompt Engineering | Prompt Versioning | Use GPT or other prompt based models to get structured output. Join our discord for Prompt-Engineering, LLMs and other latest research

Python★ 4,636⑂ 364
View promptslab/Promptify
65

serenakeyitan / awesome-notebookLM-prompts

A curated collection of the strongest NotebookLM slide prompts sourced from the real creative underground . Your go-to resource for AI powerpoint :P

★ 4,618⑂ 677
View serenakeyitan/awesome-notebookLM-prompts
66

kyegomez / tree-of-thoughts

Plug in and Play Implementation of Tree of Thoughts: Deliberate Problem Solving with Large Language Models that Elevates Model Reasoning by atleast 70%

Python★ 4,588⑂ 376
View kyegomez/tree-of-thoughts
67

conorbronsdon / avoid-ai-writing

Skill that audits and rewrites content to remove AI writing patterns. Use it with your favorite agents including Claude Code, OpenClaw, Codex, and Hermes.

JavaScript★ 4,556⑂ 399
View conorbronsdon/avoid-ai-writing
68

luban-agi / Awesome-AIGC-Tutorials

Curated tutorials and resources for Large Language Models, AI Painting, and more.

★ 4,549⑂ 298
View luban-agi/Awesome-AIGC-Tutorials
69

Jia-Ethan / codex-keysmith

Versioned Codex instruction deployment with preview, ownership manifests, hook isolation, scenario evaluation, and recovery.

Python★ 4,546⑂ 714
View Jia-Ethan/codex-keysmith
70

IBM / mcp-context-forge

An AI Gateway, registry, and proxy that sits in front of any MCP, A2A, or REST/gRPC APIs, exposing a unified endpoint with centralized discovery, guardrails and management.

Python★ 4,504⑂ 877
View IBM/mcp-context-forge
71

0xNyk / council-of-high-intelligence

Structured multi-perspective deliberation for hard decisions. Run full councils, focused triads, or duo debates across Claude Code, Codex, Gemini CLI, and OpenCode.

Shell★ 4,335⑂ 409
View 0xNyk/council-of-high-intelligence
72

algorithmicsuperintelligence / optillm

Optimizing inference proxy for LLMs

Python★ 4,299⑂ 389
View algorithmicsuperintelligence/optillm
73

UditAkhourii / adhd

ADHD — a skill for coding agents. Tree-of-thought with pruning, built on the Claude & Codex Agent SDK.

TypeScript★ 4,246⑂ 292
View UditAkhourii/adhd
74

microsoft / apm

Agent Package Manager

Python★ 3,855⑂ 364
View microsoft/apm
75

urchade / GLiNER

Generalist and Lightweight Model for Named Entity Recognition (Extract any entity types from texts)

Python★ 3,848⑂ 307
View urchade/GLiNER
76

Hunyuan-PromptEnhancer / PromptEnhancer

[CVPR 2026] PromptEnhancer is a prompt-rewriting tool, refining prompts into clearer, structured versions for better image generation.

Python★ 3,773⑂ 326
View Hunyuan-PromptEnhancer/PromptEnhancer
77

zou-group / textgrad

TextGrad: Automatic ''Differentiation'' via Text -- using large language models to backpropagate textual gradients. Published in Nature.

Python★ 3,738⑂ 293
View zou-group/textgrad
78

atfortes / Awesome-LLM-Reasoning

From Chain-of-Thought prompting to OpenAI o1 and DeepSeek-R1 🍓

★ 3,685⑂ 213
View atfortes/Awesome-LLM-Reasoning
79

Open-Less / openless

Hold a key, speak, release — AI-polished text appears at your cursor in any app. Open-source voice input for macOS & Windows. (按住快捷键说话,松开即得润色后的文字)

Rust★ 3,604⑂ 332
View Open-Less/openless
80

filipecalegario / awesome-generative-ai

A curated list of Generative AI tools, works, models, and references

★ 3,540⑂ 885
View filipecalegario/awesome-generative-ai
81

Leonxlnx / unlazy

Anti-laziness skill for AI agents. Core: the Depth Tree method, which splits a task N layers deep and gives every leaf the full time budget of the whole task, so effort multiplies with depth.

JavaScript★ 3,473⑂ 244
View Leonxlnx/unlazy
82

DSXiangLi / DecryptPrompt

总结Prompt&LLM论文,开源数据&模型,AIGC应用

★ 3,440⑂ 318
View DSXiangLi/DecryptPrompt
83

foryourhealth111-pixel / Vibe-Skills

Intelligent Skill routing and workflow orchestration for AI agents — +21.12 pp reward, −29.6% tokens on SkillsBench with DeepSeekV4Flash-VE.

Python★ 3,371⑂ 292
View foryourhealth111-pixel/Vibe-Skills
84

pezzolabs / pezzo

🕹️ Open-source, developer-first LLMOps platform designed to streamline prompt design, version management, instant delivery, collaboration, troubleshooting, observability and more.

TypeScript★ 3,273⑂ 279
View pezzolabs/pezzo
85

protectai / llm-guard

The Security Toolkit for LLM Interactions

Python★ 3,209⑂ 461
View protectai/llm-guard
86

L1Xu4n / Awesome-ChatGPT-prompts-ZH_CN

如何将ChatGPT调教成一只猫娘

★ 3,185⑂ 169
View L1Xu4n/Awesome-ChatGPT-prompts-ZH_CN
87

yanliudesign / mono-color-skill

One-ink editorial print image skill — warm paper, halftone photography, active negative space, and restrained typography.

Python★ 3,175⑂ 85
View yanliudesign/mono-color-skill
88

Forward-Future / loopy

A library of practical AI-agent loops and an installable skill for finding, adapting, and designing repeatable agent workflows.

JavaScript★ 3,146⑂ 278
View Forward-Future/loopy
89

PenglongHuang / chinese-novelist-skill

🎭 AI 写小说:从零生成 10-50 章完整中文小说,三层问答 · 创作记忆 · 悬念钩子 · 自动校验,长篇网文连载皆宜|开源免费,适配主流 coding agent|AI novel writing skill

Python★ 3,124⑂ 455
View PenglongHuang/chinese-novelist-skill
90

wquguru / harness-books

📚 Two books on harness engineering — the design philosophies behind Claude Code & Codex: constraints, query loops, context governance, multi-agent verification. harness-books.agentway.dev

Python★ 3,120⑂ 372
View wquguru/harness-books
91

phodal / prompt-patterns

Prompt 编写模式:如何将思维框架赋予机器,以设计模式的形式来思考 prompt

★ 3,091⑂ 198
View phodal/prompt-patterns
92

KhazP / vibe-coding-prompt-template

Templates and workflow for generating PRDs, Tech Designs, and MVP and more using LLMs for AI IDEs

TypeScript★ 3,088⑂ 381
View KhazP/vibe-coding-prompt-template
93

hegelai / prompttools

Open-source tools for prompt testing and experimentation, with support for both LLMs (e.g. OpenAI, LLaMA) and vector databases (e.g. Chroma, Weaviate, LanceDB).

Python★ 3,055⑂ 256
View hegelai/prompttools
94

ianarawjo / ChainForge

An open-source visual programming environment for battle-testing prompts to LLMs.

TypeScript★ 3,030⑂ 257
View ianarawjo/ChainForge
95

wesammustafa / Claude-Code-Everything-You-Need-to-Know

A practical Claude Code guide with clear mental models and copy-paste examples — setup, prompt engineering, slash commands, skills, hooks, subagents, agent teams, and MCP servers.

Python★ 3,027⑂ 344
View wesammustafa/Claude-Code-Everything-You-Need-to-Know
96

Eladlev / AutoPrompt

A framework for prompt tuning using Intent-based Prompt Calibration

Python★ 3,019⑂ 264
View Eladlev/AutoPrompt
97

sergebulaev / linkedin-skills

Claude skills for LinkedIn. 11 Claude Code and Codex skills that write human-sounding LinkedIn posts, craft comments that get noticed, analyze your feed, and build a publishing cadence

Python★ 2,930⑂ 517
View sergebulaev/linkedin-skills
98

spcl / graph-of-thoughts

Official Implementation of "Graph of Thoughts: Solving Elaborate Problems with Large Language Models"

Python★ 2,841⑂ 216
View spcl/graph-of-thoughts
99

OpenPipe / OpenPipe

Turn expensive prompts into cheap fine-tuned models

TypeScript★ 2,830⑂ 177
View OpenPipe/OpenPipe
100

microsoftarchive / promptbench

A unified evaluation framework for large language models

Python★ 2,821⑂ 222
View microsoftarchive/promptbench

Prompt Optimization & Compression

Prompt Libraries, Templates & Collections

Prompt Tuning & Evaluation

Guardrails, Structured Output & Testing

More AI Trending Projects

1

cloudflare / security-audit-skill

JavaScript★ 17,418⑂ 959▲ 3,155 stars
2

affaan-m / ECC

JavaScript★ 263,267⑂ 39,395▲ 1,012 stars
3

Tencent / BrowserSkill

TypeScript★ 5,916⑂ 417▲ 612 stars
4

vercel-labs / json-render

TypeScript★ 16,987⑂ 905▲ 585 stars
5

addyosmani / agent-skills

JavaScript★ 97,396⑂ 10,274▲ 556 stars
6

anthropics / claude-code

TypeScript★ 146,908⑂ 23,988▲ 483 stars

Iterative prompting: a loop, not a magic sentence

Iterative prompt refinement means treating a prompt as something you test rather than something you write once. The loop is short: write a first version, run it against a fixed set of inputs you care about, look at where it fails, change one thing, run it again. What makes that engineering rather than tinkering is the fixed set — a small evaluation file kept next to the prompt, so every change is measured against the same cases instead of your memory of last week.

Real examples of iteration look banal on purpose: moving an instruction from the middle of a long prompt to the end because the model honours the last constraint more reliably; replacing a vague adjective with an explicit output format; splitting one prompt that does three jobs into two that do one each. The optimiser projects below automate the search, but the loop is the technique — they simply run it faster than you can by hand.

What prompt compression does

Prompt compression is the family of techniques that shorten the input without changing the answer. It splits into two kinds. Token-level pruning deletes text the model demonstrably does not need — stop-words, duplicated instructions, boilerplate from a retrieved document — usually with a small trained model scoring each span. Prompt rewriting restates a long instruction more densely, often by letting a model do the shortening and then verifying that behaviour did not change. Both pay off when the same long prefix is sent on every request: cost scales with input tokens, so a prefix that is 40% smaller is a bill that is roughly 40% smaller, and long prompts also raise latency and lower reliability.

Prompt hygiene: versioning, evaluation, structure

Prompt hygiene is what keeps a working prompt working after three people have edited it. Concretely: keep prompts in version control as files rather than as strings buried in application code; keep a small evaluation set and run it before merging a prompt change, the same way you would run tests before merging code; state the output format explicitly and validate it, so a malformed answer surfaces as an error instead of a parse failure three services downstream. The structured-output and guardrail projects in the last section exist because asking politely for JSON is not a validation strategy.

Prompt engineering questions

What is iterative prompt refinement? Repeatedly revising a prompt and re-measuring it against a fixed set of test inputs, changing one variable per round, until the failure rate on those inputs is acceptable.

Which is an example of iteration in prompt engineering? Adding an explicit output schema after seeing the model return prose, re-running the same evaluation set, and confirming that parse failures dropped — then keeping the version that passed.

What is prompt compression? Reducing the number of tokens a prompt consumes while preserving the model behaviour you rely on, usually by pruning unnecessary spans or rewriting instructions more densely.

What is prompt hygiene? Treating prompts as versioned artefacts: stored in files, reviewed, tested against an evaluation set, with an explicit output contract and validation around them.

Do I need a framework to do this? No. A folder of prompt files, a script that runs them over a JSONL of cases and a diff of the results covers most of it. Frameworks start to pay off once the number of prompts and models grows.

Related AI Trending Lists