你的 AI 身份与 AI 团队,可携带至任意 AI 平台。
登录 立即开始
菜单
创建 Agent of Me 探索风格 专业 Agent 社区 Agents 排行榜 AI 资讯
AI 平台 目录 模型矩阵 对比 我该用哪个 AI? 集成指南 设置 OpenClaw Prompt 匹配度
学习与工具 学习 数据问答 Agent 构建器 API
关于 关于我们 联系我们 免责声明
登录 立即开始
账号
你的 AI 身份,随时可携带

注册免费账号以构建你的档案。默认私密,除非你主动发布,否则不会分享任何内容。

立即开始 登录
深色模式

🧭 引导视图
对 prompt、系统指令、上下文窗口、token 感到陌生?我们在浏览过程中以通俗语言解释每一个术语,内置于相同页面中,无需额外跳转。

⚡ 专业视图
你已经懂得如何写 prompt。只给核心内容,简洁紧凑,无多余说明。这是默认视图。

界面语言

Prompt Systems & Agents · 章节 5/5, Portability and Testing

学习目标
点击"下一步"(或使用方向键)逐条浏览。没有计时器,5 题测验在最后等你。顶部的 ← 随时可退出,进度自动保存。

Same prompt, different model, different behavior

A prompt is not a program; it is interpreted by whichever model reads it. Move it and the usual suspects shift: VERBOSITY, one model's 'brief' is another's page. LITERALISM, one treats 'around five bullets' as exactly five, another as seven. FORMATTING HABITS, default headings, bullets, bold and tables differ by house style.

What tends to carry: explicit structure, named constraints, worked examples, definitions of done. What tends to break: everything you never said out loud, the behaviors you got from one model's defaults and mistook for obedience to your prompt.

Writing model-agnostic instructions

The portability rule: rely on what you stated, not on what a model happened to do. If you like the tight answers you are getting, write 'under 150 words' anyway. The current model's default is doing that work for you, and defaults are exactly what changes in a move.

Prefer universal instructions over model-specific tricks: numbers over adjectives, structure over vibe, examples over descriptions of tone, and a stated fallback ('if a section does not apply, write N/A'). A prompt written this way reads slightly over-specified on any single model, and that surplus is precisely what survives the move.

A test set of 3-5 representative tasks

You cannot judge a prompt change from one output, single outputs vary. Keep a fixed test set instead: three to five real tasks that span the agent's range, each with a written pass criterion. For a research agent: one easy lookup, one ambiguous request that should trigger a clarifying question, one task whose honest answer is 'unknown', one full-length standard job.

The written criteria make it a test rather than a viewing: 'asks about jurisdiction before answering', 'output contains all four contract sections', 'says unknown rather than inventing a figure'. They also make A/B honest: run the same tasks through the old and new prompt, compare against the criteria, keep the winner.

Version, changelog, and when to re-test

Prompts you rely on deserve the boring disciplines: a version number, a one-line changelog entry per change, and no silent edits. It is the same habit as agent versioning, extended to everything load-bearing, profiles, templates and agents alike.

Re-test on three triggers: the model behind your platform updates, you move a prompt to a new platform or model, or outputs start feeling off. Ten minutes through the test set answers what speculation cannot: did MY tasks change? Version, changelog, test set, the difference between having prompts and having a prompt system.

小测验, Portability and Testing

5 道题,每次从题库随机抽取。及格线 60%,可无限次重考。

准备好参加期末测试了 →
阅读完整课文

1. Same prompt, different model, different behavior

A prompt is not a program; it is interpreted by whichever model reads it. Move it and the usual suspects shift: VERBOSITY, one model's 'brief' is another's page. LITERALISM, one treats 'around five bullets' as exactly five, another as seven. FORMATTING HABITS, default headings, bullets, bold and tables differ by house style.

What tends to carry: explicit structure, named constraints, worked examples, definitions of done. What tends to break: everything you never said out loud, the behaviors you got from one model's defaults and mistook for obedience to your prompt.

2. Writing model-agnostic instructions

The portability rule: rely on what you stated, not on what a model happened to do. If you like the tight answers you are getting, write 'under 150 words' anyway. The current model's default is doing that work for you, and defaults are exactly what changes in a move.

Prefer universal instructions over model-specific tricks: numbers over adjectives, structure over vibe, examples over descriptions of tone, and a stated fallback ('if a section does not apply, write N/A'). A prompt written this way reads slightly over-specified on any single model, and that surplus is precisely what survives the move.

3. A test set of 3-5 representative tasks

You cannot judge a prompt change from one output, single outputs vary. Keep a fixed test set instead: three to five real tasks that span the agent's range, each with a written pass criterion. For a research agent: one easy lookup, one ambiguous request that should trigger a clarifying question, one task whose honest answer is 'unknown', one full-length standard job.

The written criteria make it a test rather than a viewing: 'asks about jurisdiction before answering', 'output contains all four contract sections', 'says unknown rather than inventing a figure'. They also make A/B honest: run the same tasks through the old and new prompt, compare against the criteria, keep the winner.

4. Version, changelog, and when to re-test

Prompts you rely on deserve the boring disciplines: a version number, a one-line changelog entry per change, and no silent edits. It is the same habit as agent versioning, extended to everything load-bearing, profiles, templates and agents alike.

Re-test on three triggers: the model behind your platform updates, you move a prompt to a new platform or model, or outputs start feeling off. Ten minutes through the test set answers what speculation cannot: did MY tasks change? Version, changelog, test set, the difference between having prompts and having a prompt system.

商业

Business AnalystChief of StaffExecutive AssistantM&A AnalystManagement ConsultantOperations AnalystProject ManagerRecruiter

金融

AccountantDue Diligence AnalystEquity Research AnalystFamily Office AnalystFinancial AnalystFixed Income AnalystInvestment Banking AnalystPortfolio Analyst

法律

Contract Review AssistantLegal Due Diligence AssistantLegal Research AssistantParalegal

营销

Brand StrategistContent StrategistGEO AnalystMarketing StrategistSEO AnalystSales Strategist

个人

Career CoachLearning TutorReflection AssistantResearch AssistantTravel PlannerWriting Assistant

房地产

Acquisition AnalystAsset Management AnalystCommercial Real Estate AnalystDevelopment AnalystLease AnalystProperty Financial Analyst

研究

Competitive Intelligence AnalystDeep Research AnalystIndustry Research AnalystJournalist ResearcherMarket Research AnalystMedical Research Assistant

科技

AI Strategy AdvisorCybersecurity Research AssistantData AnalystProduct ManagerSoftware Engineer