AI identity dan AI team-mu, portabel ke setiap platform AI.
Masuk Mulai
Menu
Buat Agent of Me Jelajahi Gaya Professional Agents Community Agents Leaderboard Berita AI
Platform AI Direktori Model Matrix Bandingkan AI mana yang sebaiknya aku gunakan? Panduan Integrasi Siapkan OpenClaw Prompt Fit
Pelajari & Alat Pelajari Tanya Data Agent Builder API
Tentang Tentang kami Kontak Disclaimers
Masuk Mulai
Akun
Identitas AI Anda, portabel

Buat akun gratis untuk membangun profil Anda. Privat secara default. Tidak ada yang dibagikan kecuali Anda mempublikasikannya.

Mulai Masuk
Mode Gelap

🧭 Tampilan Terpandu
Belum familiar dengan prompt, system instruction, context window, token? Kami menjelaskan setiap istilah saat kamu menjelajah, dalam bahasa yang mudah dipahami. Halaman yang sama, dengan panduan terintegrasi.

⚡ Tampilan Ahli
Kamu sudah paham cara prompting bekerja. Cukup isinya, ringkas dan to the point, tanpa penjelasan tambahan. Ini adalah tampilan default.

Bahasa antarmuka

Prompt Systems & Agents · Bagian 5/5, Portability and Testing

Tujuan pembelajaran
Ketuk Berikutnya (atau gunakan tombol panah) untuk berpindah satu ide sekaligus. Tidak ada timer, kuis 5 pertanyaan menunggu di akhir. Tombol ← di atas untuk keluar kapan saja; progres tersimpan.

Same prompt, different model, different behavior

A prompt is not a program; it is interpreted by whichever model reads it. Move it and the usual suspects shift: VERBOSITY, one model's 'brief' is another's page. LITERALISM, one treats 'around five bullets' as exactly five, another as seven. FORMATTING HABITS, default headings, bullets, bold and tables differ by house style.

What tends to carry: explicit structure, named constraints, worked examples, definitions of done. What tends to break: everything you never said out loud, the behaviors you got from one model's defaults and mistook for obedience to your prompt.

Writing model-agnostic instructions

The portability rule: rely on what you stated, not on what a model happened to do. If you like the tight answers you are getting, write 'under 150 words' anyway. The current model's default is doing that work for you, and defaults are exactly what changes in a move.

Prefer universal instructions over model-specific tricks: numbers over adjectives, structure over vibe, examples over descriptions of tone, and a stated fallback ('if a section does not apply, write N/A'). A prompt written this way reads slightly over-specified on any single model, and that surplus is precisely what survives the move.

A test set of 3-5 representative tasks

You cannot judge a prompt change from one output, single outputs vary. Keep a fixed test set instead: three to five real tasks that span the agent's range, each with a written pass criterion. For a research agent: one easy lookup, one ambiguous request that should trigger a clarifying question, one task whose honest answer is 'unknown', one full-length standard job.

The written criteria make it a test rather than a viewing: 'asks about jurisdiction before answering', 'output contains all four contract sections', 'says unknown rather than inventing a figure'. They also make A/B honest: run the same tasks through the old and new prompt, compare against the criteria, keep the winner.

Version, changelog, and when to re-test

Prompts you rely on deserve the boring disciplines: a version number, a one-line changelog entry per change, and no silent edits. It is the same habit as agent versioning, extended to everything load-bearing, profiles, templates and agents alike.

Re-test on three triggers: the model behind your platform updates, you move a prompt to a new platform or model, or outputs start feeling off. Ten minutes through the test set answers what speculation cannot: did MY tasks change? Version, changelog, test set, the difference between having prompts and having a prompt system.

Mini kuis, Portability and Testing

5 pertanyaan, diambil segar dari bank setiap percobaan. Nilai lulus 60%. Pengulangan tak terbatas.

Siap untuk tes akhir →
Baca teks pelajaran lengkap

1. Same prompt, different model, different behavior

A prompt is not a program; it is interpreted by whichever model reads it. Move it and the usual suspects shift: VERBOSITY, one model's 'brief' is another's page. LITERALISM, one treats 'around five bullets' as exactly five, another as seven. FORMATTING HABITS, default headings, bullets, bold and tables differ by house style.

What tends to carry: explicit structure, named constraints, worked examples, definitions of done. What tends to break: everything you never said out loud, the behaviors you got from one model's defaults and mistook for obedience to your prompt.

2. Writing model-agnostic instructions

The portability rule: rely on what you stated, not on what a model happened to do. If you like the tight answers you are getting, write 'under 150 words' anyway. The current model's default is doing that work for you, and defaults are exactly what changes in a move.

Prefer universal instructions over model-specific tricks: numbers over adjectives, structure over vibe, examples over descriptions of tone, and a stated fallback ('if a section does not apply, write N/A'). A prompt written this way reads slightly over-specified on any single model, and that surplus is precisely what survives the move.

3. A test set of 3-5 representative tasks

You cannot judge a prompt change from one output, single outputs vary. Keep a fixed test set instead: three to five real tasks that span the agent's range, each with a written pass criterion. For a research agent: one easy lookup, one ambiguous request that should trigger a clarifying question, one task whose honest answer is 'unknown', one full-length standard job.

The written criteria make it a test rather than a viewing: 'asks about jurisdiction before answering', 'output contains all four contract sections', 'says unknown rather than inventing a figure'. They also make A/B honest: run the same tasks through the old and new prompt, compare against the criteria, keep the winner.

4. Version, changelog, and when to re-test

Prompts you rely on deserve the boring disciplines: a version number, a one-line changelog entry per change, and no silent edits. It is the same habit as agent versioning, extended to everything load-bearing, profiles, templates and agents alike.

Re-test on three triggers: the model behind your platform updates, you move a prompt to a new platform or model, or outputs start feeling off. Ten minutes through the test set answers what speculation cannot: did MY tasks change? Version, changelog, test set, the difference between having prompts and having a prompt system.

Bisnis

Business AnalystChief of StaffExecutive AssistantM&A AnalystManagement ConsultantOperations AnalystProject ManagerRecruiter

Keuangan

AccountantDue Diligence AnalystEquity Research AnalystFamily Office AnalystFinancial AnalystFixed Income AnalystInvestment Banking AnalystPortfolio Analyst

Hukum

Contract Review AssistantLegal Due Diligence AssistantLegal Research AssistantParalegal

Pemasaran

Brand StrategistContent StrategistGEO AnalystMarketing StrategistSEO AnalystSales Strategist

Personal

Career CoachLearning TutorReflection AssistantResearch AssistantTravel PlannerWriting Assistant

Properti

Acquisition AnalystAsset Management AnalystCommercial Real Estate AnalystDevelopment AnalystLease AnalystProperty Financial Analyst

Riset

Competitive Intelligence AnalystDeep Research AnalystIndustry Research AnalystJournalist ResearcherMarket Research AnalystMedical Research Assistant

Teknologi

AI Strategy AdvisorCybersecurity Research AssistantData AnalystProduct ManagerSoftware Engineer