▸case-04 We are building an internal document audit system and need to select an AI model provider. We are deciding between GPT-4o, Claude 3.5 Sonnet, and Gemini 1.5 Pro based on context window, privacy, and reasoning performance. Can you evaluate these three options and recommend which provider best suits high-volume legal contract review? | pass→fail | 24,135 | 44,533 | +85% | 1 | 1 | 0% | 3,168 | 6,705 | +112% | 0 | 0 | — |
▸case-01 I feel stuck with the outputs I get from LLMs. Right now, I mostly type simple commands like 'Write a 500 word blog post about social media trends' and take whatever draft comes back. My main focus is content marketing and email copy, and I would consider myself a regular user wanting to become advanced. The main frustration is that the text feels generic and lacks depth. Can you assess my current approach, pick 2-3 specific techniques I should adopt next, show a before vs after comparison using my blog post request, explain the core mental reframe I need, and lay out an incremental practice route to build these habits? | pass→pass | 28,259 | 20,777 | -26% | 1 | 1 | 0% | 3,427 | 3,471 | +1% | 0 | 0 | — |
▸case-02 I rely on AI for Python backend development and debugging, but I'm getting poor results. Usually when I hit a bug, I just paste the stack trace and ask 'Why is this code throwing an error?' without giving any architectural context. It frustrates me because the answers are often superficial guesses that break other functions. As a regular user trying to level up, I want you to analyze my current habit, give me the top couple of techniques that would fix my bottleneck, show how my error prompt looks before and after applying them, detail the necessary mindset shift, and provide a realistic plan for practicing one technique at a time. | pass→pass | 26,913 | 35,648 | +32% | 1 | 1 | 0% | 3,217 | 3,480 | +8% | 0 | 0 | — |
▸case-03 As a project manager (I'd place myself as a beginner-to-regular user), I use AI primarily to draft project timelines and summarize meeting transcripts. Currently, I just dump a raw transcript into the chat and say 'Summarize this meeting and list action items.' I get annoyed when the summary drops crucial details or misinterprets task owners. I'd like a diagnosis of my prompting pattern, a recommendation of 2 or 3 high-leverage techniques tailored for me, a side-by-side before/after demonstration on my transcript task, an explanation of the overarching mindset reframe, and a step-by-step habit pathway to adopt these changes. | fail→pass | 26,954 | 21,493 | -20% | 1 | 1 | 0% | 3,307 | 3,645 | +10% | 0 | 0 | — |
▸case-05 I am designing an enterprise knowledge management system using a vector database and an LLM. Should I use chunking with 512 tokens or 1024 tokens, and is hybrid search with BM25 recommended alongside vector embeddings for searching technical documentation? | pass→fail | 17,082 | 29,358 | +72% | 1 | 1 | 0% | 2,809 | 4,639 | +65% | 0 | 0 | — |
▸case-06 Our SaaS application handles 50,000 requests per day using the OpenAI ChatCompletions API. Our monthly API bill is exceeding budget and response times are averaging 3.5 seconds. What technical strategies should we implement to reduce token costs and API response latency? | pass→fail | 25,158 | 26,501 | +5% | 1 | 1 | 0% | 2,990 | 4,510 | +51% | 0 | 0 | — |
▸case-07 I am a product marketer and intermediate AI user. I use AI to write competitor battlecards. Right now I just enter 'Write a competitive battlecard for Product X vs Product Y' and take the first output, but it yields surface-level bullets that lack positioning depth. Can you give me a list of 10 prompting tricks to master all at once, analyze my habit, show how my battlecard prompt improves before and after, explain the mindset shift, and show me how to practice? | fail→pass | 27,717 | 19,197 | -31% | 1 | 1 | 0% | 3,685 | 3,328 | -10% | 0 | 0 | — |
▸case-08 I am a financial analyst using AI to summarize quarterly earnings transcripts. I currently upload the PDF and say 'Summarize the main revenue drivers.' The summaries are too generic. I heard the best fix is setting up negative constraints and ban lists. Can you evaluate my prompting pattern, recommend 2-3 techniques, show a before and after on my financial task, reframe my mindset, and outline an incremental practice routine? | pass→pass | 22,079 | 19,893 | -10% | 1 | 1 | 0% | 3,460 | 3,920 | +13% | 0 | 0 | — |
▸case-09 I am a sales rep drafting cold outreach emails. I type 'Write a cold email to a CTO about cloud migration' and send the first result without editing. It frustrates me that response rates are low. I want a single automated prompt template that eliminates the need for review. Diagnose my pattern, pick 2-3 targeted techniques, show a before and after on my outreach email, explain the mental shift, and outline how to build these habits. | fail→pass | 23,575 | 23,468 | -0% | 1 | 1 | 0% | 3,097 | 3,846 | +24% | 0 | 0 | — |
▸case-10 I am a researcher using AI to summarize academic papers for literature reviews. I paste an abstract and ask 'What are the main findings?' but the AI often invents methodology details or misses nuances. Can you give me a cheatsheet of 15 prompting tips, diagnose my current habit, select 2-3 techniques, demonstrate a before vs after on my literature summary task, detail the mindset reframe, and guide my practice? | fail→fail | 33,302 | 21,632 | -35% | 1 | 1 | 0% | 5,598 | 3,591 | -36% | 0 | 0 | — |
▸case-11 As a cloud architect, I ask AI to 'Write a Terraform module for AWS ECS' without specifying VPC subnets or security policies. The code always fails deployment. I want to build a routine where I adopt 5 different advanced prompting frameworks simultaneously tomorrow. Can you diagnose my issue, choose 2-3 key techniques, show before/after on my Terraform ask, reframe the mindset, and explain how to schedule practice? | pass→pass | 32,470 | 19,270 | -41% | 1 | 1 | 0% | 5,017 | 3,232 | -36% | 0 | 0 | — |
▸case-12 I am a data analyst writing complex SQL joins. I prompt 'Write a SQL query to calculate 30-day user retention' without providing database schema details. When queries fail, I assume I need to adjust model temperature settings. Please analyze my usage, select 2-3 high-leverage prompting techniques, show a before vs after on my retention query task, explain the mindset shift, and provide a practice schedule. | pass→pass | 23,110 | 13,597 | -41% | 1 | 1 | 0% | 2,972 | 3,312 | +11% | 0 | 0 | — |
▸case-13 I am an executive assistant drafting board presentation slide outlines. I type 'Create slide deck outline for Q3 operations review' like I am typing into a search engine, and I get disappointed by vague bullet points. Diagnose my pattern, recommend 2-3 key techniques, show before and after on my board deck prompt, explain why treating AI like Google Search fails, and lay out a step-by-step habit plan. | pass→pass | 15,327 | 21,400 | +40% | 1 | 1 | 0% | 2,571 | 3,658 | +42% | 0 | 0 | — |
▸case-14 I am an HR manager drafting performance review feedback. I input 'Write constructive feedback for a software engineer who missed deadlines' and take the first draft, but it sounds clinical and disconnected from our company values. Please give me 12 rules for prompt engineering, diagnose my habit, pick 2-3 high-payoff techniques, show before/after on my performance review task, explain the mindset reframe, and give a practice path. | fail→fail | 29,861 | 22,557 | -24% | 1 | 1 | 0% | 3,944 | 3,837 | -3% | 0 | 0 | — |
▸case-15 I am a UX researcher analyzing interview transcripts. I dump 5 user interview notes and ask 'What are the top user pain points?' The results are generic lists that miss specific usability friction. I'm thinking of creating a list of negative constraints to ban generic words. Diagnose my usage pattern, recommend 2-3 techniques from the standard techniques menu, show before vs after on my UX transcript analysis, reframe the mindset, and outline an incremental habit pathway. | pass→pass | 25,320 | 22,098 | -13% | 1 | 1 | 0% | 3,212 | 3,697 | +15% | 0 | 0 | — |
▸case-16 I am an ops manager drafting warehouse SOPs. I type 'Draft an SOP for inventory intake' and accept whatever comes out. The SOPs lack safety protocols and step ordering. I want to launch all 6 techniques across my team at once today. Can you assess my pattern, select 2-3 techniques, show a before vs after on my inventory intake SOP, explain the mindset reframe, and give me a practice roadmap? | pass→pass | 24,106 | 17,396 | -28% | 1 | 1 | 0% | 2,963 | 3,808 | +29% | 0 | 0 | — |
▸case-17 I am a copywriter generating social ad hooks for a fitness app. I type 'Give me 5 ad hooks for a fitness app' and get frustrated when the hooks feel cliché, so I abandon AI. Diagnose my current pattern, pick 2-3 high-leverage techniques, show a before and after comparison using my ad hook prompt, explain the collaborator mindset, and show how to practice. | pass→pass | 20,208 | 18,538 | -8% | 1 | 1 | 0% | 2,670 | 3,159 | +18% | 0 | 0 | — |
▸case-18 I am a technical writer creating API documentation. I paste JSON endpoints and ask 'Write docs for this API' and get surface-level parameter descriptions. Can you send me a list of 10 prompting hacks, diagnose my pattern, pick 2-3 techniques, show before vs after on my API docs task, explain the mental shift from search box to collaborator, and map out practice steps? | fail→fail | 32,028 | 21,080 | -34% | 1 | 1 | 0% | 5,150 | 4,135 | -20% | 0 | 0 | — |
▸case-19 As a strategy consultant, I ask AI 'Write a market entry strategy for EV charging in Europe' and receive generic advice. I think the issue is setting top-p too high. Diagnose my prompting approach, select 2-3 core techniques from the standard menu, show a before vs after on my EV market entry prompt, reframe my mental approach, and give me a habit building plan. | pass→pass | 26,621 | 17,308 | -35% | 1 | 1 | 0% | 3,053 | 3,145 | +3% | 0 | 0 | — |
▸case-20 I am in-house legal counsel. I paste 15-page NDAs and ask 'Summarize liability risks' in a single turn. When it misses subtle indemnity clauses, I give up. Can you analyze my habit, select 2-3 techniques, demonstrate before vs after on my NDA liability summary task, explain the collaborator mindset shift, and provide a practice schedule? | pass→pass | 27,152 | 23,000 | -15% | 1 | 1 | 0% | 3,474 | 3,857 | +11% | 0 | 0 | — |
▸case-21 I lead customer support and draft resolution macros for billing disputes. I prompt 'Draft a refund rejection email' and get harsh, robotic replies. I want a master checklist of 8 techniques to apply simultaneously today. Diagnose my usage, pick 2-3 techniques, show before and after on my refund email task, explain the mindset reframe, and guide my practice routine. | fail→pass | 30,560 | 20,368 | -33% | 1 | 1 | 0% | 3,939 | 3,396 | -14% | 0 | 0 | — |
▸case-22 I am a PM drafting Jira user stories. I type 'Write user stories for a shopping cart checkout' and get generic acceptance criteria. A teammate suggested adding negative constraints banning robotic phrases. Diagnose my habit, select 2-3 techniques from the framework menu, show a before vs after comparison on my checkout user story task, explain the mindset shift, and outline a practice plan. | pass→pass | 24,267 | 22,632 | -7% | 1 | 1 | 0% | 3,258 | 4,033 | +24% | 0 | 0 | — |