63 of them, word for word from real job posts, each with what that client has spent on freelancers. Ask yours below. 56 of the answers open something you can run right here.
$1,827,106 spent on freelancers, and then asked this Describe an AI/ML application you personally built. What did you design and code yourself? read the answer → $1,144,267 spent on freelancers, and then asked this Walk me through a real search or recommendation system you built, not a CRUD app. What was the ranking logic, and what broke first at scale? read the answer → $1,144,267 spent on freelancers, and then asked this If an ecommerce store's revenue suddenly dropped 30% while ad spend stayed the same, how would you diagnose the cause before making changes? read the answer → $1,144,267 spent on freelancers, and then asked this Our catalog has one physical part mapped to 15+ vehicles via a fitment table. Design, out loud, a filter system that stays fast as that table grows… read the answer → $1,144,267 spent on freelancers, and then asked this Share 3 to 5 specific rules from a CLAUDE.md you actively use, with the reasoning behind each. What project-specific issue or Claude behaviour… read the answer → $1,144,267 spent on freelancers, and then asked this Claude Code proposes a 12-file change to fix a slow filter query. Before you approve it, what do you personally check, and what would make you reject… read the answer → $1,144,267 spent on freelancers, and then asked this You personally have Claude Max right now. What tier, and what's the last thing Claude Code did that surprised you this week, good or bad? read the answer → $1,144,267 spent on freelancers, and then asked this Do you use Claude Code daily? You will be told in an interview to share your screen, and the interviewer is a Claude Code specialist. read the answer → $1,115,131 spent on freelancers, and then asked this Describe a real-time voice agent you've personally built or tuned. What was your end-to-end latency, and what got it there? read the answer → $1,115,131 spent on freelancers, and then asked this How would you approach barge-in, the user starting to talk while the bot is speaking, in a browser where echo cancellation can't see the TTS audio? read the answer → $1,115,131 spent on freelancers, and then asked this prove them with before/after measurements: latency numbers, and recorded test conversations. read the answer → $1,101,306 spent on freelancers, and then asked this Describe one production n8n workflow you built, including integrations and how you validated its reliability. read the answer → $988,860 spent on freelancers, and then asked this Experience building MCP servers and tools, a concrete, checkable skill, not just 'uses AI tools' read the answer → $935,596 spent on freelancers, and then asked this Tell us about 2 to 3 AI or automation solutions you've personally built. For each: the problem, the stack, what you built, and the measurable result. read the answer → $935,596 spent on freelancers, and then asked this a portfolio of real work you can walk us through read the answer → $665,013 spent on freelancers, and then asked this Name one production system where 5+ coding agents opened real PRs against one large codebase with Linear as the source of intent. What did agents… read the answer → $665,013 spent on freelancers, and then asked this Delegation silently did nothing while the webhook endpoint looked healthy. What's your first suspect, how do you confirm it inside an hour, and what… read the answer → $665,013 spent on freelancers, and then asked this Linear redelivers webhooks and agents retry. A duplicate session event double-dispatched a task and produced two PRs. Where does idempotency live,… read the answer → $665,013 spent on freelancers, and then asked this A conductor splits an epic into ~20 dependency-ordered work packages built by parallel agents. Model it in Linear. What breaks at 50 agent-driven… read the answer → $609,912 spent on freelancers, and then asked this Contributions to anthropic-sdk-python or claude-agent-sdk-python/typescript, or filed reproducible bug reports against them. read the answer → $609,912 spent on freelancers, and then asked this any open-source contributions or repos we can review read the answer → $494,665 spent on freelancers, and then asked this Describe the most complex voice AI agent you have shipped to production. What was the workflow, how did you handle interruptions and confirmations,… read the answer → $494,665 spent on freelancers, and then asked this Describe a voice interaction that failed in production: noise, an accent, silence, or an action that did not commit, and how you handled it. read the answer → $483,589 spent on freelancers, and then asked this Describe an MCP server, AI agent integration, or LLM tool-calling system you personally built. What APIs did it expose, how did you enforce what the… read the answer → $483,589 spent on freelancers, and then asked this Describe a production API you built that handled untrusted or automated callers. How did you handle authentication, rate limiting, input validation… read the answer → $483,589 spent on freelancers, and then asked this Published or open-source MCP servers. Include links in your application. read the answer → $376,190 spent on freelancers, and then asked this three examples of AI systems you personally built and a short video walking us through your work read the answer → $373,434 spent on freelancers, and then asked this A short 5 to 10 line story of a bug you fixed using Claude Code, including one thing the agent got wrong and how you caught it. read the answer → $373,434 spent on freelancers, and then asked this You use Claude Code daily to ship real code, not just chat. Explain your setup: CLAUDE.md, subagents, plan mode, hooks, how you review what the agent… read the answer → $373,434 spent on freelancers, and then asked this Your pre-deploy checklist for a production release, as a bullet list. read the answer → $339,494 spent on freelancers, and then asked this How has AI changed the way you do data work in the last year? Be specific: which tools, what you've built, and what you stopped doing by hand. read the answer → $325,374 spent on freelancers, and then asked this Have you fine-tuned a coding LLM before? Send benchmark/results. read the answer → $1,827,106 spent on freelancers, and then asked this Describe an AI/ML application you personally built. What did you design and code yourself? read the answer → $1,144,267 spent on freelancers, and then asked this Walk me through a real search or recommendation system you built, not a CRUD app. What was the ranking logic, and what broke first at scale? read the answer → $1,144,267 spent on freelancers, and then asked this If an ecommerce store's revenue suddenly dropped 30% while ad spend stayed the same, how would you diagnose the cause before making changes? read the answer → $1,144,267 spent on freelancers, and then asked this Our catalog has one physical part mapped to 15+ vehicles via a fitment table. Design, out loud, a filter system that stays fast as that table grows… read the answer → $1,144,267 spent on freelancers, and then asked this Share 3 to 5 specific rules from a CLAUDE.md you actively use, with the reasoning behind each. What project-specific issue or Claude behaviour… read the answer → $1,144,267 spent on freelancers, and then asked this Claude Code proposes a 12-file change to fix a slow filter query. Before you approve it, what do you personally check, and what would make you reject… read the answer → $1,144,267 spent on freelancers, and then asked this You personally have Claude Max right now. What tier, and what's the last thing Claude Code did that surprised you this week, good or bad? read the answer → $1,144,267 spent on freelancers, and then asked this Do you use Claude Code daily? You will be told in an interview to share your screen, and the interviewer is a Claude Code specialist. read the answer → $1,115,131 spent on freelancers, and then asked this Describe a real-time voice agent you've personally built or tuned. What was your end-to-end latency, and what got it there? read the answer → $1,115,131 spent on freelancers, and then asked this How would you approach barge-in, the user starting to talk while the bot is speaking, in a browser where echo cancellation can't see the TTS audio? read the answer → $1,115,131 spent on freelancers, and then asked this prove them with before/after measurements: latency numbers, and recorded test conversations. read the answer → $1,101,306 spent on freelancers, and then asked this Describe one production n8n workflow you built, including integrations and how you validated its reliability. read the answer → $988,860 spent on freelancers, and then asked this Experience building MCP servers and tools, a concrete, checkable skill, not just 'uses AI tools' read the answer → $935,596 spent on freelancers, and then asked this Tell us about 2 to 3 AI or automation solutions you've personally built. For each: the problem, the stack, what you built, and the measurable result. read the answer → $935,596 spent on freelancers, and then asked this a portfolio of real work you can walk us through read the answer → $665,013 spent on freelancers, and then asked this Name one production system where 5+ coding agents opened real PRs against one large codebase with Linear as the source of intent. What did agents… read the answer → $665,013 spent on freelancers, and then asked this Delegation silently did nothing while the webhook endpoint looked healthy. What's your first suspect, how do you confirm it inside an hour, and what… read the answer → $665,013 spent on freelancers, and then asked this Linear redelivers webhooks and agents retry. A duplicate session event double-dispatched a task and produced two PRs. Where does idempotency live,… read the answer → $665,013 spent on freelancers, and then asked this A conductor splits an epic into ~20 dependency-ordered work packages built by parallel agents. Model it in Linear. What breaks at 50 agent-driven… read the answer → $609,912 spent on freelancers, and then asked this Contributions to anthropic-sdk-python or claude-agent-sdk-python/typescript, or filed reproducible bug reports against them. read the answer → $609,912 spent on freelancers, and then asked this any open-source contributions or repos we can review read the answer → $494,665 spent on freelancers, and then asked this Describe the most complex voice AI agent you have shipped to production. What was the workflow, how did you handle interruptions and confirmations,… read the answer → $494,665 spent on freelancers, and then asked this Describe a voice interaction that failed in production: noise, an accent, silence, or an action that did not commit, and how you handled it. read the answer → $483,589 spent on freelancers, and then asked this Describe an MCP server, AI agent integration, or LLM tool-calling system you personally built. What APIs did it expose, how did you enforce what the… read the answer → $483,589 spent on freelancers, and then asked this Describe a production API you built that handled untrusted or automated callers. How did you handle authentication, rate limiting, input validation… read the answer → $483,589 spent on freelancers, and then asked this Published or open-source MCP servers. Include links in your application. read the answer → $376,190 spent on freelancers, and then asked this three examples of AI systems you personally built and a short video walking us through your work read the answer → $373,434 spent on freelancers, and then asked this A short 5 to 10 line story of a bug you fixed using Claude Code, including one thing the agent got wrong and how you caught it. read the answer → $373,434 spent on freelancers, and then asked this You use Claude Code daily to ship real code, not just chat. Explain your setup: CLAUDE.md, subagents, plan mode, hooks, how you review what the agent… read the answer → $373,434 spent on freelancers, and then asked this Your pre-deploy checklist for a production release, as a bullet list. read the answer → $339,494 spent on freelancers, and then asked this How has AI changed the way you do data work in the last year? Be specific: which tools, what you've built, and what you stopped doing by hand. read the answer → $325,374 spent on freelancers, and then asked this Have you fine-tuned a coding LLM before? Send benchmark/results. read the answer →
$325,374 spent on freelancers, and then asked this How would you create 10,000 executable repo-level coding tasks?" / "How would you prevent SWE-bench data contamination?" / "SFT, DPO or RL first, and… read the answer → $325,374 spent on freelancers, and then asked this Success will be measured using actual test-passing repository tasks, not training loss. read the answer → $318,518 spent on freelancers, and then asked this Tell us about a technical decision you made in the last two years that you later regretted. What did you do differently afterwards? read the answer → $318,518 spent on freelancers, and then asked this A student has spent 100 hours talking to our life coach over six months. How would you architect the memory? read the answer → $309,310 spent on freelancers, and then asked this You have an 8-hour chest-mounted body-camera video and need to detect 20 behaviours, some lasting 10 seconds and some 30+ minutes. How would you… read the answer → $309,310 spent on freelancers, and then asked this How would you measure false positives, false negatives, precision, recall and overall performance of a video LLM system? read the answer → $309,310 spent on freelancers, and then asked this When would you use prompting vs RAG vs fine-tuning for this problem? read the answer → $288,651 spent on freelancers, and then asked this Describe a production system you built with Claude Code, skills, MCP, subagents, hooks and/or gates. What made it reliable, and what broke along the… read the answer → $288,651 spent on freelancers, and then asked this Walk through a specific agent/LLM behaviour bug you diagnosed from logs or traces: the symptom, the real root cause, and how you proved the fix… read the answer → $288,651 spent on freelancers, and then asked this How do you decide which model and reasoning settings to use for a given agent task, and how do you design gates to stop an agent doing the wrong… read the answer → $288,651 spent on freelancers, and then asked this Measure before and after with evidence, not vibes. read the answer → $275,137 spent on freelancers, and then asked this How comfortable are you with a Python debugger? With PyTest? read the answer → $274,389 spent on freelancers, and then asked this Tell us about a tool or automation you built that broke after launch. What caused it, how did you fix it, and what did you change to reduce the chance… read the answer → $226,195 spent on freelancers, and then asked this Describe your approach to testing and improving QA read the answer → $90,310 spent on freelancers, and then asked this Walk through how you'd design accept/reject decisioning that must respond under 200ms at 50 requests a second. read the answer → $84,145 spent on freelancers, and then asked this How would you prevent the AI from presenting unsupported assumptions as verified research findings?" / "How would you retain sources so every… read the answer → $44,970 spent on freelancers, and then asked this How would you prevent Claude from accidentally deploying changes to the wrong client's live store?" / "How would you securely manage credentials for… read the answer → Mexico FDE spent on freelancers, and then asked this Describe a Python backend service you built that had to run unattended: how did you handle queues, retries, idempotency and failure alerts? read the answer → A client spent on freelancers, and then asked this Describe a scraper you built that ran in production. What broke it, and how did you find out it had broken? read the answer → Upwork itself, avg bid $67.92, max $350 spent on freelancers, and then asked this Describe one GRC automation you personally built and shipped. What manual process did it replace, what was the stack, roughly how much time did it… read the answer → A client spent on freelancers, and then asked this Explain how you would detect your own scraper silently breaking." with the client adding: *the second half is the part we care about most* read the answer → A client spent on freelancers, and then asked this Given a 60-case fixture set at 48% pass, how do you avoid overfitting fixes to the fixtures? read the answer → A client spent on freelancers, and then asked this Where would you put the check that prevents a false 'I moved it' claim, and why there rather than in the prompt? read the answer → A client spent on freelancers, and then asked this Give an example of an AI agent or LLM vulnerability you found and how you reproduced it. read the answer → Upwork GRC spent on freelancers, and then asked this How do you make AI-generated compliance output defensible to an auditor? read the answer → German SaaS, 97,351 hours spent on freelancers, and then asked this Which AI coding tools do you use daily, and one specific trick or setup that makes you faster than the average user. read the answer → A client spent on freelancers, and then asked this Do you have experience designing reproducible test suites or benchmarks? read the answer → a literal screening field on at least 4 jobs spent on freelancers, and then asked this Include a link to your GitHub profile and/or website read the answer → German SaaS spent on freelancers, and then asked this A link or screenshots of something real you shipped: repo, staging URL, screen recording, anything we can actually look at. read the answer → Mexico FDE spent on freelancers, and then asked this Please include one link (repo, demo or write-up) that shows a messaging agent or backend service you built and operated. read the answer → Songsterr, avg bid $102.15 spent on freelancers, and then asked this Start your proposal with the strongest thing you have personally shipped to production, with measurable results and your specific role. read the answer → $325,374 spent on freelancers, and then asked this How would you create 10,000 executable repo-level coding tasks?" / "How would you prevent SWE-bench data contamination?" / "SFT, DPO or RL first, and… read the answer → $325,374 spent on freelancers, and then asked this Success will be measured using actual test-passing repository tasks, not training loss. read the answer → $318,518 spent on freelancers, and then asked this Tell us about a technical decision you made in the last two years that you later regretted. What did you do differently afterwards? read the answer → $318,518 spent on freelancers, and then asked this A student has spent 100 hours talking to our life coach over six months. How would you architect the memory? read the answer → $309,310 spent on freelancers, and then asked this You have an 8-hour chest-mounted body-camera video and need to detect 20 behaviours, some lasting 10 seconds and some 30+ minutes. How would you… read the answer → $309,310 spent on freelancers, and then asked this How would you measure false positives, false negatives, precision, recall and overall performance of a video LLM system? read the answer → $309,310 spent on freelancers, and then asked this When would you use prompting vs RAG vs fine-tuning for this problem? read the answer → $288,651 spent on freelancers, and then asked this Describe a production system you built with Claude Code, skills, MCP, subagents, hooks and/or gates. What made it reliable, and what broke along the… read the answer → $288,651 spent on freelancers, and then asked this Walk through a specific agent/LLM behaviour bug you diagnosed from logs or traces: the symptom, the real root cause, and how you proved the fix… read the answer → $288,651 spent on freelancers, and then asked this How do you decide which model and reasoning settings to use for a given agent task, and how do you design gates to stop an agent doing the wrong… read the answer → $288,651 spent on freelancers, and then asked this Measure before and after with evidence, not vibes. read the answer → $275,137 spent on freelancers, and then asked this How comfortable are you with a Python debugger? With PyTest? read the answer → $274,389 spent on freelancers, and then asked this Tell us about a tool or automation you built that broke after launch. What caused it, how did you fix it, and what did you change to reduce the chance… read the answer → $226,195 spent on freelancers, and then asked this Describe your approach to testing and improving QA read the answer → $90,310 spent on freelancers, and then asked this Walk through how you'd design accept/reject decisioning that must respond under 200ms at 50 requests a second. read the answer → $84,145 spent on freelancers, and then asked this How would you prevent the AI from presenting unsupported assumptions as verified research findings?" / "How would you retain sources so every… read the answer → $44,970 spent on freelancers, and then asked this How would you prevent Claude from accidentally deploying changes to the wrong client's live store?" / "How would you securely manage credentials for… read the answer → Mexico FDE spent on freelancers, and then asked this Describe a Python backend service you built that had to run unattended: how did you handle queues, retries, idempotency and failure alerts? read the answer → A client spent on freelancers, and then asked this Describe a scraper you built that ran in production. What broke it, and how did you find out it had broken? read the answer → Upwork itself, avg bid $67.92, max $350 spent on freelancers, and then asked this Describe one GRC automation you personally built and shipped. What manual process did it replace, what was the stack, roughly how much time did it… read the answer → A client spent on freelancers, and then asked this Explain how you would detect your own scraper silently breaking." with the client adding: *the second half is the part we care about most* read the answer → A client spent on freelancers, and then asked this Given a 60-case fixture set at 48% pass, how do you avoid overfitting fixes to the fixtures? read the answer → A client spent on freelancers, and then asked this Where would you put the check that prevents a false 'I moved it' claim, and why there rather than in the prompt? read the answer → A client spent on freelancers, and then asked this Give an example of an AI agent or LLM vulnerability you found and how you reproduced it. read the answer → Upwork GRC spent on freelancers, and then asked this How do you make AI-generated compliance output defensible to an auditor? read the answer → German SaaS, 97,351 hours spent on freelancers, and then asked this Which AI coding tools do you use daily, and one specific trick or setup that makes you faster than the average user. read the answer → A client spent on freelancers, and then asked this Do you have experience designing reproducible test suites or benchmarks? read the answer → a literal screening field on at least 4 jobs spent on freelancers, and then asked this Include a link to your GitHub profile and/or website read the answer → German SaaS spent on freelancers, and then asked this A link or screenshots of something real you shipped: repo, staging URL, screen recording, anything we can actually look at. read the answer → Mexico FDE spent on freelancers, and then asked this Please include one link (repo, demo or write-up) that shows a messaging agent or backend service you built and operated. read the answer → Songsterr, avg bid $102.15 spent on freelancers, and then asked this Start your proposal with the strongest thing you have personally shipped to production, with measurable results and your specific role. read the answer →