
In Progress
Posted
Rubric Task Author — Project Bulls Eye (Contract, Paid Per Task) About the Role AI models are already being used for real professional work — writing reports, answering questions, building analyses. But they're not reliable enough yet: they still get things wrong in ways that matter, and right now nobody can consistently find out where. We're looking for domain experts to help find those breaking points. You'll build realistic, professional-grade deliverables in your area of expertise, then build the rubric that grades them — creating prompts hard enough that today's best AI models still get wrong, but realistic enough that a working professional would genuinely need them answered correctly. This work spans 67 occupations across 16 domain groups, including Healthcare, Finance & Accounting, Legal & Compliance, Engineering, Supply Chain & Procurement, and Business Operations. What You'll Do Design a realistic professional task from your field — the kind of question, report, or analysis a working professional would actually be asked to produce. Write the prompt itself: the realistic, hard question or scenario a professional would need answered. Build the deliverable yourself, to the standard a qualified professional would expect. Build a rubric that precisely and objectively grades that deliverable — capturing exactly what separates a correct, professional-quality answer from a flawed one. Stress-test the task against current AI models to confirm it's genuinely hard — the model should get it wrong or fall short in a meaningful way. Pass the project assessment required to qualify for task work. Submit your task for review and validation. What We're Looking For Real professional experience in one or more of the 16 domain groups (Healthcare, Finance & Accounting, Legal & Compliance, Engineering, Supply Chain & Procurement, Business Ops, or similar). Deep enough expertise to know what a genuinely hard, realistic professional scenario looks like in your field — not a textbook question, but something with the ambiguity, edge cases, or judgment calls real work involves. Ability to write clearly and define objective, defensible grading criteria — you're not just producing an answer, you're specifying what "correct" means. Comfort working independently through a structured, guided workflow. Requirements to Get Started Complete and pass the Project Bulls Eye project assessment to qualify for task work. For each task, write an original prompt and build out the full deliverable + rubric as described above. Pay Structure $30 per task, paid per task rather than by the hour. Payment is issued only after your task is validated and passes review — meaning it clears the required quality checks and is confirmed to meet the project's standards before it's accepted. Tasks that don't pass validation are not eligible for payment; you're welcome to revise and resubmit where applicable. Engagement Details Contract / freelance, task-based. Work through the guided workflow in the Project Bulls Eye learning hub, then pass the project assessment before submitting your first task. Best suited for professionals who want flexible, self-directed contract work tied to their existing domain expertise. To apply, tell us which of the 16 domain groups and occupations best match your background, and briefly describe a real, hard professional scenario from your field that you think would stump an AI model today.
Project ID: 40674292
6 proposals
Remote project
Active 6 days ago
Set your budget and timeframe
Get paid for your work
Outline your proposal
It's free to sign up and bid on jobs

I am bidding on the Rubric Task Author position, specializing in Engineering and Business Operations. I bring direct AI evaluation experience from working on the Fireweed task on CrowdGen, where I authored complex prompts, stress-tested LLM outputs, and built objective grading rubrics. This is backed by a background in Physics, finance & accounting and structural planning, giving me the technical depth required to craft realistic, high-difficulty benchmarks. To stump current models, I will design a task requiring the application of the variational principle to estimate the ground state energy of a non-standard, custom-defined potential well. Top LLMs regularly fail these scenarios because they attempt to pattern-match standard textbook solutions rather than executing the required step-by-step calculus. I am ready to pass the project assessment and begin producing validated tasks through your workflow immediately.
$4 USD in 40 days
0.0
0.0
6 freelancers are bidding on average $5 USD/hour for this job

Hi, a realistic engineering task, the deliverable, then a rubric that grades it. Writing a hard question is easy. It's the grading criteria that fall apart. Which domain group do you want first? I build production AI systems for banks, so my scenarios come from real failures. Happy to write one task first so you can judge it. Regards, Zohaib
$8 USD in 40 days
2.6
2.6

Hi, My strongest fit is Business Operations / Engineering, with 13+ years of hands-on software engineering across AI, automation, full-stack systems, data extraction, and technical product development. A realistic hard scenario I could contribute is evaluating an AI-generated solution for a production automation system where requirements conflict, edge cases are incomplete, external APIs behave inconsistently, and the solution must still be reliable and maintainable. The task could require the model to make practical engineering decisions and justify them, rather than simply produce textbook code. I’m comfortable creating the prompt, professional deliverable, objective rubric, and stress-testing the task against AI models. I can also complete the required assessment before starting task work. Let's connect and get started soon. Best regards, Binaya T.
$5 USD in 40 days
0.0
0.0

Hi, I’m interested in the Project Bulls Eye role. My strongest areas are Software Engineering, AI/ML, Data Analysis, and Business Technology. I can create realistic professional tasks, complete high-quality deliverables, and build clear, objective rubrics that test AI reasoning rather than simple knowledge. I’m comfortable working independently and following structured evaluation guidelines.
$2 USD in 40 days
0.0
0.0

I have experience creating clear, detailed rubrics that align with project objectives and AI training goals. This ensures consistent task evaluation and improves AI learning accuracy. Previously, I developed rubrics for educational and AI datasets that enhanced model performance and reviewer consistency. My work focused on clarity, relevance, and ease of use. I approach rubric writing by closely aligning criteria with project goals and testing for clarity with reviewers to ensure practical application. Are you looking for tasks focused on a specific subject area or a variety of topics?
$5 USD in 7 days
0.0
0.0

Hi, I’m interested in the Business Operations / Technology & Engineering side of Project Bulls Eye. My background includes full-stack web development, AI automation, APIs, eCommerce systems, databases, and workflow automation with tools such as n8n. A realistic scenario I’d use is: designing an end-to-end automation for a business that receives leads from multiple sources, validates and deduplicates them, enriches the records, routes them to the correct workflow, and reports failures without losing data. The challenge would include conflicting data, duplicate records, API failures, rate limits, incomplete fields, and edge cases where the automation needs human intervention. I’m comfortable creating the task, professional deliverable, and an objective rubric that clearly separates a robust production solution from an answer that only looks correct on the surface. I’m also comfortable following a structured assessment and review process and revising work based on validation feedback. Thanks, Haseeb
$5 USD in 40 days
0.0
0.0

atlanta, United States
Payment method verified
Member since Oct 24, 2019
$2-8 USD / hour
$2-8 USD / hour
$2-8 USD / hour
$2-8 USD / hour
$2-8 USD / hour
₹1500-12500 INR
$250-750 USD
$250-750 USD
₹1500-12500 INR
₹12500-37500 INR
$30-250 USD
₹1500-12500 INR
$250-750 AUD
$3000-5000 USD
$250-750 USD
₹12500-37500 INR
$750-1500 AUD
₹150000-250000 INR
₹750-1250 INR / hour
$250-750 USD
₹600-1500 INR
$250-750 USD
₹600-1500 INR
₹600-1500 INR
$8-15 USD / hour