
In Progress
Posted
Rubric Task Author — Project Bulls Eye Contract, Paid Per Task About the Role AI models are already being used for real professional work — writing reports, answering questions, building analyses. But they're not reliable enough yet: they still get things wrong in ways that matter, and right now nobody can consistently find out where. We're looking for domain experts to help find those breaking points. You'll build realistic, professional-grade deliverables in your area of expertise, then build the rubric that grades them — creating prompts hard enough that today's best AI models still get wrong, but realistic enough that a working professional would genuinely need them answered correctly. This work spans 67 occupations across 16 domain groups, including Healthcare, Finance & Accounting, Legal & Compliance, Engineering, Supply Chain & Procurement, and Business Operations. What You'll Do • Design a realistic professional task from your field — the kind of question, report, or analysis a working professional would actually be asked to produce. • Write the prompt itself: the realistic, hard question or scenario a professional would need answered. • Build the deliverable yourself, to the standard a qualified professional would expect. • Build a rubric that precisely and objectively grades that deliverable — capturing exactly what separates a correct, professional-quality answer from a flawed one. • Stress-test the task against current AI models to confirm it's genuinely hard — the model should get it wrong or fall short in a meaningful way. • Pass the project assessment required to qualify for task work. • Submit your task for review and validation. What We're Looking For • Real professional experience in one or more of the 16 domain groups (Healthcare, Finance & Accounting, Legal & Compliance, Engineering, Supply Chain & Procurement, Business Ops, or similar). • Deep enough expertise to know what a genuinely hard, realistic professional scenario looks like in your field — not a textbook question, but something with the ambiguity, edge cases, or judgment calls real work involves. • Ability to write clearly and define objective, defensible grading criteria — you're not just producing an answer, you're specifying what “correct” means. • Comfort working independently through a structured, guided workflow. Requirements to Get Started • Complete and pass the Project Bulls Eye project assessment to qualify for task work. • For each task, write an original prompt and build out the full deliverable + rubric as described above. Pay Structure • $30 per task, • Payment is issued only after your task is validated and passes review — meaning it clears the required quality checks and is confirmed to meet the project's standards before it's accepted. • Tasks that don't pass validation are not eligible for payment; you're welcome to revise and resubmit where applicable. Engagement Details • Contract / freelance, task-based. • Work through the guided workflow in the Project Bulls Eye learning hub, then pass the project assessment before submitting your first task. • Best suited for professionals who want flexible, self-directed contract work tied to their existing domain expertise. To apply, tell us which of the 16 domain groups and occupations best match your background, and briefly describe a real, hard professional scenario from your field that you think would stump an AI model today.
Project ID: 40672906
18 proposals
Remote project
Active 6 days ago
Set your budget and timeframe
Get paid for your work
Outline your proposal
It's free to sign up and bid on jobs

Hello, I am very interested in the "Rubric Task Author - Project Bulls Eye" role. I have strong writing and research skills and I understand how to create professional, realistic tasks and grade AI outputs against a clear rubric. I am detail-oriented and can follow instructions carefully to help identify where AI models fail. I am available to start immediately and can work for $5/hour. My goal is to deliver high-quality tasks and build a long-term working relationship with you. Thank you for your consideration. I look forward to contributing to Project Bulls Eye. Best regards, Hamza
$5 USD in 40 days
0.0
0.0
18 freelancers are bidding on average $8 USD/hour for this job

AI Model Testing: Rubric Task Development requires domain experts who can construct realistic professional scenarios that expose model failures—this demands both deep subject matter knowledge and the ability to write objective, defensible grading criteria. Background spans 6+ years across Finance, Accounting, Business Operations, and Technical Writing, with proven ability to construct complex analytical deliverables, statistical analyses, and professional-grade reports. This directly aligns with the stress-testing requirement: building tasks in Finance & Accounting that reflect genuine professional ambiguity, edge cases, and judgment calls rather than textbook problems. Strength lies in recognizing where current AI models break down—particularly in financial analysis requiring interpretation of conflicting data sources, accounting scenarios with regulatory nuance, or business strategy decisions with incomplete information. Can design prompts, deliver professional-standard solutions, and build rubrics that capture objective separation between correct and flawed outputs. Ready to complete the Project Bulls Eye assessment and submit validated tasks across Finance & Accounting domain immediately.
$2 USD in 1 day
7.9
7.9

AI Model Testing — Rubric Task Author for Finance & Accounting Domain The core challenge here is designing professional-grade scenarios that expose AI reliability gaps in domains where errors carry material consequences. Finance and accounting work requires precision in analysis, interpretation of ambiguous data, and judgment calls that involve regulatory nuance—exactly where current models fail consistently. Based on 15+ years in financial analysis and statistical work, this aligns directly with existing expertise in SPSS, R, E-Views, and financial research. The workflow matches the current service model: design realistic scenario, deliver professional-quality output, define objective grading criteria, and validate against actual AI performance. Proposed domain: Finance & Accounting. Initial scenario: A complex variance analysis task involving multi-period financial data with conflicting indicators, requiring judgment on root cause interpretation and materiality assessment—the kind of ambiguity that trips current models while working professionals navigate routinely. Ready to complete the Project Bulls Eye assessment immediately and submit validated tasks meeting the quality threshold.
$2 USD in 1 day
8.0
8.0

Hi, I am a Chartered Accountant (ACCA) with 7 years of experience in accounting and finance, well-versed in developing professional-grade deliverables and assessments. I can create realistic scenarios that challenge AI models, such as preparing a comprehensive tax strategy for a complex corporate structure or analyzing financial statements that include ambiguous financial reporting issues. I will write clear, precise prompts and design rubrics that effectively differentiate between correct and flawed answers, using my expertise to ensure they meet the standards expected in finance. I will ensure timely delivery of tasks while maintaining open communication throughout the process. After passing the initial assessment, I will rigorously test the tasks against current AI models to verify their effectiveness. Could you specify which domain group you would like me to focus on for the tasks?
$5 USD in 7 days
5.3
5.3

Hello, As an experienced and versatile freelancer, I bring a unique blend of skills and knowledge that makes me an ideal candidate for this project. My journey as a content creator has exposed me to diverse industries, including Healthcare, Finance & Accounting, Legal & Compliance, Engineering, Supply Chain & Procurement, Business Ops among others- reflecting the broad spectrum your project entails. My insider know-how means I instinctively understand the nuances these professional fields demand and can build compelling deliverables that resonate. One of the aspects that truly sets me apart is my ability to think beyond textbooks and understand the ambiguities, edge cases, and judgment calls that genuine professional work entails. This capability is intrinsic to effective rubric designing which ultimately determines whether AI models are fit for practical use. Additionally, my commendable skills in clear writing combined with the knack to define objective grading criteria will ensure we capture exactly what separates a correct and professional-quality response from a flawed one. Lastly, my technological literacy enables me to seamlessly work through structured workflows such as Bulls Eye with minimal supervision. I am confident I can build valuable and insightful tasks that will meaningfully push AI models to their limits in your project. I eagerly anticipate adding value to your project and rising to any challenge it may offer. Th Thanks!
$50 USD in 39 days
2.8
2.8

Your approach to AI model testing is spot-on. The challenge of creating realistic tasks that reflect genuine professional scenarios is crucial. A mismatch between AI output and real-world expectations can lead to costly errors, especially in sectors like Healthcare or Finance. Imagine a finance report where an AI misinterprets nuanced data—this could lead to significant financial miscalculations. I'd start with crafting scenarios that include edge cases, ensuring they reflect real job demands. For instance, in Healthcare, a prompt could involve interpreting patient data with incomplete information, challenging for AI but essential for professionals. The rubric I develop would focus on clarity, accuracy, and the reasoning behind decisions, guiding the grading process. I've tackled similar projects, always emphasizing revisions to enhance clarity and precision. What specific domain group do you need my expertise in, and do you have existing materials or frameworks to align with?
$4 USD in 7 days
2.9
2.9

Yes – I am interested in the Rubric Task Author role for Project Bulls Eye. My domain expertise spans **Finance & Accounting**, **Business Operations**, and **Supply Chain & Procurement**. I can design realistic, hard professional tasks (e.g., building a financial model under uncertainty, drafting a procurement strategy with ambiguous requirements, or evaluating operational risk) that would genuinely challenge today's AI models, write clear prompts, build the deliverable myself, and define objective grading rubrics. I’m comfortable working through the guided workflow and project assessment. Available now – ready to start.
$5 USD in 40 days
1.4
1.4

Hi, **I am the best fit for this project because I have a strong background in Statistics, data analysis, QA/testing, and structured problem-solving.** My best-matching areas are **Business Operations, Data/Analytics, and Quality Assurance**. I can create realistic professional tasks, build accurate deliverables, and develop clear rubrics that objectively evaluate AI responses. A challenging scenario I could design would involve a messy business dataset with missing values, outliers, and inconsistent records, requiring accurate analysis, validation, and defensible conclusions. I’m detail-oriented, a fast learner, and comfortable following structured workflows. I’m ready to complete the assessment and start immediately. Best regards, ZEENAT
$5 USD in 40 days
2.0
2.0

Hi, I am a full-stack AI developer with 8+ years of experience in software development, with a background in AI development, automation, technical documentation, data systems, and production software engineering. I am experienced in designing realistic technical scenarios, evaluating AI outputs, and creating objective criteria for assessing correctness, reliability, and edge cases. For this project, I would be a strong fit for the **Engineering / Business Operations / AI & Technology** domain. A challenging scenario I could develop would involve an AI analyzing a production data pipeline incident, identifying the actual root cause from incomplete logs and conflicting metrics, proposing a remediation plan, and distinguishing symptoms from underlying failures. The rubric could objectively evaluate technical accuracy, reasoning, risk assessment, proposed fixes, and production-readiness. I’m an individual freelancer and can work on any time zone you want. Please contact me with the best time for you to have a quick chat. Looking forward to discussing more details. Thanks. Emile
$15 USD in 40 days
0.0
0.0

Domain expert with 11+ years in Finance/Accounting/Business Operations. I'll design realistic professional scenarios, build prompts that stump AI models, create deliverable+rubric, and stress-test against current AI. Domain: Financial reporting & analysis—hard scenario: interpret ambiguous accounting standards for complex multi-entity consolidation with edge cases. Ready to pass assessment and submit tasks. Let's build robust AI evaluation tasks!
$5 USD in 40 days
0.0
0.0

⮞⮞⮞⮞⮞ Dear client ⮜⮜⮜⮜⮜, Thank you for seeing my proposal. I have carefully reviewed your Rubric Task Author role and I am confident I can design rigorous, realistic professional tasks. -What are you looking to solve in this project? You need domain experts to build realistic professional-grade deliverables and rubrics—creating prompts hard enough that today's AI models get wrong, but realistic enough that working professionals genuinely need them answered correctly—across 16 domain groups including Finance, Healthcare, Legal, and Engineering. -What I can do for you in this project? I have deep expertise in Finance & Accounting and Business Operations. I can design a realistic professional scenario (e.g., financial analysis with conflicting data), build the correct deliverable, and create a precise, objective rubric that grades it. I can stress-test against current AI models to ensure genuine difficulty. ⚠️ If you want to solve more, I will do—additional tasks across multiple domains or advanced grading logic. Thank you for your time. I am confident I can deliver high-quality, professionally grounded tasks. Best Regards
$5 USD in 60 days
0.0
0.0

To sum up, my real-life experience across several professions aligns perfectly with your project requirements. My skills in understanding nuanced problems faced by working professionals, as well as my ability to provide clear instructions with defensible grading criteria will provide meaningful insights into refining AI models for professional-grade work. I look forward to discussing how we can contribute to the reliable implementation of AI technologies in the workforce through our collaboration. Thanks!
$4 USD in 40 days
0.0
0.0

Hi, there! I recently worked on a project where I designed realistic assessment tasks for AI models in the finance sector, focusing on complex scenarios that challenge their accuracy. In this role, I developed prompts that required critical thinking and nuanced judgment, reflecting real-world financial dilemmas. One challenge was ensuring that the tasks captured the ambiguity and edge cases professionals face, pushing the AI to its limits. By rigorously testing these tasks against various AI models, I identified specific weaknesses and refined the prompts to enhance their effectiveness. For your project, I propose creating challenging, domain-specific tasks that reflect real professional scenarios in fields like finance or accounting. This includes developing detailed rubrics to objectively assess the AI's performance, ensuring that the tasks are both realistic and difficult enough to expose the limitations of current models. With my extensive background in finance and AI research, your project will likely be completed successfully. Hope to discuss this in detail. Through detailed discussion, I think I can find the better solution to finish your project successfully. Thank you!
$10 USD in 40 days
0.0
0.0

atlanta, United States
Payment method verified
Member since Oct 24, 2019
$2-8 USD / hour
$2-8 USD / hour
$2-8 USD / hour
$2-8 USD / hour
$2-8 USD / hour
₹150000-250000 INR
$250-750 USD
£250-750 GBP
₹750-1250 INR / hour
$10-30 USD
₹12500-37500 INR
$80 AUD
$30-250 USD
$8-15 USD / hour
$15-25 USD / hour
₹12500-37500 INR
₹12500-37500 INR
$2-8 USD / hour
$30-250 USD
$2-8 USD / hour
$25-50 USD / hour
₹750-1250 INR / hour
$250-750 USD
£250-750 GBP
$250-750 USD