
Closed
Posted
Paid on delivery
I’m running a series of tightly-scoped tasks under the Seal India initiative and need a partner who can take each assignment from prompt design through final submission without hand-holding. Every task pays $30 on successful delivery, and the quicker you can turn them around, the more I can queue up for you. You’ll be working with large-language models—ChatGPT-5.5, Claude, or any equivalent you already subscribe to—probing them for search and reasoning failures. I supply the overall objective; you create the cleverly-crafted prompts, document outcomes, and package the results exactly as the task spec requires. My ideal collaborator already: • Uses advanced AI tools daily and is comfortable chaining prompts, applying system messages, and analysing token-level behaviour. • Runs a reliable residential proxy so IP rotation or geo-targeting never slows you down. • Has previous hands-on experience with model debugging, AI model training, and AI-powered web scraping, so edge-case thinking comes naturally. Deliverables for each task 1. A brief rationale for your prompt strategy. 2. The full prompt set (system, user, follow-ups) with any temperature / max-token tweaks noted. 3. Raw model output plus a concise analysis highlighting failure modes. 4. A final, well-formatted submission ready for my portal. Acceptance criteria • The model failure exposed must match the objective I provide. • All prompts and outputs must reproduce on my end when I validate them behind a different proxy. • Turnaround per task: 24 hours or less. If this workflow sounds like second nature to you, let’s start with the first assignment immediately and keep the pipeline moving.
Project ID: 40635218
14 proposals
Remote project
Active 1 day ago
Set your budget and timeframe
Get paid for your work
Outline your proposal
It's free to sign up and bid on jobs
14 freelancers are bidding on average ₹24,464 INR for this job

100% doable. Been building and testing AI agents hands-on, including deploying agents for a law firm handling intake and follow-ups end to end, so probing model behavior and documenting failure modes is a natural extension of that work. I'll design targeted prompt chains around whatever objective you give me, document the full prompt set with system messages and any parameter tweaks, then package the raw output with a clear breakdown of where the model failed and why. Reproducible on your end, formatted exactly to your portal spec. Can turn the first task around well within your 24-hour window and keep pace as you queue more. Feel free to DM me for case studies. Or check the projects on my profile. Let's do it.
₹25,000 INR in 7 days
3.7
3.7

Hi, there. This work is mainly about controlled experimentation rather than simply writing prompts. Each task needs a reproducible setup where the prompt sequence, model settings, outputs, and observed failure mode can be clearly connected to the objective. I’d approach each assignment by first translating the objective into a concrete failure hypothesis, then designing system/user/follow-up prompts that isolate the behavior being tested. I’d keep model parameters and environmental assumptions documented, capture the raw outputs, and separate observations from interpretation so the final submission is easy to validate. For cases involving search, reasoning, or web-based behavior, I’d also test edge cases and alternative prompt paths rather than relying on a single successful failure. Reproducibility would be important, especially when validation happens from a different network environment. The 24-hour turnaround makes a consistent workflow useful: objective → hypothesis → prompt set → controlled tests → failure analysis → formatted submission. For the first assignment, will you provide a fixed model/environment and task objective, or should I also determine the testing setup and model configuration? Thanks, Jaroslav Caprata
₹15,000 INR in 3 days
3.4
3.4

I can start with a paid pilot and turn each objective into a reproducible eval packet within 24 hours: prompt rationale, complete system, user, and follow-up sequence, model and decoding settings, raw outputs, failure analysis, and portal-ready formatting. My strength is building controlled evaluations instead of hunting for one-off screenshots. I vary one factor at a time, keep a run log, separate search failures from reasoning failures, and retest across a second session before delivery. I work daily with GPT and Claude-class models, structured prompt chains, tool-using agents, and evidence-backed QA. I can support the India-targeted environment and follow your proxy and reproducibility requirements. For the first assignment, please provide the target model, permitted tools, success definition, and portal schema. I will return the first complete packet inside 24 hours. This bid assumes a five-task batch. I am happy to begin with one $30 milestone and scale only after you reproduce the result.
₹12,500 INR in 5 days
0.0
0.0

Namaste I understand you need prompt design testing and documentation for LLM failure probing delivered within 24 hours each. My daily work with ChatGPT and Claude plus experience building AI pipelines on AWS lets me craft system and user prompts tweak temperature and token limits and analyse outputs quickly. I will provide rationale full prompt set raw output and concise failure analysis ready for your portal reproducible behind any proxy I can start on the first task today and keep the flow steady Do you have a preferred model version for the initial assignment Are there any token or max‑output constraints you want enforced Will the proxy need a specific geographic region for validation
₹25,000 INR in 1 day
0.0
0.0

With an extensive nine-year history in web and mobile development, I've honed a unique set of skills that I believe make me perfect for your AI-powered task list. My command over multiple programming languages such as Java, PHP, and Python - combined with an expert understanding of HTML/HTML5/CSS - will allow me to optimize and craft effective prompts with ease. In addition, my proficiency with AI tools is a daily practice for me which includes debugging models, training them and using AI-powered web scraping techniques. Consequently, I am accustomed to thinking in edge-cases and chaining prompts - two skills that are crucial for the success of your tasks under the Seal India initiative. Lastly, hosting and maintaining websites have always been a part of my work ethos, making your stringent acceptance criteria of reproducing prompts and outputs on-demand behind a different proxy an assurance from my end. With all-inclusive expertise from developing your prompt strategies to delivering the final well-formatted submissions within 24 hours or less, I guarantee proficiency, reliability, and efficiency throughout the project.
₹25,000 INR in 7 days
0.0
0.0

Hi there! I read your task requirements carefully—probing models like ChatGPT-5.5 & Claude for search/reasoning failures and delivering reproducable, structured logs within 24 hours is right up my alley. I am equipped with residential proxies and ready to take on Assignment #1 immediately. Why I am a great fit for Seal India: • Prompt Engineering & Failure Analysis: Experienced in crafting complex system/user prompts, tweaking parameters (temperature/top_p), and identifying edge-case failure modes. • Proxy & Environment Ready: I use residential proxy configurations to ensure seamless reproduction of results across different geo-locations. • Fast Turnaround: I can consistently deliver structured task files within 24 hours to keep your pipeline moving smoothly. What I will deliver per task: Clear prompt strategy rationale. Complete prompt set (system, user, follow-up parameters). Raw model outputs + concise analysis highlighting failure modes. Clean submission package ready for your portal. Quick Questions: What format do you prefer for portal submissions (JSON, Markdown, or PDF)? Do you have a specific primary LLM model you want to focus on for the first assignment? Ready to start immediately upon award! Best regards, Asma
₹25,000 INR in 7 days
0.0
0.0

Hello, I have hands-on experience using advanced LLMs including ChatGPT and Claude for prompt engineering, technical research, automation, and workflow analysis. I am comfortable designing structured test cases to identify reasoning, retrieval, and search-related failure modes. For each assignment, I can provide: • Prompt strategy and testing rationale • Complete prompt chain including system and follow-up prompts • Reproducible outputs and observations • Failure-mode analysis with supporting evidence • Final submission formatted according to your specifications My background in cloud computing and technical validation allows me to approach AI evaluation systematically and document results clearly. I am comfortable working under tight turnaround requirements and delivering detailed reports within 24 hours. I would be interested in reviewing the first assignment and discussing the expected success criteria.
₹25,000 INR in 7 days
0.0
0.0

Senior AI Prompt Engineer: With 15+ years of experience, I specialize in leveraging advanced AI tools to develop effective prompt strategies tailored for large-language models. I understand your need for efficient, reliable execution without the need for close oversight. Proposed Solution: - Craft intricate prompts that probe for nuanced search and reasoning failures in models like ChatGPT-5.5 and Claude. - Utilize advanced AI tools daily for chaining prompts and analyzing token-level behavior to ensure optimal performance. - Implement a reliable residential proxy for seamless geo-targeting and IP rotation, ensuring no disruptions in workflow. Scope of Delivery: - A comprehensive rationale outlining the prompt strategy for each task. - A complete set of prompts, including system and user messages, with adjustments for temperature and max tokens. - Raw model outputs accompanied by a concise analysis pinpointing failure modes. - A professionally formatted final submission ready for your portal. Success Metrics: - Ensure that model failures identified match the provided objectives for validation. - All outputs and prompts must reproduce seamlessly under different proxies, ensuring reliability and consistency. Delivery Timeline: Each task will be completed within 24 hours, with documentation provided to guarantee clarity and reproducibility. Best Regards, Karthik B Resonite Tech
₹55,000 INR in 7 days
0.0
0.0

I have hands-on experience with AI tools for content creation, image generation, creative design, research, and workflow automation. I’m comfortable using AI to generate and improve content, create visual concepts, write prompts, research information, and speed up repetitive tasks. I also have experience exploring AI automation tools and finding practical ways to use AI for real-world projects. I’m a quick learner and can adapt to new AI tools and workflows based on the project requirements. I focus on delivering useful, clean, and practical results rather than simply generating AI content.
₹25,000 INR in 7 days
0.0
0.0

I would love to work with you you can give me a sample work if you like my work we can work on future projects too
₹25,000 INR in 7 days
0.0
0.0

I have hands-on experience working with LLMs, prompt engineering, and AI research. I’m comfortable designing structured prompt experiments to identify reasoning, retrieval, and search-related failure modes. For each assignment, I can provide: • Prompt strategy and testing rationale • Complete system, user, and follow-up prompts • Model settings and outputs • Failure analysis and reproducible documentation • Final submission formatted to your requirements. I can start immediately and work within the 24-hour turnaround.
₹15,000 INR in 4 days
0.0
0.0

Bharatpur, India
Member since Oct 16, 2023
$8-15 USD / hour
₹600-1500 INR
$15-25 AUD / hour
$10-30 USD
$250-750 USD
₹12500-37500 INR
₹12500-37500 INR
$250-750 USD
₹750-1250 INR / hour
$250-750 USD
$500 AUD
$250-750 SGD
$250-750 USD
$250-750 USD
₹1500-12500 INR
₹12500-37500 INR
$30-250 USD
$15-25 USD / hour
$25-50 USD / hour
₹12500-37500 INR