Cyber Security Engineer Jobs in Jordan
3724 Jobs Found
<h2 class="h5">Job description</h2>
<div class="t-break" data-jb-field="description">
<span>Please submit your CV in English and indicate your level of English proficiency.<br> Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems.<br> Participation is project-based, not permanent employment.<br>What this opportunity involves: We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks.<br> You'll create challenging tasks and evaluation criteria within realistic simulated environments: Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history Design tasks from intermediate states of these environments - craft the prompt, define what "solved" means, and ensure the task is solvable by an AI agent Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust What this is NOT: Not data labeling Not prompt engineering Not writing code from scratch - the agent writes most of the code; you guide and evaluate What we look for: 8+ years in software development Core stack: Python (FastAPI), JavaScript/TypeScript (React), Docker, Postgres, Kafka, Redis Experience writing tests (functional, integration) English proficiency - B2+ Why this is hard: Frontier models are already good at coding.<br> Creating a task that genuinely challenges the best models is non-trivial.<br> You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution.<br> Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<br> How it works Apply → Pass qualification(s) → Join a project → Complete tasks → Get paidEffort estimate Tasks for this project are estimated to take 30 hours to complete, depending on complexity.<br> This is an estimate and not a schedule requirement; you choose when and how to work.<br> Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<br> Compensation: Up to $150/hr equivalent , depending on level and pace.<br> Tasks are estimated at ~30 hours each; you set your own schedule.<br></span> </div>
<p><h4>Description<\/h4>\n<p>Please submit your CV in English and indicate your level of English proficiency.<\/p>\n<p>Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment.<\/p>\n<h4>What this opportunity involves<\/h4>\n<p>We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks.<\/p>\n<p>You'll create challenging tasks and evaluation criteria within realistic simulated environments:<\/p>\n<ul>\n<li>Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history<\/li>\n<li>Design tasks from intermediate states of these environments - craft the prompt, define what \"solved\" means, and ensure the task is solvable by an AI agent<\/li>\n<li>Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient<\/li>\n<li>Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust<\/li>\n<\/ul>\n<h4>What this is not<\/h4>\n<ul>\n<li>Not data labeling<\/li>\n<li>Not prompt engineering<\/li>\n<li>Not writing code from scratch - the agent writes most of the code; you guide and evaluate<\/li>\n<\/ul>\n<h4>What we look for<\/h4>\n<ul>\n<li>5+ years in software development<\/li>\n<li>Core stack: Python (FastAPI), JavaScript\/TypeScript (React), Docker, Postgres, Kafka, Redis<\/li>\n<li>Experience writing tests (functional, integration)<\/li>\n<li>English proficiency - B2+<\/li>\n<\/ul>\n<h4>Why this is hard<\/h4>\n<p>Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<\/p>\n<h4>How it works<\/h4>\n<p>Apply ? Pass qualification(s) ? Join a project ? Complete tasks ? Get paid<\/p>\n<h4>Effort estimate<\/h4>\n<p>Tasks for this project are estimated to take 20 hours to complete, depending on complexity. This is an estimate and not a schedule requirement; you choose when and how to work. Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<\/p>\n<h4>Compensation<\/h4>\n<p>Up to $50\/hr equivalent, depending on level and pace. Tasks are estimated at ~20 hours each; you set your own schedule.<\/p><\/p><p><\/p>
<p><h4>Description<\/h4>\n<p>Please submit your CV in English and indicate your level of English proficiency.<\/p>\n<p>Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment.<\/p>\n<h4>What this opportunity involves<\/h4>\n<p>We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks.<\/p>\n<p>You'll create challenging tasks and evaluation criteria within realistic simulated environments:<\/p>\n<ul>\n <li>Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history<\/li>\n <li>Design tasks from intermediate states of these environments - craft the prompt, define what \"solved\" means, and ensure the task is solvable by an AI agent<\/li>\n <li>Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient<\/li>\n <li>Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust<\/li>\n<\/ul>\n<h4>What this is not<\/h4>\n<ul>\n <li>Not data labeling<\/li>\n <li>Not prompt engineering<\/li>\n <li>Not writing code from scratch - the agent writes most of the code; you guide and evaluate<\/li>\n<\/ul>\n<h4>What we look for<\/h4>\n<ul>\n <li>5+ years in software development<\/li>\n <li>Core stack: Python (FastAPI), JavaScript\/TypeScript (React), Docker, Postgres, Kafka, Redis<\/li>\n <li>Experience writing tests (functional, integration)<\/li>\n <li>English proficiency - B2+<\/li>\n<\/ul>\n<h4>Why this is hard<\/h4>\n<p>Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<\/p>\n<h4>How it works<\/h4>\n<p>Apply? Pass qualification(s)? Join a project? Complete tasks? Get paid<\/p>\n<h4>Effort estimate<\/h4>\n<p>Tasks for this project are estimated to take 20 hours to complete, depending on complexity. This is an estimate and not a schedule requirement; you choose when and how to work. Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<\/p>\n<h4>Compensation<\/h4>\n<p>Up to $50\/hr equivalent, depending on level and pace. Tasks are estimated at approximately 20 hours each; you set your own schedule.<\/p><\/p><p><\/p>
<p><h4>Description<\/h4>\n<p>Please submit your CV in English and indicate your level of English proficiency.<\/p>\n<p>Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment.<\/p>\n<h4>What this opportunity involves<\/h4>\n<p>We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks.<\/p>\n<p>You'll create challenging tasks and evaluation criteria within realistic simulated environments:<\/p>\n<ul>\n <li>Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history<\/li>\n <li>Design tasks from intermediate states of these environments - craft the prompt, define what \"solved\" means, and ensure the task is solvable by an AI agent<\/li>\n <li>Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient<\/li>\n <li>Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust<\/li>\n<\/ul>\n<h4>What this is not<\/h4>\n<ul>\n <li>Not data labeling<\/li>\n <li>Not prompt engineering<\/li>\n <li>Not writing code from scratch - the agent writes most of the code; you guide and evaluate<\/li>\n<\/ul>\n<h4>What we look for<\/h4>\n<ul>\n <li>5+ years in software development<\/li>\n <li>Core stack: Python (FastAPI), JavaScript\/TypeScript (React), Docker, Postgres, Kafka, Redis<\/li>\n <li>Experience writing tests (functional, integration)<\/li>\n <li>English proficiency - B2+<\/li>\n<\/ul>\n<h4>Why this is hard<\/h4>\n<p>Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<\/p>\n<h4>How it works<\/h4>\n<p>Apply? Pass qualification(s)? Join a project? Complete tasks? Get paid<\/p>\n<h4>Effort estimate<\/h4>\n<p>Tasks for this project are estimated to take 20 hours to complete, depending on complexity. This is an estimate and not a schedule requirement; you choose when and how to work. Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<\/p>\n<h4>Compensation<\/h4>\n<p>Up to $50\/hr equivalent, depending on level and pace. Tasks are estimated at approximately 20 hours each; you set your own schedule.<\/p><\/p><p><\/p>
<p><h4>Description<\/h4>\n<p>Please submit your CV in English and indicate your level of English proficiency.<\/p>\n<p>Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment.<\/p>\n<h4>What this opportunity involves<\/h4>\n<p>We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks.<\/p>\n<p>You'll create challenging tasks and evaluation criteria within realistic simulated environments:<\/p>\n<ul>\n<li>Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history<\/li>\n<li>Design tasks from intermediate states of these environments - craft the prompt, define what \"solved\" means, and ensure the task is solvable by an AI agent<\/li>\n<li>Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient<\/li>\n<li>Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust<\/li>\n<\/ul>\n<h4>What this is not<\/h4>\n<ul>\n<li>Not data labeling<\/li>\n<li>Not prompt engineering<\/li>\n<li>Not writing code from scratch - the agent writes most of the code; you guide and evaluate<\/li>\n<\/ul>\n<h4>What we look for<\/h4>\n<ul>\n<li>5+ years in software development<\/li>\n<li>Core stack: Python (FastAPI), JavaScript\/TypeScript (React), Docker, Postgres, Kafka, Redis<\/li>\n<li>Experience writing tests (functional, integration)<\/li>\n<li>English proficiency - B2+<\/li>\n<\/ul>\n<h4>Why this is hard<\/h4>\n<p>Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<\/p>\n<h4>How it works<\/h4>\n<p>Apply ? Pass qualification(s) ? Join a project ? Complete tasks ? Get paid<\/p>\n<h4>Effort estimate<\/h4>\n<p>Tasks for this project are estimated to take 20 hours to complete, depending on complexity. This is an estimate and not a schedule requirement; you choose when and how to work. Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<\/p>\n<h4>Compensation<\/h4>\n<p>Up to $50\/hr equivalent, depending on level and pace. Tasks are estimated at ~20 hours each; you set your own schedule.<\/p><\/p><p><\/p>
<p><h4>Description<\/h4>\n<p>Please submit your CV in English and indicate your level of English proficiency.<\/p>\n<p>Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment.<\/p>\n<h4>What this opportunity involves<\/h4>\n<p>We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks.<\/p>\n<p>You'll create challenging tasks and evaluation criteria within realistic simulated environments:<\/p>\n<ul>\n<li>Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history<\/li>\n<li>Design tasks from intermediate states of these environments - craft the prompt, define what \"solved\" means, and ensure the task is solvable by an AI agent<\/li>\n<li>Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient<\/li>\n<li>Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust<\/li>\n<\/ul>\n<h4>What this is not<\/h4>\n<ul>\n<li>Not data labeling<\/li>\n<li>Not prompt engineering<\/li>\n<li>Not writing code from scratch - the agent writes most of the code; you guide and evaluate<\/li>\n<\/ul>\n<h4>What we look for<\/h4>\n<ul>\n<li>5+ years in software development<\/li>\n<li>Core stack: Python (FastAPI), JavaScript\/TypeScript (React), Docker, Postgres, Kafka, Redis<\/li>\n<li>Experience writing tests (functional, integration)<\/li>\n<li>English proficiency - B2+<\/li>\n<\/ul>\n<h4>Why this is hard<\/h4>\n<p>Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<\/p>\n<h4>How it works<\/h4>\n<p>Apply ? Pass qualification(s) ? Join a project ? Complete tasks ? Get paid<\/p>\n<h4>Effort estimate<\/h4>\n<p>Tasks for this project are estimated to take 20 hours to complete, depending on complexity. This is an estimate and not a schedule requirement; you choose when and how to work. Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<\/p>\n<h4>Compensation<\/h4>\n<p>Up to $50\/hr equivalent, depending on level and pace. Tasks are estimated at ~20 hours each; you set your own schedule.<\/p><\/p><p><\/p>
<p><h4>Description<\/h4>\n<p>Please submit your CV in English and indicate your level of English proficiency.<\/p>\n<p>Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment.<\/p>\n<h4>What this opportunity involves<\/h4>\n<p>We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks.<\/p>\n<p>You'll create challenging tasks and evaluation criteria within realistic simulated environments:<\/p>\n<ul>\n<li>Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history<\/li>\n<li>Design tasks from intermediate states of these environments - craft the prompt, define what \"solved\" means, and ensure the task is solvable by an AI agent<\/li>\n<li>Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient<\/li>\n<li>Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust<\/li>\n<\/ul>\n<h4>What this is not<\/h4>\n<ul>\n<li>Not data labeling<\/li>\n<li>Not prompt engineering<\/li>\n<li>Not writing code from scratch - the agent writes most of the code; you guide and evaluate<\/li>\n<\/ul>\n<h4>What we look for<\/h4>\n<ul>\n<li>5+ years in software development<\/li>\n<li>Core stack: Python (FastAPI), JavaScript\/TypeScript (React), Docker, Postgres, Kafka, Redis<\/li>\n<li>Experience writing tests (functional, integration)<\/li>\n<li>English proficiency - B2+<\/li>\n<\/ul>\n<h4>Why this is hard<\/h4>\n<p>Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<\/p>\n<h4>How it works<\/h4>\n<p>Apply ? Pass qualification(s) ? Join a project ? Complete tasks ? Get paid<\/p>\n<h4>Effort estimate<\/h4>\n<p>Tasks for this project are estimated to take 20 hours to complete, depending on complexity. This is an estimate and not a schedule requirement; you choose when and how to work. Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<\/p>\n<h4>Compensation<\/h4>\n<p>Up to $50\/hr equivalent, depending on level and pace. Tasks are estimated at approximately 20 hours each; you set your own schedule.<\/p><\/p><p><\/p>
<p><h4>Description<\/h4>\n<p>Please submit your CV in English and indicate your level of English proficiency.<\/p>\n<p>Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment.<\/p>\n<h4>What this opportunity involves<\/h4>\n<p>We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks.<\/p>\n<p>You'll create challenging tasks and evaluation criteria within realistic simulated environments:<\/p>\n<ul>\n <li>Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history<\/li>\n <li>Design tasks from intermediate states of these environments - craft the prompt, define what \"solved\" means, and ensure the task is solvable by an AI agent<\/li>\n <li>Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient<\/li>\n <li>Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust<\/li>\n<\/ul>\n<h4>What this is not<\/h4>\n<ul>\n <li>Not data labeling<\/li>\n <li>Not prompt engineering<\/li>\n <li>Not writing code from scratch - the agent writes most of the code; you guide and evaluate<\/li>\n<\/ul>\n<h4>What we look for<\/h4>\n<ul>\n <li>5+ years in software development<\/li>\n <li>Core stack: Python (FastAPI), JavaScript\/TypeScript (React), Docker, Postgres, Kafka, Redis<\/li>\n <li>Experience writing tests (functional, integration)<\/li>\n <li>English proficiency - B2+<\/li>\n<\/ul>\n<h4>Why this is hard<\/h4>\n<p>Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<\/p>\n<h4>How it works<\/h4>\n<p>Apply \u2192 Pass qualification(s) \u2192 Join a project \u2192 Complete tasks \u2192 Get paid<\/p>\n<h4>Effort estimate<\/h4>\n<p>Tasks for this project are estimated to take 20 hours to complete, depending on complexity. This is an estimate and not a schedule requirement; you choose when and how to work. Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<\/p>\n<h4>Compensation<\/h4>\n<p>Up to $50\/hr equivalent, depending on level and pace. Tasks are estimated at approximately 20 hours each; you set your own schedule.<\/p><\/p><p><\/p>
<p><h4>Description<\/h4>\n<p>Please submit your CV in English and indicate your level of English proficiency.<\/p>\n<p>Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment.<\/p>\n<h4>What this opportunity involves<\/h4>\n<p>We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks.<\/p>\n<p>You'll create challenging tasks and evaluation criteria within realistic simulated environments:<\/p>\n<ul>\n <li>Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history<\/li>\n <li>Design tasks from intermediate states of these environments - craft the prompt, define what \"solved\" means, and ensure the task is solvable by an AI agent<\/li>\n <li>Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient<\/li>\n <li>Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust<\/li>\n<\/ul>\n<h4>What this is not<\/h4>\n<ul>\n <li>Not data labeling<\/li>\n <li>Not prompt engineering<\/li>\n <li>Not writing code from scratch - the agent writes most of the code; you guide and evaluate<\/li>\n<\/ul>\n<h4>What we look for<\/h4>\n<ul>\n <li>5+ years in software development<\/li>\n <li>Core stack: Python (FastAPI), JavaScript\/TypeScript (React), Docker, Postgres, Kafka, Redis<\/li>\n <li>Experience writing tests (functional, integration)<\/li>\n <li>English proficiency - B2+<\/li>\n<\/ul>\n<h4>Why this is hard<\/h4>\n<p>Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<\/p>\n<h4>How it works<\/h4>\n<p>Apply \u2192 Pass qualification(s) \u2192 Join a project \u2192 Complete tasks \u2192 Get paid<\/p>\n<h4>Effort estimate<\/h4>\n<p>Tasks for this project are estimated to take 20 hours to complete, depending on complexity. This is an estimate and not a schedule requirement; you choose when and how to work. Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<\/p>\n<h4>Compensation<\/h4>\n<p>Up to $50\/hr equivalent, depending on level and pace. Tasks are estimated at approximately 20 hours each; you set your own schedule.<\/p><\/p><p><\/p>
<h2 class="h5">Job description</h2>
<div class="t-break" data-jb-field="description">
<span>Please submit your CV in English and indicate your level of English proficiency.<br> Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems.<br> Participation is project-based, not permanent employment.<br>What this opportunity involves: We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks.<br> You'll create challenging tasks and evaluation criteria within realistic simulated environments: Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history Design tasks from intermediate states of these environments - craft the prompt, define what "solved" means, and ensure the task is solvable by an AI agent Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust What this is NOT: Not data labeling Not prompt engineering Not writing code from scratch - the agent writes most of the code; you guide and evaluate What we look for: 8+ years in software development Core stack: Python (FastAPI), JavaScript/TypeScript (React), Docker, Postgres, Kafka, Redis Experience writing tests (functional, integration) English proficiency - B2+ Why this is hard: Frontier models are already good at coding.<br> Creating a task that genuinely challenges the best models is non-trivial.<br> You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution.<br> Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<br> How it works Apply → Pass qualification(s) → Join a project → Complete tasks → Get paidEffort estimate Tasks for this project are estimated to take 30 hours to complete, depending on complexity.<br> This is an estimate and not a schedule requirement; you choose when and how to work.<br> Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<br> Compensation: Up to $150/hr equivalent , depending on level and pace.<br> Tasks are estimated at ~30 hours each; you set your own schedule.<br></span> </div>
<h2 class="h5">Job description</h2>
<div class="t-break" data-jb-field="description">
<p>We are seeking a dynamic and very motivated Sales Representative to join our team in the accounting sector. This role is pivotal in driving our sales initiatives and fostering strong relationships with clients. As a Sales Representative, you will be at the forefront of our efforts to expand our market presence and enhance customer satisfaction. You will have the opportunity to work in a collaborative environment where your contributions will directly impact our growth and success.</p><p>At our company, we pride ourselves on cultivating a culture of continuous learning and professional development. As a Sales Representative, you will receive comprehensive training and mentorship from experienced professionals in the field. This is not just a job; it’s a stepping stone for your career. You will have opportunities to hone your skills in negotiation, customer engagement, and market analysis, all of which are essential for your growth in the accounting industry.</p><p>Moreover, we believe in recognizing and rewarding talent. As you demonstrate your capabilities and contribute to our success, you will have access to various career advancement opportunities. Whether you aspire to move into a senior sales role or branch into management, we are committed to supporting your career progression every step of the way. Join us and be a part of a team that values innovation, teamwork, and excellence.</p><p><b>Responsibilities:</b></p><ol><li>Develop and maintain strong relationships with clients by understanding their needs and providing tailored solutions, utilizing CRM software to track interactions and progress.</li><li>Conduct market research to identify potential clients and generate leads, employing analytical tools to assess market trends and customer preferences.</li><li>Present and demonstrate our accounting services to prospective clients, using persuasive communication techniques to highlight the benefits and value of our offerings.</li><li>Negotiate contracts and close sales while ensuring compliance with company policies and industry regulations, leveraging negotiation skills to achieve favorable outcomes.</li><li>Collaborate with the marketing team to develop promotional materials and campaigns that resonate with target audiences, ensuring alignment with overall business objectives.</li><li>Provide exceptional customer service by addressing client inquiries and resolving issues promptly, utilizing problem-solving skills to enhance client satisfaction.</li><li>Prepare regular sales reports and forecasts to track performance against targets, employing data analysis to inform strategic decision-making.</li><li>Participate in team meetings and training sessions to share best practices and insights, fostering a culture of collaboration and continuous improvement.</li><li>Stay updated on industry trends and competitor activities, using this knowledge to adjust sales strategies and maintain a competitive edge.</li></ol> </div><h2 class="h5">Skills</h2>
<div data-jb-field="skills"><ul><li>Excellent communication skills are essential for building relationships and effectively conveying product value to clients.</li><li>Strong negotiation abilities to secure favorable terms and close sales successfully.</li><li>Proficiency in CRM software to manage client interactions and sales pipelines efficiently.</li><li>Analytical skills to conduct market research and assess customer needs accurately.</li><li>Ability to work collaboratively within a team to achieve common goals and share insights.</li><li>Customer service orientation to ensure client satisfaction and loyalty.</li><li>Adaptability to stay current with industry trends and adjust strategies accordingly.</li></ul></div>
<h2 class="h5">Job description</h2>
<div class="t-break" data-jb-field="description">
<span>Please submit your CV in English and indicate your level of English proficiency.<br> Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems.<br> Participation is project-based, not permanent employment.<br> What this opportunity involves We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks.<br> You'll create challenging tasks and evaluation criteria within realistic simulated environments: Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history Design tasks from intermediate states of these environments - craft the prompt, define what "solved" means, and ensure the task is solvable by an AI agent Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust What this is NOT Not data labeling Not prompt engineering Not writing code from scratch - the agent writes most of the code; you guide and evaluate What we look for 5+ years in software development Core stack: Python (FastAPI), JavaScript/TypeScript (React), Docker, Postgres, Kafka, Redis Experience writing tests (functional, integration) English proficiency - B2+ Why this is hard Frontier models are already good at coding.<br> Creating a task that genuinely challenges the best models is non-trivial.<br> You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution.<br> Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<br> How it works Apply → Pass qualification(s) → Join a project → Complete tasks → Get paid Effort estimate Tasks for this project are estimated to take 20 hours to complete, depending on complexity.<br> This is an estimate and not a schedule requirement; you choose when and how to work.<br> Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<br> Compensation Up to $50/hr equivalent , depending on level and pace.<br> Tasks are estimated at ~20 hours each; you set your own schedule.<br></span> </div>
<h2 class="h5">Job description</h2>
<div class="t-break" data-jb-field="description">
<span>We're looking for a highly technical AI Solutions Architect who is passionate about solving complex business problems using AI.<br> This is a hands-on role for someone who enjoys understanding how businesses operate, designing intelligent AI-powered solutions, and bringing those solutions to life through automation and software development.<br> You'll work across a variety of AI initiatives, from understanding business challenges and designing solution architectures to building scalable, production-ready systems that create measurable business value.<br> We're looking for someone who doesn't just execute ideas, but actively contributes them.<br> Someone who is naturally curious, constantly exploring new AI capabilities, and proactively identifying opportunities to improve how we work, solve problems, and create value.<br> What You'll Do ❏ Analyze business processes and identify opportunities where AI and automation can create measurable value.<br> ❏ Design end-to-end AI solution architectures, workflows, and automation strategies before implementation.<br> ❏ Translate business requirements into scalable technical solutions.<br> ❏ Recommend the most appropriate technologies and implementation approach for each solution.<br> ❏ Build AI-powered applications, intelligent assistants, and workflow automations.<br> ❏ Design and build solutions using no-code platforms, low-code tools, custom software, or a combination of them, depending on what best fits the problem.<br> ❏ Integrate AI solutions with business platforms using APIs, automation technologies, and custom software.<br> ❏ Rapidly prototype ideas and validate concepts before full implementation.<br> ❏ Continuously identify opportunities where AI and automation can improve business operations, productivity, and decision-making.<br> ❏ Proactively propose, prototype, and implement new AI initiatives that create measurable business value.<br> ❏ Challenge existing processes and recommend better ways of solving problems using AI and automation.<br> ❏ Research, evaluate, and implement emerging AI technologies where they create real business value.<br> ❏ Continuously optimize and improve existing AI systems and automations.<br> ❏ Document solution architectures, technical decisions, and implementation approaches.<br> ❏ Collaborate with cross-functional teams throughout discovery, design, implementation, and optimization.<br> You're Our Match If You.<br>.. ❏Have a proven track record of designing, architecting, and building AI-powered applications, intelligent systems, or workflow automations.<br> ❏ Can analyze business problems and translate them into practical AI solutions.<br> ❏ Think in systems and enjoy designing solution architectures before writing code.<br> ❏ Can evaluate multiple implementation approaches and select the one best suited for the problem.<br> ❏ Are proficient in Python and have a strong software engineering foundation.<br> ❏ Have experience integrating APIs and third-party platforms.<br> ❏ Have experience building workflow automations using no-code platforms, low-code tools, custom software, or a combination of them.<br> ❏ Understand modern AI technologies and know when and how to apply them effectively.<br> ❏ Can balance technical feasibility with business objectives.<br> ❏ Naturally identify opportunities where AI can simplify, improve, or transform the way people work.<br> ❏ Think like a product builder, constantly looking for new ideas worth exploring and validating.<br> ❏ Challenge assumptions and recommend better approaches rather than simply executing requirements.<br> ❏ Have strong analytical, problem-solving, and communication skills.<br> ❏ Are proactive, take ownership, and enjoy turning ideas into working solutions with minimal supervision.<br></span> </div>
<h2 class="h5">Job description</h2>
<div class="t-break" data-jb-field="description">
<span>Please submit your CV in English and indicate your level of English proficiency.<br> Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems.<br> Participation is project-based, not permanent employment.<br> What this opportunity involves We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks.<br> You'll create challenging tasks and evaluation criteria within realistic simulated environments: Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history Design tasks from intermediate states of these environments - craft the prompt, define what "solved" means, and ensure the task is solvable by an AI agent Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust What this is NOT Not data labeling Not prompt engineering Not writing code from scratch - the agent writes most of the code; you guide and evaluate What we look for 5+ years in software development Core stack: Python (FastAPI), JavaScript/TypeScript (React), Docker, Postgres, Kafka, Redis Experience writing tests (functional, integration) English proficiency - B2+ Why this is hard Frontier models are already good at coding.<br> Creating a task that genuinely challenges the best models is non-trivial.<br> You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution.<br> Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<br> How it works Apply → Pass qualification(s) → Join a project → Complete tasks → Get paid Effort estimate Tasks for this project are estimated to take 20 hours to complete, depending on complexity.<br> This is an estimate and not a schedule requirement; you choose when and how to work.<br> Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<br> Compensation Up to $50/hr equivalent , depending on level and pace.<br> Tasks are estimated at ~20 hours each; you set your own schedule.<br></span> </div>
<h2 class="h5">Job description</h2>
<div class="t-break" data-jb-field="description">
<span>Please submit your CV in English and indicate your level of English proficiency.<br> Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems.<br> Participation is project-based, not permanent employment.<br> What this opportunity involves We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks.<br> You'll create challenging tasks and evaluation criteria within realistic simulated environments: Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history Design tasks from intermediate states of these environments - craft the prompt, define what "solved" means, and ensure the task is solvable by an AI agent Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust What this is NOT Not data labeling Not prompt engineering Not writing code from scratch - the agent writes most of the code; you guide and evaluate What we look for 5+ years in software development Core stack: Python (FastAPI), JavaScript/TypeScript (React), Docker, Postgres, Kafka, Redis Experience writing tests (functional, integration) English proficiency - B2+ Why this is hard Frontier models are already good at coding.<br> Creating a task that genuinely challenges the best models is non-trivial.<br> You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution.<br> Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<br> How it works Apply → Pass qualification(s) → Join a project → Complete tasks → Get paid Effort estimate Tasks for this project are estimated to take 20 hours to complete, depending on complexity.<br> This is an estimate and not a schedule requirement; you choose when and how to work.<br> Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<br> Compensation Up to $50/hr equivalent , depending on level and pace.<br> Tasks are estimated at ~20 hours each; you set your own schedule.<br></span> </div>
<p><h4>Company description</h4>
<p>Why work for Accor?<br>
We are far more than a worldwide leader. We welcome you as you are and you can find a job and brand that matches your personality. We support you to grow and learn every day, making sure that work brings purpose to your life, so that during your journey with us, you can continue to explore Accor’s limitless possibilities.<br>
By joining Accor, every chapter of your story is yours to write and together we can imagine tomorrow's hospitality.<br>
Do what you love, care for the world, dare to challenge the status quo! #BELIMITLESS</p>
<h4>Job description</h4>
<h4>The role</h4>
<p>To supervise and control the operation of the building management system and monitor all plant activities of the headquarters building services.<br>
To provide accurate administration support to the engineering team.<br>
To interact with housekeeping and front office team to ensure our guests receive a high level of service.<br>
To provide service that is sincere, warm and enthusiastic whilst ensuring guest satisfaction.</p>
<h4>Key deliverables and responsibilities</h4>
<h4>Planning & organizing:</h4>
<ul>
<li>Is aware of the daily activities and has product knowledge of all the hotel facilities.</li>
<li>To be punctual on duty and ensure the same of your employees.</li>
<li>Assist the director of engineering and management in planning for the investment, project and replacement budget on a yearly basis.</li>
</ul>
<h4>Operations:</h4>
<ul>
<li>Displays good teamwork.</li>
<li>Recommends improvements in methods, systems and procedures to improve efficiency and maximize customer satisfaction.</li>
<li>Be willing to train and instruct other members of the staff by passing along knowledge and skills.</li>
<li>Perform as per operational standards manual (OSM) and in line with company’s values and core behaviour.</li>
<li>Attends training classes as per schedule.</li>
<li>Ensure all relevant checklists are completed as required.</li>
<li>To have a complete understanding of and to adhere to Mövenpick Hotels & Resorts policy relating to fire, hygiene, health and safety.</li>
<li>Schedules active maintenance with engineering team to ensure work is complete, and then communicates details with the engineering supervisor.</li>
<li>Ensure all work orders are acknowledged in the Dyna Win system.</li>
<li>To carry out any other reasonable duties and responsibilities as assigned.</li>
</ul>
<h4>Administration:</h4>
<ul>
<li>Provide and follow up on the engineering supervisor on planned preventive maintenance (PPM) schedules for implementation.</li>
<li>Analyze the behaviour of the plants with respect to the set point and set up the various control parameters.</li>
<li>Perform daily auditing of plant performance via the building management system (BMS) e.g. data logging, plant review, alarm manager status, etc.</li>
<li>Analyze reasons of failure of plants and provide specific diagnostic, problem analysis and specify the necessary remedial actions for the engineering supervisor's approval.</li>
<li>Use the historical logged data of various plants to review the system and propose improvements in order to enhance system efficiency and/or reduce operation cost as well as improve the services provided to guests.</li>
<li>Program the BMS and PPM system and other plants (e.g. standby power generator, diesel driven firefighting pump, fire alarm system, etc.) for automatic periodic testing.</li>
<li>Follow the grooming standards and maintain a friendly and cheerful disposition at all times.</li>
<li>To comply with all hotel rules and regulations as outlined in the staff handbook.</li>
<li>To anticipate the needs of the guest whenever possible, to enhance quality service and in turn enhance guest satisfaction.</li>
<li>Ensure all filing systems are accurate and up to date.</li>
<li>Provide administration support to the engineering team.</li>
<li>Assist with the completion of the department vacation plan and any associated HR paperwork.</li>
<li>To give full co-operation to any colleague requiring assistance in a prompt, caring and helpful manner. To be flexible in assisting in other areas of the hotel in response to business and guest needs.</li>
<li>Is familiar with all related company documentation and especially with the relevant operational standards manual for his/her field of responsibility.</li>
</ul>
<h4>Qualifications</h4>
<ul>
<li><strong>Education:</strong> Mechanic or electric college/university education (preferred but not required).</li>
<li><strong>Experience:</strong> Minimum 3 years.</li>
<li><strong>Other:</strong> Computer skills.</li>
</ul>
<h4>Additional information</h4>
<p>Mövenpick Hotels & Resorts is in the “moments” business. We’re intimately involved in important times in our guest’s lives. And you never know when a moment can be made. A simple smile in the lobby can create the positivity that turns a business trip into a new business celebration. An insider tip on the best way to spend a day can make an entire holiday. A romantic dinner for two can lead to a longer term partnership.<br>
We understand that this vision cannot be achieved without great people who create and support work environments designed to produce exceptional results.</p></p><p></p>
<section><p class="heading jdMain">Job Description</p><p class="heading">Roles & Responsibilities</p><div class="paragraph"><p>Do you want to love what you do at work? Do you want to make a difference, an impact, and transform peoples lives? Do you want to work with a team that believes in disrupting the normal, boring, and average? If yes, then this is the job you are looking for , webook.com is Saudi s #1 event ticketing and experience booking platform in terms of technology, features, agility, revenue serving some of the largest mega events in the Kingdom surpassing over 2 billion in sales.</p><p>Key Responsibilities</p><ul><li>Develop scalable, secure and functional mobile applications.</li><li>Work on the maintenance and improvement of existing mobile applications.</li><li>Utilizing React Native to design and develop UI components for mobile apps.</li><li>Writing effective, scalable, and reusable code to create interchangeable modules.</li><li>Monitor and optimize application performance and ensure responsiveness for a seamless user experience.</li><li>Develop ideas for new programs, products, or features by monitoring industry developments and trends.</li><li>Compile and analyze data, processes, and codes to troubleshoot problems and identify areas for improvement.</li><li>Work with the product and operations team to provide guidance and execution on the product roadmap to build highly-scalable applications.</li></ul><p>Key Skills</p><ul><li>Extensive experience in developing mobile applications using React Native.</li><li>Experience with iOS and/or Android application architecture and design.</li><li>Strong knowledge of Mobile development design patterns, components, tools, concepts, best practices, standards and systems of record is strongly desired.</li><li>Strong understanding of RESTful JSON web API design principles</li><li>Familiarity with native mobile development (Android/iOS) and related technologies (Java/Kotlin, Swift/Objective-C) is a plus.</li><li>Experience or exposure with back-end web engineering, its not your expertise, but any exposure is a great perspective.</li><li>Experience in using version control systems like Git</li><li>Experience with state management libraries (e.g.. MobX) and asynchronous programming.</li><li>Solid experience in documenting software solutions using diagrams and flowcharts.</li><li>Experience in release management, online digital analytics, conversion optimization test execution like a/b testing and multivariate testing, search engine optimization/marketing, social media marketing management, strategies and execution is an added advantage.</li><li>Strong multi-tasking, organization and time management skills to manage multiple projects, deadlines and priorities within budget requirements</li><li>Exceptional collaborative and interpersonal skills; dynamic flexible team player with the ability to lead or be led.</li><li>Excellent communication and customer service skills for brainstorming, critical thinking and discussions with the ability to take direction to meet creative and business needs with extreme attention to detail and consistency.</li><li>Knowledge of user centered design, usability engineering and user experience principles, guidelines, deliverables, methods and processes.</li><li>Ability and confidence to communicate and present to senior management</li><li>Strong written and oral communication skills</li><li>Self-motivated, organized and accountable</li><li>B.S. Computer Science or related field</li><li>Minimum 3-4 years of experience with React Native.</li></ul></div></section><section><p class="heading">Desired Candidate Profile</p><p class="paragraph"></p><p>Extensive experience in developing mobile applications using React Native.</p><p>Experience with iOS and/or Android application architecture and design.</p><p>Strong knowledge of Mobile development design patterns, components, tools, concepts, best practices, standards and systems of record is strongly desired.</p><p>Strong understanding of RESTful JSON web API design principles</p><p>Familiarity with native mobile development (Android/iOS) and related technologies (Java/Kotlin, Swift/Objective-C) is a plus.</p><p>Experience or exposure with back-end web engineering, its not your expertise, but any exposure is a great perspective.</p><p>Experience in using version control systems like Git</p><p>Experience with state management libraries (e.g.. MobX) and asynchronous programming.</p><p>Solid experience in documenting software solutions using diagrams and flowcharts.</p><p>Experience in release management, online digital analytics, conversion optimization test execution like a/b testing and multivariate testing, search engine optimization/marketing, social media marketing management, strategies and execution is an added advantage.</p><p>Strong multi-tasking, organization and time management skills to manage multiple projects, deadlines and priorities within budget requirements</p><p>Exceptional collaborative and interpersonal skills; dynamic flexible team player with the ability to lead or be led.</p><p>Excellent communication and customer service skills for brainstorming, critical thinking and discussions with the ability to take direction to meet creative and business needs with extreme attention to detail and consistency.</p><p>Knowledge of user centered design, usability engineering and user experience principles, guidelines, deliverables, methods and processes.</p><p>Ability and confidence to communicate and present to senior management</p><p>Strong written and oral communication skills</p><p>Self-motivated, organized and accountable</p><p>B.S. Computer Science or related field</p><p>Minimum 3-4 years of experience with React Native.</p><p></p></section>
<h2 class="h5">Job description</h2>
<div class="t-break" data-jb-field="description">
<span>Please submit your CV in English and indicate your level of English proficiency.<br> Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems.<br> Participation is project-based, not permanent employment.<br> What this opportunity involves We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks.<br> You'll create challenging tasks and evaluation criteria within realistic simulated environments: Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history Design tasks from intermediate states of these environments - craft the prompt, define what "solved" means, and ensure the task is solvable by an AI agent Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust What this is NOT Not data labeling Not prompt engineering Not writing code from scratch - the agent writes most of the code; you guide and evaluate What we look for 5+ years in software development Core stack: Python (FastAPI), JavaScript/TypeScript (React), Docker, Postgres, Kafka, Redis Experience writing tests (functional, integration) English proficiency - B2+ Why this is hard Frontier models are already good at coding.<br> Creating a task that genuinely challenges the best models is non-trivial.<br> You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution.<br> Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<br> How it works Apply → Pass qualification(s) → Join a project → Complete tasks → Get paid Effort estimate Tasks for this project are estimated to take 20 hours to complete, depending on complexity.<br> This is an estimate and not a schedule requirement; you choose when and how to work.<br> Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<br> Compensation Up to $50/hr equivalent , depending on level and pace.<br> Tasks are estimated at ~20 hours each; you set your own schedule.<br></span> </div>
<p><h4>Description<\/h4>\n<p>Please submit your CV in English and indicate your level of English proficiency.<\/p>\n<p>Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment.<\/p>\n<h4>What this opportunity involves<\/h4>\n<p>We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks.<\/p>\n<p>You'll create challenging tasks and evaluation criteria within realistic simulated environments:<\/p>\n<ul>\n<li>Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history<\/li>\n<li>Design tasks from intermediate states of these environments - craft the prompt, define what \"solved\" means, and ensure the task is solvable by an AI agent<\/li>\n<li>Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient<\/li>\n<li>Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust<\/li>\n<\/ul>\n<h4>What this is not<\/h4>\n<ul>\n<li>Not data labeling<\/li>\n<li>Not prompt engineering<\/li>\n<li>Not writing code from scratch - the agent writes most of the code; you guide and evaluate<\/li>\n<\/ul>\n<h4>What we look for<\/h4>\n<ul>\n<li>5+ years in software development<\/li>\n<li>Core stack: Python (FastAPI), JavaScript\/TypeScript (React), Docker, Postgres, Kafka, Redis<\/li>\n<li>Experience writing tests (functional, integration)<\/li>\n<li>English proficiency - B2+<\/li>\n<\/ul>\n<h4>Why this is hard<\/h4>\n<p>Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<\/p>\n<h4>How it works<\/h4>\n<p>Apply ? Pass qualification(s) ? Join a project ? Complete tasks ? Get paid<\/p>\n<h4>Effort estimate<\/h4>\n<p>Tasks for this project are estimated to take 20 hours to complete, depending on complexity. This is an estimate and not a schedule requirement; you choose when and how to work. Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<\/p>\n<h4>Compensation<\/h4>\n<p>Up to $50\/hr equivalent, depending on level and pace. Tasks are estimated at ~20 hours each; you set your own schedule.<\/p><\/p><p><\/p>
<p><h4>Description<\/h4>\n<p>Please submit your CV in English and indicate your level of English proficiency.<\/p>\n<p>Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment.<\/p>\n<h4>What this opportunity involves<\/h4>\n<p>We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks.<\/p>\n<p>You'll create challenging tasks and evaluation criteria within realistic simulated environments:<\/p>\n<ul>\n<li>Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history<\/li>\n<li>Design tasks from intermediate states of these environments - craft the prompt, define what \"solved\" means, and ensure the task is solvable by an AI agent<\/li>\n<li>Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient<\/li>\n<li>Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust<\/li>\n<\/ul>\n<h4>What this is not<\/h4>\n<ul>\n<li>Not data labeling<\/li>\n<li>Not prompt engineering<\/li>\n<li>Not writing code from scratch - the agent writes most of the code; you guide and evaluate<\/li>\n<\/ul>\n<h4>What we look for<\/h4>\n<ul>\n<li>5+ years in software development<\/li>\n<li>Core stack: Python (FastAPI), JavaScript\/TypeScript (React), Docker, Postgres, Kafka, Redis<\/li>\n<li>Experience writing tests (functional, integration)<\/li>\n<li>English proficiency - B2+<\/li>\n<\/ul>\n<h4>Why this is hard<\/h4>\n<p>Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<\/p>\n<h4>How it works<\/h4>\n<p>Apply ? Pass qualification(s) ? Join a project ? Complete tasks ? Get paid<\/p>\n<h4>Effort estimate<\/h4>\n<p>Tasks for this project are estimated to take 20 hours to complete, depending on complexity. This is an estimate and not a schedule requirement; you choose when and how to work. Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<\/p>\n<h4>Compensation<\/h4>\n<p>Up to $50\/hr equivalent, depending on level and pace. Tasks are estimated at approximately 20 hours each; you set your own schedule.<\/p><\/p><p><\/p>