Fire Fighting EngineerJobs
2666 Jobs Found
<section><p class="heading jdMain">Job Description</p><p class="heading">Roles & Responsibilities</p><div class="paragraph"><p>Develop, maintain, and enhance customer-facing web applications and digital platforms. </p>
<p>Build and optimize backend systems and frontend interfaces using modern development frameworks and technologies. </p>
<p>Create internal tools and dashboards to support design, operations, and business performance monitoring. </p>
<p>Collaborate with cross-functional teams including Product, Design, Operations, and Leadership to deliver impactful product features. </p>
<p>Contribute to the development and launch of new mobile applications and digital experiences. </p>
<p>Work with cloud infrastructure and deployment environments, including AWS services and related technologies. </p>
<p>Support the continuous improvement of system performance, scalability, and user experience. </p>
<p>Participate in technical discussions, architecture planning, and product decision-making processes. </p>
<p>Deliver high-quality code while balancing speed, scalability, and business priorities within a startup environment. </p>
<p>Communicate technical concepts effectively with both technical and non-technical stakeholders. </p>
<p>Proactively identify challenges, recommend solutions, and take ownership of product and technical outcomes. </p>
<p>Collaborate closely with Saudi-based teams and support cross-regional business initiatives.</p></div></section><section><p class="heading">Desired Candidate Profile</p><p class="paragraph"></p><p><b>Qualifications & Requirements:</b></p>
<ul>
<li>4+ years of professional software engineering experience within fast-paced or startup environments. </li>
<li>Strong full-stack development experience with frontend and backend technologies. </li>
<li>Experience building scalable web applications, internal tools, and customer-facing digital products. </li>
<li>Product-oriented mindset with the ability to work effectively in ambiguous and rapidly changing environments. </li>
<li>Strong problem-solving skills with the ability to work independently and take ownership of projects. </li>
<li>Fluent in both Arabic and English with strong communication and collaboration skills. </li>
<li>Based in Amman, Jordan, with the ability to work primarily from the office environment. </li>
<li>Experience with Ruby on Rails, Next.js, or similar modern frameworks is highly preferred. </li>
<li>Familiarity with AWS cloud infrastructure including EC2, S3, and RDS is considered an advantage. </li>
<li>Experience within e-commerce, marketplaces, PropTech, or technology-driven businesses is preferred. </li>
<li>Strong adaptability, learning agility, and willingness to contribute beyond core responsibilities in a startup setting. </li>
<li>Ability to explain technical concepts and decisions clearly to cross-functional teams and stakeholders.</li>
</ul><p></p></section>
<p><h4>Description<\/h4>\n<p>Please submit your CV in English and indicate your level of English proficiency.<\/p>\n<p>Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment.<\/p>\n<h4>What this opportunity involves<\/h4>\n<p>We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks.<\/p>\n<p>You'll create challenging tasks and evaluation criteria within realistic simulated environments:<\/p>\n<ul>\n<li>Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history<\/li>\n<li>Design tasks from intermediate states of these environments - craft the prompt, define what \"solved\" means, and ensure the task is solvable by an AI agent<\/li>\n<li>Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient<\/li>\n<li>Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust<\/li>\n<\/ul>\n<h4>What this is not<\/h4>\n<ul>\n<li>Not data labeling<\/li>\n<li>Not prompt engineering<\/li>\n<li>Not writing code from scratch - the agent writes most of the code; you guide and evaluate<\/li>\n<\/ul>\n<h4>What we look for<\/h4>\n<ul>\n<li>5+ years in software development<\/li>\n<li>Core stack: Python (FastAPI), JavaScript\/TypeScript (React), Docker, Postgres, Kafka, Redis<\/li>\n<li>Experience writing tests (functional, integration)<\/li>\n<li>English proficiency - B2+<\/li>\n<\/ul>\n<h4>Why this is hard<\/h4>\n<p>Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<\/p>\n<h4>How it works<\/h4>\n<p>Apply ? Pass qualification(s) ? Join a project ? Complete tasks ? Get paid<\/p>\n<h4>Effort estimate<\/h4>\n<p>Tasks for this project are estimated to take 20 hours to complete, depending on complexity. This is an estimate and not a schedule requirement; you choose when and how to work. Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<\/p>\n<h4>Compensation<\/h4>\n<p>Up to $50\/hr equivalent, depending on level and pace. Tasks are estimated at ~20 hours each; you set your own schedule.<\/p><\/p><p><\/p>
<p><h4>Description<\/h4>\n<p>Please submit your CV in English and indicate your level of English proficiency.<\/p>\n<p>Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment.<\/p>\n<h4>What this opportunity involves<\/h4>\n<p>We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks.<\/p>\n<p>You'll create challenging tasks and evaluation criteria within realistic simulated environments:<\/p>\n<ul>\n <li>Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history<\/li>\n <li>Design tasks from intermediate states of these environments - craft the prompt, define what \"solved\" means, and ensure the task is solvable by an AI agent<\/li>\n <li>Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient<\/li>\n <li>Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust<\/li>\n<\/ul>\n<h4>What this is not<\/h4>\n<ul>\n <li>Not data labeling<\/li>\n <li>Not prompt engineering<\/li>\n <li>Not writing code from scratch - the agent writes most of the code; you guide and evaluate<\/li>\n<\/ul>\n<h4>What we look for<\/h4>\n<ul>\n <li>5+ years in software development<\/li>\n <li>Core stack: Python (FastAPI), JavaScript\/TypeScript (React), Docker, Postgres, Kafka, Redis<\/li>\n <li>Experience writing tests (functional, integration)<\/li>\n <li>English proficiency - B2+<\/li>\n<\/ul>\n<h4>Why this is hard<\/h4>\n<p>Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<\/p>\n<h4>How it works<\/h4>\n<p>Apply? Pass qualification(s)? Join a project? Complete tasks? Get paid<\/p>\n<h4>Effort estimate<\/h4>\n<p>Tasks for this project are estimated to take 20 hours to complete, depending on complexity. This is an estimate and not a schedule requirement; you choose when and how to work. Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<\/p>\n<h4>Compensation<\/h4>\n<p>Up to $50\/hr equivalent, depending on level and pace. Tasks are estimated at approximately 20 hours each; you set your own schedule.<\/p><\/p><p><\/p>
<p><h4>Description<\/h4>\n<p>Please submit your CV in English and indicate your level of English proficiency.<\/p>\n<p>Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment.<\/p>\n<h4>What this opportunity involves<\/h4>\n<p>We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks.<\/p>\n<p>You'll create challenging tasks and evaluation criteria within realistic simulated environments:<\/p>\n<ul>\n <li>Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history<\/li>\n <li>Design tasks from intermediate states of these environments - craft the prompt, define what \"solved\" means, and ensure the task is solvable by an AI agent<\/li>\n <li>Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient<\/li>\n <li>Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust<\/li>\n<\/ul>\n<h4>What this is not<\/h4>\n<ul>\n <li>Not data labeling<\/li>\n <li>Not prompt engineering<\/li>\n <li>Not writing code from scratch - the agent writes most of the code; you guide and evaluate<\/li>\n<\/ul>\n<h4>What we look for<\/h4>\n<ul>\n <li>5+ years in software development<\/li>\n <li>Core stack: Python (FastAPI), JavaScript\/TypeScript (React), Docker, Postgres, Kafka, Redis<\/li>\n <li>Experience writing tests (functional, integration)<\/li>\n <li>English proficiency - B2+<\/li>\n<\/ul>\n<h4>Why this is hard<\/h4>\n<p>Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<\/p>\n<h4>How it works<\/h4>\n<p>Apply? Pass qualification(s)? Join a project? Complete tasks? Get paid<\/p>\n<h4>Effort estimate<\/h4>\n<p>Tasks for this project are estimated to take 20 hours to complete, depending on complexity. This is an estimate and not a schedule requirement; you choose when and how to work. Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<\/p>\n<h4>Compensation<\/h4>\n<p>Up to $50\/hr equivalent, depending on level and pace. Tasks are estimated at approximately 20 hours each; you set your own schedule.<\/p><\/p><p><\/p>
<p><h4>Description<\/h4>\n<p>Please submit your CV in English and indicate your level of English proficiency.<\/p>\n<p>Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment.<\/p>\n<h4>What this opportunity involves<\/h4>\n<p>We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks.<\/p>\n<p>You'll create challenging tasks and evaluation criteria within realistic simulated environments:<\/p>\n<ul>\n<li>Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history<\/li>\n<li>Design tasks from intermediate states of these environments - craft the prompt, define what \"solved\" means, and ensure the task is solvable by an AI agent<\/li>\n<li>Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient<\/li>\n<li>Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust<\/li>\n<\/ul>\n<h4>What this is not<\/h4>\n<ul>\n<li>Not data labeling<\/li>\n<li>Not prompt engineering<\/li>\n<li>Not writing code from scratch - the agent writes most of the code; you guide and evaluate<\/li>\n<\/ul>\n<h4>What we look for<\/h4>\n<ul>\n<li>5+ years in software development<\/li>\n<li>Core stack: Python (FastAPI), JavaScript\/TypeScript (React), Docker, Postgres, Kafka, Redis<\/li>\n<li>Experience writing tests (functional, integration)<\/li>\n<li>English proficiency - B2+<\/li>\n<\/ul>\n<h4>Why this is hard<\/h4>\n<p>Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<\/p>\n<h4>How it works<\/h4>\n<p>Apply ? Pass qualification(s) ? Join a project ? Complete tasks ? Get paid<\/p>\n<h4>Effort estimate<\/h4>\n<p>Tasks for this project are estimated to take 20 hours to complete, depending on complexity. This is an estimate and not a schedule requirement; you choose when and how to work. Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<\/p>\n<h4>Compensation<\/h4>\n<p>Up to $50\/hr equivalent, depending on level and pace. Tasks are estimated at ~20 hours each; you set your own schedule.<\/p><\/p><p><\/p>
<p><h4>Description<\/h4>\n<p>Please submit your CV in English and indicate your level of English proficiency.<\/p>\n<p>Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment.<\/p>\n<h4>What this opportunity involves<\/h4>\n<p>We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks.<\/p>\n<p>You'll create challenging tasks and evaluation criteria within realistic simulated environments:<\/p>\n<ul>\n<li>Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history<\/li>\n<li>Design tasks from intermediate states of these environments - craft the prompt, define what \"solved\" means, and ensure the task is solvable by an AI agent<\/li>\n<li>Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient<\/li>\n<li>Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust<\/li>\n<\/ul>\n<h4>What this is not<\/h4>\n<ul>\n<li>Not data labeling<\/li>\n<li>Not prompt engineering<\/li>\n<li>Not writing code from scratch - the agent writes most of the code; you guide and evaluate<\/li>\n<\/ul>\n<h4>What we look for<\/h4>\n<ul>\n<li>5+ years in software development<\/li>\n<li>Core stack: Python (FastAPI), JavaScript\/TypeScript (React), Docker, Postgres, Kafka, Redis<\/li>\n<li>Experience writing tests (functional, integration)<\/li>\n<li>English proficiency - B2+<\/li>\n<\/ul>\n<h4>Why this is hard<\/h4>\n<p>Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<\/p>\n<h4>How it works<\/h4>\n<p>Apply ? Pass qualification(s) ? Join a project ? Complete tasks ? Get paid<\/p>\n<h4>Effort estimate<\/h4>\n<p>Tasks for this project are estimated to take 20 hours to complete, depending on complexity. This is an estimate and not a schedule requirement; you choose when and how to work. Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<\/p>\n<h4>Compensation<\/h4>\n<p>Up to $50\/hr equivalent, depending on level and pace. Tasks are estimated at ~20 hours each; you set your own schedule.<\/p><\/p><p><\/p>
<p><h4>Description<\/h4>\n<p>Please submit your CV in English and indicate your level of English proficiency.<\/p>\n<p>Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment.<\/p>\n<h4>What this opportunity involves<\/h4>\n<p>We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks.<\/p>\n<p>You'll create challenging tasks and evaluation criteria within realistic simulated environments:<\/p>\n<ul>\n<li>Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history<\/li>\n<li>Design tasks from intermediate states of these environments - craft the prompt, define what \"solved\" means, and ensure the task is solvable by an AI agent<\/li>\n<li>Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient<\/li>\n<li>Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust<\/li>\n<\/ul>\n<h4>What this is not<\/h4>\n<ul>\n<li>Not data labeling<\/li>\n<li>Not prompt engineering<\/li>\n<li>Not writing code from scratch - the agent writes most of the code; you guide and evaluate<\/li>\n<\/ul>\n<h4>What we look for<\/h4>\n<ul>\n<li>5+ years in software development<\/li>\n<li>Core stack: Python (FastAPI), JavaScript\/TypeScript (React), Docker, Postgres, Kafka, Redis<\/li>\n<li>Experience writing tests (functional, integration)<\/li>\n<li>English proficiency - B2+<\/li>\n<\/ul>\n<h4>Why this is hard<\/h4>\n<p>Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<\/p>\n<h4>How it works<\/h4>\n<p>Apply ? Pass qualification(s) ? Join a project ? Complete tasks ? Get paid<\/p>\n<h4>Effort estimate<\/h4>\n<p>Tasks for this project are estimated to take 20 hours to complete, depending on complexity. This is an estimate and not a schedule requirement; you choose when and how to work. Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<\/p>\n<h4>Compensation<\/h4>\n<p>Up to $50\/hr equivalent, depending on level and pace. Tasks are estimated at approximately 20 hours each; you set your own schedule.<\/p><\/p><p><\/p>
<p><h4>Description<\/h4>\n<p>Please submit your CV in English and indicate your level of English proficiency.<\/p>\n<p>Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment.<\/p>\n<h4>What this opportunity involves<\/h4>\n<p>We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks.<\/p>\n<p>You'll create challenging tasks and evaluation criteria within realistic simulated environments:<\/p>\n<ul>\n <li>Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history<\/li>\n <li>Design tasks from intermediate states of these environments - craft the prompt, define what \"solved\" means, and ensure the task is solvable by an AI agent<\/li>\n <li>Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient<\/li>\n <li>Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust<\/li>\n<\/ul>\n<h4>What this is not<\/h4>\n<ul>\n <li>Not data labeling<\/li>\n <li>Not prompt engineering<\/li>\n <li>Not writing code from scratch - the agent writes most of the code; you guide and evaluate<\/li>\n<\/ul>\n<h4>What we look for<\/h4>\n<ul>\n <li>5+ years in software development<\/li>\n <li>Core stack: Python (FastAPI), JavaScript\/TypeScript (React), Docker, Postgres, Kafka, Redis<\/li>\n <li>Experience writing tests (functional, integration)<\/li>\n <li>English proficiency - B2+<\/li>\n<\/ul>\n<h4>Why this is hard<\/h4>\n<p>Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<\/p>\n<h4>How it works<\/h4>\n<p>Apply \u2192 Pass qualification(s) \u2192 Join a project \u2192 Complete tasks \u2192 Get paid<\/p>\n<h4>Effort estimate<\/h4>\n<p>Tasks for this project are estimated to take 20 hours to complete, depending on complexity. This is an estimate and not a schedule requirement; you choose when and how to work. Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<\/p>\n<h4>Compensation<\/h4>\n<p>Up to $50\/hr equivalent, depending on level and pace. Tasks are estimated at approximately 20 hours each; you set your own schedule.<\/p><\/p><p><\/p>
<p><h4>Description<\/h4>\n<p>Please submit your CV in English and indicate your level of English proficiency.<\/p>\n<p>Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment.<\/p>\n<h4>What this opportunity involves<\/h4>\n<p>We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks.<\/p>\n<p>You'll create challenging tasks and evaluation criteria within realistic simulated environments:<\/p>\n<ul>\n <li>Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history<\/li>\n <li>Design tasks from intermediate states of these environments - craft the prompt, define what \"solved\" means, and ensure the task is solvable by an AI agent<\/li>\n <li>Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient<\/li>\n <li>Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust<\/li>\n<\/ul>\n<h4>What this is not<\/h4>\n<ul>\n <li>Not data labeling<\/li>\n <li>Not prompt engineering<\/li>\n <li>Not writing code from scratch - the agent writes most of the code; you guide and evaluate<\/li>\n<\/ul>\n<h4>What we look for<\/h4>\n<ul>\n <li>5+ years in software development<\/li>\n <li>Core stack: Python (FastAPI), JavaScript\/TypeScript (React), Docker, Postgres, Kafka, Redis<\/li>\n <li>Experience writing tests (functional, integration)<\/li>\n <li>English proficiency - B2+<\/li>\n<\/ul>\n<h4>Why this is hard<\/h4>\n<p>Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<\/p>\n<h4>How it works<\/h4>\n<p>Apply \u2192 Pass qualification(s) \u2192 Join a project \u2192 Complete tasks \u2192 Get paid<\/p>\n<h4>Effort estimate<\/h4>\n<p>Tasks for this project are estimated to take 20 hours to complete, depending on complexity. This is an estimate and not a schedule requirement; you choose when and how to work. Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<\/p>\n<h4>Compensation<\/h4>\n<p>Up to $50\/hr equivalent, depending on level and pace. Tasks are estimated at approximately 20 hours each; you set your own schedule.<\/p><\/p><p><\/p>
<h2 class="h5">Job description</h2>
<div class="t-break" data-jb-field="description">
<span>Please submit your CV in English and indicate your level of English proficiency.<br> Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems.<br> Participation is project-based, not permanent employment.<br>What this opportunity involves: We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks.<br> You'll create challenging tasks and evaluation criteria within realistic simulated environments: Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history Design tasks from intermediate states of these environments - craft the prompt, define what "solved" means, and ensure the task is solvable by an AI agent Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust What this is NOT: Not data labeling Not prompt engineering Not writing code from scratch - the agent writes most of the code; you guide and evaluate What we look for: 8+ years in software development Core stack: Python (FastAPI), JavaScript/TypeScript (React), Docker, Postgres, Kafka, Redis Experience writing tests (functional, integration) English proficiency - B2+ Why this is hard: Frontier models are already good at coding.<br> Creating a task that genuinely challenges the best models is non-trivial.<br> You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution.<br> Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<br> How it works Apply → Pass qualification(s) → Join a project → Complete tasks → Get paidEffort estimate Tasks for this project are estimated to take 30 hours to complete, depending on complexity.<br> This is an estimate and not a schedule requirement; you choose when and how to work.<br> Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<br> Compensation: Up to $150/hr equivalent , depending on level and pace.<br> Tasks are estimated at ~30 hours each; you set your own schedule.<br></span> </div>
<h2 class="h5">Job description</h2>
<div class="t-break" data-jb-field="description">
<span>We are seeking an experienced Technical Architect to lead the design and implementation of technology solutions that support business objectives and align with enterprise architecture standards.<br> The ideal candidate will combine strong technical expertise with strategic thinking to deliver scalable, secure, and high-performing solutions while providing technical leadership across projects.<br> Key Responsibilities Design end-to-end technical architectures for enterprise systems, applications, and integrations.<br> Develop architecture blueprints, technical specifications, and solution design documentation.<br> Evaluate, recommend, and select appropriate technologies, frameworks, platforms, and tools based on business and technical requirements.<br> Provide technical leadership and guidance to development teams throughout the software development lifecycle.<br> Ensure solutions are scalable, secure, maintainable, and aligned with industry best practices and organizational standards.<br> Conduct architecture and code reviews to ensure compliance with technical standards and quality requirements.<br> Collaborate with business analysts, project managers, product owners, and key stakeholders to translate business requirements into effective technical solutions.<br> Identify technical risks, dependencies, and constraints, and develop mitigation strategies.<br> Design and oversee system integration approaches to ensure seamless interoperability between applications and platforms.<br> Stay up to date with emerging technologies, industry trends, and architectural best practices to drive innovation.<br> Mentor and support junior architects, developers, and technical team members by promoting knowledge sharing and continuous improvement.<br> Bachelor's degree in Computer Science, Information Technology, Software Engineering, or a related field.<br> Proven experience in solution or technical architecture, software design, and enterprise application development.<br> Strong understanding of cloud platforms, system integration, APIs, microservices, and enterprise architecture principles.<br> Experience with modern development frameworks, databases, and DevOps practices.<br> Excellent analytical, problem-solving, and communication skills.<br> Ability to collaborate effectively with cross-functional teams and influence technical decisions.<br> Relevant architecture or cloud certifications are considered an advantage.<br></span> </div>
<h2 class="h5">Job description</h2>
<div class="t-break" data-jb-field="description">
<span>Please submit your CV in English and indicate your level of English proficiency.<br> Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems.<br> Participation is project-based, not permanent employment.<br> What this opportunity involves We're building a dataset to evaluate AI coding agents — how well a model handles real-world developer tasks.<br> You'll create challenging tasks and evaluation criteria within realistic simulated environments: Build virtual companies following a high-level plan - codebase, infrastructure, and context (conversations, documentation, tickets) that form a realistic environment with development history Assemble and calibrate tasks from intermediate states of the virtual company: craft the prompt, define evaluation criteria, and ensure the task is solvable and the evaluation is fair Design tasks set in isolated environments - emulations of a developer's workstation: a Linux machine with development tools (terminal, CLI), MCP servers (repository, task tracker, messenger, documentation, etc.<br>), and a real web application codebase Write tests that accept all correct solutions and reject incorrect ones - neither too strict (breaking on valid approaches) nor too lenient (passing bad ones) Iterate with an AI agent on tests - verifying they catch real problems, don't miss bad solutions, and don't break on good ones Review code written by agents, analyze why an agent failed or succeeded, and design edge cases and adversarial scenarios Iterate based on feedback from expert QA reviewers who score your work on quality criteria What this is NOT Not data labeling Not prompt engineering Not writing code from scratch - the agent writes most of the code; you guide and evaluate A significant part of the work is done together with AI - it's very hard to create tasks that challenge frontier models without using frontier models.<br> What we look for This opportunity is a good fit for experienced developers, software engineers, and/or test automation specialists open to part-time, non-permanent projects.<br> Ideally, contributors will have: Degree in Computer Science, Software Engineering, or related fields 5+ years in software development, primarily Python (FastAPI, pytest, async/await, subprocess, file operations) Background in full-stack development, with experience building React-based interfaces (JavaScript/TypeScript) and robust back-end systems Experience writing tests (functional, integration — not just running them) Docker containers, and familiarity with infrastructure tools (Postgres, Kafka, Redis) CI/CD understanding (GitHub Actions as a user: triggers, labels, reading results) English proficiency - B2 You don't need to be an expert in every item, but you should be comfortable reading and reasoning about code across the stack.<br> Why this is hard Frontier models are already good at coding.<br> Creating a task that genuinely challenges the best models is non-trivial.<br> You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution.<br> Tasks have many valid solutions.<br> Writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<br> How it works Apply → Pass qualification(s) → Join a project → Complete tasks → Get paid Effort estimate Tasks for this project are estimated to take 20 hours to complete, depending on complexity.<br> This is an estimate and not a schedule requirement; you choose when and how to work.<br> Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<br> Compensation On this project, contributors can earn up to $40 per hour equivalent , depending on their level and pace of contribution.<br> Compensation varies across projects depending on scope, complexity, and required expertise.<br> Please note that other projects on the platform may offer different earning levels based on their requirements.<br></span> </div>
<h2 class="h5">Job description</h2>
<div class="t-break" data-jb-field="description">
<span>Please submit your CV in English and indicate your level of English proficiency.<br> Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems.<br> Participation is project-based, not permanent employment.<br> What this opportunity involves We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks.<br> You'll create challenging tasks and evaluation criteria within realistic simulated environments: Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history Design tasks from intermediate states of these environments - craft the prompt, define what "solved" means, and ensure the task is solvable by an AI agent Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust What this is NOT Not data labeling Not prompt engineering Not writing code from scratch - the agent writes most of the code; you guide and evaluate What we look for 5+ years in software development Core stack: Python (FastAPI), JavaScript/TypeScript (React), Docker, Postgres, Kafka, Redis Experience writing tests (functional, integration) English proficiency - B2+ Why this is hard Frontier models are already good at coding.<br> Creating a task that genuinely challenges the best models is non-trivial.<br> You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution.<br> Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<br> How it works Apply → Pass qualification(s) → Join a project → Complete tasks → Get paid Effort estimate Tasks for this project are estimated to take 20 hours to complete, depending on complexity.<br> This is an estimate and not a schedule requirement; you choose when and how to work.<br> Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<br> Compensation Up to $50/hr equivalent , depending on level and pace.<br> Tasks are estimated at ~20 hours each; you set your own schedule.<br></span> </div>
<h2 class="h5">Job description</h2>
<div class="t-break" data-jb-field="description">
<span>Job Summary Muller's Solutions is seeking an experienced Aconex Solution Architect (Technical Lead) to lead the architecture, implementation, deployment, and optimization of Oracle Aconex solutions.<br> The successful candidate will be responsible for designing scalable solutions, leading technical delivery, configuring the Aconex platform, and ensuring successful implementation throughout the project lifecycle.<br> This role requires a strong blend of solution architecture, technical leadership, stakeholder engagement, and hands-on Oracle Aconex expertise.<br> Key Responsibilities Lead the end-to-end implementation and deployment of Oracle Aconex solutions.<br> Design scalable Aconex solution architectures based on business and project requirements.<br> Gather and analyze business and technical requirements and translate them into solution designs.<br> Configure Oracle Aconex projects, document structures, workflows, forms, correspondence, user roles, permissions, and security settings.<br> Develop implementation strategies, deployment plans, and solution roadmaps.<br> Lead technical workshops and provide expert guidance to business and project stakeholders.<br> Design and oversee integrations between Oracle Aconex and enterprise applications using available APIs and integration methods.<br> Plan and execute document and data migration activities from legacy systems to Oracle Aconex.<br> Lead system configuration, testing (SIT/UAT), go-live activities, and post-implementation support.<br> Ensure compliance with document control standards, governance policies, and industry best practices.<br> Troubleshoot complex technical issues and recommend effective solutions.<br> Prepare solution architecture documents, technical specifications, deployment guides, and implementation documentation.<br> Provide technical leadership, mentoring, and support to project teams throughout the implementation lifecycle.<br> Required Qualifications Bachelor's degree in Information Technology, Computer Science, Engineering, Construction Management, or a related field.<br> 8–10+ years of relevant experience , including extensive hands-on experience with Oracle Aconex.<br> Proven experience leading Oracle Aconex implementation and deployment projects from initiation through go-live.<br> Strong expertise in Aconex configuration, document control, workflows, correspondence management, forms, and security.<br> Experience designing enterprise document management solutions and implementation strategies.<br> Experience integrating Oracle Aconex with enterprise systems and external applications.<br> Strong understanding of document management, project information management, and engineering/construction workflows.<br> Excellent analytical, problem-solving, communication, and stakeholder management skills.<br> Experience leading technical teams and coordinating implementation activities.<br> Preferred Qualifications Oracle Aconex certifications are highly desirable.<br> Experience in engineering, construction, infrastructure, real estate, or capital projects.<br> Knowledge of document control best practices and project information management.<br> Experience with APIs, system integrations, and data migration.<br> Familiarity with cloud-based enterprise applications and digital transformation initiatives.<br> Experience working in remote and cross-functional project environments.<br> Key Skills Oracle Aconex Architecture Oracle Aconex Implementation & Deployment Solution Design Technical Leadership Document Control & Document Management Workflow Configuration System Integration Data & Document Migration User & Security Configuration Stakeholder Management Technical Documentation Testing (SIT/UAT) & Go-Live Support Problem Solving & Troubleshooting</span> </div>
<h2 class="h5">Job description</h2>
<div class="t-break" data-jb-field="description">
<p>We are seeking a dynamic and very motivated Sales Representative to join our team in the accounting sector. This role is pivotal in driving our sales initiatives and fostering strong relationships with clients. As a Sales Representative, you will be at the forefront of our efforts to expand our market presence and enhance customer satisfaction. You will have the opportunity to work in a collaborative environment where your contributions will directly impact our growth and success.</p><p>At our company, we pride ourselves on cultivating a culture of continuous learning and professional development. As a Sales Representative, you will receive comprehensive training and mentorship from experienced professionals in the field. This is not just a job; it’s a stepping stone for your career. You will have opportunities to hone your skills in negotiation, customer engagement, and market analysis, all of which are essential for your growth in the accounting industry.</p><p>Moreover, we believe in recognizing and rewarding talent. As you demonstrate your capabilities and contribute to our success, you will have access to various career advancement opportunities. Whether you aspire to move into a senior sales role or branch into management, we are committed to supporting your career progression every step of the way. Join us and be a part of a team that values innovation, teamwork, and excellence.</p><p><b>Responsibilities:</b></p><ol><li>Develop and maintain strong relationships with clients by understanding their needs and providing tailored solutions, utilizing CRM software to track interactions and progress.</li><li>Conduct market research to identify potential clients and generate leads, employing analytical tools to assess market trends and customer preferences.</li><li>Present and demonstrate our accounting services to prospective clients, using persuasive communication techniques to highlight the benefits and value of our offerings.</li><li>Negotiate contracts and close sales while ensuring compliance with company policies and industry regulations, leveraging negotiation skills to achieve favorable outcomes.</li><li>Collaborate with the marketing team to develop promotional materials and campaigns that resonate with target audiences, ensuring alignment with overall business objectives.</li><li>Provide exceptional customer service by addressing client inquiries and resolving issues promptly, utilizing problem-solving skills to enhance client satisfaction.</li><li>Prepare regular sales reports and forecasts to track performance against targets, employing data analysis to inform strategic decision-making.</li><li>Participate in team meetings and training sessions to share best practices and insights, fostering a culture of collaboration and continuous improvement.</li><li>Stay updated on industry trends and competitor activities, using this knowledge to adjust sales strategies and maintain a competitive edge.</li></ol> </div><h2 class="h5">Skills</h2>
<div data-jb-field="skills"><ul><li>Excellent communication skills are essential for building relationships and effectively conveying product value to clients.</li><li>Strong negotiation abilities to secure favorable terms and close sales successfully.</li><li>Proficiency in CRM software to manage client interactions and sales pipelines efficiently.</li><li>Analytical skills to conduct market research and assess customer needs accurately.</li><li>Ability to work collaboratively within a team to achieve common goals and share insights.</li><li>Customer service orientation to ensure client satisfaction and loyalty.</li><li>Adaptability to stay current with industry trends and adjust strategies accordingly.</li></ul></div>
<section><p class="heading jdMain">Job Description</p><p class="heading">Roles & Responsibilities</p><div class="paragraph"><p>We believe in bold ideas, diverse perspectives, and the drive to transform knowledge into impact. Here, your curiosity fuels progress, your voice shapes innovation, and your ambition helps redefine what s possible within science and learning. We are a culture that obsesses over impact, challenges, and drives what s next to power infinite possibilities for our customers, colleagues and society at large.</p><p>About the Role:</p><p>You will work closely with Product Managers, Software Engineers, and other Quality Engineers to build and maintain high-quality solutions across the Content Creators ecosystem , which includes:</p><ul><li>Article Submission (Research Exchange / ReX)</li><li>Article Transfer (ReX Transfer)</li><li>License Signing and Open Access Payments (Author Services)</li><li>Open Access Fund Management (Wiley Open Access Dashboard (WOAD) and Oable)</li></ul><p>As part of an Agile Scrum team, you will contribute to the development of automated testing frameworks, support continuous quality improvements, and help ensure our platforms deliver an exceptional user experience.</p><p>How You Will Make an Impact</p><ul><li>Design, develop, and maintain automated test suites for web applications and services.</li><li>Collaborate with Product Owners, Developers, and Quality Engineers within Agile Scrum teams.</li><li>Define and implement effective testing strategies to improve product quality and reliability.</li><li>Partner with developers to determine the most efficient and scalable testing approaches.</li><li>Analyze business requirements and translate them into comprehensive test scenarios and test cases.</li><li>Support sprint planning, estimation, and delivery activities to achieve team objectives.</li><li>Contribute to quarterly Program Increment (PI) planning and cross-functional initiatives.</li><li>Perform exploratory and manual testing when required to validate new functionality and identify potential risks.</li><li>Champion quality best practices, including shift-left testing and continuous improvement initiatives.</li><li>Provide recommendations to enhance testing frameworks, development processes, and overall system quality.</li></ul><p>What We Look For</p><p>Required Qualifications</p><ul><li>Experience in both manual and automated software testing.</li><li>Strong knowledge of Java, including:</li><ul><li>Object-Oriented Programming (OOP)</li><li>Collections Framework</li><li>Multithreading fundamentals</li><li>Lambda expressions</li><li>Stream API</li></ul><li>Experience testing RESTful APIs and strong understanding of HTTP protocols.</li><li>Strong SQL skills with hands-on experience using PostgreSQL, including:</li><ul><li>Multi-table joins</li><li>Subqueries</li><li>Aggregations</li><li>Test data preparation and backend validation</li></ul><li>Proficiency with Git, including branching, merging, rebasing, and conflict resolution.</li><li>Understanding of software development lifecycles, testing methodologies, and quality assurance practices.</li><li>Passion for shift-left testing principles and continuous quality improvement.</li><li>Understanding of modern web application architecture and behavior.</li><li>Ability to effectively work with and interpret English-language technical documentation.</li></ul><p>Technical Environment</p><p>Our technology stack includes:</p><ul><li>Java 17</li><li>Selenide for UI test automation</li><li>RestAssured for API testing</li><li>JUnit and TestNG</li><li>TestContainers for integration testing</li><li>GitHub and GitHub Actions</li><li>Allure and Report Portal for test reporting</li><li>Microservices architecture</li><li>HTTP and Kafka-based communication</li><li>Kibana and Grafana for monitoring and observability</li></ul><p>Preferred Qualifications</p><ul><li>Knowledge of Spring and Spring Boot.</li><li>Experience working with SQL and NoSQL databases.</li><li>Understanding of event-driven architectures and messaging platforms such as Kafka or RabbitMQ.</li><li>Experience working with AWS cloud services and infrastructure.</li><li>Strong written and verbal English communication skills.</li></ul><p>We power infinite possibilities. For more than 200 years, we've transformed knowledge into discoveries that shape the world. Today, our global team of innovators, creators, and experts is driving what s next in science, education, and publishing creating impact that reaches everywhere. We're not just observers of progress. We're the ones accelerating scientific breakthroughs, advancing learning, and sparking innovation that redefines entire fields and improves lives. Here, your talent matters. Your ideas have room to grow. And your work creates breakthroughs that can change everything.</p><p>Wiley is an equal opportunity/affirmative action employer. We evaluate all qualified applicants and treat all qualified applicants and employees without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, disability, protected veteran status, genetic information, or based on any individual's status in any group or class protected by applicable federal, state or local laws. Wiley is also committed to providing reasonable accommodation to applicants and employees with disabilities. Applicants who require accommodation to participate in the job application process may contact EMAIL_ADDRESS for assistance.</p><p>We are proud that our workplace promotes continual learning and internal mobility. Our values support courageous teammates, needle movers, and learning champions all while striving to support the health and well-being of all employees. We offer meeting-free Friday afternoons allowing more time for heads down work and professional development, and through a robust body of employee programing we facilitate a wide range of opportunities to foster community, learn, and grow. We are committed to fair, transparent pay, and we strive to provide competitive compensation in addition to a comprehensive benefits package. It is anticipated that most qualified candidates will fall within the range, however the ultimate salary offered for this role may be higher or lower and will be set based on a variety of non-discriminatory factors, including but not limited to, geographic location, skills, and competencies. Wiley proactively displays target base pay range for United Kingdom, Canada and USA based roles. When applying, please attach your resume/CV to be considered. #LI-SC1</p></div></section><section><p class="heading">Desired Candidate Profile</p><p class="paragraph"></p><ul><li>Experience in both manual and automated software testing.</li><li>Strong knowledge of Java, including:<ul><li>Object-Oriented Programming (OOP)</li><li>Collections Framework</li><li>Multithreading fundamentals</li><li>Lambda expressions</li><li>Stream API</li></ul></li><li>Experience testing RESTful APIs and strong understanding of HTTP protocols.</li><li>Strong SQL skills with hands-on experience using PostgreSQL, including:<ul><li>Multi-table joins</li><li>Subqueries</li><li>Aggregations</li><li>Test data preparation and backend validation</li></ul></li><li>Proficiency with Git, including branching, merging, rebasing, and conflict resolution.</li><li>Understanding of software development lifecycles, testing methodologies, and quality assurance practices.</li><li>Passion for shift-left testing principles and continuous quality improvement.</li><li>Understanding of modern web application architecture and behavior.</li><li>Ability to effectively work with and interpret English-language technical documentation.</li></ul><p>Preferred Qualifications</p><ul><li>Knowledge of Spring and Spring Boot.</li><li>Experience working with SQL and NoSQL databases.</li><li>Understanding of event-driven architectures and messaging platforms such as Kafka or RabbitMQ.</li><li>Experience working with AWS cloud services and infrastructure.</li><li>Strong written and verbal English communication skills.</li></ul><p></p></section>
<h2 class="h5">Job description</h2>
<div class="t-break" data-jb-field="description">
<span>Please submit your CV in English and indicate your level of English proficiency.<br> Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems.<br> Participation is project-based, not permanent employment.<br> What this opportunity involves We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks.<br> You'll create challenging tasks and evaluation criteria within realistic simulated environments: Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history Design tasks from intermediate states of these environments - craft the prompt, define what "solved" means, and ensure the task is solvable by an AI agent Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust What this is NOT Not data labeling Not prompt engineering Not writing code from scratch - the agent writes most of the code; you guide and evaluate What we look for 5+ years in software development Core stack: Python (FastAPI), JavaScript/TypeScript (React), Docker, Postgres, Kafka, Redis Experience writing tests (functional, integration) English proficiency - B2+ Why this is hard Frontier models are already good at coding.<br> Creating a task that genuinely challenges the best models is non-trivial.<br> You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution.<br> Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<br> How it works Apply → Pass qualification(s) → Join a project → Complete tasks → Get paid Effort estimate Tasks for this project are estimated to take 20 hours to complete, depending on complexity.<br> This is an estimate and not a schedule requirement; you choose when and how to work.<br> Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<br> Compensation Up to $50/hr equivalent , depending on level and pace.<br> Tasks are estimated at ~20 hours each; you set your own schedule.<br></span> </div>
<section><p class="heading jdMain">Job Description</p><p class="heading">Roles & Responsibilities</p><div class="paragraph"><p>Do you want to love what you do at work? Do you want to make a difference, an impact, and transform peoples lives? Do you want to work with a team that believes in disrupting the normal, boring, and average? If yes, then this is the job you are looking for , webook.com is Saudi s #1 event ticketing and experience booking platform in terms of technology, features, agility, revenue serving some of the largest mega events in the Kingdom surpassing over 2 billion in sales.</p><p>Key Responsibilities</p><ul><li>Develop scalable, secure and functional mobile applications.</li><li>Work on the maintenance and improvement of existing mobile applications.</li><li>Utilizing React Native to design and develop UI components for mobile apps.</li><li>Writing effective, scalable, and reusable code to create interchangeable modules.</li><li>Monitor and optimize application performance and ensure responsiveness for a seamless user experience.</li><li>Develop ideas for new programs, products, or features by monitoring industry developments and trends.</li><li>Compile and analyze data, processes, and codes to troubleshoot problems and identify areas for improvement.</li><li>Work with the product and operations team to provide guidance and execution on the product roadmap to build highly-scalable applications.</li></ul><p>Key Skills</p><ul><li>Extensive experience in developing mobile applications using React Native.</li><li>Experience with iOS and/or Android application architecture and design.</li><li>Strong knowledge of Mobile development design patterns, components, tools, concepts, best practices, standards and systems of record is strongly desired.</li><li>Strong understanding of RESTful JSON web API design principles</li><li>Familiarity with native mobile development (Android/iOS) and related technologies (Java/Kotlin, Swift/Objective-C) is a plus.</li><li>Experience or exposure with back-end web engineering, its not your expertise, but any exposure is a great perspective.</li><li>Experience in using version control systems like Git</li><li>Experience with state management libraries (e.g.. MobX) and asynchronous programming.</li><li>Solid experience in documenting software solutions using diagrams and flowcharts.</li><li>Experience in release management, online digital analytics, conversion optimization test execution like a/b testing and multivariate testing, search engine optimization/marketing, social media marketing management, strategies and execution is an added advantage.</li><li>Strong multi-tasking, organization and time management skills to manage multiple projects, deadlines and priorities within budget requirements</li><li>Exceptional collaborative and interpersonal skills; dynamic flexible team player with the ability to lead or be led.</li><li>Excellent communication and customer service skills for brainstorming, critical thinking and discussions with the ability to take direction to meet creative and business needs with extreme attention to detail and consistency.</li><li>Knowledge of user centered design, usability engineering and user experience principles, guidelines, deliverables, methods and processes.</li><li>Ability and confidence to communicate and present to senior management</li><li>Strong written and oral communication skills</li><li>Self-motivated, organized and accountable</li><li>B.S. Computer Science or related field</li><li>Minimum 3-4 years of experience with React Native.</li></ul></div></section><section><p class="heading">Desired Candidate Profile</p><p class="paragraph"></p><p>Extensive experience in developing mobile applications using React Native.</p><p>Experience with iOS and/or Android application architecture and design.</p><p>Strong knowledge of Mobile development design patterns, components, tools, concepts, best practices, standards and systems of record is strongly desired.</p><p>Strong understanding of RESTful JSON web API design principles</p><p>Familiarity with native mobile development (Android/iOS) and related technologies (Java/Kotlin, Swift/Objective-C) is a plus.</p><p>Experience or exposure with back-end web engineering, its not your expertise, but any exposure is a great perspective.</p><p>Experience in using version control systems like Git</p><p>Experience with state management libraries (e.g.. MobX) and asynchronous programming.</p><p>Solid experience in documenting software solutions using diagrams and flowcharts.</p><p>Experience in release management, online digital analytics, conversion optimization test execution like a/b testing and multivariate testing, search engine optimization/marketing, social media marketing management, strategies and execution is an added advantage.</p><p>Strong multi-tasking, organization and time management skills to manage multiple projects, deadlines and priorities within budget requirements</p><p>Exceptional collaborative and interpersonal skills; dynamic flexible team player with the ability to lead or be led.</p><p>Excellent communication and customer service skills for brainstorming, critical thinking and discussions with the ability to take direction to meet creative and business needs with extreme attention to detail and consistency.</p><p>Knowledge of user centered design, usability engineering and user experience principles, guidelines, deliverables, methods and processes.</p><p>Ability and confidence to communicate and present to senior management</p><p>Strong written and oral communication skills</p><p>Self-motivated, organized and accountable</p><p>B.S. Computer Science or related field</p><p>Minimum 3-4 years of experience with React Native.</p><p></p></section>
<p><h4>Description<\/h4>\n<p>Please submit your CV in English and indicate your level of English proficiency.<\/p>\n<p>Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment.<\/p>\n<h4>What this opportunity involves<\/h4>\n<p>We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks.<\/p>\n<p>You'll create challenging tasks and evaluation criteria within realistic simulated environments:<\/p>\n<ul>\n<li>Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history<\/li>\n<li>Design tasks from intermediate states of these environments - craft the prompt, define what \"solved\" means, and ensure the task is solvable by an AI agent<\/li>\n<li>Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient<\/li>\n<li>Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust<\/li>\n<\/ul>\n<h4>What this is not<\/h4>\n<ul>\n<li>Not data labeling<\/li>\n<li>Not prompt engineering<\/li>\n<li>Not writing code from scratch - the agent writes most of the code; you guide and evaluate<\/li>\n<\/ul>\n<h4>What we look for<\/h4>\n<ul>\n<li>5+ years in software development<\/li>\n<li>Core stack: Python (FastAPI), JavaScript\/TypeScript (React), Docker, Postgres, Kafka, Redis<\/li>\n<li>Experience writing tests (functional, integration)<\/li>\n<li>English proficiency - B2+<\/li>\n<\/ul>\n<h4>Why this is hard<\/h4>\n<p>Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<\/p>\n<h4>How it works<\/h4>\n<p>Apply ? Pass qualification(s) ? Join a project ? Complete tasks ? Get paid<\/p>\n<h4>Effort estimate<\/h4>\n<p>Tasks for this project are estimated to take 20 hours to complete, depending on complexity. This is an estimate and not a schedule requirement; you choose when and how to work. Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<\/p>\n<h4>Compensation<\/h4>\n<p>Up to $50\/hr equivalent, depending on level and pace. Tasks are estimated at ~20 hours each; you set your own schedule.<\/p><\/p><p><\/p>
<p><h4>Description<\/h4>\n<p>Please submit your CV in English and indicate your level of English proficiency.<\/p>\n<p>Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment.<\/p>\n<h4>What this opportunity involves<\/h4>\n<p>We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks.<\/p>\n<p>You'll create challenging tasks and evaluation criteria within realistic simulated environments:<\/p>\n<ul>\n<li>Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history<\/li>\n<li>Design tasks from intermediate states of these environments - craft the prompt, define what \"solved\" means, and ensure the task is solvable by an AI agent<\/li>\n<li>Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient<\/li>\n<li>Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust<\/li>\n<\/ul>\n<h4>What this is not<\/h4>\n<ul>\n<li>Not data labeling<\/li>\n<li>Not prompt engineering<\/li>\n<li>Not writing code from scratch - the agent writes most of the code; you guide and evaluate<\/li>\n<\/ul>\n<h4>What we look for<\/h4>\n<ul>\n<li>5+ years in software development<\/li>\n<li>Core stack: Python (FastAPI), JavaScript\/TypeScript (React), Docker, Postgres, Kafka, Redis<\/li>\n<li>Experience writing tests (functional, integration)<\/li>\n<li>English proficiency - B2+<\/li>\n<\/ul>\n<h4>Why this is hard<\/h4>\n<p>Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds.<\/p>\n<h4>How it works<\/h4>\n<p>Apply ? Pass qualification(s) ? Join a project ? Complete tasks ? Get paid<\/p>\n<h4>Effort estimate<\/h4>\n<p>Tasks for this project are estimated to take 20 hours to complete, depending on complexity. This is an estimate and not a schedule requirement; you choose when and how to work. Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.<\/p>\n<h4>Compensation<\/h4>\n<p>Up to $50\/hr equivalent, depending on level and pace. Tasks are estimated at approximately 20 hours each; you set your own schedule.<\/p><\/p><p><\/p>