Writing.io Jobs

Find the best remote jobs. Answer a few questions and we'll deploy a powerful assistant to help you search, create alerts, and more.

1 What roles are you open to?

2 Experience level

3 Work style

Did you know? If memory is enabled, Writing.io can remember your job search preferences and help you to improve your resume, craft customized outreach and more.

Trainer Information Extraction Engineer at DEF CON

Extracts and validates structured facts from documents, cites source passages, reviews model-assisted outputs, and flags ambiguity for reliable AI systems.

Mid Remote Posted 1 day ago RemoteFirstJobs Product
What this role involves

ABOUT DEFCON AI

RESILIENCE IN THE FACE OF DISRUPTION. DEFCON AI is an insights company that leverages artificial intelligence, mathematical optimization, data analytics, and software engineering for resilient optimization of complex systems.

In today’s dynamically changing world, DEFCON AI’s technology aligns outcomes with operational goals, better decision making, and empowers customers to anticipate assess, and mitigate the impacts of disruptions.

About the Role

You’ll pull structured facts out of long, free-text documents, and keep every one of them tied to the exact passage it came from. On this platform, nothing gets asserted without a citation a human reviewer can check, and that traceability is yours to build and protect.

You’ll join the analytics and AI engineering team behind a system that ingests records from dozens of disparate sources, resolves them to the right entity, and shows a human reviewer the exact source text behind every answer. Operating within a secure government cloud environment, the platform depends on extraction a reviewer can trust precisely because it can be checked against the original document.

You’ll join an established extraction practice rather than building one from scratch. The approach is set and reviewed by a senior practitioner on the team. Your contribution is sustained, consistent output at volume against long, messy, narrative documents that were never written with extraction in mind, and the judgment to flag what doesn’t fit rather than force it.

This is a fully remote role, with a later start date than our most urgent openings. We’re sourcing now so the seat is filled when the work begins.

Key Responsibilities

  • Extract structured facts from long, free-text, narrative, and semi-structured documents

  • Maintain a traceable pointer from every extracted fact back to the exact source passage it came from

  • Validate that every cited passage actually supports the extracted fact, and treat a missing or unsupported citation as a defect

  • Distinguish a stated total from a derived sum in financial narratives, retaining the contributing items and the calculation for verification

  • Consume the recorded rule context supplied to model-assisted extraction, keeping cited facts and uncertainty separate from any validity decision

  • Apply and extend the team’s established extraction approach consistently across new document types

  • Handle ambiguous, contradictory, or incomplete source text without forcing a single reading; flag it rather than guessing

  • Work with entity resolution and retrieval teammates on what a usable extracted fact needs to carry downstream

Required Qualifications

  • 3+ years extracting structured facts or data from unstructured, narrative text in a production setting

  • Experience keeping extracted output traceable to its source

  • Strong Python, with hands-on NLP or information-extraction tooling, whether rule-based, statistical, or LLM-based

  • US Citizenship Required

  • Active US Secret clearance

Preferred Qualifications

  • Extraction experience in a government, investigative, legal, financial-crime, or intelligence-adjacent setting

  • LLM-based extraction with structured output schemas and validation against source text

  • Experience where extracted facts fed a downstream matching, scoring, or retrieval system

  • Active Top Secret clearance

What Success Looks Like

  • Facts extracted at volume, with every claim traceable to its source passage

  • Extraction that holds up when a subject-matter expert checks it against the original document

  • Sustained throughput across new document types without re-deriving the approach for each one

What We Offer

  • A fully remote, results-based environment

  • Competitive salary, bonus, and equity package

  • 100% employer paid, comprehensive health insurance including medical, dental, and vision for you and your family

  • Unlimited PTO, with your manager’s approval

  • Flexible work environment where you manage your work day

  • 14 weeks of fully-paid parental leave

Salary Range: $130,000—$160,000. This represents the typical salary range for this position based on experience, skills, and other factors.

We’re an Equal Opportunity Employer: You’ll receive consideration for employment without regard to race, sex, color, religion, sexual orientation, gender identity, national origin, protected veteran status, or on the basis of disability.

Applicant Data Disclosure

By submitting an application, you acknowledge that Defcon AI uses third-party service providers to facilitate its recruitment and hiring processes. These providers include applicant tracking systems, candidate verification platforms, and fraud detection tools (collectively, “Hiring Platforms”). Your application materials, including your résumé, cover letter, work samples, responses to application questions, and any other information you submit, may be transmitted to and processed by these Hiring Platforms for the following purposes:

  • Managing and administering your application throughout the hiring process;
  • Verifying the accuracy and authenticity of application materials, including by cross-referencing information you provide against publicly available sources and proprietary databases;
  • Identifying indicators of potentially fraudulent, fabricated, or materially misleading application content, including but not limited to discrepancies between submitted materials and publicly available professional profiles, geographic anomalies, and fabricated work histories.

Applications that are flagged through this process as containing indicators of fraud or material misrepresentation may be declined from further consideration. If you have questions about the status of your application or the evaluation process, please contactrecruiting@defconai.com.

Defcon AI requires its Hiring Platform providers to process your information solely for the purposes described above and in accordance with applicable law. Your information will be retained only for as long as necessary to fulfill these purposes and any applicable legal obligations, after which it will be deleted in accordance with Defcon AI’s data retention policies.

For more information about how your data is used, please refer to our Privacy Policy and Applicant Privacy Notice .

Read the full description
Trainer Information Extraction Engineer at DEF CON

Extracts and validates facts from complex documents, preserves source citations, and reviews model-assisted outputs for accuracy and uncertainty.

Mid Remote Posted 1 day ago RemoteFirstJobs Product
What this role involves

ABOUT DEFCON AI

RESILIENCE IN THE FACE OF DISRUPTION. DEFCON AI is an insights company that leverages artificial intelligence, mathematical optimization, data analytics, and software engineering for resilient optimization of complex systems.

In today’s dynamically changing world, DEFCON AI’s technology aligns outcomes with operational goals, better decision making, and empowers customers to anticipate assess, and mitigate the impacts of disruptions.

About the Role

You’ll pull structured facts out of long, free-text documents, and keep every one of them tied to the exact passage it came from. On this platform, nothing gets asserted without a citation a human reviewer can check, and that traceability is yours to build and protect.

You’ll join the analytics and AI engineering team behind a system that ingests records from dozens of disparate sources, resolves them to the right entity, and shows a human reviewer the exact source text behind every answer. Operating within a secure government cloud environment, the platform depends on extraction a reviewer can trust precisely because it can be checked against the original document.

You’ll join an established extraction practice rather than building one from scratch. The approach is set and reviewed by a senior practitioner on the team. Your contribution is sustained, consistent output at volume against long, messy, narrative documents that were never written with extraction in mind, and the judgment to flag what doesn’t fit rather than force it.

This is a fully remote role, with a later start date than our most urgent openings. We’re sourcing now so the seat is filled when the work begins.

Key Responsibilities

  • Extract structured facts from long, free-text, narrative, and semi-structured documents

  • Maintain a traceable pointer from every extracted fact back to the exact source passage it came from

  • Validate that every cited passage actually supports the extracted fact, and treat a missing or unsupported citation as a defect

  • Distinguish a stated total from a derived sum in financial narratives, retaining the contributing items and the calculation for verification

  • Consume the recorded rule context supplied to model-assisted extraction, keeping cited facts and uncertainty separate from any validity decision

  • Apply and extend the team’s established extraction approach consistently across new document types

  • Handle ambiguous, contradictory, or incomplete source text without forcing a single reading; flag it rather than guessing

  • Work with entity resolution and retrieval teammates on what a usable extracted fact needs to carry downstream

Required Qualifications

  • 3+ years extracting structured facts or data from unstructured, narrative text in a production setting

  • Experience keeping extracted output traceable to its source

  • Strong Python, with hands-on NLP or information-extraction tooling, whether rule-based, statistical, or LLM-based

  • US Citizenship Required

  • Active US Secret clearance

Preferred Qualifications

  • Extraction experience in a government, investigative, legal, financial-crime, or intelligence-adjacent setting

  • LLM-based extraction with structured output schemas and validation against source text

  • Experience where extracted facts fed a downstream matching, scoring, or retrieval system

  • Active Top Secret clearance

What Success Looks Like

  • Facts extracted at volume, with every claim traceable to its source passage

  • Extraction that holds up when a subject-matter expert checks it against the original document

  • Sustained throughput across new document types without re-deriving the approach for each one

What We Offer

  • A fully remote, results-based environment

  • Competitive salary, bonus, and equity package

  • 100% employer paid, comprehensive health insurance including medical, dental, and vision for you and your family

  • Unlimited PTO, with your manager’s approval

  • Flexible work environment where you manage your work day

  • 14 weeks of fully-paid parental leave

Salary Range: $130,000—$160,000. This represents the typical salary range for this position based on experience, skills, and other factors.

We’re an Equal Opportunity Employer: You’ll receive consideration for employment without regard to race, sex, color, religion, sexual orientation, gender identity, national origin, protected veteran status, or on the basis of disability.

Applicant Data Disclosure

By submitting an application, you acknowledge that Defcon AI uses third-party service providers to facilitate its recruitment and hiring processes. These providers include applicant tracking systems, candidate verification platforms, and fraud detection tools (collectively, “Hiring Platforms”). Your application materials, including your résumé, cover letter, work samples, responses to application questions, and any other information you submit, may be transmitted to and processed by these Hiring Platforms for the following purposes:

  • Managing and administering your application throughout the hiring process;
  • Verifying the accuracy and authenticity of application materials, including by cross-referencing information you provide against publicly available sources and proprietary databases;
  • Identifying indicators of potentially fraudulent, fabricated, or materially misleading application content, including but not limited to discrepancies between submitted materials and publicly available professional profiles, geographic anomalies, and fabricated work histories.

Applications that are flagged through this process as containing indicators of fraud or material misrepresentation may be declined from further consideration. If you have questions about the status of your application or the evaluation process, please contactrecruiting@defconai.com.

Defcon AI requires its Hiring Platform providers to process your information solely for the purposes described above and in accordance with applicable law. Your information will be retained only for as long as necessary to fulfill these purposes and any applicable legal obligations, after which it will be deleted in accordance with Defcon AI’s data retention policies.

For more information about how your data is used, please refer to our Privacy Policy and Applicant Privacy Notice .

Read the full description
Trainer AI Product Quality Specialist at Human Agency

Evaluates and improves AI product quality through testing, feedback, and responsible implementation of AI systems.

Mid Remote Posted 3 days ago RemoteFirstJobs Product
What this role involves

Location: US or Canada

Type: Independent Contractor (potential for conversion to full-time hire)

**A note before you read further**

The world is changing fast, and so is this role. We can’t fully tell you what this job looks like in six months. That’s true of every role at Human Agency right now. What we need from you, no matter what: accountability, curiosity, and a builder-first mindset. If that sounds exciting, keep reading. If it sounds intimidating but you’re still interested, let’s talk. If none of that sounds like you, that’s okay. This probably isn’t your spot.

About Human Agency

We’re scaling rapidly and have a growing pipeline of opportunities that demand exceptional talent across disciplines. Our mission is to bring on individuals, from creative producers to technical experts to entrepreneurial leaders, who can help us realize this next chapter of growth.

We are a company of doers. Leaders roll up their sleeves, teams work flat, and everyone contributes to what ships. Titles don’t insulate us from feedback or basics. We invite critique, learn quickly, and keep raising the bar. The best ideas win here, no matter where they come from, because clients trust us to deliver the strongest outcomes every time.

Our clients’ missions, products, and bottom lines are sacred. We immerse ourselves in their world, becoming stewards of their goals and partners in solving big problems. Every product, strategy, or asset we create must be both beautiful and functional; practical, usable, and designed for real-world impact.

Humans are our most valuable resource, and we only grow by hiring people who push us forward. Across strategy, engineering, design, data, and operations, we seek out teammates who raise the bar and make us better. Always hire up, never down.

We partner with organizations of all sizes to explore, design, and implement AI strategies that are secure, scalable, and human-centered. We believe AI should amplify human potential, not replace it, and we build with that conviction in every engagement. From advisory and tooling to implementation and education, we meet clients where they are and help them integrate AI in ways that align with their mission and values. Our goal is to empower teams to work smarter, move faster, and unlock new possibilities through thoughtful, responsible innovation.

And through it all, we lead with purpose, love, and adventure. We do meaningful work with people we care about, and we make the ride an adventure worth taking. Because at Human Agency, who we are and how we work are one and the same.

The Opportunity

Our senior team is shipping fast. Too fast to spend cycles on bugs, friction points, and the steady accumulation of product rough edges that slow users down and erode trust. That’s where you come in. This role exists because we need someone with sharp product instincts who can triage what actually matters to users, resolve issues quickly, and exercise real judgment about what to fix, what to escalate, and what to ignore.

This is not a ticket-processing job. Anyone can read feedback and route it somewhere. We need someone who reads feedback, understands what users are actually trying to do, and can tell the difference between a bug that’s blocking real work and one that’s noise. You’ll be the person who keeps product quality tight so leadership can stay focused on building what’s next.

This role can grow well beyond where it starts. Strong judgment and consistent output can open the door to more: building products directly with us, and eventually client-facing solutions work, depending on performance, business need, and demonstrated ability. We hire for trajectory, not just task completion, but that trajectory has to be earned and the timeline isn’t fixed.

What You’ll Own

  • Bug triage and resolution: Work through product issues reported by users with speed and precision. You’ll regularly work in the codebase, submit fixes, and validate them before release.
  • User feedback interpretation: Read between the lines of what users report to identify the real problem, not just the surface complaint, and prioritize accordingly.
  • Product quality monitoring: Maintain a running pulse on the user experience across our products, flagging patterns and systemic issues before they compound. You’ll own testing your changes, checking relevant edge cases, and confirming the original reported workflow works before considering an issue resolved.
  • AI-assisted workflows: Use Claude Code and other AI tools to move faster and resolve more, without sacrificing the judgment that makes the output actually useful.
  • Cross-functional communication: Keep relevant teammates informed with crisp, specific updates (what’s broken, what’s fixed, what still needs attention).
  • Documentation: Leave things better than you found them by documenting fixes, edge cases, and recurring issues in a way that builds institutional knowledge over time.

Who You Are

Experience & Skills

  • Hands-on experience working on software products in a QA, product, or technical support role. You know what good looks like and can spot when something falls short.
  • You have built or experimented with Claude Code and understand how to apply AI tooling to accelerate real work, not just demo it.
  • You have a track record of resolving product issues quickly and independently, without needing someone to hold your hand through every step.
  • You can read user feedback and extract signal from noise. You understand enough about UX to know when a complaint reflects a real problem and when it doesn’t.
  • You communicate clearly in writing and know how to escalate an issue in a way that gives product leads exactly what they need to act.
  • You’re comfortable working without a detailed process in place. You can create your own system for managing volume and staying on top of what matters.

Mindset & Traits

  • You move fast and hate letting things sit. A backlog of unresolved issues bothers you at a gut level.
  • You have real product taste. You’ve spent enough time thinking about user experience that good and bad feel obvious to you, not just theoretical.
  • You take ownership. When something is your responsibility, it gets done. You don’t wait for reminders or check-ins.
  • You’re honest about what you know and what you don’t, and you’d rather ask a sharp question than guess wrong and waste everyone’s time.
  • You see this as a starting point, not a ceiling. You want to earn more responsibility and are motivated by what this role could grow into.

Compensation & Logistics

  • This is a contract role with an hourly rate commensurate with experience. We’re targeting $15-30/hr.
  • Work is fully remote; US or Canada-based candidates only.
  • Hours are flexible but output is not. You’ll be expected to move quickly and stay responsive during active sprints.
  • We’re looking to bring great candidates on immediately.

Equal Opportunity Commitment

Human Agency is an Equal Opportunity Employer. We value diverse perspectives and are committed to building inclusive, high-performing teams where everyone can do their best work.

Read the full description
Trainer AI Product Quality Specialist at Human Agency

Evaluates and improves AI product quality through testing, feedback, and responsible implementation of AI systems.

Mid Remote Posted 3 days ago RemoteFirstJobs Product
What this role involves

Location: US or Canada

Type: Independent Contractor (potential for conversion to full-time hire)

**A note before you read further**

The world is changing fast, and so is this role. We can’t fully tell you what this job looks like in six months. That’s true of every role at Human Agency right now. What we need from you, no matter what: accountability, curiosity, and a builder-first mindset. If that sounds exciting, keep reading. If it sounds intimidating but you’re still interested, let’s talk. If none of that sounds like you, that’s okay. This probably isn’t your spot.

About Human Agency

We’re scaling rapidly and have a growing pipeline of opportunities that demand exceptional talent across disciplines. Our mission is to bring on individuals, from creative producers to technical experts to entrepreneurial leaders, who can help us realize this next chapter of growth.

We are a company of doers. Leaders roll up their sleeves, teams work flat, and everyone contributes to what ships. Titles don’t insulate us from feedback or basics. We invite critique, learn quickly, and keep raising the bar. The best ideas win here, no matter where they come from, because clients trust us to deliver the strongest outcomes every time.

Our clients’ missions, products, and bottom lines are sacred. We immerse ourselves in their world, becoming stewards of their goals and partners in solving big problems. Every product, strategy, or asset we create must be both beautiful and functional; practical, usable, and designed for real-world impact.

Humans are our most valuable resource, and we only grow by hiring people who push us forward. Across strategy, engineering, design, data, and operations, we seek out teammates who raise the bar and make us better. Always hire up, never down.

We partner with organizations of all sizes to explore, design, and implement AI strategies that are secure, scalable, and human-centered. We believe AI should amplify human potential, not replace it, and we build with that conviction in every engagement. From advisory and tooling to implementation and education, we meet clients where they are and help them integrate AI in ways that align with their mission and values. Our goal is to empower teams to work smarter, move faster, and unlock new possibilities through thoughtful, responsible innovation.

And through it all, we lead with purpose, love, and adventure. We do meaningful work with people we care about, and we make the ride an adventure worth taking. Because at Human Agency, who we are and how we work are one and the same.

The Opportunity

Our senior team is shipping fast. Too fast to spend cycles on bugs, friction points, and the steady accumulation of product rough edges that slow users down and erode trust. That’s where you come in. This role exists because we need someone with sharp product instincts who can triage what actually matters to users, resolve issues quickly, and exercise real judgment about what to fix, what to escalate, and what to ignore.

This is not a ticket-processing job. Anyone can read feedback and route it somewhere. We need someone who reads feedback, understands what users are actually trying to do, and can tell the difference between a bug that’s blocking real work and one that’s noise. You’ll be the person who keeps product quality tight so leadership can stay focused on building what’s next.

This role can grow well beyond where it starts. Strong judgment and consistent output can open the door to more: building products directly with us, and eventually client-facing solutions work, depending on performance, business need, and demonstrated ability. We hire for trajectory, not just task completion, but that trajectory has to be earned and the timeline isn’t fixed.

What You’ll Own

  • Bug triage and resolution: Work through product issues reported by users with speed and precision. You’ll regularly work in the codebase, submit fixes, and validate them before release.
  • User feedback interpretation: Read between the lines of what users report to identify the real problem, not just the surface complaint, and prioritize accordingly.
  • Product quality monitoring: Maintain a running pulse on the user experience across our products, flagging patterns and systemic issues before they compound. You’ll own testing your changes, checking relevant edge cases, and confirming the original reported workflow works before considering an issue resolved.
  • AI-assisted workflows: Use Claude Code and other AI tools to move faster and resolve more, without sacrificing the judgment that makes the output actually useful.
  • Cross-functional communication: Keep relevant teammates informed with crisp, specific updates (what’s broken, what’s fixed, what still needs attention).
  • Documentation: Leave things better than you found them by documenting fixes, edge cases, and recurring issues in a way that builds institutional knowledge over time.

Who You Are

Experience & Skills

  • Hands-on experience working on software products in a QA, product, or technical support role. You know what good looks like and can spot when something falls short.
  • You have built or experimented with Claude Code and understand how to apply AI tooling to accelerate real work, not just demo it.
  • You have a track record of resolving product issues quickly and independently, without needing someone to hold your hand through every step.
  • You can read user feedback and extract signal from noise. You understand enough about UX to know when a complaint reflects a real problem and when it doesn’t.
  • You communicate clearly in writing and know how to escalate an issue in a way that gives product leads exactly what they need to act.
  • You’re comfortable working without a detailed process in place. You can create your own system for managing volume and staying on top of what matters.

Mindset & Traits

  • You move fast and hate letting things sit. A backlog of unresolved issues bothers you at a gut level.
  • You have real product taste. You’ve spent enough time thinking about user experience that good and bad feel obvious to you, not just theoretical.
  • You take ownership. When something is your responsibility, it gets done. You don’t wait for reminders or check-ins.
  • You’re honest about what you know and what you don’t, and you’d rather ask a sharp question than guess wrong and waste everyone’s time.
  • You see this as a starting point, not a ceiling. You want to earn more responsibility and are motivated by what this role could grow into.

Compensation & Logistics

  • This is a contract role with an hourly rate commensurate with experience. We’re targeting $20-30/hr.
  • Work is fully remote; US or Canada-based candidates only.
  • Hours are flexible but output is not. You’ll be expected to move quickly and stay responsive during active sprints.
  • We’re looking to bring great candidates on immediately.

Equal Opportunity Commitment

Human Agency is an Equal Opportunity Employer. We value diverse perspectives and are committed to building inclusive, high-performing teams where everyone can do their best work.

Read the full description
Trainer AI Product Quality Specialist at Human Agency

Evaluates and improves AI product quality through testing, feedback, and human-centered assessment of model outputs.

Mid Remote Posted 3 days ago RemoteFirstJobs Product
What this role involves

Location: US or Canada

Type: Independent Contractor (potential for conversion to full-time hire)

**A note before you read further**

The world is changing fast, and so is this role. We can’t fully tell you what this job looks like in six months. That’s true of every role at Human Agency right now. What we need from you, no matter what: accountability, curiosity, and a builder-first mindset. If that sounds exciting, keep reading. If it sounds intimidating but you’re still interested, let’s talk. If none of that sounds like you, that’s okay. This probably isn’t your spot.

About Human Agency

We’re scaling rapidly and have a growing pipeline of opportunities that demand exceptional talent across disciplines. Our mission is to bring on individuals, from creative producers to technical experts to entrepreneurial leaders, who can help us realize this next chapter of growth.

We are a company of doers. Leaders roll up their sleeves, teams work flat, and everyone contributes to what ships. Titles don’t insulate us from feedback or basics. We invite critique, learn quickly, and keep raising the bar. The best ideas win here, no matter where they come from, because clients trust us to deliver the strongest outcomes every time.

Our clients’ missions, products, and bottom lines are sacred. We immerse ourselves in their world, becoming stewards of their goals and partners in solving big problems. Every product, strategy, or asset we create must be both beautiful and functional; practical, usable, and designed for real-world impact.

Humans are our most valuable resource, and we only grow by hiring people who push us forward. Across strategy, engineering, design, data, and operations, we seek out teammates who raise the bar and make us better. Always hire up, never down.

We partner with organizations of all sizes to explore, design, and implement AI strategies that are secure, scalable, and human-centered. We believe AI should amplify human potential, not replace it, and we build with that conviction in every engagement. From advisory and tooling to implementation and education, we meet clients where they are and help them integrate AI in ways that align with their mission and values. Our goal is to empower teams to work smarter, move faster, and unlock new possibilities through thoughtful, responsible innovation.

And through it all, we lead with purpose, love, and adventure. We do meaningful work with people we care about, and we make the ride an adventure worth taking. Because at Human Agency, who we are and how we work are one and the same.

The Opportunity

Our senior team is shipping fast. Too fast to spend cycles on bugs, friction points, and the steady accumulation of product rough edges that slow users down and erode trust. That’s where you come in. This role exists because we need someone with sharp product instincts who can triage what actually matters to users, resolve issues quickly, and exercise real judgment about what to fix, what to escalate, and what to ignore.

This is not a ticket-processing job. Anyone can read feedback and route it somewhere. We need someone who reads feedback, understands what users are actually trying to do, and can tell the difference between a bug that’s blocking real work and one that’s noise. You’ll be the person who keeps product quality tight so leadership can stay focused on building what’s next.

This role can grow well beyond where it starts. Strong judgment and consistent output can open the door to more: building products directly with us, and eventually client-facing solutions work, depending on performance, business need, and demonstrated ability. We hire for trajectory, not just task completion, but that trajectory has to be earned and the timeline isn’t fixed.

What You’ll Own

  • Bug triage and resolution: Work through product issues reported by users with speed and precision. You’ll regularly work in the codebase, submit fixes, and validate them before release.
  • User feedback interpretation: Read between the lines of what users report to identify the real problem, not just the surface complaint, and prioritize accordingly.
  • Product quality monitoring: Maintain a running pulse on the user experience across our products, flagging patterns and systemic issues before they compound. You’ll own testing your changes, checking relevant edge cases, and confirming the original reported workflow works before considering an issue resolved.
  • AI-assisted workflows: Use Claude Code and other AI tools to move faster and resolve more, without sacrificing the judgment that makes the output actually useful.
  • Cross-functional communication: Keep relevant teammates informed with crisp, specific updates (what’s broken, what’s fixed, what still needs attention).
  • Documentation: Leave things better than you found them by documenting fixes, edge cases, and recurring issues in a way that builds institutional knowledge over time.

Who You Are

Experience & Skills

  • Hands-on experience working on software products in a QA, product, or technical support role. You know what good looks like and can spot when something falls short.
  • You have built or experimented with Claude Code and understand how to apply AI tooling to accelerate real work, not just demo it.
  • You have a track record of resolving product issues quickly and independently, without needing someone to hold your hand through every step.
  • You can read user feedback and extract signal from noise. You understand enough about UX to know when a complaint reflects a real problem and when it doesn’t.
  • You communicate clearly in writing and know how to escalate an issue in a way that gives product leads exactly what they need to act.
  • You’re comfortable working without a detailed process in place. You can create your own system for managing volume and staying on top of what matters.

Mindset & Traits

  • You move fast and hate letting things sit. A backlog of unresolved issues bothers you at a gut level.
  • You have real product taste. You’ve spent enough time thinking about user experience that good and bad feel obvious to you, not just theoretical.
  • You take ownership. When something is your responsibility, it gets done. You don’t wait for reminders or check-ins.
  • You’re honest about what you know and what you don’t, and you’d rather ask a sharp question than guess wrong and waste everyone’s time.
  • You see this as a starting point, not a ceiling. You want to earn more responsibility and are motivated by what this role could grow into.

Compensation & Logistics

  • This is a contract role with an hourly rate commensurate with experience. We’re targeting $15-30/hr.
  • Work is fully remote; US or Canada-based candidates only.
  • Hours are flexible but output is not. You’ll be expected to move quickly and stay responsive during active sprints.
  • We’re looking to bring great candidates on immediately.

Equal Opportunity Commitment

Human Agency is an Equal Opportunity Employer. We value diverse perspectives and are committed to building inclusive, high-performing teams where everyone can do their best work.

Read the full description
Trainer AI Product Quality Specialist at Human Agency

Evaluates and improves AI product quality through testing, feedback, and responsible implementation of AI tools.

Mid Remote Posted 3 days ago RemoteFirstJobs Product
What this role involves

Location: US or Canada

Type: Independent Contractor (potential for conversion to full-time hire)

**A note before you read further**

The world is changing fast, and so is this role. We can’t fully tell you what this job looks like in six months. That’s true of every role at Human Agency right now. What we need from you, no matter what: accountability, curiosity, and a builder-first mindset. If that sounds exciting, keep reading. If it sounds intimidating but you’re still interested, let’s talk. If none of that sounds like you, that’s okay. This probably isn’t your spot.

About Human Agency

We’re scaling rapidly and have a growing pipeline of opportunities that demand exceptional talent across disciplines. Our mission is to bring on individuals, from creative producers to technical experts to entrepreneurial leaders, who can help us realize this next chapter of growth.

We are a company of doers. Leaders roll up their sleeves, teams work flat, and everyone contributes to what ships. Titles don’t insulate us from feedback or basics. We invite critique, learn quickly, and keep raising the bar. The best ideas win here, no matter where they come from, because clients trust us to deliver the strongest outcomes every time.

Our clients’ missions, products, and bottom lines are sacred. We immerse ourselves in their world, becoming stewards of their goals and partners in solving big problems. Every product, strategy, or asset we create must be both beautiful and functional; practical, usable, and designed for real-world impact.

Humans are our most valuable resource, and we only grow by hiring people who push us forward. Across strategy, engineering, design, data, and operations, we seek out teammates who raise the bar and make us better. Always hire up, never down.

We partner with organizations of all sizes to explore, design, and implement AI strategies that are secure, scalable, and human-centered. We believe AI should amplify human potential, not replace it, and we build with that conviction in every engagement. From advisory and tooling to implementation and education, we meet clients where they are and help them integrate AI in ways that align with their mission and values. Our goal is to empower teams to work smarter, move faster, and unlock new possibilities through thoughtful, responsible innovation.

And through it all, we lead with purpose, love, and adventure. We do meaningful work with people we care about, and we make the ride an adventure worth taking. Because at Human Agency, who we are and how we work are one and the same.

The Opportunity

Our senior team is shipping fast. Too fast to spend cycles on bugs, friction points, and the steady accumulation of product rough edges that slow users down and erode trust. That’s where you come in. This role exists because we need someone with sharp product instincts who can triage what actually matters to users, resolve issues quickly, and exercise real judgment about what to fix, what to escalate, and what to ignore.

This is not a ticket-processing job. Anyone can read feedback and route it somewhere. We need someone who reads feedback, understands what users are actually trying to do, and can tell the difference between a bug that’s blocking real work and one that’s noise. You’ll be the person who keeps product quality tight so leadership can stay focused on building what’s next.

This role can grow well beyond where it starts. Strong judgment and consistent output can open the door to more: building products directly with us, and eventually client-facing solutions work, depending on performance, business need, and demonstrated ability. We hire for trajectory, not just task completion, but that trajectory has to be earned and the timeline isn’t fixed.

What You’ll Own

  • Bug triage and resolution: Work through product issues reported by users with speed and precision. You’ll regularly work in the codebase, submit fixes, and validate them before release.
  • User feedback interpretation: Read between the lines of what users report to identify the real problem, not just the surface complaint, and prioritize accordingly.
  • Product quality monitoring: Maintain a running pulse on the user experience across our products, flagging patterns and systemic issues before they compound. You’ll own testing your changes, checking relevant edge cases, and confirming the original reported workflow works before considering an issue resolved.
  • AI-assisted workflows: Use Claude Code and other AI tools to move faster and resolve more, without sacrificing the judgment that makes the output actually useful.
  • Cross-functional communication: Keep relevant teammates informed with crisp, specific updates (what’s broken, what’s fixed, what still needs attention).
  • Documentation: Leave things better than you found them by documenting fixes, edge cases, and recurring issues in a way that builds institutional knowledge over time.

Who You Are

Experience & Skills

  • Hands-on experience working on software products in a QA, product, or technical support role. You know what good looks like and can spot when something falls short.
  • You have built or experimented with Claude Code and understand how to apply AI tooling to accelerate real work, not just demo it.
  • You have a track record of resolving product issues quickly and independently, without needing someone to hold your hand through every step.
  • You can read user feedback and extract signal from noise. You understand enough about UX to know when a complaint reflects a real problem and when it doesn’t.
  • You communicate clearly in writing and know how to escalate an issue in a way that gives product leads exactly what they need to act.
  • You’re comfortable working without a detailed process in place. You can create your own system for managing volume and staying on top of what matters.

Mindset & Traits

  • You move fast and hate letting things sit. A backlog of unresolved issues bothers you at a gut level.
  • You have real product taste. You’ve spent enough time thinking about user experience that good and bad feel obvious to you, not just theoretical.
  • You take ownership. When something is your responsibility, it gets done. You don’t wait for reminders or check-ins.
  • You’re honest about what you know and what you don’t, and you’d rather ask a sharp question than guess wrong and waste everyone’s time.
  • You see this as a starting point, not a ceiling. You want to earn more responsibility and are motivated by what this role could grow into.

Compensation & Logistics

  • This is a contract role with an hourly rate commensurate with experience. We’re targeting $20-30/hr.
  • Work is fully remote; US or Canada-based candidates only.
  • Hours are flexible but output is not. You’ll be expected to move quickly and stay responsive during active sprints.
  • We’re looking to bring great candidates on immediately.

Equal Opportunity Commitment

Human Agency is an Equal Opportunity Employer. We value diverse perspectives and are committed to building inclusive, high-performing teams where everyone can do their best work.

Read the full description
Trainer AI Product Quality Specialist at Human Agency

Evaluates and improves AI product quality by reviewing outputs, identifying issues, and supporting responsible AI implementation.

Mid Remote Posted 3 days ago RemoteFirstJobs Product
What this role involves

Location: US or Canada

Type: Independent Contractor (potential for conversion to full-time hire)

**A note before you read further**

The world is changing fast, and so is this role. We can’t fully tell you what this job looks like in six months. That’s true of every role at Human Agency right now. What we need from you, no matter what: accountability, curiosity, and a builder-first mindset. If that sounds exciting, keep reading. If it sounds intimidating but you’re still interested, let’s talk. If none of that sounds like you, that’s okay. This probably isn’t your spot.

About Human Agency

We’re scaling rapidly and have a growing pipeline of opportunities that demand exceptional talent across disciplines. Our mission is to bring on individuals, from creative producers to technical experts to entrepreneurial leaders, who can help us realize this next chapter of growth.

We are a company of doers. Leaders roll up their sleeves, teams work flat, and everyone contributes to what ships. Titles don’t insulate us from feedback or basics. We invite critique, learn quickly, and keep raising the bar. The best ideas win here, no matter where they come from, because clients trust us to deliver the strongest outcomes every time.

Our clients’ missions, products, and bottom lines are sacred. We immerse ourselves in their world, becoming stewards of their goals and partners in solving big problems. Every product, strategy, or asset we create must be both beautiful and functional; practical, usable, and designed for real-world impact.

Humans are our most valuable resource, and we only grow by hiring people who push us forward. Across strategy, engineering, design, data, and operations, we seek out teammates who raise the bar and make us better. Always hire up, never down.

We partner with organizations of all sizes to explore, design, and implement AI strategies that are secure, scalable, and human-centered. We believe AI should amplify human potential, not replace it, and we build with that conviction in every engagement. From advisory and tooling to implementation and education, we meet clients where they are and help them integrate AI in ways that align with their mission and values. Our goal is to empower teams to work smarter, move faster, and unlock new possibilities through thoughtful, responsible innovation.

And through it all, we lead with purpose, love, and adventure. We do meaningful work with people we care about, and we make the ride an adventure worth taking. Because at Human Agency, who we are and how we work are one and the same.

The Opportunity

Our senior team is shipping fast. Too fast to spend cycles on bugs, friction points, and the steady accumulation of product rough edges that slow users down and erode trust. That’s where you come in. This role exists because we need someone with sharp product instincts who can triage what actually matters to users, resolve issues quickly, and exercise real judgment about what to fix, what to escalate, and what to ignore.

This is not a ticket-processing job. Anyone can read feedback and route it somewhere. We need someone who reads feedback, understands what users are actually trying to do, and can tell the difference between a bug that’s blocking real work and one that’s noise. You’ll be the person who keeps product quality tight so leadership can stay focused on building what’s next.

This role can grow well beyond where it starts. Strong judgment and consistent output can open the door to more: building products directly with us, and eventually client-facing solutions work, depending on performance, business need, and demonstrated ability. We hire for trajectory, not just task completion, but that trajectory has to be earned and the timeline isn’t fixed.

What You’ll Own

  • Bug triage and resolution: Work through product issues reported by users with speed and precision. You’ll regularly work in the codebase, submit fixes, and validate them before release.
  • User feedback interpretation: Read between the lines of what users report to identify the real problem, not just the surface complaint, and prioritize accordingly.
  • Product quality monitoring: Maintain a running pulse on the user experience across our products, flagging patterns and systemic issues before they compound. You’ll own testing your changes, checking relevant edge cases, and confirming the original reported workflow works before considering an issue resolved.
  • AI-assisted workflows: Use Claude Code and other AI tools to move faster and resolve more, without sacrificing the judgment that makes the output actually useful.
  • Cross-functional communication: Keep relevant teammates informed with crisp, specific updates (what’s broken, what’s fixed, what still needs attention).
  • Documentation: Leave things better than you found them by documenting fixes, edge cases, and recurring issues in a way that builds institutional knowledge over time.

Who You Are

Experience & Skills

  • Hands-on experience working on software products in a QA, product, or technical support role. You know what good looks like and can spot when something falls short.
  • You have built or experimented with Claude Code and understand how to apply AI tooling to accelerate real work, not just demo it.
  • You have a track record of resolving product issues quickly and independently, without needing someone to hold your hand through every step.
  • You can read user feedback and extract signal from noise. You understand enough about UX to know when a complaint reflects a real problem and when it doesn’t.
  • You communicate clearly in writing and know how to escalate an issue in a way that gives product leads exactly what they need to act.
  • You’re comfortable working without a detailed process in place. You can create your own system for managing volume and staying on top of what matters.

Mindset & Traits

  • You move fast and hate letting things sit. A backlog of unresolved issues bothers you at a gut level.
  • You have real product taste. You’ve spent enough time thinking about user experience that good and bad feel obvious to you, not just theoretical.
  • You take ownership. When something is your responsibility, it gets done. You don’t wait for reminders or check-ins.
  • You’re honest about what you know and what you don’t, and you’d rather ask a sharp question than guess wrong and waste everyone’s time.
  • You see this as a starting point, not a ceiling. You want to earn more responsibility and are motivated by what this role could grow into.

Compensation & Logistics

  • This is a contract role with an hourly rate commensurate with experience. We’re targeting $15-30/hr.
  • Work is fully remote; US or Canada-based candidates only.
  • Hours are flexible but output is not. You’ll be expected to move quickly and stay responsive during active sprints.
  • We’re looking to bring great candidates on immediately.

Equal Opportunity Commitment

Human Agency is an Equal Opportunity Employer. We value diverse perspectives and are committed to building inclusive, high-performing teams where everyone can do their best work.

Read the full description
Trainer Management Consultant

Evaluates and trains AI systems using management consulting expertise on high-impact AI projects.

Mid Remote Posted 3 days ago Himalayas
What this role involves
Role Title: Management ConsultantRole Type: ContractorLocation: RemoteAbout the Role micro1 is partnering with a leading AI lab to bring on former management consultants for high-impact AI training and evaluation projects.
Read the full description
Trainer AI Technical Mentor – Independent Contractor (US Canada, Europe, MENA, APAC)

Mentor and guide learners through AI technical concepts and projects as an independent contractor across multiple global regions.

Mid Remote Posted 18 days ago Jobicy AI
What this role involves
About Us Udacity is now an Accenture company, and exciting things are happening! 🚀 We are on a mission of forging futures in tech through radical talent transformation in digital...
Read the full description
Trainer AI Tutor - Video (Weekend) at xAI

Labels and annotates video content to train AI models on video understanding, using expertise in video editing, motion graphics, and VFX.

Mid Remote Posted 18 days ago RemoteFirstJobs Product
What this role involves

SpaceXAI’s mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. Our team is small, highly motivated, and focused on engineering excellence. This organization is for individuals who appreciate challenging themselves and thrive on curiosity. We operate with a flat organizational structure. All employees are expected to be hands-on and to contribute directly to the company’s mission. Leadership is given to those who show initiative and consistently deliver excellence. Work ethic and strong prioritization skills are important. All employees are expected to have strong communication skills. They should be able to concisely and accurately share knowledge with their teammates.

ABOUT THE ROLE:

As a Video Specialist, you will contribute to SpaceXAI’s mission by training and refining Grok’s ability to interpret and generate video content with precision and technical depth. Key to this role is expertise in video editing, motion graphics, or VFX, a track record of producing high quality video work, and a strong understanding of how moving images are constructed in post-production. This is a weekend-focused role with set working hours on Saturday and Sunday. Full-time positions typically follow a Saturday through Wednesday or Thursday through Monday schedule.

RESPONSIBILITIES:

  • Use proprietary software to provide labels, annotations, and inputs on projects involving video and multimedia elements.
  • Deliver high-quality curated data that captures the nuances of motion, timing, transitions, and visual effects, enhancing Grok’s understanding of video as a medium.

BASIC QUALIFICATIONS:

  • Portfolio displaying excellence in video work - edited shorts, motion graphics, VFX breakdowns, compositing reels, or similar (personal website, YouTube, Vimeo, X account, ArtStation, etc.).
  • Strong skills in video editing, pacing, color grading, and narrative flow.
  • Hands-on proficiency with tools such as Premiere Pro, DaVinci Resolve, After Effects, Nuke, or similar.
  • Ability to critically analyze and articulate what makes a video sequence work or fail at a technical level.
  • Strong communication and analytical skills.
  • Strong written and verbal English skills.

PREFERRED SKILLS AND EXPERIENCE:

  • Experience with compositing, rotoscoping, or motion tracking workflows.
  • Familiarity with AI video generation tools (Grok Imagine, Runway, Kling, Sora, Veo, or similar).
  • Experience with 3D integration in video pipelines (matchmoving, CG compositing).
  • Familiarity with scripting or automation in post-production (Python, plugins, etc.).

LOCATION AND OTHER EXPECTATIONS:

  • This role has set working hours on Saturday and Sunday.
  • Tutor roles may be offered as full-time, part-time, or contractor positions, depending on role needs and candidate fit.
  • Tutor roles may be performed remotely from any location worldwide, subject to legal eligibility, time-zone compatibility, and role specific needs.
  • For US based candidates, please note we are unable to hire in the states of Wyoming and Illinois at this time.
  • We are unable to provide visa sponsorship.
  • For those who will be working from a personal device, your computer must be a Chromebook, Mac with MacOS 11.0 or later, or Windows 10 or later.

COMPENSATION AND BENEFITS:

US based candidates: $40/hour - $75/hour depending on factors including relevant experience, skills, education, geographic location, and qualifications. International candidates: Information will be provided to you during the recruitment process.

Benefits vary based on employment type, location and jurisdiction. Benefits for eligible U.S. based positions include health insurance, 401(k) plan, and paid sick leave. Specific details and role specific information will be provided to you during the interview process.

SpaceXAI is an equal opportunity employer. For details on data processing, view our Recruitment Privacy Notice.

Read the full description
Trainer AI Tutor - Video (Weekend) at xAI

Video specialist provides labeled training data and annotations on video content to improve an AI model's understanding of motion, visual effects, and video production.

Mid Remote Posted 18 days ago RemoteFirstJobs Product
What this role involves

SpaceXAI’s mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. Our team is small, highly motivated, and focused on engineering excellence. This organization is for individuals who appreciate challenging themselves and thrive on curiosity. We operate with a flat organizational structure. All employees are expected to be hands-on and to contribute directly to the company’s mission. Leadership is given to those who show initiative and consistently deliver excellence. Work ethic and strong prioritization skills are important. All employees are expected to have strong communication skills. They should be able to concisely and accurately share knowledge with their teammates.

ABOUT THE ROLE:

As a Video Specialist, you will contribute to SpaceXAI’s mission by training and refining Grok’s ability to interpret and generate video content with precision and technical depth. Key to this role is expertise in video editing, motion graphics, or VFX, a track record of producing high quality video work, and a strong understanding of how moving images are constructed in post-production. This is a weekend-focused role with set working hours on Saturday and Sunday. Full-time positions typically follow a Saturday through Wednesday or Thursday through Monday schedule.

RESPONSIBILITIES:

  • Use proprietary software to provide labels, annotations, and inputs on projects involving video and multimedia elements.
  • Deliver high-quality curated data that captures the nuances of motion, timing, transitions, and visual effects, enhancing Grok’s understanding of video as a medium.

BASIC QUALIFICATIONS:

  • Portfolio displaying excellence in video work - edited shorts, motion graphics, VFX breakdowns, compositing reels, or similar (personal website, YouTube, Vimeo, X account, ArtStation, etc.).
  • Strong skills in video editing, pacing, color grading, and narrative flow.
  • Hands-on proficiency with tools such as Premiere Pro, DaVinci Resolve, After Effects, Nuke, or similar.
  • Ability to critically analyze and articulate what makes a video sequence work or fail at a technical level.
  • Strong communication and analytical skills.
  • Strong written and verbal English skills.

PREFERRED SKILLS AND EXPERIENCE:

  • Experience with compositing, rotoscoping, or motion tracking workflows.
  • Familiarity with AI video generation tools (Grok Imagine, Runway, Kling, Sora, Veo, or similar).
  • Experience with 3D integration in video pipelines (matchmoving, CG compositing).
  • Familiarity with scripting or automation in post-production (Python, plugins, etc.).

LOCATION AND OTHER EXPECTATIONS:

  • This role has set working hours on Saturday and Sunday.
  • Tutor roles may be offered as full-time, part-time, or contractor positions, depending on role needs and candidate fit.
  • Tutor roles may be performed remotely from any location worldwide, subject to legal eligibility, time-zone compatibility, and role specific needs.
  • For US based candidates, please note we are unable to hire in the states of Wyoming and Illinois at this time.
  • We are unable to provide visa sponsorship.
  • For those who will be working from a personal device, your computer must be a Chromebook, Mac with MacOS 11.0 or later, or Windows 10 or later.

COMPENSATION AND BENEFITS:

US based candidates: $40/hour - $75/hour depending on factors including relevant experience, skills, education, geographic location, and qualifications. International candidates: Information will be provided to you during the recruitment process.

Benefits vary based on employment type, location and jurisdiction. Benefits for eligible U.S. based positions include health insurance, 401(k) plan, and paid sick leave. Specific details and role specific information will be provided to you during the interview process.

SpaceXAI is an equal opportunity employer. For details on data processing, view our Recruitment Privacy Notice.

Read the full description
Trainer Korean Transcription Expert - Fully Remote | Upto $33/hr

Transcribes Korean audio content to create training data for AI models and language systems.

Mid Remote Posted 18 days ago Himalayas
What this role involves
About the jobMercor connects elite creative and technical talent with leading AI research labs.
Read the full description
Trainer Voice Actor - Freelance AI Trainer Project

Record voice samples and provide feedback to train AI voice models and improve synthetic speech quality.

Mid Remote Posted 20 days ago Himalayas
What this role involves
Are you an experienced voice actor eager to shape the future of AI?
Read the full description
Trainer Puzzle Solver (Coding)

Solves coding puzzles and algorithmic problems to train AI models on reasoning and debugging tasks.

Mid Remote Posted 21 days ago Himalayas
What this role involves
Puzzle Solver (Coding)Pay: $50–$80/hour Location: Global, fully remote Job Type: Contractor (~15 hours per week) Schedule: Flexible—you choose the hours and days you work, including weekends if desired We are looking for highly skilled coding puzzle solvers to contribute to an AI training project involving algorithmic reasoning, debugging, constrained technical problems, and reproducible software solutions.
Read the full description
Trainer Toptal : Professional Photoshop Artists

Professional Photoshop artist completes creative editing tasks while recording sessions to build an AI training dataset of real production work.

Mid Remote Posted 21 days ago We Work Remotely — Programming
What this role involves

Headquarters: Remote
URL: https://www.toptal.com/

About the Role

Our client is building an open dataset of real, professional Photoshop work, published on HuggingFace under CC-BY-4.0, to evaluate how well AI systems understand creative editing. Not synthetic tasks. Not staged demos. Actual production work, done the way you already do it.

We're looking for working Photoshop artists across six disciplines to complete client-style briefs while a screen recorder captures the session. You work from a brief, you open Photoshop, you do the job. Each task runs roughly 1–3 hours.

Your process is the product here. The layer decisions, the masking approach, the moment you undo three steps and take a different route - that's exactly the signal this dataset is built to capture. If you've ever thought AI evaluation of creative work is shallow because the people building it have never actually retouched a beauty shot or built a matte painting, this is your chance to fix that.

Rate: $25/hr USD

Timeline: 

This is an extremely urgent role with work starting immediately. If you are interested and apply, please expect to be contacted and requested to complete your account setup on September 9th (Friday), 10th (Saturday), or the 11th (Sunday) to be considered for this role. 

What You'll Do

  • Complete Photoshop tasks from a client-provided creative brief, working at your normal professional standard

  • Run a client-provided session recorder for the full duration of each task

  • Deliver layered, production-quality PSDs along with final output

  • Work independently, at your own pace, within the scope of each brief

  • Contribute work that will be published openly for research use

What You Bring

We're recruiting across six specializations and aiming for balanced coverage. Tell us which one is your home turf:

  1. Matte painting / digital environments

  2. Key visual / campaign compositing

  3. Game & UI assets

  4. Print & packaging pre-press

  5. Portrait / beauty retouching

  6. E-commerce / product retouching

Hard requirements - please read carefully:

  • Apple Silicon Mac (M1 or later) with your own licensed copy of Photoshop. Intel Macs and Windows machines cannot be used for this engagement.

  • Photoshop only. No Illustrator, Nuke, Blender, Procreate, or iPad workflows in the captured session.

  • No generative AI of any kind - no Generative Fill, no Firefly, no third-party generative plugins. All work must be manually executed.

  • Individual contributor. No teams, studios, or subcontracted work.

  • Comfortable publishing openly under CC-BY-4.0, meaning the resulting work and session data will be freely available for anyone to use with attribution.

  • Demonstrable professional portfolio in at least one of the six categories above.

To apply: https://weworkremotely.com/remote-jobs/toptal-professional-photoshop-artists

Read the full description
Trainer ML Annotation QA Engineer at Gather AI

Analyzes and ensures quality of annotated training data for computer vision models, making judgment calls on annotation accuracy and identifying data drift or model regressions.

Mid Posted 27 days ago RemoteFirstJobs Product
What this role involves

Job Title: ML Annotation QA Engineer

About Us

Are you ready to build the future of supply chain? At Gather AI, we’re not just creating software, we’re pioneering a new era of warehouse intelligence. We’ve developed a groundbreaking, vision-powered platform that uses autonomous drones and existing equipment to capture real-time data, completely digitizing workflows that have historically been manual and error-prone. This means facilities operate smarter, safer, and more efficiently, ultimately redefining “on-time, in full” delivery.

If you’re looking for an opportunity to contribute to truly transformative technology and make a significant impact in a vital industry, Gather AI is the place for you. We’re leading the charge in the rapidly evolving robotics industry, and we invite you to join us in reshaping the global supply chain, one intelligent warehouse at a time.

About the Team

Our engineering organization spans autonomy, computer vision and machine learning, embedded and hardware systems, full-stack, and cloud, all working in parallel across multiple active product lines tied to live customer deployments. It’s a technically deep, fast-moving team where individual contributors carry real accountability and the work shows up directly in customer operations. Ground truth quality sits at the centre of that — the annotated data this role owns is what our models are trained and measured against.

About the Role

We are looking for an ML Annotation QA Engineer to own the quality of annotated data across our computer vision and machine learning programs. This role is responsible for the judgment-heavy analysis that cannot be reliably outsourced, for the decision rules behind it, and for turning annotation output into an ongoing read on how our systems are actually performing in the field.

We work with an external annotation partner at production volume, and that continues. What we need in house is someone who can analyse the annotated data, make and defend the calls the vendor cannot make consistently, and build the aggregate view that shows which facilities and equipment are degrading and why. Annotation drift, a model regression, a tool bug, and genuine field degradation all look similar in a chart and require completely different responses — telling them apart is the core of the job.

You will work closely with Machine Learning Engineers, QA, and Engineering, and the first assignment is our warehouse forklift vision program, where barcode readability and localisation analysis are the immediate need. From there the remit grows with us: drone imagery annotation today, and new task types as customer-driven capabilities come online. Success in this role requires a combination of analytical rigor, sound judgment under ambiguity, and clear written communication.

What You’ll Do

  • Own the judgment-heavy quality analysis on annotated data that cannot be reliably outsourced — working the daily review queue and producing verdicts and root cause in house.
  • Own, version, and refine the verdict taxonomy, decision rules, and quality guidelines for the categories you cover.
  • Build and maintain performance trackers over annotated data — error rates by facility, site, equipment, and data format over time, against an agreed baseline.
  • Detect anomalies against that baseline and flag them the day they appear rather than weeks later.
  • Run root cause analysis on flagged anomalies, distinguishing annotation error from model or system error from genuine degradation in the field.
  • Report findings to engineering and ML with reproducible evidence and stated confidence, fast enough that the issue is still observable.
  • Identify systematic failure patterns rather than one-off misses, and maintain a documented pattern library others can use.
  • Query and analyse annotation data directly with Python and SQL to test hypotheses, without waiting on extracts from anyone.
  • Feed annotation-quality findings back as concrete SOP and instruction changes when the root cause is labeling rather than system behaviour.
  • Specify annotation tool improvements — what the tool should surface so analysis stops requiring manual work — and validate the fixes.
  • Stand up quality analysis and reporting for new annotation programs as they come online.
  • Track work across Jira and contribute to pre-release validation for the behaviours you cover.

First 90 Days

Within your first three months, you will be expected to:

Own the Quality Analysis

  • Take over barcode and location root cause analysis from the annotation vendor.
  • Move from supervised review to owning the daily queue at the agreed review threshold, inside the expected time budget.
  • Become the primary owner of verdicts and root cause for the categories you cover.

Become Fluent in the Annotation Pipeline

Develop working expertise in the data and the domain behind it:

  • What is captured, at what grain, and where the known quality limits are
  • Racks, locations, levels, bins and bin types; LPN, SKU, AWB and other code formats; exceptions and exception types; OCR versus barcode and the failure modes of each

Build the Performance Tracker

  • Stand up a facility performance tracker with an agreed baseline, thresholds, and reporting cadence.
  • Establish anomaly detection against it, and get the team to the point of trusting and using it.
  • Investigate flagged anomalies independently, with a root cause hypothesis and stated confidence.

Own Documentation

Take ownership of:

  • Barcode and location verdict taxonomy and decision rules
  • The pattern library of known failure modes

Responsibilities include:

  • Version control
  • Closing documentation gaps
  • Adding edge-case guidance
  • Feeding changes back to the annotation vendor’s SOPs where the root cause is labeling

Required Technical Skills

  • BS in Computer Science/Engineering, Electrical Engineering, or equivalent experience
  • Experience working with annotated datasets for CV/ML, including assessing label quality
  • Strong understanding of statistics, and able to work with data
  • Root cause analysis — generating competing hypotheses and naming the evidence that separates them
  • Familiarity with Python and SQL, or equivalent, for querying and analysing data independently
  • Writing quality guidelines, decision rules, and labeling taxonomies
  • Excellent documentation, communication, and collaboration skills

Nice-to-Have Skills

  • 2+ years in quality assurance, data quality, or product support
  • 1+ years experience with enterprise-grade ticketing systems (e.g. Jira)
  • Computer vision annotation experience as a reviewer or auditor — video event labeling, bounding boxes, polygon segmentation, counting, or classification
  • Annotation quality methodology — gold-set validation, inter-annotator agreement, sampling design
  • Strong spatial and geometric reasoning — relevant wherever labels describe position in physical space
  • Experience in warehouse automation, robotics, or computer vision applications
  • Dashboarding or BI tooling for recurring reports
  • Small team experience

Qualifications

Required

  • BS in Computer Science/Engineering, Electrical Engineering, or equivalent experience
  • 2–5 years of experience in ML QA, annotation quality, data quality, or analytics at an AI/ML company
  • Hands-on experience with annotated ML datasets, including assessing label quality
  • Demonstrated experience finding, diagnosing, and reporting data anomalies to a technical audience
  • Able to get to a defensible answer from unfamiliar data without someone preparing it first
  • Understanding of data privacy and confidentiality requirements when working with customer operational data

Preferred

  • Experience in warehouse automation, robotics, or computer vision applications
  • Experience with annotation platforms and quality tooling
  • Experience authoring annotation guidelines or standard operating procedures

Required Technologies

  • Python
  • SQL, or an equivalent query language
  • Jira

Success Traits

We’re looking for someone who is:

  • Analytical, able to tell a real pattern from noise and to say honestly when the data will not settle a question
  • Detail-oriented — a verdict or a tracker that is subtly wrong is worse than none at all
  • Comfortable with ambiguity, and willing to present competing hypotheses rather than forcing a single answer
  • Persistent in chasing a cause across sites, programs, and builds
  • Resourceful and comfortable building processes and reports where they don’t yet exist
  • A strong written communicator — the output of this role is reports other people act on
  • Collaborative, respectful, and an excellent cross-functional partner

What Success Looks Like

By the end of your first six months, you will have:

  • Established yourself as the primary owner of annotation quality analysis and root cause
  • A performance tracker the team relies on, running on a regular cadence
  • Closed the reporting cycle to same day, so issues are raised while still observable
  • Standardized the verdict taxonomy and decision rules for the categories you cover
  • Built a documented pattern library of known failure modes that others can use
  • Improved the annotation tool in at least one way that measurably reduces manual effort
  • Given ML and engineering a clearer picture of what error patterns mean for model accuracy and field performance
  • Created a repeatable method for standing up annotation quality on the next program
Read the full description