it job board logo
  • Home
  • Find IT Jobs
  • Register CV
  • Register as Employer
  • Contact us
  • Career Advice
  • Recruiting? Post a job
  • Sign in
  • Sign up
  • Home
  • Find IT Jobs
  • Register CV
  • Register as Employer
  • Contact us
  • Career Advice
Sorry, that job is no longer available. Here are some results that may be similar to the job you were looking for.

198 jobs found

Email me jobs like this
Refine Search
Current Search
network communications systems specialist
Kannada Linguistic QA Specialist (Remote)
Braintrust Bellville, Texas
Job description About this role In this hourly, remote contractor role, you will work as a Kannada Quality Assurance Lead (QAL) to oversee quality, consistency, and trainer performance across Kannada AI training projects. You will review AI-generated Kannada content and trainer/QA work, evaluate output quality against project guidelines, provide precise written feedback, and ensure that all contributors follow the expected quality standards. You will assess work for accuracy, fluency, grammar, spelling, tone, cultural appropriateness, meaning preservation, instruction-following, formatting, and adherence to project-specific rubrics. You will spot recurring quality issues, communicate updates to trainers and QAs, support onboarding, maintain documentation, and help activate contributors who are not working consistently. This role requires strong Kannada and English skills, excellent attention to detail, structured communication, and the ability to manage quality workflows across remote teams. This role is with SME Careers, a fast-growing AI Data Services company and subsidiary of SuperAnnotate, delivering training data for many of the world's largest AI companies and foundation-model labs. Your Kannada quality leadership will directly help improve the world's premier AI models by ensuring that Kannada training data is natural, accurate, culturally appropriate, well-documented, and aligned with client expectations. Your profile Bachelor's or Master's degree in Kannada, Linguistics, Translation, Communications, Journalism, English, Education, Quality Assurance, or a relevant domain/related field. Native or near-native Kannada proficiency with strong reading and writing skills. Strong grasp of the English language to follow project guidelines, communicate with teams, and provide clear feedback in English. 3 years of professional experience in Kannada writing, editing, translation, localization, content QA, AI training, education, annotation, or related language-review workflows. Strong understanding of Kannada grammar, spelling conventions, punctuation, tone, register, regional variation, and cultural context. Ability to evaluate Kannada content against detailed rubrics and identify issues such as mistranslation, literal phrasing, unnatural tone, hallucinated claims, ambiguity, or inconsistent terminology. Experience leading or supporting remote teams of trainers, annotators, reviewers, editors, or QAs is strongly preferred. Comfortable working in fast-moving remote environments using tools such as Discord, Google Sheets, Google Docs, trackers, dashboards, and project management systems. Highly detail-oriented and organized, with the ability to maintain style guides, FAQs, trackers, onboarding materials, honeypots, and other quality documentation. Experience with AI training, data annotation, large language models, prompt/response evaluation, or rubric-based LLM QA is a strong plus. Key responsibilities Quality monitoring: Spot-check Kannada items, identify quality issues, provide ongoing feedback through DMs, and escalate recurring or critical issues. Trainer and QA communication: Update trainers and QAs on Discord about new item guidelines, project changes, workflow updates, and quality expectations. Question handling: Respond to trainer/QA questions clearly and promptly, especially around Kannada wording, register, translation fidelity, cultural context, regional variation, and edge cases. Trainer/QA activation management: DM contributors who are inactive or not working, encourage activation, track follow-ups, and flag availability issues when needed. Documentation: Create and maintain Kannada project documentation, including style guides, trackers, FAQs, quality notes, examples, honeypots, and onboarding materials. Onboarding and training: Schedule and run onboarding/training calls with trainers and QAs to explain project expectations, workflows, rubrics, quality standards, and Kannada-specific style requirements. Quality alignment: Ensure all trainers and QAs apply Kannada language guidelines consistently and understand updates as projects evolve. Process improvement: Identify recurring quality gaps, propose workflow improvements, and help build scalable QA processes for Kannada-language projects. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
08/01/2026
Full time
Job description About this role In this hourly, remote contractor role, you will work as a Kannada Quality Assurance Lead (QAL) to oversee quality, consistency, and trainer performance across Kannada AI training projects. You will review AI-generated Kannada content and trainer/QA work, evaluate output quality against project guidelines, provide precise written feedback, and ensure that all contributors follow the expected quality standards. You will assess work for accuracy, fluency, grammar, spelling, tone, cultural appropriateness, meaning preservation, instruction-following, formatting, and adherence to project-specific rubrics. You will spot recurring quality issues, communicate updates to trainers and QAs, support onboarding, maintain documentation, and help activate contributors who are not working consistently. This role requires strong Kannada and English skills, excellent attention to detail, structured communication, and the ability to manage quality workflows across remote teams. This role is with SME Careers, a fast-growing AI Data Services company and subsidiary of SuperAnnotate, delivering training data for many of the world's largest AI companies and foundation-model labs. Your Kannada quality leadership will directly help improve the world's premier AI models by ensuring that Kannada training data is natural, accurate, culturally appropriate, well-documented, and aligned with client expectations. Your profile Bachelor's or Master's degree in Kannada, Linguistics, Translation, Communications, Journalism, English, Education, Quality Assurance, or a relevant domain/related field. Native or near-native Kannada proficiency with strong reading and writing skills. Strong grasp of the English language to follow project guidelines, communicate with teams, and provide clear feedback in English. 3 years of professional experience in Kannada writing, editing, translation, localization, content QA, AI training, education, annotation, or related language-review workflows. Strong understanding of Kannada grammar, spelling conventions, punctuation, tone, register, regional variation, and cultural context. Ability to evaluate Kannada content against detailed rubrics and identify issues such as mistranslation, literal phrasing, unnatural tone, hallucinated claims, ambiguity, or inconsistent terminology. Experience leading or supporting remote teams of trainers, annotators, reviewers, editors, or QAs is strongly preferred. Comfortable working in fast-moving remote environments using tools such as Discord, Google Sheets, Google Docs, trackers, dashboards, and project management systems. Highly detail-oriented and organized, with the ability to maintain style guides, FAQs, trackers, onboarding materials, honeypots, and other quality documentation. Experience with AI training, data annotation, large language models, prompt/response evaluation, or rubric-based LLM QA is a strong plus. Key responsibilities Quality monitoring: Spot-check Kannada items, identify quality issues, provide ongoing feedback through DMs, and escalate recurring or critical issues. Trainer and QA communication: Update trainers and QAs on Discord about new item guidelines, project changes, workflow updates, and quality expectations. Question handling: Respond to trainer/QA questions clearly and promptly, especially around Kannada wording, register, translation fidelity, cultural context, regional variation, and edge cases. Trainer/QA activation management: DM contributors who are inactive or not working, encourage activation, track follow-ups, and flag availability issues when needed. Documentation: Create and maintain Kannada project documentation, including style guides, trackers, FAQs, quality notes, examples, honeypots, and onboarding materials. Onboarding and training: Schedule and run onboarding/training calls with trainers and QAs to explain project expectations, workflows, rubrics, quality standards, and Kannada-specific style requirements. Quality alignment: Ensure all trainers and QAs apply Kannada language guidelines consistently and understand updates as projects evolve. Process improvement: Identify recurring quality gaps, propose workflow improvements, and help build scalable QA processes for Kannada-language projects. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
Kannada Linguistic QA Specialist (Remote)
Braintrust Seattle, Washington
Job description About this role In this hourly, remote contractor role, you will work as a Kannada Quality Assurance Lead (QAL) to oversee quality, consistency, and trainer performance across Kannada AI training projects. You will review AI-generated Kannada content and trainer/QA work, evaluate output quality against project guidelines, provide precise written feedback, and ensure that all contributors follow the expected quality standards. You will assess work for accuracy, fluency, grammar, spelling, tone, cultural appropriateness, meaning preservation, instruction-following, formatting, and adherence to project-specific rubrics. You will spot recurring quality issues, communicate updates to trainers and QAs, support onboarding, maintain documentation, and help activate contributors who are not working consistently. This role requires strong Kannada and English skills, excellent attention to detail, structured communication, and the ability to manage quality workflows across remote teams. This role is with SME Careers, a fast-growing AI Data Services company and subsidiary of SuperAnnotate, delivering training data for many of the world's largest AI companies and foundation-model labs. Your Kannada quality leadership will directly help improve the world's premier AI models by ensuring that Kannada training data is natural, accurate, culturally appropriate, well-documented, and aligned with client expectations. Your profile Bachelor's or Master's degree in Kannada, Linguistics, Translation, Communications, Journalism, English, Education, Quality Assurance, or a relevant domain/related field. Native or near-native Kannada proficiency with strong reading and writing skills. Strong grasp of the English language to follow project guidelines, communicate with teams, and provide clear feedback in English. 3 years of professional experience in Kannada writing, editing, translation, localization, content QA, AI training, education, annotation, or related language-review workflows. Strong understanding of Kannada grammar, spelling conventions, punctuation, tone, register, regional variation, and cultural context. Ability to evaluate Kannada content against detailed rubrics and identify issues such as mistranslation, literal phrasing, unnatural tone, hallucinated claims, ambiguity, or inconsistent terminology. Experience leading or supporting remote teams of trainers, annotators, reviewers, editors, or QAs is strongly preferred. Comfortable working in fast-moving remote environments using tools such as Discord, Google Sheets, Google Docs, trackers, dashboards, and project management systems. Highly detail-oriented and organized, with the ability to maintain style guides, FAQs, trackers, onboarding materials, honeypots, and other quality documentation. Experience with AI training, data annotation, large language models, prompt/response evaluation, or rubric-based LLM QA is a strong plus. Key responsibilities Quality monitoring: Spot-check Kannada items, identify quality issues, provide ongoing feedback through DMs, and escalate recurring or critical issues. Trainer and QA communication: Update trainers and QAs on Discord about new item guidelines, project changes, workflow updates, and quality expectations. Question handling: Respond to trainer/QA questions clearly and promptly, especially around Kannada wording, register, translation fidelity, cultural context, regional variation, and edge cases. Trainer/QA activation management: DM contributors who are inactive or not working, encourage activation, track follow-ups, and flag availability issues when needed. Documentation: Create and maintain Kannada project documentation, including style guides, trackers, FAQs, quality notes, examples, honeypots, and onboarding materials. Onboarding and training: Schedule and run onboarding/training calls with trainers and QAs to explain project expectations, workflows, rubrics, quality standards, and Kannada-specific style requirements. Quality alignment: Ensure all trainers and QAs apply Kannada language guidelines consistently and understand updates as projects evolve. Process improvement: Identify recurring quality gaps, propose workflow improvements, and help build scalable QA processes for Kannada-language projects. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
08/01/2026
Full time
Job description About this role In this hourly, remote contractor role, you will work as a Kannada Quality Assurance Lead (QAL) to oversee quality, consistency, and trainer performance across Kannada AI training projects. You will review AI-generated Kannada content and trainer/QA work, evaluate output quality against project guidelines, provide precise written feedback, and ensure that all contributors follow the expected quality standards. You will assess work for accuracy, fluency, grammar, spelling, tone, cultural appropriateness, meaning preservation, instruction-following, formatting, and adherence to project-specific rubrics. You will spot recurring quality issues, communicate updates to trainers and QAs, support onboarding, maintain documentation, and help activate contributors who are not working consistently. This role requires strong Kannada and English skills, excellent attention to detail, structured communication, and the ability to manage quality workflows across remote teams. This role is with SME Careers, a fast-growing AI Data Services company and subsidiary of SuperAnnotate, delivering training data for many of the world's largest AI companies and foundation-model labs. Your Kannada quality leadership will directly help improve the world's premier AI models by ensuring that Kannada training data is natural, accurate, culturally appropriate, well-documented, and aligned with client expectations. Your profile Bachelor's or Master's degree in Kannada, Linguistics, Translation, Communications, Journalism, English, Education, Quality Assurance, or a relevant domain/related field. Native or near-native Kannada proficiency with strong reading and writing skills. Strong grasp of the English language to follow project guidelines, communicate with teams, and provide clear feedback in English. 3 years of professional experience in Kannada writing, editing, translation, localization, content QA, AI training, education, annotation, or related language-review workflows. Strong understanding of Kannada grammar, spelling conventions, punctuation, tone, register, regional variation, and cultural context. Ability to evaluate Kannada content against detailed rubrics and identify issues such as mistranslation, literal phrasing, unnatural tone, hallucinated claims, ambiguity, or inconsistent terminology. Experience leading or supporting remote teams of trainers, annotators, reviewers, editors, or QAs is strongly preferred. Comfortable working in fast-moving remote environments using tools such as Discord, Google Sheets, Google Docs, trackers, dashboards, and project management systems. Highly detail-oriented and organized, with the ability to maintain style guides, FAQs, trackers, onboarding materials, honeypots, and other quality documentation. Experience with AI training, data annotation, large language models, prompt/response evaluation, or rubric-based LLM QA is a strong plus. Key responsibilities Quality monitoring: Spot-check Kannada items, identify quality issues, provide ongoing feedback through DMs, and escalate recurring or critical issues. Trainer and QA communication: Update trainers and QAs on Discord about new item guidelines, project changes, workflow updates, and quality expectations. Question handling: Respond to trainer/QA questions clearly and promptly, especially around Kannada wording, register, translation fidelity, cultural context, regional variation, and edge cases. Trainer/QA activation management: DM contributors who are inactive or not working, encourage activation, track follow-ups, and flag availability issues when needed. Documentation: Create and maintain Kannada project documentation, including style guides, trackers, FAQs, quality notes, examples, honeypots, and onboarding materials. Onboarding and training: Schedule and run onboarding/training calls with trainers and QAs to explain project expectations, workflows, rubrics, quality standards, and Kannada-specific style requirements. Quality alignment: Ensure all trainers and QAs apply Kannada language guidelines consistently and understand updates as projects evolve. Process improvement: Identify recurring quality gaps, propose workflow improvements, and help build scalable QA processes for Kannada-language projects. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
Kannada Linguistic QA Specialist (Remote)
Braintrust
Job description About this role In this hourly, remote contractor role, you will work as a Kannada Quality Assurance Lead (QAL) to oversee quality, consistency, and trainer performance across Kannada AI training projects. You will review AI-generated Kannada content and trainer/QA work, evaluate output quality against project guidelines, provide precise written feedback, and ensure that all contributors follow the expected quality standards. You will assess work for accuracy, fluency, grammar, spelling, tone, cultural appropriateness, meaning preservation, instruction-following, formatting, and adherence to project-specific rubrics. You will spot recurring quality issues, communicate updates to trainers and QAs, support onboarding, maintain documentation, and help activate contributors who are not working consistently. This role requires strong Kannada and English skills, excellent attention to detail, structured communication, and the ability to manage quality workflows across remote teams. This role is with SME Careers, a fast-growing AI Data Services company and subsidiary of SuperAnnotate, delivering training data for many of the world's largest AI companies and foundation-model labs. Your Kannada quality leadership will directly help improve the world's premier AI models by ensuring that Kannada training data is natural, accurate, culturally appropriate, well-documented, and aligned with client expectations. Your profile Bachelor's or Master's degree in Kannada, Linguistics, Translation, Communications, Journalism, English, Education, Quality Assurance, or a relevant domain/related field. Native or near-native Kannada proficiency with strong reading and writing skills. Strong grasp of the English language to follow project guidelines, communicate with teams, and provide clear feedback in English. 3 years of professional experience in Kannada writing, editing, translation, localization, content QA, AI training, education, annotation, or related language-review workflows. Strong understanding of Kannada grammar, spelling conventions, punctuation, tone, register, regional variation, and cultural context. Ability to evaluate Kannada content against detailed rubrics and identify issues such as mistranslation, literal phrasing, unnatural tone, hallucinated claims, ambiguity, or inconsistent terminology. Experience leading or supporting remote teams of trainers, annotators, reviewers, editors, or QAs is strongly preferred. Comfortable working in fast-moving remote environments using tools such as Discord, Google Sheets, Google Docs, trackers, dashboards, and project management systems. Highly detail-oriented and organized, with the ability to maintain style guides, FAQs, trackers, onboarding materials, honeypots, and other quality documentation. Experience with AI training, data annotation, large language models, prompt/response evaluation, or rubric-based LLM QA is a strong plus. Key responsibilities Quality monitoring: Spot-check Kannada items, identify quality issues, provide ongoing feedback through DMs, and escalate recurring or critical issues. Trainer and QA communication: Update trainers and QAs on Discord about new item guidelines, project changes, workflow updates, and quality expectations. Question handling: Respond to trainer/QA questions clearly and promptly, especially around Kannada wording, register, translation fidelity, cultural context, regional variation, and edge cases. Trainer/QA activation management: DM contributors who are inactive or not working, encourage activation, track follow-ups, and flag availability issues when needed. Documentation: Create and maintain Kannada project documentation, including style guides, trackers, FAQs, quality notes, examples, honeypots, and onboarding materials. Onboarding and training: Schedule and run onboarding/training calls with trainers and QAs to explain project expectations, workflows, rubrics, quality standards, and Kannada-specific style requirements. Quality alignment: Ensure all trainers and QAs apply Kannada language guidelines consistently and understand updates as projects evolve. Process improvement: Identify recurring quality gaps, propose workflow improvements, and help build scalable QA processes for Kannada-language projects. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
08/01/2026
Full time
Job description About this role In this hourly, remote contractor role, you will work as a Kannada Quality Assurance Lead (QAL) to oversee quality, consistency, and trainer performance across Kannada AI training projects. You will review AI-generated Kannada content and trainer/QA work, evaluate output quality against project guidelines, provide precise written feedback, and ensure that all contributors follow the expected quality standards. You will assess work for accuracy, fluency, grammar, spelling, tone, cultural appropriateness, meaning preservation, instruction-following, formatting, and adherence to project-specific rubrics. You will spot recurring quality issues, communicate updates to trainers and QAs, support onboarding, maintain documentation, and help activate contributors who are not working consistently. This role requires strong Kannada and English skills, excellent attention to detail, structured communication, and the ability to manage quality workflows across remote teams. This role is with SME Careers, a fast-growing AI Data Services company and subsidiary of SuperAnnotate, delivering training data for many of the world's largest AI companies and foundation-model labs. Your Kannada quality leadership will directly help improve the world's premier AI models by ensuring that Kannada training data is natural, accurate, culturally appropriate, well-documented, and aligned with client expectations. Your profile Bachelor's or Master's degree in Kannada, Linguistics, Translation, Communications, Journalism, English, Education, Quality Assurance, or a relevant domain/related field. Native or near-native Kannada proficiency with strong reading and writing skills. Strong grasp of the English language to follow project guidelines, communicate with teams, and provide clear feedback in English. 3 years of professional experience in Kannada writing, editing, translation, localization, content QA, AI training, education, annotation, or related language-review workflows. Strong understanding of Kannada grammar, spelling conventions, punctuation, tone, register, regional variation, and cultural context. Ability to evaluate Kannada content against detailed rubrics and identify issues such as mistranslation, literal phrasing, unnatural tone, hallucinated claims, ambiguity, or inconsistent terminology. Experience leading or supporting remote teams of trainers, annotators, reviewers, editors, or QAs is strongly preferred. Comfortable working in fast-moving remote environments using tools such as Discord, Google Sheets, Google Docs, trackers, dashboards, and project management systems. Highly detail-oriented and organized, with the ability to maintain style guides, FAQs, trackers, onboarding materials, honeypots, and other quality documentation. Experience with AI training, data annotation, large language models, prompt/response evaluation, or rubric-based LLM QA is a strong plus. Key responsibilities Quality monitoring: Spot-check Kannada items, identify quality issues, provide ongoing feedback through DMs, and escalate recurring or critical issues. Trainer and QA communication: Update trainers and QAs on Discord about new item guidelines, project changes, workflow updates, and quality expectations. Question handling: Respond to trainer/QA questions clearly and promptly, especially around Kannada wording, register, translation fidelity, cultural context, regional variation, and edge cases. Trainer/QA activation management: DM contributors who are inactive or not working, encourage activation, track follow-ups, and flag availability issues when needed. Documentation: Create and maintain Kannada project documentation, including style guides, trackers, FAQs, quality notes, examples, honeypots, and onboarding materials. Onboarding and training: Schedule and run onboarding/training calls with trainers and QAs to explain project expectations, workflows, rubrics, quality standards, and Kannada-specific style requirements. Quality alignment: Ensure all trainers and QAs apply Kannada language guidelines consistently and understand updates as projects evolve. Process improvement: Identify recurring quality gaps, propose workflow improvements, and help build scalable QA processes for Kannada-language projects. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
Telugu Linguistic QA Specialist (Remote)
Braintrust Seattle, Washington
Job description About this role In this hourly, remote contractor role, you will work as a Telugu Quality Assurance Lead (QAL) to oversee quality, consistency, and trainer performance across Telugu AI training projects. You will review AI-generated Telugu content and trainer/QA work, evaluate output quality against project guidelines, provide precise written feedback, and ensure that all contributors follow the expected quality standards. You will assess work for accuracy, fluency, grammar, spelling, tone, cultural appropriateness, meaning preservation, instruction-following, formatting, and adherence to project-specific rubrics. You will spot recurring quality issues, communicate updates to trainers and QAs, support onboarding, maintain documentation, and help activate contributors who are not working consistently. This role requires strong Telugu and English skills, excellent attention to detail, structured communication, and the ability to manage quality workflows across remote teams. This role is with SME Careers, a fast-growing AI Data Services company and subsidiary of SuperAnnotate, delivering training data for many of the world's largest AI companies and foundation-model labs. Your Telugu quality leadership will directly help improve the world's premier AI models by ensuring that Telugu training data is natural, accurate, culturally appropriate, well-documented, and aligned with client expectations. Selection process involves an AI interview, a domain-specific task, and an interview with a recruiter. Your profile Bachelor's or Master's degree in Telugu, Linguistics, Translation, Communications, Journalism, English, Education, Quality Assurance, or a relevant domain/related field. Native or near-native Telugu proficiency with strong reading and writing skills. Strong grasp of the English language to follow project guidelines, communicate with teams, and provide clear feedback in English. 3 years of professional experience in Telugu writing, editing, translation, localization, content QA, AI training, education, annotation, or related language-review workflows. Strong understanding of Telugu grammar, spelling conventions, punctuation, tone, register, regional variation, and cultural context. Ability to evaluate Telugu content against detailed rubrics and identify issues such as mistranslation, literal phrasing, unnatural tone, hallucinated claims, ambiguity, or inconsistent terminology. Experience leading or supporting remote teams of trainers, annotators, reviewers, editors, or QAs is strongly preferred. Comfortable working in fast-moving remote environments using tools such as Discord, Google Sheets, Google Docs, trackers, dashboards, and project management systems. Highly detail-oriented and organized, with the ability to maintain style guides, FAQs, trackers, onboarding materials, honeypots, and other quality documentation. Experience with AI training, data annotation, large language models, prompt/response evaluation, or rubric-based LLM QA is a strong plus. Key responsibilities Quality monitoring: Spot-check Telugu items, identify quality issues, provide ongoing feedback through DMs, and escalate recurring or critical issues. Trainer and QA communication: Update trainers and QAs on Discord about new item guidelines, project changes, workflow updates, and quality expectations. Question handling: Respond to trainer/QA questions clearly and promptly, especially around Telugu wording, register, translation fidelity, cultural context, Andhra Pradesh vs Telangana usage, and edge cases. Trainer/QA activation management: DM contributors who are inactive or not working, encourage activation, track follow-ups, and flag availability issues when needed. Documentation: Create and maintain Telugu project documentation, including style guides, trackers, FAQs, quality notes, examples, honeypots, and onboarding materials. Onboarding and training: Schedule and run onboarding/training calls with trainers and QAs to explain project expectations, workflows, rubrics, quality standards, and Telugu-specific style requirements. Quality alignment: Ensure all trainers and QAs apply Telugu language guidelines consistently and understand updates as projects evolve. Process improvement: Identify recurring quality gaps, propose workflow improvements, and help build scalable QA processes for Telugu-language projects. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
08/01/2026
Full time
Job description About this role In this hourly, remote contractor role, you will work as a Telugu Quality Assurance Lead (QAL) to oversee quality, consistency, and trainer performance across Telugu AI training projects. You will review AI-generated Telugu content and trainer/QA work, evaluate output quality against project guidelines, provide precise written feedback, and ensure that all contributors follow the expected quality standards. You will assess work for accuracy, fluency, grammar, spelling, tone, cultural appropriateness, meaning preservation, instruction-following, formatting, and adherence to project-specific rubrics. You will spot recurring quality issues, communicate updates to trainers and QAs, support onboarding, maintain documentation, and help activate contributors who are not working consistently. This role requires strong Telugu and English skills, excellent attention to detail, structured communication, and the ability to manage quality workflows across remote teams. This role is with SME Careers, a fast-growing AI Data Services company and subsidiary of SuperAnnotate, delivering training data for many of the world's largest AI companies and foundation-model labs. Your Telugu quality leadership will directly help improve the world's premier AI models by ensuring that Telugu training data is natural, accurate, culturally appropriate, well-documented, and aligned with client expectations. Selection process involves an AI interview, a domain-specific task, and an interview with a recruiter. Your profile Bachelor's or Master's degree in Telugu, Linguistics, Translation, Communications, Journalism, English, Education, Quality Assurance, or a relevant domain/related field. Native or near-native Telugu proficiency with strong reading and writing skills. Strong grasp of the English language to follow project guidelines, communicate with teams, and provide clear feedback in English. 3 years of professional experience in Telugu writing, editing, translation, localization, content QA, AI training, education, annotation, or related language-review workflows. Strong understanding of Telugu grammar, spelling conventions, punctuation, tone, register, regional variation, and cultural context. Ability to evaluate Telugu content against detailed rubrics and identify issues such as mistranslation, literal phrasing, unnatural tone, hallucinated claims, ambiguity, or inconsistent terminology. Experience leading or supporting remote teams of trainers, annotators, reviewers, editors, or QAs is strongly preferred. Comfortable working in fast-moving remote environments using tools such as Discord, Google Sheets, Google Docs, trackers, dashboards, and project management systems. Highly detail-oriented and organized, with the ability to maintain style guides, FAQs, trackers, onboarding materials, honeypots, and other quality documentation. Experience with AI training, data annotation, large language models, prompt/response evaluation, or rubric-based LLM QA is a strong plus. Key responsibilities Quality monitoring: Spot-check Telugu items, identify quality issues, provide ongoing feedback through DMs, and escalate recurring or critical issues. Trainer and QA communication: Update trainers and QAs on Discord about new item guidelines, project changes, workflow updates, and quality expectations. Question handling: Respond to trainer/QA questions clearly and promptly, especially around Telugu wording, register, translation fidelity, cultural context, Andhra Pradesh vs Telangana usage, and edge cases. Trainer/QA activation management: DM contributors who are inactive or not working, encourage activation, track follow-ups, and flag availability issues when needed. Documentation: Create and maintain Telugu project documentation, including style guides, trackers, FAQs, quality notes, examples, honeypots, and onboarding materials. Onboarding and training: Schedule and run onboarding/training calls with trainers and QAs to explain project expectations, workflows, rubrics, quality standards, and Telugu-specific style requirements. Quality alignment: Ensure all trainers and QAs apply Telugu language guidelines consistently and understand updates as projects evolve. Process improvement: Identify recurring quality gaps, propose workflow improvements, and help build scalable QA processes for Telugu-language projects. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
Gujarati Linguistic QA Specialist (Remote)
Braintrust Boston, Massachusetts
Job description About this role In this hourly, remote contractor role, you will work as a Gujarati Quality Assurance Lead (QAL) to oversee quality, consistency, and trainer performance across Gujarati AI training projects. You will review AI-generated Gujarati content and trainer/QA work, evaluate output quality against project guidelines, provide precise written feedback, and help ensure that all contributors follow the expected standards. You will assess work for accuracy, fluency, grammar, spelling, tone, cultural appropriateness, meaning preservation, instruction-following, formatting, and adherence to project-specific rubrics. You will spot recurring quality issues, communicate updates to trainers and QAs, support onboarding, maintain documentation, and help activate contributors who are not working consistently. This role requires strong Gujarati and English skills, excellent attention to detail, structured communication, and the ability to manage quality workflows across remote teams. This role is with SME Careers, a fast-growing AI Data Services company and subsidiary of SuperAnnotate, delivering training data for many of the world's largest AI companies and foundation-model labs. Your Gujarati quality leadership will directly help improve the world's premier AI models by ensuring that Gujarati training data is natural, accurate, culturally appropriate, well-documented, and aligned with client expectations. Selection process involves an AI interview, a domain-specific task, and an interview with a recruiter. Your profile Bachelor's or Master's degree in Gujarati, Linguistics, Translation, Communications, Journalism, English, Education, Quality Assurance, or a relevant domain/related field. Native or near-native Gujarati proficiency with strong reading and writing skills. Strong grasp of the English language to follow project guidelines, communicate with teams, and provide clear feedback in English. 3 years of professional experience in Gujarati writing, editing, translation, localization, content QA, AI training, education, annotation, or related language-review workflows. Strong understanding of Gujarati grammar, spelling conventions, punctuation, tone, register, and cultural context. Ability to evaluate Gujarati content against detailed rubrics and identify issues such as mistranslation, literal phrasing, unnatural tone, hallucinated claims, ambiguity, or inconsistent terminology. Experience leading or supporting remote teams of trainers, annotators, reviewers, editors, or QAs is strongly preferred. Comfortable working in fast-moving remote environments using tools such as Discord, Google Sheets, Google Docs, trackers, dashboards, and project management systems. Highly detail-oriented and organized, with the ability to maintain style guides, FAQs, trackers, onboarding materials, honeypots, and other quality documentation. Experience with AI training, data annotation, large language models, prompt/response evaluation, or rubric-based LLM QA is a strong plus. Key responsibilities Quality monitoring: Spot-check Gujarati items, identify quality issues, provide ongoing feedback through DMs, and escalate recurring or critical issues. Trainer and QA communication: Update trainers and QAs on Discord about new item guidelines, project changes, workflow updates, and quality expectations. Question handling: Respond to trainer/QA questions clearly and promptly, especially around Gujarati wording, register, translation fidelity, cultural context, and edge cases. Trainer/QA activation management: DM contributors who are inactive or not working, encourage activation, track follow-ups, and flag availability issues when needed. Documentation: Create and maintain Gujarati project documentation, including style guides, trackers, FAQs, quality notes, examples, honeypots, and onboarding materials. Onboarding and training: Schedule and run onboarding/training calls with trainers and QAs to explain project expectations, workflows, rubrics, quality standards, and Gujarati-specific style requirements. Quality alignment: Ensure all trainers and QAs apply Gujarati language guidelines consistently and understand updates as projects evolve. Process improvement: Identify recurring quality gaps, propose workflow improvements, and help build scalable QA processes for Gujarati-language projects. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
08/01/2026
Full time
Job description About this role In this hourly, remote contractor role, you will work as a Gujarati Quality Assurance Lead (QAL) to oversee quality, consistency, and trainer performance across Gujarati AI training projects. You will review AI-generated Gujarati content and trainer/QA work, evaluate output quality against project guidelines, provide precise written feedback, and help ensure that all contributors follow the expected standards. You will assess work for accuracy, fluency, grammar, spelling, tone, cultural appropriateness, meaning preservation, instruction-following, formatting, and adherence to project-specific rubrics. You will spot recurring quality issues, communicate updates to trainers and QAs, support onboarding, maintain documentation, and help activate contributors who are not working consistently. This role requires strong Gujarati and English skills, excellent attention to detail, structured communication, and the ability to manage quality workflows across remote teams. This role is with SME Careers, a fast-growing AI Data Services company and subsidiary of SuperAnnotate, delivering training data for many of the world's largest AI companies and foundation-model labs. Your Gujarati quality leadership will directly help improve the world's premier AI models by ensuring that Gujarati training data is natural, accurate, culturally appropriate, well-documented, and aligned with client expectations. Selection process involves an AI interview, a domain-specific task, and an interview with a recruiter. Your profile Bachelor's or Master's degree in Gujarati, Linguistics, Translation, Communications, Journalism, English, Education, Quality Assurance, or a relevant domain/related field. Native or near-native Gujarati proficiency with strong reading and writing skills. Strong grasp of the English language to follow project guidelines, communicate with teams, and provide clear feedback in English. 3 years of professional experience in Gujarati writing, editing, translation, localization, content QA, AI training, education, annotation, or related language-review workflows. Strong understanding of Gujarati grammar, spelling conventions, punctuation, tone, register, and cultural context. Ability to evaluate Gujarati content against detailed rubrics and identify issues such as mistranslation, literal phrasing, unnatural tone, hallucinated claims, ambiguity, or inconsistent terminology. Experience leading or supporting remote teams of trainers, annotators, reviewers, editors, or QAs is strongly preferred. Comfortable working in fast-moving remote environments using tools such as Discord, Google Sheets, Google Docs, trackers, dashboards, and project management systems. Highly detail-oriented and organized, with the ability to maintain style guides, FAQs, trackers, onboarding materials, honeypots, and other quality documentation. Experience with AI training, data annotation, large language models, prompt/response evaluation, or rubric-based LLM QA is a strong plus. Key responsibilities Quality monitoring: Spot-check Gujarati items, identify quality issues, provide ongoing feedback through DMs, and escalate recurring or critical issues. Trainer and QA communication: Update trainers and QAs on Discord about new item guidelines, project changes, workflow updates, and quality expectations. Question handling: Respond to trainer/QA questions clearly and promptly, especially around Gujarati wording, register, translation fidelity, cultural context, and edge cases. Trainer/QA activation management: DM contributors who are inactive or not working, encourage activation, track follow-ups, and flag availability issues when needed. Documentation: Create and maintain Gujarati project documentation, including style guides, trackers, FAQs, quality notes, examples, honeypots, and onboarding materials. Onboarding and training: Schedule and run onboarding/training calls with trainers and QAs to explain project expectations, workflows, rubrics, quality standards, and Gujarati-specific style requirements. Quality alignment: Ensure all trainers and QAs apply Gujarati language guidelines consistently and understand updates as projects evolve. Process improvement: Identify recurring quality gaps, propose workflow improvements, and help build scalable QA processes for Gujarati-language projects. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
Kannada Linguistic QA Specialist (Remote)
Braintrust Boston, Massachusetts
Job description About this role In this hourly, remote contractor role, you will work as a Kannada Quality Assurance Lead (QAL) to oversee quality, consistency, and trainer performance across Kannada AI training projects. You will review AI-generated Kannada content and trainer/QA work, evaluate output quality against project guidelines, provide precise written feedback, and ensure that all contributors follow the expected quality standards. You will assess work for accuracy, fluency, grammar, spelling, tone, cultural appropriateness, meaning preservation, instruction-following, formatting, and adherence to project-specific rubrics. You will spot recurring quality issues, communicate updates to trainers and QAs, support onboarding, maintain documentation, and help activate contributors who are not working consistently. This role requires strong Kannada and English skills, excellent attention to detail, structured communication, and the ability to manage quality workflows across remote teams. This role is with SME Careers, a fast-growing AI Data Services company and subsidiary of SuperAnnotate, delivering training data for many of the world's largest AI companies and foundation-model labs. Your Kannada quality leadership will directly help improve the world's premier AI models by ensuring that Kannada training data is natural, accurate, culturally appropriate, well-documented, and aligned with client expectations. Your profile Bachelor's or Master's degree in Kannada, Linguistics, Translation, Communications, Journalism, English, Education, Quality Assurance, or a relevant domain/related field. Native or near-native Kannada proficiency with strong reading and writing skills. Strong grasp of the English language to follow project guidelines, communicate with teams, and provide clear feedback in English. 3 years of professional experience in Kannada writing, editing, translation, localization, content QA, AI training, education, annotation, or related language-review workflows. Strong understanding of Kannada grammar, spelling conventions, punctuation, tone, register, regional variation, and cultural context. Ability to evaluate Kannada content against detailed rubrics and identify issues such as mistranslation, literal phrasing, unnatural tone, hallucinated claims, ambiguity, or inconsistent terminology. Experience leading or supporting remote teams of trainers, annotators, reviewers, editors, or QAs is strongly preferred. Comfortable working in fast-moving remote environments using tools such as Discord, Google Sheets, Google Docs, trackers, dashboards, and project management systems. Highly detail-oriented and organized, with the ability to maintain style guides, FAQs, trackers, onboarding materials, honeypots, and other quality documentation. Experience with AI training, data annotation, large language models, prompt/response evaluation, or rubric-based LLM QA is a strong plus. Key responsibilities Quality monitoring: Spot-check Kannada items, identify quality issues, provide ongoing feedback through DMs, and escalate recurring or critical issues. Trainer and QA communication: Update trainers and QAs on Discord about new item guidelines, project changes, workflow updates, and quality expectations. Question handling: Respond to trainer/QA questions clearly and promptly, especially around Kannada wording, register, translation fidelity, cultural context, regional variation, and edge cases. Trainer/QA activation management: DM contributors who are inactive or not working, encourage activation, track follow-ups, and flag availability issues when needed. Documentation: Create and maintain Kannada project documentation, including style guides, trackers, FAQs, quality notes, examples, honeypots, and onboarding materials. Onboarding and training: Schedule and run onboarding/training calls with trainers and QAs to explain project expectations, workflows, rubrics, quality standards, and Kannada-specific style requirements. Quality alignment: Ensure all trainers and QAs apply Kannada language guidelines consistently and understand updates as projects evolve. Process improvement: Identify recurring quality gaps, propose workflow improvements, and help build scalable QA processes for Kannada-language projects. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
08/01/2026
Full time
Job description About this role In this hourly, remote contractor role, you will work as a Kannada Quality Assurance Lead (QAL) to oversee quality, consistency, and trainer performance across Kannada AI training projects. You will review AI-generated Kannada content and trainer/QA work, evaluate output quality against project guidelines, provide precise written feedback, and ensure that all contributors follow the expected quality standards. You will assess work for accuracy, fluency, grammar, spelling, tone, cultural appropriateness, meaning preservation, instruction-following, formatting, and adherence to project-specific rubrics. You will spot recurring quality issues, communicate updates to trainers and QAs, support onboarding, maintain documentation, and help activate contributors who are not working consistently. This role requires strong Kannada and English skills, excellent attention to detail, structured communication, and the ability to manage quality workflows across remote teams. This role is with SME Careers, a fast-growing AI Data Services company and subsidiary of SuperAnnotate, delivering training data for many of the world's largest AI companies and foundation-model labs. Your Kannada quality leadership will directly help improve the world's premier AI models by ensuring that Kannada training data is natural, accurate, culturally appropriate, well-documented, and aligned with client expectations. Your profile Bachelor's or Master's degree in Kannada, Linguistics, Translation, Communications, Journalism, English, Education, Quality Assurance, or a relevant domain/related field. Native or near-native Kannada proficiency with strong reading and writing skills. Strong grasp of the English language to follow project guidelines, communicate with teams, and provide clear feedback in English. 3 years of professional experience in Kannada writing, editing, translation, localization, content QA, AI training, education, annotation, or related language-review workflows. Strong understanding of Kannada grammar, spelling conventions, punctuation, tone, register, regional variation, and cultural context. Ability to evaluate Kannada content against detailed rubrics and identify issues such as mistranslation, literal phrasing, unnatural tone, hallucinated claims, ambiguity, or inconsistent terminology. Experience leading or supporting remote teams of trainers, annotators, reviewers, editors, or QAs is strongly preferred. Comfortable working in fast-moving remote environments using tools such as Discord, Google Sheets, Google Docs, trackers, dashboards, and project management systems. Highly detail-oriented and organized, with the ability to maintain style guides, FAQs, trackers, onboarding materials, honeypots, and other quality documentation. Experience with AI training, data annotation, large language models, prompt/response evaluation, or rubric-based LLM QA is a strong plus. Key responsibilities Quality monitoring: Spot-check Kannada items, identify quality issues, provide ongoing feedback through DMs, and escalate recurring or critical issues. Trainer and QA communication: Update trainers and QAs on Discord about new item guidelines, project changes, workflow updates, and quality expectations. Question handling: Respond to trainer/QA questions clearly and promptly, especially around Kannada wording, register, translation fidelity, cultural context, regional variation, and edge cases. Trainer/QA activation management: DM contributors who are inactive or not working, encourage activation, track follow-ups, and flag availability issues when needed. Documentation: Create and maintain Kannada project documentation, including style guides, trackers, FAQs, quality notes, examples, honeypots, and onboarding materials. Onboarding and training: Schedule and run onboarding/training calls with trainers and QAs to explain project expectations, workflows, rubrics, quality standards, and Kannada-specific style requirements. Quality alignment: Ensure all trainers and QAs apply Kannada language guidelines consistently and understand updates as projects evolve. Process improvement: Identify recurring quality gaps, propose workflow improvements, and help build scalable QA processes for Kannada-language projects. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
Kannada Linguistic QA Specialist (Remote)
Braintrust
Job description About this role In this hourly, remote contractor role, you will work as a Kannada Quality Assurance Lead (QAL) to oversee quality, consistency, and trainer performance across Kannada AI training projects. You will review AI-generated Kannada content and trainer/QA work, evaluate output quality against project guidelines, provide precise written feedback, and ensure that all contributors follow the expected quality standards. You will assess work for accuracy, fluency, grammar, spelling, tone, cultural appropriateness, meaning preservation, instruction-following, formatting, and adherence to project-specific rubrics. You will spot recurring quality issues, communicate updates to trainers and QAs, support onboarding, maintain documentation, and help activate contributors who are not working consistently. This role requires strong Kannada and English skills, excellent attention to detail, structured communication, and the ability to manage quality workflows across remote teams. This role is with SME Careers, a fast-growing AI Data Services company and subsidiary of SuperAnnotate, delivering training data for many of the world's largest AI companies and foundation-model labs. Your Kannada quality leadership will directly help improve the world's premier AI models by ensuring that Kannada training data is natural, accurate, culturally appropriate, well-documented, and aligned with client expectations. Your profile Bachelor's or Master's degree in Kannada, Linguistics, Translation, Communications, Journalism, English, Education, Quality Assurance, or a relevant domain/related field. Native or near-native Kannada proficiency with strong reading and writing skills. Strong grasp of the English language to follow project guidelines, communicate with teams, and provide clear feedback in English. 3 years of professional experience in Kannada writing, editing, translation, localization, content QA, AI training, education, annotation, or related language-review workflows. Strong understanding of Kannada grammar, spelling conventions, punctuation, tone, register, regional variation, and cultural context. Ability to evaluate Kannada content against detailed rubrics and identify issues such as mistranslation, literal phrasing, unnatural tone, hallucinated claims, ambiguity, or inconsistent terminology. Experience leading or supporting remote teams of trainers, annotators, reviewers, editors, or QAs is strongly preferred. Comfortable working in fast-moving remote environments using tools such as Discord, Google Sheets, Google Docs, trackers, dashboards, and project management systems. Highly detail-oriented and organized, with the ability to maintain style guides, FAQs, trackers, onboarding materials, honeypots, and other quality documentation. Experience with AI training, data annotation, large language models, prompt/response evaluation, or rubric-based LLM QA is a strong plus. Key responsibilities Quality monitoring: Spot-check Kannada items, identify quality issues, provide ongoing feedback through DMs, and escalate recurring or critical issues. Trainer and QA communication: Update trainers and QAs on Discord about new item guidelines, project changes, workflow updates, and quality expectations. Question handling: Respond to trainer/QA questions clearly and promptly, especially around Kannada wording, register, translation fidelity, cultural context, regional variation, and edge cases. Trainer/QA activation management: DM contributors who are inactive or not working, encourage activation, track follow-ups, and flag availability issues when needed. Documentation: Create and maintain Kannada project documentation, including style guides, trackers, FAQs, quality notes, examples, honeypots, and onboarding materials. Onboarding and training: Schedule and run onboarding/training calls with trainers and QAs to explain project expectations, workflows, rubrics, quality standards, and Kannada-specific style requirements. Quality alignment: Ensure all trainers and QAs apply Kannada language guidelines consistently and understand updates as projects evolve. Process improvement: Identify recurring quality gaps, propose workflow improvements, and help build scalable QA processes for Kannada-language projects. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
08/01/2026
Full time
Job description About this role In this hourly, remote contractor role, you will work as a Kannada Quality Assurance Lead (QAL) to oversee quality, consistency, and trainer performance across Kannada AI training projects. You will review AI-generated Kannada content and trainer/QA work, evaluate output quality against project guidelines, provide precise written feedback, and ensure that all contributors follow the expected quality standards. You will assess work for accuracy, fluency, grammar, spelling, tone, cultural appropriateness, meaning preservation, instruction-following, formatting, and adherence to project-specific rubrics. You will spot recurring quality issues, communicate updates to trainers and QAs, support onboarding, maintain documentation, and help activate contributors who are not working consistently. This role requires strong Kannada and English skills, excellent attention to detail, structured communication, and the ability to manage quality workflows across remote teams. This role is with SME Careers, a fast-growing AI Data Services company and subsidiary of SuperAnnotate, delivering training data for many of the world's largest AI companies and foundation-model labs. Your Kannada quality leadership will directly help improve the world's premier AI models by ensuring that Kannada training data is natural, accurate, culturally appropriate, well-documented, and aligned with client expectations. Your profile Bachelor's or Master's degree in Kannada, Linguistics, Translation, Communications, Journalism, English, Education, Quality Assurance, or a relevant domain/related field. Native or near-native Kannada proficiency with strong reading and writing skills. Strong grasp of the English language to follow project guidelines, communicate with teams, and provide clear feedback in English. 3 years of professional experience in Kannada writing, editing, translation, localization, content QA, AI training, education, annotation, or related language-review workflows. Strong understanding of Kannada grammar, spelling conventions, punctuation, tone, register, regional variation, and cultural context. Ability to evaluate Kannada content against detailed rubrics and identify issues such as mistranslation, literal phrasing, unnatural tone, hallucinated claims, ambiguity, or inconsistent terminology. Experience leading or supporting remote teams of trainers, annotators, reviewers, editors, or QAs is strongly preferred. Comfortable working in fast-moving remote environments using tools such as Discord, Google Sheets, Google Docs, trackers, dashboards, and project management systems. Highly detail-oriented and organized, with the ability to maintain style guides, FAQs, trackers, onboarding materials, honeypots, and other quality documentation. Experience with AI training, data annotation, large language models, prompt/response evaluation, or rubric-based LLM QA is a strong plus. Key responsibilities Quality monitoring: Spot-check Kannada items, identify quality issues, provide ongoing feedback through DMs, and escalate recurring or critical issues. Trainer and QA communication: Update trainers and QAs on Discord about new item guidelines, project changes, workflow updates, and quality expectations. Question handling: Respond to trainer/QA questions clearly and promptly, especially around Kannada wording, register, translation fidelity, cultural context, regional variation, and edge cases. Trainer/QA activation management: DM contributors who are inactive or not working, encourage activation, track follow-ups, and flag availability issues when needed. Documentation: Create and maintain Kannada project documentation, including style guides, trackers, FAQs, quality notes, examples, honeypots, and onboarding materials. Onboarding and training: Schedule and run onboarding/training calls with trainers and QAs to explain project expectations, workflows, rubrics, quality standards, and Kannada-specific style requirements. Quality alignment: Ensure all trainers and QAs apply Kannada language guidelines consistently and understand updates as projects evolve. Process improvement: Identify recurring quality gaps, propose workflow improvements, and help build scalable QA processes for Kannada-language projects. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
Senior Software Engineer - Agent Evaluation - Freelance/Remote 100 openings
Braintrust Boston, Massachusetts
Job description Open to candidates in the North America, South America, Asia and Europe. Please submit your CV in English and indicate your level of English proficiency. Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment. What this opportunity involves We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks. You'll create challenging tasks and evaluation criteria within realistic simulated environments: Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history Design tasks from intermediate states of these environments - craft the prompt, define what "solved" means, and ensure the task is solvable by an AI agent Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust What this is NOT Not data labeling Not prompt engineering Not writing code from scratch - the agent writes most of the code; you guide and evaluate What we look for 5 years in software development Core stack: Python (FastAPI), JavaScript/TypeScript (React), Docker, Postgres, Kafka, Redis Experience writing tests (functional, integration) English proficiency - B2 Why this is hard Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds. How it works Apply Pass qualification(s) Join a project Complete tasks Get paid Effort estimate Tasks for this project are estimated to take 20 hours to complete, depending on complexity. This is an estimate and not a schedule requirement; you choose when and how to work. Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted. Hiring & Onboarding Process The process is designed to move quickly and typically includes the following steps: App review and invitation to a virtual project introduction session (approximately 30 minutes) Platform registration and identity verification Technical assessment (approximately 35 minutes) Background check (completed at no cost to candidates) Onboarding and project-specific training tasks Begin production work! Additional Requirements Willingness to complete identity verification as part of the onboarding process. Ability to complete a technical assessment. Willingness to join and participate in Discord, which will be used for project communication and updates. Successful completion of a background check is required prior to onboarding. Reliable internet connection and ability to communicate effectively in a remote environment. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
08/01/2026
Full time
Job description Open to candidates in the North America, South America, Asia and Europe. Please submit your CV in English and indicate your level of English proficiency. Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment. What this opportunity involves We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks. You'll create challenging tasks and evaluation criteria within realistic simulated environments: Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history Design tasks from intermediate states of these environments - craft the prompt, define what "solved" means, and ensure the task is solvable by an AI agent Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust What this is NOT Not data labeling Not prompt engineering Not writing code from scratch - the agent writes most of the code; you guide and evaluate What we look for 5 years in software development Core stack: Python (FastAPI), JavaScript/TypeScript (React), Docker, Postgres, Kafka, Redis Experience writing tests (functional, integration) English proficiency - B2 Why this is hard Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds. How it works Apply Pass qualification(s) Join a project Complete tasks Get paid Effort estimate Tasks for this project are estimated to take 20 hours to complete, depending on complexity. This is an estimate and not a schedule requirement; you choose when and how to work. Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted. Hiring & Onboarding Process The process is designed to move quickly and typically includes the following steps: App review and invitation to a virtual project introduction session (approximately 30 minutes) Platform registration and identity verification Technical assessment (approximately 35 minutes) Background check (completed at no cost to candidates) Onboarding and project-specific training tasks Begin production work! Additional Requirements Willingness to complete identity verification as part of the onboarding process. Ability to complete a technical assessment. Willingness to join and participate in Discord, which will be used for project communication and updates. Successful completion of a background check is required prior to onboarding. Reliable internet connection and ability to communicate effectively in a remote environment. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
Senior Software Engineer - Agent Evaluation - Freelance/Remote 100 openings
Braintrust Bellville, Texas
Job description Open to candidates in the North America, South America, Asia and Europe. Please submit your CV in English and indicate your level of English proficiency. Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment. What this opportunity involves We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks. You'll create challenging tasks and evaluation criteria within realistic simulated environments: Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history Design tasks from intermediate states of these environments - craft the prompt, define what "solved" means, and ensure the task is solvable by an AI agent Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust What this is NOT Not data labeling Not prompt engineering Not writing code from scratch - the agent writes most of the code; you guide and evaluate What we look for 5 years in software development Core stack: Python (FastAPI), JavaScript/TypeScript (React), Docker, Postgres, Kafka, Redis Experience writing tests (functional, integration) English proficiency - B2 Why this is hard Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds. How it works Apply Pass qualification(s) Join a project Complete tasks Get paid Effort estimate Tasks for this project are estimated to take 20 hours to complete, depending on complexity. This is an estimate and not a schedule requirement; you choose when and how to work. Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted. Hiring & Onboarding Process The process is designed to move quickly and typically includes the following steps: App review and invitation to a virtual project introduction session (approximately 30 minutes) Platform registration and identity verification Technical assessment (approximately 35 minutes) Background check (completed at no cost to candidates) Onboarding and project-specific training tasks Begin production work! Additional Requirements Willingness to complete identity verification as part of the onboarding process. Ability to complete a technical assessment. Willingness to join and participate in Discord, which will be used for project communication and updates. Successful completion of a background check is required prior to onboarding. Reliable internet connection and ability to communicate effectively in a remote environment. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
08/01/2026
Full time
Job description Open to candidates in the North America, South America, Asia and Europe. Please submit your CV in English and indicate your level of English proficiency. Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment. What this opportunity involves We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks. You'll create challenging tasks and evaluation criteria within realistic simulated environments: Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history Design tasks from intermediate states of these environments - craft the prompt, define what "solved" means, and ensure the task is solvable by an AI agent Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust What this is NOT Not data labeling Not prompt engineering Not writing code from scratch - the agent writes most of the code; you guide and evaluate What we look for 5 years in software development Core stack: Python (FastAPI), JavaScript/TypeScript (React), Docker, Postgres, Kafka, Redis Experience writing tests (functional, integration) English proficiency - B2 Why this is hard Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds. How it works Apply Pass qualification(s) Join a project Complete tasks Get paid Effort estimate Tasks for this project are estimated to take 20 hours to complete, depending on complexity. This is an estimate and not a schedule requirement; you choose when and how to work. Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted. Hiring & Onboarding Process The process is designed to move quickly and typically includes the following steps: App review and invitation to a virtual project introduction session (approximately 30 minutes) Platform registration and identity verification Technical assessment (approximately 35 minutes) Background check (completed at no cost to candidates) Onboarding and project-specific training tasks Begin production work! Additional Requirements Willingness to complete identity verification as part of the onboarding process. Ability to complete a technical assessment. Willingness to join and participate in Discord, which will be used for project communication and updates. Successful completion of a background check is required prior to onboarding. Reliable internet connection and ability to communicate effectively in a remote environment. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
Developer - (Bot Developer) Freelance/Remote 100 openings
Braintrust
Job description Open to candidates in the North America, South America, Asia and Europe. Description Mindrift (Toloka) is looking for skilled Bot Developers (WhatsApp Business API, Telegram Bot API, Discord API) to build conversational bots and messaging-platform integrations within our hybrid AI human environment. In this role, as an AI Pilot - that's how we refer to this position at Mindrift - you'll collaborate with Agents that handle repetitive tasks, while you provide bot engineering expertise, conversational design judgment, and quality control to ensure bots are reliable, useful, and ready for real users. This part-time remote opportunity is ideal for professionals with hands-on experience building messaging bots, working with platform APIs and webhooks, and implementing conversational logic. What We Do The Mindrift platform connects specialists with AI projects from major tech innovators. Our mission is to unlock the potential of Generative AI by tapping into real-world expertise from across the globe. About the Role As a Bot Developer, you'll design, build, and refine messaging bots across WhatsApp, Telegram, Discord, and similar platforms - for use cases such as customer service, appointment booking, order taking, content delivery, moderation, and automated notifications. Key Responsibilities Build bots for WhatsApp (Business API / Cloud API), Telegram (Bot API), Discord, and similar messaging platforms. Design and implement conversational flows, dialogue state, and fallback handling. Integrate bots with LLMs (OpenAI, Anthropic, or similar) for natural language responses where appropriate. Connect bots to backend services, databases, CRMs, and third-party APIs (booking systems, payment, content sources). Handle webhooks, rate limits, and platform-specific message formats (interactive messages, buttons, media, templates). Evaluate AI-generated bot code and refactor it for correctness, reliability, and graceful error handling. Implement logging, monitoring, and recovery so bots stay healthy in production. Requirements and benefits Educational qualifications At least 3 years of relevant experience backend, integration, automation, or bot development experience (required). Bachelor's or Master's Degree in Computer Science, Engineering, Information Technology, or related technical fields is a plus. Academic and/or Professional Experience Candidates should have a strong foundation in bot development, messaging platform integrations, and building reliable conversational workflows. We are looking for specialists who can design and maintain production-ready bots, work confidently with APIs, webhooks, and backend services, and refine AI-assisted output into stable, user-friendly experiences. Strong problem-solving skills, attention to detail, and the ability to work independently are essential. Technical Skills (Essential) At least 1 year of hands-on experience building bots for major messaging platforms (WhatsApp, Telegram, Discord, Slack, or similar) is required Strong command of Python or Node.js for backend bot logic. Solid experience with REST APIs, webhooks, OAuth, and async request handling. Experience with relational or NoSQL databases for storing conversation state and user data. Familiarity with LLM APIs (OpenAI, Anthropic) and prompt design for conversational use is a strong plus. Understanding of platform-specific limits, message templates, and approval flows (e.g., WhatsApp template messages). Experience with hosting and deployment (Docker, serverless, VPS, or PaaS) Additional requirements Strong attention to detail and commitment to bot reliability - no silent failures, no broken flows. Self-directed work ethic with the ability to design and ship complete bots independently. Portfolio or examples of bots you've built (required). English proficiency: Upper-intermediate (B2) or above (required). Project time expectations For this project, tasks are estimated to require around 10-20 hours per week during active phases, based on project requirements. This is an estimate, not a guaranteed workload, and applies only while the project is active. Hiring & Onboarding Process The process is designed to move quickly and typically includes the following steps: App review and invitation to a virtual project introduction session (approximately 30 minutes) Platform registration and identity verification Technical assessment (approximately 35 minutes) Background check (completed at no cost to candidates) Onboarding and project-specific training tasks Begin production work! Additional Requirements Willingness to complete identity verification as part of the onboarding process. Ability to complete a technical assessment. Willingness to join and participate in Discord, which will be used for project communication and updates. Successful completion of a background check is required prior to onboarding. Reliable internet connection and ability to communicate effectively in a remote environment. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
08/01/2026
Full time
Job description Open to candidates in the North America, South America, Asia and Europe. Description Mindrift (Toloka) is looking for skilled Bot Developers (WhatsApp Business API, Telegram Bot API, Discord API) to build conversational bots and messaging-platform integrations within our hybrid AI human environment. In this role, as an AI Pilot - that's how we refer to this position at Mindrift - you'll collaborate with Agents that handle repetitive tasks, while you provide bot engineering expertise, conversational design judgment, and quality control to ensure bots are reliable, useful, and ready for real users. This part-time remote opportunity is ideal for professionals with hands-on experience building messaging bots, working with platform APIs and webhooks, and implementing conversational logic. What We Do The Mindrift platform connects specialists with AI projects from major tech innovators. Our mission is to unlock the potential of Generative AI by tapping into real-world expertise from across the globe. About the Role As a Bot Developer, you'll design, build, and refine messaging bots across WhatsApp, Telegram, Discord, and similar platforms - for use cases such as customer service, appointment booking, order taking, content delivery, moderation, and automated notifications. Key Responsibilities Build bots for WhatsApp (Business API / Cloud API), Telegram (Bot API), Discord, and similar messaging platforms. Design and implement conversational flows, dialogue state, and fallback handling. Integrate bots with LLMs (OpenAI, Anthropic, or similar) for natural language responses where appropriate. Connect bots to backend services, databases, CRMs, and third-party APIs (booking systems, payment, content sources). Handle webhooks, rate limits, and platform-specific message formats (interactive messages, buttons, media, templates). Evaluate AI-generated bot code and refactor it for correctness, reliability, and graceful error handling. Implement logging, monitoring, and recovery so bots stay healthy in production. Requirements and benefits Educational qualifications At least 3 years of relevant experience backend, integration, automation, or bot development experience (required). Bachelor's or Master's Degree in Computer Science, Engineering, Information Technology, or related technical fields is a plus. Academic and/or Professional Experience Candidates should have a strong foundation in bot development, messaging platform integrations, and building reliable conversational workflows. We are looking for specialists who can design and maintain production-ready bots, work confidently with APIs, webhooks, and backend services, and refine AI-assisted output into stable, user-friendly experiences. Strong problem-solving skills, attention to detail, and the ability to work independently are essential. Technical Skills (Essential) At least 1 year of hands-on experience building bots for major messaging platforms (WhatsApp, Telegram, Discord, Slack, or similar) is required Strong command of Python or Node.js for backend bot logic. Solid experience with REST APIs, webhooks, OAuth, and async request handling. Experience with relational or NoSQL databases for storing conversation state and user data. Familiarity with LLM APIs (OpenAI, Anthropic) and prompt design for conversational use is a strong plus. Understanding of platform-specific limits, message templates, and approval flows (e.g., WhatsApp template messages). Experience with hosting and deployment (Docker, serverless, VPS, or PaaS) Additional requirements Strong attention to detail and commitment to bot reliability - no silent failures, no broken flows. Self-directed work ethic with the ability to design and ship complete bots independently. Portfolio or examples of bots you've built (required). English proficiency: Upper-intermediate (B2) or above (required). Project time expectations For this project, tasks are estimated to require around 10-20 hours per week during active phases, based on project requirements. This is an estimate, not a guaranteed workload, and applies only while the project is active. Hiring & Onboarding Process The process is designed to move quickly and typically includes the following steps: App review and invitation to a virtual project introduction session (approximately 30 minutes) Platform registration and identity verification Technical assessment (approximately 35 minutes) Background check (completed at no cost to candidates) Onboarding and project-specific training tasks Begin production work! Additional Requirements Willingness to complete identity verification as part of the onboarding process. Ability to complete a technical assessment. Willingness to join and participate in Discord, which will be used for project communication and updates. Successful completion of a background check is required prior to onboarding. Reliable internet connection and ability to communicate effectively in a remote environment. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
Senior Software Engineer (Python) - Agent Evaluation - Freelance/Remote 100 openings
Braintrust Bellville, Texas
Job description This opportunity is intended for experienced Senior Python Engineers only Open to candidates in the North America, South America, Asia and Europe. Please submit your CV in English and indicate your level of English proficiency. Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment. What this opportunity involves We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks. You'll create challenging tasks and evaluation criteria within realistic simulated environments: Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history Design tasks from intermediate states of these environments - craft the prompt, define what "solved" means, and ensure the task is solvable by an AI agent Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust What this is NOT Not data labeling Not prompt engineering Not writing code from scratch - the agent writes most of the code; you guide and evaluate What we look for - You must meet all the requirements in order to be considered for this project: 5 years of professional experience with Python. Strong experience with FastAPI , pytest , and async/await . Hands-on experience with Docker , PostgreSQL , and CI/CD pipelines. Proven experience writing and maintaining automated tests (not just executing them). Full-stack experience with React and TypeScript is a plus. English proficiency at B2 level or higher . Availability to work 30 hours per week . Why this is hard Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds. How it works Apply Pass qualification(s) Join a project Complete tasks Get paid Hiring & Onboarding Process The process is designed to move quickly and typically includes the following steps: App review and invitation to a virtual project introduction session (approximately 30 minutes) Platform registration and identity verification Technical assessment (approximately 35 minutes) Background check (completed at no cost to candidates) Onboarding and project-specific training tasks Begin production work! Additional Requirements Willingness to complete identity verification as part of the onboarding process. Ability to complete a technical assessment. Willingness to join and participate in Discord, which will be used for project communication and updates. Successful completion of a background check is required prior to onboarding. Reliable internet connection and ability to communicate effectively in a remote environment. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
08/01/2026
Full time
Job description This opportunity is intended for experienced Senior Python Engineers only Open to candidates in the North America, South America, Asia and Europe. Please submit your CV in English and indicate your level of English proficiency. Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment. What this opportunity involves We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks. You'll create challenging tasks and evaluation criteria within realistic simulated environments: Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history Design tasks from intermediate states of these environments - craft the prompt, define what "solved" means, and ensure the task is solvable by an AI agent Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust What this is NOT Not data labeling Not prompt engineering Not writing code from scratch - the agent writes most of the code; you guide and evaluate What we look for - You must meet all the requirements in order to be considered for this project: 5 years of professional experience with Python. Strong experience with FastAPI , pytest , and async/await . Hands-on experience with Docker , PostgreSQL , and CI/CD pipelines. Proven experience writing and maintaining automated tests (not just executing them). Full-stack experience with React and TypeScript is a plus. English proficiency at B2 level or higher . Availability to work 30 hours per week . Why this is hard Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds. How it works Apply Pass qualification(s) Join a project Complete tasks Get paid Hiring & Onboarding Process The process is designed to move quickly and typically includes the following steps: App review and invitation to a virtual project introduction session (approximately 30 minutes) Platform registration and identity verification Technical assessment (approximately 35 minutes) Background check (completed at no cost to candidates) Onboarding and project-specific training tasks Begin production work! Additional Requirements Willingness to complete identity verification as part of the onboarding process. Ability to complete a technical assessment. Willingness to join and participate in Discord, which will be used for project communication and updates. Successful completion of a background check is required prior to onboarding. Reliable internet connection and ability to communicate effectively in a remote environment. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
Senior Software Engineer - Agent Evaluation - Freelance/Remote 100 openings
Braintrust
Job description Open to candidates in the North America, South America, Asia and Europe. Please submit your CV in English and indicate your level of English proficiency. Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment. What this opportunity involves We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks. You'll create challenging tasks and evaluation criteria within realistic simulated environments: Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history Design tasks from intermediate states of these environments - craft the prompt, define what "solved" means, and ensure the task is solvable by an AI agent Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust What this is NOT Not data labeling Not prompt engineering Not writing code from scratch - the agent writes most of the code; you guide and evaluate What we look for 5 years in software development Core stack: Python (FastAPI), JavaScript/TypeScript (React), Docker, Postgres, Kafka, Redis Experience writing tests (functional, integration) English proficiency - B2 Why this is hard Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds. How it works Apply Pass qualification(s) Join a project Complete tasks Get paid Effort estimate Tasks for this project are estimated to take 20 hours to complete, depending on complexity. This is an estimate and not a schedule requirement; you choose when and how to work. Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted. Hiring & Onboarding Process The process is designed to move quickly and typically includes the following steps: App review and invitation to a virtual project introduction session (approximately 30 minutes) Platform registration and identity verification Technical assessment (approximately 35 minutes) Background check (completed at no cost to candidates) Onboarding and project-specific training tasks Begin production work! Additional Requirements Willingness to complete identity verification as part of the onboarding process. Ability to complete a technical assessment. Willingness to join and participate in Discord, which will be used for project communication and updates. Successful completion of a background check is required prior to onboarding. Reliable internet connection and ability to communicate effectively in a remote environment. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
08/01/2026
Full time
Job description Open to candidates in the North America, South America, Asia and Europe. Please submit your CV in English and indicate your level of English proficiency. Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment. What this opportunity involves We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks. You'll create challenging tasks and evaluation criteria within realistic simulated environments: Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history Design tasks from intermediate states of these environments - craft the prompt, define what "solved" means, and ensure the task is solvable by an AI agent Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust What this is NOT Not data labeling Not prompt engineering Not writing code from scratch - the agent writes most of the code; you guide and evaluate What we look for 5 years in software development Core stack: Python (FastAPI), JavaScript/TypeScript (React), Docker, Postgres, Kafka, Redis Experience writing tests (functional, integration) English proficiency - B2 Why this is hard Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds. How it works Apply Pass qualification(s) Join a project Complete tasks Get paid Effort estimate Tasks for this project are estimated to take 20 hours to complete, depending on complexity. This is an estimate and not a schedule requirement; you choose when and how to work. Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted. Hiring & Onboarding Process The process is designed to move quickly and typically includes the following steps: App review and invitation to a virtual project introduction session (approximately 30 minutes) Platform registration and identity verification Technical assessment (approximately 35 minutes) Background check (completed at no cost to candidates) Onboarding and project-specific training tasks Begin production work! Additional Requirements Willingness to complete identity verification as part of the onboarding process. Ability to complete a technical assessment. Willingness to join and participate in Discord, which will be used for project communication and updates. Successful completion of a background check is required prior to onboarding. Reliable internet connection and ability to communicate effectively in a remote environment. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
Senior Software Engineer (Python) - Agent Evaluation - Freelance/Remote 100 openings
Braintrust Seattle, Washington
Job description This opportunity is intended for experienced Senior Python Engineers only Open to candidates in the North America, South America, Asia and Europe. Please submit your CV in English and indicate your level of English proficiency. Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment. What this opportunity involves We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks. You'll create challenging tasks and evaluation criteria within realistic simulated environments: Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history Design tasks from intermediate states of these environments - craft the prompt, define what "solved" means, and ensure the task is solvable by an AI agent Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust What this is NOT Not data labeling Not prompt engineering Not writing code from scratch - the agent writes most of the code; you guide and evaluate What we look for - You must meet all the requirements in order to be considered for this project: 5 years of professional experience with Python. Strong experience with FastAPI , pytest , and async/await . Hands-on experience with Docker , PostgreSQL , and CI/CD pipelines. Proven experience writing and maintaining automated tests (not just executing them). Full-stack experience with React and TypeScript is a plus. English proficiency at B2 level or higher . Availability to work 30 hours per week . Why this is hard Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds. How it works Apply Pass qualification(s) Join a project Complete tasks Get paid Hiring & Onboarding Process The process is designed to move quickly and typically includes the following steps: App review and invitation to a virtual project introduction session (approximately 30 minutes) Platform registration and identity verification Technical assessment (approximately 35 minutes) Background check (completed at no cost to candidates) Onboarding and project-specific training tasks Begin production work! Additional Requirements Willingness to complete identity verification as part of the onboarding process. Ability to complete a technical assessment. Willingness to join and participate in Discord, which will be used for project communication and updates. Successful completion of a background check is required prior to onboarding. Reliable internet connection and ability to communicate effectively in a remote environment. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
08/01/2026
Full time
Job description This opportunity is intended for experienced Senior Python Engineers only Open to candidates in the North America, South America, Asia and Europe. Please submit your CV in English and indicate your level of English proficiency. Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment. What this opportunity involves We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks. You'll create challenging tasks and evaluation criteria within realistic simulated environments: Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history Design tasks from intermediate states of these environments - craft the prompt, define what "solved" means, and ensure the task is solvable by an AI agent Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust What this is NOT Not data labeling Not prompt engineering Not writing code from scratch - the agent writes most of the code; you guide and evaluate What we look for - You must meet all the requirements in order to be considered for this project: 5 years of professional experience with Python. Strong experience with FastAPI , pytest , and async/await . Hands-on experience with Docker , PostgreSQL , and CI/CD pipelines. Proven experience writing and maintaining automated tests (not just executing them). Full-stack experience with React and TypeScript is a plus. English proficiency at B2 level or higher . Availability to work 30 hours per week . Why this is hard Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds. How it works Apply Pass qualification(s) Join a project Complete tasks Get paid Hiring & Onboarding Process The process is designed to move quickly and typically includes the following steps: App review and invitation to a virtual project introduction session (approximately 30 minutes) Platform registration and identity verification Technical assessment (approximately 35 minutes) Background check (completed at no cost to candidates) Onboarding and project-specific training tasks Begin production work! Additional Requirements Willingness to complete identity verification as part of the onboarding process. Ability to complete a technical assessment. Willingness to join and participate in Discord, which will be used for project communication and updates. Successful completion of a background check is required prior to onboarding. Reliable internet connection and ability to communicate effectively in a remote environment. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
Senior Software Engineer (Python) - Agent Evaluation - Freelance/Remote 100 openings
Braintrust
Job description This opportunity is intended for experienced Senior Python Engineers only Open to candidates in the North America, South America, Asia and Europe. Please submit your CV in English and indicate your level of English proficiency. Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment. What this opportunity involves We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks. You'll create challenging tasks and evaluation criteria within realistic simulated environments: Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history Design tasks from intermediate states of these environments - craft the prompt, define what "solved" means, and ensure the task is solvable by an AI agent Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust What this is NOT Not data labeling Not prompt engineering Not writing code from scratch - the agent writes most of the code; you guide and evaluate What we look for - You must meet all the requirements in order to be considered for this project: 5 years of professional experience with Python. Strong experience with FastAPI , pytest , and async/await . Hands-on experience with Docker , PostgreSQL , and CI/CD pipelines. Proven experience writing and maintaining automated tests (not just executing them). Full-stack experience with React and TypeScript is a plus. English proficiency at B2 level or higher . Availability to work 30 hours per week . Why this is hard Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds. How it works Apply Pass qualification(s) Join a project Complete tasks Get paid Hiring & Onboarding Process The process is designed to move quickly and typically includes the following steps: App review and invitation to a virtual project introduction session (approximately 30 minutes) Platform registration and identity verification Technical assessment (approximately 35 minutes) Background check (completed at no cost to candidates) Onboarding and project-specific training tasks Begin production work! Additional Requirements Willingness to complete identity verification as part of the onboarding process. Ability to complete a technical assessment. Willingness to join and participate in Discord, which will be used for project communication and updates. Successful completion of a background check is required prior to onboarding. Reliable internet connection and ability to communicate effectively in a remote environment. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
08/01/2026
Full time
Job description This opportunity is intended for experienced Senior Python Engineers only Open to candidates in the North America, South America, Asia and Europe. Please submit your CV in English and indicate your level of English proficiency. Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment. What this opportunity involves We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks. You'll create challenging tasks and evaluation criteria within realistic simulated environments: Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history Design tasks from intermediate states of these environments - craft the prompt, define what "solved" means, and ensure the task is solvable by an AI agent Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust What this is NOT Not data labeling Not prompt engineering Not writing code from scratch - the agent writes most of the code; you guide and evaluate What we look for - You must meet all the requirements in order to be considered for this project: 5 years of professional experience with Python. Strong experience with FastAPI , pytest , and async/await . Hands-on experience with Docker , PostgreSQL , and CI/CD pipelines. Proven experience writing and maintaining automated tests (not just executing them). Full-stack experience with React and TypeScript is a plus. English proficiency at B2 level or higher . Availability to work 30 hours per week . Why this is hard Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds. How it works Apply Pass qualification(s) Join a project Complete tasks Get paid Hiring & Onboarding Process The process is designed to move quickly and typically includes the following steps: App review and invitation to a virtual project introduction session (approximately 30 minutes) Platform registration and identity verification Technical assessment (approximately 35 minutes) Background check (completed at no cost to candidates) Onboarding and project-specific training tasks Begin production work! Additional Requirements Willingness to complete identity verification as part of the onboarding process. Ability to complete a technical assessment. Willingness to join and participate in Discord, which will be used for project communication and updates. Successful completion of a background check is required prior to onboarding. Reliable internet connection and ability to communicate effectively in a remote environment. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
Telugu Linguistic QA Specialist (Remote)
Braintrust Boston, Massachusetts
Job description About this role In this hourly, remote contractor role, you will work as a Telugu Quality Assurance Lead (QAL) to oversee quality, consistency, and trainer performance across Telugu AI training projects. You will review AI-generated Telugu content and trainer/QA work, evaluate output quality against project guidelines, provide precise written feedback, and ensure that all contributors follow the expected quality standards. You will assess work for accuracy, fluency, grammar, spelling, tone, cultural appropriateness, meaning preservation, instruction-following, formatting, and adherence to project-specific rubrics. You will spot recurring quality issues, communicate updates to trainers and QAs, support onboarding, maintain documentation, and help activate contributors who are not working consistently. This role requires strong Telugu and English skills, excellent attention to detail, structured communication, and the ability to manage quality workflows across remote teams. This role is with SME Careers, a fast-growing AI Data Services company and subsidiary of SuperAnnotate, delivering training data for many of the world's largest AI companies and foundation-model labs. Your Telugu quality leadership will directly help improve the world's premier AI models by ensuring that Telugu training data is natural, accurate, culturally appropriate, well-documented, and aligned with client expectations. Selection process involves an AI interview, a domain-specific task, and an interview with a recruiter. Your profile Bachelor's or Master's degree in Telugu, Linguistics, Translation, Communications, Journalism, English, Education, Quality Assurance, or a relevant domain/related field. Native or near-native Telugu proficiency with strong reading and writing skills. Strong grasp of the English language to follow project guidelines, communicate with teams, and provide clear feedback in English. 3 years of professional experience in Telugu writing, editing, translation, localization, content QA, AI training, education, annotation, or related language-review workflows. Strong understanding of Telugu grammar, spelling conventions, punctuation, tone, register, regional variation, and cultural context. Ability to evaluate Telugu content against detailed rubrics and identify issues such as mistranslation, literal phrasing, unnatural tone, hallucinated claims, ambiguity, or inconsistent terminology. Experience leading or supporting remote teams of trainers, annotators, reviewers, editors, or QAs is strongly preferred. Comfortable working in fast-moving remote environments using tools such as Discord, Google Sheets, Google Docs, trackers, dashboards, and project management systems. Highly detail-oriented and organized, with the ability to maintain style guides, FAQs, trackers, onboarding materials, honeypots, and other quality documentation. Experience with AI training, data annotation, large language models, prompt/response evaluation, or rubric-based LLM QA is a strong plus. Key responsibilities Quality monitoring: Spot-check Telugu items, identify quality issues, provide ongoing feedback through DMs, and escalate recurring or critical issues. Trainer and QA communication: Update trainers and QAs on Discord about new item guidelines, project changes, workflow updates, and quality expectations. Question handling: Respond to trainer/QA questions clearly and promptly, especially around Telugu wording, register, translation fidelity, cultural context, Andhra Pradesh vs Telangana usage, and edge cases. Trainer/QA activation management: DM contributors who are inactive or not working, encourage activation, track follow-ups, and flag availability issues when needed. Documentation: Create and maintain Telugu project documentation, including style guides, trackers, FAQs, quality notes, examples, honeypots, and onboarding materials. Onboarding and training: Schedule and run onboarding/training calls with trainers and QAs to explain project expectations, workflows, rubrics, quality standards, and Telugu-specific style requirements. Quality alignment: Ensure all trainers and QAs apply Telugu language guidelines consistently and understand updates as projects evolve. Process improvement: Identify recurring quality gaps, propose workflow improvements, and help build scalable QA processes for Telugu-language projects. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
08/01/2026
Full time
Job description About this role In this hourly, remote contractor role, you will work as a Telugu Quality Assurance Lead (QAL) to oversee quality, consistency, and trainer performance across Telugu AI training projects. You will review AI-generated Telugu content and trainer/QA work, evaluate output quality against project guidelines, provide precise written feedback, and ensure that all contributors follow the expected quality standards. You will assess work for accuracy, fluency, grammar, spelling, tone, cultural appropriateness, meaning preservation, instruction-following, formatting, and adherence to project-specific rubrics. You will spot recurring quality issues, communicate updates to trainers and QAs, support onboarding, maintain documentation, and help activate contributors who are not working consistently. This role requires strong Telugu and English skills, excellent attention to detail, structured communication, and the ability to manage quality workflows across remote teams. This role is with SME Careers, a fast-growing AI Data Services company and subsidiary of SuperAnnotate, delivering training data for many of the world's largest AI companies and foundation-model labs. Your Telugu quality leadership will directly help improve the world's premier AI models by ensuring that Telugu training data is natural, accurate, culturally appropriate, well-documented, and aligned with client expectations. Selection process involves an AI interview, a domain-specific task, and an interview with a recruiter. Your profile Bachelor's or Master's degree in Telugu, Linguistics, Translation, Communications, Journalism, English, Education, Quality Assurance, or a relevant domain/related field. Native or near-native Telugu proficiency with strong reading and writing skills. Strong grasp of the English language to follow project guidelines, communicate with teams, and provide clear feedback in English. 3 years of professional experience in Telugu writing, editing, translation, localization, content QA, AI training, education, annotation, or related language-review workflows. Strong understanding of Telugu grammar, spelling conventions, punctuation, tone, register, regional variation, and cultural context. Ability to evaluate Telugu content against detailed rubrics and identify issues such as mistranslation, literal phrasing, unnatural tone, hallucinated claims, ambiguity, or inconsistent terminology. Experience leading or supporting remote teams of trainers, annotators, reviewers, editors, or QAs is strongly preferred. Comfortable working in fast-moving remote environments using tools such as Discord, Google Sheets, Google Docs, trackers, dashboards, and project management systems. Highly detail-oriented and organized, with the ability to maintain style guides, FAQs, trackers, onboarding materials, honeypots, and other quality documentation. Experience with AI training, data annotation, large language models, prompt/response evaluation, or rubric-based LLM QA is a strong plus. Key responsibilities Quality monitoring: Spot-check Telugu items, identify quality issues, provide ongoing feedback through DMs, and escalate recurring or critical issues. Trainer and QA communication: Update trainers and QAs on Discord about new item guidelines, project changes, workflow updates, and quality expectations. Question handling: Respond to trainer/QA questions clearly and promptly, especially around Telugu wording, register, translation fidelity, cultural context, Andhra Pradesh vs Telangana usage, and edge cases. Trainer/QA activation management: DM contributors who are inactive or not working, encourage activation, track follow-ups, and flag availability issues when needed. Documentation: Create and maintain Telugu project documentation, including style guides, trackers, FAQs, quality notes, examples, honeypots, and onboarding materials. Onboarding and training: Schedule and run onboarding/training calls with trainers and QAs to explain project expectations, workflows, rubrics, quality standards, and Telugu-specific style requirements. Quality alignment: Ensure all trainers and QAs apply Telugu language guidelines consistently and understand updates as projects evolve. Process improvement: Identify recurring quality gaps, propose workflow improvements, and help build scalable QA processes for Telugu-language projects. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
Telugu Linguistic QA Specialist (Remote)
Braintrust
Job description About this role In this hourly, remote contractor role, you will work as a Telugu Quality Assurance Lead (QAL) to oversee quality, consistency, and trainer performance across Telugu AI training projects. You will review AI-generated Telugu content and trainer/QA work, evaluate output quality against project guidelines, provide precise written feedback, and ensure that all contributors follow the expected quality standards. You will assess work for accuracy, fluency, grammar, spelling, tone, cultural appropriateness, meaning preservation, instruction-following, formatting, and adherence to project-specific rubrics. You will spot recurring quality issues, communicate updates to trainers and QAs, support onboarding, maintain documentation, and help activate contributors who are not working consistently. This role requires strong Telugu and English skills, excellent attention to detail, structured communication, and the ability to manage quality workflows across remote teams. This role is with SME Careers, a fast-growing AI Data Services company and subsidiary of SuperAnnotate, delivering training data for many of the world's largest AI companies and foundation-model labs. Your Telugu quality leadership will directly help improve the world's premier AI models by ensuring that Telugu training data is natural, accurate, culturally appropriate, well-documented, and aligned with client expectations. Selection process involves an AI interview, a domain-specific task, and an interview with a recruiter. Your profile Bachelor's or Master's degree in Telugu, Linguistics, Translation, Communications, Journalism, English, Education, Quality Assurance, or a relevant domain/related field. Native or near-native Telugu proficiency with strong reading and writing skills. Strong grasp of the English language to follow project guidelines, communicate with teams, and provide clear feedback in English. 3 years of professional experience in Telugu writing, editing, translation, localization, content QA, AI training, education, annotation, or related language-review workflows. Strong understanding of Telugu grammar, spelling conventions, punctuation, tone, register, regional variation, and cultural context. Ability to evaluate Telugu content against detailed rubrics and identify issues such as mistranslation, literal phrasing, unnatural tone, hallucinated claims, ambiguity, or inconsistent terminology. Experience leading or supporting remote teams of trainers, annotators, reviewers, editors, or QAs is strongly preferred. Comfortable working in fast-moving remote environments using tools such as Discord, Google Sheets, Google Docs, trackers, dashboards, and project management systems. Highly detail-oriented and organized, with the ability to maintain style guides, FAQs, trackers, onboarding materials, honeypots, and other quality documentation. Experience with AI training, data annotation, large language models, prompt/response evaluation, or rubric-based LLM QA is a strong plus. Key responsibilities Quality monitoring: Spot-check Telugu items, identify quality issues, provide ongoing feedback through DMs, and escalate recurring or critical issues. Trainer and QA communication: Update trainers and QAs on Discord about new item guidelines, project changes, workflow updates, and quality expectations. Question handling: Respond to trainer/QA questions clearly and promptly, especially around Telugu wording, register, translation fidelity, cultural context, Andhra Pradesh vs Telangana usage, and edge cases. Trainer/QA activation management: DM contributors who are inactive or not working, encourage activation, track follow-ups, and flag availability issues when needed. Documentation: Create and maintain Telugu project documentation, including style guides, trackers, FAQs, quality notes, examples, honeypots, and onboarding materials. Onboarding and training: Schedule and run onboarding/training calls with trainers and QAs to explain project expectations, workflows, rubrics, quality standards, and Telugu-specific style requirements. Quality alignment: Ensure all trainers and QAs apply Telugu language guidelines consistently and understand updates as projects evolve. Process improvement: Identify recurring quality gaps, propose workflow improvements, and help build scalable QA processes for Telugu-language projects. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
08/01/2026
Full time
Job description About this role In this hourly, remote contractor role, you will work as a Telugu Quality Assurance Lead (QAL) to oversee quality, consistency, and trainer performance across Telugu AI training projects. You will review AI-generated Telugu content and trainer/QA work, evaluate output quality against project guidelines, provide precise written feedback, and ensure that all contributors follow the expected quality standards. You will assess work for accuracy, fluency, grammar, spelling, tone, cultural appropriateness, meaning preservation, instruction-following, formatting, and adherence to project-specific rubrics. You will spot recurring quality issues, communicate updates to trainers and QAs, support onboarding, maintain documentation, and help activate contributors who are not working consistently. This role requires strong Telugu and English skills, excellent attention to detail, structured communication, and the ability to manage quality workflows across remote teams. This role is with SME Careers, a fast-growing AI Data Services company and subsidiary of SuperAnnotate, delivering training data for many of the world's largest AI companies and foundation-model labs. Your Telugu quality leadership will directly help improve the world's premier AI models by ensuring that Telugu training data is natural, accurate, culturally appropriate, well-documented, and aligned with client expectations. Selection process involves an AI interview, a domain-specific task, and an interview with a recruiter. Your profile Bachelor's or Master's degree in Telugu, Linguistics, Translation, Communications, Journalism, English, Education, Quality Assurance, or a relevant domain/related field. Native or near-native Telugu proficiency with strong reading and writing skills. Strong grasp of the English language to follow project guidelines, communicate with teams, and provide clear feedback in English. 3 years of professional experience in Telugu writing, editing, translation, localization, content QA, AI training, education, annotation, or related language-review workflows. Strong understanding of Telugu grammar, spelling conventions, punctuation, tone, register, regional variation, and cultural context. Ability to evaluate Telugu content against detailed rubrics and identify issues such as mistranslation, literal phrasing, unnatural tone, hallucinated claims, ambiguity, or inconsistent terminology. Experience leading or supporting remote teams of trainers, annotators, reviewers, editors, or QAs is strongly preferred. Comfortable working in fast-moving remote environments using tools such as Discord, Google Sheets, Google Docs, trackers, dashboards, and project management systems. Highly detail-oriented and organized, with the ability to maintain style guides, FAQs, trackers, onboarding materials, honeypots, and other quality documentation. Experience with AI training, data annotation, large language models, prompt/response evaluation, or rubric-based LLM QA is a strong plus. Key responsibilities Quality monitoring: Spot-check Telugu items, identify quality issues, provide ongoing feedback through DMs, and escalate recurring or critical issues. Trainer and QA communication: Update trainers and QAs on Discord about new item guidelines, project changes, workflow updates, and quality expectations. Question handling: Respond to trainer/QA questions clearly and promptly, especially around Telugu wording, register, translation fidelity, cultural context, Andhra Pradesh vs Telangana usage, and edge cases. Trainer/QA activation management: DM contributors who are inactive or not working, encourage activation, track follow-ups, and flag availability issues when needed. Documentation: Create and maintain Telugu project documentation, including style guides, trackers, FAQs, quality notes, examples, honeypots, and onboarding materials. Onboarding and training: Schedule and run onboarding/training calls with trainers and QAs to explain project expectations, workflows, rubrics, quality standards, and Telugu-specific style requirements. Quality alignment: Ensure all trainers and QAs apply Telugu language guidelines consistently and understand updates as projects evolve. Process improvement: Identify recurring quality gaps, propose workflow improvements, and help build scalable QA processes for Telugu-language projects. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
Telugu Linguistic QA Specialist (Remote)
Braintrust Bellville, Texas
Job description About this role In this hourly, remote contractor role, you will work as a Telugu Quality Assurance Lead (QAL) to oversee quality, consistency, and trainer performance across Telugu AI training projects. You will review AI-generated Telugu content and trainer/QA work, evaluate output quality against project guidelines, provide precise written feedback, and ensure that all contributors follow the expected quality standards. You will assess work for accuracy, fluency, grammar, spelling, tone, cultural appropriateness, meaning preservation, instruction-following, formatting, and adherence to project-specific rubrics. You will spot recurring quality issues, communicate updates to trainers and QAs, support onboarding, maintain documentation, and help activate contributors who are not working consistently. This role requires strong Telugu and English skills, excellent attention to detail, structured communication, and the ability to manage quality workflows across remote teams. This role is with SME Careers, a fast-growing AI Data Services company and subsidiary of SuperAnnotate, delivering training data for many of the world's largest AI companies and foundation-model labs. Your Telugu quality leadership will directly help improve the world's premier AI models by ensuring that Telugu training data is natural, accurate, culturally appropriate, well-documented, and aligned with client expectations. Selection process involves an AI interview, a domain-specific task, and an interview with a recruiter. Your profile Bachelor's or Master's degree in Telugu, Linguistics, Translation, Communications, Journalism, English, Education, Quality Assurance, or a relevant domain/related field. Native or near-native Telugu proficiency with strong reading and writing skills. Strong grasp of the English language to follow project guidelines, communicate with teams, and provide clear feedback in English. 3 years of professional experience in Telugu writing, editing, translation, localization, content QA, AI training, education, annotation, or related language-review workflows. Strong understanding of Telugu grammar, spelling conventions, punctuation, tone, register, regional variation, and cultural context. Ability to evaluate Telugu content against detailed rubrics and identify issues such as mistranslation, literal phrasing, unnatural tone, hallucinated claims, ambiguity, or inconsistent terminology. Experience leading or supporting remote teams of trainers, annotators, reviewers, editors, or QAs is strongly preferred. Comfortable working in fast-moving remote environments using tools such as Discord, Google Sheets, Google Docs, trackers, dashboards, and project management systems. Highly detail-oriented and organized, with the ability to maintain style guides, FAQs, trackers, onboarding materials, honeypots, and other quality documentation. Experience with AI training, data annotation, large language models, prompt/response evaluation, or rubric-based LLM QA is a strong plus. Key responsibilities Quality monitoring: Spot-check Telugu items, identify quality issues, provide ongoing feedback through DMs, and escalate recurring or critical issues. Trainer and QA communication: Update trainers and QAs on Discord about new item guidelines, project changes, workflow updates, and quality expectations. Question handling: Respond to trainer/QA questions clearly and promptly, especially around Telugu wording, register, translation fidelity, cultural context, Andhra Pradesh vs Telangana usage, and edge cases. Trainer/QA activation management: DM contributors who are inactive or not working, encourage activation, track follow-ups, and flag availability issues when needed. Documentation: Create and maintain Telugu project documentation, including style guides, trackers, FAQs, quality notes, examples, honeypots, and onboarding materials. Onboarding and training: Schedule and run onboarding/training calls with trainers and QAs to explain project expectations, workflows, rubrics, quality standards, and Telugu-specific style requirements. Quality alignment: Ensure all trainers and QAs apply Telugu language guidelines consistently and understand updates as projects evolve. Process improvement: Identify recurring quality gaps, propose workflow improvements, and help build scalable QA processes for Telugu-language projects. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
08/01/2026
Full time
Job description About this role In this hourly, remote contractor role, you will work as a Telugu Quality Assurance Lead (QAL) to oversee quality, consistency, and trainer performance across Telugu AI training projects. You will review AI-generated Telugu content and trainer/QA work, evaluate output quality against project guidelines, provide precise written feedback, and ensure that all contributors follow the expected quality standards. You will assess work for accuracy, fluency, grammar, spelling, tone, cultural appropriateness, meaning preservation, instruction-following, formatting, and adherence to project-specific rubrics. You will spot recurring quality issues, communicate updates to trainers and QAs, support onboarding, maintain documentation, and help activate contributors who are not working consistently. This role requires strong Telugu and English skills, excellent attention to detail, structured communication, and the ability to manage quality workflows across remote teams. This role is with SME Careers, a fast-growing AI Data Services company and subsidiary of SuperAnnotate, delivering training data for many of the world's largest AI companies and foundation-model labs. Your Telugu quality leadership will directly help improve the world's premier AI models by ensuring that Telugu training data is natural, accurate, culturally appropriate, well-documented, and aligned with client expectations. Selection process involves an AI interview, a domain-specific task, and an interview with a recruiter. Your profile Bachelor's or Master's degree in Telugu, Linguistics, Translation, Communications, Journalism, English, Education, Quality Assurance, or a relevant domain/related field. Native or near-native Telugu proficiency with strong reading and writing skills. Strong grasp of the English language to follow project guidelines, communicate with teams, and provide clear feedback in English. 3 years of professional experience in Telugu writing, editing, translation, localization, content QA, AI training, education, annotation, or related language-review workflows. Strong understanding of Telugu grammar, spelling conventions, punctuation, tone, register, regional variation, and cultural context. Ability to evaluate Telugu content against detailed rubrics and identify issues such as mistranslation, literal phrasing, unnatural tone, hallucinated claims, ambiguity, or inconsistent terminology. Experience leading or supporting remote teams of trainers, annotators, reviewers, editors, or QAs is strongly preferred. Comfortable working in fast-moving remote environments using tools such as Discord, Google Sheets, Google Docs, trackers, dashboards, and project management systems. Highly detail-oriented and organized, with the ability to maintain style guides, FAQs, trackers, onboarding materials, honeypots, and other quality documentation. Experience with AI training, data annotation, large language models, prompt/response evaluation, or rubric-based LLM QA is a strong plus. Key responsibilities Quality monitoring: Spot-check Telugu items, identify quality issues, provide ongoing feedback through DMs, and escalate recurring or critical issues. Trainer and QA communication: Update trainers and QAs on Discord about new item guidelines, project changes, workflow updates, and quality expectations. Question handling: Respond to trainer/QA questions clearly and promptly, especially around Telugu wording, register, translation fidelity, cultural context, Andhra Pradesh vs Telangana usage, and edge cases. Trainer/QA activation management: DM contributors who are inactive or not working, encourage activation, track follow-ups, and flag availability issues when needed. Documentation: Create and maintain Telugu project documentation, including style guides, trackers, FAQs, quality notes, examples, honeypots, and onboarding materials. Onboarding and training: Schedule and run onboarding/training calls with trainers and QAs to explain project expectations, workflows, rubrics, quality standards, and Telugu-specific style requirements. Quality alignment: Ensure all trainers and QAs apply Telugu language guidelines consistently and understand updates as projects evolve. Process improvement: Identify recurring quality gaps, propose workflow improvements, and help build scalable QA processes for Telugu-language projects. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
Telugu Linguistic QA Specialist (Remote)
Braintrust
Job description About this role In this hourly, remote contractor role, you will work as a Telugu Quality Assurance Lead (QAL) to oversee quality, consistency, and trainer performance across Telugu AI training projects. You will review AI-generated Telugu content and trainer/QA work, evaluate output quality against project guidelines, provide precise written feedback, and ensure that all contributors follow the expected quality standards. You will assess work for accuracy, fluency, grammar, spelling, tone, cultural appropriateness, meaning preservation, instruction-following, formatting, and adherence to project-specific rubrics. You will spot recurring quality issues, communicate updates to trainers and QAs, support onboarding, maintain documentation, and help activate contributors who are not working consistently. This role requires strong Telugu and English skills, excellent attention to detail, structured communication, and the ability to manage quality workflows across remote teams. This role is with SME Careers, a fast-growing AI Data Services company and subsidiary of SuperAnnotate, delivering training data for many of the world's largest AI companies and foundation-model labs. Your Telugu quality leadership will directly help improve the world's premier AI models by ensuring that Telugu training data is natural, accurate, culturally appropriate, well-documented, and aligned with client expectations. Selection process involves an AI interview, a domain-specific task, and an interview with a recruiter. Your profile Bachelor's or Master's degree in Telugu, Linguistics, Translation, Communications, Journalism, English, Education, Quality Assurance, or a relevant domain/related field. Native or near-native Telugu proficiency with strong reading and writing skills. Strong grasp of the English language to follow project guidelines, communicate with teams, and provide clear feedback in English. 3 years of professional experience in Telugu writing, editing, translation, localization, content QA, AI training, education, annotation, or related language-review workflows. Strong understanding of Telugu grammar, spelling conventions, punctuation, tone, register, regional variation, and cultural context. Ability to evaluate Telugu content against detailed rubrics and identify issues such as mistranslation, literal phrasing, unnatural tone, hallucinated claims, ambiguity, or inconsistent terminology. Experience leading or supporting remote teams of trainers, annotators, reviewers, editors, or QAs is strongly preferred. Comfortable working in fast-moving remote environments using tools such as Discord, Google Sheets, Google Docs, trackers, dashboards, and project management systems. Highly detail-oriented and organized, with the ability to maintain style guides, FAQs, trackers, onboarding materials, honeypots, and other quality documentation. Experience with AI training, data annotation, large language models, prompt/response evaluation, or rubric-based LLM QA is a strong plus. Key responsibilities Quality monitoring: Spot-check Telugu items, identify quality issues, provide ongoing feedback through DMs, and escalate recurring or critical issues. Trainer and QA communication: Update trainers and QAs on Discord about new item guidelines, project changes, workflow updates, and quality expectations. Question handling: Respond to trainer/QA questions clearly and promptly, especially around Telugu wording, register, translation fidelity, cultural context, Andhra Pradesh vs Telangana usage, and edge cases. Trainer/QA activation management: DM contributors who are inactive or not working, encourage activation, track follow-ups, and flag availability issues when needed. Documentation: Create and maintain Telugu project documentation, including style guides, trackers, FAQs, quality notes, examples, honeypots, and onboarding materials. Onboarding and training: Schedule and run onboarding/training calls with trainers and QAs to explain project expectations, workflows, rubrics, quality standards, and Telugu-specific style requirements. Quality alignment: Ensure all trainers and QAs apply Telugu language guidelines consistently and understand updates as projects evolve. Process improvement: Identify recurring quality gaps, propose workflow improvements, and help build scalable QA processes for Telugu-language projects. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
08/01/2026
Full time
Job description About this role In this hourly, remote contractor role, you will work as a Telugu Quality Assurance Lead (QAL) to oversee quality, consistency, and trainer performance across Telugu AI training projects. You will review AI-generated Telugu content and trainer/QA work, evaluate output quality against project guidelines, provide precise written feedback, and ensure that all contributors follow the expected quality standards. You will assess work for accuracy, fluency, grammar, spelling, tone, cultural appropriateness, meaning preservation, instruction-following, formatting, and adherence to project-specific rubrics. You will spot recurring quality issues, communicate updates to trainers and QAs, support onboarding, maintain documentation, and help activate contributors who are not working consistently. This role requires strong Telugu and English skills, excellent attention to detail, structured communication, and the ability to manage quality workflows across remote teams. This role is with SME Careers, a fast-growing AI Data Services company and subsidiary of SuperAnnotate, delivering training data for many of the world's largest AI companies and foundation-model labs. Your Telugu quality leadership will directly help improve the world's premier AI models by ensuring that Telugu training data is natural, accurate, culturally appropriate, well-documented, and aligned with client expectations. Selection process involves an AI interview, a domain-specific task, and an interview with a recruiter. Your profile Bachelor's or Master's degree in Telugu, Linguistics, Translation, Communications, Journalism, English, Education, Quality Assurance, or a relevant domain/related field. Native or near-native Telugu proficiency with strong reading and writing skills. Strong grasp of the English language to follow project guidelines, communicate with teams, and provide clear feedback in English. 3 years of professional experience in Telugu writing, editing, translation, localization, content QA, AI training, education, annotation, or related language-review workflows. Strong understanding of Telugu grammar, spelling conventions, punctuation, tone, register, regional variation, and cultural context. Ability to evaluate Telugu content against detailed rubrics and identify issues such as mistranslation, literal phrasing, unnatural tone, hallucinated claims, ambiguity, or inconsistent terminology. Experience leading or supporting remote teams of trainers, annotators, reviewers, editors, or QAs is strongly preferred. Comfortable working in fast-moving remote environments using tools such as Discord, Google Sheets, Google Docs, trackers, dashboards, and project management systems. Highly detail-oriented and organized, with the ability to maintain style guides, FAQs, trackers, onboarding materials, honeypots, and other quality documentation. Experience with AI training, data annotation, large language models, prompt/response evaluation, or rubric-based LLM QA is a strong plus. Key responsibilities Quality monitoring: Spot-check Telugu items, identify quality issues, provide ongoing feedback through DMs, and escalate recurring or critical issues. Trainer and QA communication: Update trainers and QAs on Discord about new item guidelines, project changes, workflow updates, and quality expectations. Question handling: Respond to trainer/QA questions clearly and promptly, especially around Telugu wording, register, translation fidelity, cultural context, Andhra Pradesh vs Telangana usage, and edge cases. Trainer/QA activation management: DM contributors who are inactive or not working, encourage activation, track follow-ups, and flag availability issues when needed. Documentation: Create and maintain Telugu project documentation, including style guides, trackers, FAQs, quality notes, examples, honeypots, and onboarding materials. Onboarding and training: Schedule and run onboarding/training calls with trainers and QAs to explain project expectations, workflows, rubrics, quality standards, and Telugu-specific style requirements. Quality alignment: Ensure all trainers and QAs apply Telugu language guidelines consistently and understand updates as projects evolve. Process improvement: Identify recurring quality gaps, propose workflow improvements, and help build scalable QA processes for Telugu-language projects. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
Senior Software Engineer (Python) - Agent Evaluation - Freelance/Remote 100 openings
Braintrust
Job description This opportunity is intended for experienced Senior Python Engineers only Open to candidates in the North America, South America, Asia and Europe. Please submit your CV in English and indicate your level of English proficiency. Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment. What this opportunity involves We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks. You'll create challenging tasks and evaluation criteria within realistic simulated environments: Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history Design tasks from intermediate states of these environments - craft the prompt, define what "solved" means, and ensure the task is solvable by an AI agent Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust What this is NOT Not data labeling Not prompt engineering Not writing code from scratch - the agent writes most of the code; you guide and evaluate What we look for - You must meet all the requirements in order to be considered for this project: 5 years of professional experience with Python. Strong experience with FastAPI , pytest , and async/await . Hands-on experience with Docker , PostgreSQL , and CI/CD pipelines. Proven experience writing and maintaining automated tests (not just executing them). Full-stack experience with React and TypeScript is a plus. English proficiency at B2 level or higher . Availability to work 30 hours per week . Why this is hard Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds. How it works Apply Pass qualification(s) Join a project Complete tasks Get paid Hiring & Onboarding Process The process is designed to move quickly and typically includes the following steps: App review and invitation to a virtual project introduction session (approximately 30 minutes) Platform registration and identity verification Technical assessment (approximately 35 minutes) Background check (completed at no cost to candidates) Onboarding and project-specific training tasks Begin production work! Additional Requirements Willingness to complete identity verification as part of the onboarding process. Ability to complete a technical assessment. Willingness to join and participate in Discord, which will be used for project communication and updates. Successful completion of a background check is required prior to onboarding. Reliable internet connection and ability to communicate effectively in a remote environment. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
08/01/2026
Full time
Job description This opportunity is intended for experienced Senior Python Engineers only Open to candidates in the North America, South America, Asia and Europe. Please submit your CV in English and indicate your level of English proficiency. Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment. What this opportunity involves We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks. You'll create challenging tasks and evaluation criteria within realistic simulated environments: Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history Design tasks from intermediate states of these environments - craft the prompt, define what "solved" means, and ensure the task is solvable by an AI agent Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust What this is NOT Not data labeling Not prompt engineering Not writing code from scratch - the agent writes most of the code; you guide and evaluate What we look for - You must meet all the requirements in order to be considered for this project: 5 years of professional experience with Python. Strong experience with FastAPI , pytest , and async/await . Hands-on experience with Docker , PostgreSQL , and CI/CD pipelines. Proven experience writing and maintaining automated tests (not just executing them). Full-stack experience with React and TypeScript is a plus. English proficiency at B2 level or higher . Availability to work 30 hours per week . Why this is hard Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds. How it works Apply Pass qualification(s) Join a project Complete tasks Get paid Hiring & Onboarding Process The process is designed to move quickly and typically includes the following steps: App review and invitation to a virtual project introduction session (approximately 30 minutes) Platform registration and identity verification Technical assessment (approximately 35 minutes) Background check (completed at no cost to candidates) Onboarding and project-specific training tasks Begin production work! Additional Requirements Willingness to complete identity verification as part of the onboarding process. Ability to complete a technical assessment. Willingness to join and participate in Discord, which will be used for project communication and updates. Successful completion of a background check is required prior to onboarding. Reliable internet connection and ability to communicate effectively in a remote environment. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
Senior Software Engineer (Python) - Agent Evaluation - Freelance/Remote 100 openings
Braintrust Boston, Massachusetts
Job description This opportunity is intended for experienced Senior Python Engineers only Open to candidates in the North America, South America, Asia and Europe. Please submit your CV in English and indicate your level of English proficiency. Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment. What this opportunity involves We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks. You'll create challenging tasks and evaluation criteria within realistic simulated environments: Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history Design tasks from intermediate states of these environments - craft the prompt, define what "solved" means, and ensure the task is solvable by an AI agent Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust What this is NOT Not data labeling Not prompt engineering Not writing code from scratch - the agent writes most of the code; you guide and evaluate What we look for - You must meet all the requirements in order to be considered for this project: 5 years of professional experience with Python. Strong experience with FastAPI , pytest , and async/await . Hands-on experience with Docker , PostgreSQL , and CI/CD pipelines. Proven experience writing and maintaining automated tests (not just executing them). Full-stack experience with React and TypeScript is a plus. English proficiency at B2 level or higher . Availability to work 30 hours per week . Why this is hard Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds. How it works Apply Pass qualification(s) Join a project Complete tasks Get paid Hiring & Onboarding Process The process is designed to move quickly and typically includes the following steps: App review and invitation to a virtual project introduction session (approximately 30 minutes) Platform registration and identity verification Technical assessment (approximately 35 minutes) Background check (completed at no cost to candidates) Onboarding and project-specific training tasks Begin production work! Additional Requirements Willingness to complete identity verification as part of the onboarding process. Ability to complete a technical assessment. Willingness to join and participate in Discord, which will be used for project communication and updates. Successful completion of a background check is required prior to onboarding. Reliable internet connection and ability to communicate effectively in a remote environment. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386
08/01/2026
Full time
Job description This opportunity is intended for experienced Senior Python Engineers only Open to candidates in the North America, South America, Asia and Europe. Please submit your CV in English and indicate your level of English proficiency. Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based, not permanent employment. What this opportunity involves We're building a dataset to evaluate AI coding agents - how well a model handles real-world developer tasks. You'll create challenging tasks and evaluation criteria within realistic simulated environments: Build realistic developer environments - a virtual company with codebase, infrastructure, and context (tickets, docs, conversations) that forms a believable development history Design tasks from intermediate states of these environments - craft the prompt, define what "solved" means, and ensure the task is solvable by an AI agent Write tests that verify agent solutions - accept all valid approaches and reject incorrect ones, neither too strict nor too lenient Iterate on tasks and tests based on QA feedback - review agent solutions, analyze failures, and refine until the evaluation is fair and robust What this is NOT Not data labeling Not prompt engineering Not writing code from scratch - the agent writes most of the code; you guide and evaluate What we look for - You must meet all the requirements in order to be considered for this project: 5 years of professional experience with Python. Strong experience with FastAPI , pytest , and async/await . Hands-on experience with Docker , PostgreSQL , and CI/CD pipelines. Proven experience writing and maintaining automated tests (not just executing them). Full-stack experience with React and TypeScript is a plus. English proficiency at B2 level or higher . Availability to work 30 hours per week . Why this is hard Frontier models are already good at coding. Creating a task that genuinely challenges the best models is non-trivial. You need to deeply understand where models fail and what scenarios reveal the difference between a good and a bad solution. Tasks have many valid solutions - writing tests that accept all correct solutions and reject incorrect ones is harder than it sounds. How it works Apply Pass qualification(s) Join a project Complete tasks Get paid Hiring & Onboarding Process The process is designed to move quickly and typically includes the following steps: App review and invitation to a virtual project introduction session (approximately 30 minutes) Platform registration and identity verification Technical assessment (approximately 35 minutes) Background check (completed at no cost to candidates) Onboarding and project-specific training tasks Begin production work! Additional Requirements Willingness to complete identity verification as part of the onboarding process. Ability to complete a technical assessment. Willingness to join and participate in Discord, which will be used for project communication and updates. Successful completion of a background check is required prior to onboarding. Reliable internet connection and ability to communicate effectively in a remote environment. Company Braintrust is a global talent network that connects top independent professionals with leading companies for high-quality, flexible work. We help organizations hire skilled talent faster while giving professionals access to vetted opportunities with innovative teams.5c143e31-5e48-4549-b2d185386

Modal Window

  • Home
  • Contact
  • About Us
  • FAQs
  • Terms & Conditions
  • Privacy
  • Employer
  • Post a Job
  • Search Resumes
  • Sign in
  • Job Seeker
  • Find Jobs
  • Create Resume
  • Sign in
  • IT blog
  • Facebook
  • Twitter
  • LinkedIn
  • Youtube
© 2008-2026 IT Job Board