at Terac
Terac is hiring multilingual data contributors on a contract basis to find and submit public, legally usable PDF documents in Telugu, Odia, Gujarati, Malayalam, Japanese, and Korean to support AI training.
WHAT WE'RE RESEARCHING
We're running a paid study on multilingual document sourcing to improve AI text recognition and generation. High-quality, legally usable PDFs in various languages are essential for training robust machine learning models. Your contributions will directly support the development of better language processing tools.
HOW IT WORKS
You will work asynchronously to find and submit public, legally usable PDF documents in your designated language. During this process, you will verify that each document meets our quality and licensing requirements. You will upload the files through our secure platform and provide basic metadata for each submission. We will review your uploaded documents to ensure they match the project guidelines before approving the task.
WHO THIS IS FOR
We are hiring fluent readers of Telugu, Odia, Gujarati, Malayalam, Japanese, and Korean who know how to source public documents online. Ideal candidates are detail-oriented individuals comfortable navigating digital archives, public records, or open-source repositories. We welcome data annotators, researchers, and general language contributors who understand basic copyright and licensing rules.
WHAT YOU'LL DO
WHO SHOULD APPLY
COMPENSATION
$150 per task
READY TO PARTICIPATE?
Start your paid interview now https://terac.com/interview/start/r/8dea52b6-365c-449f-9c7f-e4c456d53652?utm_source=ashby_listing_description
ABOUT TERAC
Terac is building the world's largest pool of vetted human experts for AI. Researchers, AI labs, and product teams use Terac to recruit, screen, and pay study participants across industries, languages, and skill sets.
Learn more at terac.com https://terac.com or on YouTube at @jointerac https://www.youtube.com/@jointerac.
Responsibilities
Requirements