We are looking for PDF Annotation & Transcription Experts – Japanese candidates for a project delivered through Mercor.
What you'll do
- Open and check a task: pages are provided, so you do not source documents yourself. We find the PDFs and upload them for you. Before annotating, confirm the page is in Japanese, is legible, has real content, and shows no personal details
- Annotate structure: identify and bound every meaningful region of the page - document title, section heading, paragraph, list, table, figure, diagram, caption, formula, question, answer field - and assign each a component type and a reading-order index
- Record relationships: link each region to the figure or table it belongs to through a parent component identifier
- Transcribe faithfully: reproduce all text exactly as it appears, including kanji, hiragana, katakana, furigana and handwritten content, flagging any region where the source is not legible
- Capture page metadata: language, document type, source, page dimensions, and flags for tables, formulas and handwriting
- Review a colleague's work: every task is reviewed end to end by a second Japanese expert, and experienced annotators take on that review
What you need
- Native fluency in Japanese, including full command of kanji, hiragana and katakana, is required for this position. All annotation and transcription work is performed in Japanese.
- You are a native Japanese speaker with full command of kanji, hiragana and katakana, including furigana and variant character forms
- You have professional experience in interpretation, journalism, transcription, translation, editorial work, or comparable document-intensive work
- You are exact: character-level accuracy matters more here than speed, and a single wrong character is a defect
- You are systematic: you apply a taxonomy consistently across hundreds of pages rather than improvising per document
- You are comfortable with unfamiliar layouts: vertical text, multi-column newspapers, exam papers, handwritten forms
Nice to have
- AI training data: annotation, labeling, grading, or bilingual evaluation for training datasets
- Transcription and localization: MTPE, subtitling, bilingual QA, OCR correction or post-editing
- Document production: typesetting, copy-editing, proofreading, or digitization of Japanese-language material
- Script and encoding: Unicode normalization, Japanese input methods, full-width and half-width forms, and kanji variant handling
Who you work with
Project and contracting process: Mercor. Applications continue on the provider's website.

