Summary

Arlo Barnes

“ A scan of a work is uploaded (like all other media files) . An index is created with the same filename as the scanned file, but in the Index: namespace. The index contains some bibliographic metadata, but primarily exists to give the ProofreadPage extension a mapping between the File: and the pages in the Page: namespace, and to provide a mapping between the physical pages of the PDF or DjVu file and the "logical" page numbers of the work (think of a book where the first chapter is "Page 1", but where the actual first page of the PDF is the cover, or title page, etc.) ”
Source: Wikisource

Arlo Barnes

“ Based on the Index: the ProofreadPage extension creates (initially empty/redlinked) virtual pages in Page: namespace for every physical page in the DjVu or PDF. These wiki pages is where you put the wikimarkup needed to roughly recreate on-wiki the text of the original page. If the original DjVu or PDF file contains a OCR text layer, ProofreadPage helps you out by extracting that OCR text and putting it into the edit field for an empty page. For the above example, these pages will live at Page:TheGreatNovel.pdf/1, Page:TheGreatNovel.pdf/2, etc. ”
Source: Wikisource

Arlo Barnes

“ I would recommend skipping complicated pages like title pages and pages with chapter headings etc. for now, and focus on the pages that mainly contain prose text. There are usually very few of those special pages relative to the whole work, but you can end up spending an inordinate amount of time on figuring out how to do them; so unless you feel a page is relatively straightforward it's better to skip it for now or ask the community for help. ”
Source: Wikisource

Get perspective with Kwize: daily news enlightened by great literature