Jump to content
Main menu
Main menu
move to sidebar
hide
Navigation
Main page
Current events
Recent changes
Random page
Help
Categories
LemonWiki共筆
Search
Search
Appearance
Log in
Personal tools
Log in
Pages for logged out editors
learn more
Contributions
Talk
Editing
PDF to Markdown Conversion
(section)
Page
Discussion
English
Read
Edit
View history
Tools
Tools
move to sidebar
hide
Actions
Read
Edit
View history
General
What links here
Related changes
Special pages
Page information
Appearance
move to sidebar
hide
Warning:
You are not logged in. Your IP address will be publicly visible if you make any edits. If you
log in
or
create an account
, your edits will be attributed to your username, along with other benefits.
Anti-spam check. Do
not
fill this in!
=== Prompt === Prompt <pre> I will provide a PDF file. Convert its contents into Markdown-formatted text. Please follow these requirements: # If the PDF contains images, charts, screenshots, scanned pages, diagrams, or other graphical content, inspect the visual content and extract '''all readable text contained within it'''. Transcribe that text into the Markdown output. Do not omit text simply because it is embedded in an image rather than stored in the PDF text layer. # If a graphical element contains no text, such as a photograph, icon, illustration, or purely visual diagram, do not skip it. Represent it using the following format: <code>[Image: Brief description of its content or purpose]</code> # Preserve the original document's structure and hierarchy as closely as possible, including: * Headings (<code>#</code>, <code>##</code>, <code>###</code>, etc.) * Paragraphs * Ordered and unordered lists * Tables * '''Bold text''' '' ''Italic text* * Blockquotes * Code blocks * Other meaningful structural formatting # Convert tables into standard Markdown table syntax whenever possible. If a table is too complex to reproduce accurately in Markdown, explain the limitation and preserve the relationships between rows, columns, headers, and values as faithfully as possible. # If a page contains both textual and graphical content, preserve the original '''reading order'''. Insert transcribed image content or image descriptions at the position where the corresponding visual element appears in the document. # Do not add, rewrite, correct, infer, summarize, or guess information that does not appear in the original PDF. The output should be a transcription and structural conversion, not an interpretation of the document. # If any text is blurry, obscured, incomplete, or otherwise impossible to identify reliably, use: <code>[Unrecognizable]</code> Do not guess the missing text. # When both the PDF text layer and the visual rendering contain the same text, avoid unnecessary duplication. Prefer the version that most accurately represents the document while preserving visually significant information that may not exist in the text layer. # Pay particular attention to visual elements that commonly contain important text, including: * Screenshots * Scanned documents * Infographics * Flowcharts * Charts and graphs * Forms * Presentation slides embedded in the PDF * Captions and annotations * Labels inside diagrams * Text rendered as part of an image </pre>
Summary:
Please note that all contributions to LemonWiki共筆 are considered to be released under the Creative Commons Attribution-NonCommercial-ShareAlike (see
LemonWiki共筆:Copyrights
for details). If you do not want your writing to be edited mercilessly and redistributed at will, then do not submit it here.
You are also promising us that you wrote this yourself, or copied it from a public domain or similar free resource.
Do not submit copyrighted work without permission!
Cancel
Editing help
(opens in new window)