You can think of this class as a virtual document, such as an HTML page, a PDF file, or a text file.
你可以把这个类想象成代表了一个实际的文档,比如一个html页面,一个PD f文档,或者一个文本文件。
Output file types include, but aren't limited to, Microsoft Office, PDF, rich text and HTML.
输出文件的类型包括MicrosoftOffice、PDF、富文本和HTML,但不局限于这几种。
The PDDocument class represents the PDF file, and the PDFTextStripper knows how to extract text from the PDF files.
PDDocument类代表了PDF文件,而 PDFTextStripper知道怎样从 PDF 文件中获取文本。
Based on PDF file structure analysis, this article gives a proposal on how to recover the text content and file information, and introduces the application and effect of this proposal.
在分析PD F文件结构的基础上,提出了一套实现受损pdf的文本内容和文件信息修复的方案,并介绍了该技术的应用前景和效果。
This paper introduces the structure of PDF documents, and shows the procedures for file parsing and text extraction from the parsed content streams.
介绍了PDF的文件结构,在此基础上,给出了PD F文件的解析流程,以及从解析后的内容流中提取文本内容的方法。
Also it should allow to add all types of components in PDF file like text, image and different types of charts.
也应该允许添加所有类型的组件在PD F文件如文本,图像和不同类型的。
Also it should allow to add all types of components in PDF file like text, image and different types of charts.
也应该允许添加所有类型的组件在PD F文件如文本,图像和不同类型的。
应用推荐