Sponsored Content
Top Forums UNIX for Beginners Questions & Answers Limitations of 'pdftotext' in Linux... Post 303041313 by apmcd47 on Thursday 21st of November 2019 06:12:50 AM
Old 11-21-2019
This is probably a stupid question, but is your PDF rendering text or an image, such a a scanned-in page, or text rendered to an image before being imported into the PDF? I'm not quite sure how you can check that. If your copy of Acrobat has the capability, can you perform an OCR on the document?

Andrew
 

10 More Discussions You Might Find Interesting

1. UNIX for Dummies Questions & Answers

mkdir limitations

What characters can't be used with a mkdir? Any limits on length of name? Thank you, Randy M. Zeitman http://www.StoneRoseDesign.com (12 Replies)
Discussion started by: flignar
12 Replies

2. UNIX for Dummies Questions & Answers

csplit limitations

I am trying to use the csplit file on a file that contains records that have more than 2048 characters on a line. The resultant split file seems to ignore the rest of the line and I lose the data. Is there any way that csplit can handle record lengths greater than 2048? Thanks (0 Replies)
Discussion started by: ravagga
0 Replies

3. HP-UX

pdftotext / PDF conversion to .txt binaries

Good day, I've been trying to look for a way to compile the Xpdf sources in our HP-UX server, but have been failing to do so because there is no GCC installed, and I don't have privileges to install GCC. I was looking for a functionality to convert PDF files to .txt, which is exactly like the... (2 Replies)
Discussion started by: mike_s_6
2 Replies

4. UNIX and Linux Applications

gnuplot limitations

I'm running a simulation (programmed in C) which makes calls to gnuplot periodically to plot data I have stored. First I open a pipe to gnuplot and set it to multiplot: FILE * pipe = popen("gnuplot", "w"); fprintf(pipe, "set multiplot\n"); fflush(pipe); (this pipe stays open until the... (0 Replies)
Discussion started by: sedavidw
0 Replies

5. Red Hat

Limitations on the partition of linux

Hi, I need a documentation about limitations on the linux partition. On how many primary and extended I could create. And also on different type of storage, how many big capacity I can create. Thanks. (3 Replies)
Discussion started by: itik
3 Replies

6. UNIX for Dummies Questions & Answers

Basic problem with pdftotext

Hi, I have used pdftotext with good results in the past, but today for some reason I keep getting the same error message. My command is as follows: And the error message is I am using Vmware player with Ubuntu server, but I don't think that is causing this issue as I have been using... (2 Replies)
Discussion started by: Joq
2 Replies

7. Red Hat

Eth0 Limitations

Hi, I have noticed some performance issues on my RHEL5 server but the memory and CPU utilization on the box is fine. I have a 1G full duplexed eth0 card and I am suspicious that this may be causing the problem. My eth0 settings are as follows: Settings for eth0: Supported ports: ... (12 Replies)
Discussion started by: Duffs22
12 Replies

8. Solaris

Solaris limitations

Hi, I recently started working with Solaris, and what I noticed is that a lot of commands I used to regularly use don't work, like sed -i and grep -r. I have found work arounds for these problems though but it's a pain in the ass. I'm just wondering why they decided not to include these handy... (4 Replies)
Discussion started by: Subbeh
4 Replies

9. Linux

Linux partitions and limitations

In recently reading an article on linux basics before I embark and my personal installation project I came across this passage - IDE drives have three types of partition: primary, logical, and extended. The partition table is located in the master boot record (MBR) of a disk. The MBR is the... (12 Replies)
Discussion started by: Synchlavier
12 Replies

10. UNIX for Dummies Questions & Answers

Pdftotext from multiple pdf files to a single text file

I have a directory having a number of pdf files. I want to convert all the files to text, stored in a single text file The following creates multiple text files ls *.pdf | xargs -n1 pdftotext (1 Reply)
Discussion started by: kristinu
1 Replies
PDFDRAW(1)						      General Commands Manual							PDFDRAW(1)

NAME
pdfdraw - render PDF documents SYNOPSIS
pdfdraw [options] input.pdf [pages] DESCRIPTION
pdfdraw will render a PDF document to image files. The supported image formats are: pgm, ppm, pam and png. Select the pages to be ren- dered by specifying a comma separated list of ranges and individual page numbers (for example: 1,5,10-15). In no pages are specified all the pages will be rendered. OPTIONS
-o output The image format is deduced from the output file name. Embed %d in the name to indicate the page number (for example: "page%d.png"). -p password Use the specified password if the file is encrypted. -r resolution Render the page at the specified resolution. The default resolution is 72 dpi. -R angle Rotate clockwise by given number of degrees. -a Save the alpha channel. The default behavior is to render each page with a white background. With this option, the page background is transparent. Only supported for pam and png output formats. -g Render in grayscale. The default is to render a full color RGB image. If the output format is pgm or ppm this option is ignored. -m Show timing information. Take the time it takes for each page to render and print a summary at the end. -5 Print an MD5 checksum of the rendered image data for each page. -t Print the text contents of each page in UTF-8 encoding. Give the option twice to print detailed information about the location of each character in XML format. -x Print the display list used to render each page. -A Disable the use of accelerated functions. -G gamma Gamma correct the output image. Some typical values are 0.7 or 1.4 to thin or darken text rendering. -I Invert the output image colors. pages Comma separated list of ranges to render. SEE ALSO
mupdf(1), pdfclean(1). pdfshow(1). AUTHOR
MuPDF was written by Tor Andersson <tor@ghostscript.com>. MuPDF is Copyright 2006-2010 Artifex Software, Inc. September 4, 2011 PDFDRAW(1)
All times are GMT -4. The time now is 05:21 AM.
Unix & Linux Forums Content Copyright 1993-2022. All Rights Reserved.
Privacy Policy