Logo Questions Linux Laravel Mysql Ubuntu Git Menu
 

C# PDFSharp: Examples of how to strip text from PDF?

Tags:

c#

text

pdfsharp

I have a fairly simple task: I need to read a PDF file and write out its image contents while ignoring its text contents. So essentially I need to do the complement of "save as text".

Ideally, I would prefer to avoid any sort of re-compression of the image contents but if it's not possible, it's ok too.

Are the examples of how to do it?

Thanks!

like image 209
I Z Avatar asked Mar 06 '12 21:03

I Z


People also ask

What C is used for?

C programming language is a machine-independent programming language that is mainly used to create many types of applications and operating systems such as Windows, and other complicated programs such as the Oracle database, Git, Python interpreter, and games and is considered a programming foundation in the process of ...

What is C in C language?

What is C? C is a general-purpose programming language created by Dennis Ritchie at the Bell Laboratories in 1972. It is a very popular language, despite being old. C is strongly associated with UNIX, as it was developed to write the UNIX operating system.

What is the full name of C?

In the real sense it has no meaning or full form. It was developed by Dennis Ritchie and Ken Thompson at AT&T bell Lab. First, they used to call it as B language then later they made some improvement into it and renamed it as C and its superscript as C++ which was invented by Dr.

Is C language easy?

Compared to other languages—like Java, PHP, or C#—C is a relatively simple language to learn for anyone just starting to learn computer programming because of its limited number of keywords.


1 Answers

Extracting text from a PDF file with PDFsharp is not a simple task.

It was discussed recently in this thread: https://stackoverflow.com/a/9161732/162529

like image 171
I liked the old Stack Overflow Avatar answered Sep 17 '22 22:09

I liked the old Stack Overflow