Archive for June 28th, 2008

Batch Extract XMP from PDF to XML

Here is a requirement that want to batch dump xmp from PDF to xml file or Database,

I’d like to know if you have developed or if you can develop an application for extracting customize XMP from PDF documents.

I’ll try to be more relevant: I customized a specific card for additional metadata in Acrobat Professional. If I save the xmp properties in xml format, I obtain the value that I inserted, after that I import xml file in Database. I’d like to know if is possible to develop an application that can extract xmp customized value from a group of PDF files.

And what is XMP?

Adobe’s Extensible Metadata Platform (XMP) is a labeling technology that allows you to embed data about a file, known as metadata, into the file itself. With XMP, desktop applications and back-end publishing systems gain a common method for capturing, sharing, and leveraging this valuable metadata — opening the door for more efficient job processing, workflow automation, and rights management, among many other possibilities. With XMP, Adobe has taken the “heavy lifting” out of metadata integration, offering content creators an easy way to embed meaningful information about their projects and providing industry partners with standards-based building blocks to develop optimized workflow solutions.

Finally, I used iTextSharp(of course iText also ok) to batch extract XMP from PDF, and save it to XML.

Share and Enjoy:
  • Digg
  • del.icio.us
  • Netvouz
  • DZone
  • ThisNext
  • MisterWong
  • Wists
  • BlinkList
  • blogmarks
  • blogtercimlap
  • connotea
  • DotNetKicks
  • Fark
  • Fleck
  • Gwar
  • Haohao
  • IndianPad
  • Internetmedia
  • LinkaGoGo
  • MyShare
  • Netscape
  • NewsVine
  • Rec6
  • Reddit
  • Scoopeo
  • Slashdot
  • StumbleUpon
  • Technorati
  • Webride

Found super fast tools to divide A4 to 2 A5 pages

It is just a article talks about my two softwares posted by JimmyZou on http://www.mobileread.com/forums/archive/index.php/t-10515.html

After google a lot, I find super fast tool to cut A4 pdf into double pages A5 file.

I did one, and attach it behide, you guys can check it out.

Using DOS command to do it, and it done very fast, I split the 210K size pdf in 2 seconds, and the outcome file is 240K only, very effective!

You can find the software here:

http://www.rubypdf.com/

There are two softwares needed: PDFRotate and PDFDivide

First copy all software and the PDFs into one directory, the 2 simple steps:
1.First use PDFRoate
\PDFRoate 1-A4.pdf 1-A4-90.pdf 90
Rotate pdf 90 degree first, prepare to divide it.
2.USE PDFDivide
\Divide 1-A4-90.PDF 1-2A5.PDF

The only none-beatiful thing is that it cut directly, so in some pages there are 1 line letters been cut into 2 parts.

But anyway, it’s nice and fast.

Share and Enjoy:
  • Digg
  • del.icio.us
  • Netvouz
  • DZone
  • ThisNext
  • MisterWong
  • Wists
  • BlinkList
  • blogmarks
  • blogtercimlap
  • connotea
  • DotNetKicks
  • Fark
  • Fleck
  • Gwar
  • Haohao
  • IndianPad
  • Internetmedia
  • LinkaGoGo
  • MyShare
  • Netscape
  • NewsVine
  • Rec6
  • Reddit
  • Scoopeo
  • Slashdot
  • StumbleUpon
  • Technorati
  • Webride