IBM webMethods Hybrid Integration

IBM webMethods Hybrid Integration

Join this online group to communicate across IBM product users and experts by sharing advice and best practices with peers and staying up to date regarding product enhancements.



#Automation


#Applicationintegration
#webMethods
#Integration
 View Only
  • 1.  problem with two pdf files

    Posted 03/08/05 03:49 PM

    Hi,

    i have a problem with these two pdf files:
    the Tamino Non-XML indexer returns empty , i have tried with others pdf files and i have no problem.
    Can someone help me?


    Regards


    #Tamino
    #webMethods
    #API-Management


  • 2.  RE: problem with two pdf files

    Posted 03/08/05 03:51 PM

    And here there is the other file


    #Tamino
    #API-Management
    #webMethods


  • 3.  RE: problem with two pdf files

    Posted 03/09/05 01:27 PM

    Hi Andrea,
    I found this note in the documentation for the Non-XML Indexer and I think it explains your problem:
    The Tamino Non-XML Indexer supports the extraction of content and metadata from PDF files. However, in some circumstances, for example if the PDF file contains LZW compressed objects, no content information is extracted.
    I loaded two pdf files; one of yours and one supplied with Adobe Acrobat Reader (the Help “Reader.pdf” file). If you open Reader.pdf with an editor like Ultra-Edit, the indexed content appears as readable plain text inside the pdf file. In the case of your file, there is no readable plain text. So I think all of your content is compressed and therefore is not indexable.
    Does this help?

    [This message was edited by Bill Leeney on 09 March 2005 at 13:12.]


    #API-Management
    #Tamino
    #webMethods


  • 4.  RE: problem with two pdf files

    Posted 03/09/05 01:48 PM

    thanks a lot for your help,
    i think it can be this the problem

    best regards


    #webMethods
    #API-Management
    #Tamino