Uploaded image for project: 'Tika'
  1. Tika
  2. TIKA-1191

ForkParser / ClassLoaderProxy does not define package

    Details

    • Type: Bug
    • Status: Resolved
    • Priority: Major
    • Resolution: Fixed
    • Affects Version/s: 1.4, 1.5
    • Fix Version/s: 1.18
    • Component/s: parser
    • Labels:
      None

      Description

      ForkParser will throw an Exception in some cases :

      org.apache.tika.exception.TikaException: Invalid embedded resource
      at org.apache.tika.parser.microsoft.AbstractPOIFSExtractor.handleEmbeddedOfficeDoc(AbstractPOIFSExtractor.java:189)
      at org.apache.tika.parser.microsoft.WordExtractor.parse(WordExtractor.java:135)
      at org.apache.tika.parser.microsoft.OfficeParser.parse(OfficeParser.java:186)
      at org.apache.tika.parser.microsoft.OfficeParser.parse(OfficeParser.java:161)
      at sun.reflect.NativeMethodAccessorImpl.invoke0(Native Method)
      at sun.reflect.NativeMethodAccessorImpl.invoke(NativeMethodAccessorImpl.java:39)
      at sun.reflect.DelegatingMethodAccessorImpl.invoke(DelegatingMethodAccessorImpl.java:25)
      at java.lang.reflect.Method.invoke(Method.java:597)
      at org.apache.tika.fork.ForkServer.call(ForkServer.java:144)
      at org.apache.tika.fork.ForkServer.processRequests(ForkServer.java:124)
      at org.apache.tika.fork.ForkServer.main(ForkServer.java:69)
      Caused by: java.lang.NullPointerException
      at org.apache.tika.mime.MimeTypesFactory.create(MimeTypesFactory.java:136)
      at org.apache.tika.mime.MimeTypes.getDefaultMimeTypes(MimeTypes.java:499)
      at org.apache.tika.config.TikaConfig.getDefaultMimeTypes(TikaConfig.java:60)
      at org.apache.tika.config.TikaConfig.<init>(TikaConfig.java:169)
      at org.apache.tika.config.TikaConfig.getDefaultConfig(TikaConfig.java:268)
      at org.apache.tika.parser.microsoft.AbstractPOIFSExtractor.getTikaConfig(AbstractPOIFSExtractor.java:72)
      at org.apache.tika.parser.microsoft.AbstractPOIFSExtractor.getDetector(AbstractPOIFSExtractor.java:79)
      at org.apache.tika.parser.microsoft.AbstractPOIFSExtractor.handleEmbeddedOfficeDoc(AbstractPOIFSExtractor.java:176)
      ... 10 more

      A patch will follow

        Attachments

        1. ClassLoaderProxy.java.patch
          1 kB
          Nicolas Belisle
        2. Test.java
          1 kB
          Nicolas Belisle
        3. test.eml
          12 kB
          Nicolas Belisle

          Issue Links

            Activity

              People

              • Assignee:
                Unassigned
                Reporter:
                nicolas.belisle Nicolas Belisle
              • Votes:
                0 Vote for this issue
                Watchers:
                8 Start watching this issue

                Dates

                • Created:
                  Updated:
                  Resolved: