Uses of Interface
org.apache.tika.parser.microsoft.chm.ChmAccessor

Packages that use ChmAccessor
Package
Description
 
  • Uses of ChmAccessor in org.apache.tika.parser.microsoft.chm

    Modifier and Type
    Class
    Description
    class 
    The Header 0000: char[4] 'ITSF' 0004: DWORD 3 (Version number) 0008: DWORD Total header length, including header section table and following data. 000C: DWORD 1 (unknown) 0010: DWORD a timestamp 0014: DWORD Windows Language ID 0018: GUID {7C01FD10-7BAA-11D0-9E0C-00A0-C922-E6EC} 0028: GUID {7C01FD11-7BAA-11D0-9E0C-00A0-C922-E6EC} Note: a GUID is $10 bytes, arranged as 1 DWORD, 2 WORDs, and 8 BYTEs. 0000: QWORD Offset of section from beginning of file 0008: QWORD Length of section Following the header section table is 8 bytes of additional header data.
    class 
    Directory header The directory starts with a header; its format is as follows: 0000: char[4] 'ITSP' 0004: DWORD Version number 1 0008: DWORD Length of the directory header 000C: DWORD $0a (unknown) 0010: DWORD $1000 Directory chunk size 0014: DWORD "Density" of quickref section, usually 2 0018: DWORD Depth of the index tree - 1 there is no index, 2 if there is one level of PMGI chunks 001C: DWORD Chunk number of root index chunk, -1 if there is none (though at least one file has 0 despite there being no index chunk, probably a bug) 0020: DWORD Chunk number of first PMGL (listing) chunk 0024: DWORD Chunk number of last PMGL (listing) chunk 0028: DWORD -1 (unknown) 002C: DWORD Number of directory chunks (total) 0030: DWORD Windows language ID 0034: GUID {5D02926A-212E-11D0-9DF9-00A0C922E6EC} 0044: DWORD $54 (This is the length again) 0048: DWORD -1 (unknown) 004C: DWORD -1 (unknown) 0050: DWORD -1 (unknown)
    class 
    ::DataSpace/Storage//ControlData This file contains $20 bytes of information on the compression.
    class 
    LZXC reset table For ensuring a decompression.
    class 
    Description Note: not always exists An index chunk has the following format: 0000: char[4] 'PMGI' 0004: DWORD Length of quickref/free area at end of directory chunk 0008: Directory index entries (to quickref/free area) The quickref area in an PMGI is the same as in an PMGL The format of a directory index entry is as follows: BYTE: length of name BYTEs: name (UTF-8 encoded) ENCINT: directory listing chunk which starts with name Encoded Integers aka ENCINT An ENCINT is a variable-length integer.
    class 
    Description There are two types of directory chunks -- index chunks, and listing chunks.
    Methods in org.apache.tika.parser.microsoft.chm with parameters of type ChmAccessor
    Modifier and Type
    Method
    Description
    static final void
    ChmAssert.assertChmAccessorNotNull(ChmAccessor<?> chmAccessor)
    Checks if ChmAccessor is not null In case of null throws exception
    static final void
    ChmAssert.assertChmAccessorParameters(byte[] data, ChmAccessor<?> chmAccessor, int count)
    Checks validity of ChmAccessor parameters