[dna-issues] [JBoss JIRA] Updated: (DNA-114) Create MS Office file sequencer

Michael Trezzi (JIRA) jira-events at lists.jboss.org
Wed Jun 11 02:22:15 EDT 2008


     [ http://jira.jboss.com/jira/browse/DNA-114?page=all ]

Michael Trezzi updated DNA-114:
-------------------------------

    Attachment: dna-sequencer-msoffice.zip

Ok, I am attaching the current version of the sequencer. It has test, however it waits for the SequencingContext to be ready for recognizing the format of the document. Also, as I am still learning the output.setProperty statements might be all wrong :)

> Create MS Office file sequencer
> -------------------------------
>
>                 Key: DNA-114
>                 URL: http://jira.jboss.com/jira/browse/DNA-114
>             Project: DNA
>          Issue Type: Task
>          Components: Sequencers
>            Reporter: Randall Hauch
>             Fix For: 0.2
>
>         Attachments: dna-sequencer-msoffice.zip
>
>
> Create a single sequencer that is capable of sequencing the MS Office files, including MS Word, MS Excel, and MS PowerPoint.  All of the files' standard metadata (author, title, word count, page count, etc.) should be extracted, as should metadata specific to the different kinds of files.  For example, the sequencer should extract all of the content of the Excel spreadsheets, while it should extract at least the slide titles (and ideally thumbnails).

-- 
This message is automatically generated by JIRA.
-
If you think it was sent incorrectly contact one of the administrators: http://jira.jboss.com/jira/secure/Administrators.jspa
-
For more information on JIRA, see: http://www.atlassian.com/software/jira

        



More information about the dna-issues mailing list