Merge multiple word documents into one Open Xml

前端 未结 4 891
余生分开走
余生分开走 2020-11-30 01:57

I have around 10 word documents which I generate using open xml and other stuff. Now I would like to create another word document and one by one I would like to join them in

4条回答
  •  攒了一身酷
    2020-11-30 02:49

    Using openXML SDK only, you can use AltChunk element to merge the multiple document into one.

    This link the-easy-way-to-assemble-multiple-word-documents and this one How to Use altChunk for Document Assembly provide some samples.

    EDIT 1

    Based on your code that uses altchunk in the updated question (update#1), here is the VB.Net code I have tested and that works like a charm for me:

    Using myDoc = DocumentFormat.OpenXml.Packaging.WordprocessingDocument.Open("D:\\Test.docx", True)
            Dim altChunkId = "AltChunkId" + DateTime.Now.Ticks.ToString().Substring(0, 2)
            Dim mainPart = myDoc.MainDocumentPart
            Dim chunk = mainPart.AddAlternativeFormatImportPart(
                DocumentFormat.OpenXml.Packaging.AlternativeFormatImportPartType.WordprocessingML, altChunkId)
            Using fileStream As IO.FileStream = IO.File.Open("D:\\Test1.docx", IO.FileMode.Open)
                chunk.FeedData(fileStream)
            End Using
            Dim altChunk = New DocumentFormat.OpenXml.Wordprocessing.AltChunk()
            altChunk.Id = altChunkId
            mainPart.Document.Body.InsertAfter(altChunk, mainPart.Document.Body.Elements(Of DocumentFormat.OpenXml.Wordprocessing.Paragraph).Last())
            mainPart.Document.Save()
    End Using
    

    EDIT 2

    The second issue (update#2)

    This code is appending the Test2 data twice, in place of Test1 data as well.

    is related to altchunkid.

    For each document you want to merge in the main document, you need to:

    1. add an AlternativeFormatImportPart in the mainDocumentPart with an Id which must to be unique. This element contains the Inserted data
    2. add in the body an Altchunk element in which you set the id to reference the previous AlternativeFormatImportPart.

    In your code, you are using the same Id for all the AltChunks. It's why you see many time the same text.

    I am not sure the altchunkid will be unique with your code: string altChunkId = "AltChunkId" + DateTime.Now.Ticks.ToString().Substring(0, 2);

    If you don't need to set a specific value, I recommend you to not set explicitly the AltChunkId when you add the AlternativeFormatImportPart. Instead, you get one generated by the SDK like this:

    VB.Net

    Dim chunk As AlternativeFormatImportPart = mainPart.AddAlternativeFormatImportPart(DocumentFormat.OpenXml.Packaging.AlternativeFormatImportPartType.WordprocessingML)
    Dim altchunkid As String = mainPart.GetIdOfPart(chunk)
    

    C#

    AlternativeFormatImportPart chunk = mainPart.AddAlternativeFormatImportPart(DocumentFormat.OpenXml.Packaging.AlternativeFormatImportPartType.WordprocessingML);
    string altchunkid = mainPart.GetIdOfPart(chunk);
    

提交回复
热议问题