博客
关于我
C# 读取Word文本框中的文本、图片和表格(附VB.NET代码)
阅读量:410 次
发布时间:2019-03-06

本文共 6444 字,大约阅读时间需要 21 分钟。

【概述】

Word中可插入文本框,在文本框中可添加文本、图片、表格等内容。本篇文章通过C#程序代码介绍如何来读取文本框中的文本、图片和表格等内容。附VB.NET代码,有需要可作参考。

【程序环境】

程序中所需必要的程序集文件,及其他相关dll文件(见下文)。

用于测试的Word源文档如图:

 

【程序代码】

1.读取文本框中的文本

所需程序集:

【C#】

using Spire.Doc;using Spire.Doc.Documents;using Spire.Doc.Fields;using System;using System.IO;using System.Text;namespace ExtractText{    class Program    {        static void Main(string[] args)        {            //加载Word源文档            Document doc = new Document();            doc.LoadFromFile("test.docx");            //获取文本框            TextBox textbox = doc.TextBoxes[0];            //创建StringBuilder类的对象            StringBuilder sb = new StringBuilder();            //遍历文本框中的对象,获取文本            foreach (object obj in textbox.Body.ChildObjects)            {                if (obj is Paragraph)                {                    String text = ((Paragraph)obj).Text;                    sb.AppendLine(text);                }            }            //保存写入的txt文档到指定路径            File.WriteAllText("ExtractedText.txt", sb.ToString());            System.Diagnostics.Process.Start("ExtractedText.txt");        }    }}

【vb.net】

Imports Spire.DocImports Spire.Doc.DocumentsImports Spire.Doc.FieldsImports System.IOImports System.TextNamespace ExtractText    Class Program        Private Shared Sub Main(args As String())            '加载Word源文档            Dim doc As New Document()            doc.LoadFromFile("test.docx")            '获取文本框            Dim textbox As TextBox = doc.TextBoxes(0)            '创建StringBuilder类的对象            Dim sb As New StringBuilder()            '遍历文本框中的对象,获取文本            For Each obj As Object In textbox.Body.ChildObjects                If TypeOf obj Is Paragraph Then                    Dim text As [String] = DirectCast(obj, Paragraph).Text                    sb.AppendLine(text)                End If            Next            '保存写入的txt文档到指定路径            File.WriteAllText("ExtractedText.txt", sb.ToString())            System.Diagnostics.Process.Start("ExtractedText.txt")        End Sub    End ClassEnd Namespace

文本读取结果:

2.读取文本框中的图片

所需程序集:

【C#】

using Spire.Doc;using Spire.Doc.Documents;using Spire.Doc.Fields;using System;namespace ExtractImg{    class Program    {        static void Main(string[] args)        {            //加载Word源文档            Document doc = new Document();            doc.LoadFromFile("test.docx");            //获取文本框            TextBox textbox = doc.TextBoxes[0];                int index = 0 ;            //遍历文本框中所有段落            for (int i = 0 ; i < textbox.Body.Paragraphs.Count;i++)            {                Paragraph paragraph = textbox.Body.Paragraphs[i];                //遍历段落中的所有子对象                for (int j = 0; j < paragraph.ChildObjects.Count; j++)                {                    object obj = paragraph.ChildObjects[j];                                        //判定对象是否为图片                    if (obj is DocPicture)                    {                        //获取图片                        DocPicture picture = (DocPicture) obj;                        String imageName = String.Format("Image-{0}.png", index);                        picture.Image.Save(imageName, System.Drawing.Imaging.ImageFormat.Png);                        index++;                    }                }            }                         }    }}

【vb.net】

Imports Spire.DocImports Spire.Doc.DocumentsImports Spire.Doc.FieldsNamespace ExtractImg    Class Program        Private Shared Sub Main(args As String())            '加载Word源文档            Dim doc As New Document()            doc.LoadFromFile("test.docx")            '获取文本框            Dim textbox As TextBox = doc.TextBoxes(0)            Dim index As Integer = 0            '遍历文本框中所有段落            For i As Integer = 0 To textbox.Body.Paragraphs.Count - 1                Dim paragraph As Paragraph = textbox.Body.Paragraphs(i)                '遍历段落中的所有子对象                For j As Integer = 0 To paragraph.ChildObjects.Count - 1                    Dim obj As Object = paragraph.ChildObjects(j)                    '判定对象是否为图片                    If TypeOf obj Is DocPicture Then                        '获取图片                        Dim picture As DocPicture = DirectCast(obj, DocPicture)                        Dim imageName As [String] = [String].Format("Image-{0}.png", index)                        picture.Image.Save(imageName, System.Drawing.Imaging.ImageFormat.Png)                        index += 1                    End If                Next            Next        End Sub    End ClassEnd Namespace

图片读取结果:

3.读取文本框中的表格

所需程序集:

【C#】

using Spire.Doc;using Spire.Doc.Documents;using Spire.Doc.Fields;using System.IO;using System.Text;namespace ExtractTable{    class Program    {        static void Main(string[] args)        {            //加载Word文档            Document doc = new Document();            doc.LoadFromFile("test.docx");            //获取文本框            TextBox textbox = doc.TextBoxes[0];            //获取文本框中表格            Table table = textbox.Body.Tables[0] as Table;            StringBuilder sb = new StringBuilder();            //遍历表格中的段落并提取文本            foreach (TableRow row in table.Rows)            {                foreach (TableCell cell in row.Cells)                {                    foreach (Paragraph paragraph in cell.Paragraphs)                    {                        sb.AppendLine(paragraph.Text);                    }                }            }            File.WriteAllText("ExtractedTable.txt", sb.ToString());        }    }}

【vb.net】

Imports Spire.DocImports Spire.Doc.DocumentsImports Spire.Doc.FieldsImports System.IOImports System.TextNamespace ExtractTable    Class Program        Private Shared Sub Main(args As String())            '加载Word文档            Dim doc As New Document()            doc.LoadFromFile("test.docx")            '获取文本框            Dim textbox As TextBox = doc.TextBoxes(0)            '获取文本框中表格            Dim table As Table = TryCast(textbox.Body.Tables(0), Table)            Dim sb As New StringBuilder()            '遍历表格中的段落并提取文本            For Each row As TableRow In table.Rows                For Each cell As TableCell In row.Cells                    For Each paragraph As Paragraph In cell.Paragraphs                        sb.AppendLine(paragraph.Text)                    Next                Next            Next            File.WriteAllText("ExtractedTable.txt", sb.ToString())        End Sub    End ClassEnd Namespace

表格数据读取结果:

 

【最后】

以上是本文关于通过C#程序读取Word中的文本框的方法。另推荐阅读《

 

(本文完,如需转载,请务必注明出处!!)

 

你可能感兴趣的文章
mudbox卸载/完美解决安装失败/如何彻底卸载清除干净mudbox各种残留注册表和文件的方法...
查看>>
mysql 1264_关于mysql 出现 1264 Out of range value for column 错误的解决办法
查看>>
mysql 1593_Linux高可用(HA)之MySQL主从复制中出现1593错误码的低级错误
查看>>
mysql 5.6 修改端口_mysql5.6.24怎么修改端口号
查看>>
MySQL 8.0 恢复孤立文件每表ibd文件
查看>>
MySQL 8.0开始Group by不再排序
查看>>
mysql ansi nulls_SET ANSI_NULLS ON SET QUOTED_IDENTIFIER ON 什么意思
查看>>
multi swiper bug solution
查看>>
MySQL Binlog 日志监听与 Spring 集成实战
查看>>
MySQL binlog三种模式
查看>>
multi-angle cosine and sines
查看>>
Mysql Can't connect to MySQL server
查看>>
mysql case when 乱码_Mysql CASE WHEN 用法
查看>>
Multicast1
查看>>
mysql client library_MySQL数据库之zabbix3.x安装出现“configure: error: Not found mysqlclient library”的解决办法...
查看>>
MySQL Cluster 7.0.36 发布
查看>>
Multimodal Unsupervised Image-to-Image Translation多通道无监督图像翻译
查看>>
MySQL Cluster与MGR集群实战
查看>>
multipart/form-data与application/octet-stream的区别、application/x-www-form-urlencoded
查看>>
mysql cmake 报错,MySQL云服务器应用及cmake报错解决办法
查看>>