跳到主要内容

转变 PDF 力争卓越 Power Automate

此操作的作用

PDF4me 转变 PDF 卓越 变换 PDF 文件导入 完全可编辑的 Microsoft Excel (.xlsx) 你电脑里的电子表格 Power Automate 流程。自动提取表格、行项目、数值数据和表单回复,并将其保存为 Excel 原生单元格,而非平面图像。选择 草稿模式 快速原生应用PDF 提取(1) API 按文件调用)或 高模式 人工智能 OCR 适用于 8 种以上语言的扫描文档和复杂布局(2 API 每页调用次数)。用一个直接链接到以下流程的自动化操作,替换手动上传 Smallpdf、iLovePDF、Adobe Acrobat 或 Tabula 的操作 SharePoint、Excel Online、Power BI、Dataverse、 Dynamics 365,以及任何 Microsoft 365 工作流程。

相关博客文章(1)

验证您的身份 API 要求

每一个 PDF4me 行动 Power Automate 需要有效 PDF4me 联系用你的 API 第一次按下键 Power Automate 在所有地方重复使用它 PDF4me 流程中的操作。

您不容错过的重要事实

草稿模式还是高级模式,请根据您的需求选择合适的模式。 PDF 来源
草稿使用 1 API 每个文件调用一次,非常适合原生应用 PDFs 来自会计软件, ERPs或者 BI 工具(规模化应用时成本降低 5-10 倍)。高使用率 2 API 每页调用次数(含完整内容) OCR + 表格重建,扫描发票、拍摄的账单和基于图像的表格需要进行此操作。
原生 Excel 单元格,而非表格的平面图像
表格标题会变成 Excel 标题;数值列会显示为数字;日期列会解析为日期。单元格可排序、可筛选,并可用于公式、数据透视表和 Power BI 导入,这与屏幕截图样式不同。 PDF-到 Excel 的工具。
输出格式为 .xlsx(Office Open XML),而非传统的 .xls。
现代的 .xlsx 格式及 MIME 类型 application/vnd.openxmlformats-officedocument.spreadsheetml.sheet兼容 Excel 2007+、Excel Online 和 Excel for Mac。 Microsoft 365Power BI、Google Sheets 和 LibreOffice Calc。
Power Automate PDF4me 将 PDF 转换为 Excel 操作,显示从之前的“获取文件内容”步骤映射的文件内容、带有 .pdf 扩展名的文件名,以及设置为“草稿”(适用于原生 PDF)或“高”(适用于扫描文档)的质量类型。

从上一步映射文件内容和文件名,选择质量类型(原生文件为草稿)。 PDFs(扫描结果较高),然后运行,输出中将返回可编辑的 XLSX 文件。

参数

必需的: 必须提供文件内容、文件名和质量类型。如果没有二进制源文件,则无法进行转换。 PDF 以及优质的选择。

范围必需的它的作用例子
File ContentYesBinary content of the source PDF file. Map from any prior action that returns file bytes: SharePoint Get file content, OneDrive Get file content, Outlook email attachment, Forms file upload, HTTP Request, or Get file content using path.File Content (binary)
File NameYesFilename of the source PDF including the .pdf extension. Used for identification during processing and to derive the default output filename (replacing .pdf with .xlsx). Map dynamically from prior step variables.invoice-2026.pdf
Quality TypeYesConversion quality mode. Draft for native born-digital PDFs with selectable text (1 API call per file, fast and cost-efficient). High for scanned PDFs, image-based PDFs, or complex multi-column layouts (2 API calls per page with OCR and table reconstruction).Draft

质量类型选项

草稿默认,快速,1 API 每个文件的调用
最适合原生数字用户 PDFs 支持选择文本、从 QuickBooks/SAP/Xero/NetSuite 导出发票、BI 工具报告、浏览器打印到 -PDF 财务报表 ERP生成的文档。几秒钟内即可返回可用的 XLSX 文件。规模化运行时,成本效益比高模式高 5-10 倍。
高的准确,2 API 每页调用次数 OCR
适用于扫描发票、拍摄的银行对账单、基于图像的表格或复杂的多列财务报表。应用人工智能技术 OCR + 表格重建。支持英语、德语、法语、西班牙语、意大利语、葡萄牙语、荷兰语、波兰语 OCR

输出字段

场地类型它包含什么
File ContentBinaryConverted XLSX file as binary content. Map directly into downstream actions: SharePoint Create file, OneDrive Create file, Excel Online (Business) Add a row, Power BI dataset refresh, Outlook Send an email attachment, or Dynamics 365 Record attachment.
File NameStringOutput filename of the generated Excel spreadsheet (e.g. invoice-2026.xlsx). Derived from the input File Name with the .pdf extension replaced by .xlsx. Use directly as the filename in SharePoint or OneDrive Create file actions.

快速设置

  1. 在你的 Power Automate 流程,点击 + 新步骤 并搜索 PDF4me
  2. 选择 转变 PDF 卓越 行动来自 PDF4me 连接连接器。
  3. 选择你的 PDF4me 连接或点击 添加新连接 你的 API 钥匙。
  4. 地图 文件内容 到先前操作的二进制输出:通常 Get file contentSharePoint), Get file content using path (OneDrive) Outlook 电子邮件附件或 HTTP 响应正文。
  5. 文件名 从上一步动态获取(必须包含) .pdf 扩展):例如 triggerOutputs()?['body/Name']
  6. 选择 质量类型
  • 草稿 本地 PDFs (发票导出、BI 报表、 ERP 产出):快速且成本效益高。
  • 高的 扫描 PDFs拍摄银行对账单或复杂的多列表格,均适用。 OCR + 表格重建。
  1. 保存流程并使用样本进行测试 PDF 包含表格。
  2. 连接输出端 文件内容 接下来,请执行以下操作: SharePoint 创建文件、Excel Online 添加行、Power BI 数据集刷新 Dynamics 365 记录附件,或 Outlook 以附件形式发送至电子邮件。

工作流程示例

工作流程示例Common Power Automate flow patterns using Convert PDF to Excel.
供应商发票邮件 → Excel 明细 → 导入会计系统
  1. 一个 Outlook 当收到来自已知供应商域的电子邮件时,触发器会触发。 PDF 发票附件。
  2. 转变 PDF 将 Excel 文件导出时,质量类型设置为“草稿”(供应商发票通常为原生格式)。 PDFs 来自会计系统)。
  3. 生成的 XLSX 文件包含以 Excel 原生行形式呈现的行项目(描述、数量、单价、总计)。
  4. Excel Online 操作读取行并将每一行作为记录发布到 Dynamics 通过 365 Finance、Business Central 或您的应付账款系统 HTTP
  5. 原版 PDF 提取的 XLSX 文件归档于 SharePoint 包含元数据(供应商、金额、发票号),以便进行审计追溯。
月度财务 PDF 报表 → Excel → Power BI 仪表板刷新
  1. 每月1日,系统会按计划运行“月度报告”流程。 SharePoint 包含库 PDF 来自审计师和财务团队的财务报告。
  2. 对于每个新的 PDF, 转变 PDF 使用“质量类型 = 草稿”运行 Excel,生成一个可编辑的 XLSX 文件,其中提取了所有数据表。
  3. Excel Online 操作会将数据追加到主“合并财务报表”工作簿中。 SharePoint按报告类别和报告期划分。
  4. 在合并工作簿上触发 Power BI 数据集刷新后,执行仪表板现在会在几分钟内反映最新数据。 PDF 到达。
  5. Microsoft Teams 中的自适应卡片会通知首席财务官,月度结算数据已准备好供审核。
扫描的银行对账单 → OCR Excel → 自动对账
  1. A SharePoint 当簿记员上传扫描的银行对账单时,触发器会触发。 PDFs 添加到“待核对报表”库中。
  2. 转变 PDF 以质量类型 = 高(扫描源)运行到 Excel, OCR + 需要重新构建表格)。
  3. OCR-提取的 XLSX 文件包含每一行交易记录(日期、描述、金额、余额),以可编辑、可排序的 Excel 单元格形式呈现。
  4. A Power Automate 将每个循环中的每个事务与 Dataverse 中的记录进行匹配,或者 ERP匹配的行会被标记为“已核对”;未匹配的行会被路由到 Microsoft Teams 审批频道,供记账员调查。
  5. 已核对的 XLSX 文件已存档至“已核对报表”文件夹。 SharePoint 包含元数据(帐户、周期、匹配率)的库,用于审计。

常见问题解答

What's the difference between Draft and High quality in Convert PDF to Excel?+
Draft mode is the fast path: one API call per file, designed for native born-digital PDFs with selectable text such as invoice exports from QuickBooks, Xero, SAP, NetSuite, browser print-to-PDF financial statements, accounting software output, and BI tool reports. It returns a usable XLSX in seconds and is the right choice for most modern PDFs. High mode uses two API calls per page and applies AI-powered OCR plus table reconstruction, designed for scanned PDFs, photographed pages, image-based PDFs, or PDFs with complex multi-column financial layouts. Use High when source files came from a scanner or when Draft output has missing rows, misaligned tables, or character recognition errors.
How does Convert PDF to Excel extract tables from a PDF document?+
The action automatically detects table boundaries in the source PDF, reconstructs row and column structure, and outputs the data as native Excel cells: not flat images of cells. Table headers become Excel headers; numeric columns are typed as numbers (not strings) so SUM, AVERAGE, and other formulas work immediately; date columns are parsed as dates where possible. For PDFs with multiple tables on one page or tables spanning multiple pages, the action handles each table as a separate region while preserving row-level relationships. Use High quality mode for scanned tables, multi-column layouts, or PDFs where row alignment is critical (such as bank statements or multi-page invoices).
Can I convert scanned or image-based PDFs to editable Excel spreadsheets?+
Yes. Set Quality Type to High to apply AI-powered OCR over each page. The action extracts text and numeric values from scanned tables, reconstructs row and column structure, and produces an XLSX where the formerly image-based data is now fully editable, sortable, filterable, and ready for Excel formulas and Power BI ingestion. OCR supports English, German, French, Spanish, Italian, Portuguese, Dutch, Polish, and other major business languages without extra configuration. Ideal for digitizing legacy paper invoices, scanned bank statements, photographed receipts, historical accounting records, and any paper-origin financial content that needs to flow into modern Microsoft 365 analysis workflows.
What output format does the action produce: .xlsx or legacy .xls?+
The output is always the modern Microsoft Excel .xlsx format (Office Open XML standard, ECMA-376 / ISO 29500). This is fully compatible with Microsoft Excel 2007 and later, Excel Online (Microsoft 365), Excel for Mac, Power BI Desktop (direct import), Google Sheets (with conversion), LibreOffice Calc, Apple Numbers, and all major spreadsheet applications. The legacy binary .xls format is not produced. The XLSX file ships with proper MIME type application/vnd.openxmlformats-officedocument.spreadsheetml.sheet for correct handling by SharePoint, Outlook, Microsoft Teams, OneDrive, and Power BI dataset connections. The XLSX format also enables features that .xls does not support: million-row tables, conditional formatting, and modern pivot table engines.
How does this compare to manual Smallpdf, iLovePDF, Adobe Acrobat, or Tabula conversion?+
Online tools like Smallpdf, iLovePDF, Adobe Acrobat online, Nitro PDF, Soda PDF, and the open-source Tabula desktop application require manual file upload one PDF at a time with daily free-tier limits (typically 2-3 conversions per day on free tier) or paid subscriptions ($9-20/month per user). The PDF4me Power Automate action automates the same table extraction at scale: triggered by SharePoint file uploads, Outlook email attachments containing invoices or statements, Microsoft Forms responses, Dataverse record updates, OneDrive folder changes, or scheduled batch flows, and chains directly into downstream Microsoft 365 actions including Excel Online Add a row, Power BI dataset refresh, Dynamics 365 Finance record creation, Business Central, SharePoint metadata updates, and any accounting system via HTTP. Use this when you process dozens to thousands of PDFs per month consistently across the Microsoft 365 ecosystem rather than one-off ad-hoc conversions.

相关操作

在其他平台上执行相同任务

获取帮助