Zum Hauptinhalt springen

Lesen PDF Metadaten in Make: Dokumentinformationen abrufen

PDF4me Die Dokumenteninformationen abrufen ist eine Make Modul, das ein PDF Es liefert 11 Metadatenfelder – Titel, Autor, Seitenzahl, Dateigröße und mehr – ohne die Quelldatei zu verändern. Nutzen Sie es, um Dokumente anhand von Schlüsselwörtern zu sortieren, zu große Uploads vor der Konvertierung zu blockieren oder Archivdateien automatisch in eine Tabellenkalkulation zu katalogisieren.

Was dieses Modul bewirkt

PDF4me - Dokumentinformationen abrufen liest die Metadaten und Struktureigenschaften von beliebigen PDF ohne die Quelldatei zu verändern. Es gibt Folgendes zurück: 11 einzelne Felder - Titel, Autor, Betreff, Stichwörter, Urheber, Produzent, Erstellungsdatum, Änderungsdatum, Seitenzahl, Dateigröße und PDF-Version - als zuweisbare Token, die Sie direkt einbinden können Make Filter, Router, Datenbankmodule oder Benachrichtigungsschritte. Nutzen Sie es, um die Dokumentgröße vor einem Konvertierungsschritt vorab zu prüfen, Dateien nach Autor oder Abteilungs-Keyword zu sortieren, Katalogarchive automatisch in Airtable oder Google Sheets hochzuladen oder die Seitenzahl in einer Genehmigungs-E-Mail anzuzeigen, damit die Prüfer wissen, was sie freigeben.

Verwandte Blog-Beiträge
Zu dieser Funktion gibt es noch keinen Blogbeitrag – folgt in Kürze.
Schauen Sie sich in der Zwischenzeit im PDF4me-Blog Tutorials und Arbeitsabläufe für alle Plattformen an.
Besuchen Sie den Blog

Authentifizierung Ihres API Anfrage

Jeder PDF4me Modul in Make erfordert eine gültige VerbindungErstellen oder wählen Sie einen Eintrag aus, der Ihre Daten enthält. PDF4me API Schlüssel, damit das Szenario den Metadatendienst sicher aufrufen kann.

Wichtige Fakten, die Sie nicht verpassen sollten

Streng zerstörungsfrei

Das Modul liest die Dateistruktur und gibt Daten zurück. Keine Bytes im Original. PDF werden geschrieben, neu angeordnet oder gelöscht. Das von Ihnen eingelesene Dokument ist identisch mit dem, das weiterverarbeitet wird – es kann bedenkenlos auf Produktionsdateien ausgeführt werden, ohne dass das Risiko besteht, den Quellcode zu verändern.

11 strukturierte Felder zurückgegeben

Titel, Autor, Betreff, Stichwörter, Urheber, Produzent, Erstellungsdatum, Änderungsdatum, Seitenzahl, Dateigröße und PDF-Version sind alle als einzelne zuordnungsfähige Token im Ausgabepaket verfügbar – es ist kein Parsing-Schritt oder eine benutzerdefinierte Funktion erforderlich, um einzelne Werte zu extrahieren.

Direkt in die Routing-Logik einbinden.

Kartenseitenzahl, Autor oder Schlüsselwörter direkt in eine Make Filtern oder Routern Sie Ihr Szenario anhand von Dokumenteigenschaften. Leiten Sie übergroße Dateien an einen Ablehnungspfad weiter und senden Sie abteilungsbezogene Dokumente. PDFs in passende Ordner verschieben oder bereits katalogisierte Dokumente mit einer Duplikatsprüfung überspringen.

Das Modul „Dokumentinformationen abrufen“ von PDF4me zeigt die Verbindung zu testuser01@pdf4me.com an, die Datei ist auf „Mit Dropbox verknüpfen“ eingestellt – Option „Datei herunterladen“ sichtbar, der Dateiname wurde aus Schritt 7 übernommen und das Dokument aus Schritt 7 übernommen.

Wählen Karte unter Datei, dann Karte Dateiname Und Dokumentieren aus dem Schritt, der die heruntergeladen hat PDFDie

Parameter

Erforderlich: Verbindung, Dateiname, Und Dokumentieren. Satz Datei Zu Karte Zuerst müssen beide Eingabefelder sichtbar werden. Für das Dokument werden Binärdateien benötigt – eine URL oder ein Dateiname allein reicht nicht aus.

ParameterErforderlichWas es tutBeispielzuordnung
ConnectionRequiredPDF4me API connection used by the scenario to authenticate metadata requests. Click Add and paste your API key if connecting for the first time.Your PDF4me connection
FileRequiredDetermines how the PDF is supplied. Choose Map to wire File Name and Document from a prior module's output - the most common setup for multi-step scenarios.Map
File NameConditionalFilename with .pdf extension from the source module, required when File is set to Map. Map from the file name field of your Dropbox, Drive, or OneDrive step.7. File Name
DocumentConditionalBinary PDF content from the download step, required when File is set to Map. Must be actual file bytes - not a URL or a filename string. Map from the Data field of Dropbox or the File Content field of SharePoint.7. Data

Wie lese ich? PDF Metadaten in Make?

  1. Hinzufügen PDF4meDokumentinformationen abrufen zu Ihrem Szenario.
  2. Wählen Verbindung (oder klicken Sie hier) Hinzufügen um einen mit Ihrem API Schlüssel).
  3. Unter Datei, wählen KarteDie
  4. Karte Dateiname Fügen Sie im Dateinamenfeld aus Ihrem Download-Schritt Folgendes hinzu: .pdf Verlängerung.
  5. Karte Dokumentieren zum Binärdatenfeld aus demselben Schritt (normalerweise Daten in Dropbox oder Dateiinhalt in SharePoint/OneDrive-Modulen).
  6. Speichern und klicken Einmal ausführenErweitern Sie die Dateiinformationen Ausgabebündel – jedes der 11 Metadatenfelder ist ein separates, zuordnungsfähiges Token, das in jedes nachfolgende Modul eingebunden werden kann.

Welche Metadatenfelder gibt dieses Modul zurück?

Das Modul gibt einen Wert zurück Dateiinformationen Das Paket enthält 11 einzelne Metadatenfelder – jedes ist unabhängig zuordenbar. Titel, Autor, Thema, Schlagwörter, Urheber und Produzent stammen aus dem PDF-Dokumentinformationswörterbuch (Abschnitt 14.3.3 des PDF 32000-1:2008 Spezifikation), während Erstellungsdatum und Änderungsdatum folgen ISO 8601 Zeitstempelformatierung.

FeldTypWas es enthält und wie man es verwendet
TitleStringDocument title as stored in the PDF metadata dictionary. Empty string if not set. Map into an Airtable record or email subject line.
AuthorStringName of the person or system that created the document. Use in a Router to send files to department-specific folders based on creator.
SubjectStringSubject or description field from the document properties dialog. Often populated by enterprise document management systems.
KeywordsStringComma-separated keyword tags embedded by the author. Filter on this field in a Make Filter to route to the correct team or project.
CreatorStringApplication that originally created the document - for example, Microsoft Word, Adobe InDesign, or a PDF library name.
ProducerStringPDF conversion library or printer driver that generated the final PDF bytes. Different from Creator when a Word file was later exported to PDF.
Creation DateDateTimeISO 8601 timestamp of when the document was first created. Use in a Filter to process only documents created within a target date range.
Modification DateDateTimeISO 8601 timestamp of the last save. Compare to a stored baseline date to detect documents that have changed since last processing.
Page CountIntegerTotal number of pages. Use in a Filter to block files over a page limit or route short vs. long documents to different processing branches.
File SizeIntegerFile size in bytes. Gate on this before a conversion or email step - reject oversized files and notify the submitter to re-upload a compressed version.
PDF VersionStringPDF specification version (e.g. 1.4, 1.7, 2.0). Verify conformance before routing to a PDF/A archive that requires a specific version.

Wann sollte ich „Dokumentinformationen abrufen“ verwenden?

Typische KonfigurationenCommon Make scenario patterns that use Get Document Information to drive routing, validation, or cataloging.
Dokumente anhand von Schlüsselwörtern in Abteilungsordner weiterleiten
  1. Ein neues PDF landet in einem gemeinsam genutzten Dropbox-Posteingangsordner und löst das Szenario aus.
  2. Dokumentinformationen abrufen Extrahiert das Metadatenfeld „Schlüsselwörter“.
  3. A Make Der Router prüft den Wert des Schlüsselworts anhand der Abteilungsnamen (Finanzen, Recht, Personalwesen).
  4. Jede Niederlassung lädt die Datei in den entsprechenden Teamordner in Google Drive hoch.
  5. Eine Slack-Benachrichtigung informiert das Zielteam über den Dokumenttitel und die Seitenzahl.
Vorverarbeitungs-Validierungsgate
  1. Das Absenden eines Formulars löst das Szenario mit dem zugehörigen Anhang aus. PDFDie
  2. Dokumentinformationen abrufen Gibt Seitenzahl und Dateigröße zurück.
  3. Ein Filter verhindert, dass Dateien, die größer als 50 MB oder kleiner als 1 Seite sind, weiterverarbeitet werden.
  4. Gültige Dokumente werden an das Konvertierungs- oder E-Signatur-Modul weitergeleitet.
  5. Abgelehnte Dateien lösen eine E-Mail-Antwort aus, in der der Einreicher aufgefordert wird, eine kleinere Version erneut hochzuladen.
Automatische Katalogisierung neuer Archiv-Uploads in Airtable
  1. Ein geplantes Szenario durchläuft neue Abschnitte PDFs in einem OneDrive-Archivordner.
  2. Dokumentinformationen abrufen Extrahiert alle 11 Metadatenfelder pro Datei.
  3. Ein Airtable-Modul erstellt einen neuen Datensatz mit Titel, Autor, Erstellungsdatum und Seitenzahl.
  4. Der Datensatz verlinkt zurück zur OneDrive-Datei-URL, um den Zugriff mit einem Klick zu ermöglichen.
  5. Ein Filter überspringt Dateien, die bereits einen entsprechenden Airtable-Datensatz haben, um Duplikate zu vermeiden.

Praktische Tipps

Structural fields are always populated
Page Count, File Size, and PDF Version come from the file structure itself, not optional metadata, so they are never empty even on documents with no author-supplied properties.
Empty strings are normal, not errors
Author, Title, Subject, Keywords, Creator, and Producer return empty strings when the document creator never set them. The module does not fail on missing metadata.
File Name and Document only appear after choosing Map
The panel starts with just Connection and File visible. Select Map under File to reveal the File Name and Document mapping fields.
Run Repair PDF first on suspect files
If incoming PDFs may be corrupted or come from an unreliable source, run Repair PDF before this module so metadata reads cleanly instead of returning empty fields.
Creator and Producer are often different
Creator names the authoring application, Producer names the library or driver that generated the final PDF bytes, useful for distinguishing a Word export from a native PDF tool.

Spickzettel

FeldWert
FileMap
File Name7. File Name
Document7. Data
Always populatedPage Count, File Size, PDF Version
Can be emptyTitle, Author, Subject, Keywords, Creator, Producer

Häufig gestellte Fragen

Does this module modify the PDF in any way?+
No. Get Document Information performs a strictly read-only inspection of the file. The module reads the document dictionary and cross-reference table to extract metadata values and returns them as a structured output bundle. No bytes in the source PDF are written, reordered, or removed during this process. The file you map into this module is identical byte-for-byte to what gets passed into any downstream module - you can safely run it on production documents without any risk of altering the originals.
What if the PDF has no metadata fields set?+
Page Count, File Size, and PDF Version are always populated because they are derived from the document structure itself rather than optional metadata fields. Author, Title, Subject, Keywords, Creator, and Producer will return empty strings when the document creator did not fill them in - this is common with PDFs generated by automated systems or older software that strips document information during export. The module does not error on empty metadata; it simply returns empty strings for those specific fields.
Can I use the Page Count output to branch my scenario conditionally?+
Yes - this is one of the most practical uses of the module. Map the Page Count field into a Make Filter condition placed after Get Document Information. For example, set the Filter to pass only when Page Count is greater than 20 and route those to a PDF splitting workflow, while shorter documents continue to an email step. You can similarly use File Size to reject oversized uploads before a conversion step, or check Creation Date to process only documents created within a target date range such as the current calendar month.
What PDF versions and variants does this module support?+
The module reads metadata from all common PDF versions - PDF 1.0 through PDF 2.0 - as well as PDF/A archival variants including PDF/A-1, PDF/A-2, and PDF/A-3. The PDF Version field in the output tells you exactly which specification the document conforms to (for example 1.4, 1.7, or 2.0). This is particularly useful in compliance workflows where you need to verify that a document meets a required PDF version before it is sent to a regulated archive or long-term storage system.
How do I access a specific metadata field in a downstream module?+
After the module runs, expand the File Info output bundle in the Make mapping panel - each of the 11 fields (Title, Author, Subject, Keywords, Creator, Producer, Creation Date, Modification Date, Page Count, File Size, PDF Version) appears as its own individually selectable token. Click any token to insert it into a field in the next module - for example, map Title into the Subject field of a Gmail step, Page Count into a Filter condition, or Author into an Airtable record column. No custom function or JSON parsing is required.

Verwandte Aktionen

Dieselbe Aufgabe auf anderen Plattformen

Hilfe erhalten