BASALT · JOURNAL

How redaction actually works

Field notes on destroying content in PDFs, proving it is gone, and the ways redaction quietly fails. Written by the people who built the engine.

Software that requires an account, and software that does notWhat an account requirement means in practice for confidential document work, and where it is genuinely useful.An Acrobat alternative for Mac when redaction is the jobAn Adobe Acrobat alternative for Mac built around redaction: destroying text in the content stream, verifying the result, and proving it.What a solo practitioner actually needs from a PDF toolThe short list of operations that dominate solo and small firm document work, and how to judge a tool against it.Bates numbering in Acrobat and elsewhereWhat Acrobat's Bates feature does, what a registry adds, and why collision detection and a manifest matter on real productions.Cloud features, offline work, and what leaves your machineWhat Acrobat's cloud integration does, how to work offline, and what enforcing offline at the operating system level means.Compressing PDFs, predicted sizes and measured onesHow PDF compression actually works, why output size is hard to predict, and the guarantee worth asking for.Comparing two versions of a documentWhy layout diffs are usually useless, what a text level comparison shows instead, and what a blackline is for.Form filling, flattening, and what XFA breaksHow PDF forms actually work, why appearance regeneration matters, and the older Adobe format that most tools cannot open.Hidden layers, and content that is present but not drawnWhat optional content groups are, why a switched off layer is still in the file, and the difference between warning and removing.What native means on a Mac, and why it matters for documentsThe practical differences between a native macOS application and a cross platform one for document work, without the usual purity argument.OCR in Acrobat and on the MacHow Acrobat's OCR compares with the recognition built into macOS, what each is good at, and the redaction consequence of an OCR layer.Working out what your PDF software costs per useA short exercise for deciding whether a professional subscription is good value in your specific workflow.The redaction mistakes that happen with good softwareThe recurring errors in redaction work, none of which are caused by inadequate tools, and how each is caught.Why verification after redaction is not standardWhat most redaction tools check after performing a removal, why checking your own output is weak, and what independent verification means.Sanitize Document, and the step after redactionWhat Acrobat's Sanitize Document removes, why redaction alone is not enough, and how the two steps get separated in practice.Signing a PDF, and the two things called a signatureThe difference between a drawn signature and a cryptographic one, which you actually need, and what each proves.Acrobat's subscription against a one time licenceHow subscription and perpetual licensing actually compare over time, what each gets you, and when the subscription is the better deal.Software updates, lock in, and what happens if a vendor stopsHow update models differ between subscription and perpetual software, and the risks each carries for a professional workflow.Acrobat or Basalt, decided in five questionsA short decision procedure rather than a feature table, including the cases where Basalt is the wrong answer.Adding page numbers, headers, and footers to a PDF on a MacAdd page numbers to a PDF on a Mac: page box geometry, rotation traps, and why the number belongs in the content stream, not an annotation.Keeping PDFs readable for the long termWhat makes a document readable in twenty years, and the choices that quietly make it unreadable.Basalt and Acrobat redaction compared honestlyWhat Acrobat's redaction does well, where the two differ, and why verification before writing is the substantive distinction.Running the same PDF steps over and over without doing them by handBatch process PDF files on a Mac: why saved procedures beat manual repetition, and how to keep automation from becoming its own risk.Redacting hundreds of documents without doing it hundreds of timesWhat can safely be batched in redaction work, what cannot, and why applying one set of marks across many documents is dangerous.Bates numbering a large productionHow Bates numbering scales across thousands of pages, why collision detection matters, and what a manifest is for.Bates numbering on a Mac without AcrobatBates numbering on a Mac without Acrobat: what the numbers have to guarantee, the page geometry traps, and how to stamp a production correctly.The best PDF software for lawyers working on a MacThe best PDF software for lawyers on a Mac, judged by what a practice actually does: exhibits, Bates numbers, OCR, redaction, and provable results.A black box over text is not a redactionTo black out text in a PDF is to add a drawing instruction. The words stay in the content stream until something actually removes them.Why your PDF looks blurry after compressing itWhat compression actually did to the images, and how to get the size down without the damage.Why you cannot select text in a PDFThe three reasons text will not select, how to tell them apart, and what to do about each.What a certificate of redaction is, and what makes one worth havingA certificate of redaction should be independently verifiable evidence, not a receipt. What it must contain and how to check one without the tool.Someone sent you a redacted PDF. Here is how to check itHow to check if a PDF is really redacted: six passes that read the file instead of the rendered page, and what to do when one of them fails.What happens to your document when you upload it to an online PDF toolThe real online PDF redaction risk, explained technically: where your file goes, who can reach it, how long it lives, and when that matters.How to compare two PDFs and see exactly what changedCompare two PDFs for differences: why pixel comparison reports every page as changed, and why comparing the text is what answers the real question.Compressing a large scanned PDF without wrecking itWhich compression settings actually reduce a scanned document, which ones damage evidence, and why predicted sizes should be measured.How to compress a PDF on a Mac without wrecking the scansHow to compress a PDF on Mac: why the Reduce File Size filter destroys scans, what downsampling actually does, and how to hit a size limit safely.Getting a PDF under 10 MBA practical sequence for hitting a hard size limit without making the document unreadable.Cropping PDF pages on a MacHow cropping actually works in a PDF, why cropped content is still in the file, and when that matters.How to delete pages from a PDF on a MacThree ways to remove pages, including the free one already on your Mac, and what to check afterwards.What a discovery production actually weighsRealistic sizes for legal productions, why they are mostly scans, and how to plan machine time and disk before the deadline.Disk space, not memory, is the real cost of large PDFsWhy applications that handle large documents well use disk instead of RAM, what a working copy is, and how to avoid running out of space mid job.Duplicating and rearranging pages in a PDFRepeating pages, building a document from ranges, and assembling exhibits.Editing the bookmarks in a PDF on a MacEdit PDF bookmarks on a Mac: the document outline, named destinations, why merging breaks navigation, and what bookmark titles disclose.Extracting specific pages from a PDFHow to pull a range of pages into a new document without losing quality, and the difference between extracting and splitting.How redactions fail, and what each failure has in commonRedaction failures follow a handful of repeatable patterns. Here is the mechanism behind each one and the single property every failed redaction shares.Redacting student recordsWhat education records typically require removing before release, and the small-cohort problem that catches schools out.Filling in a PDF form on a Mac so the values stay putHow to fill out a PDF form on Mac so values do not vanish: AcroForm fields, appearance streams, NeedAppearances, and when to flatten before sending.Redacting every instance of a name across a long documentHow to search and redact all instances in PDF documents: catching split words, headers, annotations, form fields, and the OCR layer before you apply.What flattening a PDF actually does, and when you need itWhat flattening a PDF on Mac really does: form fields and annotations become page content, why it is not redaction, and when to flatten.Foxit alternatives for Mac usersWhat Foxit covers, where Mac users find the gaps, and how to choose a replacement.What free Mac PDF editors can and cannot doAn honest map of the free options on macOS, what each genuinely covers, and the specific things none of them do.PDF redaction under the GDPR: what erasure actually requiresGDPR PDF redaction explained at the file level: why a black box is not erasure, which hidden PDF structures retain personal data, and how to prove removal.What a PDF still carries after you think you cleaned itHidden data in PDF files: XMP packets, incremental updates, OCR layers, attachments, optional content groups, and thumbnails that survive a redaction.De-identifying PDFs under HIPAA without leaving tracesHIPAA PDF redaction at the file level: Safe Harbor versus Expert Determination, where PHI survives in a PDF, and how to verify removal before release.How large a PDF can you actually open on a MacWhy PDF page count matters less than file size, what actually consumes memory when a document opens, and measured numbers for a 273 MB scanned file.How to redact a PDF on a Mac so the text is actually goneHow to redact a PDF on Mac so the text is removed from the file, not just covered, and how to prove the removal before you send the document.A desktop alternative to iLovePDF for Mac usersThe trade between browser convenience and local processing, and what a desktop equivalent needs to cover.The location data hiding in your PDF's imagesPhotographs embedded in documents can carry GPS coordinates, and stripping document metadata does not remove it.Turning photos and scans into a single PDF on a MacHow to convert images to PDF on Mac: get page order, orientation, and page size right, and avoid re-encoding photos you already have as JPEG.Inserting pages into an existing PDFAdding pages from another document, inserting blanks, and keeping the result in the right order.How to see everything a PDF is actually carryingWhat is inside a PDF file beyond its pages: metadata, XMP, embedded files, hidden layers, OCR text, JavaScript, and earlier revisions.Basalt's large file numbers, and how to reproduce themThe measured performance figures for large documents, the machine they came from, and the caveats that go with a single run.Why Preview struggles with large scanned PDFsWhat makes Apple's Preview slow on big documents, when that matters, and what the limitation means for redaction work specifically.Searching a large PDF, and why it is usually fastWhat happens when you search a long document, why text search is cheap, and why searching a scanned file finds nothing until you OCR it.Why page 500 should open as fast as page 1The difference between rendering pages on demand and rendering them eagerly, and a simple test that tells you which your application does.Making a scanned PDF searchable on a MacWhat OCR adds to a scan, how to tell whether a document already has a text layer, and the redaction trap it creates.Merging many large PDFs into one productionWhat happens to size and structure when large documents are combined, why the result is often smaller than the sum, and what to preserve.How to merge PDF files on a Mac without losing anythingHow to merge PDF files on a Mac: what Preview does well, what it silently drops, and how to combine documents while keeping forms and bookmarks.Document privacy for nonprofits and small charitiesDonor data, beneficiary records and grant reporting, handled without a compliance department.Running OCR on a large scanned productionWhat OCR does to a scanned document, how long it takes on a large one, and the redaction trap that the added text layer creates.How to make a scanned PDF searchable on a MacHow to OCR a PDF on Mac: what an invisible text layer is, how it is drawn, why it leaks under redactions, and how to add one locally.Offline PDF editors for Mac that work with the network switched offAn offline PDF editor Mac users can actually verify: what offline means, how to prove an app has no network access, and how the real options compare.Before you upload a document to a PDF websiteA short checklist for deciding whether a specific document can go through a web based tool, and what to do when it cannot.How to password protect a PDF on a Mac, and what the password actually doesHow to password protect a PDF on Mac: user vs owner passwords, RC4 and AES encryption, and why permission flags are not security.Redacting in patent and IP mattersTrade secrets, prior art and the confidentiality tiers that make IP redaction unusual.Making a PDF readable by a screen readerWhat accessibility means in a PDF, why scans fail entirely, and the minimum worth doing.PDF editing software for Mac: what you can actually edit, and what you cannotPDF editing software for Mac, explained honestly: which operations are reliable, which are guesswork, and how to choose a tool for the work you do.PDF editors for Mac with a one time purchase, not a subscriptionA PDF editor Mac one time purchase means no subscription. What a perpetual license really buys, what to ask before paying, and how options compare.PDF Expert alternatives on the MacWhat PDF Expert does well, where its subscription model bites, and how to judge an alternative against the work you actually do.When a PDF refuses to get smallerWhy compression sometimes does nothing, or makes the file bigger, and what to do instead.When PDF fonts look wrong on another machineFont embedding, substitution, and why a document can look perfect on your Mac and wrong everywhere else.When a PDF form will not save your entriesWhy typed values disappear, the difference between filling and flattening, and the format that will not open at all.Hidden layers in a PDF, and why switching one off is not removing itPDF layers are optional content groups. A layer switched off is invisible on screen and in print, but the content stays and the text still extracts.Page count and file size are almost unrelatedWhy a short PDF can be heavier than a long one, what actually drives document weight, and which number to use when planning work.When a PDF password will not workThe two different passwords a PDF can carry, why the right one can still fail, and what no legitimate tool will do.When a PDF prints blank or wrongWhy a document can look correct on screen and print empty, and the fix for each cause.A redaction workflow for law firms that survives reviewA PDF redaction workflow for law firms: intake hashing, Bates endorsement order, reason codes, verification, and the privilege log that has to match.PDF redaction not working: the five things that are usually wrongPDF redaction not working usually comes down to five causes, from annotation marks to surviving OCR layers. How to tell which one you have.PDF redaction software for Mac: what to look for before you trust oneChoosing PDF redaction software for Mac: the criteria that separate real content removal from a drawn black box, and how to test any tool yourself.Your PDF is telling everyone who wrote it, and whereWhen a PDF shows author name, it comes from the info dictionary and the XMP packet. Where those fields come from and how to strip them for real.Exporting PDF pages as imagesTurning pages into PNG or JPEG, choosing a resolution, and what you lose in the conversion.When a PDF is too big to emailThe real attachment limits, which fixes actually work, and the order to try them in.What to do when a PDF is too large to openWhy large PDFs fail to open, how to tell whether the file or the application is at fault, and what to try before assuming the document is corrupt.The PDF jobs that keep interrupting your actual workPDF tools for Mac judged by the real jobs: combining exhibits, splitting a bundle, making a scan searchable, numbering a production, cleaning a file.PDF work with no internet connectionWhich PDF operations genuinely need a network, which never did, and how to work on a plane or in a secure facility.When a PDF will not open on your MacHow to tell a damaged file from a viewer problem, and the repair that works on genuinely broken documents.Free PDF apps and what they trade awayWhy some capable PDF applications are free, what that usually means, and how to evaluate one.How to permanently delete text from a PDFHow to permanently delete text from a PDF: excise it from the content stream, strip every parallel copy, rewrite the file, and verify the bytes.Why redacting a PDF in Preview is not redactionIf you redact a PDF in Preview on Mac, the text stays in the file under the black box. Here is what Preview really does and why the words survive.Preview versus a dedicated PDF applicationAn honest account of how far the built in option goes, and the specific point at which it stops.Building a privilege log from the redactions you already madeA privilege log should be a byproduct of review, not a second document. How to build one from redaction marks and keep it true to the production.Long document jobs need progress and a cancel buttonWhy the interface around a slow operation matters as much as its speed, and what to expect from a tool running a sixteen minute job.Handling redactions in a public records requestExemptions, the duty to release what remains, and why records offices are the most common source of published redaction failures.Recovering a damaged PDFWhat usually breaks in a damaged file, which repairs work, and when to stop trying.How long it takes to redact a 1,000 page PDFMeasured timings for redacting text and scanned pages, why scans are far slower, and how to plan a large production around it.Redacting a bank statement so the account number is really goneHow to redact a bank statement properly: every number is extractable text, partial masking leaks, and how to verify the account number is gone.Redacting client names from work you want to show publiclyHow to redact client names from documents for a portfolio or case study, including logos, letterheads, file names, and image metadata that identify.Redacting documents for a public records releaseHow to redact documents for FOIA releases: citing an exemption per mark, keeping withholdings consistent across a set, and logging what was removed.Redacting workplace investigation filesProtecting witnesses while producing a usable report, and what happens when the file is later disclosed.Redacting in family law mattersChildren, addresses and financial disclosure, where the safety stakes are higher than in most litigation.Redacting financial statements and bank recordsAccount numbers, balances and counterparties, and why partial masking often fails.Redacting HR and employee documentsWhat typically needs removing from personnel files before they are shared, and the parts people miss.Redacting images, photographs, and scans inside a PDFHow to redact an image in a PDF so the pixels are destroyed and re-encoded, not covered, including scans, DCTDecode JPEGs, and stale thumbnails.Redacting immigration and visa documentsPassports, sponsor details and third parties, and the specific risks of documents that identify people by status.Redacting insurance claim filesClaim files mix medical, financial and third party data, and what has to come out before each recipient.Redacting medical billing and claims dataWhere identifiers hide in billing documents, and why codes can identify a patient on their own.Redacting medical records without leaving the identifiers behindHow to redact medical records so identifiers are actually gone: what identifies beyond the name, where it hides, and how to verify removal.Redacting the same area across many pages at onceHow to redact multiple pages in a PDF at once: apply one mark to a page range for repeated headers, footers, watermarks, and table columns.What to strip from a PDF before you email it outside the firmA checklist to remove sensitive information from PDF before sending: metadata, comments, form values, attachments, and prior generations.How to redact a PDF without Adobe AcrobatHow to redact a PDF without Adobe Acrobat on a Mac: what Preview can and cannot do, why browser tools upload your file, and the procedure that works.Redacting property and conveyancing documentsClient identity, funding detail and the fraud exposure specific to property transactions.Redacting scanned documents, where the text you cannot see is the problemHow to redact a scanned PDF safely: destroy the image pixels and the invisible OCR text layer beneath them, then verify by extracting text from the output.Redacting a Word or Excel file that became a PDFHow to redact a Word document converted to PDF: document properties, surviving comments and tracked changes, hidden rows, and speaker notes.Your redacted PDF still shows the text when you copy it. Here is whyIf your redacted PDF still shows text on copy and paste, the mark never removed the words. The mechanism, and how to fix it properly.The checklist to run before you file a redacted documentA numbered redaction checklist to run before filing: what to mark, what to strip, how to verify the output, and what to keep for the record.Embedded files, the attachment inside your PDF nobody looks atHow to find and remove attachments from a PDF: embedded file streams, file attachment annotations, and why page review never surfaces them.Removing a black box from a PDF, and what it means that you canHow to remove black box from PDF marks on your own files, why it usually takes one command, and what that proves about whether the file is safe.How to remove a password from a PDF you are allowed to openHow to remove a password from a PDF on Mac when you know it, why cracking is the only alternative, and why online unlockers are a data transfer.How to remove metadata from a PDF on a MacHow to remove metadata from a PDF on Mac: what the Info dictionary, XMP packets, and incremental updates actually hold, and how to strip them for real.Removing personal information from a PDF before you share itHow to remove personal information from PDF files: names, IDs, signatures, plus the metadata and photo EXIF that survive an ordinary redaction.Reordering, rotating, and deleting PDF pages on a MacHow to reorder pages in a PDF on Mac, rotate sideways scans permanently, delete pages, and insert blanks, without losing bookmarks or form fields.Making a document process repeatableWhy the same sequence of steps done by hand goes wrong, and what to automate first.Reproducible redaction, and why a byte-identical result mattersReproducible redaction means the same file plus the same marks yields byte-identical output, so anyone can re-run the marks and confirm what was filed.Reversing the page order of a PDFFlipping a document end to end, usually because a scanner fed it backwards.How to rotate PDF pages on a Mac and make it stickRotating pages in Preview, why rotation sometimes does not save, and how to rotate a whole document at once.Why your scanned PDF is not searchableHow to confirm a document has no text layer, what OCR adds, and why searching an OCRed scan can still miss pages.Why scanned PDFs are so much larger than text PDFsThe size difference between a text PDF and a scanned one, what drives it, and which settings actually reduce it without destroying evidence.A desktop alternative to Sejda and similar toolsWhere browser based editors hit their limits, and what a local equivalent needs to cover.How to sign a PDF on a Mac, and what a signature on a page is worthHow to sign a PDF on a Mac: signature images versus cryptographic digital signatures, and which one your workflow actually needs.An offline alternative to Smallpdf and the browser toolsWhat happens to a file you upload to a browser based PDF tool, when that matters, and what an offline equivalent looks like.Splitting a large PDF by size rather than page countWhy page count is the wrong unit when splitting scanned documents, how size based splitting works, and when splitting beats compression.Splitting a PDF into equal partsSplitting by page count, by size, and at chapter boundaries, and which to use when.How to split a PDF on a Mac, five ways that actually come upHow to split a PDF on Mac: extract page ranges, break at chapters, split every N pages, or stay under an email size limit, without cloud uploads.Reading a PDF in chunks instead of loading itWhat it means to stream a document rather than buffer it, why the choice decides your maximum file size, and how to verify which one an application does.A checklist for moving off AcrobatHow to work out whether an alternative covers what you actually use, and which Acrobat features have no equivalent worth pretending about.Removing restrictions from a PDF you ownThe difference between encryption and permissions, and how to remove restrictions legitimately.How to check whether a PDF was really redactedHow to check if a PDF is redacted: the copy test, extraction with an independent parser, metadata inspection, and reading the update chain.Verifying a digital signature on a PDFWhat a signature actually proves, what it does not, and how to check one you have been sent.Verifying a redaction on a very large fileWhy verification is cheap even on huge documents, what a verifier actually checks, and why it should run before the file is written.Adding a watermark to a PDF on a Mac that survives the printAdd a watermark to a PDF on a Mac: why annotation watermarks vanish on paper, and how page content watermarks survive printing.Why PDF applications run out of memory on large filesThe architectural reason large PDFs exhaust memory, the difference between buffering and streaming a file, and how to tell which one an application does.