A — Access Control
Access control is the set of rules that decide who can view, edit, or share a specific document. Instead of leaving every file open to an entire team, access control limits sensitive records, such as contracts or financial statements, to the people who genuinely need them. This reduces the risk of a document being changed, forwarded, or seen by someone who shouldn't have reached it in the first place.
B — Bates Numbering
Bates numbering is the practice of stamping each page of a document with a unique, sequential identifier, most often used in legal and litigation document management. It allows large sets of documents, such as case files or discovery materials, to be referenced precisely and consistently across a legal team, a court filing, or an audit, without confusion over which version or page is being discussed.
C — Cloud Storage
Cloud storage means keeping documents on secure remote servers instead of a single computer or office server. Files are accessed over the internet rather than a local network, so a business isn't limited to one machine or one location. For a document management system, cloud storage is usually what makes remote access, backup, and multi-office collaboration possible without extra hardware.
D — Document Management System (DMS)
A document management system, or DMS, is software that helps a business store, organize, and retrieve digital documents from one centralized platform. It replaces scattered folders, personal drives, and email attachments with a structured system built around how documents are actually used, covering everything from everyday filing to access control, search, and how long records are kept before they're archived or deleted.
E — Encryption
Encryption converts a document into a coded format that can only be read by someone with the correct key or credentials. It protects files both while they're stored and while they're being transferred between a device and a server, so a document intercepted or accessed without authorization can't simply be opened and read. It's one part of a broader approach to document security.
F — Folder Structure
Folder structure is the hierarchy of folders and subfolders used to organize documents, typically by category, client, project, or department. A consistent folder structure makes it possible for anyone on a team to predict where a document should live, rather than relying on memory or personal habit. In a document management system, this structure is often applied automatically as files are added.
G — Governance
Governance, sometimes called information governance, refers to the policies a business sets for how documents are created, organized, accessed, retained, and eventually deleted. It ties access control, retention, and security together into one consistent approach, so decisions about how a document is handled don't depend on whichever employee happens to be managing it that day.
H — Hash Value
A hash value is a short string of characters generated from a document's contents, acting like a fingerprint for that exact file. If even a single character in the document changes, its hash value changes too. Document management systems sometimes use hash values to confirm a file hasn't been altered since it was stored, which matters for records that need to stay verifiably unchanged.
I — Indexing
Indexing is the process of tagging documents with searchable details, such as their title, category, date, or client, so they can be found quickly later. Rather than requiring someone to open and read every file, indexing lets a document management system return the right result the moment a term is searched, turning a large, unsorted library of files into something genuinely usable.
J — Journaling
Journaling is the practice of automatically capturing and retaining a copy of business communications or documents as they're created, often for compliance or recordkeeping purposes. Rather than depending on individual employees to save or forward copies, journaling ensures a consistent record exists independently, which can matter for regulated industries or for resolving disputes about what was actually sent or agreed.
K — Knowledge Base
A knowledge base is an organized collection of reference documents, guides, or answers that a business maintains so information doesn't live only in one person's head. Inside a document management system, a knowledge base often sits alongside working documents like contracts and records, giving a team a consistent place to find explanations, policies, and answers to recurring questions.
L — Document Lifecycle
The document lifecycle describes the stages a document passes through, from creation and review to storage, active use, and eventual archiving or deletion. Thinking about the full lifecycle, rather than just where a file is stored today, helps a business plan for retention rules and access changes, instead of letting old documents accumulate indefinitely with no clear end point.
M — Metadata
Metadata is descriptive information attached to a document, such as its category, client, author, or date, rather than the content of the document itself. It's what allows a document management system to organize and search for a file without needing to open it. Good metadata is often applied automatically on upload, so consistency doesn't depend on someone filling in details by hand.
N — Network Drive
A network drive is a shared storage location on a local office network, traditionally used to store business files before cloud-based systems became common. Network drives require a connection to that specific network, which makes remote access difficult and gives little structure beyond whatever folder habits a team happens to follow, a limitation many businesses eventually replace with a proper document management system.
O — OCR (Optical Character Recognition)
OCR, or optical character recognition, is technology that converts a scanned image of a paper document into machine-readable, searchable text. Without OCR, a scanned contract or invoice is just a picture, unsearchable and hard to reuse. With it, the same document becomes something a document management system can index, search, and even extract information from automatically.
P — Permissions
Permissions are the specific settings that define what a particular person can do with a document, such as view it, edit it, or share it with someone else. They're the mechanism behind broader access control policies, applied at the level of an individual file or folder, so two people can have very different levels of access to the exact same set of documents.
Q — Query
A query is the search request someone enters into a document management system to find a specific file, such as a client name, document type, or date range. A well-indexed system can match a query against metadata and content alike, returning the right document in seconds rather than requiring someone to manually browse through folders one at a time.
R — Retention Policy
A retention policy sets the rules for how long a document is kept before it's archived or permanently deleted. Retention policies often reflect legal, financial, or industry requirements, and applying them consistently across a document management system reduces the risk of keeping records too long, disposing of them too early, or handling different document types inconsistently across a business.
S — Search
Search is the ability to locate a document by name, keyword, category, or content, rather than browsing folder by folder to find it. In a document management system, search typically works against both a file's metadata and its indexed content, which is why documents that are properly categorized and tagged tend to surface faster and more reliably than ones dropped into a generic folder.
T — Taxonomy
A taxonomy is the overall system of categories and subcategories used to classify documents, such as organizing files by department, client, project, or document type. A clear taxonomy is what makes a folder structure predictable and consistent as a business grows, rather than letting categories multiply informally every time someone isn't sure exactly where a new document belongs.
U — Unstructured Data
Unstructured data refers to information that doesn't fit neatly into rows, columns, or a fixed format, such as contracts, emails, scanned letters, and reports. Most business documents are unstructured data, which is why generic databases handle them poorly. A document management system is built specifically to organize, tag, and search this kind of file, rather than the structured records a typical database expects.
V — Version Control
Version control is the practice of tracking changes to a document over time, so a team always knows which copy is current instead of comparing several similarly named files. Without it, edits often get made to the wrong copy, or two people unknowingly work on separate versions of the same document, creating confusion about which one should actually be considered final.
W — Workflow
A workflow is the defined sequence of steps a document moves through, such as drafting, review, approval, and filing, often involving more than one person. Document management systems can support workflows by routing a file to the right person automatically and tracking where it currently stands, replacing informal chains of emails and manual follow-ups with a visible, repeatable process.
X — XML (Extensible Markup Language)
XML, short for Extensible Markup Language, is a text-based format used to structure data so it can be read by both people and software. Some document management systems use XML behind the scenes to store metadata or exchange document information with other systems, since its structured tags make it easier for different software to interpret the same data consistently.
Y — Year-End Archiving
Year-end archiving is the practice of moving a year's completed documents, such as invoices, contracts, or financial records, into long-term storage once that period closes. It keeps active folders focused on current work while still preserving older records for as long as a business's retention policy requires, rather than leaving every past document mixed in with what a team is using day to day.
Z — ZIP Archive
A ZIP archive is a single compressed file that bundles multiple documents together, making them easier to store, transfer, or download as one unit rather than many separate files. Businesses sometimes export a batch of records as a ZIP archive when sharing a large set of documents externally, such as during an audit, a legal request, or a client handoff.