APPX email management for state archives

Appraise, process, and publish email—without taking people out of the process.

QEPS turns collected email into searchable archives, working collections, and controlled public access. Automation handles the repetitive first pass. Archivists review the decisions that matter.

Runs where your records belong: inside your secure network, on infrastructure you manage.

QEPS operating model Archive to access
  1. 01
    Archive Ingest, scan, checksum, parse, index
  2. 02
    Collect Choose accounts, packages, and email
  3. 03
    Process Classify, restrict, retain, label, redact
  4. 04
    Publish Release approved collections for search

One chain of custody. The original email, extracted evidence, processing results, and review history stay connected.

EML MBOX PST ZIP 7ZIP TAR GZIP Gmail metadata

What QEPS actually does

Four parts of the email-records job, connected.

QEPS separates the source archive from the collection being processed. That gives archivists room to survey what arrived, select what belongs together, apply policy, and release only what is ready.

01

Searchable email archives

Receive and survey what the agency sent.

Build an archive around entities, email accounts, ingested packages, and individual messages. Search and inspect the source before deciding how it should be processed.

  • Desktop, server, shared drive, S3, and URL sources
  • Virus checking and SHA-2 checksums
  • Metadata and attachment extraction
  • Solr indexing for keyword and faceted search
02

Processing collections

Define the body of email that belongs together.

Create collections with groups, subgroups, and members, then associate the archives, accounts, packages, or individual messages that belong in each processing group.

  • Reusable processing libraries
  • Collection-level rules and overrides
  • Background processing with progress tracking
  • Review queues for unresolved results
03

Published collections

Make approved email available to the public.

Publish or unpublish a completed collection without opening the working repository. The public sees searchable, approved records and access copies.

  • Keyword, field, facet, and natural-language search
  • Labels, dates, addresses, domains, and entity tags
  • PDF display for email records
  • Attachment access copies, download, and print
04 Roadmap

Preservation export

Prepare permanent email for preservation.

The QEPS roadmap includes export of permanent-retention email, originals and normalized email, separate attachments, and structured metadata for preservation workflows.

  • Original PST, MBOX, EML, and MSG sources
  • Normalized EML packaged as MBOX
  • Attachments as separate files
  • Preservation metadata in XML

Keep the source. Shape the work.

An archive is what arrived.
A collection is what you decide to process.

QEPS keeps those ideas separate, then links them. An account or package can be surveyed in the archive and brought into the right collection without losing its source context.

Archive hierarchy
  1. Entity Agency or organization
  2. Email account Collected account
  3. Ingested package Transfer or container
  4. Individual email Message and attachments
Collection hierarchy
  1. Collection Processing objective
  2. Group / subgroup Organizational structure
  3. Member Selected sources
  4. Email Processed record

Inside the email archive

Bring an entity’s email together.
Make it discoverable.

An email archive is a searchable repository defined for one entity. It contains the emails sent and received through every email account owned by that entity. The entity is usually a person, but it can also be an organization or another defined subject of archival work.

Archive scope

One entity. Every account. One searchable repository.

Keep the entity’s correspondence together while preserving account and source context across the archive.

Person Organization Other entity
What it contains
  1. Email accountsEvery account owned by the entity
  2. Sent + received messagesCorrespondence across those accounts
  3. Searchable evidenceMessage content, metadata, attachments, and threads

Ingest the source

Five steps from files to a searchable archive.

To create an archive, the entity’s email must first be ingested. QEPS guides the work from a desktop upload through extraction, validation, parsing, and indexing.

  1. 01 — Upload

    Build a source list and send it to the server.

    A user-friendly drag-and-drop interface lets staff select files from any storage device accessible from their desktop, review the list, and initiate the upload.

    PSTMBOXEMLZIP
  2. 02 — Extract

    Turn containers into individual email.

    QEPS extracts messages from ZIP, PST, and MBOX files and saves each email in EML format.

  3. 03 — Validate

    Check files and establish fixity.

    Extracted email is virus checked and assigned checksums for fixity purposes.

  4. 04 — Parse

    Capture the evidence that makes email useful.

    The system extracts metadata and attachments, and identifies the relationships that form email threads.

  5. 05 — Index

    Make the archive ready for discovery.

    Extracted metadata and email content are used to build a searchable index of all email for the entity.

Archival discovery

Explore what arrived before you decide what it means.

The archive search is designed for archivists conducting discovery. Results provide a focused path into the correspondence, its attachments, and the relationships between messages.

SearchKeywords in subject lines and message bodies

NarrowDate ranges, correspondent addresses, and facets

Review resultsA relevance-ordered table of selected emails

Open evidenceView individual emails and their attachments

Related-message exploration

When an archivist opens an email, QEPS also shows a table of the 50 additional archive emails most like it. Any of those emails can be selected and opened to continue discovery.

Inside the email collection

Group related email.
Process it as one body of work.

An email collection brings together related email from one or more archives for processing as a group. When every email is complete, the collection may be published, exported for preservation, or both.

Collection structure

One member—or many members in an organizational hierarchy.

Each member can draw selected email from one or more email archives, keeping the processing objective separate from the source archive.

One or more archivesMember hierarchySelected email
Release gate

Process every email before it leaves the collection.

A collection cannot be published or exported for preservation until all of its email has been processed.

Processed collectionPublish and/or preserve

The processing objective

Review, weed, classify, protect, and prepare each email.

  • Review and weed
  • Redact when necessary
  • Add classification labels
  • Assign record class
  • Assign access restriction
  • Assign record retention

Processing options

Apply the policy signals that matter to your program.

QEPS offers multiple processing methods. Each can make a policy assignment and require an archivist’s review when the rule calls for it.

01 — Pattern matching

Find defined sensitive patterns.

Search for user-defined patterns such as credit-card, Social Security, and driver-license numbers. A match can also be automatically redacted.

02 — Lexicon processing

Find the words and phrases that carry meaning.

Search against a user-defined lexicon—such as a legal lexicon—and act when an email exceeds the specified relevance threshold.

03 — Address matching

Find specific email addresses.

Match designated addresses within an email. The matching address can be automatically redacted when policy requires it.

04 — Domain matching

Find domains wherever they appear.

Search for domain names in email addresses, URLs, and hyperlinks. A matching domain can be automatically redacted.

05 — Hyperlink matching

Find specific links.

Search for user-defined hyperlinks and automatically redact the matching hyperlink when required.

06 - Record Class Assignment

Assign a Record Class from the email’s content.

Analyze content to assign Record Class. That value can also assign Access Restriction, Record Retention, and labels—and require review.

07 - Access Restriction Assignment

Assign an Access Restriction from the email’s content.

Analyze content to assign Access Restriction. That value can also assign Record Class, Record Retention, and labels—and require review.

08 - Record Retention Assignment

Assign a Record Retention from the email’s content.

Analyze content to assign Record Retention. That value can also assign Record Class, Access Restriction, and labels—and require review.

What a processing result can do

Pattern, lexicon, address, domain, and hyperlink processing can assign a record class, access restriction, record retention, or label—and can require manual review. Where supported, the matched value can also be redacted automatically.

  • Record class
  • Access restriction
  • Record retention
  • Label
  • Manual review
  • Automatic redaction

Content-based decisions

Use reviewed examples or local AI to assign policy values.

Record Class, Access Restriction, and Record Retention assignment use the same content-analysis methods. Each targets its associated value and any related action values.

Keyword similarity

Start with a reviewed random sample.

An archivist assigns the target value to a random sample. QEPS uses emails with a given value to find emails with similar content and assign the same value.

Preferred method — local AI

Read the email against user-defined rules.

A local AI model analyzes the email’s content and applies the program’s rules to assign the appropriate Record Class, Access Restriction, or Record Retention value, plus related actions.

Resolve the result

Evaluate all assigned actions together.

QEPS evaluates the actions from its processing sources and determines the appropriate values for the email.

Completion and human control

Automate the first pass. Keep final authority with the archivist.

Complete means complete

An email is complete only when it has a record class, access restriction, and record retention—and every mandated review has been performed.

Manual decisions prevail

Archivists may review any email, assign those values directly, and redact its content. A manual action always takes precedence over a system action.

Attachment access copies

QEPS can automatically create attachment access copies based on media type. Archivists can review and redact those copies as needed.

Processing with rules, context, and review

Teach QEPS your policy.
Review what cannot be settled automatically.

Teams define the vocabularies and actions that matter to their program. QEPS applies them consistently across a collection and calls attention to conflicts, exceptions, and uncertain results.

What your team defines

Record classes Labels Access restrictions Retention Patterns Lexicons Email addresses Domains Attachment formats

What QEPS examines

  • Header fields and body sections
  • Addresses and domain names
  • Embedded hyperlinks
  • Patterns and lexicon matches
  • Attachments and file types
  • Meaning and similarity

What comes out

  • Record-class assignments
  • Access restrictions
  • Retention assignments
  • Labels and redactions
  • Attachment access copies
  • Emails flagged for review

Human intelligence + local analysis

Automation does the volume work. People make the archival decisions.

QEPS first pass

Repeat the rules across the collection.

Pattern matching, lexicon analysis, and semantic comparison help find sensitive content, related messages, and likely processing actions without reading every message from a blank page.

Archivist review

Resolve the cases that need judgment.

Reviewers validate results, settle conflicting assignments, change individual records, control redactions, and approve the collection before publication.

Private by design

Keep email out of public AI services.

QEPS can run its semantic model inside the deployment. Message text does not need to be sent to a public AI provider for meaning-based search.

Local AI for email collections

Process every email.
Review the ones that matter.

QEPS runs a local large language model on the QEPS server. It reads each email against the rules your team defines, then returns a decision a reviewer can inspect, validate, or change.

Your processing policy

Define rules once. Apply them across the collection.

  • Record classRecord · Non-Record
  • Access restrictionPublic · Private · Protected · Controlled · Exempt
  • Record retentionNone · Short Term · Long Term · Permanent
  • Redaction and reviewTerms, patterns, exceptions, and confidence thresholds
Local model reviewOne email · Explainable result

AI reads the message, applies relevant rules, and records why.

Record classRecord

Access restrictionPublic

Record retentionPermanent

Confidence and justification

High confidence. The result cites the message content and the user-defined rules that support each assignment.

Review queue

Flag low-confidence, conflicting, or likely-redaction results for an archivist.

01

Keep email on your server.

Message text, attachments, and processing context stay within the QEPS deployment instead of being sent to an online AI service.

02

Make costs predictable.

A local model avoids per-message or per-token charges as collections grow, letting teams plan around infrastructure they manage.

03

Process at collection scale.

Local processing avoids repeated transfers, provider queues, and request limits when QEPS is reviewing large volumes of email.

Human review stays in control

QEPS identifies messages that need manual redaction review or result validation. With well-defined rules and confidence thresholds, AI can reduce manual review and classification to as few as 1% of a collection.

Installation and deployment

Build a QEPS server for collection-scale work.

QEPS is not a hosted solution. It is installed on a Linux server inside your secure network, keeping email processing, local-model inference, search, and archive data in your environment for security and privacy.

Baseline QEPS serverLinux deployment

Compute, memory, and fast working storage for the QEPS stack.

Memory64 GB RAM minimum

CPU16+ high-performance cores

GPUNVIDIA RTX 509032 GB VRAM minimum

Working storage2 TB+ high-speed storage

Software stack

The services that make local processing work.

  • QEPS system
  • Solr search engine
  • Python
  • LM Studio
  • vLLM
Archive storage

Plan separately for the email you will retain.

Archive and collection storage varies with the number of email accounts, the volume collected, and the processing work retained. A reasonable initial deployment starts at 16 TB; long-term requirements can approach 128 TB or more.

These estimates do not include storage required for system backups.

Optional APPX service

Bring your own server—or let APPX configure and manage it.

Discuss your deployment

Public access

Publish the collection, not the working repository.

QEPS creates a separate searchable index for approved collections. Staff can publish, unpublish, or withhold a collection while the original archive and internal processing data remain under administrative control.

Search by keyword, natural-language prompt, subject, date, sender, or recipient.

Narrow with labels, entity tags, dates, addresses, and domains.

Read the approved email as PDF and open the permitted attachment copy.

QEPS Search page for a published email collection, with keyword, natural-language, and email-address searches; collection filters; and email results.

Email in its archival context

QEPS works alongside AXAEM.

A QEPS collection can carry AXAEM references for its bibliographic record, agency, entity, and contact. That connects processed email to the broader records-management and archival description around it.

Explore AXAEM ↗

QEPS resources

Keep learning.
Start with the work.

These CoSA ShopTalk recordings explore the challenges of state-government email and the practical role of automation and AI. Additional QEPS reference materials will be collected here as they become available.

CoSA ShopTalk · 2026 Recording

Using AI to Process State Government Emails

State archives face immense volumes of email and strict records-management duties. This session examines how AI can identify official records and non-records, flag sensitive information such as PII, and apply restrictions, retention schedules, and descriptive metadata—improving consistency while keeping long-term access in view.

Watch the recording ↗
CoSA ShopTalk · 2025 Recording

Practical Uses of AI at State Archives

An introduction to practical ways AI can support state archives. The session considers where automated analysis can assist staff, the policy and review controls that keep people accountable, and opportunities to make archival work more consistent and scalable.

Watch the recording ↗
CoSA ShopTalk · 2024 Recording

Meeting the Challenges of Processing Emails

State-government email creates a large, complex workload before public access or preservation. This session presents a hybrid approach: use computer algorithms for repetitive work and direct human review to the exceptions and decisions that need judgment.

Watch the recording ↗

See the real workflow

Bring a mailbox.
Walk through it with us.

A useful QEPS demonstration starts with the kind of email your team receives and the decisions you need to make. Ask about a guided demo or a trial run with representative material.

Product contact Steve Frizzell steve@appx.com 1-800-879-2779 ext. 2481

Request a demo or trial run

Tell us a little about your email collections. We'll follow up personally.

We'll use these details only to respond to your request.