Skip to content
  • Solutions
    Lightning IQ up the darkness light bulg
     
    Lightning IQ™ is a leader in data intelligence software helping organizations better understand, manage and control the costs and processes associated with risk, compliance and regulatory matters.
     
  • Features
    tall and pretty crop
     
    The Lightning IQ platform makes the impossible, possible, with its revolutionary design and performance - capable of mapping and indexing petabytes of metadata and full-text in hours.

    Blistering speed

    Scan and index petabytes of metadata and full-text in hours. Transform your data driven initiatives and go from data chaos to data quality across your entire enterprise.
     

    Innovative architecture

    Distributed, event-driven, asynchronous and stateless micro-services deliver best-in-class performance and provide a platform for rapid innovation.
     

     

    High-risk data identification

    Quickly identify ROT, find PII, and/or the location of confidential or critical intellectual property.
     

    Effortless deployment

    Fully automated deployment both on-prem or in the cloud eliminates the time-consuming, error-prone "configuration" typically associated with data scanning.
     

     

    AI-Readiness 

    Lightning IQ scans petabytes of enterprise data to uncover hidden risks - across file shares, cloud drives, and legacy systems - before they become AI liabilities.
     

    Terabytes in minutes; petabytes in hours

    Scale up and down - horizontally and vertically - as needed. No drive too small, no network cluster too big for Lightning IQ.      
     

     

  • Partners
  • Company
    Our Vision
    To transform how enterprises understand and manage their unstructured data, delivering total visibility, control, and confidence. We empower organizations to operate securely, reduce risk, and maximize the value of their enterprise data.
     

     

    Our Mission
    1. To illuminate and classify enterprise data at petabyte scale
    2. To reduce data risk through actionable insights and intelligent control
    3. To enable confident, compliant decisions across every corner of the business
     

     

Search icon
Shiny blue binary code on black background
Lightning IQ 2.0 Generally available July 31, 2026

Scan. Classify. Act.

Lightning IQ 2.0 delivers the interface and features so you can act on your data.
Terabytes per minute. Petabytes per hour. Version 2.0 closes the loop from petabyte-scale intelligence to defensible action - agentless, in place, with the cost of inaction priced in dollars.
LIGHTNING IQ 2.0 / IN-PLACE SCAN SCANNING
2,417,908
Files / min
3.75 PB
Indexed
IN PLACE
Agentless
PIIPHIDUPLICATEROTPST OWNERGPS / CAMERARETENTION
100B
records scanned per day
45+ PB
indexed in 24 hours
39
new legal export metadata fields
$1M+
saved per petabyte of ROT
The platform shift

One closed loop, no exports in between.

In 1.0, findings ended in a report. In 2.0, findings flow into classification and straight into action — the whole loop runs inside the platform, against data that never moves.
01

Scan

Agentless, in-place scanning across NFS, Azure Blob, and AWS S3. A redesigned template-first workflow gets a scan running in minutes, with Create Custom Scan for power users.

02

Classify

Reusable search patterns tag every file by policy — PII, PHI, sensitivity, retention schedule — with duplicate detection inline and costs attached per category.

03

Act

Move, copy, or delete in bulk directly from the results. One decision drives action across millions of files, with hash re-verification before each one.

New in 2.0 — redesigned scan workflow Redesigned scan workflow · template-first
image (27)
Speed and scale

Terabytes per minute. Petabytes per hour.

2.0 keeps the engine that made 1.0 the fastest enterprise storage indexing platform, and puts action on the other side of it. Scale up and down, horizontally and vertically — no drive too small, no network cluster too big.
Metadata index rate 1 TB / minute
Full-text extraction 1 PB / hour
Daily scan ceiling 100B records · 45+ PB
PROVEN AT SCALE / CUSTOMER RUN
5.13B
items indexed in a single customer run
3.75 PB
of unstructured data scanned
5 hrs
start to finish, in place
/san/vol04/finance/archive/2019/q3_close.xlsx PII HIT
/san/vol04/legal/custodian_87/mail.pst PST OWNER FOUND
/azure/blob/media/dsc_10442.mov GPS + CAMERA
/s3/eng-backups/2017/build-artifacts.tar ROT · OBSOLETE
/san/vol02/marketing/logo_final_v7.ai DUPLICATE
/san/vol09/hr/enrollment_2024.pdf PHI HIT
New in 2.0 — Costs

Every category of junk data now carries a price tag.

ROT Analysis gains a dedicated Costs section: a color-coded panel with cost, file counts, and data sizes per category, sub-categorized by search name. Top Categories by Metadata brings size and count cards onto the Overview page. The business case builds itself.
Duplicates 312 TB $421,000
Obsolete — last modified pre-2018 486 TB $655,000
Trivial — system and cache files 94 TB $127,000
Orphaned PST archives 160 TB $217,000
$1,420,000 annual cost of ROT in this estate
New UI — ROT Analysis · Costs ROT Analysis → Costs
costs
Investigation for everyone

If you can point and click, you can run a petabyte-scale investigation.

The 2.0 dynamic query builder composes complex Lucene searches visually — no syntax required. New users stop needing an expert; experienced users stop making syntax errors.
New UI — Dynamic Query Builder Dynamic Query Builder
Query Builder LIQ

Dynamic query builder

Compose complex Lucene searches visually, with no syntax to memorize and nothing to mistype.

Custom and extrapolated metadata search

Filter on custom fields and machine-generated signals — classify against your retention schedule by computed properties, not filenames.

Scan completion notifications

Real-time lifecycle updates pushed to Slack, email, and other channels. No more checking on scans manually.

Redesigned scan workflow

Template-first setup for repeatable runs, plus Create Custom Scan when a matter needs something bespoke.

Richer intelligence

Every file, fully characterized.

Text extraction, language detection with confidence scores, MIME type, full A/C/M timestamps, custom hashing — on a schema that adapts to the scan config. 2.0 adds five new classes of signal on top.
OWNERSHIP

PST ownership identification

Owners auto-derived from mail store properties, so abandoned and orphaned PST archives surface with a name attached.

MEDIA

Expanded media metadata

GPS coordinates, camera make and model, video timestamps, and audio tags captured during the scan.

GEOGRAPHY

Geographic distribution analytics

Map where a scan's data was created, how much of it, and when — useful for data residency and sovereignty questions.

STRUCTURE

Directory and file statistics

Largest directories by count and size, widest fanout, and deepest subdirectory trees across the estate.

TIME

File distribution by year

File count and file size by last-modified year — the fastest way to see how much of the estate stopped changing a decade ago.

New in 2.0 — Remediation

Act at scale. Prove you acted correctly.

Move, copy, and delete data directly from scan results — one decision file driving bulk action across storage. Before Lightning IQ acts on a file it re-verifies the file's hash against the scan; if the content changed after the decision was made, the file is skipped. Duplicate identification mode flags duplicates while keeping them in the pipeline, so nothing disappears without a decision.
38
Succeeded
2
Skipped
0
Failed
HASH VERIFICATION / PRE-ACTION
/san/vol04/archive/2019/q3_close.xlsx HASH MATCH · DELETED
/s3/eng-backups/2017/build-artifacts.tar HASH MATCH · MOVED
! /san/vol02/design/brand_kit.zip CHANGED · SKIPPED
/azure/blob/media/dsc_10442.mov HASH MATCH · COPIED
/san/vol09/hr/enrollment_2024.pdf HASH MATCH · MOVED

Every session returns succeeded, failed, and skipped counts with throughput and duration by action type.

Your intelligence layer, not another silo

Turn a petabyte scan into a dashboard.

Native CSV, PARQUET, and JSONL output drops straight into Power BI and any BI tool. Legal export gains 39 new metadata fields and full family binaries. We don't ask you to rip anything out — we feed what you already run.
PARQUET · CSV · JSONL

Analytics exports

Structured output that drops straight into Power BI for self-service exploration — charts, slicers, drill-down into paths and categories.

+39 FIELDS

Expanded legal export metadata

Email headers, document properties, and geolocation in .dat files, plus full family binaries rather than hits alone.

REPORTING

Transparent remediation reporting

Succeeded, failed, and skipped counts with throughput and duration by action type — an auditable record of what happened.

TEMPLATES

Reusable scan and export templates

Save a scan and export configuration once, then run it again at any scale with identical settings.

Scan and remediate in place across NFS / SMBAzure BlobAWS S3Google Cloud StorageOn-prem SAN
Deploy cleaner

Less friction between you and the first scan.

Deployment friction in 2.0 is measurably lower than what you would remember from a 1.0 evaluation.

Workload Identity and ADC for GCS

Authenticate Google Cloud Storage without long-lived service account keys.

Surfaced upload metrics and failure reasons

Tell configuration, authentication, and transient failures apart at a glance.

AD1 forensic support

Unpack AccessData Logical Image Files alongside live storage in the same scan.

Improved cloud storage handling

Leading-slash blob keys no longer error out, and the Storages panel scrolls properly at depth.

Release notes

Everything in 2.0.

The full 2.0 feature list, grouped by the job it does. Version 2.0 is generally available July 31, 2026.
01

Act on your data

Remediation — move, copy, and delete data directly from scan results.
Duplicate identification mode — mark duplicates with a flag while preserving them in the pipeline.
02

Prove the cost

ROT Analysis — Costs section — color-coded panel showing cost, file counts, and data sizes per category, sub-categorized by search name.
Top categories by metadata — size and count cards on the Overview page.
03

Investigate faster

Dynamic query builder — build complex Lucene searches visually, no syntax required.
Advanced search on custom and extrapolated metadata — filter on custom fields and machine-generated signals.
Scan completion notifications — real-time lifecycle updates via Slack, email, and other channels.
Redesigned scan workflow — template-first, with Create Custom Scan for power users.
04

Richer intelligence

PST ownership identification — auto-derive PST owners from mail store properties to find abandoned or orphaned PSTs.
Expanded media metadata capture — GPS coordinates, camera make and model, video timestamps, audio tags.
Geographic distribution analytics — map where a scan's data was created, how much, and when.
Enhanced directory and file statistics — largest directories by count and size, widest fanout, deepest subdirectories.
File distribution by year — file count and file size by last-modified year.
05

Feed your stack

Analytics exports — CSV, PARQUET, and JSONL for BI and visualization platforms.
Expanded legal export metadata — 39 new fields including email headers, document properties, and geolocation in .dat files.
Full family binaries in legal export — full families included, not just hits.
06

Deploy cleaner

Workload Identity / ADC for GCS — authenticate Google Cloud Storage without long-lived service account keys.
Surfaced upload metrics and failure reasons — diagnose configuration, authentication, or transient failures at a glance.
AD1 forensic support — unpack AccessData Logical Image Files.
Improved cloud storage handling — leading-slash blob keys no longer error out.
Under-the-hood fixes — text extraction, Lucene dot handling, Storages panel scrollbar, and S3 legal export corruption.
“Lightning IQ exists to answer one simple question: What's really in your data?

Daniel Pidutti · CEO, Lightning IQ

See the three things that didn't exist last time.

Twenty minutes is all it takes to watch a petabyte scan turn into defensible action. Lightning IQ 2.0 is generally available July 31, 2026.