Storage & Cost

File-Level Archiving Is GA — And It Quietly Pulls Files Out of Copilot

Microsoft 365 Archive went file-level in July 2026, cutting cold storage to $0.05/GB/month. Most coverage treats the Copilot exclusion as a bonus. It is a tradeoff, it is silent, and it turns a billing decision into an AI-relevance decision.

SharePoint Support TeamOctober 8, 202615 min read
File-Level Archiving Is GA — And It Quietly Pulls Files Out of Copilot - Storage & Cost guide by SharePoint Support
File-Level Archiving Is GA — And It Quietly Pulls Files Out of Copilot - Expert Storage & Cost guidance from SharePoint Support

What changed in July

Microsoft 365 Archive can now archive individual files and folders instead of only whole sites. Public preview landed March 30, 2026; general availability rolled out from late June through late July 2026, enabled by default in tenants where Microsoft 365 Archive was already active.

SharePoint governance framework showing policies, roles, and compliance
SharePoint governance model with policies and compliance controls

The old whole-site model was close to useless for the situation almost every enterprise is actually in: an active site carrying years of inactive content. You could not archive the 2018 project folders without archiving the site that the 2026 project is running in. File-level granularity fixes that, and it is a genuine improvement.

It also introduces a tradeoff that most of the coverage is treating as a bonus feature. We will get to that.

---

The money, stated plainly

| Tier | Rate | Annualized |

|---|---|---|

| Base tenant pool | 1 TB + 10 GB per licensed user | included |

| Overage (standard) | $0.20 / GB / month | ~$2,400 / TB / year |

| Microsoft 365 Archive | $0.05 / GB / month | ~$600 / TB / year |

That is a 75% per-GB reduction on archived content. On 10 TB of inactive overage, roughly $24,000/year becomes roughly $6,000/year. Retrieval is metered separately, so your effective savings depend on reactivation behavior — a point we will come back to, because it is where naive archive programs lose their gains.

---

The reality check: archiving does not free quota

This is the single most expensive misunderstanding about the feature, and it catches experienced admins.

Archived content still counts against your storage footprint. You are not reclaiming space. You are paying a lower rate on the same bytes.

If you are sitting at a quota ceiling and expecting an archive pass to create headroom, it will not. The things that create headroom are different:

  • Version history trimming — often the single largest recoverable footprint in a mature tenant
  • Recycle bins — both stages, including the site collection second-stage bin
  • Preservation Hold Library — retention policy residue that silently accumulates and that most admins have never looked at
  • Actual deletion of content that has no retention obligation

Archiving lowers the bill. Cleanup creates room. They are different projects with different outcomes, and conflating them is how a storage program delivers a smaller invoice while the quota alarm keeps firing.

---

The part nobody is pricing in: Copilot goes blind on archived content

Archived files are excluded from search and from Copilot's index.

Nearly every write-up on this feature frames that as an upside — archiving removes stale documents, so Copilot stops surfacing the 2019 version of the expense policy. And sometimes that is exactly right. Noise reduction in Copilot grounding is a real problem and archiving is a real lever against it.

But framing it purely as a benefit is a mistake, because the same mechanism runs the other way:

If you archive content Copilot needed, answer quality degrades and nothing tells you.

There is no warning. No "this answer would have been better with archived content" flag. No degradation alert. Copilot simply answers from a smaller corpus and sounds exactly as confident as it did before. The failure is silent, and it surfaces weeks later as "Copilot has gotten worse lately" with no obvious cause and no changelog to point at.

Why "inactive" is the wrong test

Storage tooling ranks candidates by last-accessed date. That is the correct signal for cost. It is a poor signal for AI relevance, because the two properties are close to uncorrelated.

Consider what scores as inactive but is high-value to Copilot grounding:

  • Signed contracts and MSAs. Nobody opens them for years. They are the authoritative answer to most commercial questions.
  • Closed-project post-mortems and design decisions. Rarely read. Precisely the institutional memory Copilot is supposed to make retrievable.
  • Superseded-but-referenced policy. The current version is what people open; the prior version is what explains why a control exists.
  • Regulatory submissions and audit evidence. Touched once, then dormant, then urgently needed.
  • Completed engagement documentation. The exact corpus an internal "how did we solve this before" prompt depends on.

Every one of those is a top-ranked archive candidate by last-accessed date, and several of them are the highest-value grounding content in the tenant.

Now the inverse — content that is genuinely safe to archive:

  • Superseded drafts where a canonical final exists
  • Build artifacts, exports, and generated reports
  • Duplicated attachments and mail-drop residue
  • Personal working copies in departed users' areas
  • Media assets from completed campaigns

The distinction is not activity. It is whether the content is authoritative for a question someone might ask.

A test that actually works

Before archiving a content class, ask: *if someone asked Copilot a question this content answers, would we want the archived version in the answer?*

  • Yes → do not archive, or archive only after confirming a canonical active copy exists elsewhere
  • No → archive freely
  • Unsure → sample ten files and read them. The answer becomes obvious fast, and ten minutes here is worth more than any automated scoring pass.

This reframes the whole exercise. Storage tooling is optimizing for one variable. You need to optimize for two, and only one of them shows up on the invoice.

---

Retrieval is a plan, not an afterthought

Reactivation timing varies by how long content has been archived, and retrieval carries its own cost. Treating archiving as a one-way door is how programs get into trouble.

Before any bulk pass, settle:

  • Who can trigger reactivation, and does that require a ticket, an admin, or a self-service path
  • Expected latency per retention tier, tested rather than assumed
  • Which content is off-limits because delayed retrieval would break a legal hold, a regulatory response window, or an operational SLA
  • What retrieval costs at your expected volume, and whether a chatty department could erase the savings

That last one is not hypothetical. An archive pass that saves $18,000 a year and then triggers constant reactivation because someone archived an actively-referenced library nets out to work you did for nothing.

Legal hold content deserves specific attention. If a matter is live or reasonably anticipated, archiving anything in scope introduces retrieval latency into a process that has deadlines set by someone other than you. Coordinate with legal before, not after.

---

A rollout that will not hurt

Step 1 — Find out what you are actually paying. Establish current overage in GB and dollars. If you are not over your allocation, archiving saves you nothing and this entire project is premature.

Step 2 — Clean before you archive. Version history, both recycle bins, Preservation Hold Library. This is where headroom comes from, and it is free. Archiving dirty data means paying to store duplicates at a discount.

Step 3 — Classify by grounding value, not activity. Run the Copilot test above against your largest inactive content classes. Produce two lists: safe-to-archive and grounding-critical.

Step 4 — Confirm the default-on state. If Microsoft 365 Archive was already active in your tenant, file-level archiving likely arrived during the June-to-July window without an opt-in. Verify who has access to it and whether anyone has already used it.

Step 5 — Pilot on one content class. Something unambiguously safe — completed campaign media, build artifacts. Measure the billing delta and test a retrieval end to end, including latency.

Step 6 — Write the retrieval runbook before widening. Who, how, how long, how much.

Step 7 — Widen by class, never by date sweep. A date-based bulk archive is the mechanism by which organizations accidentally remove their institutional memory from Copilot in a single afternoon.

---

Our read

File-level archiving is a good feature and the economics are real — 75% off inactive overage is not a rounding error at enterprise scale.

But it has quietly changed what an archiving decision *is*. It used to be a pure infrastructure call: cold storage is cheaper, move the cold things. Now every archive decision also decides what your AI can see, and the two objectives do not point the same direction. Cost optimization says archive by last-accessed date. Copilot quality says keep anything authoritative, however dormant.

Organizations that run this as a storage project, using storage tooling, optimizing a storage metric, will hit their savings target and degrade their Copilot deployment at the same time — and because the degradation is silent, they will not connect the two.

Run it as a content governance project that happens to save money. Same savings, and you keep the corpus your AI investment depends on.

---

Need help separating grounding-critical content from genuinely cold storage before an archive pass? See SharePoint Consulting and SharePoint & Copilot, or get in touch.

Sources

Share this article:

Written by the SharePoint Support Team

Senior SharePoint Consultants | 25+ Years Microsoft Ecosystem Experience

Our senior SharePoint consultants bring deep expertise spanning 500+ enterprise migrations and compliance implementations across HIPAA, SOC 2, and FedRAMP environments. We cover SharePoint Online, Microsoft 365, migrations, Copilot readiness, and large-scale governance.

Frequently Asked Questions

What is file-level archiving in Microsoft 365 Archive?▼
It lets you move individual files and folders into a low-cost cold storage tier while the rest of the site stays fully active. Before this, Microsoft 365 Archive worked at whole-site granularity only, which made it useless for the common case: an active site carrying years of inactive content. File-level archiving entered public preview on March 30, 2026, and general availability rolled out from late June through late July 2026.
How much does archiving actually save?▼
SharePoint storage beyond your tenant allocation bills at $0.20 per GB per month, which is roughly $2,400 per TB per year. Microsoft 365 Archive meters archived content at $0.05 per GB per month — a 75 percent per-GB reduction. On 10 TB of inactive overage that is roughly $24,000 a year down to about $6,000. Retrieval is billed separately, so the real number depends on how often you pull content back.
Does archiving free up my SharePoint storage quota?▼
No, and this is the most consequential misunderstanding about the feature. Archived content still counts against your tenant's storage footprint. You are not reclaiming quota — you are paying a lower rate on the same bytes. If you are already at a quota ceiling and expecting archiving to create headroom, it will not. Deletion, version trimming, and clearing the Preservation Hold Library are what create headroom; archiving is what lowers the bill.
Are archived files removed from Microsoft Copilot?▼
Yes. Archived files are excluded from both search and Copilot's index. Most coverage frames this as a benefit — less noise in Copilot results — and sometimes it genuinely is. But it is a tradeoff, not a free win: if you archive content Copilot needed in order to answer well, answer quality degrades and nothing warns you. The archiving decision is now simultaneously a cost decision and an AI-capability decision, and most organizations are only making the first one.
How long does it take to get an archived file back?▼
Reactivation timing varies depending on how long the content has been archived, and retrieval carries its own cost. This is why archiving needs an explicit retrieval plan rather than being treated as a one-way cost optimization. Before a bulk archive pass, confirm who can trigger reactivation, what the expected latency is for your retention tiers, and which content categories are effectively off-limits because a delayed retrieval would break a legal, regulatory, or operational commitment.
Is file-level archiving enabled automatically?▼
The GA rollout was enabled by default in tenants where Microsoft 365 Archive was already active. If you had Microsoft 365 Archive turned on, the file-level capability likely appeared in your tenant during the June-to-July 2026 window without a separate opt-in. That is worth verifying rather than assuming, because a capability that arrives by default is a capability your admins may already be using without governance around it.

Need Expert Help?

Our SharePoint consultants are ready to help you implement these strategies in your organization.

Continue Reading in Storage & Cost

Architecture

SharePoint Online Limits in 2026: What Actually Breaks at Scale

Everyone can recite the 5,000-item threshold. Almost nobody can tell you what genuinely fails first in a large tenant — and in 2026 the answer changed, because agents and Copilot now query your lists too.

Emergency Support

Your SharePoint Farm Is Now Unpatched: The Emergency Support Playbook

SharePoint Server 2016 and 2019 hit end of support on July 14, 2026 — and unlike Windows Server, there is no Extended Security Updates program to buy. Here is what changes operationally, and how our 24/7/365 emergency response works when it does.

AI & Copilot

Copilot in SharePoint Just Got Live Dashboards and One-Click AI Buttons

The August 2026 release turns SharePoint lists, Excel and CSV files into dashboards that stay connected to their source data, and lets site owners drop saved Copilot prompts onto pages as buttons. Both are genuinely useful. Both need governance before you turn them loose.

AI & Copilot

Microsoft Copilot for SharePoint: The Guide for 2025

Everything you need to know about Microsoft Copilot integration with SharePoint, from setup to advanced automation strategies.

Migration

SharePoint Migration Best Practices: 15 Expert...

Learn the proven migration strategies used by leading organizations to migrate to SharePoint Online without disruption.

Governance

Building an Enterprise SharePoint Governance Framework...

Create a governance framework that balances security with productivity, enabling self-service while maintaining control.