Explore our AI courses, practical training for non-technical teamsExplore courses Explore AI courses
OperationsInventoryHow-To

How to Use AI for Inventory Management Without Buying a Platform

Demand forecasting is the thing every inventory tool leads with and the thing your business is least ready for. The useful work is quieter, and you can start it this week.

TLDR: Use AI to reconcile messy SKU data, surface dead stock, draft reorder logic and argue with your buying plan. Treat any forecasting accuracy claim as unverified until someone names the baseline.
92.5%Of M5 competition teams failed to beat a simple forecasting benchmark
65%Of store inventory records did not match the physical count
55%Of surveyed SMBs hold at least 20% excess stock

Share this article

The Short Version

In the M5 forecasting competition, 5,507 teams tried to beat a set of benchmarks on 42,840 series of real Walmart sales data. Only 7.5% beat a simple exponential smoothing method, and fewer than half beat a plain naive forecast. That is the honest ceiling on AI demand forecasting, and it assumes clean history you probably do not have: a landmark retail study found 65% of nearly 370,000 store inventory records did not match the physical count. So the biggest wins for a small ops team are the unglamorous ones. Reconcile your SKU list, find the stock that has not moved in a year, get your reorder logic written down where you can read it, and make the model argue with your assumptions before you spend money on a platform.

What you are actually trying to fix

Inventory management is two problems wearing one coat. Either you have cash sitting on a shelf that nobody wants, or you have a customer standing in front of an empty one. Every tool in this category, AI or otherwise, is selling you a better trade-off between those two.

The money involved is large enough to feel abstract. IHL Group, which has tracked this measure for 18 years, puts the global cost of inventory distortion (out-of-stocks plus overstocks) at $1.73 trillion a year, equal to 6.5% of global retail sales, even after retailers spent $172 billion on improvements in the previous twelve months.[1] North America accounts for $415 billion of it, and supply chain disruption alone for $301 billion.[1]

None of that tells you anything about your business. This might: across US retail as a whole, the Census Bureau’s inventories-to-sales ratio sat at 1.25 in May 2026, meaning retailers were holding roughly one and a quarter months of merchandise against a month of sales.[2] If your number is meaningfully above that, you have a cash problem before you have a forecasting problem.

The closest thing to a picture of businesses your size comes from Netstock, which combined anonymised platform data from more than 2,400 customers with a survey of over 130 users at companies under $250M in revenue. It found 55% holding at least 20% excess stock, up from 48% the year before, 46% reporting that 5% or more of inventory counts as dead stock, and 17% carrying more than 10% dead stock, up from 12%.[3]

Worth sitting with for a second: every business in that sample had already bought inventory planning software.

What the forecasting claims actually rest on

I used to open workshops with the demand forecasting demo, because it lands. You paste in two years of sales, the model draws a confident line into next quarter, and the room goes quiet in a good way. I stopped doing it, because I could not answer the one question a sharp ops manager always asked: better than what?

There is a clean public answer to that question and almost nobody in this market quotes it. The M5 competition, run on Kaggle and written up in the International Journal of Forecasting, gave 5,507 teams (7,092 people across 101 countries) five and a half years of real Walmart unit sales, 42,840 series covering 3,049 products across 10 stores, and asked them to forecast 28 days ahead.[4] These were motivated people competing for a $50,000 prize pool with modern machine learning.

The scoreboard against the organisers’ benchmarks:

Share of M5 teams that beat each simple benchmark

Beat a plain naive forecast48.4%
Beat a seasonal naive forecast35.8%
Beat simple exponential smoothing7.5%

Of 5,507 teams forecasting 42,840 series of Walmart sales data, the share whose final submission outperformed each benchmark method (M5 Accuracy competition results).[4]

More than half the field could not beat a forecast that says “tomorrow looks like today.” Only 415 teams, 7.5%, beat exponential smoothing applied at the product-store level and summed upward, a method you could implement in a spreadsheet.[4]

The honest other half of that story: the teams at the top did win, and they won properly. Every one of the top 50 improved on that exponential smoothing benchmark by more than 14%, the top five by more than 20%, and the winning team by 22.4%.[4] Machine learning is genuinely better here. The catch is that “genuinely better” means roughly a fifth better than a method with no AI in it, achieved by specialists on immaculate data, and 92.5% of the people who tried could not get there at all.

So when a vendor tells you their AI is “30% more accurate,” the only useful reply is: more accurate than which baseline, at what level of aggregation, over what horizon? Stephan Kolassa, who spent years producing automated forecasts for large European retail chains, wrote a whole piece in Foresight pulling apart the published accuracy surveys people quote at each other, and concluded that “the quest for external forecasting benchmarks is futile.”[5] His own working number is worth knowing: in grocery retail, he saw normal one-week-ahead errors ranging from 20% to 60% MAPE across different companies, depending on how many fast sellers they had, how much fresh produce, and how good their data was.[5]

That range is the real answer to “what accuracy should I expect.” It’s a three-fold spread, and the thing that moves you within it is mostly not the algorithm. If you want the longer version of this argument applied to revenue rather than units, we covered it in our guide to using AI for sales forecasting.

The precondition that never makes the demo

A few months ago I sat with an operations manager at a homeware wholesaler who wanted help choosing between two forecasting tools. Before we looked at either, I asked her to export her product list. It had 1,900 rows. One planter appeared three times: once as “Terracotta Pot LG”, once as “TERRACOTTA-POT-LARGE”, and once as “TP-LG-01”, each with its own stock count and its own sales history.

No model fixes that. Every forecast it produces will be a forecast of one third of the demand for a planter, three times over, and it will look completely plausible.

This is not a small-business embarrassment, it’s the normal condition of retail data. The reference study here is DeHoratius and Raman in Management Science, who examined nearly 370,000 inventory records across 37 stores of a single retailer and found 65% of them inaccurate, with an average absolute deviation of almost five units per SKU, about 35% of the average actual stock on the shelf.[6]

What makes that worth acting on rather than just despairing about is the sales evidence. An ECR Retail Loss study ran matched pairs of stores across several retailers, tracked sales and inventory records for 12 weeks, then performed a stocktake and corrected the records in the test stores only, leaving the control stores alone, and tracked another 12 weeks. Across roughly 233,000 SKUs, correcting inventory records grew sales by 4% to 8%.[7]

Before you forecast anything, check four things: that each product appears exactly once in your list, that units of measure are consistent (a case of 12 is not 12), that discontinued items are marked as discontinued rather than sitting at zero, and that your on-hand numbers have been physically counted within the last quarter. AI is genuinely good at the first two. It cannot do the fourth for you.

Here’s the reframe I now use in every session. Data cleaning isn’t the boring prerequisite to the AI work, it’s the AI work with the highest confirmed payoff attached to it, and it’s the part a language model handles well. Pattern-matching three spellings of “terracotta pot” is exactly what these tools do without complaint at two in the afternoon on a Thursday.

Five unglamorous jobs AI does well

These are the five uses I’d defend to a sceptical finance director. None of them involve the model predicting the future.

1. Reconciling the same product listed three ways

Paste your product export into a general-purpose assistant and ask it to group probable duplicates, showing its reasoning for each group. The prompt phrasing that works: “Group rows that are likely the same physical product under different names or SKU formats. For each group, list every row you included and say what evidence you used. Flag anything you are unsure about separately.” That last sentence matters more than it looks, because without it you get a confident merge list with no way to audit it. Our walkthrough on using AI to analyse a spreadsheet covers the mechanics, and the same technique transfers straight across to product records.

2. Finding the stock that has stopped moving

Dead stock hides because nobody has a report for it. Give the model your SKU list with last sale date, units on hand and unit cost, and ask for every item with no sale in 180 days, sorted by capital tied up, with a running total. You’ll get a number for how much cash is sitting still. In the Netstock survey, 11% of businesses admitted they had no strategy at all for reducing excess, and 69% relied on promotions while only 29% redistributed stock between locations.[3] Knowing which items to promote is the whole game, and it’s a sorting problem.

3. Drafting reorder logic you can actually read

Ask for a reorder point calculation per SKU, written as a spreadsheet formula, using your average daily sales, your supplier lead time and a safety stock buffer you choose. The output should be a formula in a cell, visible and editable. The moment your reorder logic lives inside a model rather than inside your spreadsheet, you’ve lost the ability to explain a purchase order to your accountant.

4. Turning supplier chaos into a table

Forward it your last twenty supplier confirmation emails and ask for a table of supplier, product, quantity, promised date, actual date, and days late. Netstock’s respondents named lead time variability their top supplier problem at 68%, ahead of long lead times at 58% and cost at 48%.[3] You cannot buffer for variability you have never measured, and most small teams have the data for it sitting unread in an inbox.

5. Arguing with your own buying plan

This is the one people skip and the one I’d keep if I could only keep one. Paste your draft purchase order alongside last year’s sales for those items and ask: what is this order assuming about demand that isn’t written down? Which of these lines would I regret if sales came in 20% under plan? Which am I buying because of a minimum order quantity rather than because I need it? On that last point, 45% of surveyed businesses said minimum order quantities force them to buy more than they need.[3]

The pattern across all five: the model works on documents, records and assumptions, and the arithmetic stays somewhere you can see it. Every one of these produces an output you could check by hand in ten minutes if you wanted to, which is exactly why they’re safe to act on.

Most of this is arithmetic, and you should see it

Something I notice in almost every session: people expect inventory to be harder than it is. Then we write the reorder point on a whiteboard and there’s a small silence.

Average daily demand, multiplied by supplier lead time in days, plus a safety buffer. That’s the reorder point. When stock falls to that number, you order. The safety buffer is where judgement lives, and judgement is the part that should stay yours.

Lead time is the input that quietly destroys people, because most businesses use the number the supplier quoted rather than the number the supplier delivers. The spread is enormous. In Netstock’s benchmark data, the top quartile of businesses held average lead times of 20 days or less, while the bottom quartile faced delays stretching beyond 80 days.[3] Same industries, same regions, four times the wait. If you’re buffering against a quoted 20 days and actually receiving 60, no forecasting model on earth will save your service level.

So the sequence that works is: measure your real lead times from your own emails and receipts, feed those into the reorder point, and only then worry about whether your demand estimate is clever. The supplier conversation matters more than the algorithm, and the receipts that prove your real lead times are usually sitting in accounts payable, which is where our guide to using AI for invoice processing picks the thread up.

My honest take after years of teaching this to non-technical teams: use AI to write the formula, explain the formula, and pressure-test the inputs. Let the spreadsheet do the arithmetic. Language models are excellent at reasoning about structure and unreliable at long chains of calculation, and there is no reason to make them do the one thing a spreadsheet has done perfectly since 1985.

What these tools actually cost

Most operations managers who ask me which inventory platform to buy need a clean spreadsheet and two good prompts. A minority genuinely need software, and the tell is that they’ve already run the manual process twice and keep hitting the same wall. Here’s what the market looks like if you’re in the second group, with every price taken from the vendor’s own pricing page in August 2026.

Published pricing for tools a small business would actually consider

ToolPublished priceWorth knowing
Zoho InventoryFree plan (50 orders, 1 user, 1 location); Standard $29/organisation/month billed annually, up to Enterprise at $249Priced per organisation rather than per user, which suits small teams
KatanaFree plan; Core from $299/monthNo per-user fees; built around manufacturing and assembly
Inventory Planner by SageFree to install, then quote only4.4 out of 5 across 130 reviews on Shopify’s app store
Stocky by ShopifyIncluded with a Shopify POS Pro subscriptionRated 3.0 out of 5 across 188 reviews on Shopify’s own app store
NetstockFrom $900/month, annual subscriptionImplementation typically takes 6 to 10 weeks and requires an ERP to plug into
Microsoft 365 Copilot$30/user/month, paid yearly, on top of a qualifying Microsoft 365 licencePuts Copilot inside Excel, where your inventory data probably already lives

Prices as published on each vendor’s own pricing or listing page, checked August 2026. Ratings are the vendors’ own app store listings.[8,9,10,11,12,13]

Two things stand out to me in that table. The first is the gap between $29 and $900, which is not a gap in quality so much as a gap in what the product assumes about you: Netstock is built to sit on top of an ERP and needs six to ten weeks of implementation before it does anything.[12] If you don’t have an ERP, that price isn’t for you yet.

The second is the Stocky rating. Shopify’s own inventory app, free with POS Pro, sits at 3.0 out of 5 across 188 reviews on Shopify’s own store.[11] Free bundled tooling is worth trying precisely because it costs nothing, and it’s also worth checking the reviews before you build a process around it.

A first pass you can run in one afternoon

No new software, no integration, no project plan. Four hours and an export.

  1. Export three files. Your product list with on-hand quantity and unit cost, your sales by SKU for the last 24 months, and your purchase orders for the same period. If any of those doesn’t exist in exportable form, that’s your finding for the day and it’s a real one.
  2. Deduplicate the product list. Ask for probable duplicate groups with evidence and an explicit “unsure” pile. Fix them at source, in your actual system, not in the copy.
  3. Rank your dead stock. No sale in 180 days, sorted by capital tied up, with a running total. Then decide what you’re doing with the top ten lines this month.
  4. Calculate real lead times. Promised date against actual receipt date for every purchase order, by supplier, with an average and a worst case.
  5. Write reorder points for your top 20 SKUs. Real lead times, real average daily sales, a buffer you chose deliberately. As formulas in a spreadsheet you own.
  6. Make the model attack it. Paste the whole thing back and ask what assumptions are unstated and where the plan breaks if demand drops 20%.

Two rules while you run it. Don’t paste supplier contract terms, pricing agreements or anything commercially sensitive into a consumer chatbot; use product codes and aggregate figures, or a tool where your organisation has a contract and training on your data is switched off. And write the date next to every assumption, because in six months you’ll want to know whether your plan was wrong or whether the world moved, and those need different fixes.

Run this twice, a quarter apart. If the second run surfaces the same problems as the first and you’re still fixing them by hand, that’s the moment a platform starts making sense.

Where this goes wrong

Four failure patterns, in the order I see them.

Forecasting on top of unreconciled data. Covered above, and still the most common. The output looks fine, which is the entire problem. A forecast built on three spellings of one product is wrong in a way that no accuracy metric will reveal to you.

Buying the platform to avoid making the decision. Software will tell you what your data says. It will not tell you whether to carry a slow-moving line because a good customer expects it. In Netstock’s survey, 62% of businesses fell into what it labelled insufficient forward planning, sitting on slow stock or replenishing excess without cutting purchase orders, and that share went up, not down, year over year.[3] Those are businesses that already own the software.

Letting the model do arithmetic out of sight. If you cannot point at the cell where a reorder quantity came from, you cannot defend the purchase order, and eventually somebody will ask you to.

Chasing accuracy when the constraint is cash. Reliance on credit for inventory fell from 53% to 47% between 2024 and 2025 in that same survey, cash usage from 54% to 44%, and 27% now report no defined financing strategy at all.[3] If capital is your binding constraint, a 5% better forecast is worth far less than clearing the dead stock you already know about.

The single thing I’d have you take from all of this: AI is genuinely useful here, but the usefulness sits in getting your own numbers into a state where a decision is possible rather than in the demand curve. That work is unglamorous and checkable, and a small team can do it in an afternoon. The teams getting value from AI for inventory management are rarely the ones with the best model, they’re the ones who finally looked at what they own.

Frequently Asked Questions

Can AI accurately forecast demand for a small business?

Less well than the marketing suggests. In the M5 competition, 5,507 teams forecast 42,840 series of real Walmart sales and only 7.5% beat a simple exponential smoothing benchmark, while fewer than half beat a plain naive forecast. The best teams did beat it, by around 22%, but they were specialists working on clean data. For a small business with patchy history, expect a modest improvement over a sensible spreadsheet method, not a transformation.

How much sales history do I need before AI forecasting is worth trying?

Two full years is the practical minimum, because you need at least two cycles of your seasonality to distinguish a pattern from a one-off. More important than length is consistency: one product per row, consistent units of measure, discontinued lines marked as such, and physical counts done within the last quarter. A short clean history beats a long messy one every time.

What is the single biggest mistake when using AI for inventory management?

Forecasting on top of data you have not reconciled. A study of nearly 370,000 store inventory records at one retailer found 65% did not match the physical count, with an average error of about 35% of the stock actually on the shelf. Duplicate SKUs, inconsistent units and stale counts all produce forecasts that look completely plausible and are quietly wrong.

Do I need dedicated inventory software or will a spreadsheet and AI do?

Most small teams get further with a clean spreadsheet and a general-purpose assistant than with a platform they have not defined a process for. Published entry prices range widely: Zoho Inventory has a free plan and a $29 per organisation per month tier, Katana starts at $299 a month, and Netstock starts at $900 a month with a six to ten week implementation. Run your manual process twice, then buy if you keep hitting the same wall.

Is it safe to put my inventory data into an AI tool?

Product codes, quantities and dates are usually low risk. Supplier contract terms, negotiated pricing and customer-identifiable order data are not, and should not go into a consumer chatbot. Use aggregate figures and internal codes where you can, and reserve anything commercially sensitive for a tool where your organisation holds a contract and model training on your data is disabled.

About This Article

Every number here was checked against its original source in August 2026, and prices come only from each vendor's own pricing or listing page. Several widely circulated figures were left out because they could not be traced to a primary source, including the “carrying cost is 20 to 30% of inventory value” rule of thumb, which appears everywhere and originates nowhere I could verify. I also excluded the widely repeated claim that Shopify is retiring Stocky on a specific date, because the only source for it is a Shopify help page that would not load for me. The Netstock figures come from a vendor's own customer base and are labelled as such in the text. This is general operational guidance, not financial advice.

Sources

  1. IHL Group, “Retail Inventory Crisis Persists Despite $172 Billion in Improvements,” September 2025 (global analysis of retail inventory distortion). https://www.ihlservices.com/news/analyst-corner/2025/09/retail-inventory-crisis-persists-despite-172-billion-in-improvements/
  2. U.S. Census Bureau, Retailers: Inventories to Sales Ratio (RETAILIRSA), retrieved from FRED, Federal Reserve Bank of St. Louis; May 2026 observation, released 16 July 2026. https://fred.stlouisfed.org/series/RETAILIRSA
  3. Netstock, 2025 Benchmark Report: The State of Supply Chain Planning (platform data from 2,400+ customers plus a survey of 130+ users at businesses under $250M revenue), October 2025. https://www.netstock.com/research/supply-chain-planning-report/
  4. Makridakis, Spiliotis & Assimakopoulos, “The M5 Accuracy competition: Results, findings and conclusions” (5,507 teams, 42,840 Walmart series, 28-day horizon). https://statmodeling.stat.columbia.edu/wp-content/uploads/2021/10/M5_accuracy_competition.pdf
  5. Stephan Kolassa, “Can We Obtain Valid Benchmarks from Published Surveys of Forecast Accuracy?” Foresight: The International Journal of Applied Forecasting, Issue 11, Fall 2008. https://forecasters.org/wp-content/uploads/Valid-Benchmarks-from-Published-Surveys-of-Forecast-Accuracy_Foresight11.pdf
  6. Nicole DeHoratius & Ananth Raman, “Inventory Record Inaccuracy: An Empirical Analysis,” Management Science 54(4), 2008 (nearly 370,000 records across 37 stores). https://ideas.repec.org/a/inm/ormnsc/v54y2008i4p627-641.html
  7. ECR Retail Loss, “Measuring the Sales Impact of Improving Inventory Records” (matched test and control stores, approximately 233,000 SKUs, 12 weeks before and after correction). https://ecrloss.com/research-paper/improving-inventory-records/
  8. Zoho, Zoho Inventory plans and pricing page (US), checked August 2026. https://www.zoho.com/us/inventory/pricing/
  9. Katana, Katana pricing page, checked August 2026. https://katanamrp.com/pricing/
  10. Shopify App Store, Inventory Planner by Sage listing (4.4 rating, 130 reviews), checked August 2026. https://apps.shopify.com/inventory-planner
  11. Shopify App Store, Stocky listing (free with Shopify POS Pro, 3.0 rating, 188 reviews), checked August 2026. https://apps.shopify.com/stocky
  12. Netstock, pricing page and FAQ (pricing starts at $900/month; implementation typically 6 to 10 weeks), checked August 2026. https://www.netstock.com/pricing/
  13. Microsoft, Microsoft 365 Copilot plans and pricing for enterprise, checked August 2026. https://www.microsoft.com/en-us/microsoft-365-copilot/pricing/enterprise
Sana Mian
Sana Mian, Co-Founder of Future Factors AI

Sana is an AI educator and learning designer specialising in making complex ideas stick for non-technical professionals. She has trained 2,000+ learners across corporate teams, bootcamps, and keynote stages. Future Factors offers AI Bootcamps, Corporate Workshops, and Speaking & Consulting for businesses ready to adopt AI without the overwhelm.

More about Sana →

Psst, Hey You!

(Yeah, You!)

Want helpful AI tips flying Into your inbox?

Weekly tips. Real examples. Practical help for busy professionals.

We care about your data, check out our privacy policy.