Home · Blog · Culling & AI

How to test AI culling accuracy on your own shoot

Osmel Contreras · Founder, Kepla · August 12, 2026 · 8 min read
The sharpest frame of a burst
Culling & AI

Every culling tool publishes an accuracy number, and not one of them can tell you whether it would have picked your gallery. The only test that answers that question uses a shoot you already delivered as the answer key. It takes about half an hour of your attention, it costs nothing, and here is exactly how to run it.

01 · WHY TEST IT YOURSELF

A published accuracy number is not about your work

Accuracy sounds like a property of the software. It is not. It is a score against an answer key, and the answer key belongs to whoever ran the test. When a vendor says its model is accurate, what it means is that the model agreed with a set of judgments made by people you have never met, on photos you did not shoot, for clients you do not have.

Your question is different and much narrower. You want to know whether this tool would have found the frames you delivered last spring. Nobody can answer that but you.

Be especially careful with cross-brand comparisons. FilterPixel publishes accuracy figures for its own product alongside figures for its rivals. That is a company testing itself and grading its competition, published as marketing. It may well be sincere. It is still not evidence, and neither is the equivalent chart from anyone else.

The wider argument about whether these models can be trusted with creative judgment at all is a separate piece, and we wrote it in can you trust AI to pick your photos. This article assumes you have stopped reading opinions and want a number of your own.

02 · THE ANSWER KEY

A shoot you have already delivered is a labeled data set. You went through every frame at full attention, you made a decision on each one, and a client paid you for the result. That is a better answer key than anything a vendor could build, because it encodes your taste, your standard, and the things you know about that family.

Pick your test shoot with a little care:

If you have never written down what you are actually selecting for, do that before you test anything, because you cannot grade a tool against a standard you have not stated. That exercise is its own job and we walked through it in building your culling criteria.

03 · THE SETUP

What you need before you start the clock

Four things: the complete original folder, a list of the filenames you delivered, a trial account, and a pass mark you write down before you look at anything.

The trial accounts are the easy part. Every serious tool in this category will let you test it properly for free, which is the whole reason this test is worth running rather than arguing about.

ToolWhat the free trial gives youWhere it runs
Aftershoot Select30 days, unlimited images, no card requiredYour own computer
Narrative SelectFull trial of the top Ultra tier, no card requiredYour own computer
Imagen2 culling projects freeCloud, photos upload first
FilterPixel14 days, first 10,000 photos, no card requiredYour own computer

Trial terms checked August 2026 from each vendor's own page: Aftershoot, Narrative, Imagen, FilterPixel.

Note the third column, because it changes how you plan the half hour. A local tool starts working the moment you point it at a folder. A cloud tool has to receive the photos first, and on an ordinary home connection a full wedding is not a coffee break. That is a real part of the cost of using it, so let it show up in your test rather than starting your timer after the upload finishes.

Write your pass mark down first

Decide, in advance and in writing, what result would make you buy and what result would make you walk away. Do it before you see a single pick. Everyone who skips this step finds a way to be pleased with whatever number they get, because by then they have spent an afternoon on it and they want it to have been worth doing.

04 · THE METHOD

The half hour test, step by step

The thirty minutes is your attention, not the computer's. The analysis run happens without you.

  1. Copy the folder. Work on a duplicate of the original shoot. Not because these tools are dangerous, but because a test you are nervous about is a test you will cut short.
  2. Build your answer key. Get the filenames of the images you actually delivered into a list. Your gallery host will export a file list, or you can pull the names straight from your delivered folder. Put them in column A of a spreadsheet.
  3. Run the tool on everything. The whole shoot, every frame, at default settings. Do not hand it a curated subset, and do not skip the getting-ready coverage because it is boring. Start it and walk away.
  4. Export what it picked. Every one of these tools can give you its selection as a list of filenames, or as flags you can filter in Lightroom and then export. Put that list in column B.
  5. Match the two lists. In column C, next to each delivered filename, use =IF(COUNTIF(B:B,A2)>0,"picked","missed") and fill it down. That gives you your misses in one pass instead of an hour of squinting.
  6. Look at every miss with your own eyes. This is the step that matters and it is the step people skip. The count tells you almost nothing. The frames tell you everything.

Run it a second time with the tool's sensitivity turned up, if it has that control. Most do, and the default is a setting somebody chose for an average photographer who is not you. A tool that fails at default and passes when you loosen it is a tool that works, with a note in the manual.

05 · THE SCORING SHEET

The scoring sheet, blank

Four boxes. Fill them in from your spreadsheet, for your shoot.

You delivered itYou did not deliver it
The tool picked itAgreed keepers: ____Extras: ____
The tool did not pick itMisses: ____Agreed passes: ____

One division gives you the only headline figure worth having. Take your agreed keepers, divide by agreed keepers plus misses, and you have the share of your real gallery the tool found on its own. Call it keeper recall. It answers the question you actually asked.

The second number, extras divided by everything the tool picked, tells you how much padding you have to read past. It is worth writing down but it is worth far less worry, and that asymmetry is the most important thing on this page.

The two errors do not cost the same. An extra costs you about a second: you glance at a frame, you do not pick it, you move on. A miss costs you a photograph, and if you have trusted the shortlist and never opened the rest, it costs you that photograph permanently, in a gallery a client keeps forever. A tool that hands you a slightly bloated shortlist is doing its job. A tool that quietly loses the shot of the groom's father is not, however good its headline percentage looks.

06 · READING THE RESULT

How to read your own numbers

There is no industry pass mark, and anyone who quotes you one is quoting their own marketing. That is exactly why you wrote yours down in advance. Compare your result to that, not to a number from a chart.

Then put the percentage aside and go through the missed frames one at a time, asking three questions:

Now look at the other side. Go through the extras, the frames it picked that you did not deliver. Some will be duplicates you had already chosen between, and you can ignore those. But a few will be genuinely good frames you walked past at midnight with three hundred to go, and finding one of those is worth more than any percentage on your sheet. It is not a defect in the tool. It is a defect in tired human attention, which is the actual thing you are trying to buy your way out of.

One more read of the sheet: your extras pile is also a report on your own consistency. If the tool keeps offering you frames you cannot articulate a reason for skipping, your standard may be looser than you think.

07 · WHERE IT GOES WRONG

Five ways this test quietly lies to you

The method is simple. Fooling yourself with it is also simple.

When you have your sheet filled in for two shoots, you have something almost nobody in this market has: an accuracy figure for your own work, produced by a method you can explain. That is the basis for a decision, and it turns choosing culling software from a matter of taste into a matter of evidence.

One tool you cannot test yet

Ours. Kepla's picking app for iPhone, iPad and Mac is still being built, so there is nothing for you to point at a folder today and we are not going to pretend otherwise. Two things we will commit to now: it picks and never touches your files, so everything it passes over stays exactly where it is and stays one tap away, and we would rather you ran this test on us than took our word for it. The booking page is live and free today while we build the rest.

08 · COMMON QUESTIONS

FAQ

How long does it take to test an AI culling tool?

About thirty minutes of your attention per shoot, plus machine time you can walk away from. Copying the folder, exporting your delivered filenames and matching the two lists in a spreadsheet is roughly ten minutes. Looking properly at every frame the tool missed takes the other twenty, and it is the part worth doing slowly. Cloud tools add upload time on top, which can be substantial for a full wedding.

Can I test AI culling software for free?

Yes, and you should never pay to run this test. Checked in August 2026, Aftershoot gives 30 days with no card, Narrative gives a full trial of its top tier with no card, FilterPixel gives 14 days and your first 10,000 photos with no card, and Imagen includes two free culling projects. That is enough free access to test three or four tools on the same shoot and compare them fairly.

What is a good accuracy score for AI culling?

There is no published industry benchmark worth quoting, and vendor figures are graded against their own answer keys rather than your gallery. That is why the method here asks you to write down your own pass mark before you run the test. What matters more than the percentage is the character of the misses: losing a few interchangeable frames is a shrug, losing one photograph you would have fought for is not.

Should I test on a wedding or a portrait session?

Test on both if you shoot both, because they stress different things. A wedding tests volume, mixed light and long burst sequences. A portrait or family session tests fine judgment between very similar frames, where the difference between the keeper and the reject is one expression. A tool can be strong at one and ordinary at the other, and you will only see that if you try both.

Is my client work at risk during a test like this?

Work on a copy of the folder and the question mostly goes away. The real thing to check is where the photos go: desktop tools analyze on your own machine and nothing leaves it, while cloud tools require you to upload the shoot first, which may matter for your client contracts. Whichever you test, note that these tools mark and flag rather than discard, so your originals stay put.

How many shoots should I test before deciding?

Two at minimum, chosen to be as different from each other as your work gets. One shoot tells you how a tool handled one set of conditions, which is not the same as knowing how it will behave in your season. If the two results are close, you have a reliable picture. If they are far apart, test a third and look for what the weak one had in common with it.

FOUNDING COHORT · 100 SEATS

Never sort photos at 1 AM again.

Kepla for Mac clears the obvious misses from a card, names the reason on every frame it sets aside, and leaves the choosing to you. Nothing is ever deleted, moved or renamed. Free through the private preview · the first hundred photographers keep it at $99 a year.

Request Founding Access
KEEP READING