The attributes and their allowed values are agreed once, then applied to every product as the file changes. Coverage, completeness and accuracy are measured and reported rather than asserted.
Automated coding is only worth having if it stays correct as the file moves. Three things keep it there.
The attributes and their allowed values are agreed up front and applied to every product the same way. Nothing is re-interpreted batch to batch, so one cycle is comparable to the next and to the one before it.
New products are coded as they arrive rather than in an annual push. The file is current when someone queries it, instead of current as of whenever the last project finished.
Accuracy is scored against a held-out set the coding pass never saw. That is the difference between a quality claim and a quality measurement, and it is the number that survives a stakeholder asking how you know.
A different engagement, on a retailer file where the product label was often the only description a product carried.
A coded file is infrastructure. What it is worth is whatever the systems downstream can finally do once every product is described the same way.
Take maintained attributes into search, audiences, space planning and reporting, on your own schema.
Thirty minutes on your categories. How the record gets built, what it holds, and the questions your team could put to it.