Reference Glossary
Caption writing standards
Caption writing standards, commonly AP style, define a concise two-sentence structure — the first covering who, what, where, and when, the second giving context on why the image matters.
Why it matters in a DAM
When captions follow a consistent structure and are stored in the IPTC Description field, a DAM can support syndication, search, and automated caption display across channels without editors rewriting captions per platform. Inconsistent free-text captions — some missing the date, some missing the location — break any workflow that tries to parse or search that field programmatically, and force manual review before every republish instead of a clean automated feed.
A worked example
Common mistake
Captions get written inconsistently across contributors — some omit the date or location entirely, some bury the who/what in the second sentence — so when the DAM later needs to auto-populate a search filter or feed the caption into a syndication export, a large share of assets have unusable or incomplete structured caption data.
Caption writing standards, most widely codified through AP style, exist to make a caption predictable enough that an editor scanning dozens of images can extract the essential facts in seconds. The convention calls for two sentences: the first establishes who is in the photo, what is happening (in present tense), and where and when it was taken; the second, optional but common, explains why the moment is newsworthy or significant. That predictability isn’t just an editorial nicety — it’s what allows captions to be reliably parsed, searched, and reused across different publication contexts without a human rewriting them each time.
- Sentence 1 (who/what/where/when): present tense, full names and titles, city and state, date
- Sentence 2 (why): concise context on why the moment or image matters
In a DAM, caption text almost always lives in the IPTC Description field, and a consistent writing standard is what makes that field actually useful as structured data rather than just readable prose. When every caption reliably includes a date and location in a predictable position, a DAM can support date-range or location-based search against caption text, feed captions into a syndication export without manual review, and auto-populate accessibility alt text from the same source. When captions are written inconsistently — location sometimes present, sometimes missing, date format varying by contributor — none of that automation is reliable, and someone ends up manually checking captions before every republish.
Frequently asked
How long should a standard news caption be?
AP style guidance calls for roughly two concise sentences — the first establishing who, what, where, and when, the second providing context on why the image is significant.
What tense should the first sentence of a caption be written in?
Present tense for describing the action in the photo, which reflects the convention that a caption describes what is happening in the frozen moment rather than narrating it as a past event.
Should a caption repeat information already in the headline or article?
Some overlap with the article is normal since captions are often read independently of the body text, but a caption shouldn't be purely redundant — it should add specific detail (who exactly, when exactly) that a headline typically omits.
Where does caption text get stored so it travels with the image?
In the IPTC Description (also called Caption-Abstract) field embedded in the image file's metadata, so the caption stays attached to the file even when it moves outside the originating system.
Does caption writing style differ between print, wire, and digital-native newsrooms?
The core who/what/where/when structure is broadly shared, but specific conventions (date format, inclusion of a subject's age or hometown, length limits) can vary by outlet style guide, which is why a DAM benefits from documenting which house style its contributors should follow.
Can caption text be auto-generated and then edited by a human?
Increasingly yes for a first draft — AI-assisted caption drafting is common — but factual details like names, locations, and dates in a caption should always be human-verified before publication, since an inaccurate caption is a factual error, not just a style issue.
Sources
- AP style captions call for roughly two sentences: the first including who, what (in present tense), where, and when, with a second sentence providing context on why the photo matters. checked 2026-08-07 — AP style captions summary, GSU Photojournalism course notes