Intent: Why You Press the Shutter

Subject: Who the Photograph Is About

Twelve chapters on arranging the frame. This one asks the earlier question — what are you actually saying?

A photograph that cannot say who its subject is will not be rescued by composition or light. Visual weight is the mechanism that decides, and it does not take instructions.
A photograph that cannot say who its subject is will not be rescued by composition or light. Visual weight is the mechanism that decides, and it does not take instructions. Dorothea Lange, 1936 · Public domain · Source

Show someone a photograph of yours and ask: “what did you see first?”

If their answer is not what you had in mind, that is not their failing.

This chapter is about closing that gap, and the first step is accepting something uncomfortable: where the viewer looks is not up to you.

The question to ask on location is this:

If only one thing could stay in the frame, which one?

Until you can answer it, it is not yet time to press the shutter.

The question is not the same as composition. Composition asks where things go; subject asks why it is this thing at all. Part two taught the first, and the first only means anything once the second has an answer — a subject you are unsure of stays unsure when you place it on a third.

Your eye is not looking at the whole photograph

The mechanism first, because it explains every rule that follows.

Only a small central patch of the retina — the fovea — resolves fine detail, covering roughly two degrees of vision, about a thumbnail at arm’s length. Resolution collapses rapidly outside it.

So “I am looking at this photograph” is an illusion. What actually happens is that the eye jumps three or four times a second, resting for two to three hundred milliseconds at a time, and the impression of a whole picture is assembled from a sequence of small patches. The jumps are called saccades, and vision is suppressed during them — you never see the movement itself.

Which makes the real question: how does the eye choose where to jump next?

It can only use the coarse information peripheral vision supplies, and peripheral vision has definite characteristics:

  • Highly sensitive to brightness differences and to motion — evolutionarily the two things that mattered.
  • Poor at detail.
  • Poor at colour too — cones are concentrated in the fovea and the periphery is nearly all rods.

That already yields a strong practical conclusion: luminance decides where the viewer looks; hue does not. The previous part’s rule about a small saturated accent carried an unstated condition — the accent must also differ in brightness, or be large enough for peripheral vision to register. A saturated patch at the same luminance as its surroundings is close to invisible out there.

Visual weight, though, only settles the first look. In the 1960s the Soviet researcher Alfred Yarbus ran a now-classic experiment: the same viewer, the same painting, but a different question each time — how wealthy is this family, how old are the people, memorise where everyone is standing. The recorded scan paths barely resemble one another.

Both halves of that matter. The first two or three jumps are set by the physical properties of the picture and are yours to control; everything after is set by the question in the viewer’s head and is not. What composition can actually do is make sure the first landing is the right one — because that landing determines which question the viewer carries into the rest of the frame.

This is worth more than it sounds, because most photographs today are not contemplated. They are met while scrolling, in the second or two before a thumb moves — a handful of saccades in total. For most of your audience, the first fixation is not the beginning of the encounter. It is very nearly the whole of it.

Faces are the one exception, and it is an exception at the hardware level. A dedicated region of the temporal lobe processes them, fast enough to finish before you are aware of it, and crude arrangements of two dots and a line will set it off — which is why you see faces in power sockets, car fronts and clouds.

The consequence is blunt: any face in the frame is the subject, whether you wanted it or not. The stranger walking past in the background counts.

Visual weight

The mechanism, arranged as something usable:

What adds weightWhy it worksHow you use itHow it hurts you
Faces, eyesDedicated neural hardware, too fast to resistMake the subject the only facePassers-by, faces on posters
The brightest regionPeripheral vision’s strongest cuePut the subject in the lightA blown corner of sky, a window
The highest-contrast edgeSaccade targets are chosen by edge densitySeparate subject from background in luminanceRailings, branches, wire fencing
The sharpest regionSharp things get foveal confirmation firstFocus only on the subjectSomething behind the subject is sharper
The most saturated colourStrong in the fovea, weak in the peripheryMake the subject the only saturated thingSignage, bins, advertising
Whatever differs mostDifference is itself a signalOne red note in grey, one diagonal among verticalsThe accidental exception — a crooked sign
Gaze and implied motionWe automatically follow what others look atHave people in frame look at the subjectThe subject looks out of frame and takes the viewer with them

The “no clear subject” complaint is almost always one row of this table landing in the wrong place. And note the right-hand column: every tool here is simultaneously a trap.

Nearly all the visual weight in this photograph stacks in one place: the mother's face is the brightest, the highest in contrast and the sharpest thing in the frame. The children turning away does two jobs at once — it removes two competing faces, and their posture pushes attention towards her. They are not supporting cast; they are the device that moves your eye.
Nearly all the visual weight in this photograph stacks in one place: the mother's face is the brightest, the highest in contrast and the sharpest thing in the frame. The children turning away does two jobs at once — it removes two competing faces, and their posture pushes attention towards her. They are not supporting cast; they are the device that moves your eye. Dorothea Lange, 1936 · Public domain · Source

Letting the subject win

There are only two families of method: add weight to it, or take weight off everything else.

To add weight to the subjectTo take it off everything else
Put it in the brightest lightLet the background fall into shade
Separate it from the background in luminanceOpen the aperture and blur the background
Focus on it aloneMove closer and exclude the clutter
Lead the eye to it with linesWait for the clutter to leave
Give it space around itFind a clean wall to put behind it

The two columns are not equally effective. Subtraction almost always wins, because visual weight is relative — darkening the background by a stop is equivalent to brightening the subject by one, without touching the subject’s texture or skin.

And within subtraction, moving closer is the most underrated move there is. Beginners photograph from the middle distance, where the frame contains a little of everything and enough of nothing. Three steps forward solves several problems at once: clutter leaves the frame, the subject’s relative area grows, and the fall-off of the light becomes visible. One action, three fixes, no equipment.

A counter-example. Nothing is technically wrong — exposure is right, detail is clean, colour is normal. Now try to say who it is about. You cannot: stalls, signs, people and produce all carry similar weight and none of them has been singled out by light or luminance. The frame is rich and has no protagonist. Rich and clear are different things, and they frequently work against each other.
A counter-example. Nothing is technically wrong — exposure is right, detail is clean, colour is normal. Now try to say who it is about. You cannot: stalls, signs, people and produce all carry similar weight and none of them has been singled out by light or luminance. The frame is rich and has no protagonist. Rich and clear are different things, and they frequently work against each other. Dietmar Rabich, 2018 · CC BY-SA 4.0 · Source

The subject need not be large

The figure is small and off-centre. Look at what surrounds him: row upon row of empty chairs, a desk at the right with papers and an inkwell on it, and the lower half of another person standing at the top of the frame. None of it competes — every one of those things is lower in brightness and contrast than his face — so the isolation itself becomes the weight. The empty chairs are not background; they are saying this place was meant to be full.
The figure is small and off-centre. Look at what surrounds him: row upon row of empty chairs, a desk at the right with papers and an inkwell on it, and the lower half of another person standing at the top of the frame. None of it competes — every one of those things is lower in brightness and contrast than his face — so the isolation itself becomes the weight. The empty chairs are not background; they are saying this place was meant to be full. John Vachon, 1940 · Public domain · Source

A small subject surrounded by emptiness can carry enormous weight, because isolation is a form of contrast. Peripheral vision sees a large unvarying field, and the single variation in the middle of it becomes the only place worth jumping to.

The extreme version. The subject is reduced to a mark and everything else is sky and sea. What this picture is about is not the figure but **the ratio between the figure and the emptiness** — the ratio is the content. Crop half the emptiness away and nothing is left.
The extreme version. The subject is reduced to a mark and everything else is sky and sea. What this picture is about is not the figure but **the ratio between the figure and the emptiness** — the ratio is the content. Crop half the emptiness away and nothing is left. Dietmar Rabich, 2022 · CC BY-SA 4.0 · Source

The test is not what fraction of the frame the subject occupies. It is whether the relationship between subject and surroundings is what the photograph is about. If the surroundings are helping to say something, keep them; if they merely happen to be there, cut them.

This is also where empty space is most often misused. Space is not “nothing beside the subject”. For it to work, the emptiness has to mean something — exposure, isolation, waiting, insignificance. Emptiness without a meaning is just frame you paid for and did not use.

The subject need not be a thing

Everything so far has assumed the subject is an object — a person, a building, a bird. Many good photographs have subjects that are not objects:

  • A relationship. The distance between two people, the ratio of a person to a building. No single element is the protagonist here; the protagonist is the space between them, and that space has to be kept clean, with nothing else allowed into it.
  • A quality of light. The whole of the part on light. Who the light falls on is secondary; what matters is the shape, edge and direction of the light itself.
  • A condition. Rain just stopped, the crowd just left, the fog has not lifted. This kind of subject has no boundary — it is distributed across the whole frame.

This has a practical consequence: when the subject is a relationship or a condition, “move closer” stops working and can actively damage the picture. Moving in severs the relationship and crops away the surroundings the condition depends on. Which returns to the same sentence: decide what you are saying first, and only then can a rule help you.

The focus point is not the subject

A common confusion worth separating: focusing is a technical act, choosing a subject is a narrative one. They usually coincide. They are not the same level of decision.

Shallow depth of field does add visual weight — sharp things get foveal confirmation first, as the table said. But it has two side effects.

First, it discards the information in the surroundings along with the clutter. If the photograph is about a person’s relationship to a place, blurring the place past recognition deletes half the subject. Wide apertures often make pictures look more professional because they force a subtraction; that subtraction is not necessarily the one you wanted.

Second, it lets you off the hook. Blurring a messy background at f/1.4 is much easier than walking twenty steps to find a clean wall, and the results differ: blurred clutter is still clutter, now in the form of coloured, bright blobs that go on holding weight in peripheral vision. A genuinely clean background contains nothing, not something you cannot make out.

A suggested order of operations: solve the background with your feet, and give the aperture only what is left.

One photograph, one thing

The last principle, and the hardest: a photograph says one thing.

Trying to convey expression, atmosphere, scale and weather at once usually leaves all four vague. Better to make four photographs — which is exactly what the next two chapters are about.

How do you know whether what you want to say is one thing? A serviceable test: say it in one sentence that contains no “and”. “A mother is worried” is one thing. “A mother is worried, and the camp is squalid, and this is the Depression” is three. They are all true, and one photograph will not hold three.

Cameron says one thing here: this man's face. The background carries no information, the body dissolves into black, and even the detail of the clothing has been given up. The more completely everything else is abandoned, the louder the one remaining thing becomes.
Cameron says one thing here: this man's face. The background carries no information, the body dissolves into black, and even the detail of the clothing has been given up. The more completely everything else is abandoned, the louder the one remaining thing becomes. Julia Margaret Cameron / Adam Cuerden, 1867 · Public domain · Source

Six common symptoms and their causes

Always: the picture shows X → the usual cause is Y → change Z.

  1. The first thing people see is not your subject → one row of the table has landed in the wrong place, usually “the brightest region” → find the brightest patch in the frame; if it is not the subject, darken it or exclude it.
  2. A stranger in the background has taken the photograph over → faces are a hardware-level priority and do not distinguish lead from extra → there is no compromise: wait for them to leave, change angle, or blur them past being recognisable as a face.
  3. The picture looks rich but says nothing in particular → too many elements at similar weight, none clearly winning → subtract: move closer, use a longer lens, wait for some of it to leave. Adding will not fix this.
  4. The subject is large and still does not stand out → size is not the main source of visual weight; luminance and contrast are → separate subject and background in brightness, or find a cleaner background.
  5. Plenty of empty space and it reads as merely empty → the emptiness has no meaning and is not part of the content → ask what it is saying; if you cannot answer, crop it and use the close version instead.
  6. The subject looks out of frame and the viewer drifts out too → implied gaze is a powerful lead and it is currently pointing outside → leave space on the side they are looking towards, or wait for them to turn back. Cropping tight on the gaze side makes the frame tense — a deliberate effect, not a default.

Lange made six frames that day

The hero of this chapter is Dorothea Lange’s Migrant Mother, 1936. It is close to a textbook case of all the visual weight stacking in one place, but what is worth learning is how she arrived at that frame.

She had stopped at a pea-pickers’ camp at Nipomo, California. By her own later account she made five exposures — the Library of Congress holds six, and further frames survive in other collections — and whether five or six, they are a very short and visibly converging sequence: from a wider frame including the tent and its surroundings, progressively closer, until the last one is the composition we know — the mother and three children, with the background almost entirely excluded.

That process is this chapter’s two rules being executed:

  • Move closer. Every step removed information about the setting and increased the mother’s relative weight.
  • One photograph, one thing. The wider frames are simultaneously about the camp, the condition of migrant labour, and this family. None of them is bad. None of them is as certain about who it is about.

One further detail is worth knowing. Those photographs were made for the federal Resettlement Administration — reorganised in 1937 into the Farm Security Administration everyone now names — whose photographic section was directed by Roy Stryker, who issued his photographers shooting scripts — not “go and make some pictures” but explicit lists of what was to be documented. In that programme, “what are you actually saying” was written down before anyone picked up a camera. The order of this chapter — decide who first, arrange the frame second — is not a modern teaching device. It is how those photographs were made.

(And because they were the output of a government programme, they are in the public domain today, which is why this site can use them — and why the same names keep appearing throughout this part.)

A table for the field

What to confirmHow to check on the spotWhat to do if it fails
Is the subject the brightest thing?Squint until detail goes; see which patch stays brightestDarken the rest, or wait for the light to move
How many faces are in frame?Count them, including background and postersMore than one: get it down to one
Is the subject clipped?Check all four edgesStep back half a pace, or turn the camera
Does the empty space mean something?Say aloud what it is doingIf you cannot, move closer
Where does the gaze go?Follow the direction people in frame are lookingPointing out of frame: leave space on that side
What is this about?Say it in one sentenceIf it takes two, make it two photographs

Further watching

  • Story of "Migrant Mother" Photograph by Dorothea Lange - American Artifacts

    C-SPAN · 4 min

    Why this one This chapter’s closing section says Lange started that day with a wider frame that included the whole tent and its surroundings, and worked closer. In four minutes a Library of Congress curator lays that sequence out in order: the distant tent at 0:24, the tent with the family at 1:36, the back of the truck at 2:00, the tent mouth from closer in at 3:24, the mother nursing at 3:36, and only at 3:48 the frame everyone knows. At 3:29 she says it plainly: "There’s a series of pictures showing the teenage girl out in front of their tent… and then gradually she gets closer and closer and makes the famous photograph." At 0:32 there is a detail this chapter does not carry: Lange drove past that camp, made some pictures, got back in the car, and turned around partway home because "I didn’t do what I was supposed to do. I didn’t get the picture. I didn’t say what needed to be said" — which is the question this chapter opens with. At 2:48 there is a photograph of Lange herself holding a camera. It ends on a C-SPAN programme card.

  • What it means if you can see faces in objects - Susan G. Wardle

    TED-Ed · 5 min

    Why this one This chapter calls faces a hardware-level exception, fast enough to finish before you are aware of it. This gives that claim its numbers: at 1:15, recognising most objects takes about a quarter of a second while a face takes a tenth. At 1:31 it goes further — people reported seeing faces in over 35% of pure-noise images, a system that would rather false-alarm than miss. At 2:53 comes the line most useful here: the brain works out the face is illusory within about a quarter of a second, and you go on seeing it anyway. So a spare face in your frame stealing attention is not a matter of discipline — knowing it is irrelevant does not make it stop. The whole thing is illustrated animation: no photographs, and not one sentence about photography. Take the mechanism, not the examples.