Many pictures at once

One square-on picture ties the two groups a third row only steadies

A wall surveyed from two groups of cameras either side of it is likeliest to come out twisted, and the twist is the two groups pitching in opposite senses. A third row of marks steadies it: the reported twist falls from ±5.05 to ±2.31 millimetres. One picture taken square to the wall, which belongs to neither group, ties them: ±1.25, from a third of the extra views, and a five-millimetre twist is then seen ninety-eight times in a hundred instead of fifty-six. The picture need not be square — twenty degrees off is nearly as good — and it can be taken from four times as far away and still beat the row.

Worth reading first: Another picture of the same sweep · The track and the scene together.

A wall’s twist can be tested against the error the adjustment reports took a flatness survey of a facade — six cameras in two groups, sixty degrees either side of the wall’s normal, every pose and every mark found in one bundle adjustment — and asked whether the adjustment’s own covariance could be trusted to say when the reconstructed wall’s twist was real. It could. A flat wall’s fitted twist cleared twice its reported error one time in twenty at every layout tried. What the layout decided was how large a real twist had to be before it was seen: ten marks in two rows reported the twist to ±5.05 millimetres and saw a five-millimetre twist seventeen times in a hundred; fifteen marks in three rows reported ±2.31 and saw it fifty-six.

The third row helps, the essay before that one found, because the two groups are tied to each other only through the marks both see, and a third row tightens the tie. It ended by naming the other way to tighten it. A few pictures taken square to the wall see every mark the two oblique groups see, and tie each group to the middle rather than to each other. The question was whether three square-on pictures, which cost a photographer a few seconds, buy what the third row buys — so that a plain wall with only a string course and a cornice to mark could still be surveyed for a saddle.

They buy more than the row does, and one picture buys almost all of it.

One picture against a row

The survey is the earlier essays’ throughout: a wall 6 metres wide, marked in two rows of five from half a metre to two and a half metres up; three cameras either side, 55, 60 and 65 degrees off the normal, 8 metres from the wall’s middle, with 50-degree lenses; every mark read in every picture to half a pixel; seven numbers held to fix the gauge, as seven numbers no picture can name requires. The twist is the saddle: the marked corners pushed out of the wall’s plane, opposite corners the same way, the middle left where it is. Its reported error is the standard deviation the adjustment’s covariance implies for it.

One picture taken square to the wall reports its twist to ±1.25 mm, where a third row of marks reaches ±2.31 and the two groups alone ±5.05A facade photographed from 8 m by two groups of three cameras 60° either side of its normal, every pose and mark found in one bundle adjustment, marks read to half a pixel: the standard deviation the adjustment reports for the wall's twist (the marked corners pushed out of the plane in a saddle). two groups, ten marks in two rows: ±5.05 mm, from 60 views; a third row: fifteen marks: ±2.31 mm, from 90 views; one square-on picture: ±1.25 mm, from 70 views; three square-on pictures: ±1.13 mm, from 90 views; three square-on and a third row: ±1.12 mm, from 135 views. One square-on picture adds ten views where a third row adds thirty, and in the currency an adjustment counts in — the reciprocal of the twist's variance — it adds 4.1 times what the row adds. Once a square-on picture is in, a third row changes the twist's error in the third decimal place: the two fixes are for one weakness, and the picture is the stronger.two groups, ten marks in two rows±5.05 mma third row: fifteen marks±2.31 mmone square-on picture±1.25 mmthree square-on pictures±1.13 mmthree square-on and a third row±1.12 mmthe twist's reported standard deviation±60° groups of three, marks to 0.5 px
Fig. 1 The twist’s reported standard deviation for the two groups alone, with a third row of marks, and with one or three pictures taken square to the wall.

The two groups alone report the twist to ±5.05 millimetres from sixty views — six cameras, ten marks. A third row of five marks adds thirty views and brings it to ±2.31. One picture taken square to the wall, from the same 8 metres, adds ten views and brings it to ±1.25. Three square-on pictures spread over ten degrees bring it to ±1.13, and adding the third row as well changes the third decimal place: ±1.12.

In the currency the adjustment itself counts in — information, the reciprocal of the variance — the single picture adds 4.1 times what the row adds, from a third as many views. And the two fixes are not additive. Once a square-on picture is in, the third row has almost nothing left to do, which says the two were remedies for one and the same weakness and the picture is the stronger remedy.

The twist is the two groups pitching against each other

The adjustment’s covariance says what the weakness is. Every quantity it fits is correlated with every other, and the twist’s correlations with the cameras’ poses are the place to look.

In the two-group adjustment the fitted twist moves with every camera's pitch (correlation 0.94–0.95, opposite signs on the two sides); with one square-on picture, at most 0.12The correlation, in the bundle's own covariance, between the wall's fitted twist and each camera's pitch — its rotation about its own horizontal axis — for the two ±60° groups alone and with one picture added square to the wall; ten marks in two rows read to half a pixel. Two groups: left −65° -0.94, left −60° -0.95, left −55° -0.95, right 55° 0.95, right 60° 0.95, right 65° 0.94. With the square-on picture: left −65° -0.03, left −60° -0.08, left −55° -0.12, right 55° 0.12, right 60° 0.08, right 65° 0.03, square-on 0° -0.00. The twist the adjustment cannot pin is the left group pitching one way and the right group the other: each group then reads depth across the wall slightly sheared, in opposite senses, and the two shears meet in a saddle. A picture taken square-on belongs to neither group and is pitched with neither, and every mark it sees is drawn where it is across and up the wall whatever its depth, so the counter-pitch no longer fits its marks.left −65°, two groups−0.94left −60°, two groups−0.95left −55°, two groups−0.95right 55°, two groups+0.95right 60°, two groups+0.95right 65°, two groups+0.94left −65°, with the square-on−0.03left −60°, with the square-on−0.08left −55°, with the square-on−0.12right 55°, with the square-on+0.12right 60°, with the square-on+0.08right 65°, with the square-on+0.03square-on 0°, with the square-on−0.00correlation of the twist with each camera's pitchsize shown, sign printed
Fig. 2 The correlation between the wall’s fitted twist and each camera’s pitch, for the two groups alone and with one square-on picture added.

In the two-group adjustment the fitted twist moves with every camera’s pitch — its rotation about its own horizontal axis — with a correlation of 0.94 or 0.95, negative for the three cameras on the left and positive for the three on the right. The twist is the left group tilting down a little and the right group tilting up, or the reverse. Each group reads the wall’s depth through its own oblique view, and a camera turned 60 degrees from the normal that is pitched by a hair reads its marks’ depths sheared across the wall: higher marks pushed one way, lower marks the other. The two groups pitched in opposite senses shear in opposite senses, and two opposite shears seen from opposite sides meet in a saddle. The marks both groups see constrain this only weakly, because a mark’s depth is the quantity each group reads worst.

A picture taken square to the wall belongs to neither group. It is not pitched with the left group or with the right, and in it every mark is drawn where it is across and up the wall almost regardless of its depth — so it says, with its full precision, where each mark is in the wall’s plane. A counter-pitch of the two groups moves the marks’ reconstructed positions in that plane as well as in depth — more, in fact: read off the same covariance, each millimetre of twist at the corners carries the marks 2 to 3.5 millimetres across the wall, the lower row one way and the upper row the other, and up to 1.6 millimetres up or down. Those positions no longer agree with what the square-on picture drew, and it reads them to a fraction of a millimetre. With one such picture added the correlations fall to at most 0.12. The loose mode is gone, and the twist is left with whatever error the readings themselves set.

That also explains why the third row helps less. A third row gives both groups more marks through which to hold their relative pitch, but each of those marks is again read in depth by two oblique groups. It tightens the tie through the same weak channel. The square-on picture ties through a different one.

The first picture does the work

If the square-on picture is a tie rather than a reading, the first one should do nearly everything and the rest little.

The first square-on picture takes the twist's error from ±5.05 to ±1.25 mm; five more take it only to ±1.07Square-on pictures added to the two ±60° groups, spread 10° wide about the wall's normal at 8 m, ten marks in two rows read to half a pixel: the twist's reported standard deviation 0 pictures, ±5.049 mm; 1 picture, ±1.245 mm; 2 pictures, ±1.167 mm; 3 pictures, ±1.125 mm; 4 pictures, ±1.101 mm; 6 pictures, ±1.072 mm. The dashed line is a third row of marks with no square-on picture, ±2.314 mm. The twist mode's share of the wall's flatness variance falls from 71% to 17% with the first picture and stays there. The first picture removes the weakness and the rest only add readings, and ten more readings are worth little beside the sixty the two groups already took.024012346square-on pictures added to the two groupsthe twist's reported standard deviation, mma third row insteadsquare-on picturesten marks in two rows, 0.5 px±60° groups of three at 8 m
Fig. 3 The twist’s reported standard deviation against the number of square-on pictures added; the dashed line is a third row of marks instead.

So it is. The first picture takes the reported twist from ±5.05 to ±1.25 millimetres. The second takes it to ±1.17, the third to ±1.13, and six reach ±1.07. The twist’s share of the wall’s whole flatness variance — the measure the bundle keeps the needles turned used to show that one shape dominated the error — falls from 71 per cent to 17 with the first picture and stays there with the rest. After the first picture, no single shape dominates: the error is spread over the wall’s departures the way the readings spread it.

The same picture closes most of the distance between the bundle and a survey whose camera poses were known beforehand. With the two groups alone, the bundle’s root-mean-square flatness error was 4.23 millimetres against 2.38 with every pose known; with one square-on picture, 2.28 against 2.21. The adjustment then knows the wall almost as well as a survey with a calibrated rig — which is the ceiling, since where the adjustment stops is that recovering poses from the pictures can only add error.

The third group need only be well clear of the other two

The argument was that a square-on picture breaks the counter-pitch because it belongs to neither group. If that is right, what matters is not squareness but separation, and a third group standing some way off the normal should work almost as well.

The third group need not be square-on, only well clear of the other two: three pictures 20° off the normal report ±1.28 mm, and at 40° they still beat a third rowA third group added to the two ±60° groups — three pictures spread 10° wide, or one picture — standing 0, 10, 20, 30, 40, 50° off the wall's normal on one side, at 8 m; ten marks in two rows read to half a pixel. The twist's reported standard deviation, three pictures: ±1.13, ±1.16, ±1.28, ±1.53, ±2.02, ±3.04 mm; one picture: ±1.25, ±1.30, ±1.46, ±1.82, ±2.53, ±3.92 mm. The dashed line is a third row of marks instead, ±2.31 mm, which three pictures beat up to 40° off the normal and one picture up to 30°. As the third group turns towards one of the two, it becomes a fourth and fifth camera of that group, pitched with it, and stops tying the groups together.0123401020304050how far the third group stands off the wall's normal, degrees (the others at 60°)the twist's reported standard deviation, mmthree picturesone picturea third row insteadten marks in two rows, 0.5 pxthe third group on one side
Fig. 4 The twist’s reported standard deviation as the third group stands further off the wall’s normal, towards one of the two groups at 60°; three pictures and one.

Three pictures standing 10 degrees off the normal report the twist to ±1.16 millimetres; 20 degrees off, ±1.28; 30, ±1.53; 40, ±2.02, still better than the third row. At 50 degrees, ten degrees from the group at 60, they report ±3.04 and have become a fourth, fifth and sixth camera of that group, pitched with it. A single picture follows the same curve a little higher and beats the row up to 30 degrees off.

So the practical instruction is loose. A photographer who has walked round two sides of a wall at sixty degrees should take a picture or two from anywhere within twenty or thirty degrees of square-on, and the instruction it replaces — mark a third row — is both more work and less effective.

From four times as far, still better than a row

The square-on picture need not be taken from where the groups stand either.

Taken from four times as far, the square-on picture still reports the twist to ±2.29 mm — better than a third row of marks at ±2.31One picture taken square to the wall from 8, 12, 16, 24, 32 m, added to the two ±60° groups at 8 m, its lens the groups' 50° so the wall shrinks in it with distance; ten marks in two rows read to half a pixel. The twist's reported standard deviation: 8 m, ±1.25 mm; 12 m, ±1.42 mm; 16 m, ±1.59 mm; 24 m, ±1.94 mm; 32 m, ±2.29 mm. The bundle's flatness error: 2.28, 2.40, 2.48, 2.63, 2.78 mm, against 2.21, 2.30, 2.34, 2.36, 2.37 with every pose known. The dashed line is a third row instead. A distant picture draws the wall small and reads its marks' positions less finely, but it still belongs to neither group, and the tie it makes does not depend on its resolution until the marks are a few pixels apart. Closer than the groups the wall leaves this lens's frame, and a closer square-on picture needs a wider lens.0123812162432how far from the wall the square-on picture is taken, m (the groups at 8 m)the twist's reported standard deviation, mma third row insteadone square-on pictureten marks in two rows, 0.5 pxthe same 50° lens throughout
Fig. 5 The twist’s reported standard deviation for one square-on picture taken from 8 to 32 m, with the groups’ 50° lens; the dashed line is a third row of marks instead.

With the same 50-degree lens, taken from 12 metres the picture reports the twist to ±1.42 millimetres, from 16 to ±1.59, from 24 to ±1.94, and from 32 metres — where the wall is a quarter the size in the frame and every mark’s position is read four times less finely in metres — to ±2.29, still a hair better than the third row. A distant picture is a weaker reading, but it is no less a member of neither group, and the tie it makes survives the loss of resolution until the marks are only a few pixels apart. A picture from across a street, taken because the photographer could not stand any nearer, is worth taking.

The other direction is closed by the lens rather than the geometry. Nearer than 8 metres, a 50-degree lens no longer holds the six-metre wall in its frame: at 6 metres six of the fifteen marks of a three-row wall fall outside it, and at 3 metres fourteen. A closer square-on picture needs a wider lens, and was not measured here.

What a surveyor now sees

The point of reporting the twist’s error was to say how large a real twist has to be before it is seen. With the error reported honestly — which the earlier essay established at every layout — the answer is the two tails of a normal distribution at each reported error.

A 5 mm twist is called real 98 times in a hundred with one square-on picture, 56 with a third row and 16 with neither; a 3 mm twist, 66, 24 and 8The share of reconstructions in which the fitted twist clears twice the error the adjustment reports, against the true twist, for a reported error of ±5.05, ±2.31, ±1.25 mm — two groups, two rows, a third row, one square-on picture. two groups, two rows: 5% 5% 6% 8% 12% 16% 21% 34% 49%; a third row: 5% 7% 13% 24% 39% 56% 72% 93% 99%; one square-on picture: 5% 12% 35% 66% 89% 98% 100% 100% 100% at 0, 1, 2, 3, 4, 5, 6, 8, 10 mm. The earlier measurement found the reported error honest at every layout — a flat wall clears twice it one time in twenty — so the curves are the normal's two tails at each reported error, and the layouts differ only in how large a real twist must be before it is seen.00.2500.5000.75010123456810how far the wall is truly twisted at its marked corners, mmshare of reconstructions calling the twist realtwo groups, two rowsa third rowone square-on picturethe test at twice the reported errordashed: one in twenty
Fig. 6 The share of reconstructions calling a twist real, against the true twist, for the two groups alone, with a third row, and with one square-on picture.

A 5-millimetre twist is called real 98 times in 100 with one square-on picture, 56 with a third row and 16 with neither. A 3-millimetre twist, 66, 24 and 8. A 2-millimetre twist, 35, 13 and 6. On a flat wall every layout calls a twist real 5 times in 100, as a test at twice its reported error should. The plain wall of the earlier question, with only a string course and a cornice to mark — two rows — was a wall on which a five-millimetre saddle went unseen five times in six; with one more picture, taken square from wherever the photographer can stand, it is seen every time but twice in a hundred.

Neither kind of picture surveys the wall alone

It would be a misreading to conclude that the square-on pictures are the survey and the groups a formality. Three pictures taken square to the wall and nothing else measure its flatness to a root-mean-square 37 millimetres with their poses known, and 94 with their poses found from the pictures — useless for finding a saddle of five. A square-on camera reads each mark’s depth along its own line of sight, which is the direction a single camera reads worst, and three cameras ten degrees apart barely separate their sight lines. Six of them spread over twenty degrees still give 16 and 25 millimetres.

The two groups are what measure depth. Their sight lines cross the wall’s normal at sixty degrees from either side, so between them they read each mark’s departure from the wall across their lines of sight, which is the direction every camera reads well. What they cannot do is hold themselves to each other. The square-on picture does exactly that and nothing else of consequence: it reads where each mark is in the wall’s plane, which is the one thing the counter-pitch disturbs that the groups read badly. The survey’s flatness comes from the groups, and its freedom from a saddle comes from the picture. Each kind of picture supplies what the other lacks, and the combination reaches 2.28 millimetres, within a tenth of a millimetre of what the same cameras give with every pose known in advance.

Opening the groups does less than one picture

The earlier essay found a second lever: opening the two groups further from the normal shrinks the twist, because steeper views read each mark’s depth with more of the reading’s precision. It is worth setting that against the picture. Groups at 45 degrees either side report the twist to ±5.83 millimetres; at 60, ±5.05; at 75, ±3.30. With one square-on picture added, the same three openings give ±1.71, ±1.25 and ±1.03. At every opening the picture does more than opening the groups by thirty degrees does, and the two levers combine: a steep pair of groups with a square-on picture between them is the best of the arrangements tried. A third row at the same three openings gives ±3.38, ±2.31 and ±1.47 — always between the groups alone and the groups with a picture.

That ordering is the same at every opening because the weakness is the same at every opening. However steep the two groups, they are two directions, and two directions leave a counter-rotation free; only the steepness of the views changes how much each rotation costs in the marks’ depths. A third direction removes the freedom rather than raising its price.

Why a picture is cheaper than a mark

There is a general point here about what a survey’s design is spending. Split the track and the needles turn found that splitting cameras into two groups either side of a wall turns each mark’s error needle across the wall, which is what makes the two-group survey good at flatness in the first place. The price is that the groups see the wall from two directions only, and anything that rotates one direction against the other is weakly held. Marks add constraints within those two directions. A camera from a third direction adds a direction.

That is why a single picture outperforms a whole row of marks. The row’s thirty new views are thirty more readings of the same two kinds; the picture’s ten are readings of a third kind. A survey is trusted at its own accuracy, unless its error has a shape warned that a reported accuracy is a fair summary only when the error is spread evenly; the two-group survey’s error had a shape — 71 per cent of it in one saddle — and the square-on picture is what removes the shape rather than shrinking it.

The linear world the numbers live in

Every number is from the adjustment’s covariance at the true solution. That is the first-order description of a least-squares fit, exact for small reading errors. Half a pixel is small for these cameras, but the earlier essay drew its reconstructions from the same linear model, so neither it nor this one checks what larger errors do to the adjustment’s own nonlinearity.

Every picture sees every mark. The square-on picture is placed where the wall fits its frame, and the groups at 8 metres see all marks. A picture that sees only part of the wall ties only the part it sees, and a third group that sees only the middle of a long wall would leave its ends to the two groups’ counter-pitch.

The marks are matched correctly in every picture. A square-on view of a regular facade is the view most prone to matching a mark to its neighbour, the failure a wrong match is not a small error is about; the earlier essay showed a mismatch big enough to fake a twist shows in its own residual, and the same test applies to the square-on picture’s marks.

The lens is the same in every picture and known. A square-on picture taken with another camera, of another focal length, adds its own unknown calibration, and whether the tie survives an unknown focal length in the picture that makes it was not measured.

Still open: whether the square-on picture can come from a different camera

A photographer surveying a wall with a calibrated camera may not have it for the last picture, or may take the square-on view with a phone from across the street. That picture’s focal length and principal point are then unknowns of the adjustment too, and a focal length is exactly the kind of number that trades against a depth.

The measurement that settles what such a picture is worth adds one square-on picture whose focal length — and, separately, whose principal point — the adjustment must find for itself, and asks how much of the tie survives: whether a picture square to the wall, which reads the marks’ positions in the wall’s plane, fixes its own focal length from the wall’s known width between the two groups’ readings, or whether the unknown focal length lets the picture scale itself along with the counter-pitch, so that the twist returns. The question with a number in it is how far the reported twist climbs back from ±1.25 millimetres towards ±5.05, and whether a second uncalibrated picture from another place, or a known distance between two marks, recovers it.

Shares its objects with

Essays that name at least two of the same things, and that neither author linked.

Named objects

A flat tag is an object no other essay names yet.

bundle adjustmentCovariancedegrees of freedomgauge freedomleast squaresResidual