The user wants me to proofread OCR output of a historical table about vessels entered at Hong Kong ports in 1880. The OCR has many errors: broken words, misaligned columns, garbled characters, missing data, etc. I need to reconstruct the table in Markdown format.
First, I need to understand the structure. The table has columns: NATIONALITY OF VESSELS, then three sections: WITH CARGOES, IN BALLAST, TOTAL. Each section has Vessels, Tons, Crews. So total columns: 1 (Nationality) + 3*3 = 10 columns? Actually: With Cargoes: Vessels, Tons, Crews; In Ballast: Vessels, Tons, Crews; Total: Vessels, Tons, Crews. That's 10 columns.
But the OCR output is messy. Let's parse line by line.
The header lines:
"(S6)
No. 3.-NUMBER, TONNAGE, and CREWS of VESSELS of EACH NÀTION ENTERED at Ports in the Colony of Hongkong, in the Year 1880.
ENTERED.
NATIONALITY OF VESSELS.
With CaroOKS.
IN BALLAST.
TOTAL.
Versels.
Tons.
Crews Vescis.
Tons. Crews. Vessels. Tons. Crews."
So the header rows:
Row1: NATIONALITY OF VESSELS | With Cargoes | In Ballast | Total
Row2: Vessels | Tons | Crews | Vessels | Tons | Crews | Vessels | Tons | Crews
But the OCR shows "With CaroOKS." likely "With Cargoes." "Versels." likely "Vessels." "Vescis." likely "Vessels." So we have 9 data columns plus nationality.
Now data rows. Let's list each nationality line as appears:
180 117,88 3,713
701.א
158
130 126,133 3,869"
This is messy. Probably:
American: With Cargoes: Vessels=180, Tons=117,888? Wait "117,88" maybe 117,888? But later total shows 130 126,133 3,869. Let's see: The line "180 117,88 3,713" then "701.א 158" then "130 126,133 3,869". Might be: With Cargoes: 180 vessels, 117,888 tons, 3,713 crews? But 117,88 could be 117,888? Actually typical numbers: 117,888 tons? But then In Ballast: 701 vessels? That seems high. Wait "701.א" maybe "701" and "א" is a glitch. Then "158" maybe crews? Then Total: 130 vessels? That doesn't add up: 180+701=881, not 130. So maybe the columns are misaligned.
Let's look at the total row at the end: "TOTAL 19,421 3,724,593 336,554 7,380 461,252 85,333 20,801 4,185,845 421,887". That gives totals: With Cargoes: Vessels=19,421, Tons=3,724,593, Crews=336,554; In Ballast: Vessels=7,380, Tons=461,252, Crews=85,333; Total: Vessels=20,801, Tons=4,185,845, Crews=421,887.
So the sum of vessels with cargoes across nationalities = 19,421. In ballast = 7,380. Total = 20,801.
Now we need to parse each nationality row to match these totals.
Let's go through each nationality line as they appear in OCR:
That's 20 nationalities.
We need to extract numbers for each. The OCR lines are broken. Let's try to reconstruct by looking at patterns.
The OCR text after header:
"American,
180 117,88 3,713
701.א
158
130 126,133 3,869
Annamese,
17
:
1
78
17
Austrian,
3.170
90
:
3,179
90
Belgian,
1,125
34
Bolivian,
740
15
:
:.
:
1
1,125
34
1
740
15
British,
2,043 1,851,701 85,422
30
20,534 850
2,073 1,878,325 86,272
Chinese,
150 124,107) 7,370
5
2,000 171
155 126,287 7,547
Chinese Junks,
16,610 1,240,088, 219,701
7,310 409,270 83,625 23,920 1,650,258 303,386
Danish,
द्म
000 5,390
1
783
24
-
6,173] 246
Dutch,
16 12.200 807
3
1,354
58
19 13.60
865
French,
113
165,152 9,422
1
300
11
114
165,452 9,433
German,
225
134,085 4,440
21
12,156
438
240
140,841 4,887
Honduras,
1
510
13
1
510
13
J
Japanese,........
28 20,076 2,077
28
26,076 2,077
Norwegian,
3,681
87
Portuguese,......
1
6321
18
:
:
:
S,681
87
1
832
18
Russian.
2,274
145
10
2,274
145
Siamese,
30 14,905 802
90
14,905 882
Spanish,
53 19,038 2,005
Swedish,
578
19
:
⠀
:
59
19,038 2,005
2
578
19
F.
TOTAL 19,421 3,724,593 336,554
7,380 461,252 85,333 20,801 4,185,845 421,887"
We need to parse each. It seems each nationality block has lines for With Cargoes, In Ballast, Total. But the OCR has line breaks and missing delimiters.
Let's attempt to parse by assuming each nationality has three rows: With Cargoes (Vessels, Tons, Crews), In Ballast (Vessels, Tons, Crews), Total (Vessels, Tons, Crews). But the OCR sometimes merges them.
Better approach: The original table likely had columns aligned. The OCR read line by line, but the columns are not separated by tabs. We need to infer from numbers.
We have the total row at the end, which gives us the column totals. We can use that to check each nationality's numbers.
Let's list nationalities and try to assign numbers.
First, American:
Lines:
"American,
180 117,88 3,713
701.א
158
130 126,133 3,869"
Probably:
With Cargoes: Vessels=180, Tons=117,888? But "117,88" could be 117,888? However, the total tons for American is 126,133? Wait the last line "130 126,133 3,869" likely is Total: Vessels=130, Tons=126,133, Crews=3,869. But 130 vessels total? Then With Cargoes 180? That can't be. Maybe the first line is "180 117,888 3,713" but 180 vessels with cargoes, 117,888 tons, 3,713 crews. Then In Ballast: "701 158"? Actually "701.א 158" maybe "701 158"? But need three numbers: Vessels, Tons, Crews. Could be "701 158 ?" missing tons? Or "701 158" maybe Vessels=701, Tons=158? That seems too low tons. Or "701 158" could be Vessels=701, Crews=158? But tons missing. Then Total: 130 126,133 3,869. That doesn't add: 180+701=881, not 130. So maybe the first line is not American? Wait the header says "With CaroOKS." then "IN BALLAST." then "TOTAL." Then "Versels. Tons. Crews Vescis. Tons. Crews. Vessels. Tons. Crews." So the first data row after header might be American with cargoes: 180 vessels, 117,888 tons, 3,713 crews. Then next line "701.א 158" might be In Ballast: 701 vessels, 158 tons? But 158 tons for 701 vessels is unrealistic. Could be 701 vessels, 158,000 tons? But written as "158"? Or maybe "701 158" is actually "701 158,000"? But the OCR didn't capture commas.
Let's look at other nationalities for pattern.
Annamese:
"Annamese,
17
:
1
78
17"
Probably: With Cargoes: 17 vessels, ? tons, ? crews. In Ballast: 1 vessel, 78 tons, 17 crews? Total: 17? Not sure.
Austrian:
"Austrian,
3.170
90
:
3,179
90"
Maybe: With Cargoes: 3 vessels? "3.170" could be 3,170 tons? But need vessels, tons, crews. "3.170 90" maybe 3 vessels, 1,170 tons, 90 crews? Then In Ballast: none? Then Total: 3,179 90? That seems like tons and crews.
Belgian:
"Belgian,
1,125
34
Bolivian,
740
15
:
:.
:
1
1,125
34
1
740
15"
This is messy. It seems Belgian and Bolivian are interleaved. Probably:
Belgian: With Cargoes: 1 vessel? 1,125 tons, 34 crews. In Ballast: none? Total: 1,125 34.
Bolivian: With Cargoes: 1 vessel, 740 tons, 15 crews. Total: 1,740 15.
But the lines show "1 1,125 34" and "1 740 15" at the end.
British:
"British,
2,043 1,851,701 85,422
30
20,534 850
2,073 1,878,325 86,272"
This looks clear: With Cargoes: 2,043 vessels, 1,851,701 tons, 85,422 crews. In Ballast: 30 vessels, 20,534 tons, 850 crews. Total: 2,073 vessels, 1,878,325 tons, 86,272 crews. Good.
Chinese:
"Chinese,
150 124,107) 7,370
5
2,000 171
155 126,287 7,547"
With Cargoes: 150 vessels, 124,107 tons, 7,370 crews. In Ballast: 5 vessels, 2,000 tons, 171 crews. Total: 155 vessels, 126,287 tons, 7,547 crews. Good.
Chinese Junks:
"Chinese Junks,
16,610 1,240,088, 219,701
7,310 409,270 83,625 23,920 1,650,258 303,386"
With Cargoes: 16,610 vessels, 1,240,088 tons, 219,701 crews. In Ballast: 7,310 vessels, 409,270 tons, 83,625 crews. Total: 23,920 vessels, 1,650,258 tons, 303,386 crews. Good.
Danish:
"Danish,
द्म
000 5,390
1
783
24
-
6,173] 246"
This is garbled. "द्म" is Devanagari? Probably OCR error. "000 5,390" maybe 1,000? Or 5,390 tons? "1 783 24" maybe In Ballast: 1 vessel, 783 tons, 24 crews. Total: "6,173] 246" maybe 6,173 tons, 246 crews? But need vessels. Could be With Cargoes: 1 vessel? Actually Danish might have 1 vessel with cargoes? Let's see total tons for Danish: 6,173? And crews 246? But we need to match totals.
Dutch:
"Dutch,
16 12.200 807
3
1,354
58
19 13.60
865"
With Cargoes: 16 vessels, 12,200 tons, 807 crews. In Ballast: 3 vessels, 1,354 tons, 58 crews. Total: 19 vessels, 13,600 tons, 865 crews. But "13.60" likely 13,600. Good.
French:
"French,
113
165,152 9,422
1
300
11
114
165,452 9,433"
With Cargoes: 113 vessels, 165,152 tons, 9,422 crews. In Ballast: 1 vessel, 300 tons, 11 crews. Total: 114 vessels, 165,452 tons, 9,433 crews. Good.
German:
"German,
225
134,085 4,440
21
12,156
438
240
140,841 4,887"
With Cargoes: 225 vessels, 134,085 tons, 4,440 crews. In Ballast: 21 vessels, 12,156 tons, 438 crews. Total: 240 vessels, 140,841 tons, 4,887 crews. Good.
Honduras:
"Honduras,
1
510
13
1
510
13"
With Cargoes: 1 vessel, 510 tons, 13 crews. In Ballast: none? Total: 1 vessel, 510 tons, 13 crews. But there are two lines "1 510 13" repeated. Probably With Cargoes and Total same.
Japanese:
"J
Japanese,........
28 20,076 2,077
28
26,076 2,077"
With Cargoes: 28 vessels, 20,076 tons, 2,077 crews. In Ballast: none? Total: 28 vessels, 26,076 tons, 2,077 crews? But total tons 26,076 > 20,076, so maybe In Ballast has some tons but same vessels? Actually "28 26,076 2,077" might be total. But then In Ballast missing. Could be In Ballast: 0 vessels, 6,000 tons? Not sure.
Norwegian:
"Norwegian,
3,681
87
Portuguese,......
1
6321
18
:
:
:
S,681
87
1
832
18"
This is interleaved Norwegian and Portuguese. Let's separate.
Norwegian:
"Norwegian,
3,681
87
...
S,681
87"
Probably With Cargoes: 1 vessel? "3,681" tons, 87 crews. In Ballast: none? Total: 3,681 tons, 87 crews. But "S,681" maybe 3,681? So Norwegian: 1 vessel, 3,681 tons, 87 crews.
Portuguese:
"Portuguese,......
1
6321
18
:
:
:
1
832
18"
With Cargoes: 1 vessel, 6,321 tons, 18 crews. In Ballast: none? Total: 1 vessel, 832 tons? That doesn't match. Maybe "6321" is 6,321? But total shows 832. Could be OCR error: 6321 vs 832. Might be 832 tons. Let's see: "1 6321 18" then later "1 832 18". Probably the first is With Cargoes: 1 vessel, 6,321 tons? But total 832? That's inconsistent. Maybe the first is "1 832 18" and OCR misread 832 as 6321? Actually "6321" could be "832" with a smudge. But we'll need to decide.
Russian:
"Russian.
2,274
145
10
2,274
145"
With Cargoes: 10 vessels? "2,274 145" maybe tons and crews. Then "10" maybe vessels? Then Total: 2,274 145. So Russian: 10 vessels, 2,274 tons, 145 crews.
Siamese:
"Siamese,
30 14,905 802
90
14,905 882"
With Cargoes: 30 vessels, 14,905 tons, 802 crews. In Ballast: 90? Wait "90" alone. Then Total: 14,905 882. That seems off. Maybe In Ballast: 0 vessels? But "90" could be crews? Actually "30 14,905 802" then "90" then "14,905 882". Could be With Cargoes: 30 vessels, 14,905 tons, 802 crews. In Ballast: 0 vessels, 0 tons, 90 crews? No. Total: 30 vessels, 14,905 tons, 882 crews? But 802+90=892, not 882. Hmm.
Spanish:
"Spanish,
53 19,038 2,005
Swedish,
578
19
:
⠀
:
59
19,038 2,005
2
578
19"
Spanish and Swedish interleaved.
Spanish: "53 19,038 2,005" then later "59 19,038 2,005"? Actually after Swedish lines, we see "59 19,038 2,005". So Spanish: With Cargoes: 53 vessels, 19,038 tons, 2,005 crews. In Ballast: 6 vessels? 59-53=6? But no tons/crews for ballast. Total: 59 vessels, 19,038 tons, 2,005 crews. So In Ballast: 6 vessels, 0 tons, 0 crews? Unlikely.
Swedish: "578 19" then "2 578 19". So Swedish: With Cargoes: 2 vessels, 578 tons, 19 crews. In Ballast: none? Total: 2 vessels, 578 tons, 19 crews.
Now, we have the total row: With Cargoes total vessels = 19,421. Let's sum the With Cargoes vessels we have so far (from clear ones):
British: 2,043
Chinese: 150
Chinese Junks: 16,610
Dutch: 16
French: 113
German: 225
Honduras: 1
Japanese: 28
Norwegian: 1? (if 1 vessel)
Portuguese: 1
Russian: 10
Siamese: 30
Spanish: 53
Swedish: 2
American: 180? (from first line)
Annamese: 17
Austrian: 3? (from "3.170" maybe 3 vessels)
Belgian: 1
Bolivian: 1
Danish: ? maybe 1
Sum these: 2,043+150=2,193; +16,610=18,803; +16=18,819; +113=18,932; +225=19,157; +1=19,158; +28=19,186; +1=19,187; +1=19,188; +10=19,198; +30=19,228; +53=19,281; +2=19,283; +180=19,463; +17=19,480; +3=19,483; +1=19,484; +1=19,485; +? Danish. That's already 19,485, exceeding 19,421. So my assumptions are off. Probably American is not 180. Let's recalc using the total row as guide.
We need to parse each nationality correctly. The OCR is messy but we can try to reconstruct the table by aligning columns based on the total row and the clear rows (British, Chinese, Chinese Junks, Dutch, French, German, etc.). The unclear rows (American, Annamese, Austrian, Belgian, Bolivian, Danish, Japanese, Norwegian, Portuguese, Russian, Siamese, Spanish, Swedish) need to be inferred.
Maybe the OCR output is actually a single string per nationality with numbers separated by spaces, but line breaks are arbitrary. Let's look at the raw text again and try to split by nationality.
The text after "F." (maybe "F." is a footer) then "TOTAL ...". But before that, each nationality seems to be separated by a blank line? In the OCR, there are blank lines between some.
Let's copy the OCR text and split by double newline? But the user gave a single block. I'll treat each nationality as a block separated by a blank line in the OCR. However, the OCR may not have consistent blank lines.
Better: I'll write a script mentally to parse. But as a language model, I can manually parse.
Let's list the lines sequentially with line numbers:
17.
24.
31.
47.
53.
59.
63.
72.
80.
89.
98.
106.
112.
128.
135.
140.
154.
Now, we need to assign each nationality 9 numbers (Vessels_cargo, Tons_cargo, Crews_cargo, Vessels_ballast, Tons_ballast, Crews_ballast, Vessels_total, Tons_total, Crews_total). Some nationalities may have no ballast entries (i.e., zeros). The table likely includes all three sections for each.
From the clear ones (British, Chinese, Chinese Junks, Dutch, French, German, Honduras), we see pattern: each nationality block has 3 lines: first line: With Cargoes (three numbers), second line: In Ballast (three numbers), third line: Total (three numbers). But in OCR, these lines are split across multiple lines.
For British: line 49: "2,043 1,851,701 85,422" -> With Cargoes. Line 50-51: "30" and "20,534 850" -> In Ballast: 30 vessels, 20,534 tons, 850 crews. Line 52: "2,073 1,878,325 86,272" -> Total.
For Chinese: line 55: "150 124,107) 7,370" -> With Cargoes. Line 56-57: "5" and "2,000 171" -> In Ballast: 5 vessels, 2,000 tons, 171 crews. Line 58: "155 126,287 7,547" -> Total.
For Chinese Junks: line 61: "16,610 1,240,088, 219,701" -> With Cargoes. Line 62: "7,310 409,270 83,625 23,920 1,650,258 303,386" -> This line contains both In Ballast and Total? Actually it has six numbers: 7,310 409,270 83,625 (In Ballast) and 23,920 1,650,258 303,386 (Total). So line 62 combines both.
For Dutch: line 74: "16 12.200 807" -> With Cargoes. Line 75-77: "3", "1,354", "58" -> In Ballast: 3, 1,354, 58. Line 78-79: "19 13.60", "865" -> Total: 19, 13,600, 865.
For French: line 82: "113" -> only vessels? Then line 83: "165,152 9,422" -> tons and crews. So With Cargoes: 113, 165,152, 9,422. Line 84-86: "1", "300", "11" -> In Ballast: 1, 300, 11. Line 87-88: "114", "165,452 9,433" -> Total: 114, 165,452, 9,433.
For German: line 91: "225" -> vessels. Line 92: "134,085 4,440" -> tons, crews. Line 93-95: "21", "12,156", "438" -> In Ballast. Line 96-97: "240", "140,841 4,887" -> Total.
For Honduras: line 100: "1", line 101: "510", line 102: "13" -> With Cargoes: 1, 510, 13. Line 103-105: "1", "510", "13" -> Total (same). In Ballast missing (probably zeros).
For Japanese: line 109: "28 20,076 2,077" -> With Cargoes. Line 110: "28" -> maybe In Ballast vessels? Line 111: "26,076 2,077" -> Total tons and crews? But Total vessels should be 28? Actually line 111 has two numbers: 26,076 and 2,077. So Total: 28 vessels, 26,076 tons, 2,077 crews. In Ballast: 0 vessels? But line 110 "28" might be In Ballast vessels? That would duplicate. Probably line 110 is a stray. Actually Japanese might have no ballast. So With Cargoes: 28, 20,076, 2,077. Total: 28, 26,076, 2,077. But tons differ. That suggests In Ballast has some tonnage but zero vessels? Unlikely. Maybe the With Cargoes tons is 20,076 and In Ballast tons is 6,000? But no vessels. Let's check other nationalities: sometimes In Ballast has vessels but zero? Not sure.
For Norwegian and Portuguese interleaved: lines 113-127. Need to separate.
Norwegian: line 114: "3,681", line 115: "87". Then later line 123: "S,681", line 124: "87". So Norwegian might have With Cargoes: 1 vessel? 3,681 tons, 87 crews. In Ballast: none. Total: 3,681 tons, 87 crews. But "S,681" likely "3,681". So Norwegian: 1 vessel, 3,681 tons, 87 crews.
Portuguese: line 117: "1", line 118: "6321", line 119: "18". Then line 125: "1", line 126: "832", line 127: "18". So Portuguese: With Cargoes: 1 vessel, 6,321 tons, 18 crews? But total shows 832 tons. Could be OCR error: 6321 vs 832. Maybe the correct is 832 tons. The "6321" might be "832" with a smudge. Or maybe With Cargoes: 1 vessel, 832 tons, 18 crews. And the "6321" is a misread of "832" with a '6' and '3'? Actually "832" could be misread as "6321" if the '8' looks like '6' and '3'? Not sure. We'll need to decide based on totals.
Russian: line 130: "2,274", line 131: "145", line 132: "10", line 133: "2,274", line 134: "145". So With Cargoes: 10 vessels, 2,274 tons, 145 crews. In Ballast: none. Total: same.
Siamese: line 137: "30 14,905 802" -> With Cargoes: 30, 14,905, 802. Line 138: "90" -> maybe In Ballast crews? Line 139: "14,905 882" -> Total tons and crews. So Total: 30 vessels, 14,905 tons, 882 crews. In Ballast: 0 vessels, 0 tons, 80 crews? But 802+80=882. So In Ballast: 0 vessels, 0 tons, 80 crews? That seems odd. Maybe In Ballast: 0 vessels, 0 tons, 0 crews, and the 90 is a misprint. Or maybe In Ballast: 90 crews? But no vessels. Could be that the 90 is actually the In Ballast vessels? But then total vessels would be 120. Not matching.
Spanish and Swedish interleaved: lines 141-153.
Spanish: line 142: "53 19,038 2,005" -> With Cargoes: 53, 19,038, 2,005.
Then line 149: "59 19,038 2,005" -> Total: 59, 19,038, 2,005.
So In Ballast: 6 vessels, 0 tons, 0 crews? Or maybe In Ballast: 6 vessels, some tons/crews but not recorded? The total tons same as cargo, so ballast tons=0.
Swedish: line 144: "578", line 145: "19" -> With Cargoes: 2 vessels? Actually line 151: "2", line 152: "578", line 153: "19" -> Total: 2, 578, 19. So With Cargoes: 2, 578, 19. In Ballast: none.
Now American, Annamese, Austrian, Belgian, Bolivian, Danish are messy.
Let's compute totals from clear nationalities to see what's left for the messy ones.
We'll sum With Cargoes vessels, tons, crews for clear nationalities:
Clear nationalities (with confident numbers):
Sum vessels: 2,043+150=2,193; +16,610=18,803; +16=18,819; +113=18,932; +225=19,157; +1=19,158; +28=19,186; +1=19,187; +1=19,188; +10=19,198; +30=19,228; +53=19,281; +2=19,283.
Total With Cargoes vessels from total row: 19,421. So remaining vessels for American, Annamese, Austrian, Belgian, Bolivian, Danish = 19,421 - 19,283 = 138 vessels.
Now sum tons for these clear nationalities:
British: 1,851,701
Chinese: 124,107
Chinese Junks: 1,240,088
Dutch: 12,200
French: 165,152
German: 134,085
Honduras: 510
Japanese: 20,076
Norwegian: 3,681
Portuguese: 832
Russian: 2,274
Siamese: 14,905
Spanish: 19,038
Swedish: 578
Sum = let's calculate:
1,851,701
+124,107 = 1,975,808
+1,240,088 = 3,215,896
+12,200 = 3,228,096
+165,152 = 3,393,248
+134,085 = 3,527,333
+510 = 3,527,843
+20,076 = 3,547,919
+3,681 = 3,551,600
+832 = 3,552,432
+2,274 = 3,554,706
+14,905 = 3,569,611
+19,038 = 3,588,649
+578 = 3,589,227
Total With Cargoes tons from total row: 3,724,593. Difference = 3,724,593 - 3,589,227 = 135,366 tons for the remaining 6 nationalities.
Sum crews for clear:
British: 85,422
Chinese: 7,370
Chinese Junks: 219,701
Dutch: 807
French: 9,422
German: 4,440
Honduras: 13
Japanese: 2,077
Norwegian: 87
Portuguese: 18
Russian: 145
Siamese: 802
Spanish: 2,005
Swedish: 19
Sum = 85,422+7,370=92,792; +219,701=312,493; +807=313,300; +9,422=322,722; +4,440=327,162; +13=327,175; +2,077=329,252; +87=329,339; +18=329,357; +145=329,502; +802=330,304; +2,005=332,309; +19=332,328.
Total With Cargoes crews from total row: 336,554. Difference = 336,554 - 332,328 = 4,226 crews for remaining 6 nationalities.
Now we have 6 nationalities: American, Annamese, Austrian, Belgian, Bolivian, Danish. They account for 138 vessels, 135,366 tons, 4,226 crews.
Now let's parse each from OCR.
American block (lines 12-16):
"American,
180 117,88 3,713
701.א
158
130 126,133 3,869"
We need to extract three rows: With Cargoes, In Ballast, Total. The last line "130 126,133 3,869" looks like Total: 130 vessels, 126,133 tons, 3,869 crews. That would be the total for American. Then With Cargoes and In Ballast should sum to that.
If Total vessels = 130, then With Cargoes + In Ballast = 130. The first line "180 117,88 3,713" has 180 vessels, which exceeds 130. So maybe the first line is not With Cargoes. Could be that the first line is actually the With Cargoes for American? But 180 > 130. Unless the total is not sum? But total should be sum. So perhaps the first line is not American's With Cargoes. Maybe the OCR misaligned: the line "180 117,88 3,713" might belong to another nationality? But it's right after "American,".
Let's look at the numbers: 180, 117,88 (maybe 117,888), 3,713. Then "701.א 158" - 701 and 158. Then total 130, 126,133, 3,869. If With Cargoes = 180, In Ballast = 701, total would be 881, not 130. So that's not it.
Maybe the columns are shifted: The header shows "With Cargoes", "In Ballast", "Total". But the OCR might have read the columns vertically? No.
Another possibility: The table might have multiple sub-columns? No.
Let's check the Annamese block: lines 18-23.
"Annamese,
17
:
1
78
17"
If Total for Annamese is 17 vessels? The last number 17. The first 17 might be With Cargoes vessels. Then ":" maybe separator. Then "1" might be In Ballast vessels. Then "78" tons? Then "17" crews? But total crews 17? Not sure.
Austrian: lines 25-30.
"Austrian,
3.170
90
:
3,179
90"
"3.170" could be 3,170 tons? "90" crews. Then ":" then "3,179" tons, "90" crews. So With Cargoes: 3,170 tons, 90 crews. Total: 3,179 tons, 90 crews. Vessels? Not given. Maybe 1 vessel? Or 3 vessels? The "3.170" might be "3 170" meaning 3 vessels, 170 tons? But 170 tons is low. Could be 3 vessels, 1,170 tons? The dot might be a thousands separator? In European notation, dot is thousands separator. So "3.170" = 3,170. That is tons. So vessels missing. Maybe the number of vessels is in the previous line? But there is none. Could be that Austrian has 1 vessel? But then tons 3,170. That's plausible for a steamship.
Belgian and Bolivian interleaved: lines 32-46.
"Belgian,
1,125
34
Bolivian,
740
15
:
:.
:
1
1,125
34
1
740
15"
This suggests:
Belgian: With Cargoes: 1,125 tons, 34 crews. In Ballast: none. Total: 1,125 tons, 34 crews. Vessels? Probably 1 vessel (since "1" appears later). The line "1 1,125 34" appears at line 41-43. So Belgian: 1 vessel, 1,125 tons, 34 crews.
Bolivian: With Cargoes: 740 tons, 15 crews. Total: 740 tons, 15 crews. Vessels: 1 (line 44-46: "1 740 15").
Danish: lines 64-71.
"Danish,
द्म
000 5,390
1
783
24
-
6,173] 246"
"द्म" is garbage. "000 5,390" maybe 5,390 tons? "1 783 24" could be In Ballast: 1 vessel, 783 tons, 24 crews. "6,173] 246" could be Total: 6,173 tons, 246 crews. With Cargoes vessels? Maybe 1 vessel? But then total vessels would be 2. But we need to fit into the remaining 138 vessels.
Now, we have 6 nationalities: American, Annamese, Austrian, Belgian, Bolivian, Danish. Their total vessels = 138. We have some clues:
Maybe my clear nationalities sum is off because I missed some nationalities? Let's list all nationalities from the table: The OCR shows: American, Annamese, Austrian, Belgian, Bolivian, British, Chinese, Chinese Junks, Danish, Dutch, French, German, Honduras, Japanese, Norwegian, Portuguese, Russian, Siamese, Spanish, Swedish. That's 20. I have 14 clear (British, Chinese, Chinese Junks, Dutch, French, German, Honduras, Japanese, Norwegian, Portuguese, Russian, Siamese, Spanish, Swedish) = 14. The remaining 6 are American, Annamese, Austrian, Belgian, Bolivian, Danish. So 14+6=20.
My clear sum vessels = 19,283. Total = 19,421. Difference = 138. So the 6 nationalities sum to 138 vessels.
Now, from the OCR, American total line says 130 vessels. That would leave 8 for the other 5. But Annamese first number is 17. So maybe the American total line is not 130 vessels. Let's re-examine American block.
Lines:
12: American,
13: 180 117,88 3,713
14: 701.א
15: 158
16: 130 126,133 3,869
Perhaps the numbers are arranged in columns but OCR read them row-wise incorrectly. The original table likely had columns: Nationality | With Cargoes (Vessels, Tons, Crews) | In Ballast (Vessels, Tons, Crews) | Total (Vessels, Tons, Crews). The OCR might have read the first row of data (American) across the columns but line breaks inserted.
Maybe the line "180 117,88 3,713" corresponds to With Cargoes: Vessels=180, Tons=117,888, Crews=3,713.
Then "701.א 158" corresponds to In Ballast: Vessels=701, Tons=158? But 158 tons for 701 vessels is impossible. Could be 701 vessels, 158,000 tons? But written as 158? Or maybe "701.א" is "701" and "א" is a symbol for something else. "158" could be crews. Then In Ballast: 701 vessels, ? tons, 158 crews. Tons missing.
Then "130 126,133 3,869" corresponds to Total: Vessels=130, Tons=126,133, Crews=3,869.
But 180+701=881, not 130. So maybe the With Cargoes vessels is not 180. Could be that "180" is actually the In Ballast vessels? And "701" is With Cargoes? Let's test: If With Cargoes vessels = 701, In Ballast = 180, total = 881. Not 130.
Maybe the numbers are not vessels but tons? No, the header says Vessels, Tons, Crews.
Let's look at the total row for American from the final total? The final total is for all nationalities. We need to extract American's contribution from the total by subtracting others. But we don't have others fully.
Maybe the OCR has merged two lines: The line "180 117,88 3,713" might be for American With Cargoes, and the line "701.א 158" might be for Annamese? But Annamese starts at line 18.
Wait, line 14: "701.א" and line 15: "158" are between American and Annamese. Could be that "701.א 158" is actually the In Ballast for American? But then Annamese starts at line 18 with "Annamese,".
Let's check the original scan layout. Perhaps the table has two pages? The "(S6)" might indicate page 6. The table might be split across pages. But we have only this OCR.
Another approach: The total row at the end gives totals for all. We can try to reconstruct each nationality by solving a system? But we have 20 nationalities, each with 9 numbers, too many unknowns.
But we have the OCR text which contains the numbers, albeit scrambled. We need to clean up the OCR errors and present the table as best as possible in Markdown.
The user wants the proofread text in Markdown, reconstructing the table. We should output a Markdown table with the data as it appears in the original, correcting obvious OCR errors.
We need to decide on the correct numbers for each nationality. Since this is a historical record, we might be able to infer from known data? But we must not add external knowledge. We must only correct OCR errors based on internal consistency.
We have the total row which is likely correct. We have several nationalities that are clear. For the messy ones, we can try to parse the OCR lines as they are, correcting obvious misreads (like "द्म" -> maybe "1" or something, "701.א" -> "701", "117,88" -> "117,888", "3.170" -> "3,170", "12.200" -> "12,200", "13.60" -> "13,600", "6321" -> "832"? etc.)
We also need to preserve the structure: the table has three sections: With Cargoes, In Ballast, Total. Each with Vessels, Tons, Crews.
We should produce a Markdown table with columns: Nationality, With Cargoes Vessels, With Cargoes Tons, With Cargoes Crews, In Ballast Vessels, In Ballast Tons, In Ballast Crews, Total Vessels, Total Tons, Total Crews.
We'll fill in the numbers as best we can.
Let's go through each nationality systematically, using the OCR lines and the clear pattern.
We'll create a dictionary for each nationality.
First, list nationalities in order as they appear:
Now, for each, extract numbers from the OCR lines.
We'll write a parsing algorithm mentally.
British (clear):
With Cargoes: 2,043 vessels, 1,851,701 tons, 85,422 crews
In Ballast: 30 vessels, 20,534 tons, 850 crews
Total: 2,073 vessels, 1,878,325 tons, 86,272 crews
Chinese (clear):
With Cargoes: 150 vessels, 124,107 tons, 7,370 crews
In Ballast: 5 vessels, 2,000 tons, 171 crews
Total: 155 vessels, 126,287 tons, 7,547 crews
Chinese Junks (clear):
With Cargoes: 16,610 vessels, 1,240,088 tons, 219,701 crews
In Ballast: 7,310 vessels, 409,270 tons, 83,625 crews
Total: 23,920 vessels, 1,650,258 tons, 303,386 crews
Dutch (clear):
With Cargoes: 16 vessels, 12,200 tons, 807 crews
In Ballast: 3 vessels, 1,354 tons, 58 crews
Total: 19 vessels, 13,600 tons, 865 crews
French (clear):
With Cargoes: 113 vessels, 165,152 tons, 9,422 crews
In Ballast: 1 vessel, 300 tons, 11 crews
Total: 114 vessels, 165,452 tons, 9,433 crews
German (clear):
With Cargoes: 225 vessels, 134,085 tons, 4,440 crews
In Ballast: 21 vessels, 12,156 tons, 438 crews
Total: 240 vessels, 140,841 tons, 4,887 crews
Honduras (clear):
With Cargoes: 1 vessel, 510 tons, 13 crews
In Ballast: 0 vessels, 0 tons, 0 crews (not shown, but total same)
Total: 1 vessel, 510 tons, 13 crews
Japanese (lines 107-111):
"J
Japanese,........
28 20,076 2,077
28
26,076 2,077"
Interpretation: With Cargoes: 28 vessels, 20,076 tons, 2,077 crews.
In Ballast: 0 vessels, 0 tons, 0 crews? But there is a line "28" alone. Could be In Ballast vessels = 28? But then total vessels would be 56. The total line "26,076 2,077" has only two numbers, missing vessels. Probably the total vessels is 28 (same as cargo). The "28" on line 110 might be a stray or the In Ballast vessels (0). Actually, looking at other nationalities, the In Ballast line often has three numbers. Here line 110 is just "28". Could be that the In Ballast entry is missing and the "28" is actually the Total vessels? But then total tons 26,076. So maybe With Cargoes: 28, 20,076, 2,077; In Ballast: 0, 6,000, 0? But no. Let's check the total tons for Japanese in the grand total? We can't know.
Given the pattern, I think the Japanese entry is: With Cargoes: 28, 20,076, 2,077; In Ballast: 0, 0, 0; Total: 28, 26,076, 2,077. But why would total tons be higher than cargo tons if no ballast? That doesn't make sense. Unless the "20,076" is tons of cargo, and "26,076" is total tonnage (including ballast?). But ballast tonnage is usually the ship's tonnage when empty? Actually "Tonnage" likely refers to vessel tonnage (net register tonnage), not cargo weight. So a vessel has a fixed tonnage. The "With Cargoes" and "In Ballast" refer to whether the vessel entered with cargo or in ballast. The tonnage is the vessel's tonnage. So the same vessel has the same tonnage regardless. So the total tonnage should be the sum of tonnages of vessels entered with cargo plus those in ballast. So if a vessel enters with cargo, its tonnage is counted under With Cargoes. If it enters in ballast, its tonnage is counted under In Ballast. The total tonnage is the sum. So for Japanese, if 28 vessels entered with cargo (total tonnage 20,076), and some entered in ballast (adding tonnage), total tonnage would be higher. But the total vessels would be 28 + ballast vessels. The total line shows "26,076 2,077" with no vessel count. The vessel count might be missing. The line "28" before that might be the ballast vessel count? If ballast vessels = 28, then total vessels = 56. But then total crews 2,077 same as cargo crews? That seems unlikely.
Let's look at other nationalities: For British, cargo vessels 2,043, ballast 30, total 2,073. Cargo tons 1,851,701, ballast 20,534, total 1,878,325. Cargo crews 85,422, ballast 850, total 86,272. So ballast adds vessels, tons, crews.
For Japanese, if cargo: 28 vessels, 20,076 tons, 2,077 crews. If ballast: 28 vessels, 6,000 tons, 0 crews? But total crews 2,077 suggests ballast crews = 0. But ballast vessels would have crews. So not.
Maybe the line "28 20,076 2,077" is actually the total? And the cargo and ballast are not shown? But the table should have all three.
Given the difficulty, I'll assume the OCR for Japanese is garbled and we should present what we can: perhaps the data is: With Cargoes: 28, 20,076, 2,077; In Ballast: 0, 0, 0; Total: 28, 20,076, 2,077? But the total line shows 26,076. Could be a typo: 20,076 vs 26,076. The OCR might have misread "20,076" as "26,076" due to a smudge. Or the "28" on line 110 is actually the In Ballast vessels, and the total vessels is 56, but the total line omitted the vessel count. However, the total line for other nationalities always has three numbers. For Japanese, the total line has two numbers. So likely the vessel count is missing. The "28" on line 110 might be the total vessels. Then the total tons 26,076, crews 2,077. Then cargo line "28 20,076 2,077" would be identical except tons. So maybe cargo tons 20,076, ballast tons 6,000, total 26,076. But ballast vessels? If total vessels 28, and cargo vessels 28, then ballast vessels 0. But then ballast tons 6,000 with 0 vessels impossible. So maybe cargo vessels 20? Not given.
Given the confusion, I'll look at the grand total for In Ballast vessels: 7,380. Sum of In Ballast vessels from clear nationalities: British 30, Chinese 5, Chinese Junks 7,310, Dutch 3, French 1, German 21, Honduras 0, Japanese ?, Norwegian ?, Portuguese ?, Russian 0, Siamese ?, Spanish ?, Swedish 0. Sum so far: 30+5=35; +7,310=7,345; +3=7,348; +1=7,349; +21=7,370. So remaining In Ballast vessels for other nationalities = 7,380 - 7,370 = 10. Those 10 must be distributed among American, Annamese, Austrian, Belgian, Bolivian, Danish, Japanese, Norwegian, Portuguese, Russian, Siamese, Spanish, Swedish. But Russian, Swedish have 0. Belgian, Bolivian likely 0. So American, Annamese, Austrian, Danish, Japanese, Norwegian, Portuguese, Siamese, Spanish share 10 ballast vessels.
Similarly, In Ballast tons total = 461,252. Clear sum: British 20,534, Chinese 2,000, Chinese Junks 409,270, Dutch 1,354, French 300, German 12,156, others? Sum = 20,534+2,000=22,534; +409,270=431,804; +1,354=433,158; +300=433,458; +12,156=445,614. Remaining = 461,252 - 445,614 = 15,638 tons for the other nationalities.
In Ballast crews total = 85,333. Clear sum: British 850, Chinese 171, Chinese Junks 83,625, Dutch 58, French 11, German 438. Sum = 850+171=1,021; +83,625=84,646; +58=84,704; +11=84,715; +438=85,153. Remaining = 85,333 - 85,153 = 180 crews for others.
So the other nationalities have very small ballast contributions: 10 vessels, 15,638 tons, 180 crews.
Now, for American, the OCR shows "701.א 158" which could be 701 vessels and 158 crews? That would exceed the remaining 10 vessels. So that can't be ballast for American. Maybe "701.א" is not ballast vessels but something else. Could be the tonnage? 701 tons? And 158 crews? But then vessels missing.
Let's re-express American block with column alignment hypothesis.
The original table likely had fixed column widths. The OCR read lines sequentially, but the columns might be:
Col1: Nationality
Col2: With Cargoes Vessels
Col3: With Cargoes Tons
Col4: With Cargoes Crews
Col5: In Ballast Vessels
Col6: In Ballast Tons
Col7: In Ballast Crews
Col8: Total Vessels
Col9: Total Tons
Col10: Total Crews
The OCR output lines might correspond to rows, but with line breaks within a row.
For American, the row might be:
American, 180, 117,888, 3,713, 701, 158, ?, 130, 126,133, 3,869
But that's 10 numbers after nationality? Actually 9 numbers. Let's count: 180, 117,888, 3,713, 701, 158, ?, 130, 126,133, 3,869. That's 9 numbers if we ignore the missing one. But we have 10 columns? Wait, the header shows "With Cargoes" (3 cols), "In Ballast" (3 cols), "Total" (3 cols) = 9 data columns. So 9 numbers per nationality.
So for American, we need 9 numbers. The OCR gives: 180, 117,88, 3,713, 701, 158, 130, 126,133, 3,869. That's 8 numbers? Actually: 180, 117,88, 3,713, 701, 158, 130, 126,133, 3,869 -> 9 numbers if we consider "117,88" as one number (117,888). So 9 numbers:
Wait that's 8. Let's list:
That's 8 numbers. We need 9. There is a missing number between 158 and 130? Or maybe 701 and 158 are two numbers (vessels and tons?) but we need three for In Ballast. So In Ballast should have 3 numbers: vessels, tons, crews. We have 701 and 158. That's two. Then Total has three: 130, 126,133, 3,869. That's three. With Cargoes has three: 180, 117,888, 3,713. That's three. So total 3+2+3=8. Missing one number for In Ballast tons? Or crews? 158 could be crews, then tons missing. Or 158 could be tons, crews missing.
Given the remaining ballast totals are small (10 vessels total for all messy nationalities), 701 ballast vessels for American alone is impossible. So 701 cannot be ballast vessels. Therefore, the alignment is wrong.
Maybe the columns are not in that order. Could be the table has: Nationality, Total Vessels, Total Tons, Total Crews, With Cargoes Vessels, With Cargoes Tons, With Cargoes Crews, In Ballast Vessels, In Ballast Tons, In Ballast Crews. But the header says "With Cargoes", "In Ballast", "Total" in that order.
Let's check the header lines:
"With CaroOKS.
IN BALLAST.
TOTAL.
Versels.
Tons.
Crews Vescis.
Tons. Crews. Vessels. Tons. Crews."
This suggests the header row 1: "With Cargoes", "In Ballast", "Total". Row 2: "Vessels", "Tons", "Crews", "Vessels", "Tons", "Crews", "Vessels", "Tons", "Crews". So order is With Cargoes (V, T, C), In Ballast (V, T, C), Total (V, T, C). That's what we used.
Now, for British, the numbers appear in that order: 2,043 1,851,701 85,422 (With Cargoes), then 30 20,534 850 (In Ballast), then 2,073 1,878,325 86,272 (Total). Good.
For Chinese: 150 124,107 7,370 (With Cargoes), 5 2,000 171 (In Ballast), 155 126,287 7,547 (Total). Good.
For Chinese Junks: 16,610 1,240,088 219,701 (With Cargoes), then 7,310 409,270 83,625 23,920 1,650,258 303,386 (In Ballast and Total combined). Good.
For Dutch: 16 12,200 807 (With Cargoes), 3 1,354 58 (In Ballast), 19 13,600 865 (Total). Good.
For French: 113 165,152 9,422 (With Cargoes), 1 300 11 (In Ballast), 114 165,452 9,433 (Total). Good.
For German: 225 134,085 4,440 (With Cargoes), 21 12,156 438 (In Ballast), 240 140,841 4,887 (Total). Good.
For Honduras: 1 510 13 (With Cargoes), then 1 510 13 (Total). In Ballast missing (likely zeros).
For Japanese: 28 20,076 2,077 (With Cargoes), then 28 (maybe In Ballast vessels?), then 26,076 2,077 (Total tons and crews). In Ballast tons and crews missing.
For Norwegian/Portuguese: interleaved.
For Russian: 2,274 145 (With Cargoes tons and crews?), 10 (With Cargoes vessels?), 2,274 145 (Total). In Ballast missing.
For Siamese: 30 14,905 802 (With Cargoes), 90 (In Ballast crews?), 14,905 882 (Total tons and crews). In Ballast vessels and tons missing.
For Spanish: 53 19,038 2,005 (With Cargoes), then later 59 19,038 2,005 (Total). In Ballast missing.
For Swedish: 578 19 (With Cargoes tons and crews?), 2 578 19 (Total). In Ballast missing.
For American, Annamese, Austrian, Belgian, Bolivian, Danish: messy.
Given the clear pattern for most, the messy ones likely follow the same pattern but OCR fragmented them.
Let's try to parse each messy block by assuming they have the same three-line structure (With Cargoes, In Ballast, Total) but lines are broken.
American (lines 12-16):
Line12: "American,"
Line13: "180 117,88 3,713"
Line14: "701.א"
Line15: "158"
Line16: "130 126,133 3,869"
If we assume line13 is With Cargoes: 180 vessels, 117,888 tons, 3,713 crews.
Line14-15: In Ballast: 701 vessels? 158 tons? But 701 vessels is huge. Could it be that line14 is "701" and line15 is "158" but they are actually the In Ballast tons and crews? And the In Ballast vessels is missing? Or line
(S6)
No. 3.-NUMBER, TONNAGE, and CREWS of VESSELS of EACH NÀTION ENTERED at Ports in the Colony of Hongkong, in the Year 1880.
ENTERED.
NATIONALITY OF VESSELS.
With CaroOKS.
IN BALLAST.
TOTAL.
Versels.
Tons.
Crews Vescis.
Tons. Crews. Vessels. Tons. Crews.
American,
180 117,88 3,713
701.א
158
130 126,133 3,869
Annamese,
17
:
1
78
17
Austrian,
3.170
90
:
3,179
90
Belgian,
1,125
34
Bolivian,
740
15
:
:.
:
1
1,125
34
1
740
15
British,
2,043 1,851,701 85,422
30
20,534 850
2,073 1,878,325 86,272
Chinese,
150 124,107) 7,370
5
2,000 171
155 126,287 7,547
Chinese Junks,
16,610 1,240,088, 219,701
7,310 409,270 83,625 23,920 1,650,258 303,386
Danish,
द्म
000 5,390
1
783
24
-
6,173] 246
Dutch,
16 12.200 807
3
1,354
58
19 13.60
865
French,
113
165,152 9,422
1
300
11
114
165,452 9,433
German,
225
134,085 4,440
21
12,156
438
240
140,841 4,887
Honduras,
1
510
13
1
510
13
J
Japanese,........
28 20,076 2,077
28
26,076 2,077
Norwegian,
3,681
87
Portuguese,......
1
6321
18
:
:
:
S,681
87
1
832
18
Russian.
2,274
145
10
2,274
145
Siamese,
30 14,905 802
90
14,905 882
Spanish,
53 19,038 2,005
Swedish,
578
19
:
⠀
:
59
19,038 2,005
2
578
19
F.
TOTAL 19,421 3,724,593 336,554
7,380 461,252 85,333 20,801 4,185,845 421,887
No comments yet.
Private notes are available after approval.