1873 — Page 148

Blue Books 香港計冊 All AI Reviewed

The user wants me to proofread OCR output of a historical table about vessels entered at ports in Hong Kong. The text appears to be a statistical table showing number, tonnage, and crews of vessels entered at ports, categorized by countries whence arrived, with British and Foreign vessels, with cargoes and in ballast.

I need to reconstruct this as a proper Markdown table. The OCR is quite messy with misaligned columns, broken numbers, and special characters. Let me parse the data carefully.

First, let me understand the structure:

  • Columns: Countries whence arrived, then British (With Cargoes, In Ballast, Total), then Foreign (With Cargoes, In Ballast, Total)
  • Each sub-column has Vessels, Tons, Crews

The header row shows:

"No. 1.-NUMBer, tonnagE, and CREWS of VESSELS ENTERED at PORTS"

"COUNTRIES WHENCE ARRIVED."

"WITH CARGOES."

"BRITISH."

"IN BALLAST."

"TOTAL."

"FOREIGN."

"WITH CARGOES."

Then the data rows. Let me parse each country:

  1. Australia and New Zealand
  2. British Columbia
  3. Coast of China and Formosa
  4. Cochin China
  5. Continent of Europe
  6. Great Britain
  7. India and Singapore
  8. Japan
  9. Java and other Islands in the Indian Archipelago
  10. Macao
  11. Pellew Islands
  12. Philippine Islands
  13. Sandwich Islands
  14. Siam
  15. South America
  16. United States of America
  17. West Indies
  18. TOTAL

The OCR has many artifacts: "·", ":", "", "", "", "||", "|||", "+", "---", "⠀", "C#", etc.

I need to reconstruct the numbers. Let me go row by row.

Row 1: Australia and New Zealand

British With Cargoes: 53 vessels, 28,507 tons, 881 crews

British In Ballast: appears blank/zero? The OCR shows "·" and ":" and "**" and ":" - likely zeros or dashes

British Total: 53 vessels, 28,507 tons, 881 crews (same as with cargoes)

Foreign With Cargoes: 99 vessels, 14,835 tons, 476 crews

Foreign In Ballast: not shown? Maybe zeros

Foreign Total: not shown? The row seems incomplete.

Wait, the table has 6 main columns: Countries, British (With Cargoes, In Ballast, Total), Foreign (With Cargoes, In Ballast, Total). Each with 3 sub-columns (Vessels, Tons, Crews). So 1 + 33 + 33 = 19 columns? Actually: Countries (1) + British With Cargoes (3) + British In Ballast (3) + British Total (3) + Foreign With Cargoes (3) + Foreign In Ballast (3) + Foreign Total (3) = 19 columns.

But the OCR only shows partial data. Let me look at the header again: "WITH CARGOES. BRITISH. IN BALLAST. TOTAL. FOREIGN. WITH CARGOES." It seems the header is truncated. Probably the full header is: Countries | British With Cargoes (Vessels, Tons, Crews) | British In Ballast (Vessels, Tons, Crews) | British Total (Vessels, Tons, Crews) | Foreign With Cargoes (Vessels, Tons, Crews) | Foreign In Ballast (Vessels, Tons, Crews) | Foreign Total (Vessels, Tons, Crews)

The OCR shows "WITH CARGOES." then "BRITISH." then "IN BALLAST." then "TOTAL." then "FOREIGN." then "WITH CARGOES." - likely the Foreign In Ballast and Foreign Total headers are missing from OCR.

Now parsing each row. I'll need to infer missing values (likely zeros) where OCR shows dots or blanks.

Let me transcribe row by row:

  1. Australia and New Zealand

British With Cargoes: 53, 28,507, 881

British In Ballast: 0, 0, 0 (OCR shows "·", ":", "**", ":" - likely zeros)

British Total: 53, 28,507, 881

Foreign With Cargoes: 99, 14,835, 476

Foreign In Ballast: ? Not shown. Probably 0,0,0

Foreign Total: ? Not shown. Probably 99, 14,835, 476

But the OCR shows after Foreign With Cargoes: "99 14,835 476" then next line "British Columbin,." So likely Foreign In Ballast and Total are omitted because zero? Or maybe the table only shows Foreign With Cargoes and Foreign Total? Let's check other rows.

Row 2: British Columbia

British With Cargoes: 1, 805, 19 (OCR shows "1 805 19")

British In Ballast: 0? OCR shows ":" and "1 805) 10" - wait "805) 10" maybe 805 tons, 10 crews? But that seems like In Ballast? Actually the OCR: "1 805 19 : 1 805) 10 659 15". Let's parse: "British Columbin,." then "1 805 19" (With Cargoes), then ":" (separator), then "1 805) 10" (In Ballast? but vessels=1, tons=805, crews=10), then "659 15" (Foreign With Cargoes? vessels=659? That seems huge). Wait, "659 15" might be Foreign With Cargoes: 659 vessels? No, that can't be. Let's look at the original alignment.

The OCR text is messy. Let me copy the raw text and try to align columns manually.

Raw text lines:

"( 146 )

No. 1.-NUMBer, tonnagE, and CREWS of VESSELS ENTERED at PORTS

COUNTRIES WHENCE ARRIVED.

WITH CARGOES.

BRITISH.

IN BALLAST.

TOTAL.

FOREIGN.

WITH CARGOES.

Australia and New

Zealand,

·

Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons, Crews.

53 28,507

881

**

:.

53

28,507 881

99 14,835 476

British Columbin,.

1

805 19

:

1

805) 10

659 15

Coast of China aud

Formosa,

788

520,845 34,808 41

27,460 916

*820

548,011 35,724 17,687 1,363,296 240,229

Cochin China,......

71

60,021 2,124 1

385 14

72

00,406 2,138 00

37,632 1,602

Continent of En-

rope,

C#

1,802 61

*

3

1,802 51 42

69,457 5,210

Great Britain,..... 120 140,226|| 4,613

3

3,083 104

120

140,300 4,717

43

38,378 1,095

India und Singa

pore,...

82 110,752 0,302

82

17

110,752||| 6,332||

8

4,335 120

Japan,

02

63,505 3,158

1

845

30

01

04,140 3,194

52

39,343 4,517

Java and other

Islands in the In- dinn Archipelago,

Macao,

+

Pellew Islands,

8

3,116 101 1

1,380

45

0

4,505 230

1,076 71

311

101,470 12,508 1

80

20

312

101,556 12,618 1,467 211,054 86,957

**

***

Philippine Islands,

11

5,137 308

Sandwich Islands,

Siam,

  1. 8,451 382

South America,....

...

United States of

America,

14

20,706

770

West Indies,

:

...

*

:

:

1

176 12

11

5,137

308

62

20,800 1,634

:

1,493

40

17 8,454

382

65

28,700 2,429

1,312

28

14

20,705 770

24

74,597 2,238

---

+

1,113,689 67,370 19,4821,008,1 TOTAL,...... 1,544 1,080,435 60,241 48 33,254 1,135 1,592 1,119,089 67,376 19,4821,008,181 296,679"

This is extremely messy. I need to reconstruct the table logically. Since it's a known historical document (likely Hong Kong Blue Book or similar), the table probably has standard columns. I'll assume the columns are:

Countries British With Cargoes (Vessels, Tons, Crews) British In Ballast (Vessels, Tons, Crews) British Total (Vessels, Tons, Crews) Foreign With Cargoes (Vessels, Tons, Crews) Foreign In Ballast (Vessels, Tons, Crews) Foreign Total (Vessels, Tons, Crews)

But the OCR only shows up to Foreign With Cargoes for some rows. However, the TOTAL row at the end shows many numbers: "1,544 1,080,435 60,241 48 33,254 1,135 1,592 1,119,089 67,376 19,4821,008,181 296,679". That's 11 numbers? Let's count: 1,544; 1,080,435; 60,241; 48; 33,254; 1,135; 1,592; 1,119,089; 67,376; 19,4821,008,181; 296,679. That's 11 numbers. But we expect 18 numbers (6 categories 3). Actually 6 categories 3 = 18. But the total row might only show British and Foreign totals? Let's see: British With Cargoes total, British In Ballast total, British Total total, Foreign With Cargoes total, Foreign In Ballast total, Foreign Total total. That's 6*3=18 numbers. The OCR shows 11 numbers. Something off.

Maybe the table is simpler: Only British (With Cargoes, In Ballast, Total) and Foreign (With Cargoes, Total) - omitting Foreign In Ballast? Or maybe the table has: Countries, British (With Cargoes, In Ballast, Total), Foreign (With Cargoes, In Ballast, Total). That's 1 + 9 + 9 = 19 columns. The total row would have 18 numbers (excluding country name). The OCR total row has "1,544 1,080,435 60,241 48 33,254 1,135 1,592 1,119,089 67,376 19,4821,008,181 296,679". That's 11 numbers. Could be that some numbers are merged: "19,4821,008,181" looks like three numbers concatenated: 19,482; 1,008,181? Actually "19,4821,008,181" might be "19,482 1,008,181"? But there's a missing space. And "296,679" at end.

Let's parse the total row as it appears: "1,544 1,080,435 60,241 48 33,254 1,135 1,592 1,119,089 67,376 19,4821,008,181 296,679". If we split:

  1. 1,544
  2. 1,080,435
  3. 60,241
  4. 48
  5. 33,254
  6. 1,135
  7. 1,592
  8. 1,119,089
  9. 67,376
  10. 19,4821,008,181 (garbled)
  11. 296,679

Maybe the correct totals are:

British With Cargoes: 1,544 vessels, 1,080,435 tons, 60,241 crews

British In Ballast: 48 vessels, 33,254 tons, 1,135 crews

British Total: 1,592 vessels, 1,119,089 tons, 67,376 crews

Foreign With Cargoes: 19,482 vessels, 1,008,181 tons, 296,679 crews? But that's only 3 numbers for Foreign With Cargoes, missing Foreign In Ballast and Foreign Total.

Alternatively, maybe the table only has British and Foreign With Cargoes and Total, and In Ballast is only for British? The header says "BRITISH. IN BALLAST. TOTAL. FOREIGN. WITH CARGOES." It doesn't mention Foreign In Ballast or Foreign Total. But logically there should be.

Given the difficulty, I should reconstruct the table as best as possible from the OCR, preserving the data that is legible, and using ... for missing/illegible parts.

But the user expects a Markdown table. I need to produce a clean table. Since the OCR is so messy, I might need to infer the structure from typical Hong Kong port returns. This looks like "Number, Tonnage, and Crews of Vessels Entered at Ports" from Hong Kong Blue Book 1860s-1880s.

I'll attempt to parse each country row by row, using the numbers that appear in sequence.

Let me write a script mentally to parse. But as an AI, I'll do manually.

First, the header row in OCR: "Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons, Crews." That's 4 groups of (Vessels, Tons, Crews) = 12 columns. Plus country column = 13 columns. But we expect 19. So maybe the table only shows: British With Cargoes, British In Ballast, British Total, Foreign With Cargoes. That's 4 categories * 3 = 12 columns. That matches the header! The header shows four "Vessels Tons. Crews." groups. So the table has only four categories: British With Cargoes, British In Ballast, British Total, Foreign With Cargoes. No Foreign In Ballast, no Foreign Total. That matches the header text: "BRITISH. IN BALLAST. TOTAL. FOREIGN. WITH CARGOES." So Foreign only With Cargoes is shown. That makes sense for a partial table.

Thus columns:

  1. Countries
  2. British With Cargoes - Vessels
  3. British With Cargoes - Tons
  4. British With Cargoes - Crews
  5. British In Ballast - Vessels
  6. British In Ballast - Tons
  7. British In Ballast - Crews
  8. British Total - Vessels
  9. British Total - Tons
  10. British Total - Crews
  11. Foreign With Cargoes - Vessels
  12. Foreign With Cargoes - Tons
  13. Foreign With Cargoes - Crews

Total row should have 12 numbers (excluding country). The OCR total row has 11 numbers? Let's count again: "1,544 1,080,435 60,241 48 33,254 1,135 1,592 1,119,089 67,376 19,4821,008,181 296,679". That's 11 groups. But we need 12. The last group "19,4821,008,181 296,679" might be two numbers: 19,482 and 1,008,181 and 296,679? Actually "19,4821,008,181" could be "19,482 1,008,181" missing space. Then "296,679" is the third. So that would be three numbers for Foreign With Cargoes: Vessels=19,482, Tons=1,008,181, Crews=296,679. Then the previous numbers: 1,544; 1,080,435; 60,241 (British With Cargoes); 48; 33,254; 1,135 (British In Ballast); 1,592; 1,119,089; 67,376 (British Total). That's 9 numbers. Plus 3 for Foreign With Cargoes = 12. Perfect! So the total row is:

British With Cargoes: 1,544 vessels, 1,080,435 tons, 60,241 crews

British In Ballast: 48 vessels, 33,254 tons, 1,135 crews

British Total: 1,592 vessels, 1,119,089 tons, 67,376 crews

Foreign With Cargoes: 19,482 vessels, 1,008,181 tons, 296,679 crews

Great. Now I can parse each country row similarly: each row should have 12 numbers (3 per category). Some may be zero (shown as blanks, dots, dashes).

Let's parse each country row from the OCR text. I'll go through the raw text sequentially.

The text after header: "Australia and New Zealand, · Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons, Crews. 53 28,507 881 ** :. 53 28,507 881 99 14,835 476"

But the "Vessels Tons. Crews..." is the header repeated. Then data: "53 28,507 881" (British With Cargoes). Then "** :." likely zeros for British In Ballast (three zeros). Then "53 28,507 881" (British Total). Then "99 14,835 476" (Foreign With Cargoes). So row 1:

Australia and New Zealand: BWC: 53, 28507, 881; BIB: 0,0,0; BT: 53, 28507, 881; FWC: 99, 14835, 476.

Row 2: "British Columbin,. 1 805 19 : 1 805) 10 659 15"

"British Columbin,." (British Columbia). Then "1 805 19" (BWC). Then ":" (separator). Then "1 805) 10" - this looks like BIB: 1 vessel, 805 tons, 10 crews? But the parenthesis is artifact. Then "659 15" - only two numbers? Should be three for British Total? Actually after BIB we expect British Total (3 numbers) then Foreign With Cargoes (3 numbers). But we only see "659 15". Maybe the British Total is same as BWC because BIB is zero? But here BIB is non-zero (1,805,10). So British Total should be BWC+BIB: vessels=2, tons=1610, crews=29. But the OCR shows "659 15". That seems like Foreign With Cargoes: 659 vessels, 15? tons? 15 crews? But 659 vessels for British Columbia seems high. Let's check the next row: "Coast of China aud Formosa, 788 520,845 34,808 41 27,460 916 *820 548,011 35,724 17,687 1,363,296 240,229". That row has many numbers.

Maybe the "British Columbin,." row is actually two lines: "British Columbin,." and then "1 805 19 : 1 805) 10" and then "659 15" belongs to next row? But "Coast of China" starts after. Let's see the raw: "British Columbin,. 1 805 19 : 1 805) 10 659 15 Coast of China aud Formosa, 788 520,845 34,808 41 27,460 916 *820 548,011 35,724 17,687 1,363,296 240,229"

Perhaps "659 15" is the start of Coast of China row? But Coast of China starts with "788". Hmm.

Let's look at the pattern: Each country row seems to have the country name, then 12 numbers. The OCR splits lines arbitrarily.

I'll try to extract all numbers in order from the text after the header, ignoring country names, and then assign 12 numbers per country. There are 17 countries (including TOTAL). 17*12 = 204 numbers. But the OCR may have missing numbers.

Better approach: Since this is a known historical document, I could search my knowledge for "No. 1.-NUMBER, TONNAGE, AND CREWS OF VESSELS ENTERED AT PORTS" Hong Kong. But as an AI, I should work with given OCR.

Given the complexity, perhaps the user expects me to clean up the OCR into a readable Markdown table, preserving the data as it appears, with ... for illegible parts. But the instruction says: "Reconstruct Sentences and Paragraphs" but this is a table. For tables: "if data is clearly in table format, reconstruct the table using Markdown table syntax."

So I must produce a Markdown table. I'll do my best to parse the data into a table with the 13 columns (Country + 12 data columns). I'll use the total row as anchor to verify column meanings.

Let me list the countries in order as they appear:

  1. Australia and New Zealand
  2. British Columbia
  3. Coast of China and Formosa
  4. Cochin China
  5. Continent of Europe
  6. Great Britain
  7. India and Singapore
  8. Japan
  9. Java and other Islands in the Indian Archipelago
  10. Macao
  11. Pellew Islands
  12. Philippine Islands
  13. Sandwich Islands
  14. Siam
  15. South America
  16. United States of America
  17. West Indies
  18. TOTAL

That's 18 rows (including TOTAL). The OCR shows 17 countries + TOTAL = 18.

Now, I'll go through the OCR text and extract numbers for each country.

I'll copy the OCR text into a linear sequence, removing line breaks, but keeping country names as delimiters.

Text after header (starting from "Australia and New Zealand"):

"Australia and New Zealand, · Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons, Crews. 53 28,507 881 * :. 53 28,507 881 99 14,835 476 British Columbin,. 1 805 19 : 1 805) 10 659 15 Coast of China aud Formosa, 788 520,845 34,808 41 27,460 916 820 548,011 35,724 17,687 1,363,296 240,229 Cochin China,...... 71 60,021 2,124 1 385 14 72 00,406 2,138 00 37,632 1,602 Continent of En- rope, C# 1,802 61 3 1,802 51 42 69,457 5,210 Great Britain,..... 120 140,226|| 4,613 3 3,083 104 120 140,300 4,717 43 38,378 1,095 India und Singa pore,... 82 110,752 0,302 82 17 110,752||| 6,332|| 8 4,335 120 Japan, 02 63,505 3,158 1 845 30 01 04,140 3,194 52 39,343 4,517 Java and other Islands in the In- dinn Archipelago, Macao, + Pellew Islands, 8 3,116 101 1 1,380 45 0 4,505 230 1,076 71 311 101,470 12,508 1 80 20 312 101,556 12,618 1,467 211,054 86,957 Philippine Islands, 11 5,137 308 Sandwich Islands, Siam, 1. 8,451 382 South America,.... ... United States of America, 14 20,706 770 West Indies, : ... * : ⠀ : 1 176 12 11 5,137 308 62 20,800 1,634 : 1,493 40 17 8,454 382 65 28,700 2,429 1,312 28 14 • 20,705 770 24 74,597 2,238 --- + 1,113,689 67,370 19,4821,008,1 TOTAL,...... 1,544 1,080,435 60,241 48 33,254 1,135 1,592 1,119,089 67,376 19,4821,008,181 296,679"

This is a mess. But I can see patterns: each country name appears, then numbers. The numbers are grouped in 3s.

Let me write a parser mentally. I'll split by country names (known list). But the OCR has typos: "British Columbin,.", "Coast of China aud Formosa", "Continent of En- rope", "India und Singa pore", "In- dinn Archipelago", "Philippine Islands", "Sandwich Islands", "Siam", "South America", "United States of America", "West Indies", "TOTAL".

I'll go sequentially.

Start: "Australia and New Zealand" then numbers: 53, 28507, 881, (then ** :. likely 0,0,0), 53, 28507, 881, 99, 14835, 476. That's 12 numbers? Let's count: 53,28507,881 (3), then three zeros (3), then 53,28507,881 (3), then 99,14835,476 (3) = 12. Good.

Next country: "British Columbin,." then numbers: 1, 805, 19, (then ":" maybe separator), 1, 805, 10 (from "805) 10"), then 659, 15? But we need 12 numbers. After that, next country "Coast of China aud Formosa" appears. So maybe the numbers for British Columbia are only 8 numbers? Let's see: "1 805 19 : 1 805) 10 659 15". That's 1,805,19,1,805,10,659,15 = 7 numbers. Missing 5 numbers. Could be that British In Ballast is 1,805,10; British Total is 2,1610,29; Foreign With Cargoes is 659,15,? But 659,15 only two numbers. Maybe the third is missing. Or maybe the "659 15" belongs to Coast of China? But Coast of China starts with 788.

Look at Coast of China: "788 520,845 34,808 41 27,460 916 *820 548,011 35,724 17,687 1,363,296 240,229". That's many numbers: let's count: 788, 520845, 34808, 41, 27460, 916, 820, 548011, 35724, 17687, 1363296, 240229 = 12 numbers! Perfect. So Coast of China has 12 numbers.

Thus British Columbia should have 12 numbers before Coast of China. The text between "British Columbin,." and "Coast of China aud Formosa" is: "1 805 19 : 1 805) 10 659 15". That's only 7 numbers. Perhaps the OCR missed some numbers (like zeros for British Total and Foreign With Cargoes?). But the total row shows British Columbia likely has some entries.

Maybe the "659 15" is actually the start of Coast of China? But Coast of China starts with 788. Could be that "659 15" are the last two numbers of British Columbia (Foreign With Cargoes: 659 vessels, 15 tons? but missing crews). And the first number of Coast of China is 788. But then Coast of China would have 13 numbers? Let's check: if we take "659 15 788 520,845 34,808 41 27,460 916 *820 548,011 35,724 17,687 1,363,296 240,229" that's 14 numbers. Too many.

Alternatively, the "659 15" might be "659 15" for British Columbia's Foreign With Cargoes (vessels=659, tons=15? no). That seems unlikely.

Let's look at the original image? Not available. I need to make a best guess.

Given the total row, we can work backwards? But too many unknowns.

Perhaps the table is not uniform; some countries have only British entries, others only Foreign. But the header suggests all four categories for each country.

Another idea: The OCR might have merged two lines: "British Columbin,." might be two entries: "British Columbia" and "Coast of China"? No.

Let's search memory: This looks like the "Table No. 1" from Hong Kong Blue Book 1870s. I recall a table: "Number, Tonnage, and Crews of Vessels Entered at Ports in the Colony of Hong Kong from each Country". The countries include: Great Britain, British Colonies (Australia, British Columbia, etc.), Foreign Countries (USA, China, Japan, etc.). The table usually has columns: Countries, British (With Cargoes, In Ballast, Total), Foreign (With Cargoes, In Ballast, Total). But the header here only shows Foreign With Cargoes. Maybe the original table had two pages: one for British, one for Foreign? But the OCR shows both.

Given the difficulty, I think the best is to present the table as reconstructed from the OCR with clear markings for uncertain data. But the user wants a clean Markdown table. I'll try to produce a table with the data I can confidently parse, and for rows that are garbled, I'll put ... or note.

But the instruction: "Preserve original content — do not add or remove substantive words. Do not rephrase, rewrite, or change the author's style, tone, or word choice. Only correct what is clearly wrong." For a table, this means I should keep the numbers as they appear, but fix obvious OCR errors (like "00,406" -> "60,406"?). And align columns.

I'll create a Markdown table with the 13 columns. I'll parse each country row by extracting numbers in groups of 3 for each of the 4 categories. I'll use the country names as row labels.

Let me attempt to parse all numbers in order after each country name, assuming 12 numbers per country (except TOTAL which has 12). I'll write a quick mental script.

I'll go through the text token by token (numbers and country names). But doing manually for 18 rows is tedious but doable.

I'll list the text tokens (numbers and country names) in sequence.

First, define country names as they appear in OCR:

  1. "Australia and New Zealand"
  2. "British Columbin,."
  3. "Coast of China aud Formosa"
  4. "Cochin China,......"
  5. "Continent of En- rope"
  6. "Great Britain,....."
  7. "India und Singa pore"
  8. "Japan"
  9. "Java and other Islands in the In- dinn Archipelago"
  10. "Macao"
  11. "Pellew Islands"
  12. "Philippine Islands"
  13. "Sandwich Islands"
  14. "Siam"
  15. "South America"
  16. "United States of America"
  17. "West Indies"
  18. "TOTAL"

Now, the text between them. I'll split the raw text by these country names (approximate). But the raw text has line breaks and OCR artifacts.

Let me write the raw text as a single string and then split by country names (using the cleaned names). But I'll do manually.

I'll copy the raw text into a variable and simulate.

Raw text (from "Australia and New Zealand" to end):

"Australia and New Zealand, · Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons, Crews. 53 28,507 881 * :. 53 28,507 881 99 14,835 476 British Columbin,. 1 805 19 : 1 805) 10 659 15 Coast of China aud Formosa, 788 520,845 34,808 41 27,460 916 820 548,011 35,724 17,687 1,363,296 240,229 Cochin China,...... 71 60,021 2,124 1 385 14 72 00,406 2,138 00 37,632 1,602 Continent of En- rope, C# 1,802 61 3 1,802 51 42 69,457 5,210 Great Britain,..... 120 140,226|| 4,613 3 3,083 104 120 140,300 4,717 43 38,378 1,095 India und Singa pore,... 82 110,752 0,302 82 17 110,752||| 6,332|| 8 4,335 120 Japan, 02 63,505 3,158 1 845 30 01 04,140 3,194 52 39,343 4,517 Java and other Islands in the In- dinn Archipelago, Macao, + Pellew Islands, 8 3,116 101 1 1,380 45 0 4,505 230 1,076 71 311 101,470 12,508 1 80 20 312 101,556 12,618 1,467 211,054 86,957 Philippine Islands, 11 5,137 308 Sandwich Islands, Siam, 1. 8,451 382 South America,.... ... United States of America, 14 20,706 770 West Indies, : ... * : ⠀ : 1 176 12 11 5,137 308 62 20,800 1,634 : 1,493 40 17 8,454 382 65 28,700 2,429 1,312 28 14 • 20,705 770 24 74,597 2,238 --- + 1,113,689 67,370 19,4821,008,1 TOTAL,...... 1,544 1,080,435 60,241 48 33,254 1,135 1,592 1,119,089 67,376 19,4821,008,181 296,679"

Now, I'll split by country names. But note: "Java and other Islands in the In- dinn Archipelago, Macao, + Pellew Islands," appears as three countries in one line? Actually "Java and other Islands in the Indian Archipelago", "Macao", "Pellew Islands" are three separate countries. The OCR runs them together: "Java and other Islands in the In- dinn Archipelago, Macao, + Pellew Islands,". So they are separate rows but the numbers follow for each? The numbers after that: "8 3,116 101 1 1,380 45 0 4,505 230 1,076 71 311 101,470 12,508 1 80 20 312 101,556 12,618 1,467 211,054 86,957 *" That's many numbers. Likely each country gets 12 numbers. So three countries = 36 numbers. Let's count numbers in that block:

8, 3116, 101, 1, 1380, 45, 0, 4505, 230, 1076, 71, 311, 101470, 12508, 1, 80, 20, 312, 101556, 12618, 1467, 211054, 86957. That's 23 numbers. Not 36. Then " *" maybe artifacts. Then "Philippine Islands, 11 5,137 308" etc.

This is too messy. Perhaps the OCR has merged multiple table rows into one line.

Given the time, I think the best is to output a Markdown table that represents the structure, with the data that is clearly readable, and use ... for missing/illegible cells. But the user might expect a complete table.

Given the instruction: "You are a professional editor specializing in proofreading OCR output of historical records relating to Hong Kong." So I should use my knowledge of Hong Kong historical records to reconstruct the table accurately. I can recall or approximate the data? But I shouldn't invent data.

Maybe I can find a similar table online in my training data. The table "No. 1.-NUMBER, TONNAGE, AND CREWS OF VESSELS ENTERED AT PORTS" appears in Hong Kong Blue Books. For example, Blue Book 1875, 1880, etc. The data might match a specific year. The total row: British With Cargoes 1,544 vessels, 1,080,435 tons, 60,241 crews; British In Ballast 48 vessels, 33,254 tons, 1,135 crews; British Total 1,592 vessels, 1,119,089 tons, 67,376 crews; Foreign With Cargoes 19,482 vessels, 1,008,181 tons, 296,679 crews. This looks like a specific year. Could be 1879 or 1880.

But I cannot be sure. However, as an AI, I have access to a lot of historical data. I could reconstruct the table from known data? But the instruction says not to add or remove substantive content. So I must stick to the OCR.

Given the difficulty, I'll produce a Markdown table with the columns as identified, and fill in the numbers as they appear in the OCR for each country, using the sequence of numbers. I'll do my best to align.

Let me attempt to parse each country's numbers by taking the numbers that appear after the country name until the next country name. I'll write a quick parser in my mind.

I'll go through the raw text and extract numbers (including those with commas, pipes, etc.) and country names.

I'll tokenize by spaces and newlines, but keep numbers with commas as single tokens.

Let me list tokens sequentially:

Tokens (split by whitespace, but keep punctuation attached? Better to split by whitespace and then cleanly).

I'll copy the raw text into a list of words.

But perhaps I can use the fact that the total row at the end has 12 numbers (after splitting the garbled one). And the total row matches the sum of each column? Not necessarily, but likely.

Given the time constraints, I'll produce a table with the header and the rows that are clearly parseable: Australia and New Zealand, Coast of China and Formosa, Cochin China, Continent of Europe, Great Britain, India and Singapore, Japan, and TOTAL. For the others, I'll put ... or partial.

But the user might want the full table. I'll try to parse all.

Let me do a systematic extraction.

I'll write a pseudo-code:

text = "Australia and New Zealand, · Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons, Crews. 53 28,507 881 * :. 53 28,507 881 99 14,835 476 British Columbin,. 1 805 19 : 1 805) 10 659 15 Coast of China aud Formosa, 788 520,845 34,808 41 27,460 916 820 548,011 35,724 17,687 1,363,296 240,229 Cochin China,...... 71 60,021 2,124 1 385 14 72 00,406 2,138 00 37,632 1,602 Continent of En- rope, C# 1,802 61 3 1,802 51 42 69,457 5,210 Great Britain,..... 120 140,226|| 4,613 3 3,083 104 120 140,300 4,717 43 38,378 1,095 India und Singa pore,... 82 110,752 0,302 82 17 110,752||| 6,332|| 8 4,335 120 Japan, 02 63,505 3,158 1 845 30 01 04,140 3,194 52 39,343 4,517 Java and other Islands in the In- dinn Archipelago, Macao, + Pellew Islands, 8 3,116 101 1 1,380 45 0 4,505 230 1,076 71 311 101,470 12,508 1 80 20 312 101,556 12,618 1,467 211,054 86,957 Philippine Islands, 11 5,137 308 Sandwich Islands, Siam, 1. 8,451 382 South America,.... ... United States of America, 14 20,706 770 West Indies, : ... * : ⠀ : 1 176 12 11 5,137 308 62 20,800 1,634 : 1,493 40 17 8,454 382 65 28,700 2,429 1,312 28 14 • 20,705 770 24 74,597 2,238 --- + 1,113,689 67,370 19,4821,008,1 TOTAL,...... 1,544 1,080,435 60,241 48 33,254 1,135 1,592 1,119,089 67,376 19,4821,008,181 296,679"

I'll split by country names (using a list). But the country names in text have variations. I'll use the following split points (by the start of each country name as it appears):

  1. "Australia and New Zealand"
  2. "British Columbin"
  3. "Coast of China aud Formosa"
  4. "Cochin China"
  5. "Continent of En- rope"
  6. "Great Britain"
  7. "India und Singa pore"
  8. "Japan"
  9. "Java and other Islands in the In- dinn Archipelago"
  10. "Macao"
  11. "Pellew Islands"
  12. "Philippine Islands"
  13. "Sandwich Islands"
  14. "Siam"
  15. "South America"
  16. "United States of America"
  17. "West Indies"
  18. "TOTAL"

Now, I'll find the indices of these in the text. But doing manually.

Let me write the text with markers.

I'll go through the text and note where each country appears.

Start: "Australia and New Zealand, · Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons, Crews. 53 28,507 881 ** :. 53 28,507 881 99 14,835 476 British Columbin,."

So after "476" comes "British Columbin,.".

Next: "British Columbin,. 1 805 19 : 1 805) 10 659 15 Coast of China aud Formosa,"

So after "15" comes "Coast of China aud Formosa,".

Next: "Coast of China aud Formosa, 788 520,845 34,808 41 27,460 916 *820 548,011 35,724 17,687 1,363,296 240,229 Cochin China,......"

After "240,229" comes "Cochin China,......".

Next: "Cochin China,...... 71 60,021 2,124 1 385 14 72 00,406 2,138 00 37,632 1,602 Continent of En- rope,"

After "1,602" comes "Continent of En- rope,".

Next: "Continent of En- rope, C# 1,802 61 * 3 1,802 51 42 69,457 5,210 Great Britain,....."

After "5,210" comes "Great Britain,.....".

Next: "Great Britain,..... 120 140,226|| 4,613 3 3,083 104 120 140,300 4,717 43 38,378 1,095 India und Singa pore,..."

After "1,095" comes "India und Singa pore,...".

Next: "India und Singa pore,... 82 110,752 0,302 82 17 110,752||| 6,332|| 8 4,335 120 Japan,"

After "120" comes "Japan,".

Next: "Japan, 02 63,505 3,158 1 845 30 01 04,140 3,194 52 39,343 4,517 Java and other Islands in the In- dinn Archipelago,"

After "4,517" comes "Java and other Islands in the In- dinn Archipelago,".

Next: "Java and other Islands in the In- dinn Archipelago, Macao, + Pellew Islands, 8 3,116 101 1 1,380 45 0 4,505 230 1,076 71 311 101,470 12,508 1 80 20 312 101,556 12,618 1,467 211,054 86,957 * Philippine Islands,"

After "86,957 *" comes "Philippine Islands,".

Next: "Philippine Islands, 11 5,137 308 Sandwich Islands,"

After "308" comes "Sandwich Islands,".

Next: "Sandwich Islands, Siam, 1. 8,451 382 South America,...."

After "382" comes "South America,....".

Next: "South America,.... ... United States of America, 14 20,706 770 West Indies,"

After "770" comes "West Indies,".

Next: "West Indies, : ... * : ⠀ : 1 176 12 11 5,137 308 62 20,800 1,634 : 1,493 40 17 8,454 382 65 28,700 2,429 1,312 28 14 • 20,705 770 24 74,597 2,238 --- + 1,113,689 67,370 19,4821,008,1 TOTAL,......"

After "19,4821,008,1" comes "TOTAL,......".

Next: "TOTAL,...... 1,544 1,080,435 60,241 48 33,254 1,135 1,592 1,119,089 67,376 19,4821,008,181 296,679"

End.

Now, for each country, the numbers between its name and the next country name are its data. But note that the header "Vessels Tons. Crews..." appears only at the start. Also, there are artifacts like "·", "*", ":.", "C#", "", "||", "|||", "+", "---", " *", "•", etc. These are not numbers. I need to extract only numeric tokens (digits with commas, possibly with pipes). Also, some numbers are split like "02" for Japan (maybe 2? but likely 2 vessels? Actually "02" could be 2). "1." for Siam (maybe 1). "8" for Java? etc.

Let's extract numeric tokens for each country.

Define a function to extract numbers from a string: tokens that match regex ^\d[\d,.]*$ or with pipes? But pipes are artifacts. I'll clean: remove any non-digit, non-comma, non-period? But tons have commas. So keep digits and commas. Also, some have trailing pipes like "140,226||" -> "140,226". "110,752|||" -> "110,752". "6,332||" -> "6,332". "00,406" -> "00,406" (maybe 60,406?). "19,4821,008,181" is garbled.

I'll process each country's segment.

I'll write the segment strings:

  1. Australia and New Zealand segment: "· Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons, Crews. 53 28,507 881 ** :. 53 28,507 881 99 14,835 476"

Numbers: 53, 28507, 881, 53, 28507, 881, 99, 14835, 476. That's 9 numbers. But we need 12. The missing three are for British In Ballast (likely 0,0,0). The "** :." might represent zeros. So we can assume British In Ballast: 0,0,0. So the 12 numbers: BWC: 53,28507,881; BIB: 0,0,0; BT: 53,28507,881; FWC: 99,14835,476.

  1. British Columbia segment: "1 805 19 : 1 805) 10 659 15"

Numbers: 1, 805, 19, 1, 805, 10, 659, 15. That's 8 numbers. Need 12. Missing 4. Could be that British In Ballast: 1,805,10; British Total: 2,1610,29; Foreign With Cargoes: 659,15,? (missing crews). Or maybe the segment includes only up to Foreign With Cargoes vessels and tons, missing crews. The next segment starts with 788 (Coast of China). So maybe the 659,15 are actually the first two numbers of Coast of China? But Coast of China segment starts with 788. So 659,15 belong to British Columbia. Let's assume British Columbia has: BWC: 1,805,19; BIB: 1,805,10; BT: 2,1610,29; FWC: 659,15,? (missing). But we have only 8 numbers. Could be that BT and FWC are combined? Not sure.

Given the total row, we can check if British Columbia appears in totals. But not now.

I'll note the numbers as they appear and fill missing with ....

  1. Coast of China and Formosa segment: "788 520,845 34,808 41 27,460 916 *820 548,011 35,724 17,687 1,363,296 240,229"

Numbers: 788, 520845, 34808, 41, 27460, 916, 820, 548011, 35724, 17687, 1363296, 240229. That's 12 numbers! Perfect.

So:

BWC: 788, 520845, 34808

BIB: 41, 27460, 916

BT: 820, 548011, 35724

FWC: 17687, 1363296, 240229

  1. Cochin China segment: "71 60,021 2,124 1 385 14 72 00,406 2,138 00 37,632 1,602"

Numbers: 71, 60021, 2124, 1, 385, 14, 72, 00406, 2138, 00, 37632, 1602. That's 12 numbers.

But "00,406" likely 60,406? "00" maybe 0? "00" as tons? Let's keep as is: 00,406 -> 00,406 (maybe 60,406). "00" -> 0.

So:

BWC: 71, 60021, 2124

BIB: 1, 385, 14

BT: 72, 00406, 2138

FWC: 00, 37632, 1602? Wait, 00 is vessels? Then 37632 tons, 1602 crews. But FWC should be three numbers: vessels, tons, crews. So 00, 37632, 1602. That works: 0 vessels? But 37632 tons with 0 vessels? That seems odd. Maybe "00" is actually 60? But it's "00". Could be "60" misread. But we'll keep as 00.

  1. Continent of Europe segment: "C# 1,802 61 * 3 1,802 51 42 69,457 5,210"

Numbers: 1802, 61, 3, 1802, 51, 42, 69457, 5210. That's 8 numbers. Need 12. Missing 4. "C#" and "" are artifacts. So we have 8 numbers. Possibly BWC: 1802,61,? missing tons? Actually 1802 might be vessels, 61 tons? But tons should be larger. 1,802 could be tons? 61 crews? Then 3 maybe 3 vessels? Let's see: "1,802 61" then "* 3" then "1,802 51" then "42 69,457 5,210". Hard to parse.

Maybe the segment is: "C# 1,802 61 * 3 1,802 51 42 69,457 5,210". Numbers: 1802, 61, 3, 1802, 51, 42, 69457, 5210. Could be: BWC: 1802 vessels, 61 tons? No. 1,802 tons, 61 crews? Then BIB: 3 vessels, 1,802 tons, 51 crews? Then BT: 42 vessels, 69,457 tons, 5,210 crews? Then FWC missing. But we have only 8 numbers.

Given the difficulty, I'll keep the numbers as they appear in groups of 3 as much as possible.

  1. Great Britain segment: "120 140,226|| 4,613 3 3,083 104 120 140,300 4,717 43 38,378 1,095"

Numbers: 120, 140226, 4613, 3, 3083, 104, 120, 140300, 4717, 43, 38378, 1095. That's 12 numbers! Good.

Clean: 140,226|| -> 140226; 140,300 -> 140300.

So:

BWC: 120, 140226, 4613

BIB: 3, 3083, 104

BT: 120, 140300, 4717

FWC: 43, 38378, 1095

  1. India and Singapore segment: "82 110,752 0,302 82 17 110,752||| 6,332|| 8 4,335 120"

Numbers: 82, 110752, 0302, 82, 17, 110752, 6332, 8, 4335, 120. That's 10 numbers. Need 12. "0,302" -> 0,302? Maybe 6,302? "110,752|||" -> 110752. "6,332||" -> 6332.

We have 10 numbers. Missing 2. Could be BWC: 82, 110752, 0302 (3); BIB: 82, 17, 110752? That doesn't make sense. Let's group: 82, 110752, 0302 (3); 82, 17, 110752 (3); 6332, 8, 4335 (3); 120 (1). Not good.

Maybe the segment includes "82 110,752 0,302" (BWC), "82 17 110,752|||" (BIB? but 82 vessels again?), "6,332|| 8 4,335 120" (BT and FWC?). Hard.

  1. Japan segment: "02 63,505 3,158 1 845 30 01 04,140 3,194 52 39,343 4,517"

Numbers: 02, 63505, 3158, 1, 845, 30, 01, 04140, 3194, 52, 39343, 4517. That's 12 numbers! Good.

Clean: 02 -> 2; 01 -> 1; 04,140 -> 4140.

So:

BWC: 2, 63505, 3158

BIB: 1, 845, 30

BT: 1, 4140, 3194

FWC: 52, 39343, 4517

  1. Java and other Islands in the Indian Archipelago segment: This is combined with Macao and Pellew Islands. The segment: "Macao, + Pellew Islands, 8 3,116 101 1 1,380 45 0 4,505 230 1,076 71 311 101,470 12,508 1 80 20 312 101,556 12,618 1,467 211,054 86,957 *"

But the country name "Java and other Islands in the In- dinn Archipelago" appears before "Macao, + Pellew Islands,". So the segment for Java is from after its name to "Macao,". But the text: "Java and other Islands in the In- dinn Archipelago, Macao, + Pellew Islands, 8 3,116 101 1 1,380 45 0 4,505 230 1,076 71 311 101,470 12,508 1 80 20 312 101,556 12,618 1,467 211,054 86,957 *"

So the numbers after "Archipelago," are for Java? But then "Macao," and "Pellew Islands," are separate countries, but they have no numbers before the numbers start? Actually the numbers start after "Pellew Islands,". So likely the three countries share the same numbers? No, each should have its own row. The OCR has merged three rows into one line. The numbers that follow are for Java, Macao, Pellew Islands sequentially? There are 23 numbers. 23/3 = 7.66, not 12 each. 23 numbers for three countries = not divisible by 12. Maybe the numbers are for Java only? But then Macao and Pellew Islands have no data? Or the numbers are for all three combined? Unlikely.

Let's count numbers in that block: "8 3,116 101 1 1,380 45 0 4,505 230 1,076 71 311 101,470 12,508 1 80 20 312 101,556 12,618 1,467 211,054 86,957"

List: 8, 3116, 101, 1, 1380, 45, 0, 4505, 230, 1076, 71, 311, 101470, 12508, 1, 80, 20, 312, 101556, 12618, 1467, 211054, 86957. That's 23 numbers.

If each country has 12 numbers, three countries need 36. So not.

Maybe the table for these three countries is different? Or the OCR has lost many numbers.

Given the complexity, I'll treat "Java and other Islands in the Indian Archipelago", "Macao", "Pellew Islands" as three rows but with missing data. I'll put the numbers that appear in the first row (Java) and leave others blank? But the numbers might be for Java only? Let's see the total row: the totals include all countries. Hard.

  1. Philippine Islands segment: "11 5,137 308" only 3 numbers. Then next country "Sandwich Islands". So Philippine Islands only has 3 numbers? Maybe only BWC? But need 12.
  1. Sandwich Islands segment: "Siam, 1. 8,451 382" Actually the segment: "Sandwich Islands, Siam, 1. 8,451 382 South America,...." So Sandwich Islands has no numbers? The numbers "1. 8,451 382" belong to Siam? Because "Siam, 1. 8,451 382". So Sandwich Islands might have zero entries.
  1. Siam segment: "1. 8,451 382" -> numbers: 1, 8451, 382. Only 3 numbers.
  1. South America segment: "..." no numbers? Then "United States of America, 14 20,706 770" So South America maybe zero.
  1. United States of America segment: "14 20,706 770" -> 3 numbers.
  1. West Indies segment: ": ... * : ⠀ : 1 176 12 11 5,137 308 62 20,800 1,634 : 1,493 40 17 8,454 382 65 28,700 2,429 1,312 28 14 • 20,705 770 24 74,597 2,238 --- + 1,113,689 67,370 19,4821,008,1"

Numbers: 1, 176, 12, 11, 5137, 308, 62, 20800, 1634, 1493, 40, 17, 8454, 382, 65, 28700, 2429, 1312, 28, 14, 20705, 770, 24, 74597, 2238, 1113689, 67370, 1948210081? That's many numbers. But this segment ends at "19,4821,008,1" which is the start of TOTAL? Actually the TOTAL segment starts after that. So West Indies segment includes numbers up to "19,4821,008,1". But that number is likely the Foreign With Cargoes total for all countries? But it's in West Indies segment. This is messy.

Given the extreme OCR corruption for the latter countries, I think the best is to produce a table with the rows that are clearly parseable (1-8) and for the rest, indicate missing data. But the user might want the full table as per OCR.

Given the instruction: "Return the proofread text in standard Markdown". I'll produce a Markdown table with the header and all 18 rows, filling in the numbers as I can parse them, using ... for missing or illegible cells. I'll also clean obvious OCR errors (like "00,406" -> "60,406"? but I'll keep as "00,406" with note? Better to correct unambiguous errors: "00,406" likely "60,406" because Cochin China BT tons should be around 60k. But I'll keep original? The rule: "Correct unambiguous OCR spelling errors". "00,406" is ambiguous. "19,4821,008,181" is clearly garbled; I'll split as "19,482 1,008,181" based on total row.

I'll make a decision: I'll reconstruct the table with the 13 columns. For each country, I'll list the 12 data points in order: BWC_V, BWC_T, BWC_C, BIB_V, BIB_T, BIB_C, BT_V, BT_T, BT_C, FWC_V, FWC_T, FWC_C. I'll extract from the segments as groups of 3. If a segment has 12 numbers, perfect. If less, I'll pad with ... at the end. If more, I'll truncate? But unlikely.

Let's do it systematically for each country using the segments defined.

I'll create a list of countries and their raw segment strings (the text between country name and next country name). Then extract numbers (cleaned) from each segment. Then group into 3s. If not enough, fill with "...".

First, define cleaned numbers: remove any non-digit, non-comma characters from each token, but keep commas. Tokens separated by whitespace. Also, tokens like "140,226||" become "140,226". "0,302" stays. "00,406" stays. "19,4821,008,181" is one token? In the TOTAL segment, it's "19,4821,008,181" as one token. I'll split that later.

Let's process each segment.

I'll write a quick mental script.

Segment 1 (Australia): "· Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons, Crews. 53 28,507 881 ** :. 53 28,507 881 99 14,835 476"

Tokens: "·", "Vessels", "Tons.", "Crews.", "Vessels", "Tons.", "Crews.", "Vessels", "Tons.", "Crews.", "Vessels", "Tons,", "Crews.", "53", "28,507", "881", "**", ":.", "53", "28,507", "881", "99", "14,835", "476"

Numbers: "53", "28,507", "881", "53", "28,507", "881", "99", "14,835", "476" -> 9 numbers.

Grouped: [53,28507,881], [53,28507,881], [99,14835,476] -> missing BIB group. So insert [0,0,0] as second group? But the original has "** :." which might indicate zeros. I'll assume BIB = 0,0,0. So final 12: 53,28507,881, 0,0,0, 53,28507,881, 99,14835,476.

Segment 2 (British Columbia): "1 805 19 : 1 805) 10 659 15"

Tokens: "1", "805", "19", ":", "1", "805)", "10", "659", "15"

Numbers: "1", "805", "19", "1", "805", "10", "659", "15" -> 8 numbers.

Grouped: [1,805,19], [1,805,10], [659,15] -> third group incomplete. Missing one number for FWC crews. Also missing BT group? Actually we have 8 numbers = 2 full groups (6) + 2. So maybe groups: BWC, BIB, then BT_V, BT_T? But BT needs 3. FWC needs 3. So not sure.

Given the pattern, perhaps the segment is missing BT and FWC crews. I'll keep as: BWC: 1,805,19; BIB: 1,805,10; BT: 659,15,...? No.

Let's look at the total row: British Columbia likely small. I'll just put the numbers as they appear in order across the 12 columns, padding with ... at the end. So 8 numbers fill first 8 columns: BWC_V=1, BWC_T=805, BWC_C=19, BIB_V=1, BIB_T=805, BIB_C=10, BT_V=659, BT_T=15, BT_C=..., FWC_V=..., FWC_T=..., FWC_C=... But that seems wrong because BT_V=659 is huge.

Alternatively, maybe the numbers are: BWC: 1,805,19; BIB: 0,0,0; BT: 1,805,19; FWC: 659,15,? But we have 1,805,10 extra.

The "1 805) 10" might be BT: 1,805,10? But BWC is 1,805,19. Then FWC: 659,15. That would be 3+3+3=9 numbers? Actually BWC(3), BT(3), FWC(3) =9, but we have 8. Missing one.

If BIB is omitted (all zero), then BWC, BT, FWC = 9 numbers. We have 8. So maybe FWC crews missing.

I'll assume the table for British Columbia has BWC, BIB, BT, FWC. The OCR gave BWC, BIB, and then FWC (two numbers). So I'll set BT = BWC+BIB? But not sure.

Given the ambiguity, I'll output the raw numbers in a single row with 12 columns, but for British Columbia, I'll put the 8 numbers in the first 8 columns and ... for rest. But that would misalign columns.

Better: Since the table has fixed columns, I must align each number to its column. I need to know which column each number belongs to. The header gives column order. The total row gives totals for each column. I could use the total row to deduce column meanings, but not per country.

Given the time, I'll produce a table with the data for the countries that have clear 12-number rows: Australia, Coast of China, Cochin China, Great Britain, Japan. For others, I'll put "

Edit History

2026-07-19 08:11:44 · NVIDIA / nvidia/nemotron-3-ultra-550b-a55b
Live
View comparison
AI Proofread
The user wants me to proofread OCR output of a historical table about vessels entered at ports in Hong Kong. The text appears to be a statistical table showing number, tonnage, and crews of vessels entered at ports, categorized by countries whence arrived, with British and Foreign vessels, with cargoes and in ballast. I need to reconstruct this as a proper Markdown table. The OCR is quite messy with misaligned columns, broken numbers, and special characters. Let me parse the data carefully. First, let me understand the structure: Columns: Countries whence arrived, then British (With Cargoes, In Ballast, Total), then Foreign (With Cargoes, In Ballast, Total) Each sub-column has Vessels, Tons, Crews The header row shows: "No. 1.-NUMBer, tonnagE, and CREWS of VESSELS ENTERED at PORTS" "COUNTRIES WHENCE ARRIVED." "WITH CARGOES." "BRITISH." "IN BALLAST." "TOTAL." "FOREIGN." "WITH CARGOES." Then the data rows. Let me parse each country: Australia and New Zealand British Columbia Coast of China and Formosa Cochin China Continent of Europe Great Britain India and Singapore Japan Java and other Islands in the Indian Archipelago Macao Pellew Islands Philippine Islands Sandwich Islands Siam South America United States of America West Indies TOTAL The OCR has many artifacts: "·", ":", "", "", "", "||", "|||", "+", "---", "⠀", "C#", etc. I need to reconstruct the numbers. Let me go row by row. Row 1: Australia and New Zealand British With Cargoes: 53 vessels, 28,507 tons, 881 crews British In Ballast: appears blank/zero? The OCR shows "·" and ":" and "**" and ":" - likely zeros or dashes British Total: 53 vessels, 28,507 tons, 881 crews (same as with cargoes) Foreign With Cargoes: 99 vessels, 14,835 tons, 476 crews Foreign In Ballast: not shown? Maybe zeros Foreign Total: not shown? The row seems incomplete. Wait, the table has 6 main columns: Countries, British (With Cargoes, In Ballast, Total), Foreign (With Cargoes, In Ballast, Total). Each with 3 sub-columns (Vessels, Tons, Crews). So 1 + 33 + 33 = 19 columns? Actually: Countries (1) + British With Cargoes (3) + British In Ballast (3) + British Total (3) + Foreign With Cargoes (3) + Foreign In Ballast (3) + Foreign Total (3) = 19 columns. But the OCR only shows partial data. Let me look at the header again: "WITH CARGOES. BRITISH. IN BALLAST. TOTAL. FOREIGN. WITH CARGOES." It seems the header is truncated. Probably the full header is: Countries | British With Cargoes (Vessels, Tons, Crews) | British In Ballast (Vessels, Tons, Crews) | British Total (Vessels, Tons, Crews) | Foreign With Cargoes (Vessels, Tons, Crews) | Foreign In Ballast (Vessels, Tons, Crews) | Foreign Total (Vessels, Tons, Crews) The OCR shows "WITH CARGOES." then "BRITISH." then "IN BALLAST." then "TOTAL." then "FOREIGN." then "WITH CARGOES." - likely the Foreign In Ballast and Foreign Total headers are missing from OCR. Now parsing each row. I'll need to infer missing values (likely zeros) where OCR shows dots or blanks. Let me transcribe row by row: Australia and New Zealand British With Cargoes: 53, 28,507, 881 British In Ballast: 0, 0, 0 (OCR shows "·", ":", "**", ":" - likely zeros) British Total: 53, 28,507, 881 Foreign With Cargoes: 99, 14,835, 476 Foreign In Ballast: ? Not shown. Probably 0,0,0 Foreign Total: ? Not shown. Probably 99, 14,835, 476 But the OCR shows after Foreign With Cargoes: "99 14,835 476" then next line "British Columbin,." So likely Foreign In Ballast and Total are omitted because zero? Or maybe the table only shows Foreign With Cargoes and Foreign Total? Let's check other rows. Row 2: British Columbia British With Cargoes: 1, 805, 19 (OCR shows "1 805 19") British In Ballast: 0? OCR shows ":" and "1 805) 10" - wait "805) 10" maybe 805 tons, 10 crews? But that seems like In Ballast? Actually the OCR: "1 805 19 : 1 805) 10 659 15". Let's parse: "British Columbin,." then "1 805 19" (With Cargoes), then ":" (separator), then "1 805) 10" (In Ballast? but vessels=1, tons=805, crews=10), then "659 15" (Foreign With Cargoes? vessels=659? That seems huge). Wait, "659 15" might be Foreign With Cargoes: 659 vessels? No, that can't be. Let's look at the original alignment. The OCR text is messy. Let me copy the raw text and try to align columns manually. Raw text lines: "( 146 ) No. 1.-NUMBer, tonnagE, and CREWS of VESSELS ENTERED at PORTS COUNTRIES WHENCE ARRIVED. WITH CARGOES. BRITISH. IN BALLAST. TOTAL. FOREIGN. WITH CARGOES. Australia and New Zealand, · Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons, Crews. 53 28,507 881 ** :. 53 28,507 881 99 14,835 476 British Columbin,. 1 805 19 : 1 805) 10 659 15 Coast of China aud Formosa, 788 520,845 34,808 41 27,460 916 *820 548,011 35,724 17,687 1,363,296 240,229 Cochin China,...... 71 60,021 2,124 1 385 14 72 00,406 2,138 00 37,632 1,602 Continent of En- rope, C# 1,802 61 * 3 1,802 51 42 69,457 5,210 Great Britain,..... 120 140,226|| 4,613 3 3,083 104 120 140,300 4,717 43 38,378 1,095 India und Singa pore,... 82 110,752 0,302 82 17 110,752||| 6,332|| 8 4,335 120 Japan, 02 63,505 3,158 1 845 30 01 04,140 3,194 52 39,343 4,517 Java and other Islands in the In- dinn Archipelago, Macao, + Pellew Islands, 8 3,116 101 1 1,380 45 0 4,505 230 1,076 71 311 101,470 12,508 1 80 20 312 101,556 12,618 1,467 211,054 86,957 ** *** Philippine Islands, 11 5,137 308 Sandwich Islands, Siam, 8,451 382 South America,.... ... United States of America, 14 20,706 770 West Indies, : ... * : ⠀ : 1 176 12 11 5,137 308 62 20,800 1,634 : 1,493 40 17 8,454 382 65 28,700 2,429 1,312 28 14 • 20,705 770 24 74,597 2,238 --- + 1,113,689 67,370 19,4821,008,1 TOTAL,...... 1,544 1,080,435 60,241 48 33,254 1,135 1,592 1,119,089 67,376 19,4821,008,181 296,679" This is extremely messy. I need to reconstruct the table logically. Since it's a known historical document (likely Hong Kong Blue Book or similar), the table probably has standard columns. I'll assume the columns are: Countries British With Cargoes (Vessels, Tons, Crews) British In Ballast (Vessels, Tons, Crews) British Total (Vessels, Tons, Crews) Foreign With Cargoes (Vessels, Tons, Crews) Foreign In Ballast (Vessels, Tons, Crews) Foreign Total (Vessels, Tons, Crews) But the OCR only shows up to Foreign With Cargoes for some rows. However, the TOTAL row at the end shows many numbers: "1,544 1,080,435 60,241 48 33,254 1,135 1,592 1,119,089 67,376 19,4821,008,181 296,679". That's 11 numbers? Let's count: 1,544; 1,080,435; 60,241; 48; 33,254; 1,135; 1,592; 1,119,089; 67,376; 19,4821,008,181; 296,679. That's 11 numbers. But we expect 18 numbers (6 categories 3). Actually 6 categories 3 = 18. But the total row might only show British and Foreign totals? Let's see: British With Cargoes total, British In Ballast total, British Total total, Foreign With Cargoes total, Foreign In Ballast total, Foreign Total total. That's 6*3=18 numbers. The OCR shows 11 numbers. Something off. Maybe the table is simpler: Only British (With Cargoes, In Ballast, Total) and Foreign (With Cargoes, Total) - omitting Foreign In Ballast? Or maybe the table has: Countries, British (With Cargoes, In Ballast, Total), Foreign (With Cargoes, In Ballast, Total). That's 1 + 9 + 9 = 19 columns. The total row would have 18 numbers (excluding country name). The OCR total row has "1,544 1,080,435 60,241 48 33,254 1,135 1,592 1,119,089 67,376 19,4821,008,181 296,679". That's 11 numbers. Could be that some numbers are merged: "19,4821,008,181" looks like three numbers concatenated: 19,482; 1,008,181? Actually "19,4821,008,181" might be "19,482 1,008,181"? But there's a missing space. And "296,679" at end. Let's parse the total row as it appears: "1,544 1,080,435 60,241 48 33,254 1,135 1,592 1,119,089 67,376 19,4821,008,181 296,679". If we split: 1,544 1,080,435 60,241 48 33,254 1,135 1,592 1,119,089 67,376 19,4821,008,181 (garbled) 296,679 Maybe the correct totals are: British With Cargoes: 1,544 vessels, 1,080,435 tons, 60,241 crews British In Ballast: 48 vessels, 33,254 tons, 1,135 crews British Total: 1,592 vessels, 1,119,089 tons, 67,376 crews Foreign With Cargoes: 19,482 vessels, 1,008,181 tons, 296,679 crews? But that's only 3 numbers for Foreign With Cargoes, missing Foreign In Ballast and Foreign Total. Alternatively, maybe the table only has British and Foreign With Cargoes and Total, and In Ballast is only for British? The header says "BRITISH. IN BALLAST. TOTAL. FOREIGN. WITH CARGOES." It doesn't mention Foreign In Ballast or Foreign Total. But logically there should be. Given the difficulty, I should reconstruct the table as best as possible from the OCR, preserving the data that is legible, and using ... for missing/illegible parts. But the user expects a Markdown table. I need to produce a clean table. Since the OCR is so messy, I might need to infer the structure from typical Hong Kong port returns. This looks like "Number, Tonnage, and Crews of Vessels Entered at Ports" from Hong Kong Blue Book 1860s-1880s. I'll attempt to parse each country row by row, using the numbers that appear in sequence. Let me write a script mentally to parse. But as an AI, I'll do manually. First, the header row in OCR: "Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons, Crews." That's 4 groups of (Vessels, Tons, Crews) = 12 columns. Plus country column = 13 columns. But we expect 19. So maybe the table only shows: British With Cargoes, British In Ballast, British Total, Foreign With Cargoes. That's 4 categories * 3 = 12 columns. That matches the header! The header shows four "Vessels Tons. Crews." groups. So the table has only four categories: British With Cargoes, British In Ballast, British Total, Foreign With Cargoes. No Foreign In Ballast, no Foreign Total. That matches the header text: "BRITISH. IN BALLAST. TOTAL. FOREIGN. WITH CARGOES." So Foreign only With Cargoes is shown. That makes sense for a partial table. Thus columns: Countries British With Cargoes - Vessels British With Cargoes - Tons British With Cargoes - Crews British In Ballast - Vessels British In Ballast - Tons British In Ballast - Crews British Total - Vessels British Total - Tons British Total - Crews Foreign With Cargoes - Vessels Foreign With Cargoes - Tons Foreign With Cargoes - Crews Total row should have 12 numbers (excluding country). The OCR total row has 11 numbers? Let's count again: "1,544 1,080,435 60,241 48 33,254 1,135 1,592 1,119,089 67,376 19,4821,008,181 296,679". That's 11 groups. But we need 12. The last group "19,4821,008,181 296,679" might be two numbers: 19,482 and 1,008,181 and 296,679? Actually "19,4821,008,181" could be "19,482 1,008,181" missing space. Then "296,679" is the third. So that would be three numbers for Foreign With Cargoes: Vessels=19,482, Tons=1,008,181, Crews=296,679. Then the previous numbers: 1,544; 1,080,435; 60,241 (British With Cargoes); 48; 33,254; 1,135 (British In Ballast); 1,592; 1,119,089; 67,376 (British Total). That's 9 numbers. Plus 3 for Foreign With Cargoes = 12. Perfect! So the total row is: British With Cargoes: 1,544 vessels, 1,080,435 tons, 60,241 crews British In Ballast: 48 vessels, 33,254 tons, 1,135 crews British Total: 1,592 vessels, 1,119,089 tons, 67,376 crews Foreign With Cargoes: 19,482 vessels, 1,008,181 tons, 296,679 crews Great. Now I can parse each country row similarly: each row should have 12 numbers (3 per category). Some may be zero (shown as blanks, dots, dashes). Let's parse each country row from the OCR text. I'll go through the raw text sequentially. The text after header: "Australia and New Zealand, · Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons, Crews. 53 28,507 881 ** :. 53 28,507 881 99 14,835 476" But the "Vessels Tons. Crews..." is the header repeated. Then data: "53 28,507 881" (British With Cargoes). Then "** :." likely zeros for British In Ballast (three zeros). Then "53 28,507 881" (British Total). Then "99 14,835 476" (Foreign With Cargoes). So row 1: Australia and New Zealand: BWC: 53, 28507, 881; BIB: 0,0,0; BT: 53, 28507, 881; FWC: 99, 14835, 476. Row 2: "British Columbin,. 1 805 19 : 1 805) 10 659 15" "British Columbin,." (British Columbia). Then "1 805 19" (BWC). Then ":" (separator). Then "1 805) 10" - this looks like BIB: 1 vessel, 805 tons, 10 crews? But the parenthesis is artifact. Then "659 15" - only two numbers? Should be three for British Total? Actually after BIB we expect British Total (3 numbers) then Foreign With Cargoes (3 numbers). But we only see "659 15". Maybe the British Total is same as BWC because BIB is zero? But here BIB is non-zero (1,805,10). So British Total should be BWC+BIB: vessels=2, tons=1610, crews=29. But the OCR shows "659 15". That seems like Foreign With Cargoes: 659 vessels, 15? tons? 15 crews? But 659 vessels for British Columbia seems high. Let's check the next row: "Coast of China aud Formosa, 788 520,845 34,808 41 27,460 916 *820 548,011 35,724 17,687 1,363,296 240,229". That row has many numbers. Maybe the "British Columbin,." row is actually two lines: "British Columbin,." and then "1 805 19 : 1 805) 10" and then "659 15" belongs to next row? But "Coast of China" starts after. Let's see the raw: "British Columbin,. 1 805 19 : 1 805) 10 659 15 Coast of China aud Formosa, 788 520,845 34,808 41 27,460 916 *820 548,011 35,724 17,687 1,363,296 240,229" Perhaps "659 15" is the start of Coast of China row? But Coast of China starts with "788". Hmm. Let's look at the pattern: Each country row seems to have the country name, then 12 numbers. The OCR splits lines arbitrarily. I'll try to extract all numbers in order from the text after the header, ignoring country names, and then assign 12 numbers per country. There are 17 countries (including TOTAL). 17*12 = 204 numbers. But the OCR may have missing numbers. Better approach: Since this is a known historical document, I could search my knowledge for "No. 1.-NUMBER, TONNAGE, AND CREWS OF VESSELS ENTERED AT PORTS" Hong Kong. But as an AI, I should work with given OCR. Given the complexity, perhaps the user expects me to clean up the OCR into a readable Markdown table, preserving the data as it appears, with ... for illegible parts. But the instruction says: "Reconstruct Sentences and Paragraphs" but this is a table. For tables: "if data is clearly in table format, reconstruct the table using Markdown table syntax." So I must produce a Markdown table. I'll do my best to parse the data into a table with the 13 columns (Country + 12 data columns). I'll use the total row as anchor to verify column meanings. Let me list the countries in order as they appear: Australia and New Zealand British Columbia Coast of China and Formosa Cochin China Continent of Europe Great Britain India and Singapore Japan Java and other Islands in the Indian Archipelago Macao Pellew Islands Philippine Islands Sandwich Islands Siam South America United States of America West Indies TOTAL That's 18 rows (including TOTAL). The OCR shows 17 countries + TOTAL = 18. Now, I'll go through the OCR text and extract numbers for each country. I'll copy the OCR text into a linear sequence, removing line breaks, but keeping country names as delimiters. Text after header (starting from "Australia and New Zealand"): "Australia and New Zealand, · Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons, Crews. 53 28,507 881 * :. 53 28,507 881 99 14,835 476 British Columbin,. 1 805 19 : 1 805) 10 659 15 Coast of China aud Formosa, 788 520,845 34,808 41 27,460 916 820 548,011 35,724 17,687 1,363,296 240,229 Cochin China,...... 71 60,021 2,124 1 385 14 72 00,406 2,138 00 37,632 1,602 Continent of En- rope, C# 1,802 61 3 1,802 51 42 69,457 5,210 Great Britain,..... 120 140,226|| 4,613 3 3,083 104 120 140,300 4,717 43 38,378 1,095 India und Singa pore,... 82 110,752 0,302 82 17 110,752||| 6,332|| 8 4,335 120 Japan, 02 63,505 3,158 1 845 30 01 04,140 3,194 52 39,343 4,517 Java and other Islands in the In- dinn Archipelago, Macao, + Pellew Islands, 8 3,116 101 1 1,380 45 0 4,505 230 1,076 71 311 101,470 12,508 1 80 20 312 101,556 12,618 1,467 211,054 86,957 Philippine Islands, 11 5,137 308 Sandwich Islands, Siam, 1. 8,451 382 South America,.... ... United States of America, 14 20,706 770 West Indies, : ... * : ⠀ : 1 176 12 11 5,137 308 62 20,800 1,634 : 1,493 40 17 8,454 382 65 28,700 2,429 1,312 28 14 • 20,705 770 24 74,597 2,238 --- + 1,113,689 67,370 19,4821,008,1 TOTAL,...... 1,544 1,080,435 60,241 48 33,254 1,135 1,592 1,119,089 67,376 19,4821,008,181 296,679" This is a mess. But I can see patterns: each country name appears, then numbers. The numbers are grouped in 3s. Let me write a parser mentally. I'll split by country names (known list). But the OCR has typos: "British Columbin,.", "Coast of China aud Formosa", "Continent of En- rope", "India und Singa pore", "In- dinn Archipelago", "Philippine Islands", "Sandwich Islands", "Siam", "South America", "United States of America", "West Indies", "TOTAL". I'll go sequentially. Start: "Australia and New Zealand" then numbers: 53, 28507, 881, (then ** :. likely 0,0,0), 53, 28507, 881, 99, 14835, 476. That's 12 numbers? Let's count: 53,28507,881 (3), then three zeros (3), then 53,28507,881 (3), then 99,14835,476 (3) = 12. Good. Next country: "British Columbin,." then numbers: 1, 805, 19, (then ":" maybe separator), 1, 805, 10 (from "805) 10"), then 659, 15? But we need 12 numbers. After that, next country "Coast of China aud Formosa" appears. So maybe the numbers for British Columbia are only 8 numbers? Let's see: "1 805 19 : 1 805) 10 659 15". That's 1,805,19,1,805,10,659,15 = 7 numbers. Missing 5 numbers. Could be that British In Ballast is 1,805,10; British Total is 2,1610,29; Foreign With Cargoes is 659,15,? But 659,15 only two numbers. Maybe the third is missing. Or maybe the "659 15" belongs to Coast of China? But Coast of China starts with 788. Look at Coast of China: "788 520,845 34,808 41 27,460 916 *820 548,011 35,724 17,687 1,363,296 240,229". That's many numbers: let's count: 788, 520845, 34808, 41, 27460, 916, 820, 548011, 35724, 17687, 1363296, 240229 = 12 numbers! Perfect. So Coast of China has 12 numbers. Thus British Columbia should have 12 numbers before Coast of China. The text between "British Columbin,." and "Coast of China aud Formosa" is: "1 805 19 : 1 805) 10 659 15". That's only 7 numbers. Perhaps the OCR missed some numbers (like zeros for British Total and Foreign With Cargoes?). But the total row shows British Columbia likely has some entries. Maybe the "659 15" is actually the start of Coast of China? But Coast of China starts with 788. Could be that "659 15" are the last two numbers of British Columbia (Foreign With Cargoes: 659 vessels, 15 tons? but missing crews). And the first number of Coast of China is 788. But then Coast of China would have 13 numbers? Let's check: if we take "659 15 788 520,845 34,808 41 27,460 916 *820 548,011 35,724 17,687 1,363,296 240,229" that's 14 numbers. Too many. Alternatively, the "659 15" might be "659 15" for British Columbia's Foreign With Cargoes (vessels=659, tons=15? no). That seems unlikely. Let's look at the original image? Not available. I need to make a best guess. Given the total row, we can work backwards? But too many unknowns. Perhaps the table is not uniform; some countries have only British entries, others only Foreign. But the header suggests all four categories for each country. Another idea: The OCR might have merged two lines: "British Columbin,." might be two entries: "British Columbia" and "Coast of China"? No. Let's search memory: This looks like the "Table No. 1" from Hong Kong Blue Book 1870s. I recall a table: "Number, Tonnage, and Crews of Vessels Entered at Ports in the Colony of Hong Kong from each Country". The countries include: Great Britain, British Colonies (Australia, British Columbia, etc.), Foreign Countries (USA, China, Japan, etc.). The table usually has columns: Countries, British (With Cargoes, In Ballast, Total), Foreign (With Cargoes, In Ballast, Total). But the header here only shows Foreign With Cargoes. Maybe the original table had two pages: one for British, one for Foreign? But the OCR shows both. Given the difficulty, I think the best is to present the table as reconstructed from the OCR with clear markings for uncertain data. But the user wants a clean Markdown table. I'll try to produce a table with the data I can confidently parse, and for rows that are garbled, I'll put ... or note. But the instruction: "Preserve original content — do not add or remove substantive words. Do not rephrase, rewrite, or change the author's style, tone, or word choice. Only correct what is clearly wrong." For a table, this means I should keep the numbers as they appear, but fix obvious OCR errors (like "00,406" -> "60,406"?). And align columns. I'll create a Markdown table with the 13 columns. I'll parse each country row by extracting numbers in groups of 3 for each of the 4 categories. I'll use the country names as row labels. Let me attempt to parse all numbers in order after each country name, assuming 12 numbers per country (except TOTAL which has 12). I'll write a quick mental script. I'll go through the text token by token (numbers and country names). But doing manually for 18 rows is tedious but doable. I'll list the text tokens (numbers and country names) in sequence. First, define country names as they appear in OCR: "Australia and New Zealand" "British Columbin,." "Coast of China aud Formosa" "Cochin China,......" "Continent of En- rope" "Great Britain,....." "India und Singa pore" "Japan" "Java and other Islands in the In- dinn Archipelago" "Macao" "Pellew Islands" "Philippine Islands" "Sandwich Islands" "Siam" "South America" "United States of America" "West Indies" "TOTAL" Now, the text between them. I'll split the raw text by these country names (approximate). But the raw text has line breaks and OCR artifacts. Let me write the raw text as a single string and then split by country names (using the cleaned names). But I'll do manually. I'll copy the raw text into a variable and simulate. Raw text (from "Australia and New Zealand" to end): "Australia and New Zealand, · Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons, Crews. 53 28,507 881 * :. 53 28,507 881 99 14,835 476 British Columbin,. 1 805 19 : 1 805) 10 659 15 Coast of China aud Formosa, 788 520,845 34,808 41 27,460 916 820 548,011 35,724 17,687 1,363,296 240,229 Cochin China,...... 71 60,021 2,124 1 385 14 72 00,406 2,138 00 37,632 1,602 Continent of En- rope, C# 1,802 61 3 1,802 51 42 69,457 5,210 Great Britain,..... 120 140,226|| 4,613 3 3,083 104 120 140,300 4,717 43 38,378 1,095 India und Singa pore,... 82 110,752 0,302 82 17 110,752||| 6,332|| 8 4,335 120 Japan, 02 63,505 3,158 1 845 30 01 04,140 3,194 52 39,343 4,517 Java and other Islands in the In- dinn Archipelago, Macao, + Pellew Islands, 8 3,116 101 1 1,380 45 0 4,505 230 1,076 71 311 101,470 12,508 1 80 20 312 101,556 12,618 1,467 211,054 86,957 Philippine Islands, 11 5,137 308 Sandwich Islands, Siam, 1. 8,451 382 South America,.... ... United States of America, 14 20,706 770 West Indies, : ... * : ⠀ : 1 176 12 11 5,137 308 62 20,800 1,634 : 1,493 40 17 8,454 382 65 28,700 2,429 1,312 28 14 • 20,705 770 24 74,597 2,238 --- + 1,113,689 67,370 19,4821,008,1 TOTAL,...... 1,544 1,080,435 60,241 48 33,254 1,135 1,592 1,119,089 67,376 19,4821,008,181 296,679" Now, I'll split by country names. But note: "Java and other Islands in the In- dinn Archipelago, Macao, + Pellew Islands," appears as three countries in one line? Actually "Java and other Islands in the Indian Archipelago", "Macao", "Pellew Islands" are three separate countries. The OCR runs them together: "Java and other Islands in the In- dinn Archipelago, Macao, + Pellew Islands,". So they are separate rows but the numbers follow for each? The numbers after that: "8 3,116 101 1 1,380 45 0 4,505 230 1,076 71 311 101,470 12,508 1 80 20 312 101,556 12,618 1,467 211,054 86,957 *" That's many numbers. Likely each country gets 12 numbers. So three countries = 36 numbers. Let's count numbers in that block: 8, 3116, 101, 1, 1380, 45, 0, 4505, 230, 1076, 71, 311, 101470, 12508, 1, 80, 20, 312, 101556, 12618, 1467, 211054, 86957. That's 23 numbers. Not 36. Then " *" maybe artifacts. Then "Philippine Islands, 11 5,137 308" etc. This is too messy. Perhaps the OCR has merged multiple table rows into one line. Given the time, I think the best is to output a Markdown table that represents the structure, with the data that is clearly readable, and use ... for missing/illegible cells. But the user might expect a complete table. Given the instruction: "You are a professional editor specializing in proofreading OCR output of historical records relating to Hong Kong." So I should use my knowledge of Hong Kong historical records to reconstruct the table accurately. I can recall or approximate the data? But I shouldn't invent data. Maybe I can find a similar table online in my training data. The table "No. 1.-NUMBER, TONNAGE, AND CREWS OF VESSELS ENTERED AT PORTS" appears in Hong Kong Blue Books. For example, Blue Book 1875, 1880, etc. The data might match a specific year. The total row: British With Cargoes 1,544 vessels, 1,080,435 tons, 60,241 crews; British In Ballast 48 vessels, 33,254 tons, 1,135 crews; British Total 1,592 vessels, 1,119,089 tons, 67,376 crews; Foreign With Cargoes 19,482 vessels, 1,008,181 tons, 296,679 crews. This looks like a specific year. Could be 1879 or 1880. But I cannot be sure. However, as an AI, I have access to a lot of historical data. I could reconstruct the table from known data? But the instruction says not to add or remove substantive content. So I must stick to the OCR. Given the difficulty, I'll produce a Markdown table with the columns as identified, and fill in the numbers as they appear in the OCR for each country, using the sequence of numbers. I'll do my best to align. Let me attempt to parse each country's numbers by taking the numbers that appear after the country name until the next country name. I'll write a quick parser in my mind. I'll go through the raw text and extract numbers (including those with commas, pipes, etc.) and country names. I'll tokenize by spaces and newlines, but keep numbers with commas as single tokens. Let me list tokens sequentially: Tokens (split by whitespace, but keep punctuation attached? Better to split by whitespace and then cleanly). I'll copy the raw text into a list of words. But perhaps I can use the fact that the total row at the end has 12 numbers (after splitting the garbled one). And the total row matches the sum of each column? Not necessarily, but likely. Given the time constraints, I'll produce a table with the header and the rows that are clearly parseable: Australia and New Zealand, Coast of China and Formosa, Cochin China, Continent of Europe, Great Britain, India and Singapore, Japan, and TOTAL. For the others, I'll put ... or partial. But the user might want the full table. I'll try to parse all. Let me do a systematic extraction. I'll write a pseudo-code: text = "Australia and New Zealand, · Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons, Crews. 53 28,507 881 * :. 53 28,507 881 99 14,835 476 British Columbin,. 1 805 19 : 1 805) 10 659 15 Coast of China aud Formosa, 788 520,845 34,808 41 27,460 916 820 548,011 35,724 17,687 1,363,296 240,229 Cochin China,...... 71 60,021 2,124 1 385 14 72 00,406 2,138 00 37,632 1,602 Continent of En- rope, C# 1,802 61 3 1,802 51 42 69,457 5,210 Great Britain,..... 120 140,226|| 4,613 3 3,083 104 120 140,300 4,717 43 38,378 1,095 India und Singa pore,... 82 110,752 0,302 82 17 110,752||| 6,332|| 8 4,335 120 Japan, 02 63,505 3,158 1 845 30 01 04,140 3,194 52 39,343 4,517 Java and other Islands in the In- dinn Archipelago, Macao, + Pellew Islands, 8 3,116 101 1 1,380 45 0 4,505 230 1,076 71 311 101,470 12,508 1 80 20 312 101,556 12,618 1,467 211,054 86,957 Philippine Islands, 11 5,137 308 Sandwich Islands, Siam, 1. 8,451 382 South America,.... ... United States of America, 14 20,706 770 West Indies, : ... * : ⠀ : 1 176 12 11 5,137 308 62 20,800 1,634 : 1,493 40 17 8,454 382 65 28,700 2,429 1,312 28 14 • 20,705 770 24 74,597 2,238 --- + 1,113,689 67,370 19,4821,008,1 TOTAL,...... 1,544 1,080,435 60,241 48 33,254 1,135 1,592 1,119,089 67,376 19,4821,008,181 296,679" I'll split by country names (using a list). But the country names in text have variations. I'll use the following split points (by the start of each country name as it appears): "Australia and New Zealand" "British Columbin" "Coast of China aud Formosa" "Cochin China" "Continent of En- rope" "Great Britain" "India und Singa pore" "Japan" "Java and other Islands in the In- dinn Archipelago" "Macao" "Pellew Islands" "Philippine Islands" "Sandwich Islands" "Siam" "South America" "United States of America" "West Indies" "TOTAL" Now, I'll find the indices of these in the text. But doing manually. Let me write the text with markers. I'll go through the text and note where each country appears. Start: "Australia and New Zealand, · Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons, Crews. 53 28,507 881 ** :. 53 28,507 881 99 14,835 476 British Columbin,." So after "476" comes "British Columbin,.". Next: "British Columbin,. 1 805 19 : 1 805) 10 659 15 Coast of China aud Formosa," So after "15" comes "Coast of China aud Formosa,". Next: "Coast of China aud Formosa, 788 520,845 34,808 41 27,460 916 *820 548,011 35,724 17,687 1,363,296 240,229 Cochin China,......" After "240,229" comes "Cochin China,......". Next: "Cochin China,...... 71 60,021 2,124 1 385 14 72 00,406 2,138 00 37,632 1,602 Continent of En- rope," After "1,602" comes "Continent of En- rope,". Next: "Continent of En- rope, C# 1,802 61 * 3 1,802 51 42 69,457 5,210 Great Britain,....." After "5,210" comes "Great Britain,.....". Next: "Great Britain,..... 120 140,226|| 4,613 3 3,083 104 120 140,300 4,717 43 38,378 1,095 India und Singa pore,..." After "1,095" comes "India und Singa pore,...". Next: "India und Singa pore,... 82 110,752 0,302 82 17 110,752||| 6,332|| 8 4,335 120 Japan," After "120" comes "Japan,". Next: "Japan, 02 63,505 3,158 1 845 30 01 04,140 3,194 52 39,343 4,517 Java and other Islands in the In- dinn Archipelago," After "4,517" comes "Java and other Islands in the In- dinn Archipelago,". Next: "Java and other Islands in the In- dinn Archipelago, Macao, + Pellew Islands, 8 3,116 101 1 1,380 45 0 4,505 230 1,076 71 311 101,470 12,508 1 80 20 312 101,556 12,618 1,467 211,054 86,957 * Philippine Islands," After "86,957 *" comes "Philippine Islands,". Next: "Philippine Islands, 11 5,137 308 Sandwich Islands," After "308" comes "Sandwich Islands,". Next: "Sandwich Islands, Siam, 1. 8,451 382 South America,...." After "382" comes "South America,....". Next: "South America,.... ... United States of America, 14 20,706 770 West Indies," After "770" comes "West Indies,". Next: "West Indies, : ... * : ⠀ : 1 176 12 11 5,137 308 62 20,800 1,634 : 1,493 40 17 8,454 382 65 28,700 2,429 1,312 28 14 • 20,705 770 24 74,597 2,238 --- + 1,113,689 67,370 19,4821,008,1 TOTAL,......" After "19,4821,008,1" comes "TOTAL,......". Next: "TOTAL,...... 1,544 1,080,435 60,241 48 33,254 1,135 1,592 1,119,089 67,376 19,4821,008,181 296,679" End. Now, for each country, the numbers between its name and the next country name are its data. But note that the header "Vessels Tons. Crews..." appears only at the start. Also, there are artifacts like "·", "*", ":.", "C#", "", "||", "|||", "+", "---", " *", "•", etc. These are not numbers. I need to extract only numeric tokens (digits with commas, possibly with pipes). Also, some numbers are split like "02" for Japan (maybe 2? but likely 2 vessels? Actually "02" could be 2). "1." for Siam (maybe 1). "8" for Java? etc. Let's extract numeric tokens for each country. Define a function to extract numbers from a string: tokens that match regex ^\d[\d,.]*$ or with pipes? But pipes are artifacts. I'll clean: remove any non-digit, non-comma, non-period? But tons have commas. So keep digits and commas. Also, some have trailing pipes like "140,226||" -> "140,226". "110,752|||" -> "110,752". "6,332||" -> "6,332". "00,406" -> "00,406" (maybe 60,406?). "19,4821,008,181" is garbled. I'll process each country's segment. I'll write the segment strings: Australia and New Zealand segment: "· Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons, Crews. 53 28,507 881 ** :. 53 28,507 881 99 14,835 476" Numbers: 53, 28507, 881, 53, 28507, 881, 99, 14835, 476. That's 9 numbers. But we need 12. The missing three are for British In Ballast (likely 0,0,0). The "** :." might represent zeros. So we can assume British In Ballast: 0,0,0. So the 12 numbers: BWC: 53,28507,881; BIB: 0,0,0; BT: 53,28507,881; FWC: 99,14835,476. British Columbia segment: "1 805 19 : 1 805) 10 659 15" Numbers: 1, 805, 19, 1, 805, 10, 659, 15. That's 8 numbers. Need 12. Missing 4. Could be that British In Ballast: 1,805,10; British Total: 2,1610,29; Foreign With Cargoes: 659,15,? (missing crews). Or maybe the segment includes only up to Foreign With Cargoes vessels and tons, missing crews. The next segment starts with 788 (Coast of China). So maybe the 659,15 are actually the first two numbers of Coast of China? But Coast of China segment starts with 788. So 659,15 belong to British Columbia. Let's assume British Columbia has: BWC: 1,805,19; BIB: 1,805,10; BT: 2,1610,29; FWC: 659,15,? (missing). But we have only 8 numbers. Could be that BT and FWC are combined? Not sure. Given the total row, we can check if British Columbia appears in totals. But not now. I'll note the numbers as they appear and fill missing with .... Coast of China and Formosa segment: "788 520,845 34,808 41 27,460 916 *820 548,011 35,724 17,687 1,363,296 240,229" Numbers: 788, 520845, 34808, 41, 27460, 916, 820, 548011, 35724, 17687, 1363296, 240229. That's 12 numbers! Perfect. So: BWC: 788, 520845, 34808 BIB: 41, 27460, 916 BT: 820, 548011, 35724 FWC: 17687, 1363296, 240229 Cochin China segment: "71 60,021 2,124 1 385 14 72 00,406 2,138 00 37,632 1,602" Numbers: 71, 60021, 2124, 1, 385, 14, 72, 00406, 2138, 00, 37632, 1602. That's 12 numbers. But "00,406" likely 60,406? "00" maybe 0? "00" as tons? Let's keep as is: 00,406 -> 00,406 (maybe 60,406). "00" -> 0. So: BWC: 71, 60021, 2124 BIB: 1, 385, 14 BT: 72, 00406, 2138 FWC: 00, 37632, 1602? Wait, 00 is vessels? Then 37632 tons, 1602 crews. But FWC should be three numbers: vessels, tons, crews. So 00, 37632, 1602. That works: 0 vessels? But 37632 tons with 0 vessels? That seems odd. Maybe "00" is actually 60? But it's "00". Could be "60" misread. But we'll keep as 00. Continent of Europe segment: "C# 1,802 61 * 3 1,802 51 42 69,457 5,210" Numbers: 1802, 61, 3, 1802, 51, 42, 69457, 5210. That's 8 numbers. Need 12. Missing 4. "C#" and "" are artifacts. So we have 8 numbers. Possibly BWC: 1802,61,? missing tons? Actually 1802 might be vessels, 61 tons? But tons should be larger. 1,802 could be tons? 61 crews? Then 3 maybe 3 vessels? Let's see: "1,802 61" then "* 3" then "1,802 51" then "42 69,457 5,210". Hard to parse. Maybe the segment is: "C# 1,802 61 * 3 1,802 51 42 69,457 5,210". Numbers: 1802, 61, 3, 1802, 51, 42, 69457, 5210. Could be: BWC: 1802 vessels, 61 tons? No. 1,802 tons, 61 crews? Then BIB: 3 vessels, 1,802 tons, 51 crews? Then BT: 42 vessels, 69,457 tons, 5,210 crews? Then FWC missing. But we have only 8 numbers. Given the difficulty, I'll keep the numbers as they appear in groups of 3 as much as possible. Great Britain segment: "120 140,226|| 4,613 3 3,083 104 120 140,300 4,717 43 38,378 1,095" Numbers: 120, 140226, 4613, 3, 3083, 104, 120, 140300, 4717, 43, 38378, 1095. That's 12 numbers! Good. Clean: 140,226|| -> 140226; 140,300 -> 140300. So: BWC: 120, 140226, 4613 BIB: 3, 3083, 104 BT: 120, 140300, 4717 FWC: 43, 38378, 1095 India and Singapore segment: "82 110,752 0,302 82 17 110,752||| 6,332|| 8 4,335 120" Numbers: 82, 110752, 0302, 82, 17, 110752, 6332, 8, 4335, 120. That's 10 numbers. Need 12. "0,302" -> 0,302? Maybe 6,302? "110,752|||" -> 110752. "6,332||" -> 6332. We have 10 numbers. Missing 2. Could be BWC: 82, 110752, 0302 (3); BIB: 82, 17, 110752? That doesn't make sense. Let's group: 82, 110752, 0302 (3); 82, 17, 110752 (3); 6332, 8, 4335 (3); 120 (1). Not good. Maybe the segment includes "82 110,752 0,302" (BWC), "82 17 110,752|||" (BIB? but 82 vessels again?), "6,332|| 8 4,335 120" (BT and FWC?). Hard. Japan segment: "02 63,505 3,158 1 845 30 01 04,140 3,194 52 39,343 4,517" Numbers: 02, 63505, 3158, 1, 845, 30, 01, 04140, 3194, 52, 39343, 4517. That's 12 numbers! Good. Clean: 02 -> 2; 01 -> 1; 04,140 -> 4140. So: BWC: 2, 63505, 3158 BIB: 1, 845, 30 BT: 1, 4140, 3194 FWC: 52, 39343, 4517 Java and other Islands in the Indian Archipelago segment: This is combined with Macao and Pellew Islands. The segment: "Macao, + Pellew Islands, 8 3,116 101 1 1,380 45 0 4,505 230 1,076 71 311 101,470 12,508 1 80 20 312 101,556 12,618 1,467 211,054 86,957 *" But the country name "Java and other Islands in the In- dinn Archipelago" appears before "Macao, + Pellew Islands,". So the segment for Java is from after its name to "Macao,". But the text: "Java and other Islands in the In- dinn Archipelago, Macao, + Pellew Islands, 8 3,116 101 1 1,380 45 0 4,505 230 1,076 71 311 101,470 12,508 1 80 20 312 101,556 12,618 1,467 211,054 86,957 *" So the numbers after "Archipelago," are for Java? But then "Macao," and "Pellew Islands," are separate countries, but they have no numbers before the numbers start? Actually the numbers start after "Pellew Islands,". So likely the three countries share the same numbers? No, each should have its own row. The OCR has merged three rows into one line. The numbers that follow are for Java, Macao, Pellew Islands sequentially? There are 23 numbers. 23/3 = 7.66, not 12 each. 23 numbers for three countries = not divisible by 12. Maybe the numbers are for Java only? But then Macao and Pellew Islands have no data? Or the numbers are for all three combined? Unlikely. Let's count numbers in that block: "8 3,116 101 1 1,380 45 0 4,505 230 1,076 71 311 101,470 12,508 1 80 20 312 101,556 12,618 1,467 211,054 86,957" List: 8, 3116, 101, 1, 1380, 45, 0, 4505, 230, 1076, 71, 311, 101470, 12508, 1, 80, 20, 312, 101556, 12618, 1467, 211054, 86957. That's 23 numbers. If each country has 12 numbers, three countries need 36. So not. Maybe the table for these three countries is different? Or the OCR has lost many numbers. Given the complexity, I'll treat "Java and other Islands in the Indian Archipelago", "Macao", "Pellew Islands" as three rows but with missing data. I'll put the numbers that appear in the first row (Java) and leave others blank? But the numbers might be for Java only? Let's see the total row: the totals include all countries. Hard. Philippine Islands segment: "11 5,137 308" only 3 numbers. Then next country "Sandwich Islands". So Philippine Islands only has 3 numbers? Maybe only BWC? But need 12. Sandwich Islands segment: "Siam, 1. 8,451 382" Actually the segment: "Sandwich Islands, Siam, 1. 8,451 382 South America,...." So Sandwich Islands has no numbers? The numbers "1. 8,451 382" belong to Siam? Because "Siam, 1. 8,451 382". So Sandwich Islands might have zero entries. Siam segment: "1. 8,451 382" -> numbers: 1, 8451, 382. Only 3 numbers. South America segment: "..." no numbers? Then "United States of America, 14 20,706 770" So South America maybe zero. United States of America segment: "14 20,706 770" -> 3 numbers. West Indies segment: ": ... * : ⠀ : 1 176 12 11 5,137 308 62 20,800 1,634 : 1,493 40 17 8,454 382 65 28,700 2,429 1,312 28 14 • 20,705 770 24 74,597 2,238 --- + 1,113,689 67,370 19,4821,008,1" Numbers: 1, 176, 12, 11, 5137, 308, 62, 20800, 1634, 1493, 40, 17, 8454, 382, 65, 28700, 2429, 1312, 28, 14, 20705, 770, 24, 74597, 2238, 1113689, 67370, 1948210081? That's many numbers. But this segment ends at "19,4821,008,1" which is the start of TOTAL? Actually the TOTAL segment starts after that. So West Indies segment includes numbers up to "19,4821,008,1". But that number is likely the Foreign With Cargoes total for all countries? But it's in West Indies segment. This is messy. Given the extreme OCR corruption for the latter countries, I think the best is to produce a table with the rows that are clearly parseable (1-8) and for the rest, indicate missing data. But the user might want the full table as per OCR. Given the instruction: "Return the proofread text in standard Markdown". I'll produce a Markdown table with the header and all 18 rows, filling in the numbers as I can parse them, using ... for missing or illegible cells. I'll also clean obvious OCR errors (like "00,406" -> "60,406"? but I'll keep as "00,406" with note? Better to correct unambiguous errors: "00,406" likely "60,406" because Cochin China BT tons should be around 60k. But I'll keep original? The rule: "Correct unambiguous OCR spelling errors". "00,406" is ambiguous. "19,4821,008,181" is clearly garbled; I'll split as "19,482 1,008,181" based on total row. I'll make a decision: I'll reconstruct the table with the 13 columns. For each country, I'll list the 12 data points in order: BWC_V, BWC_T, BWC_C, BIB_V, BIB_T, BIB_C, BT_V, BT_T, BT_C, FWC_V, FWC_T, FWC_C. I'll extract from the segments as groups of 3. If a segment has 12 numbers, perfect. If less, I'll pad with ... at the end. If more, I'll truncate? But unlikely. Let's do it systematically for each country using the segments defined. I'll create a list of countries and their raw segment strings (the text between country name and next country name). Then extract numbers (cleaned) from each segment. Then group into 3s. If not enough, fill with "...". First, define cleaned numbers: remove any non-digit, non-comma characters from each token, but keep commas. Tokens separated by whitespace. Also, tokens like "140,226||" become "140,226". "0,302" stays. "00,406" stays. "19,4821,008,181" is one token? In the TOTAL segment, it's "19,4821,008,181" as one token. I'll split that later. Let's process each segment. I'll write a quick mental script. Segment 1 (Australia): "· Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons, Crews. 53 28,507 881 ** :. 53 28,507 881 99 14,835 476" Tokens: "·", "Vessels", "Tons.", "Crews.", "Vessels", "Tons.", "Crews.", "Vessels", "Tons.", "Crews.", "Vessels", "Tons,", "Crews.", "53", "28,507", "881", "**", ":.", "53", "28,507", "881", "99", "14,835", "476" Numbers: "53", "28,507", "881", "53", "28,507", "881", "99", "14,835", "476" -> 9 numbers. Grouped: [53,28507,881], [53,28507,881], [99,14835,476] -> missing BIB group. So insert [0,0,0] as second group? But the original has "** :." which might indicate zeros. I'll assume BIB = 0,0,0. So final 12: 53,28507,881, 0,0,0, 53,28507,881, 99,14835,476. Segment 2 (British Columbia): "1 805 19 : 1 805) 10 659 15" Tokens: "1", "805", "19", ":", "1", "805)", "10", "659", "15" Numbers: "1", "805", "19", "1", "805", "10", "659", "15" -> 8 numbers. Grouped: [1,805,19], [1,805,10], [659,15] -> third group incomplete. Missing one number for FWC crews. Also missing BT group? Actually we have 8 numbers = 2 full groups (6) + 2. So maybe groups: BWC, BIB, then BT_V, BT_T? But BT needs 3. FWC needs 3. So not sure. Given the pattern, perhaps the segment is missing BT and FWC crews. I'll keep as: BWC: 1,805,19; BIB: 1,805,10; BT: 659,15,...? No. Let's look at the total row: British Columbia likely small. I'll just put the numbers as they appear in order across the 12 columns, padding with ... at the end. So 8 numbers fill first 8 columns: BWC_V=1, BWC_T=805, BWC_C=19, BIB_V=1, BIB_T=805, BIB_C=10, BT_V=659, BT_T=15, BT_C=..., FWC_V=..., FWC_T=..., FWC_C=... But that seems wrong because BT_V=659 is huge. Alternatively, maybe the numbers are: BWC: 1,805,19; BIB: 0,0,0; BT: 1,805,19; FWC: 659,15,? But we have 1,805,10 extra. The "1 805) 10" might be BT: 1,805,10? But BWC is 1,805,19. Then FWC: 659,15. That would be 3+3+3=9 numbers? Actually BWC(3), BT(3), FWC(3) =9, but we have 8. Missing one. If BIB is omitted (all zero), then BWC, BT, FWC = 9 numbers. We have 8. So maybe FWC crews missing. I'll assume the table for British Columbia has BWC, BIB, BT, FWC. The OCR gave BWC, BIB, and then FWC (two numbers). So I'll set BT = BWC+BIB? But not sure. Given the ambiguity, I'll output the raw numbers in a single row with 12 columns, but for British Columbia, I'll put the 8 numbers in the first 8 columns and ... for rest. But that would misalign columns. Better: Since the table has fixed columns, I must align each number to its column. I need to know which column each number belongs to. The header gives column order. The total row gives totals for each column. I could use the total row to deduce column meanings, but not per country. Given the time, I'll produce a table with the data for the countries that have clear 12-number rows: Australia, Coast of China, Cochin China, Great Britain, Japan. For others, I'll put "
Baseline (Original)
( 146 ) No. 1.-NUMBer, tonnagE, and CREWS of VESSELS ENTERED at PORTS COUNTRIES WHENCE ARRIVED. WITH CARGOES. BRITISH. IN BALLAST. TOTAL. FOREIGN. WITH CARGOES. Australia and New Zealand, · Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons, Crews. 53 28,507 881 ** :. 53 28,507 881 99 14,835 476 British Columbin,. 1 805 19 : 1 805) 10 659 15 Coast of China aud Formosa, 788 520,845 34,808 41 27,460 916 *820 548,011 35,724 17,687 1,363,296 240,229 Cochin China,...... 71 60,021 2,124 1 385 14 72 00,406 2,138 00 37,632 1,602 Continent of En- rope, C# 1,802 61 * 3 1,802 51 42 69,457 5,210 Great Britain,..... 120 140,226|| 4,613 3 3,083 104 120 140,300 4,717 43 38,378 1,095 India und Singa pore,... 82 110,752 0,302 82 17 110,752||| 6,332|| 8 4,335 120 Japan, 02 63,505 3,158 1 845 30 01 04,140 3,194 52 39,343 4,517 Java and other Islands in the In- dinn Archipelago, Macao, + Pellew Islands, 8 3,116 101 1 1,380 45 0 4,505 230 1,076 71 311 101,470 12,508 1 80 20 312 101,556 12,618 1,467 211,054 86,957 ** *** Philippine Islands, 11 5,137 308 Sandwich Islands, Siam, 8,451 382 South America,.... ... United States of America, 14 20,706 770 West Indies, : ... * : ⠀ : 1 176 12 11 5,137 308 62 20,800 1,634 : 1,493 40 17 8,454 382 65 28,700 2,429 1,312 28 14 • 20,705 770 24 74,597 2,238 --- + 1,113,689 67,370 19,4821,008,1 TOTAL,...... 1,544 1,080,435 60,241 48 33,254 1,135 1,592 1,119,089 67,376 19,4821,008,181 296,679
2026-07-19 08:11:44 · Baseline
View content

( 146 )

No. 1.-NUMBer, tonnagE, and CREWS of VESSELS ENTERED at PORTS

COUNTRIES WHENCE ARRIVED.

WITH CARGOES.

BRITISH.

IN BALLAST.

TOTAL.

FOREIGN.

WITH CARGOES.

Australia and New

Zealand,

·

Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons. Crews. Vessels Tons, Crews.

53 28,507

881

**

:.

53

28,507 881

99 14,835 476

British Columbin,.

1

805 19

:

1

805) 10

659 15

Coast of China aud

Formosa,

788

520,845 34,808 41

27,460 916

*820

548,011 35,724 17,687 1,363,296 240,229

Cochin China,......

71

60,021 2,124 1

385 14

72

00,406 2,138 00

37,632 1,602

Continent of En-

rope,

C#

1,802 61

*

3

1,802 51 42

69,457 5,210

Great Britain,..... 120 140,226|| 4,613

3

3,083 104

120

140,300 4,717

43

38,378 1,095

India und Singa

pore,...

82 110,752 0,302

82

17

110,752||| 6,332||

8

4,335 120

Japan,

02

63,505 3,158

1

845

30

01

04,140 3,194

52

39,343 4,517

Java and other

Islands in the In- dinn Archipelago,

Macao,

+

Pellew Islands,

8

3,116 101 1

1,380

45

0

4,505 230

1,076 71

311

101,470 12,508 1

80

20

312

101,556 12,618 1,467 211,054 86,957

**

***

Philippine Islands,

11

5,137 308

Sandwich Islands,

Siam,

  1. 8,451 382

South America,....

...

United States of

America,

14

20,706

770

West Indies,

:

...

*

:

:

1

176 12

11

5,137

308

62

20,800 1,634

:

1,493

40

17 8,454

382

65

28,700 2,429

1,312

28

14

20,705 770

24

74,597 2,238

---

+

1,113,689 67,370 19,4821,008,1 TOTAL,...... 1,544 1,080,435 60,241 48 33,254 1,135 1,592 1,119,089 67,376 19,4821,008,181 296,679

Comments

Approved members can add comments, bookmarks, and private notes.

No comments yet.

Private Research Note

Private notes are available after approval.