1932 — Page 441

Blue Books 香港計冊 All AI Reviewed

The user wants me to proofread OCR output of a historical table showing values of imports and exports by countries for Hong Kong, 1922-1932. The OCR is messy with misaligned columns, broken numbers, and text fragments. I need to reconstruct this as a proper Markdown table.

First, let me analyze the structure. The header says "VALUES OF IMPORTS AND EXPORTS (By Countries) 1922-1932." Then "IMPORTS." and "$000's omitted." So values are in thousands of dollars.

The columns appear to be: Countries, then years 1922, 1923, 1924, 1931, 1932? Wait, the header shows "1922 1923 1924 1931 1932" but the data seems to have more columns? Let me check the OCR text.

Looking at the data rows, there are many numbers. The years listed at top: 1922, 1923, 1924, 1931, 1932. But the table might have more years? The OCR shows "1922 1923 1924 1931 1932" but then data for each country has 5 numbers? Let's count.

For U.K.: "55,529 61,812 80,328 78,251 76,905" - that's 5 numbers, matches 5 years.

Australia: "5.903 5,960 8,860 6,288 12.045" - 5 numbers.

Burma: "13.243 6.559 8.936 6,999 6.256" - 5 numbers.

Canada: "2,286 3,978 6.238 5.123 6,236" - 5 numbers.

Ceylon: "$8 172 179 278 213" - first entry "$8" probably 8? But "$000's omitted" so maybe 8? Actually "$8" could be 8 (thousand). But likely a typo: maybe "8" or "58"? We'll keep as 8.

E. Africa: "295 610 404 397 412" - 5 numbers.

India: "26.381 19,552 16,989 17.583 17,649" - 5 numbers.

N. Zealand: "117 158 460 279 179" - 5 numbers.

N. Borneo: "2.271 2,323 2,356 4,503 3.589" - 5 numbers.

S. Africa: "14 310 100 149 509" - 5 numbers.

Straits: "12,461 15,294 14,387 15,197 9,112" - 5 numbers.

W. Africa: "15" only one number? Then next line "W. Indies" with "9 3 34 6" - that's 4 numbers? Actually "W. Indies 9 3 34 6" maybe missing one year? But the table has 5 years. Could be "9 3 34 6" and missing 1922? Or maybe the OCR merged lines.

Let's read the raw OCR lines:

"W. Africa

15

W. Indies

9

3

34

6

B. Empire, other

234

466

332

sal

546"

This is messy. It seems the OCR read line by line, not preserving columns. The original table likely had columns for each year. The OCR output is a vertical list of numbers. I need to reconstruct the table by aligning numbers per country across years.

Given the header: Countries, 1922, 1923, 1924, 1931, 1932. That's 5 year columns.

But the data for each country should have 5 numbers. However, the OCR has broken them into separate lines. I need to parse sequentially.

Let me list all tokens in order after "Countries 1922 1923 1924 1931 1932 $ $ $". Actually the header line: "Countries 1922 1923 1924 1931 1932 $ $ $". The "$ $ $" might be artifacts.

Then the data rows: each country name followed by 5 numbers. But the OCR has line breaks at each number? Let's see the raw text:

"U. K.

55,529

61,812

80,328

78,251

76,905

Australia

5.903

5,960

8,860

6,288

12.045

Burma

13.243

6.559

8.936

6,999

6.256

Canada

2,286

3,978

6.238

5.123

6,236

Ceylon

$8

172

179

278

213

E. Africa

295

610

404

397

412

India

26.381

19,552

16,989

17.583

17,649

N. Zealand

117

158

460

279

179

N. Borneo

2.271

2,323

2,356

4,503

3.589

S. Africa

14

310

100

149

509

Straits

12,461

15,294

14,387

15,197

9,112

W. Africa

15

W. Indies

9

3

34

6

B. Empire, other

234

466

332

sal

546

Belgium

3,366

4,477

5,923

15,018

12,920

N. China

60,228

68,998

77.052

102.501

86,642

*M. China

1,004

1,050

733

12.63!

11,692

*S. China

131

115

85.222

71.659

Cuba

2.677

5

1.336

7

1,086

C. America

9

10

50

24

$1

Denmark

107

487

561

217

186

Egypt

19

51

52

721

124

France

1,611

1.878

2.718

4,261

3.638

Indo China

75,190

$3,463

93,152

50.681

52.669

Germany

6,629

9.957

15.605

37.560

25.708

Holland

2.193

1.462

2,774

4.513

5.534

Italy

642

1.476

2.202

8.655

4.424

Japan

66,220

58,739

77.586

68 303

21.280

Kwong Chow Wan

9.910

8.346

+Macao

10.424

8.878

Norway

770

N. East Indies

48,264

849 68.266

Philippines

6.520

Portugal

3-4

Siam

42.819

4,904 62 67.978

072 80.393 5.559

683

1.669

79,950

61.620

S. America

1.162

53

Sweden

933

986

53.233 52 1.472

1.722 192 47.120

1.643

148

57,822

190 2,599

3,205

Switzerland

475

289

337

1.888

1.118

Spain

64

99

52

U. S. A.

Other

49,815 649

51.566

50,281

338 57,200

68

46.125

1.250

1.658

2.201

2.106

Total

489.707

545.200

613 321

737.740

624.048

Total British Empire

118.846

116,727

139 569

131,932

133,658

Total Foreign

370.861

428,473

473.762

£:02,808 490,390

*Not fully recorded until 193).

+Included in South Chins 1922 1924."

This is a mess. The OCR has lost column alignment. I need to reconstruct the table by assuming each country has 5 data points (for 5 years). But some countries have more or less numbers due to OCR errors.

Let's think: The original table likely had columns for years 1922, 1923, 1924, 1931, 1932. That's 5 years. But the title says "1922-1932", maybe there are more years? The header only shows 5 years. Could be a summary table with selected years.

Given the data, many countries have 5 numbers. But some have more because of line breaks merging with next country.

I need to parse sequentially: each country name appears, then 5 numbers, then next country name. But the OCR sometimes splits numbers across lines, and sometimes country names are split (e.g., "Kwong Chow Wan" then "9.910" then "8.346" then "+Macao" then "10.424" then "8.878"). Actually "+Macao" might be a separate row? The note says "+Included in South Chins 1922 1924." So Macao might be included in South China for those years. But the table might have a row for Macao separately for later years.

Similarly, "M. China" and "S. China" with asterisks.

Also "N. East Indies" appears with "48,264" then "849 68.266" - that's two numbers? Actually "849 68.266" might be two numbers: 849 and 68,266? But the OCR merged them.

"Philippines" then "6.520" only one number? Then "Portugal" "3-4" maybe 3-4? Then "Siam" "42.819" then "4,904 62 67.978" etc.

This is extremely messy. The OCR has completely lost the tabular structure. I need to do my best to reconstruct a plausible table.

Given the instruction: "If data is clearly in table format, reconstruct the table using Markdown table syntax." I must produce a Markdown table.

But the data is too corrupted to perfectly reconstruct. However, I can attempt to align based on the pattern: each country has 5 values for the 5 years. The total row at the end has 5 values: "489.707 545.200 613 321 737.740 624.048" that's 6 numbers? Actually "489.707 545.200 613 321 737.740 624.048" - that's 6 numbers. But "Total British Empire" has 5 numbers: "118.846 116,727 139 569 131,932 133,658" that's 6 numbers? Let's count: 118.846, 116,727, 139, 569, 131,932, 133,658 = 6 numbers. "Total Foreign" has "370.861 428,473 473.762 £:02,808 490,390" = 5 numbers? Actually 370.861, 428,473, 473.762, £:02,808, 490,390 = 5 numbers.

Maybe there are 6 years? The header shows 1922, 1923, 1924, 1931, 1932 (5 years). But the totals have 6 numbers? Let's check the first total: "Total 489.707 545.200 613 321 737.740 624.048" - that's 6 numbers. Could be 1922, 1923, 1924, 1930, 1931, 1932? But header only shows 5.

Wait, the header line: "Countries 1922 1923 1924 1931 1932 $ $ $". The "$ $ $" might indicate three more columns? Or maybe the table has 8 columns? But the data rows don't have that many.

Given the difficulty, perhaps the original table had columns for each year from 1922 to 1932 (11 years). But the OCR only captured some years? The title says "1922-1932". The header shows only 5 years. Could be a summary with selected years.

I think the best approach is to reconstruct the table as it appears in the OCR, with the columns: Country, 1922, 1923, 1924, 1931, 1932. And for each country, take the first 5 numbers after the country name. But the OCR has line breaks that make it ambiguous.

Let's write a script mentally to parse. But as a language model, I'll manually parse the list.

I'll go through the raw text and group by country.

The text starts after "Countries 1922 1923 1924 1931 1932 $ $ $". Then:

  1. U. K. -> 55,529; 61,812; 80,328; 78,251; 76,905
  2. Australia -> 5.903; 5,960; 8,860; 6,288; 12.045
  3. Burma -> 13.243; 6.559; 8.936; 6,999; 6.256
  4. Canada -> 2,286; 3,978; 6.238; 5.123; 6,236
  5. Ceylon -> $8; 172; 179; 278; 213
  6. E. Africa -> 295; 610; 404; 397; 412
  7. India -> 26.381; 19,552; 16,989; 17.583; 17,649
  8. N. Zealand -> 117; 158; 460; 279; 179
  9. N. Borneo -> 2.271; 2,323; 2,356; 4,503; 3.589
  10. S. Africa -> 14; 310; 100; 149; 509
  11. Straits -> 12,461; 15,294; 14,387; 15,197; 9,112
  12. W. Africa -> 15; (then next country) W. Indies -> 9; 3; 34; 6; (then) B. Empire, other -> 234; 466; 332; sal; 546

But W. Africa only has one number "15". W. Indies has four numbers: 9, 3, 34, 6. B. Empire, other has five: 234, 466, 332, sal, 546. "sal" is likely a typo for a number, maybe "381"? Or "541"? Could be "381"? We'll keep as "sal" but maybe it's "381"? The OCR might have misread "381" as "sal". But we should preserve as is? The rule: "Correct unambiguous OCR spelling errors". "sal" is not a number, likely a misread of "381" or "541". But not sure. I'll keep as "sal" but maybe replace with "..."? The rule: "if a word or phrase is clearly missing due to OCR damage, insert ...". But "sal" is present, not missing. It's likely a misrecognition. I'll keep as "sal" but note it's probably a number.

Then:

  1. Belgium -> 3,366; 4,477; 5,923; 15,018; 12,920
  2. N. China -> 60,228; 68,998; 77.052; 102.501; 86,642
  3. *M. China -> 1,004; 1,050; 733; 12.63!; 11,692
  4. S. China -> 131; 115; 85.222; 71.659; (missing fifth?) Actually only four numbers: 131, 115, 85.222, 71.659. Then next is "Cuba". So S. China might have only 4 numbers? But should have 5. Maybe the fifth is missing or merged with Cuba.
  1. Cuba -> 2.677; 5; 1.336; 7; 1,086
  2. C. America -> 9; 10; 50; 24; $1
  3. Denmark -> 107; 487; 561; 217; 186
  4. Egypt -> 19; 51; 52; 721; 124
  5. France -> 1,611; 1.878; 2.718; 4,261; 3.638
  6. Indo China -> 75,190; $3,463; 93,152; 50.681; 52.669
  7. Germany -> 6,629; 9.957; 15.605; 37.560; 25.708
  8. Holland -> 2.193; 1.462; 2,774; 4.513; 5.534
  9. Italy -> 642; 1.476; 2.202; 8.655; 4.424
  10. Japan -> 66,220; 58,739; 77.586; 68 303; 21.280
  11. Kwong Chow Wan -> 9.910; 8.346; (then +Macao) -> 10.424; 8.878; (then Norway) -> 770; (then N. East Indies) -> 48,264; 849 68.266; (then Philippines) -> 6.520; (then Portugal) -> 3-4; (then Siam) -> 42.819; 4,904 62 67.978; 072 80.393 5.559; 683; 1.669; 79,950; 61.620; (then S. America) -> 1.162; 53; (then Sweden) -> 933; 986; 53.233 52 1.472; 1.722 192 47.120; 1.643; 148; 57,822; 190 2,599; 3,205; (then Switzerland) -> 475; 289; 337; 1.888; 1.118; (then Spain) -> 64; 99; 52; (then U. S. A.) -> (then Other) -> 49,815 649; 51.566; 50,281; 338 57,200; 68; 46.125; 1.250; 1.658; 2.201; 2.106; (then Total) -> 489.707; 545.200; 613 321; 737.740; 624.048; (then Total British Empire) -> 118.846; 116,727; 139 569; 131,932; 133,658; (then Total Foreign) -> 370.861; 428,473; 473.762; £:02,808; 490,390.

This is extremely messy. The OCR has completely jumbled the rows after a certain point. It seems the original table had two sections: British Empire countries and Foreign countries. The OCR merged them.

Given the complexity, perhaps the user expects me to output the table as best as I can, using the data that is clearly aligned for the first many rows, and then for the rest, maybe just note the corruption? But the instruction says to reconstruct the table using Markdown table syntax. I must produce a Markdown table.

I think the best is to create a table with the columns: Country, 1922, 1923, 1924, 1931, 1932. And fill in the values for the countries that have clear 5 numbers. For the messy part, I'll try to reconstruct as many rows as possible.

But the user provided the OCR text as a single block. They want the proofread text in Markdown. The original document is a table. So I should output a Markdown table.

Given the difficulty, I'll attempt to parse the first 26 rows (up to Japan) which seem relatively clean. Then for the rest, I'll try to salvage.

But note: The OCR includes "Kwong Chow Wan" and "+Macao" and "Norway" etc. The note says "+Included in South Chins 1922 1924." So Macao might be part of South China for early years.

Also "N. East Indies" likely "Netherlands East Indies".

The data for Siam, Sweden, etc. are garbled.

Maybe the original table had more columns (like 1925-1930) but the OCR only captured some? The header only shows 5 years. The totals have 6 numbers for Total and Total British Empire, 5 for Total Foreign. That suggests maybe 6 years? Let's check the Total row: "489.707 545.200 613 321 737.740 624.048" - that's 6 numbers. Total British Empire: "118.846 116,727 139 569 131,932 133,658" - 6 numbers. Total Foreign: "370.861 428,473 473.762 £:02,808 490,390" - 5 numbers. Inconsistent.

Perhaps the table has 6 year columns: 1922, 1923, 1924, 1930, 1931, 1932? But header shows 1922, 1923, 1924, 1931, 1932. Could be a misprint.

Given the ambiguity, I'll assume 5 year columns as per header. For totals, I'll use the first 5 numbers for Total and Total British Empire, and all 5 for Total Foreign.

But the Total row has 6 numbers. Which 5 correspond? Probably 1922, 1923, 1924, 1931, 1932. The sixth might be a total for something else. I'll take the first 5: 489.707, 545.200, 613, 321, 737.740? But "613 321" might be two numbers: 613 and 321? Actually "613 321" could be "613,321"? The OCR often misses commas. So "613 321" might be 613,321. Similarly "737.740" and "624.048". So the Total row might be: 489,707; 545,200; 613,321; 737,740; 624,048. That's 5 numbers if we combine 613 and 321. The OCR wrote "613 321" with a space. So likely it's 613,321. Similarly, Total British Empire: "118.846 116,727 139 569 131,932 133,658" - "139 569" might be 139,569. So 5 numbers: 118,846; 116,727; 139,569; 131,932; 133,658. Total Foreign: "370.861 428,473 473.762 £:02,808 490,390" - £:02,808 might be 602,808? Or 502,808? The "£:" is likely a misread of a number. Could be "602,808". But we'll keep as is.

Thus, the table likely has 5 year columns. Good.

Now, for each country, we need 5 numbers. The OCR for early countries gives 5 numbers each. For later countries, the numbers are jumbled.

I'll write a Markdown table with the countries I can confidently parse. For the garbled ones, I'll either omit or try to approximate.

But the instruction: "Preserve original content — do not add or remove substantive words." So I should include all rows, even if messy. But I must reconstruct the table. Perhaps I can keep the raw numbers as they appear in sequence for each country, but that would be misleading.

Given the time, I'll produce a table with the clearly parsed rows (up to Japan) and then for the rest, I'll do my best to parse the remaining tokens into rows of 5 numbers each, using the country names as delimiters.

Let's list all country names in order as they appear in the OCR:

  1. U. K.
  2. Australia
  3. Burma
  4. Canada
  5. Ceylon
  6. E. Africa
  7. India
  8. N. Zealand
  9. N. Borneo
  10. S. Africa
  11. Straits
  12. W. Africa
  13. W. Indies
  14. B. Empire, other
  15. Belgium
  16. N. China
  17. *M. China
  18. *S. China
  19. Cuba
  20. C. America
  21. Denmark
  22. Egypt
  23. France
  24. Indo China
  25. Germany
  26. Holland
  27. Italy
  28. Japan
  29. Kwong Chow Wan
  30. +Macao
  31. Norway
  32. N. East Indies
  33. Philippines
  34. Portugal
  35. Siam
  36. S. America
  37. Sweden
  38. Switzerland
  39. Spain
  40. U. S. A.
  41. Other
  42. Total
  43. Total British Empire
  44. Total Foreign

That's 44 rows. The table likely had that many.

Now, for each, we need 5 numbers. The OCR provides a stream of numbers. Let's extract all numbers in order after the header.

I'll write a quick mental parser: after "Countries 1922 1923 1924 1931 1932 $ $ $", the tokens are:

U. K.

55,529

61,812

80,328

78,251

76,905

Australia

5.903

5,960

8,860

6,288

12.045

Burma

13.243

6.559

8.936

6,999

6.256

Canada

2,286

3,978

6.238

5.123

6,236

Ceylon

$8

172

179

278

213

E. Africa

295

610

404

397

412

India

26.381

19,552

16,989

17.583

17,649

N. Zealand

117

158

460

279

179

N. Borneo

2.271

2,323

2,356

4,503

3.589

S. Africa

14

310

100

149

509

Straits

12,461

15,294

14,387

15,197

9,112

W. Africa

15

W. Indies

9

3

34

6

B. Empire, other

234

466

332

sal

546

Belgium

3,366

4,477

5,923

15,018

12,920

N. China

60,228

68,998

77.052

102.501

86,642

*M. China

1,004

1,050

733

12.63!

11,692

*S. China

131

115

85.222

71.659

Cuba

2.677

5

1.336

7

1,086

C. America

9

10

50

24

$1

Denmark

107

487

561

217

186

Egypt

19

51

52

721

124

France

1,611

1.878

2.718

4,261

3.638

Indo China

75,190

$3,463

93,152

50.681

52.669

Germany

6,629

9.957

15.605

37.560

25.708

Holland

2.193

1.462

2,774

4.513

5.534

Italy

642

1.476

2.202

8.655

4.424

Japan

66,220

58,739

77.586

68 303

21.280

Kwong Chow Wan

9.910

8.346

+Macao

10.424

8.878

Norway

770

N. East Indies

48,264

849 68.266

Philippines

6.520

Portugal

3-4

Siam

42.819

4,904 62 67.978

072 80.393 5.559

683

1.669

79,950

61.620

S. America

1.162

53

Sweden

933

986

53.233 52 1.472

1.722 192 47.120

1.643

148

57,822

190 2,599

3,205

Switzerland

475

289

337

1.888

1.118

Spain

64

99

52

U. S. A.

Other

49,815 649

51.566

50,281

338 57,200

68

46.125

1.250

1.658

2.201

2.106

Total

489.707

545.200

613 321

737.740

624.048

Total British Empire

118.846

116,727

139 569

131,932

133,658

Total Foreign

370.861

428,473

473.762

£:02,808

490,390

Now, we need to assign 5 numbers to each country. For the first 28 countries (up to Japan), each has exactly 5 numbers. Good.

For Kwong Chow Wan: it has two numbers (9.910, 8.346) then the next token is "+Macao" which is a country name. So Kwong Chow Wan only has 2 numbers? But should have 5. Maybe the OCR missed the rest. Or maybe Kwong Chow Wan and Macao are combined? The note says "+Included in South Chins 1922 1924." So perhaps for 1922-1924, Macao is included in South China, and for 1931-1932, it's separate? But the table only has 5 years: 1922,1923,1924,1931,1932. So maybe Kwong Chow Wan has data for 1922,1923? And Macao for 1931,1932? But we have only 2 numbers for Kwong Chow Wan and 2 for Macao. That would make 4 numbers. Missing one.

Norway: only one number "770". Then N. East Indies: "48,264" and "849 68.266" (maybe two numbers: 849 and 68,266). Philippines: "6.520" only one. Portugal: "3-4" maybe one number. Siam: many numbers. S. America: two numbers. Sweden: many. Switzerland: five numbers (475,289,337,1.888,1.118) - good. Spain: three numbers (64,99,52). U. S. A.: zero numbers? Then Other: many numbers. Total: five numbers (if we combine 613 321). Total British Empire: five numbers (combine 139 569). Total Foreign: five numbers.

So many rows are incomplete.

Given the difficulty, perhaps the original table had more columns (like 11 years) and the OCR only captured fragments. But the header explicitly says 1922,1923,1924,1931,1932. So I'll stick to 5 columns.

I think the best is to output a Markdown table with the rows that have 5 numbers, and for the others, fill with the available numbers and use "..." for missing.

But the instruction: "Missing text — if a word or phrase is clearly missing due to OCR damage, insert ... in its place." So for missing numbers, I can insert "...".

I'll construct the table with columns: Country, 1922, 1923, 1924, 1931, 1932.

I'll go through each country and assign the first 5 numbers from the stream. For countries where the stream has fewer than 5 numbers before the next country, I'll pad with "...".

But the stream is sequential. So I can simulate: iterate through the list of countries, for each, take the next 5 numbers from the number stream. But the number stream includes numbers that belong to the next country if the previous country has fewer. That's the problem.

Better: The OCR output is essentially a list of lines. The original table had rows. The OCR read row by row but split each cell into a new line. So each row: country name, then 5 numbers. But the OCR put each on a new line. So the structure is: country, num1, num2, num3, num4, num5, next country, num1, num2, ... However, the OCR also inserted line breaks within numbers (like "68 303" split). But generally, the pattern holds for the first 28 rows.

After Japan, the pattern breaks because the OCR might have merged two columns? Or the table had sub-rows? For example, "Kwong Chow Wan" might be a sub-row under China? But it's listed as a separate row.

Given the time, I'll produce a table for the first 28 rows (up to Japan) which are clean, and then for the remaining rows, I'll do a best-effort reconstruction using the numbers that appear after each country name until the next country name. But the OCR doesn't have clear delimiters.

Alternatively, I could output the raw data as a Markdown table with the exact numbers as they appear in the OCR for each country, but that would be messy.

Given the user is a professional editor, they might want a cleaned-up version of the table as it appears in the original. Since I cannot see the original, I must infer.

I'll assume the table has 5 year columns. I'll create rows for all 44 countries, filling in numbers where clear, and using "..." for missing.

Let's manually assign for each country based on the number stream, assuming each country takes 5 numbers in order. But the number stream is not separated per country after Japan. However, we can use the country names as markers. The numbers between two country names belong to the first country. So we can split the number stream by country names.

Let's list the country names in order as they appear in the OCR text. Then the numbers between country name i and country name i+1 belong to country i. For the last country, numbers until the end.

From the raw text, the country names appear in this order (by scanning the text):

  1. U. K.
  2. Australia
  3. Burma
  4. Canada
  5. Ceylon
  6. E. Africa
  7. India
  8. N. Zealand
  9. N. Borneo
  10. S. Africa
  11. Straits
  12. W. Africa
  13. W. Indies
  14. B. Empire, other
  15. Belgium
  16. N. China
  17. *M. China
  18. *S. China
  19. Cuba
  20. C. America
  21. Denmark
  22. Egypt
  23. France
  24. Indo China
  25. Germany
  26. Holland
  27. Italy
  28. Japan
  29. Kwong Chow Wan
  30. +Macao
  31. Norway
  32. N. East Indies
  33. Philippines
  34. Portugal
  35. Siam
  36. S. America
  37. Sweden
  38. Switzerland
  39. Spain
  40. U. S. A.
  41. Other
  42. Total
  43. Total British Empire
  44. Total Foreign

Now, extract the numbers between each.

I'll go through the raw text and split.

The raw text after header:

"U. K.

55,529

61,812

80,328

78,251

76,905

Australia

5.903

5,960

8,860

6,288

12.045

Burma

13.243

6.559

8.936

6,999

6.256

Canada

2,286

3,978

6.238

5.123

6,236

Ceylon

$8

172

179

278

213

E. Africa

295

610

404

397

412

India

26.381

19,552

16,989

17.583

17,649

N. Zealand

117

158

460

279

179

N. Borneo

2.271

2,323

2,356

4,503

3.589

S. Africa

14

310

100

149

509

Straits

12,461

15,294

14,387

15,197

9,112

W. Africa

15

W. Indies

9

3

34

6

B. Empire, other

234

466

332

sal

546

Belgium

3,366

4,477

5,923

15,018

12,920

N. China

60,228

68,998

77.052

102.501

86,642

*M. China

1,004

1,050

733

12.63!

11,692

*S. China

131

115

85.222

71.659

Cuba

2.677

5

1.336

7

1,086

C. America

9

10

50

24

$1

Denmark

107

487

561

217

186

Egypt

19

51

52

721

124

France

1,611

1.878

2.718

4,261

3.638

Indo China

75,190

$3,463

93,152

50.681

52.669

Germany

6,629

9.957

15.605

37.560

25.708

Holland

2.193

1.462

2,774

4.513

5.534

Italy

642

1.476

2.202

8.655

4.424

Japan

66,220

58,739

77.586

68 303

21.280

Kwong Chow Wan

9.910

8.346

+Macao

10.424

8.878

Norway

770

N. East Indies

48,264

849 68.266

Philippines

6.520

Portugal

3-4

Siam

42.819

4,904 62 67.978

072 80.393 5.559

683

1.669

79,950

61.620

S. America

1.162

53

Sweden

933

986

53.233 52 1.472

1.722 192 47.120

1.643

148

57,822

190 2,599

3,205

Switzerland

475

289

337

1.888

1.118

Spain

64

99

52

U. S. A.

Other

49,815 649

51.566

50,281

338 57,200

68

46.125

1.250

1.658

2.201

2.106

Total

489.707

545.200

613 321

737.740

624.048

Total British Empire

118.846

116,727

139 569

131,932

133,658

Total Foreign

370.861

428,473

473.762

£:02,808

490,390"

Now, for each country, the numbers are the lines until the next country name. But note that some numbers are on the same line (e.g., "68 303", "849 68.266", "4,904 62 67.978", etc.). We'll split by whitespace.

Let's parse each country's numbers:

  1. U. K.: 55,529; 61,812; 80,328; 78,251; 76,905 -> 5 numbers.
  2. Australia: 5.903; 5,960; 8,860; 6,288; 12.045 -> 5.
  3. Burma: 13.243; 6.559; 8.936; 6,999; 6.256 -> 5.
  4. Canada: 2,286; 3,978; 6.238; 5.123; 6,236 -> 5.
  5. Ceylon: $8; 172; 179; 278; 213 -> 5.
  6. E. Africa: 295; 610; 404; 397; 412 -> 5.
  7. India: 26.381; 19,552; 16,989; 17.583; 17,649 -> 5.
  8. N. Zealand: 117; 158; 460; 279; 179 -> 5.
  9. N. Borneo: 2.271; 2,323; 2,356; 4,503; 3.589 -> 5.
  10. S. Africa: 14; 310; 100; 149; 509 -> 5.
  11. Straits: 12,461; 15,294; 14,387; 15,197; 9,112 -> 5.
  12. W. Africa: only "15" -> 1 number.
  13. W. Indies: 9; 3; 34; 6 -> 4 numbers.
  14. B. Empire, other: 234; 466; 332; sal; 546 -> 5.
  15. Belgium: 3,366; 4,477; 5,923; 15,018; 12,920 -> 5.
  16. N. China: 60,228; 68,998; 77.052; 102.501; 86,642 -> 5.
  17. *M. China: 1,004; 1,050; 733; 12.63!; 11,692 -> 5.
  18. *S. China: 131; 115; 85.222; 71.659 -> 4 numbers.
  19. Cuba: 2.677; 5; 1.336; 7; 1,086 -> 5.
  20. C. America: 9; 10; 50; 24; $1 -> 5.
  21. Denmark: 107; 487; 561; 217; 186 -> 5.
  22. Egypt: 19; 51; 52; 721; 124 -> 5.
  23. France: 1,611; 1.878; 2.718; 4,261; 3.638 -> 5.
  24. Indo China: 75,190; $3,463; 93,152; 50.681; 52.669 -> 5.
  25. Germany: 6,629; 9.957; 15.605; 37.560; 25.708 -> 5.
  26. Holland: 2.193; 1.462; 2,774; 4.513; 5.534 -> 5.
  27. Italy: 642; 1.476; 2.202; 8.655; 4.424 -> 5.
  28. Japan: 66,220; 58,739; 77.586; 68 303; 21.280 -> 5 (but "68 303" is two numbers? Actually it's one number 68,303 split. We'll treat as 68,303).
  29. Kwong Chow Wan: 9.910; 8.346 -> 2 numbers.
  30. +Macao: 10.424; 8.878 -> 2 numbers.
  31. Norway: 770 -> 1 number.
  32. N. East Indies: 48,264; 849 68.266 -> split: 48,264; 849; 68.266 -> 3 numbers.
  33. Philippines: 6.520 -> 1 number.
  34. Portugal: 3-4 -> 1 number (maybe 3-4 is a range? but likely a number 34? or 3-4? We'll keep as "3-4").
  35. Siam: 42.819; 4,904 62 67.978; 072 80.393 5.559; 683; 1.669; 79,950; 61.620 -> many numbers. Let's split: 42.819; 4,904; 62; 67.978; 072; 80.393; 5.559; 683; 1.669; 79,950; 61.620 -> 11 numbers.
  36. S. America: 1.162; 53 -> 2 numbers.
  37. Sweden: 933; 986; 53.233 52 1.472; 1.722 192 47.120; 1.643; 148; 57,822; 190 2,599; 3,205 -> split: 933; 986; 53.233; 52; 1.472; 1.722; 192; 47.120; 1.643; 148; 57,822; 190; 2,599; 3,205 -> 14 numbers.
  38. Switzerland: 475; 289; 337; 1.888; 1.118 -> 5 numbers.
  39. Spain: 64; 99; 52 -> 3 numbers.
  40. U. S. A.: (no numbers before next country "Other") -> 0 numbers.
  41. Other: 49,815 649; 51.566; 50,281; 338 57,200; 68; 46.125; 1.250; 1.658; 2.201; 2.106 -> split: 49,815; 649; 51.566; 50,281; 338; 57,200; 68; 46.125; 1.250; 1.658; 2.201; 2.106 -> 12 numbers.
  42. Total: 489.707; 545.200; 613 321; 737.740; 624.048 -> split: 489.707; 545.200; 613; 321; 737.740; 624.048 -> 6 numbers.
  43. Total British Empire: 118.846; 116,727; 139 569; 131,932; 133,658 -> split: 118.846; 116,727; 139; 569; 131,932; 133,658 -> 6 numbers.
  44. Total Foreign: 370.861; 428,473; 473.762; £:02,808; 490,390 -> 5 numbers.

Now, the table should have 5 year columns. For each country, we need to map their numbers to the 5 years. For countries with exactly 5 numbers, it's straightforward. For those with fewer, we have missing data. For those with more, maybe the table has more columns? But the header says only 5 years. Could be that the table includes 1922-1932 inclusive (11 years) but the header only shows 5? The title says "1922-1932". The header shows "1922 1923 1924 1931 1932". That's 5 years. Maybe the table only shows those 5 years. So each country should have 5 numbers.

Thus, for countries with more than 5 numbers, the extra numbers might be for other years not shown in header? Or they are misaligned. For Siam, Sweden, Other, they have many numbers. Perhaps those numbers belong to multiple rows? But the OCR didn't separate rows.

Given the difficulty, I'll assume the table has 5 columns. For each country, I'll take the first 5 numbers from their list. If they have fewer, I'll pad with "...". For countries with more, I'll only take first 5.

But for W. Africa, only 1 number; W. Indies 4; *S. China 4; Kwong Chow Wan 2; +Macao 2; Norway 1; N. East Indies 3; Philippines 1; Portugal 1; Siam 11; S. America 2; Sweden 14; Spain 3; U.S.A. 0; Other 12.

This will result in many "..." which might be acceptable.

But maybe the table originally had more columns (like 11 years) and the OCR header only captured 5. The totals have 6 numbers for Total and Total British Empire, 5 for Total Foreign. That suggests maybe 6 years? Let's check the Total row: 489.707, 545.200, 613, 321, 737.740, 624.048. If we combine 613 and 321 as 613,321, then we have 5 numbers: 489,707; 545,200; 613,321; 737,740; 624,048. That matches 5 years. Similarly Total British Empire: 118,846; 116,727; 139,569; 131,932; 133,658. Total Foreign: 370,861; 428,473; 473,762; 602,808?; 490,390. So 5 years.

Thus, 5 year columns.

Now, for each country, we need to decide which 5 numbers correspond to the 5 years. For the first 28 countries, they have exactly 5 numbers, so that's fine.

For the rest, the numbers are likely also 5 per country, but the OCR merged multiple countries' numbers because of missing country names? But we have country names. The numbers between country names are for that country. For Kwong Chow Wan, only 2 numbers. Maybe the other 3 are missing (not recorded). For +Macao, 2 numbers. For Norway, 1. For N. East Indies, 3. For Philippines, 1. For Portugal, 1. For Siam, many. For S. America, 2. For Sweden, many. For Spain, 3. For U.S.A., 0. For Other, many.

This suggests that the OCR for the latter part is completely garbled, possibly because the original table had a different layout (maybe two pages, or columns). The user might not expect perfect reconstruction for the garbled part. But they said "reconstruct the table using Markdown table syntax." I'll do my best.

Given the context, this is a historical trade statistics table. The first part (British Empire) is clear. The foreign countries part is also clear up to Japan. Then there are some additional entries like Kwong Chow Wan, Macao, Norway, Netherlands East Indies, Philippines, Portugal, Siam, South America, Sweden, Switzerland, Spain, USA, Other. These are likely also part of the table. The numbers for them might be in the original but OCR messed up.

I could try to reconstruct by assuming each of those countries has 5 numbers, and the numbers are in the stream in order. But the stream has many numbers for Siam and Sweden, which might be because the OCR combined multiple rows (maybe Siam has data for multiple years? But only 5 years). Or maybe the table has sub-columns for each year (like imports and exports)? But the title says "VALUES OF IMPORTS AND EXPORTS (By Countries) 1922-1932." This table is only for IMPORTS (as per "IMPORTS."). There might be a separate table for EXPORTS. So this is just imports.

Thus, each country has one value per year (in $000). So 5 values.

I think the best is to output a table with the 28 clean rows, and then for the remaining rows, include them with the numbers they have, and use "..." for missing. But the user might want the full table.

Given the instruction "Preserve original content — do not add or remove substantive words." I should include all rows. I'll create a Markdown table with all 44 rows, filling in the numbers as they appear in the OCR for each country, but limited to 5 columns. For countries with more than 5 numbers, I'll take the first 5. For fewer, pad with "...".

But for Siam and Sweden, they have many numbers. Which 5 are correct? Hard to know. Maybe the OCR combined several countries' numbers because the country names were not recognized. For example, after Portugal, the next country is Siam, but the numbers for Siam might include numbers for other countries that lost their names. However, the country names are present: Siam, S. America, Sweden, Switzerland, Spain, U.S.A., Other. So the numbers between them should belong to each. But the numbers for Siam are many, meaning the OCR might have inserted extra line breaks within Siam's numbers? Actually, the numbers for Siam are on multiple lines: "42.819", "4,904 62 67.978", "072 80.393 5.559", "683", "1.669", "79,950", "61.620". That's 7 lines, but each line may have multiple numbers. That's too many for 5 years.

Maybe the table has more than 5 year columns? Let's check the header again: "Countries 1922 1923 1924 1931 1932 $ $ $". The "$ $ $" might indicate three more columns? Or they are just artifacts. Could be that the table has 8 columns: 1922, 1923, 1924, 1925, 1926, 1927, 1928, 1929, 1930, 1931, 1932? But the header only shows 5.

Given the time, I'll assume 5 columns. For the messy rows, I'll take the first 5 numbers from their list. For Siam: 42.819, 4,904, 62, 67.978, 072. For Sweden: 933, 986, 53.233, 52, 1.472. For Other: 49,815, 649, 51.566, 50,281, 338. For Total: 489.707, 545.200, 613, 321, 737.740 (but 613 and 321 should be combined? I'll keep as separate columns? That would be 6 columns. I'll combine 613 and 321 as 613,321 for the third column? But the table has 5 columns. The Total row has 6 numbers if we don't combine. The header has 5 years. So likely the third year (1924) value is 613,321. The OCR split it. So I'll combine them.

Similarly for Total British Empire: combine 139 and 569 as 139,569.

For Total Foreign: £:02,808 maybe 602,808.

I'll clean up numbers: remove spaces, fix obvious OCR errors (like "sal" -> maybe "381"? but I'll keep as "sal" or replace with "..."? The rule: correct unambiguous OCR spelling errors. "sal" is not a number, likely a misread of "381" or "541". But not unambiguous. I'll keep as "sal" but maybe it's "381"? I'll keep as "sal" to preserve original.

Also "12.63!" -> likely "12,631"? Or "12.631"? I'll keep as "12.63!".

"68 303" -> "68,303".

"849 68.266" -> two numbers: 849 and 68,266. But for N. East Indies, we have 3 numbers: 48,264; 849; 68,266. That's 3 numbers for 5 years. Missing two.

"3-4" for Portugal -> maybe "34"? Or "3-4" is a range. I'll keep as "3-4".

"£:02,808" -> maybe "602,808". I'll keep as "£:02,808".

Now, I'll construct the Markdown table.

First, the header row: | Country | 1922 | 1923 | 1924 | 1931 | 1932 |

But the original header has "Countries" and then years. I'll use "Country" as column header.

Also note the title: "VALUES OF IMPORTS AND EXPORTS (By Countries) 1922-1932." and "IMPORTS." and "$000's omitted." I'll include that as text before the table.

Also the footnotes: "*Not fully recorded until 193)." and "+Included in South Chins 1922 1924." I'll include after the table.

Now, the table rows.

I'll write the Markdown.

Let's list rows with 5 values each.

  1. U. K. | 55,529 | 61,812 | 80,328 | 78,251 | 76,905
  2. Australia | 5,903 | 5,960 | 8,860 | 6,288 | 12,045 (note: 5.903 -> 5,903; 12.045 -> 12,045)
  3. Burma | 13,243 | 6,559 | 8,936 | 6,999 | 6,256
  4. Canada | 2,286 | 3,978 | 6,238 | 5,123 | 6,236
  5. Ceylon | 8 | 172 | 179 | 278 | 213 (remove $)
  6. E. Africa | 295 | 610 | 404 | 397 | 412
  7. India | 26,381 | 19,552 | 16,989 | 17,583 | 17,649
  8. N. Zealand | 117 | 158 | 460 | 279 | 179
  9. N. Borneo | 2,271 | 2,323 | 2,356 | 4,503 | 3,589
  10. S. Africa | 14 | 310 | 100 | 149 | 509
  11. Straits | 12,461 | 15,294 | 14,387 | 15,197 | 9,112
  12. W. Africa | 15 | ... | ... | ... | ...
  13. W. Indies | 9 | 3 | 34 | 6 | ...
  14. B. Empire, other | 234 | 466 | 332 | sal | 546
  15. Belgium | 3,366 | 4,477 | 5,923 | 15,018 | 12,920
  16. N. China | 60,228 | 68,998 | 77,052 | 102,501 | 86,642
  17. *M. China | 1,004 | 1,050 | 733 | 12.63! | 11,692
  18. *S. China | 131 | 115 | 85,222 | 71,659 | ...
  19. Cuba | 2,677 | 5 | 1,336 | 7 | 1,086
  20. C. America | 9 | 10 | 50 | 24 | 1 (remove $)
  21. Denmark | 107 | 487 | 561 | 217 | 186
  22. Egypt | 19 | 51 | 52 | 721 | 124
  23. France | 1,611 | 1,878 | 2,718 | 4,261 | 3,638
  24. Indo China | 75,190 | 3,463 | 93,152 | 50,681 | 52,669 (remove $ from second)
  25. Germany | 6,629 | 9,957 | 15,605 | 37,560 | 25,708
  26. Holland | 2,193 | 1,462 | 2,774 | 4,513 | 5,534
  27. Italy | 642 | 1,476 | 2,202 | 8,655 | 4,424
  28. Japan | 66,220 | 58,739 | 77,586 | 68,303 | 21,280
  29. Kwong Chow Wan | 9,910 | 8,346 | ... | ... | ...
  30. +Macao | 10,424 | 8,878 | ... | ... | ...
  31. Norway | 770 | ... | ... | ... | ...
  32. N. East Indies | 48,264 | 849 | 68,266 | ... | ...
  33. Philippines | 6,520 | ... | ... | ... | ...
  34. Portugal | 3-4 | ... | ... | ... | ...
  35. Siam | 42,819 | 4,904 | 62 | 67,978 | 72 (072 -> 72? but 072 is 72) Actually first five: 42.819, 4,904, 62, 67.978, 072. I'll use 42,819; 4,904; 62; 67,978; 72.
  36. S. America | 1,162 | 53 | ... | ... | ...
  37. Sweden | 933 | 986 | 53,233 | 52 | 1,472
  38. Switzerland | 475 | 289 | 337 | 1,888 | 1,118
  39. Spain | 64 | 99 | 52 | ... | ...
  40. U. S. A. | ... | ... | ... | ... | ...
  41. Other | 49,815 | 649 | 51,566 | 50,281 | 338
  42. Total | 489,707 | 545,200 | 613,321 | 737,740 | 624,048
  43. Total British Empire | 118,846 | 116,727 | 139,569 | 131,932 | 133,658
  44. Total Foreign | 370,861 | 428,473 | 473,762 | 602,808 | 490,390 (assuming £:02,808 -> 602,808)

I'll keep the numbers as they appear in OCR but with commas for thousands? The OCR uses commas and periods inconsistently. I'll standardize to commas for thousands and periods for decimals? But the values are in $000's, so they are whole numbers (thousands of dollars). The decimals are likely OCR errors for commas. For example, "5.903" is 5,903. "13.243" is 13,243. "2.271" is 2,271. "26.381" is 26,381. "77.052" is 77,052. "102.501" is 102,501. "12.63!" is 12,631? "85.222" is 85,222. "71.659" is 71,659. "2.677" is 2,677. "1.336" is 1,336. "75,190" already has comma. "$3,463" is 3,463. "93,152" okay. "50.681" is 50,681. "52.669" is 52,669. "9.957" is 9,957. "15.605" is 15,605. "37.560" is 37,560. "25.708" is 25,708. "2.193" is 2,193. "1.462" is 1,462. "4.513" is 4,513. "5.534" is 5,534. "1.476" is 1,476. "2.202" is 2,202. "8.655" is 8,655. "4.424" is 4,424. "77.586" is 77,586. "68 303" is 68,303. "21.280" is 21,280. "9.910" is 9,910. "8.346" is 8,346. "10.424" is 10,424. "8.878" is 8,878. "48,264" okay. "849" okay. "68.266" is 68,266. "6.520" is 6,520. "42.819" is 42,819. "4,904" okay. "62" okay. "67.978" is 67,978. "072" is 72. "80.393" is 80,393. "5.559" is 5,559. "683" okay. "1.669" is 1,669. "79,950" okay. "61.620" is 61,620. "1.162" is 1,162. "53" okay. "933" okay. "986" okay. "53.233" is 53,233. "52" okay. "1.472" is 1,472. "1.722" is 1,722. "192" okay. "47.120" is 47,120. "1.643" is 1,643. "148" okay. "57,822" okay. "190" okay. "2,599" okay. "3,205" okay. "475" okay. "289" okay. "337" okay. "1.888" is 1,888. "1.118" is 1,118. "64" okay. "99" okay. "52" okay. "49,815" okay. "649" okay. "51.566" is 51,566. "50,281" okay. "338" okay. "57,200" okay. "68" okay. "46.125" is 46,125. "1.250" is 1,250. "1.658" is 1,658. "2.201" is 2,201. "2.106" is 2,106. "489.707" is 489,707. "545.200" is 545,200. "613 321" is 613,321. "737.740" is 737,740. "624.048" is 624,048. "118.846" is 118,846. "116,727" okay. "139 569" is 139,569. "131,932" okay. "133,658" okay. "370.861" is 370,861. "428,473" okay. "473.762" is 473,762. "£:02,808" -> 602,808? I'll use

Edit History

2026-07-15 07:46:53 · NVIDIA / nvidia/nemotron-3-ultra-550b-a55b
Live
View comparison
AI Proofread
The user wants me to proofread OCR output of a historical table showing values of imports and exports by countries for Hong Kong, 1922-1932. The OCR is messy with misaligned columns, broken numbers, and text fragments. I need to reconstruct this as a proper Markdown table. First, let me analyze the structure. The header says "VALUES OF IMPORTS AND EXPORTS (By Countries) 1922-1932." Then "IMPORTS." and "$000's omitted." So values are in thousands of dollars. The columns appear to be: Countries, then years 1922, 1923, 1924, 1931, 1932? Wait, the header shows "1922 1923 1924 1931 1932" but the data seems to have more columns? Let me check the OCR text. Looking at the data rows, there are many numbers. The years listed at top: 1922, 1923, 1924, 1931, 1932. But the table might have more years? The OCR shows "1922 1923 1924 1931 1932" but then data for each country has 5 numbers? Let's count. For U.K.: "55,529 61,812 80,328 78,251 76,905" - that's 5 numbers, matches 5 years. Australia: "5.903 5,960 8,860 6,288 12.045" - 5 numbers. Burma: "13.243 6.559 8.936 6,999 6.256" - 5 numbers. Canada: "2,286 3,978 6.238 5.123 6,236" - 5 numbers. Ceylon: "$8 172 179 278 213" - first entry "$8" probably 8? But "$000's omitted" so maybe 8? Actually "$8" could be 8 (thousand). But likely a typo: maybe "8" or "58"? We'll keep as 8. E. Africa: "295 610 404 397 412" - 5 numbers. India: "26.381 19,552 16,989 17.583 17,649" - 5 numbers. N. Zealand: "117 158 460 279 179" - 5 numbers. N. Borneo: "2.271 2,323 2,356 4,503 3.589" - 5 numbers. S. Africa: "14 310 100 149 509" - 5 numbers. Straits: "12,461 15,294 14,387 15,197 9,112" - 5 numbers. W. Africa: "15" only one number? Then next line "W. Indies" with "9 3 34 6" - that's 4 numbers? Actually "W. Indies 9 3 34 6" maybe missing one year? But the table has 5 years. Could be "9 3 34 6" and missing 1922? Or maybe the OCR merged lines. Let's read the raw OCR lines: "W. Africa 15 W. Indies 9 3 34 6 B. Empire, other 234 466 332 sal 546" This is messy. It seems the OCR read line by line, not preserving columns. The original table likely had columns for each year. The OCR output is a vertical list of numbers. I need to reconstruct the table by aligning numbers per country across years. Given the header: Countries, 1922, 1923, 1924, 1931, 1932. That's 5 year columns. But the data for each country should have 5 numbers. However, the OCR has broken them into separate lines. I need to parse sequentially. Let me list all tokens in order after "Countries 1922 1923 1924 1931 1932 $ $ $". Actually the header line: "Countries 1922 1923 1924 1931 1932 $ $ $". The "$ $ $" might be artifacts. Then the data rows: each country name followed by 5 numbers. But the OCR has line breaks at each number? Let's see the raw text: "U. K. 55,529 61,812 80,328 78,251 76,905 Australia 5.903 5,960 8,860 6,288 12.045 Burma 13.243 6.559 8.936 6,999 6.256 Canada 2,286 3,978 6.238 5.123 6,236 Ceylon $8 172 179 278 213 E. Africa 295 610 404 397 412 India 26.381 19,552 16,989 17.583 17,649 N. Zealand 117 158 460 279 179 N. Borneo 2.271 2,323 2,356 4,503 3.589 S. Africa 14 310 100 149 509 Straits 12,461 15,294 14,387 15,197 9,112 W. Africa 15 W. Indies 9 3 34 6 B. Empire, other 234 466 332 sal 546 Belgium 3,366 4,477 5,923 15,018 12,920 N. China 60,228 68,998 77.052 102.501 86,642 *M. China 1,004 1,050 733 12.63! 11,692 *S. China 131 115 85.222 71.659 Cuba 2.677 5 1.336 7 1,086 C. America 9 10 50 24 $1 Denmark 107 487 561 217 186 Egypt 19 51 52 721 124 France 1,611 1.878 2.718 4,261 3.638 Indo China 75,190 $3,463 93,152 50.681 52.669 Germany 6,629 9.957 15.605 37.560 25.708 Holland 2.193 1.462 2,774 4.513 5.534 Italy 642 1.476 2.202 8.655 4.424 Japan 66,220 58,739 77.586 68 303 21.280 Kwong Chow Wan 9.910 8.346 +Macao 10.424 8.878 Norway 770 N. East Indies 48,264 849 68.266 Philippines 6.520 Portugal 3-4 Siam 42.819 4,904 62 67.978 072 80.393 5.559 683 1.669 79,950 61.620 S. America 1.162 53 Sweden 933 986 53.233 52 1.472 1.722 192 47.120 1.643 148 57,822 190 2,599 3,205 Switzerland 475 289 337 1.888 1.118 Spain 64 99 52 U. S. A. Other 49,815 649 51.566 50,281 338 57,200 68 46.125 1.250 1.658 2.201 2.106 Total 489.707 545.200 613 321 737.740 624.048 Total British Empire 118.846 116,727 139 569 131,932 133,658 Total Foreign 370.861 428,473 473.762 £:02,808 490,390 *Not fully recorded until 193). +Included in South Chins 1922 1924." This is a mess. The OCR has lost column alignment. I need to reconstruct the table by assuming each country has 5 data points (for 5 years). But some countries have more or less numbers due to OCR errors. Let's think: The original table likely had columns for years 1922, 1923, 1924, 1931, 1932. That's 5 years. But the title says "1922-1932", maybe there are more years? The header only shows 5 years. Could be a summary table with selected years. Given the data, many countries have 5 numbers. But some have more because of line breaks merging with next country. I need to parse sequentially: each country name appears, then 5 numbers, then next country name. But the OCR sometimes splits numbers across lines, and sometimes country names are split (e.g., "Kwong Chow Wan" then "9.910" then "8.346" then "+Macao" then "10.424" then "8.878"). Actually "+Macao" might be a separate row? The note says "+Included in South Chins 1922 1924." So Macao might be included in South China for those years. But the table might have a row for Macao separately for later years. Similarly, "M. China" and "S. China" with asterisks. Also "N. East Indies" appears with "48,264" then "849 68.266" - that's two numbers? Actually "849 68.266" might be two numbers: 849 and 68,266? But the OCR merged them. "Philippines" then "6.520" only one number? Then "Portugal" "3-4" maybe 3-4? Then "Siam" "42.819" then "4,904 62 67.978" etc. This is extremely messy. The OCR has completely lost the tabular structure. I need to do my best to reconstruct a plausible table. Given the instruction: "If data is clearly in table format, reconstruct the table using Markdown table syntax." I must produce a Markdown table. But the data is too corrupted to perfectly reconstruct. However, I can attempt to align based on the pattern: each country has 5 values for the 5 years. The total row at the end has 5 values: "489.707 545.200 613 321 737.740 624.048" that's 6 numbers? Actually "489.707 545.200 613 321 737.740 624.048" - that's 6 numbers. But "Total British Empire" has 5 numbers: "118.846 116,727 139 569 131,932 133,658" that's 6 numbers? Let's count: 118.846, 116,727, 139, 569, 131,932, 133,658 = 6 numbers. "Total Foreign" has "370.861 428,473 473.762 £:02,808 490,390" = 5 numbers? Actually 370.861, 428,473, 473.762, £:02,808, 490,390 = 5 numbers. Maybe there are 6 years? The header shows 1922, 1923, 1924, 1931, 1932 (5 years). But the totals have 6 numbers? Let's check the first total: "Total 489.707 545.200 613 321 737.740 624.048" - that's 6 numbers. Could be 1922, 1923, 1924, 1930, 1931, 1932? But header only shows 5. Wait, the header line: "Countries 1922 1923 1924 1931 1932 $ $ $". The "$ $ $" might indicate three more columns? Or maybe the table has 8 columns? But the data rows don't have that many. Given the difficulty, perhaps the original table had columns for each year from 1922 to 1932 (11 years). But the OCR only captured some years? The title says "1922-1932". The header shows only 5 years. Could be a summary with selected years. I think the best approach is to reconstruct the table as it appears in the OCR, with the columns: Country, 1922, 1923, 1924, 1931, 1932. And for each country, take the first 5 numbers after the country name. But the OCR has line breaks that make it ambiguous. Let's write a script mentally to parse. But as a language model, I'll manually parse the list. I'll go through the raw text and group by country. The text starts after "Countries 1922 1923 1924 1931 1932 $ $ $". Then: U. K. -> 55,529; 61,812; 80,328; 78,251; 76,905 Australia -> 5.903; 5,960; 8,860; 6,288; 12.045 Burma -> 13.243; 6.559; 8.936; 6,999; 6.256 Canada -> 2,286; 3,978; 6.238; 5.123; 6,236 Ceylon -> $8; 172; 179; 278; 213 E. Africa -> 295; 610; 404; 397; 412 India -> 26.381; 19,552; 16,989; 17.583; 17,649 N. Zealand -> 117; 158; 460; 279; 179 N. Borneo -> 2.271; 2,323; 2,356; 4,503; 3.589 S. Africa -> 14; 310; 100; 149; 509 Straits -> 12,461; 15,294; 14,387; 15,197; 9,112 W. Africa -> 15; (then next country) W. Indies -> 9; 3; 34; 6; (then) B. Empire, other -> 234; 466; 332; sal; 546 But W. Africa only has one number "15". W. Indies has four numbers: 9, 3, 34, 6. B. Empire, other has five: 234, 466, 332, sal, 546. "sal" is likely a typo for a number, maybe "381"? Or "541"? Could be "381"? We'll keep as "sal" but maybe it's "381"? The OCR might have misread "381" as "sal". But we should preserve as is? The rule: "Correct unambiguous OCR spelling errors". "sal" is not a number, likely a misread of "381" or "541". But not sure. I'll keep as "sal" but maybe replace with "..."? The rule: "if a word or phrase is clearly missing due to OCR damage, insert ...". But "sal" is present, not missing. It's likely a misrecognition. I'll keep as "sal" but note it's probably a number. Then: Belgium -> 3,366; 4,477; 5,923; 15,018; 12,920 N. China -> 60,228; 68,998; 77.052; 102.501; 86,642 *M. China -> 1,004; 1,050; 733; 12.63!; 11,692 S. China -> 131; 115; 85.222; 71.659; (missing fifth?) Actually only four numbers: 131, 115, 85.222, 71.659. Then next is "Cuba". So S. China might have only 4 numbers? But should have 5. Maybe the fifth is missing or merged with Cuba. Cuba -> 2.677; 5; 1.336; 7; 1,086 C. America -> 9; 10; 50; 24; $1 Denmark -> 107; 487; 561; 217; 186 Egypt -> 19; 51; 52; 721; 124 France -> 1,611; 1.878; 2.718; 4,261; 3.638 Indo China -> 75,190; $3,463; 93,152; 50.681; 52.669 Germany -> 6,629; 9.957; 15.605; 37.560; 25.708 Holland -> 2.193; 1.462; 2,774; 4.513; 5.534 Italy -> 642; 1.476; 2.202; 8.655; 4.424 Japan -> 66,220; 58,739; 77.586; 68 303; 21.280 Kwong Chow Wan -> 9.910; 8.346; (then +Macao) -> 10.424; 8.878; (then Norway) -> 770; (then N. East Indies) -> 48,264; 849 68.266; (then Philippines) -> 6.520; (then Portugal) -> 3-4; (then Siam) -> 42.819; 4,904 62 67.978; 072 80.393 5.559; 683; 1.669; 79,950; 61.620; (then S. America) -> 1.162; 53; (then Sweden) -> 933; 986; 53.233 52 1.472; 1.722 192 47.120; 1.643; 148; 57,822; 190 2,599; 3,205; (then Switzerland) -> 475; 289; 337; 1.888; 1.118; (then Spain) -> 64; 99; 52; (then U. S. A.) -> (then Other) -> 49,815 649; 51.566; 50,281; 338 57,200; 68; 46.125; 1.250; 1.658; 2.201; 2.106; (then Total) -> 489.707; 545.200; 613 321; 737.740; 624.048; (then Total British Empire) -> 118.846; 116,727; 139 569; 131,932; 133,658; (then Total Foreign) -> 370.861; 428,473; 473.762; £:02,808; 490,390. This is extremely messy. The OCR has completely jumbled the rows after a certain point. It seems the original table had two sections: British Empire countries and Foreign countries. The OCR merged them. Given the complexity, perhaps the user expects me to output the table as best as I can, using the data that is clearly aligned for the first many rows, and then for the rest, maybe just note the corruption? But the instruction says to reconstruct the table using Markdown table syntax. I must produce a Markdown table. I think the best is to create a table with the columns: Country, 1922, 1923, 1924, 1931, 1932. And fill in the values for the countries that have clear 5 numbers. For the messy part, I'll try to reconstruct as many rows as possible. But the user provided the OCR text as a single block. They want the proofread text in Markdown. The original document is a table. So I should output a Markdown table. Given the difficulty, I'll attempt to parse the first 26 rows (up to Japan) which seem relatively clean. Then for the rest, I'll try to salvage. But note: The OCR includes "Kwong Chow Wan" and "+Macao" and "Norway" etc. The note says "+Included in South Chins 1922 1924." So Macao might be part of South China for early years. Also "N. East Indies" likely "Netherlands East Indies". The data for Siam, Sweden, etc. are garbled. Maybe the original table had more columns (like 1925-1930) but the OCR only captured some? The header only shows 5 years. The totals have 6 numbers for Total and Total British Empire, 5 for Total Foreign. That suggests maybe 6 years? Let's check the Total row: "489.707 545.200 613 321 737.740 624.048" - that's 6 numbers. Total British Empire: "118.846 116,727 139 569 131,932 133,658" - 6 numbers. Total Foreign: "370.861 428,473 473.762 £:02,808 490,390" - 5 numbers. Inconsistent. Perhaps the table has 6 year columns: 1922, 1923, 1924, 1930, 1931, 1932? But header shows 1922, 1923, 1924, 1931, 1932. Could be a misprint. Given the ambiguity, I'll assume 5 year columns as per header. For totals, I'll use the first 5 numbers for Total and Total British Empire, and all 5 for Total Foreign. But the Total row has 6 numbers. Which 5 correspond? Probably 1922, 1923, 1924, 1931, 1932. The sixth might be a total for something else. I'll take the first 5: 489.707, 545.200, 613, 321, 737.740? But "613 321" might be two numbers: 613 and 321? Actually "613 321" could be "613,321"? The OCR often misses commas. So "613 321" might be 613,321. Similarly "737.740" and "624.048". So the Total row might be: 489,707; 545,200; 613,321; 737,740; 624,048. That's 5 numbers if we combine 613 and 321. The OCR wrote "613 321" with a space. So likely it's 613,321. Similarly, Total British Empire: "118.846 116,727 139 569 131,932 133,658" - "139 569" might be 139,569. So 5 numbers: 118,846; 116,727; 139,569; 131,932; 133,658. Total Foreign: "370.861 428,473 473.762 £:02,808 490,390" - £:02,808 might be 602,808? Or 502,808? The "£:" is likely a misread of a number. Could be "602,808". But we'll keep as is. Thus, the table likely has 5 year columns. Good. Now, for each country, we need 5 numbers. The OCR for early countries gives 5 numbers each. For later countries, the numbers are jumbled. I'll write a Markdown table with the countries I can confidently parse. For the garbled ones, I'll either omit or try to approximate. But the instruction: "Preserve original content — do not add or remove substantive words." So I should include all rows, even if messy. But I must reconstruct the table. Perhaps I can keep the raw numbers as they appear in sequence for each country, but that would be misleading. Given the time, I'll produce a table with the clearly parsed rows (up to Japan) and then for the rest, I'll do my best to parse the remaining tokens into rows of 5 numbers each, using the country names as delimiters. Let's list all country names in order as they appear in the OCR: U. K. Australia Burma Canada Ceylon E. Africa India N. Zealand N. Borneo S. Africa Straits W. Africa W. Indies B. Empire, other Belgium N. China *M. China *S. China Cuba C. America Denmark Egypt France Indo China Germany Holland Italy Japan Kwong Chow Wan +Macao Norway N. East Indies Philippines Portugal Siam S. America Sweden Switzerland Spain U. S. A. Other Total Total British Empire Total Foreign That's 44 rows. The table likely had that many. Now, for each, we need 5 numbers. The OCR provides a stream of numbers. Let's extract all numbers in order after the header. I'll write a quick mental parser: after "Countries 1922 1923 1924 1931 1932 $ $ $", the tokens are: U. K. 55,529 61,812 80,328 78,251 76,905 Australia 5.903 5,960 8,860 6,288 12.045 Burma 13.243 6.559 8.936 6,999 6.256 Canada 2,286 3,978 6.238 5.123 6,236 Ceylon $8 172 179 278 213 E. Africa 295 610 404 397 412 India 26.381 19,552 16,989 17.583 17,649 N. Zealand 117 158 460 279 179 N. Borneo 2.271 2,323 2,356 4,503 3.589 S. Africa 14 310 100 149 509 Straits 12,461 15,294 14,387 15,197 9,112 W. Africa 15 W. Indies 9 3 34 6 B. Empire, other 234 466 332 sal 546 Belgium 3,366 4,477 5,923 15,018 12,920 N. China 60,228 68,998 77.052 102.501 86,642 *M. China 1,004 1,050 733 12.63! 11,692 *S. China 131 115 85.222 71.659 Cuba 2.677 5 1.336 7 1,086 C. America 9 10 50 24 $1 Denmark 107 487 561 217 186 Egypt 19 51 52 721 124 France 1,611 1.878 2.718 4,261 3.638 Indo China 75,190 $3,463 93,152 50.681 52.669 Germany 6,629 9.957 15.605 37.560 25.708 Holland 2.193 1.462 2,774 4.513 5.534 Italy 642 1.476 2.202 8.655 4.424 Japan 66,220 58,739 77.586 68 303 21.280 Kwong Chow Wan 9.910 8.346 +Macao 10.424 8.878 Norway 770 N. East Indies 48,264 849 68.266 Philippines 6.520 Portugal 3-4 Siam 42.819 4,904 62 67.978 072 80.393 5.559 683 1.669 79,950 61.620 S. America 1.162 53 Sweden 933 986 53.233 52 1.472 1.722 192 47.120 1.643 148 57,822 190 2,599 3,205 Switzerland 475 289 337 1.888 1.118 Spain 64 99 52 U. S. A. Other 49,815 649 51.566 50,281 338 57,200 68 46.125 1.250 1.658 2.201 2.106 Total 489.707 545.200 613 321 737.740 624.048 Total British Empire 118.846 116,727 139 569 131,932 133,658 Total Foreign 370.861 428,473 473.762 £:02,808 490,390 Now, we need to assign 5 numbers to each country. For the first 28 countries (up to Japan), each has exactly 5 numbers. Good. For Kwong Chow Wan: it has two numbers (9.910, 8.346) then the next token is "+Macao" which is a country name. So Kwong Chow Wan only has 2 numbers? But should have 5. Maybe the OCR missed the rest. Or maybe Kwong Chow Wan and Macao are combined? The note says "+Included in South Chins 1922 1924." So perhaps for 1922-1924, Macao is included in South China, and for 1931-1932, it's separate? But the table only has 5 years: 1922,1923,1924,1931,1932. So maybe Kwong Chow Wan has data for 1922,1923? And Macao for 1931,1932? But we have only 2 numbers for Kwong Chow Wan and 2 for Macao. That would make 4 numbers. Missing one. Norway: only one number "770". Then N. East Indies: "48,264" and "849 68.266" (maybe two numbers: 849 and 68,266). Philippines: "6.520" only one. Portugal: "3-4" maybe one number. Siam: many numbers. S. America: two numbers. Sweden: many. Switzerland: five numbers (475,289,337,1.888,1.118) - good. Spain: three numbers (64,99,52). U. S. A.: zero numbers? Then Other: many numbers. Total: five numbers (if we combine 613 321). Total British Empire: five numbers (combine 139 569). Total Foreign: five numbers. So many rows are incomplete. Given the difficulty, perhaps the original table had more columns (like 11 years) and the OCR only captured fragments. But the header explicitly says 1922,1923,1924,1931,1932. So I'll stick to 5 columns. I think the best is to output a Markdown table with the rows that have 5 numbers, and for the others, fill with the available numbers and use "..." for missing. But the instruction: "Missing text — if a word or phrase is clearly missing due to OCR damage, insert ... in its place." So for missing numbers, I can insert "...". I'll construct the table with columns: Country, 1922, 1923, 1924, 1931, 1932. I'll go through each country and assign the first 5 numbers from the stream. For countries where the stream has fewer than 5 numbers before the next country, I'll pad with "...". But the stream is sequential. So I can simulate: iterate through the list of countries, for each, take the next 5 numbers from the number stream. But the number stream includes numbers that belong to the next country if the previous country has fewer. That's the problem. Better: The OCR output is essentially a list of lines. The original table had rows. The OCR read row by row but split each cell into a new line. So each row: country name, then 5 numbers. But the OCR put each on a new line. So the structure is: country, num1, num2, num3, num4, num5, next country, num1, num2, ... However, the OCR also inserted line breaks within numbers (like "68 303" split). But generally, the pattern holds for the first 28 rows. After Japan, the pattern breaks because the OCR might have merged two columns? Or the table had sub-rows? For example, "Kwong Chow Wan" might be a sub-row under China? But it's listed as a separate row. Given the time, I'll produce a table for the first 28 rows (up to Japan) which are clean, and then for the remaining rows, I'll do a best-effort reconstruction using the numbers that appear after each country name until the next country name. But the OCR doesn't have clear delimiters. Alternatively, I could output the raw data as a Markdown table with the exact numbers as they appear in the OCR for each country, but that would be messy. Given the user is a professional editor, they might want a cleaned-up version of the table as it appears in the original. Since I cannot see the original, I must infer. I'll assume the table has 5 year columns. I'll create rows for all 44 countries, filling in numbers where clear, and using "..." for missing. Let's manually assign for each country based on the number stream, assuming each country takes 5 numbers in order. But the number stream is not separated per country after Japan. However, we can use the country names as markers. The numbers between two country names belong to the first country. So we can split the number stream by country names. Let's list the country names in order as they appear in the OCR text. Then the numbers between country name i and country name i+1 belong to country i. For the last country, numbers until the end. From the raw text, the country names appear in this order (by scanning the text): U. K. Australia Burma Canada Ceylon E. Africa India N. Zealand N. Borneo S. Africa Straits W. Africa W. Indies B. Empire, other Belgium N. China *M. China *S. China Cuba C. America Denmark Egypt France Indo China Germany Holland Italy Japan Kwong Chow Wan +Macao Norway N. East Indies Philippines Portugal Siam S. America Sweden Switzerland Spain U. S. A. Other Total Total British Empire Total Foreign Now, extract the numbers between each. I'll go through the raw text and split. The raw text after header: "U. K. 55,529 61,812 80,328 78,251 76,905 Australia 5.903 5,960 8,860 6,288 12.045 Burma 13.243 6.559 8.936 6,999 6.256 Canada 2,286 3,978 6.238 5.123 6,236 Ceylon $8 172 179 278 213 E. Africa 295 610 404 397 412 India 26.381 19,552 16,989 17.583 17,649 N. Zealand 117 158 460 279 179 N. Borneo 2.271 2,323 2,356 4,503 3.589 S. Africa 14 310 100 149 509 Straits 12,461 15,294 14,387 15,197 9,112 W. Africa 15 W. Indies 9 3 34 6 B. Empire, other 234 466 332 sal 546 Belgium 3,366 4,477 5,923 15,018 12,920 N. China 60,228 68,998 77.052 102.501 86,642 *M. China 1,004 1,050 733 12.63! 11,692 *S. China 131 115 85.222 71.659 Cuba 2.677 5 1.336 7 1,086 C. America 9 10 50 24 $1 Denmark 107 487 561 217 186 Egypt 19 51 52 721 124 France 1,611 1.878 2.718 4,261 3.638 Indo China 75,190 $3,463 93,152 50.681 52.669 Germany 6,629 9.957 15.605 37.560 25.708 Holland 2.193 1.462 2,774 4.513 5.534 Italy 642 1.476 2.202 8.655 4.424 Japan 66,220 58,739 77.586 68 303 21.280 Kwong Chow Wan 9.910 8.346 +Macao 10.424 8.878 Norway 770 N. East Indies 48,264 849 68.266 Philippines 6.520 Portugal 3-4 Siam 42.819 4,904 62 67.978 072 80.393 5.559 683 1.669 79,950 61.620 S. America 1.162 53 Sweden 933 986 53.233 52 1.472 1.722 192 47.120 1.643 148 57,822 190 2,599 3,205 Switzerland 475 289 337 1.888 1.118 Spain 64 99 52 U. S. A. Other 49,815 649 51.566 50,281 338 57,200 68 46.125 1.250 1.658 2.201 2.106 Total 489.707 545.200 613 321 737.740 624.048 Total British Empire 118.846 116,727 139 569 131,932 133,658 Total Foreign 370.861 428,473 473.762 £:02,808 490,390" Now, for each country, the numbers are the lines until the next country name. But note that some numbers are on the same line (e.g., "68 303", "849 68.266", "4,904 62 67.978", etc.). We'll split by whitespace. Let's parse each country's numbers: U. K.: 55,529; 61,812; 80,328; 78,251; 76,905 -> 5 numbers. Australia: 5.903; 5,960; 8,860; 6,288; 12.045 -> 5. Burma: 13.243; 6.559; 8.936; 6,999; 6.256 -> 5. Canada: 2,286; 3,978; 6.238; 5.123; 6,236 -> 5. Ceylon: $8; 172; 179; 278; 213 -> 5. E. Africa: 295; 610; 404; 397; 412 -> 5. India: 26.381; 19,552; 16,989; 17.583; 17,649 -> 5. N. Zealand: 117; 158; 460; 279; 179 -> 5. N. Borneo: 2.271; 2,323; 2,356; 4,503; 3.589 -> 5. S. Africa: 14; 310; 100; 149; 509 -> 5. Straits: 12,461; 15,294; 14,387; 15,197; 9,112 -> 5. W. Africa: only "15" -> 1 number. W. Indies: 9; 3; 34; 6 -> 4 numbers. B. Empire, other: 234; 466; 332; sal; 546 -> 5. Belgium: 3,366; 4,477; 5,923; 15,018; 12,920 -> 5. N. China: 60,228; 68,998; 77.052; 102.501; 86,642 -> 5. *M. China: 1,004; 1,050; 733; 12.63!; 11,692 -> 5. *S. China: 131; 115; 85.222; 71.659 -> 4 numbers. Cuba: 2.677; 5; 1.336; 7; 1,086 -> 5. C. America: 9; 10; 50; 24; $1 -> 5. Denmark: 107; 487; 561; 217; 186 -> 5. Egypt: 19; 51; 52; 721; 124 -> 5. France: 1,611; 1.878; 2.718; 4,261; 3.638 -> 5. Indo China: 75,190; $3,463; 93,152; 50.681; 52.669 -> 5. Germany: 6,629; 9.957; 15.605; 37.560; 25.708 -> 5. Holland: 2.193; 1.462; 2,774; 4.513; 5.534 -> 5. Italy: 642; 1.476; 2.202; 8.655; 4.424 -> 5. Japan: 66,220; 58,739; 77.586; 68 303; 21.280 -> 5 (but "68 303" is two numbers? Actually it's one number 68,303 split. We'll treat as 68,303). Kwong Chow Wan: 9.910; 8.346 -> 2 numbers. +Macao: 10.424; 8.878 -> 2 numbers. Norway: 770 -> 1 number. N. East Indies: 48,264; 849 68.266 -> split: 48,264; 849; 68.266 -> 3 numbers. Philippines: 6.520 -> 1 number. Portugal: 3-4 -> 1 number (maybe 3-4 is a range? but likely a number 34? or 3-4? We'll keep as "3-4"). Siam: 42.819; 4,904 62 67.978; 072 80.393 5.559; 683; 1.669; 79,950; 61.620 -> many numbers. Let's split: 42.819; 4,904; 62; 67.978; 072; 80.393; 5.559; 683; 1.669; 79,950; 61.620 -> 11 numbers. S. America: 1.162; 53 -> 2 numbers. Sweden: 933; 986; 53.233 52 1.472; 1.722 192 47.120; 1.643; 148; 57,822; 190 2,599; 3,205 -> split: 933; 986; 53.233; 52; 1.472; 1.722; 192; 47.120; 1.643; 148; 57,822; 190; 2,599; 3,205 -> 14 numbers. Switzerland: 475; 289; 337; 1.888; 1.118 -> 5 numbers. Spain: 64; 99; 52 -> 3 numbers. U. S. A.: (no numbers before next country "Other") -> 0 numbers. Other: 49,815 649; 51.566; 50,281; 338 57,200; 68; 46.125; 1.250; 1.658; 2.201; 2.106 -> split: 49,815; 649; 51.566; 50,281; 338; 57,200; 68; 46.125; 1.250; 1.658; 2.201; 2.106 -> 12 numbers. Total: 489.707; 545.200; 613 321; 737.740; 624.048 -> split: 489.707; 545.200; 613; 321; 737.740; 624.048 -> 6 numbers. Total British Empire: 118.846; 116,727; 139 569; 131,932; 133,658 -> split: 118.846; 116,727; 139; 569; 131,932; 133,658 -> 6 numbers. Total Foreign: 370.861; 428,473; 473.762; £:02,808; 490,390 -> 5 numbers. Now, the table should have 5 year columns. For each country, we need to map their numbers to the 5 years. For countries with exactly 5 numbers, it's straightforward. For those with fewer, we have missing data. For those with more, maybe the table has more columns? But the header says only 5 years. Could be that the table includes 1922-1932 inclusive (11 years) but the header only shows 5? The title says "1922-1932". The header shows "1922 1923 1924 1931 1932". That's 5 years. Maybe the table only shows those 5 years. So each country should have 5 numbers. Thus, for countries with more than 5 numbers, the extra numbers might be for other years not shown in header? Or they are misaligned. For Siam, Sweden, Other, they have many numbers. Perhaps those numbers belong to multiple rows? But the OCR didn't separate rows. Given the difficulty, I'll assume the table has 5 columns. For each country, I'll take the first 5 numbers from their list. If they have fewer, I'll pad with "...". For countries with more, I'll only take first 5. But for W. Africa, only 1 number; W. Indies 4; *S. China 4; Kwong Chow Wan 2; +Macao 2; Norway 1; N. East Indies 3; Philippines 1; Portugal 1; Siam 11; S. America 2; Sweden 14; Spain 3; U.S.A. 0; Other 12. This will result in many "..." which might be acceptable. But maybe the table originally had more columns (like 11 years) and the OCR header only captured 5. The totals have 6 numbers for Total and Total British Empire, 5 for Total Foreign. That suggests maybe 6 years? Let's check the Total row: 489.707, 545.200, 613, 321, 737.740, 624.048. If we combine 613 and 321 as 613,321, then we have 5 numbers: 489,707; 545,200; 613,321; 737,740; 624,048. That matches 5 years. Similarly Total British Empire: 118,846; 116,727; 139,569; 131,932; 133,658. Total Foreign: 370,861; 428,473; 473,762; 602,808?; 490,390. So 5 years. Thus, 5 year columns. Now, for each country, we need to decide which 5 numbers correspond to the 5 years. For the first 28 countries, they have exactly 5 numbers, so that's fine. For the rest, the numbers are likely also 5 per country, but the OCR merged multiple countries' numbers because of missing country names? But we have country names. The numbers between country names are for that country. For Kwong Chow Wan, only 2 numbers. Maybe the other 3 are missing (not recorded). For +Macao, 2 numbers. For Norway, 1. For N. East Indies, 3. For Philippines, 1. For Portugal, 1. For Siam, many. For S. America, 2. For Sweden, many. For Spain, 3. For U.S.A., 0. For Other, many. This suggests that the OCR for the latter part is completely garbled, possibly because the original table had a different layout (maybe two pages, or columns). The user might not expect perfect reconstruction for the garbled part. But they said "reconstruct the table using Markdown table syntax." I'll do my best. Given the context, this is a historical trade statistics table. The first part (British Empire) is clear. The foreign countries part is also clear up to Japan. Then there are some additional entries like Kwong Chow Wan, Macao, Norway, Netherlands East Indies, Philippines, Portugal, Siam, South America, Sweden, Switzerland, Spain, USA, Other. These are likely also part of the table. The numbers for them might be in the original but OCR messed up. I could try to reconstruct by assuming each of those countries has 5 numbers, and the numbers are in the stream in order. But the stream has many numbers for Siam and Sweden, which might be because the OCR combined multiple rows (maybe Siam has data for multiple years? But only 5 years). Or maybe the table has sub-columns for each year (like imports and exports)? But the title says "VALUES OF IMPORTS AND EXPORTS (By Countries) 1922-1932." This table is only for IMPORTS (as per "IMPORTS."). There might be a separate table for EXPORTS. So this is just imports. Thus, each country has one value per year (in $000). So 5 values. I think the best is to output a table with the 28 clean rows, and then for the remaining rows, include them with the numbers they have, and use "..." for missing. But the user might want the full table. Given the instruction "Preserve original content — do not add or remove substantive words." I should include all rows. I'll create a Markdown table with all 44 rows, filling in the numbers as they appear in the OCR for each country, but limited to 5 columns. For countries with more than 5 numbers, I'll take the first 5. For fewer, pad with "...". But for Siam and Sweden, they have many numbers. Which 5 are correct? Hard to know. Maybe the OCR combined several countries' numbers because the country names were not recognized. For example, after Portugal, the next country is Siam, but the numbers for Siam might include numbers for other countries that lost their names. However, the country names are present: Siam, S. America, Sweden, Switzerland, Spain, U.S.A., Other. So the numbers between them should belong to each. But the numbers for Siam are many, meaning the OCR might have inserted extra line breaks within Siam's numbers? Actually, the numbers for Siam are on multiple lines: "42.819", "4,904 62 67.978", "072 80.393 5.559", "683", "1.669", "79,950", "61.620". That's 7 lines, but each line may have multiple numbers. That's too many for 5 years. Maybe the table has more than 5 year columns? Let's check the header again: "Countries 1922 1923 1924 1931 1932 $ $ $". The "$ $ $" might indicate three more columns? Or they are just artifacts. Could be that the table has 8 columns: 1922, 1923, 1924, 1925, 1926, 1927, 1928, 1929, 1930, 1931, 1932? But the header only shows 5. Given the time, I'll assume 5 columns. For the messy rows, I'll take the first 5 numbers from their list. For Siam: 42.819, 4,904, 62, 67.978, 072. For Sweden: 933, 986, 53.233, 52, 1.472. For Other: 49,815, 649, 51.566, 50,281, 338. For Total: 489.707, 545.200, 613, 321, 737.740 (but 613 and 321 should be combined? I'll keep as separate columns? That would be 6 columns. I'll combine 613 and 321 as 613,321 for the third column? But the table has 5 columns. The Total row has 6 numbers if we don't combine. The header has 5 years. So likely the third year (1924) value is 613,321. The OCR split it. So I'll combine them. Similarly for Total British Empire: combine 139 and 569 as 139,569. For Total Foreign: £:02,808 maybe 602,808. I'll clean up numbers: remove spaces, fix obvious OCR errors (like "sal" -> maybe "381"? but I'll keep as "sal" or replace with "..."? The rule: correct unambiguous OCR spelling errors. "sal" is not a number, likely a misread of "381" or "541". But not unambiguous. I'll keep as "sal" but maybe it's "381"? I'll keep as "sal" to preserve original. Also "12.63!" -> likely "12,631"? Or "12.631"? I'll keep as "12.63!". "68 303" -> "68,303". "849 68.266" -> two numbers: 849 and 68,266. But for N. East Indies, we have 3 numbers: 48,264; 849; 68,266. That's 3 numbers for 5 years. Missing two. "3-4" for Portugal -> maybe "34"? Or "3-4" is a range. I'll keep as "3-4". "£:02,808" -> maybe "602,808". I'll keep as "£:02,808". Now, I'll construct the Markdown table. First, the header row: | Country | 1922 | 1923 | 1924 | 1931 | 1932 | But the original header has "Countries" and then years. I'll use "Country" as column header. Also note the title: "VALUES OF IMPORTS AND EXPORTS (By Countries) 1922-1932." and "IMPORTS." and "$000's omitted." I'll include that as text before the table. Also the footnotes: "*Not fully recorded until 193)." and "+Included in South Chins 1922 1924." I'll include after the table. Now, the table rows. I'll write the Markdown. Let's list rows with 5 values each. U. K. | 55,529 | 61,812 | 80,328 | 78,251 | 76,905 Australia | 5,903 | 5,960 | 8,860 | 6,288 | 12,045 (note: 5.903 -> 5,903; 12.045 -> 12,045) Burma | 13,243 | 6,559 | 8,936 | 6,999 | 6,256 Canada | 2,286 | 3,978 | 6,238 | 5,123 | 6,236 Ceylon | 8 | 172 | 179 | 278 | 213 (remove $) E. Africa | 295 | 610 | 404 | 397 | 412 India | 26,381 | 19,552 | 16,989 | 17,583 | 17,649 N. Zealand | 117 | 158 | 460 | 279 | 179 N. Borneo | 2,271 | 2,323 | 2,356 | 4,503 | 3,589 S. Africa | 14 | 310 | 100 | 149 | 509 Straits | 12,461 | 15,294 | 14,387 | 15,197 | 9,112 W. Africa | 15 | ... | ... | ... | ... W. Indies | 9 | 3 | 34 | 6 | ... B. Empire, other | 234 | 466 | 332 | sal | 546 Belgium | 3,366 | 4,477 | 5,923 | 15,018 | 12,920 N. China | 60,228 | 68,998 | 77,052 | 102,501 | 86,642 *M. China | 1,004 | 1,050 | 733 | 12.63! | 11,692 *S. China | 131 | 115 | 85,222 | 71,659 | ... Cuba | 2,677 | 5 | 1,336 | 7 | 1,086 C. America | 9 | 10 | 50 | 24 | 1 (remove $) Denmark | 107 | 487 | 561 | 217 | 186 Egypt | 19 | 51 | 52 | 721 | 124 France | 1,611 | 1,878 | 2,718 | 4,261 | 3,638 Indo China | 75,190 | 3,463 | 93,152 | 50,681 | 52,669 (remove $ from second) Germany | 6,629 | 9,957 | 15,605 | 37,560 | 25,708 Holland | 2,193 | 1,462 | 2,774 | 4,513 | 5,534 Italy | 642 | 1,476 | 2,202 | 8,655 | 4,424 Japan | 66,220 | 58,739 | 77,586 | 68,303 | 21,280 Kwong Chow Wan | 9,910 | 8,346 | ... | ... | ... +Macao | 10,424 | 8,878 | ... | ... | ... Norway | 770 | ... | ... | ... | ... N. East Indies | 48,264 | 849 | 68,266 | ... | ... Philippines | 6,520 | ... | ... | ... | ... Portugal | 3-4 | ... | ... | ... | ... Siam | 42,819 | 4,904 | 62 | 67,978 | 72 (072 -> 72? but 072 is 72) Actually first five: 42.819, 4,904, 62, 67.978, 072. I'll use 42,819; 4,904; 62; 67,978; 72. S. America | 1,162 | 53 | ... | ... | ... Sweden | 933 | 986 | 53,233 | 52 | 1,472 Switzerland | 475 | 289 | 337 | 1,888 | 1,118 Spain | 64 | 99 | 52 | ... | ... U. S. A. | ... | ... | ... | ... | ... Other | 49,815 | 649 | 51,566 | 50,281 | 338 Total | 489,707 | 545,200 | 613,321 | 737,740 | 624,048 Total British Empire | 118,846 | 116,727 | 139,569 | 131,932 | 133,658 Total Foreign | 370,861 | 428,473 | 473,762 | 602,808 | 490,390 (assuming £:02,808 -> 602,808) I'll keep the numbers as they appear in OCR but with commas for thousands? The OCR uses commas and periods inconsistently. I'll standardize to commas for thousands and periods for decimals? But the values are in $000's, so they are whole numbers (thousands of dollars). The decimals are likely OCR errors for commas. For example, "5.903" is 5,903. "13.243" is 13,243. "2.271" is 2,271. "26.381" is 26,381. "77.052" is 77,052. "102.501" is 102,501. "12.63!" is 12,631? "85.222" is 85,222. "71.659" is 71,659. "2.677" is 2,677. "1.336" is 1,336. "75,190" already has comma. "$3,463" is 3,463. "93,152" okay. "50.681" is 50,681. "52.669" is 52,669. "9.957" is 9,957. "15.605" is 15,605. "37.560" is 37,560. "25.708" is 25,708. "2.193" is 2,193. "1.462" is 1,462. "4.513" is 4,513. "5.534" is 5,534. "1.476" is 1,476. "2.202" is 2,202. "8.655" is 8,655. "4.424" is 4,424. "77.586" is 77,586. "68 303" is 68,303. "21.280" is 21,280. "9.910" is 9,910. "8.346" is 8,346. "10.424" is 10,424. "8.878" is 8,878. "48,264" okay. "849" okay. "68.266" is 68,266. "6.520" is 6,520. "42.819" is 42,819. "4,904" okay. "62" okay. "67.978" is 67,978. "072" is 72. "80.393" is 80,393. "5.559" is 5,559. "683" okay. "1.669" is 1,669. "79,950" okay. "61.620" is 61,620. "1.162" is 1,162. "53" okay. "933" okay. "986" okay. "53.233" is 53,233. "52" okay. "1.472" is 1,472. "1.722" is 1,722. "192" okay. "47.120" is 47,120. "1.643" is 1,643. "148" okay. "57,822" okay. "190" okay. "2,599" okay. "3,205" okay. "475" okay. "289" okay. "337" okay. "1.888" is 1,888. "1.118" is 1,118. "64" okay. "99" okay. "52" okay. "49,815" okay. "649" okay. "51.566" is 51,566. "50,281" okay. "338" okay. "57,200" okay. "68" okay. "46.125" is 46,125. "1.250" is 1,250. "1.658" is 1,658. "2.201" is 2,201. "2.106" is 2,106. "489.707" is 489,707. "545.200" is 545,200. "613 321" is 613,321. "737.740" is 737,740. "624.048" is 624,048. "118.846" is 118,846. "116,727" okay. "139 569" is 139,569. "131,932" okay. "133,658" okay. "370.861" is 370,861. "428,473" okay. "473.762" is 473,762. "£:02,808" -> 602,808? I'll use
Baseline (Original)
(S5) VALUES OF IMPORTS AND EXPORTS (By Countries) 1922-1932. IMPORTS. $000's omitted. Countries 1922 1923 1924 1931 1932 $ $ $ U. K. 55,529 61,812 80,328 78,251 76,905 Australia 5.903 5,960 8,860 6,288 12.045 Burma 13.243 6.559 8.936 6,999 6.256 Canada 2,286 3,978 6.238 5.123 6,236 Ceylon $8 172 179 278 213 E. Africa 295 610 404 397 412 India 26.381 19,552 16,989 17.583 17,649 N. Zealand 117 158 460 279 179 N. Borneo 2.271 2,323 2,356 4,503 3.589 S. Africa 14 310 100 149 509 Straits 12,461 15,294 14,387 15,197 9,112 W. Africa 15 W. Indies 9 3 34 6 B. Empire, other 234 466 332 sal 546 Belgium 3,366 4,477 5,923 15,018 12,920 N. China 60,228 68,998 77.052 102.501 86,642 *M. China 1,004 1,050 733 12.63! 11,692 *S. China 131 115 85.222 71.659 Cuba 2.677 5 1.336 7 1,086 C. America 9 10 50 24 $1 Denmark 107 487 561 217 186 Egypt 19 51 52 721 124 France 1,611 1.878 2.718 4,261 3.638 Indo China 75,190 $3,463 93,152 50.681 52.669 Germany 6,629 9.957 15.605 37.560 25.708 Holland 2.193 1.462 2,774 4.513 5.534 Italy 642 1.476 2.202 8.655 4.424 Japan 66,220 58,739 77.586 68 303 21.280 Kwong Chow Wan 9.910 8.346 +Macao 10.424 8.878 Norway 770 N. East Indies 48,264 849 68.266 Philippines 6.520 Portugal 3-4 Siam 42.819 4,904 62 67.978 072 80.393 5.559 683 1.669 79,950 61.620 S. America 1.162 53 Sweden 933 986 53.233 52 1.472 1.722 192 47.120 1.643 148 57,822 190 2,599 3,205 Switzerland 475 289 337 1.888 1.118 Spain 64 99 52 U. S. A. Other 49,815 649 51.566 50,281 338 57,200 68 46.125 1.250 1.658 2.201 2.106 Total 489.707 545.200 613 321 737.740 624.048 Total British Empire 118.846 116,727 139 569 131,932 133,658 Total Foreign 370.861 428,473 473.762 £:02,808 490,390 *Not fully recorded until 193). +Included in South Chins 1922 1924. 499
2026-07-15 07:46:53 · Baseline
View content

(S5)

VALUES OF IMPORTS AND EXPORTS

(By Countries) 1922-1932.

IMPORTS.

$000's omitted.

Countries

1922

1923

1924

1931

1932

$

$

$

U. K.

55,529

61,812

80,328

78,251

76,905

Australia

5.903

5,960

8,860

6,288

12.045

Burma

13.243

6.559

8.936

6,999

6.256

Canada

2,286

3,978

6.238

5.123

6,236

Ceylon

$8

172

179

278

213

E. Africa

295

610

404

397

412

India

26.381

19,552

16,989

17.583

17,649

N. Zealand

117

158

460

279

179

N. Borneo

2.271

2,323

2,356

4,503

3.589

S. Africa

14

310

100

149

509

Straits

12,461

15,294

14,387

15,197

9,112

W. Africa

15

W. Indies

9

3

34

6

B. Empire, other

234

466

332

sal

546

Belgium

3,366

4,477

5,923

15,018

12,920

N. China

60,228

68,998

77.052

102.501

86,642

*M. China

1,004

1,050

733

12.63!

11,692

*S. China

131

115

85.222

71.659

Cuba

2.677

5

1.336

7

1,086

C. America

9

10

50

24

$1

Denmark

107

487

561

217

186

Egypt

19

51

52

721

124

France

1,611

1.878

2.718

4,261

3.638

Indo China

75,190

$3,463

93,152

50.681

52.669

Germany

6,629

9.957

15.605

37.560

25.708

Holland

2.193

1.462

2,774

4.513

5.534

Italy

642

1.476

2.202

8.655

4.424

Japan

66,220

58,739

77.586

68 303

21.280

Kwong Chow Wan

9.910

8.346

+Macao

10.424

8.878

Norway

770

N. East Indies

48,264

849 68.266

Philippines

6.520

Portugal

3-4

Siam

42.819

4,904 62 67.978

072 80.393 5.559

683

1.669

79,950

61.620

S. America

1.162

53

Sweden

933

986

53.233 52 1.472

1.722 192 47.120

1.643

148

57,822

190 2,599

3,205

Switzerland

475

289

337

1.888

1.118

Spain

64

99

52

U. S. A.

Other

49,815 649

51.566

50,281

338 57,200

68

46.125

1.250

1.658

2.201

2.106

Total

489.707

545.200

613 321

737.740

624.048

Total British Empire

118.846

116,727

139 569

131,932

133,658

Total Foreign

370.861

428,473

473.762

£:02,808 490,390

*Not fully recorded until 193).

+Included in South Chins 1922 1924.

499

Comments

Approved members can add comments, bookmarks, and private notes.

No comments yet.

Private Research Note

Private notes are available after approval.