The user wants me to proofread OCR output of a historical record relating to Hong Kong, specifically a table of staff appointments for the Kowloon-Canton Railway Traffic Staff. The OCR text is messy with broken lines, misaligned columns, and garbled text. I need to reconstruct the table in Markdown format.
First, I need to understand the structure. The header row appears to be:
OFFICE. Date NAME. of Appointment, Authority. Annual Salary. KOWLOON-CANTON RAILWAY,—Continued, TRAFFIC STAFF,—Continued. House or Quarters, and Allowances for Rout, Entertainment, Personal, or for any other purpose, Absence from the Colony during 1928. Date of First Appointment.
But the OCR has broken lines. The actual data rows follow. There are multiple entries with columns: Office, Date of Appointment, Name, Authority, Annual Salary, House/Quarters/Allowances, Absence from Colony during 1928, Date of First Appointment.
I need to parse the OCR lines and align them into rows. The OCR seems to have lost column alignment. Let's list the lines as they appear:
This is a mess. I need to reconstruct the table. The header suggests columns: Office, Date of Appointment, Name, Authority, Annual Salary, House/Quarters/Allowances, Absence from Colony during 1928, Date of First Appointment.
But the OCR seems to have multiple entries. Let's try to group by entries. The first entry appears to be "Booking Clerk." with "Li Yeung-shan." and "Fan Kum-fook."? Actually, "Booking Clerk." appears, then "Do.." (ditto) for next rows? The "Do.." likely means same office. So there are multiple booking clerks.
Let's parse systematically. The OCR likely came from a table with rows. Each row has 8 columns. The OCR reading order might be column by column? But the text seems to be read row by row but with line breaks within cells.
Better approach: Look for patterns. The "Date of Appointment" column seems to have dates like "1st January, 1927", "6th December, 1927", "1st January, 1926", "1st May, 1925", "1st October, 1920", "16th June, 1914", "1st March, 1926", "1st October, 1920", "1st January, 1922", "25th November, 1919", "14th November, 1919", "25th November, 1919", "1st October, 1923", "1st April, 1925", "17th January, 1928", "1st November, 1924", "10th November, 1928", "10th December, 1928", "1st January, 1915", "1st May, 1918", "1st October, 1910", "1st October, 1920".
The "Name" column: "Li Yeung-shan", "Fan Kum-fook", "Lai C'hi-ki", "Lo Yan-shau", "Cheng Chi-kwan", "Kan U-wah", "Cheng Chung-hang", "Wan Bau-kong", "Chan Yuk-slan", "Lam Han-cheong", "Lan Kam-fni", "Jp Kwan-yan", "Yong Ping-shan", "Yeung King-yan", "Ynen He-ting", "Lenny Wali-ip", "Ng Kau", "Li Pui", "Chung Hing", "Tong Yuk".
The "Authority" column: "No. 11 in 100 of 1920", "Do.", "No. 1219 of 1922", "No. 1 in 1125 of 1916", "No. 1091 of 1915", "The Manager", "Do.", "Do.", "Do.", "(1)", "(2)".
The "Annual Salary" column: "8625", "625", "625", "825", "650", "1,150", "750", "750", "825", "H25" (maybe 825?), "750", "525", "525", "288", "288", "288", "288", "523", "525", "526".
The "House/Quarters/Allowances" column: "Quarters", "Do.", "Do.", "Quarters", "Quarters", "$180 Rent Allowance", "Quarters", "Do.", "Do.", "Do.", "Do.", "Do.", "Quarters", "Do.", "Do.", "Do.", "Do.", "Quarters", "Do.", "Do.".
The "Absence from Colony during 1928" column: "10 days", "12 days", "7 days", "8 days", "7 days", "8 days", "6 days", "3 days", "10 days", "9 days", "7 days", "10 days", "10 days", "10 days", "10 days", "10 days", "10 days", "10 days", "10 days", "10 days".
The "Date of First Appointment" column: "7th December, 1925", "1st January, 1926", "1st May, 1925", "1st October, 1920", "13th October, 1914", "22nd July, 1920", "1st January, 1922", "9th November, 1912", "14th April, 1919", "14th November, 1919", "25th November, 1919", "1st October, 1913", "1st February, 1916", "17th January, 1928", "1st November, 1924", "10th November, 1928", "10th December, 1928", "1st October, 1913", "1st February, 1916", "1st October, 1910".
But there are also some stray numbers like "198", "(J 102 )" at top. That might be page number and reference. "198" could be page number. "(J 102 )" could be a file reference.
Also there is "KOWLOON-CANTON RAILWAY,—Continued, TRAFFIC STAFF,—Continued." as a header.
Now, I need to produce a Markdown table. I'll create a table with columns: Office, Date of Appointment, Name, Authority, Annual Salary, House/Quarters/Allowances, Absence from Colony (1928), Date of First Appointment.
I'll need to match each row. Let's list the rows in order as they appear in the OCR. The OCR seems to list multiple offices: Booking Clerk (multiple), Senior Goods Clerk, Goods Clerk, Relieving Goods Clerk, Signulman (Signalman?), Block Operator, Probationer, The Manager (maybe a title?), Shunter.
But the "The Manager" appears as an office? Actually "The Manager" appears in the line "The Manager." after "Yeung King-yan." and "17th January,"? Wait line 115: "Yeung King-yan." line 116: "The Manager." line 117: "288" line 118: "1928." That seems like a separate entry: Office: "The Manager"? But "The Manager" might be the authority for the previous entry? Let's see: "Yeung King-yan." "17th January," "The Manager." "288" "1928." "Do" "Ynen He-ting." ... This is confusing.
Maybe the table has a column "Authority" which sometimes says "The Manager". For example, for Yeung King-yan, the authority is "The Manager". Then the next entry "Ynen He-ting" has authority "Do." (ditto). Then "Lenny Wali-ip" authority "Do.", "Ng Kau" authority "Do.".
Then "Shunter" office, with "Li Pui" and "Chung Hing" and "Tong Yuk". The "(1)" and "(2)" might be footnote markers.
Let's reconstruct row by row based on the sequence of data. I'll write a script mentally, but better to manually parse.
I'll go through the lines and assign to columns based on expected order. The header order: Office, Date of Appointment, Name, Authority, Annual Salary, House/Quarters/Allowances, Absence from Colony during 1928, Date of First Appointment.
But the OCR might have read column by column? However, the lines seem to be in row-major order but with line breaks within cells. For example, "Booking Clerk." is office. Then "Do.." likely means same office for next rows. Then "Do.." again. Then "Do.," maybe another ditto. Then "Li Yeung-shan." name. Then "Fan Kum-fook." name. Then "1st January, 1927. Do." date of appointment and authority? Actually "1st January, 1927." is date of appointment for Li Yeung-shan? And "Do." might be authority (ditto). Then "No. 11 in 100 of 1920." authority for Fan Kum-fook? Then "8625" salary for Li? Then "Quarters." house. Then "10 days." absence. Then "7th December," "1925." date of first appointment for Li. Then "Do." maybe ditto for house? Then "625" salary for Fan? Then "Do." house. Then "12 days." absence. Then "Lai C'hi-ki." name. Then "Lo Yan-shau," name. Then "6th December," "Do." date of appointment? Then "625" salary. Then "Do." house. Then "7 days." absence. Then "1927." maybe year for date of first appointment? Then "1st January, 1926." date of first appointment for Lai? Then "1st May," "1925." date of first appointment for Lo? Then "1st October," "Do." date of appointment for next? Then "825" salary. Then "1920." year? Then "Do." authority. Then "Senior Goods Clerk," office. Then "Goods Clerk," office. Then "Relieving Goods Clerk," office. Then "Cheng Chi-kwan," name. Then "Kan U-wah," name. Then "Cheng Chung-hang." name. Then "16th June," "Do." date of appointment for Cheng Chi-kwan? Then "650" salary. Then "Quarters," house. Then "Do." house for next? Then "7 days." absence. Then "13th October," "1914." date of first appointment for Cheng Chi-kwan. Then "8 days." absence for Kan? Then "1st March," "1926." date of appointment for Kan? Then "1st October," "Do." date of appointment for Cheng Chung-hang? Then "1,150 | $180 Rent Allowance." salary and allowance. Then "6 days." absence. Then "22nd July," "1920." date of first appointment for Cheng Chung-hang. Then "1st January," "Do." date of appointment for next? Then "750" salary. Then "Quarters." house. Then "•" maybe bullet. Then "3 days." absence. Then "Wan Bau-kong." name. Then "1922. Do." date of appointment? Then "D.," maybe "Do."? Then "750" salary. Then "Do." house. Then "10 days." absence. Then "1925." year? Then "1912." year? Then "9th November." date? Then "1919. 14th April," date of first appointment for Wan? Then "Signulman," office (Signalman). Then "Do.," ditto. Then "Do.." ditto. Then "Chan Yuk-slan," name. Then "25th November." date of appointment? Then "No. 1091 of 1915." authority. Then "825" salary. Then "Do." house. Then "9 days." absence. Then "1919. 14th November," date of first appointment. Then "Lam Han-cheong," name. Then "1915." year? Then "Do." authority. Then "Do." house. Then "H25 |" maybe "825"? salary. Then "Do." house. Then "7 days." absence. Then "25th November," date of first appointment. Then "1" maybe a number. Then "Block Operator," office. Then "Don" maybe "Do."? Then "Probationer," office. Then "Lan Kam-fni." name. Then "Jp Kwan-yan." name. Then "1st October, 1923." date of appointment for Lan? Then "No. 1219 of 1922." authority. Then "750" salary. Then "Do." house. Then "10 days." absence. Then "Yong Ping-shan." name. Then "1st April, 1925." date of appointment. Then "Do." authority. Then "No. 1 in 1125 of 1916." authority? Then "525" salary. Then "10 Jays." absence (typo for days). Then "Do." house. Then "525" salary for next? Then "Quarters." house. Then "10 days." absence. Then "Yeung King-yan." name. Then "17th January," date of appointment. Then "The Manager." authority. Then "288" salary. Then "1928." year? Then "Do" authority for next. Then "Ynen He-ting." name. Then "1st November," date of appointment. Then "Do." authority. Then "288" salary. Then "1924." year? Then "Do" authority. Then "Lenny Wali-ip." name. Then "10th November," date of appointment. Then "Do." authority. Then "288" salary. Then "1928." year. Then "Do.." authority. Then "Ng Kau," name. Then "10th December," date of appointment. Then "Do." authority. Then "288" salary. Then "1928." year. Then "Shunter," office. Then "Do.," ditto. Then "Do.," ditto. Then "Li Pui." name. Then "(1)" footnote. Then "1st January," date of appointment. Then "Do." authority. Then "523" salary. Then "1915." year? Then "Chung Hing." name. Then "1st May, 1918." date of appointment. Then "Do." authority. Then "525" salary. Then "Quarters." house. Then "10 days." absence. Then "1914." year? Then "1915. 12th July," date of first appointment for Li Pui? Then "1921." year? Then "1st Angust, 1921." date of first appointment for Chung Hing? Then "1st December," date? Then "1919," year. Then "17th January, 1928." date of first appointment for Ng Kau? Then "1st November, 192H." date for Ynen? Then "10th November," date for Lenny? Then "1928." year. Then "10th December," date for Ng Kau? Then "1928," year. Then "1st October. 1913." date for Li Pui? Then "1st February," date for Chung Hing? Then "1916." year. Then "Tong Yuk." name. Then "(2)" footnote. Then "1st October," date of appointment. Then "Do." authority. Then "526" salary. Then "Do." house. Then "10 days." absence. Then "1st October," date of first appointment. Then "1910." year. Then "1920." year.
This is extremely messy. The OCR has mixed up the columns. The "Date of First Appointment" column seems to be at the end, but the OCR has many dates scattered.
Given the difficulty, perhaps the best is to reconstruct the table as best as possible, using the header and typical structure of such civil service lists. I'll create a Markdown table with the columns and fill rows based on logical grouping.
I'll assume each row corresponds to one staff member. The offices are: Booking Clerk (multiple), Senior Goods Clerk, Goods Clerk, Relieving Goods Clerk, Signalman (Signulman), Block Operator, Probationer, Shunter (multiple). Also "The Manager" might be an office but likely an authority.
Let's list staff members in order of appearance:
But there are also "The Manager" entries. Actually, Yeung King-yan, Ynen He-ting, Lenny Wali-ip, Ng Kau have authority "The Manager" and salary 288. They might be "Shunter" as well? But Shunter appears later with Li Pui, Chung Hing, Tong Yuk with salaries 523, 525, 526. So maybe Yeung King-yan etc. are "Probationer" or "Block Operator"? But Block Operator and Probationer are separate offices.
Let's check the OCR lines around Yeung King-yan: "Yeung King-yan." "17th January," "The Manager." "288" "1928." "Do" "Ynen He-ting." "1st November," "Do." "288" "1924." "Do" "Lenny Wali-ip." "10th November," "Do." "288" "1928." "Do.." "Ng Kau," "10th December," "Do." "288" "1928." Then "Shunter," "Do.," "Do.," "Li Pui." "(1)" "1st January," "Do." "523" "1915." "Chung Hing." "1st May, 1918." "Do." "525" "Quarters." "10 days." "1914." "1915. 12th July," "1921." "1st Angust, 1921." "1st December," "1919," "17th January, 1928." "1st November, 192H." "10th November," "1928." "10th December," "1928," "1st October. 1913." "1st February," "1916." "Tong Yuk." "(2)" "1st October," "Do." "526" "Do." "10 days." "1st October," "1910." "1920."
It seems the "Date of First Appointment" column for the Shunter entries are listed at the end in a block. The OCR has read the last column (Date of First Appointment) for all rows at the end? That is possible: the table might have the last column as "Date of First Appointment" and the OCR read it after all rows? But the lines show dates for earlier rows interspersed.
Actually, looking at the lines after "10 days." for Chung Hing, we see "1914." "1915. 12th July," "1921." "1st Angust, 1921." "1st December," "1919," "17th January, 1928." "1st November, 192H." "10th November," "1928." "10th December," "1928," "1st October. 1913." "1st February," "1916." These are likely the "Date of First Appointment" for the preceding staff members in order: Li Pui, Chung Hing, Tong Yuk, and maybe the previous four (Yeung King-yan, Ynen He-ting, Lenny Wali-ip, Ng Kau). But there are 7 dates for 7 staff? Let's count: Li Pui, Chung Hing, Tong Yuk = 3. Yeung King-yan, Ynen He-ting, Lenny Wali-ip, Ng Kau = 4. Total 7. The dates:
That's 10 dates. Hmm.
Maybe the table includes all staff from the beginning. The "Date of First Appointment" column might have been read in a separate block at the end of the table. The OCR output seems to have the columns interleaved.
Given the complexity, I think the best is to produce a Markdown table that captures the data as accurately as possible, using the header row and aligning each row with the available data. I'll need to make some educated guesses.
I'll structure the table with the following columns: Office, Date of Appointment, Name, Authority, Annual Salary, House/Quarters/Allowances, Absence from Colony (1928), Date of First Appointment.
I'll go through the OCR lines sequentially and assign to rows. Since the OCR likely reads row by row, but with line breaks, I'll try to parse each row as a set of 8 fields. However, the number of fields per row may vary because some cells contain multiple lines.
Alternative: The OCR might have read the table column by column. But the header line suggests the columns are in that order. The OCR text after header starts with "198 (J 102 )" which might be page number and reference. Then "Booking Clerk." which is the first office. Then "Do.." repeated for subsequent rows with same office. Then names, then dates, etc.
I'll attempt to reconstruct by simulating the table reading order: The table likely has multiple rows for each office. The first column "Office" may have "Booking Clerk" for first few rows, then "Senior Goods Clerk", "Goods Clerk", "Relieving Goods Clerk", "Signalman", "Block Operator", "Probationer", "Shunter". The "Date of Appointment" column has dates for each row. The "Name" column has names. The "Authority" column has references. The "Annual Salary" column has numbers. The "House/Quarters/Allowances" column has "Quarters" or "Do." or "$180 Rent Allowance". The "Absence from Colony during 1928" column has "X days". The "Date of First Appointment" column has dates.
I'll write a Markdown table with rows for each staff member. I'll use the data as they appear in the OCR, but I need to match them correctly.
Let's list the staff members in the order they appear in the OCR, with associated data.
From the OCR, the first office is "Booking Clerk." Then "Do.." appears three times, indicating four rows with office "Booking Clerk". The names: "Li Yeung-shan.", "Fan Kum-fook.", "Lai C'hi-ki.", "Lo Yan-shau,". So four booking clerks.
Then next office: "Senior Goods Clerk," then "Goods Clerk," then "Relieving Goods Clerk,". Three distinct offices, each one row? Names: "Cheng Chi-kwan,", "Kan U-wah,", "Cheng Chung-hang." So three rows.
Then next: "Signulman," "Do.," "Do.." three rows for Signalman. Names: "Chan Yuk-slan,", "Lam Han-cheong,". That's two names, but three rows? Maybe there are three signalmen: Chan Yuk-slan, Lam Han-cheong, and maybe another? The third "Do.." might be for a third signalman but name missing? Or "Lan Kam-fni." appears later under "Block Operator". So maybe only two signalmen.
Then "Block Operator," "Don" (maybe "Do."), "Probationer," - three offices? But "Block Operator" and "Probationer" are separate. Names: "Lan Kam-fni.", "Jp Kwan-yan.", "Yong Ping-shan." That's three names for three offices? But "Block Operator" and "Probationer" are two offices, plus maybe "Don" is "Do." for Block Operator? Actually "Don" might be a typo for "Do." meaning same office? But then "Probationer" is a different office. So maybe two Block Operators and one Probationer? But there are three names.
Then "Yeung King-yan." appears after "10 days." and "Quarters." and "525" etc. That seems like a new row. Then "The Manager." appears as authority? Then "Ynen He-ting.", "Lenny Wali-ip.", "Ng Kau." four names with authority "The Manager" and salary 288. Then "Shunter," "Do.," "Do.," three rows for Shunter. Names: "Li Pui.", "Chung Hing.", "Tong Yuk.".
So total rows: 4 (Booking Clerk) + 3 (Senior Goods Clerk, Goods Clerk, Relieving Goods Clerk) + 2 (Signalman) + 3 (Block Operator/Probationer) + 4 (The Manager? but office not clear) + 3 (Shunter) = 19 rows.
But the "The Manager" group might be "Shunter" as well? But they have different salary (288 vs 523). And they have authority "The Manager". The Shunter group has authority "(1)" and "(2)" and "Do.".
Maybe the "The Manager" group are "Probationer" or "Block Operator"? But Block Operator and Probationer already have names.
Let's examine the OCR lines for the Block Operator/Probationer section:
"1
Block Operator,
Don
Probationer,
Lan Kam-fni.
Jp Kwan-yan.
1st October, 1923.
No. 1219 of 1922.
750
Do.
10 days.
Yong Ping-shan.
1st April, 1925.
Do.
No. 1 in 1125 of 1916.
525
10 Jays.
Do.
525
Quarters.
10 days.
Yeung King-yan.
17th January,
The Manager.
288
1928.
Do
Ynen He-ting.
1st November,
Do.
288
1924.
Do
Lenny Wali-ip.
10th November,
Do.
288
1928.
Do..
Ng Kau,
10th December,
Do.
288
1928."
This suggests that after "Probationer," there are three names: Lan Kam-fni, Jp Kwan-yan, Yong Ping-shan. But there are two offices: Block Operator and Probationer. Maybe Lan Kam-fni is Block Operator, Jp Kwan-yan is also Block Operator (Don = Do.), and Yong Ping-shan is Probationer. Then Yeung King-yan is a new office? But "The Manager" appears as authority for Yeung King-yan. Then Ynen He-ting, Lenny Wali-ip, Ng Kau also have authority "The Manager". So maybe these four are "Shunter" but with lower salary? But then later "Shunter" appears with three names.
Perhaps the table continues from previous page and the "Shunter" group is separate. The "The Manager" group might be "Shunter" as well but appointed by the Manager? The authority column might be "The Manager" for those appointed by the Manager.
Given the ambiguity, I'll keep the offices as they appear in the OCR: "Booking Clerk", "Senior Goods Clerk", "Goods Clerk", "Relieving Goods Clerk", "Signalman", "Block Operator", "Probationer", "Shunter". For the four names with authority "The Manager", I'll assign office as "Shunter" as well? But they have different salary. Maybe they are "Junior Shunter" or "Shunter (Probationer)"? However, the OCR does not give an office for them explicitly. The office column might be blank (ditto) from previous? The previous office before Yeung King-yan is "Probationer" (for Yong Ping-shan). But Yeung King-yan appears after "10 days." and "Quarters." and "525" for Yong Ping-shan. So maybe Yeung King-yan is another Probationer? But then authority "The Manager" and salary 288.
Let's look at the line "Yeung King-yan." It comes after "525 Quarters. 10 days." That seems like the end of Yong Ping-shan's row. Then "Yeung King-yan." starts a new row. The next line "17th January," is likely Date of Appointment. Then "The Manager." is Authority. Then "288" Annual Salary. Then "1928." maybe House/Quarters? But 1928 is a year, not quarters. Then "Do" Authority for next? Actually "Do" might be House/Quarters? But "Do" usually means ditto for previous column. The columns: after Annual Salary comes House/Quarters/Allowances, then Absence, then Date of First Appointment. So "1928." could be House/Quarters? That doesn't make sense. Maybe "1928." is the year for Date of Appointment? But Date of Appointment already has "17th January," maybe "17th January, 1928." So "1928." is the year for that date. Then "Do" might be House/Quarters (ditto from previous row? Previous row had "Quarters"). Then "Ynen He-ting." Name. Then "1st November," Date of Appointment. Then "Do." Authority. Then "288" Salary. Then "1924." Year for Date of Appointment? Then "Do" House. Then "Lenny Wali-ip." Name. Then "10th November," Date. Then "Do." Authority. Then "288" Salary. Then "1928." Year. Then "Do.." House. Then "Ng Kau," Name. Then "10th December," Date. Then "Do." Authority. Then "288" Salary. Then "1928." Year.
Then "Shunter," Office. Then "Do.," Office ditto. Then "Do.," Office ditto. Then "Li Pui." Name. Then "(1)" maybe Authority? Then "1st January," Date. Then "Do." Authority. Then "523" Salary. Then "1915." Year? Then "Chung Hing." Name. Then "1st May, 1918." Date. Then "Do." Authority. Then "525" Salary. Then "Quarters." House. Then "10 days." Absence. Then "1914." Year? Then "1915. 12th July," Date of First Appointment for Li Pui? Then "1921." Year? Then "1st Angust, 1921." Date for Chung Hing? Then "1st December," Date? Then "1919," Year. Then "17th January, 1928." Date for Yeung King-yan? Then "1st November, 192H." Date for Ynen He-ting? Then "10th November," Date for Lenny Wali-ip? Then "1928." Year. Then "10th December," Date for Ng Kau? Then "1928," Year. Then "1st October. 1913." Date for Li Pui? Then "1st February," Date for Chung Hing? Then "1916." Year. Then "Tong Yuk." Name. Then "(2)" Authority. Then "1st October," Date. Then "Do." Authority. Then "526" Salary. Then "Do." House. Then "10 days." Absence. Then "1st October," Date of First Appointment. Then "1910." Year. Then "1920." Year.
This is extremely messy. The OCR has completely jumbled the last column (Date of First Appointment) for all rows at the end.
Given the instruction to proofread and reconstruct sentences and paragraphs, but this is a table. The instruction says: "If data is clearly in table format, reconstruct the table using Markdown table syntax." So I must produce a table.
I think the best approach is to produce a table with the columns as per header, and fill each row with the data that can be reasonably inferred. I'll need to make some assumptions.
Let's define the columns:
I'll create rows for each staff member in the order they appear in the OCR, using the data that seems to belong to them.
I'll go through the OCR text sequentially and parse row by row. Since the OCR is line-oriented, I'll try to group lines into rows based on the pattern: Office (maybe ditto), Date of Appointment, Name, Authority, Annual Salary, House/Quarters, Absence, Date of First Appointment.
But the OCR doesn't have clear delimiters. However, the original table likely had each row on a single line (or wrapped). The OCR has broken each cell into multiple lines.
Given the time, I'll produce a table with the data I can extract, leaving some cells blank if uncertain. But the instruction says "Preserve original content — do not add or remove substantive words." So I should not omit data. I need to include all data.
Maybe I can reconstruct the table as it appears in the OCR, but in a structured way. Since the OCR text is essentially a linearized version of the table, I could present it as a list of rows with the cells in order. But the columns are not aligned.
Another idea: The OCR might have read the table column by column. The header lists columns. The text after header might be column 1 (Office) for all rows, then column 2 (Date of Appointment) for all rows, etc. But the text doesn't look like that.
Let's check: After header, we have "198 (J 102 ) Booking Clerk. Do.. Do.. Do., Li Yeung-shan. Fan Kum-fook. 1st January, 1927. Do. No. 11 in 100 of 1920. 8625 Quarters. 10 days. 7th December, 1925. Do. 625 Do. 12 days. Lai C'hi-ki. Lo Yan-shau, 6th December, Do. 625 Do. 7 days. 1927. 1st January, 1926. 1st May, 1925. 1st October, Do. 825 1920. Do. Senior Goods Clerk, Goods Clerk, Relieving Goods Clerk, Cheng Chi-kwan, Kan U-wah, Cheng Chung-hang. 16th June, Do. 650 Quarters, Do. 7 days. 13th October, 1914. 8 days. 1st March, 1926. 1st October, Do. 1,150 | $180 Rent Allowance. 6 days. 22nd July, 1920. 1st January, Do. 750 Quarters. • 3 days. Wan Bau-kong. 1922. Do. D., 750 Do. 10 days. 1925. 1912. 9th November. 1919. 14th April, Signulman, Do., Do.. Chan Yuk-slan, 25th November. No. 1091 of 1915. 825 Do. 9 days. 1919. 14th November, Lam Han-cheong, 1915. Do. Do. H25 | Do. 7 days. 25th November, 1 Block Operator, Don Probationer, Lan Kam-fni. Jp Kwan-yan. 1st October, 1923. No. 1219 of 1922. 750 Do. 10 days. Yong Ping-shan. 1st April, 1925. Do. No. 1 in 1125 of 1916. 525 10 Jays. Do. 525 Quarters. 10 days. Yeung King-yan. 17th January, The Manager. 288 1928. Do Ynen He-ting. 1st November, Do. 288 1924. Do Lenny Wali-ip. 10th November, Do. 288 1928. Do.. Ng Kau, 10th December, Do. 288 1928. Shunter, Do., Do., Li Pui. (1) 1st January, Do. 523 1915. Chung Hing. 1st May, 1918. Do. 525 Quarters. 10 days. 1914. 1915. 12th July, 1921. 1st Angust, 1921. 1st December, 1919, 17th January, 1928. 1st November, 192H. 10th November, 1928. 10th December, 1928, 1st October. 1913. 1st February, 1916. Tong Yuk. (2) 1st October, Do. 526 Do. 10 days. 1st October, 1910. 1920."
If I consider that the table has 8 columns, and the OCR read row by row, but each cell may contain line breaks. The number of rows is unknown.
Given the difficulty, I'll produce a Markdown table that captures the data in a structured way, using the header and then each row as a set of cells. I'll try to match each piece of data to a column based on context.
I'll write a script in my mind to parse, but since I can't run code, I'll do it manually.
Let's list all the distinct data items in order and assign them to columns for each row.
I'll assume the table rows are in this order (based on office changes):
Row 1: Office: Booking Clerk, Date of Appointment: 1st January, 1927, Name: Li Yeung-shan, Authority: No. 11 in 100 of 1920, Annual Salary: 8625, House/Quarters: Quarters, Absence: 10 days, Date of First Appointment: 7th December, 1925
Row 2: Office: Booking Clerk (Do.), Date of Appointment: (maybe 6th December, 1927?), Name: Fan Kum-fook, Authority: Do. (or No. 11 in 100 of 1920?), Annual Salary: 625, House/Quarters: Do. (Quarters), Absence: 12 days, Date of First Appointment: 1st January, 1926
Row 3: Office: Booking Clerk (Do.), Date of Appointment: 6th December, 1927, Name: Lai C'hi-ki, Authority: Do., Annual Salary: 625, House/Quarters: Do., Absence: 7 days, Date of First Appointment: 1st May, 1925
Row 4: Office: Booking Clerk (Do.), Date of Appointment: 1st October, 1920?, Name: Lo Yan-shau, Authority: Do., Annual Salary: 825, House/Quarters: Do., Absence: ? (maybe 8 days?), Date of First Appointment: 1st October, 1920? But there is "1st October, Do. 825 1920. Do." That might be for Lo Yan-shau.
But then "Senior Goods Clerk" appears. So Row 5: Office: Senior Goods Clerk, Date of Appointment: 16th June, 1914?, Name: Cheng Chi-kwan, Authority: Do. (or No. 1219 of 1922?), Annual Salary: 650, House/Quarters: Quarters, Absence: 7 days, Date of First Appointment: 13th October, 1914
Row 6: Office: Goods Clerk, Date of Appointment: 1st March, 1926, Name: Kan U-wah, Authority: Do., Annual Salary: 1,150 | $180 Rent Allowance, House/Quarters: (maybe included in salary), Absence: 6 days, Date of First Appointment: 22nd July, 1920
Row 7: Office: Relieving Goods Clerk, Date of Appointment: 1st January, 1922, Name: Cheng Chung-hang, Authority: Do., Annual Salary: 750, House/Quarters: Quarters, Absence: 3 days, Date of First Appointment: 1st January, 1922? But there is "1st January, Do. 750 Quarters. • 3 days. Wan Bau-kong. 1922. Do. D., 750 Do. 10 days." This suggests Wan Bau-kong is another row? Maybe Wan Bau-kong is a Goods Clerk? But office not given.
Row 8: Office: (maybe Goods Clerk), Date of Appointment: 1922, Name: Wan Bau-kong, Authority: Do., Annual Salary: 750, House/Quarters: Do., Absence: 10 days, Date of First Appointment: 9th November, 1912
Row 9: Office: Signalman, Date of Appointment: 25th November, 1919, Name: Chan Yuk-slan, Authority: No. 1091 of 1915, Annual Salary: 825, House/Quarters: Do., Absence: 9 days, Date of First Appointment: 14th November, 1919
Row 10: Office: Signalman (Do.), Date of Appointment: 1915?, Name: Lam Han-cheong, Authority: Do., Annual Salary: 825 (H25), House/Quarters: Do., Absence: 7 days, Date of First Appointment: 25th November, 1919
Row 11: Office: Block Operator, Date of Appointment: 1st October, 1923, Name: Lan Kam-fni, Authority: No. 1219 of 1922, Annual Salary: 750, House/Quarters: Do., Absence: 10 days, Date of First Appointment: ? (maybe 1st October, 1913?)
Row 12: Office: Block Operator (Do.), Date of Appointment: 1st April, 1925, Name: Jp Kwan-yan, Authority: Do., Annual Salary: 525, House/Quarters: Do., Absence: 10 days, Date of First Appointment: 1st February, 1916?
Row 13: Office: Probationer, Date of Appointment: 1st April, 1925? Actually Yong Ping-shan: Date of Appointment: 1st April, 1925, Authority: No. 1 in 1125 of 1916, Annual Salary: 525, House/Quarters: Quarters, Absence: 10 days, Date of First Appointment: ?
Row 14: Office: (maybe Shunter?), Date of Appointment: 17th January, 1928, Name: Yeung King-yan, Authority: The Manager, Annual Salary: 288, House/Quarters: Do. (Quarters?), Absence: 10 days?, Date of First Appointment: 17th January, 1928?
Row 15: Office: (same), Date of Appointment: 1st November, 1924, Name: Ynen He-ting, Authority: Do., Annual Salary: 288, House/Quarters: Do., Absence: 10 days?, Date of First Appointment: 1st November, 1924
Row 16: Office: (same), Date of Appointment: 10th November, 1928, Name: Lenny Wali-ip, Authority: Do., Annual Salary: 288, House/Quarters: Do., Absence: 10 days?, Date of First Appointment: 10th November, 1928
Row 17: Office: (same), Date of Appointment: 10th December, 1928, Name: Ng Kau, Authority: Do., Annual Salary: 288, House/Quarters: Do., Absence: 10 days?, Date of First Appointment: 10th December, 1928
Row 18: Office: Shunter, Date of Appointment: 1st January, 1915, Name: Li Pui, Authority: (1), Annual Salary: 523, House/Quarters: Quarters, Absence: 10 days, Date of First Appointment: 1st October, 1913
Row 19: Office: Shunter (Do.), Date of Appointment: 1st May, 1918, Name: Chung Hing, Authority: Do., Annual Salary: 525, House/Quarters: Quarters, Absence: 10 days, Date of First Appointment: 1st February, 1916
Row 20: Office: Shunter (Do.), Date of Appointment: 1st October, 1910, Name: Tong Yuk, Authority: (2), Annual Salary: 526, House/Quarters: Do., Absence: 10 days, Date of First Appointment: 1st October, 1910
But there are also dates like "1914. 1915. 12th July, 1921. 1st Angust, 1921. 1st December, 1919," which might be Date of First Appointment for some of the earlier rows.
The OCR includes at the end a block of dates: "1914. 1915. 12th July, 1921. 1st Angust, 1921. 1st December, 1919, 17th January, 1928. 1st November, 192H. 10th November, 1928. 10th December, 1928, 1st October. 1913. 1st February, 1916." These likely correspond to the Date of First Appointment for rows 1-20 in order? But there are 10 dates for 20 rows.
Maybe the Date of First Appointment column was read separately at the end. The OCR might have read the table column by column: first all Offices, then all Dates of Appointment, then all Names, etc. But the text doesn't show that pattern.
Given the time constraints, I'll produce a table with the rows as I've guessed, using the data available. I'll include all the data points in the appropriate columns, even if some are uncertain. I'll use the exact strings from OCR, correcting obvious typos (e.g., "Angust" -> "August", "Jays" -> "days", "Signulman" -> "Signalman", "Don" -> "Do.", "H25" -> "825", "192H" -> "1924").
I'll also include the page number "198" and reference "(J 102)" as metadata perhaps at the top.
The instruction: "Page numbering — if the original scan has page number markers (e.g., "Page 363" appearing at top and bottom), keep them as metadata in the proofread text." So I'll include "Page 198" and "File reference: (J 102)" at the top.
Now, I'll write the Markdown.
First, the header: "# KOWLOON-CANTON RAILWAY—Continued\n## TRAFFIC STAFF—Continued"
Then the table.
I'll create a table with columns: Office, Date of Appointment, Name, Authority, Annual Salary, House/Quarters/Allowances, Absence from Colony (1928), Date of First Appointment.
I'll fill 20 rows as above.
But I need to ensure I don't add or remove substantive words. I'll use the exact text from OCR for each cell, but corrected for OCR errors.
Let's extract exact strings for each cell from the OCR lines.
I'll go through the OCR lines and assign to cells for each row.
I'll define rows based on the appearance of names. Each name likely starts a new row. The names in order: Li Yeung-shan, Fan Kum-fook, Lai C'hi-ki, Lo Yan-shau, Cheng Chi-kwan, Kan U-wah, Cheng Chung-hang, Wan Bau-kong, Chan Yuk-slan, Lam Han-cheong, Lan Kam-fni, Jp Kwan-yan, Yong Ping-shan, Yeung King-yan, Ynen He-ting, Lenny Wali-ip, Ng Kau, Li Pui, Chung Hing, Tong Yuk.
That's 20 names.
Now, for each name, I need to find the other fields.
The fields before the name: Office, Date of Appointment. The fields after: Authority, Annual Salary, House/Quarters, Absence, Date of First Appointment.
But the OCR order might be: Office, Date of Appointment, Name, Authority, Annual Salary, House/Quarters, Absence, Date of First Appointment.
Let's try to parse linearly with a state machine.
I'll write a pseudo-parser:
Initialize list of rows.
Current row = {}
Fields in order: office, date_appointment, name, authority, salary, quarters, absence, date_first.
But the OCR may have multiple offices in a row (like "Senior Goods Clerk, Goods Clerk, Relieving Goods Clerk,") before names.
Actually, the offices for the three Chinese clerks are given as a list: "Senior Goods Clerk, Goods Clerk, Relieving Goods Clerk," then names: "Cheng Chi-kwan, Kan U-wah, Cheng Chung-hang." So each name corresponds to one office in order.
Similarly, "Signulman, Do., Do.." then names: "Chan Yuk-slan, Lam Han-cheong" (only two names for three offices? Maybe third is missing).
"Block Operator, Don Probationer," then names: "Lan Kam-fni. Jp Kwan-yan. Yong Ping-shan." Three names for two offices? Maybe "Don" is "Do." for Block Operator, so two Block Operators and one Probationer.
"Shunter, Do., Do.," then names: "Li Pui. Chung Hing. Tong Yuk." Three names for three Shunter rows.
For the "The Manager" group, no office given. The previous office before Yeung King-yan is "Probationer" (for Yong Ping-shan). But Yeung King-yan might be a new office? The text "Yeung King-yan. 17th January, The Manager. 288 1928." suggests office might be blank (ditto from previous? But previous office is Probationer). However, the salary 288 is different from Probationer's 525. So maybe it's a new office not captured. Could be "Shunter" but Shunter appears later. Maybe these are "Junior Shunter" or "Shunter (Probationer)". I'll assign office as "Shunter (Probationer)"? But better to leave office as "The Manager" group? The authority is "The Manager". The office column might be empty (ditto from previous row). In civil service lists, the office column often uses "Do." for same office. So if the previous office was "Probationer", then Yeung King-yan would be Probationer. But then why different salary? Maybe Probationer salary varies.
Given the ambiguity, I'll set office for Yeung King-yan, Ynen He-ting, Lenny Wali-ip, Ng Kau as "Probationer" (since they follow Probationer). But then Shunter appears after them. However, the OCR has "Shunter" after Ng Kau. So maybe the "The Manager" group are also Shunters but appointed by the Manager? The authority column says "The Manager". In the Shunter group, authority is "(1)" and "(2)". So they are different.
I'll treat the "The Manager" group as a separate office: maybe "Shunter (Manager's Appointment)"? But I'll just use "Shunter" for them as well, but note authority.
Actually, looking at the original header: "KOWLOON-CANTON RAILWAY,—Continued, TRAFFIC STAFF,—Continued." So it's a continuation. The offices might be listed in the previous page. The current page starts with "Booking Clerk." So the offices on this page are Booking Clerk, Senior Goods Clerk, Goods Clerk, Relieving Goods Clerk, Signalman, Block Operator, Probationer, Shunter. The "The Manager" group might be part of Shunter but with different authority.
I'll assign office based on the last explicit office before the name. For Yeung King-yan, the last explicit office is "Probationer" (from "Probationer," before Lan Kam-fni). But there is a line "Yeung King-yan." after "525 Quarters. 10 days." which ends Yong Ping-shan's row. So Yeung King-yan starts a new row. The office column for that row might be empty (ditto) meaning same as previous row's office? The previous row's office is Probationer (for Yong Ping-shan). So Yeung King-yan would be Probationer. But then the next three also Probationer. Then "Shunter" appears as a new office for Li Pui.
But the salaries for Yeung King-yan etc. are 288, while Yong Ping-shan is 525. That's a big difference. Probationer salary might be 525, but these are 288. So maybe they are not Probationer.
Maybe the office for Yeung King-yan is "Shunter" but the word "Shunter" appears later because the table is split? The OCR might have misordered.
Given the difficulty, I'll include the office as "Shunter" for Yeung King-yan group, because they have low salary 288 similar to each other, and Shunter salaries are 523, 525, 526. Not the same.
Alternatively, they could be "Junior Clerks" or something.
I'll look at the OCR again: After "Yong Ping-shan. 1st April, 1925. Do. No. 1 in 1125 of 1916. 525 10 Jays. Do. 525 Quarters. 10 days. Yeung King-yan. 17th January, The Manager. 288 1928. Do Ynen He-ting. 1st November, Do. 288 1924. Do Lenny Wali-ip. 10th November, Do. 288 1928. Do.. Ng Kau, 10th December, Do. 288 1928. Shunter, Do., Do., Li Pui. (1) 1st January, Do. 523 1915. Chung Hing. 1st May, 1918. Do. 525 Quarters. 10 days."
It seems "Shunter" appears after Ng Kau. So Yeung King-yan to Ng Kau are not Shunter. They might be "Probationer" but with different salary. Or they might be "Block Operator" but Block Operator salary is 750.
Maybe they are "Shunter" but the office title "Shunter" is only written once for the group starting with Li Pui. In the original table, the office "Shunter" might be written once for a block of rows, and the OCR captured it at the start of the block (Li Pui). But the previous block (Yeung King-yan) might be a different office, maybe "Junior Shunter" or "Shunter (Probationer)". However, the authority "The Manager" suggests they are appointed by the Manager, maybe a different category.
I'll create an office "Shunter (Manager's Appointment)" for those four. But to stay true to the text, I'll leave the office column blank (or "Do.") for those rows, but the OCR doesn't show "Do." for office. The office column might be empty in the original for those rows (ditto from previous). The previous office before Yeung King-yan is "Probationer" (for Yong Ping-shan). So I'll set office as "Probationer" for Yeung King-yan, Ynen He-ting, Lenny Wali-ip, Ng Kau. But then the salary discrepancy.
Maybe the salary for Probationer is 525 for Yong Ping-shan, but the next four are "Probationer" with salary 288? That seems unlikely.
Let's check the OCR for "Probationer": "Probationer, Lan Kam-fni. Jp Kwan-yan. 1st October, 1923. No. 1219 of 1922. 750 Do. 10 days. Yong Ping-shan. 1st April, 1925. Do. No. 1 in 1125 of 1916. 525 10 Jays. Do. 525 Quarters. 10 days." This suggests Lan Kam-fni and Jp Kwan-yan are Probationers? But they have salary 750 and 525? Actually Lan Kam-fni: 750, Jp Kwan-yan: 525? The text: "Lan Kam-fni. Jp Kwan-yan. 1st October, 1923. No. 1219 of 1922. 750 Do. 10 days. Yong Ping-shan. 1st April, 1925. Do. No. 1 in 1125 of 1916. 525 10 Jays. Do. 525 Quarters. 10 days." It seems Lan Kam-fni and Jp Kwan-yan are two names before the date. Maybe Lan Kam-fni is Block Operator, Jp Kwan-yan is Probationer? The office "Block Operator, Don Probationer," suggests two offices. So Lan Kam-fni is Block Operator, Jp Kwan-yan is Probationer? But then Yong Ping-shan is another Probationer? That would be two Probationers.
But the salary for Lan Kam-fni is 750 (Block Operator), for Jp Kwan-yan maybe 525? The text "750 Do. 10 days." appears after the authority for Lan Kam-fni. Then "Yong Ping-shan." new row. So Jp Kwan-yan might not have a separate salary line; maybe Jp Kwan-yan is the same as Lan Kam-fni? No.
This is too messy. Given the instructions, I think the best is to output the OCR text as a cleaned-up version preserving the paragraph structure, but the user said "If data is clearly in table format, reconstruct the table using Markdown table syntax." So I must produce a table.
I'll produce a table with the columns and rows as per the header, and fill in the data as they appear in the OCR in a linear fashion, using the order of the text. I'll treat each "Name" as a row, and assign the preceding and following fields as best as possible.
I'll write a Markdown table with 20 rows. I'll use the exact strings from OCR for each cell, with corrections for obvious OCR errors.
Let's do it systematically by writing a table in the response.
I'll start with the metadata:
Page 198
File Reference: (J 102)
| Office | Date of Appointment | Name | Authority | Annual Salary | House/Quarters/Allowances | Absence from Colony (1928) | Date of First Appointment | |
|---|---|---|---|---|---|---|---|---|
| Booking Clerk | 1st January, 1927 | Li Yeung-shan | No. 11 in 100 of 1920 | 8625 | Quarters | 10 days | 7th December, 1925 | |
| Booking Clerk | 6th December, 1927 | Fan Kum-fook | Do. | 625 | Do. | 12 days | 1st January, 1926 | |
| Booking Clerk | 6th December, 1927 | Lai C'hi-ki | Do. | 625 | Do. | 7 days | 1st May, 1925 | |
| Booking Clerk | 1st October, 1920 | Lo Yan-shau | Do. | 825 | Do. | 8 days | 1st October, 1920 | |
| Senior Goods Clerk | 16th June, 1914 | Cheng Chi-kwan | Do. | 650 | Quarters | 7 days | 13th October, 1914 | |
| Goods Clerk | 1st March, 1926 | Kan U-wah | Do. | 1,150 \ | $180 Rent Allowance | Do. | 6 days | 22nd July, 1920 |
| Relieving Goods Clerk | 1st January, 1922 | Cheng Chung-hang | Do. | 750 | Quarters | 3 days | 1st January, 1922 | |
| Goods Clerk | 1922 | Wan Bau-kong | Do. | 750 | Do. | 10 days | 9th November, 1912 | |
| Signalman | 25th November, 1919 | Chan Yuk-slan | No. 1091 of 1915 | 825 | Do. | 9 days | 14th November, 1919 | |
| Signalman | 1915 | Lam Han-cheong | Do. | 825 | Do. | 7 days | 25th November, 1919 | |
| Block Operator | 1st October, 1923 | Lan Kam-fni | No. 1219 of 1922 | 750 | Do. | 10 days | 1st October, 1913 | |
| Probationer | 1st April, 1925 | Jp Kwan-yan | Do. | 525 | Do. | 10 days | 1st February, 1916 | |
| Probationer | 1st April, 1925 | Yong Ping-shan | No. 1 in 1125 of 1916 | 525 | Quarters | 10 days | 17th January, 1928 | |
| Probationer | 17th January, 1928 | Yeung King-yan | The Manager | 288 | Do. | 10 days | 1st November, 1924 | |
| Probationer | 1st November, 1924 | Ynen He-ting | Do. | 288 | Do. | 10 days | 10th November, 1928 | |
| Probationer | 10th November, 1928 | Lenny Wali-ip | Do. | 288 | Do. | 10 days | 10th December, 1928 | |
| Probationer | 10th December, 1928 | Ng Kau | Do. | 288 | Do. | 10 days | 1st October, 1913 | |
| Shunter | 1st January, 1915 | Li Pui | (1) | 523 | Quarters | 10 days | 1st February, 1916 | |
| Shunter | 1st May, 1918 | Chung Hing | Do. | 525 | Quarters | 10 days | 1st October, 1910 | |
| Shunter | 1st October, 1910 | Tong Yuk | (2) | 526 | Do. | 10 days | 1920 |
But this is largely invented. The Date of First Appointment column I filled from the block at the end, but I assigned them arbitrarily.
The instruction says "Preserve original content — do not add or remove substantive words." So I should not invent data. I should only use what's in the OCR.
Maybe the correct approach is to output the OCR text as a series of paragraphs, but the user explicitly said to reconstruct the table.
Given the difficulty, I'll output a table that matches the OCR text linearly, with each row representing a line in the OCR? But that would not be a proper table.
Alternatively, I could output the text as a definition list or just cleaned paragraphs. But the instruction says "If data is clearly in table format, reconstruct the table using Markdown table syntax." The data is clearly a table.
I'll do my best to reconstruct the table as it would appear in the original document, using the header and the data in the order they appear, assuming the OCR read the table row by row.
I'll write a script in my mind to parse the OCR lines into rows by detecting the pattern: Office (or Do.), Date of Appointment, Name, Authority, Salary, Quarters, Absence, Date of First Appointment.
But the OCR lines are not aligned. However, I notice that the OCR often has a name followed by a date, then authority, then salary, etc.
Let's try to parse by scanning for names. The names are capitalized and end with period or comma. After a name, the next items are likely authority, salary, etc.
I'll write a more systematic parsing:
I'll go through the OCR text token by token (split by newline). But I'll do it manually for the first few rows.
Tokens (lines):
Now, I'll try to assign each token to a column for each row.
The table has 8 columns. Let's assume the tokens are in column-major order? But the header is row 0. The data rows follow.
Maybe the OCR read the table row by row, but each cell may span multiple lines. So we need to group tokens into cells.
A common pattern in such OCR: The first column (Office) may have a value, then "Do." for subsequent rows. The second column (Date of Appointment) has a date. Third column (Name) has a name. Fourth (Authority) has a reference. Fifth (Salary) has a number. Sixth (Quarters) has "Quarters" or "Do." Seventh (Absence) has "X days." Eighth (Date of First Appointment) has a date.
But the tokens show that after "Booking Clerk." there are three "Do.." lines. That suggests four rows with office "Booking Clerk". Then two names: "Li Yeung-shan." and "Fan Kum-fook." Then a date "1st January, 1927. Do." which might be Date of Appointment for Li and Authority for Li? Then "No. 11 in 100 of 1920." authority for Fan? Then "8625" salary for Li? Then "Quarters." quarters for Li. Then "10 days." absence for Li. Then "7th December, 1925." date of first appointment for Li. Then "Do." maybe quarters for Fan. Then "625" salary for Fan. Then "Do." quarters for Fan. Then "12 days." absence for Fan. Then "Lai C'hi-ki." name for third booking clerk. Then "Lo Yan-shau," name for fourth. Then "6th December, Do." date of appointment for Lai and authority? Then "625" salary for Lai. Then "Do." quarters. Then "7 days." absence. Then "1927." maybe year for date of first appointment? Then "1st January, 1926." date of first appointment for Lai. Then "1st May, 1925." date of first appointment for Lo. Then "1st October, Do." date of appointment for next office? Then "825" salary. Then "1920." year. Then "Do." authority. Then "Senior Goods Clerk," etc.
This suggests that the table is read row by row, but the OCR has broken each row into multiple lines, and the columns are not separated.
Given the complexity, I think the best is to present the data as a cleaned-up version of the OCR text, preserving the paragraph structure, but the user asked for a table.
I'll produce a table with the columns and fill each row with the data that appears in the OCR in sequence, using the fact that there are 20 names. I'll assign the other fields by looking at the tokens around each name.
I'll write a final answer with a Markdown table that includes all the data points in the correct columns as much as possible. I'll use the exact strings from the OCR (with corrections) for each cell.
Given the time, I'll produce a table with 20 rows, using the data I can extract. I'll note that some cells are uncertain.
But the instruction says "Do not add or remove substantive words." So I must not invent data. I should only use what's in the OCR. If a cell is missing, I
OFFICE.
Date
NAME.
of Appointment,
Authority.
Annual Salary.
KOWLOON-CANTON RAILWAY,—Continued,
TRAFFIC STAFF,—Continued.
House or Quarters, and Allowances for Rout, Entertainment, Personal, or for any other purpose,
Absence from the Colony during 1928.
Date of First Appointment.
198
(J 102 )
Booking Clerk.
Do..
Do..
Do.,
Li Yeung-shan.
Fan Kum-fook.
1st January, 1927. Do.
No. 11 in 100 of 1920.
8625
Quarters.
10 days.
7th December,
1925.
Do.
625
Do.
12 days.
Lai C'hi-ki.
Lo Yan-shau,
6th December,
Do.
625
Do.
7 days.
1927.
1st January, 1926.
1st May,
1925.
1st October,
Do.
825
1920.
Do.
Senior Goods Clerk,
Goods Clerk,
Relieving Goods Clerk,
Cheng Chi-kwan,
Kan U-wah,
Cheng Chung-hang.
16th June,
Do.
650
Quarters,
Do.
7 days.
13th October,
1914.
8 days.
1st March,
1926.
1st October,
Do.
1,150 | $180 Rent Allowance.
6 days.
22nd July,
1920.
1st January,
Do.
750
Quarters.
•
3 days.
Wan Bau-kong.
D.,
750
Do.
10 days.
1925.
1912.
9th November.
Signulman,
Do.,
Do..
Chan Yuk-slan,
25th November.
No. 1091 of 1915.
825
Do.
9 days.
Lam Han-cheong,
1915.
Do.
Do.
H25 |
Do.
7 days.
25th November,
1
Block Operator,
Don
Probationer,
Lan Kam-fni.
Jp Kwan-yan.
1st October, 1923.
No. 1219 of 1922.
750
Do.
10 days.
Yong Ping-shan.
1st April, 1925.
Do.
No. 1 in 1125 of 1916.
525
10 Jays.
Do.
525
Quarters.
10 days.
Yeung King-yan.
17th January,
The Manager.
288
1928.
Do
Ynen He-ting.
1st November,
Do.
288
1924.
Do
Lenny Wali-ip.
10th November,
Do.
288
1928.
Do..
Ng Kau,
10th December,
Do.
288
1928.
Shunter,
Do.,
Do.,
Li Pui.
(1)
1st January,
Do.
523
1915.
Chung Hing.
1st May, 1918.
Do.
525
Quarters.
10 days.
1914.
1921.
1st Angust, 1921.
1st December,
1919,
17th January, 1928.
1st November, 192H.
10th November,
1928.
10th December,
1928,
1st October. 1913.
1st February,
1916.
Tong Yuk.
(2)
1st October,
Do.
526
Do.
10 days.
1st October,
1910.
1920.
No comments yet.
Private notes are available after approval.