Author

Xiang-Yu Wang

1 paper indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#small language model Preprint Aug 2026

The Pulse Beneath the Job Title: Monthly Readings of Requirements and Tasks from 750 Million Chinese Job Ads

How do we define an occupation? By its job title? An accountant at a small trading company keeps the books; at a listed firm the same title demands a certified-accountant licence, and the week goes to the reports that regulators and the board read. Same title, different bar, different work. What defines an occupation is who it lets in and what it asks them to do. In a rapidly changing labor market, tracking those requirements and tasks is how to take the market's pulse. Yet no instrument reads both at the speed they change. Official occupational directories like O*NET report one national average per occupation, updated every few years. Job postings are timely but unstructured. Research built on them works from job titles plus proprietary skill keywords, which blur what is asked of a candidate into what a candidate is asked to do. The blur matters, because rising requirements and changing tasks are different events with different causes. We separate them. From 752.6 million job ads posted on China's five leading recruitment platforms between 2022 and 2026, we extract the phrases employers write, unify those that name the same thing, and validate the mapping from text back to entry. By doing so we construct two catalogs, 20,721 requirements a candidate must meet and 44,479 tasks the hire will do. With the entries standardized, we annotate them further. Each task, for example, carries a score for how far a language model could absorb it. Matched back onto every ad, the catalogs read the market month by month. Two examples show what the layer beneath the job title buys. First, the occupational registry records one accountant where the ads record a staircase, the junior certificate at the bottom of the wage range and the intermediate one at the top. Second, counting occupations says the work most exposed to language models is disappearing, and counting tasks says far less of it is.

Qin Chen, Ying Fang, Xiang-Yu Wang et al. · 0 citations