JobHop v2: A Large-Scale Career Trajectory Dataset from Unstructured Resumes
Large-scale, richly annotated career trajectory data underpins workforce planning, job recommendation, and labour market analysis, yet publicly available datasets are either small, closed to independent use, or built from pre-standardized occupational codes with LLM-synthesized rather than authentic free text. We present JobHop~v2, an improved version of the publicly available JobHop dataset, constructed through end-to-end large language model (LLM) extraction from a corpus of ${\sim}440{,}000$ pseudonymized, multilingual resumes provided by VDAB, the Flemish Public Employment Service. The rel
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 49%chrisliu298/awesome-llm-unlearning →
- PossiblePossibly related (embedding) · 46%tarunlnmiit/autopilot-jobhunt →
- PossiblePossibly related (embedding) · 45%samuelhm/42Jobs →
- LinkedLinked via arxiv author · 85%Iman Johary →
“JobHop v2: A Large-Scale Career Trajectory Dataset from Unstructured Resumes”
- LinkedLinked via arxiv author · 85%Guillaume Bied →
“JobHop v2: A Large-Scale Career Trajectory Dataset from Unstructured Resumes”
- LinkedLinked via arxiv author · 85%Alexandru C. Mara →
“JobHop v2: A Large-Scale Career Trajectory Dataset from Unstructured Resumes”
- LinkedLinked via arxiv author · 85%Tijl De Bie →
“JobHop v2: A Large-Scale Career Trajectory Dataset from Unstructured Resumes”
