Postdoc Position In Multimodal Foundation Models For Document Understanding
1 day ago
Bellaterra
POSTDOC POSITION IN MULTIMODAL FOUNDATION MODELS FOR DOCUMENT UNDERSTANDINGCall reference: 20260619_MDUWe are seeking apostdocto join the Vision, Language and Reading group at the Computer Vision Center (CVC) in Barcelona, Spain.The position is initially for 1 year and linked to the project “Multimodal LLMs for Document Understanding” (Mu Doc U), financed by the Spanish Ministry of Science.The successful candidate is expected to participate in large‑scale training efforts, research on multimodal pre‑training and finetuning methods, and applications on the specific use case of Document Understanding.KEY DUTIESLead and contribute to specific Work Packages (WPs) within the project, ensuring timely delivery of objectives and nduct cutting‑edge research on multimodal large language models (LLMs), and manage the publication of research findings in top‑tier conferences and journals.Prepare and submit proposals for high‑performance computing (HPC) resources and manage allocated compute efficiently.Optimize training pipelines for scalability and ntribute to group mentoring activities such as reading groups or internal seminars.Actively collaborate with internal group members and external project partners.CANDIDATE’S PROFILEThe candidate should possess aPh D in machine learning or computer vision, or be in the final stage of their Ph D with a scheduled or imminent thesis defense. A strong publication record is required. We are looking for candidates who have publications in top conferences like ICDAR, CVPR, ECCV, ICCV, AAAI, Neur IPS.The candidate should have a strong background inLarge Language Modelsand experience in thedocument image analysisfield. Experience inindustrywill be considered a strong asset.The applicant is expected to be fluent in both oral and written communication in English. They should work well in a team while demonstrating initiative and independence. The candidate is expected to co‑supervise Ph D NDITIONSGross annual salary: €30,000Starting date: July or September 2026ABOUT CVCThe selected candidate will work in the Computer Vision Centre (CVC) in Barcelona, a research institute comprising more than 150 researchers and support staff, dedicated to computer vision research and knowledge transfer.