-
with direct industrial impact. Responsibilities The successful candidate will: Conduct cutting-edge research in Computer Vision, Video Understanding, and Multimodal AI. Design AI models for concept
-
Vision, Video Understanding, and Multimodal AI. Design AI models for concept-driven video understanding of consumer facial care behaviours. Work with PI and company to develop the AI solution Develop novel
-
largest higher education institution. We are a vibrant community of approximately 18,000 students and 1,900 employees. ... (Video unable to load from YouTube. Accept cookie and refresh page to watch video
-
43,000 students work to create knowledge for a better world. You will find more information about working at NTNU and the application process here. ... (Video unable to load from YouTube. Accept cookie and
-
. At NTNU, 9,000 employees and 43,000 students work to create knowledge for a better world. You will find more information about working at NTNU and the application process here. ... (Video unable to load
-
. You can find more information about working at NTNU and the application process here . ... (Video unable to load from YouTube. Accept cookie and refresh page to watch video, or click here to open video
-
multiple LLM-powered agents operate concurrently to explore user-generated worlds and trigger diverse contextual scenarios. Develop a unified multimodal detection pipeline that integrates video, audio, and
-
into clear, human-centred narratives that spark curiosity, accelerate understanding, and inspire action. Key Responsibilities Lead end-to-end design processes, from user research and ideation to wireframing
-
with the Principal Investigator (PI) and industry partner to ensure all project deliverables are met Design and implement evaluation frameworks and metrics for vision-language models Develop annotated video datasets
-
contextual scenarios. Develop a unified multimodal detection pipeline that integrates video, audio, and text leveraging fine-tuned Vision-Language Models (VLMs) from WP3, supporting zero-shot reasoning and