Skip to content
#

crossmodal

Here are 8 public repositories matching this topic...

Language: All
Filter by language

This project focuses on social media multimodal hate detection, addressing challenges such as image–text semantic inconsistency, implicit references, and sarcastic/ironic expressions. We develop a CLIP + LLMs based framework for multimodal representation learning and semantic alignment. Prompt Engineering (PE) optimizes LLM output, while multi-LLM

  • Updated Dec 17, 2025
  • Jupyter Notebook

PEANUT (Prompt-Enhanced Ablation with Optical Flow-Based Neural Unit) designed to enhance video restoration by combining spatial and temporal consistency with clarity optimization. The core innovation lies in Prompt-Guided Mask Self-Generation and leveraging optical flow-based neural units to generate high-fidelity video sequences

  • Updated Apr 9, 2026
  • Python

Add this topic to your repo

To associate your repository with the crossmodal topic, visit your repo's landing page and select "manage topics."

Learn more