PRISM-RAG: Multimodal Hypergraph Retrieval-Augmented Generation for Tobacco Product and Legislative Policy Reasoning
cs.CV
Submitted: 2026-09-20
Updated: 2026-09-20
Code: https://github.com/explosion/spaCy
Project page: https://manuelserna.github.io/sch-tpami-website
Terminology
Sources
- Public Health Advocacy Dataset: A Dataset of Tobacco Usage Videos from Social Media
- HyperGraphRAG: Retrieval-Augmented Generation via Hypergraph-Structured Knowledge Representation
- Language Models are Few-Shot Learners
- Generative Pre-trained Transformer: A Comprehensive Review on Enabling Technologies, Potential Applications, Emerging Challenges, and Future Directions
- GPT-4 Technical Report
- Training language models to follow instructions with human feedback
- Learning to summarize from human feedback
- LLaMA: Open and Efficient Foundation Language Models
- Llama 2: Open Foundation and Fine-Tuned Chat Models
- Instruction Tuning with GPT-4
- Visual Instruction Tuning
- Improved Baselines with Visual Instruction Tuning
- LLaVA-Mini: Efficient Image and Video Large Multimodal Models with One Vision Token
- OpenAI GPT-5 System Card
- Gemma 3 Technical Report
- A Survey on Hallucination in Large Vision-Language Models
- Retrieval-Augmented Generation with Graphs (GraphRAG)
- GRAG: Graph Retrieval-Augmented Generation
- LightRAG: Simple and Fast Retrieval-Augmented Generation
- PathRAG: Pruning Graph-based Retrieval Augmented Generation with Relational Paths
Related papers
- Loss Knows Best: Detecting Annotation Errors in Videos via Loss Trajectories
- AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories
- Benchmarking the Robustness of Foundation Models for Mammography under Domain Shift
- MambaX-Net: Dual-Input Mamba-Enhanced Cross-Attention Network for Longitudinal MRI Segmentation
- TeleOCR: Navigating Document Parsing Across Digital and Camera-Captured Documents
- A Survey on Efficient Vision-Language-Action Models