FedCoT-VQA: A Federated Learning and Unlearning Framework for Chain-of-Thought Planners in VideoQA
cs.CR
Submitted: 2026-09-18
Updated: 2026-09-18
Terminology
Sources
- Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition
- Frame-Voyager: Learning to Query Frames for Video Large Language Models
- M-LLM Based Video Frame Selection for Efficient Video Understanding
- Q-Frame: Query-aware Frame Selection and Multi-Resolution Adaptation for Video-LLMs
- Towards Sparse Video Understanding and Reasoning
- Weakly Supervised Temporal Adjacent Network for Language Grounding
- Stochastic subgradient method converges at the rate $O(k^{-1/4})$ on weakly convex functions
- Flower: A Friendly Federated Learning Research Framework
- STAR: A Benchmark for Situated Reasoning in Real-World Videos
- FeDeRA:Efficient Fine-tuning of Language Models in Federated Learning Leveraging Weight Decomposition
- InternVideo: General Video Foundation Models via Generative and Discriminative Learning
- CLIP-ViP: Adapting Pre-trained Image-Text Model to Video-Language Representation Alignment
- ChatVideo: A Tracklet-centric Multimodal and Versatile Video Understanding System
Related papers
- SoK: AI-Augmented Binary Reversing
- Relaxed Sender Anonymity for CBDC Interbank Settlement: A Zero-Knowledge Approach on Permissioned EVM
- Calibration-Family Overfit: Why Trusted Sabotage Monitors Don't Transfer Across Lineages
- Efficient Fuzzy PSI under One-Sided Assumptions
- Sealing the Audit-Runtime Gap for LLM Skills
- Token Composition: A Graph Based on EVM Logs