MARS: Malware Analysis with Rule-Based Scoring of LLM Claims
cs.CR, cs.AI
Submitted: 2026-10-07
Updated: 2026-10-07
Code: https://github.com/mandiant/flare-fakenet-ng
Terminology
Sources
- Measuring Faithfulness in Chain-of-Thought Reasoning
- FacTool: Factuality Detection in Generative AI -- A Tool Augmented Framework for Multi-Task and Multi-Domain Scenarios
- Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
- Identifying Adversary Tactics and Techniques in Malware Binaries with an LLM Agent
- Ignore Previous Prompt: Attack Techniques For Language Models
Related papers
- SoK: AI-Augmented Binary Reversing
- Relaxed Sender Anonymity for CBDC Interbank Settlement: A Zero-Knowledge Approach on Permissioned EVM
- Calibration-Family Overfit: Why Trusted Sabotage Monitors Don't Transfer Across Lineages
- Efficient Fuzzy PSI under One-Sided Assumptions
- Sealing the Audit-Runtime Gap for LLM Skills
- Token Composition: A Graph Based on EVM Logs