Combining Hierarchical Cognitive Process with Process Supervision for Interpretable Scene Safety Understanding
cs.CL
Submitted: 2026-09-22
Updated: 2026-09-22
Code: https://github.com/hiyouga/LLaMA-Factory
Project page: https://rasahq.github.io/rasa-nlu-train
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- Why Exposure Bias Matters: An Imitation Learning Perspective of Error Accumulation in Language Generation
- InternLM2 Technical Report
- Training Verifiers to Solve Math Word Problems
- LoRA: Low-Rank Adaptation of Large Language Models
- LoraHub: Efficient Cross-Task Generalization via Dynamic LoRA Composition
- Mixtral of Experts
- RiskBench: A Scenario-based Benchmark for Risk Identification
- Let's Verify Step by Step
- DoRA: Weight-Decomposed Low-Rank Adaptation
- Enhancing Vision-Language Models with Scene Graphs for Traffic Accident Understanding
- Improve Mathematical Reasoning in Language Models by Automated Process Supervision
- Rethinking the Role of Demonstrations: What Makes In-Context Learning Work?
- Gemma: Open Models Based on Gemini Research and Technology
- Llama 2: Open Foundation and Fine-Tuned Chat Models
- LoRA-Flow: Dynamic LoRA Fusion for Large Language Models in Generative Tasks
- Caption Anything: Interactive Image Description with Diverse Multimodal Controls
- Larger language models do in-context learning differently
- Mixture-of-Subspaces in Low-Rank Adaptation
- Mixture of LoRA Experts
- Baichuan 2: Open Large-scale Language Models
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering