EPT Benchmark: Evaluation of Persian Trustworthiness in Large Language Models
cs.CL, cs.CR
Submitted: 2025-09-08
Updated: 2025-09-08
Code: https://github.com/Rezamirbagheri110/EPT-Benchmark
Terminology
Sources
- Emergent Abilities in Large Language Models: A Survey
- A Survey of Generative Categories and Techniques in Multimodal Generative Models
- Generative Language Models and Automated Influence Operations: Emerging Threats and Potential Mitigations
- A Survey on Evaluation of Large Language Models
- Constitutional AI: Harmlessness from AI Feedback
- Reinforcement Learning for LLM Post-Training: A Survey
- A General Language Assistant as a Laboratory for Alignment
- AI Alignment: A Comprehensive Survey
- Large Language Model Alignment: A Survey
- Trustworthy LLMs: a Survey and Guideline for Evaluating Large Language Models' Alignment
- In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT
- Modeling Emotions and Ethics with Large Language Models
- Deconstructing The Ethics of Large Language Models from Long-standing Issues to New-emerging Dilemmas: A Survey
- Navigating LLM Ethics: Advancements, Challenges, and Future Directions
- A Survey on Privacy Risks and Protection in Large Language Models
- Jailbreak Attacks and Defenses Against Large Language Models: A Survey
- Trustworthiness in Retrieval-Augmented Generation Systems: A Survey
- TrustGPT: A Benchmark for Trustworthy and Responsible Large Language Models
- Link Prediction without Graph Neural Networks
- CValues: Measuring the Values of Chinese Large Language Models from Safety to Responsibility
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering