SourceCornell UniversityEvalSafetyGap: A Hybrid Survey and Conceptual Framework for LLM Evaluation-Safety Failures
- Publisher
- Cornell University
- Date
- Jun 29, 2026
- Trust tier
- openalex
- License
- Not specified
SourceEuropean Organization for Nuclear ResearchAI Alignment and Safety of Large Language Models: A Survey of RLHF, Constitutional AI, Red-Teaming, and Value Learning
- Publisher
- European Organization for Nuclear Research
- Date
- Jul 15, 2026
- Trust tier
- openalex
- License
- cc-by
SourceOxford University PressCan AI help reduce prejudice? Evaluating the effectiveness of AI-powered personalized persuasion on support for transgender rights
- Publisher
- Oxford University Press
- Date
- Jun 29, 2026
- Trust tier
- openalex
- License
- Not specified
SourceEuropean Organization for Nuclear ResearchAI Alignment and Safety of Large Language Models: A Survey of RLHF, Constitutional AI, Red-Teaming, and Value Learning
- Publisher
- European Organization for Nuclear Research
- Date
- Jul 15, 2026
- Trust tier
- openalex
- License
- cc-by
SourceSpringer Science and Business Media LLCFrom reactive filtering to proactive moral architecture: rethinking ethical alignment in large language models
- Publisher
- Springer Science and Business Media LLC
- Date
- Jul 6, 2026
- Trust tier
- crossref
- License
- https://www.springernature.com/gp/researchers/text-and-data-mining
SourceR. Oldenbourg VerlagGround-truthing AI energy consumption: validating CodeCarbon against external measurements
- Publisher
- R. Oldenbourg Verlag
- Date
- Jul 10, 2026
- Trust tier
- openalex
- License
- cc-by
SourceOpenAlex indexed sourceWhat Does Your AI Mean by "Flourishing"? The Case for Disclosing and Benchmarking the Values in AI Alignment (Preprint)
- Publisher
- OpenAlex indexed source
- Date
- Jul 6, 2026
- Trust tier
- openalex
- License
- cc-by
SourceEuropean Organization for Nuclear ResearchBU213 AI Alignment Regime Switching under B_U BU213 B_U 体系下 AI 对齐中的 Regime Switching
- Publisher
- European Organization for Nuclear Research
- Date
- Jul 1, 2026
- Trust tier
- openalex
- License
- cc-by
SourceFrontiers Media SATrust, transparency, and disciplinary alignment in research funding evaluations: an illustrative case from the Norwegian AI Center call (2024–2025)
- Publisher
- Frontiers Media SA
- Date
- Jun 26, 2026
- Trust tier
- crossref
- License
- https://creativecommons.org/licenses/by/4.0/
SourceNature PortfolioPerceptions of voluntary horizontal confidential food safety data sharing: an exploratory interview study with food industry leadership
- Publisher
- Nature Portfolio
- Date
- Jun 24, 2026
- Trust tier
- openalex
- License
- cc-by-nc-nd
SourceOpenAlex indexed sourceAI, Digital Platforms, and the New Systemic Risk
- Publisher
- OpenAlex indexed source
- Date
- Jun 23, 2026
- Trust tier
- openalex
- License
- cc-by-nc-nd
SourceCogitatioParty Equalization or Normalization Through Visual Generative AI in the 2025 German Federal Election
- Publisher
- Cogitatio
- Date
- Jul 2, 2026
- Trust tier
- openalex
- License
- cc-by
SourceEuropean Organization for Nuclear ResearchA Prompt-Based AI Safety Evaluation of Gender Bias in Large Language Models: A Comparative Study of ChatGPT and Gemini
- Publisher
- European Organization for Nuclear Research
- Date
- Jul 6, 2026
- Trust tier
- openalex
- License
- cc-by
SourceElsevier BVThe scales of justitia: A comprehensive survey on safety evaluation of LLMs
- Publisher
- Elsevier BV
- Date
- Jun 28, 2026
- Trust tier
- openalex
- License
- Not specified
SourceSpringer Science+Business MediaBEADS: Bias Evaluation Across Domains
- Publisher
- Springer Science+Business Media
- Date
- Jul 3, 2026
- Trust tier
- openalex
- License
- cc-by-nc
Sourcenpj Mental Health ResearchPsyEval: a comprehensive large language model evaluation benchmark for mental health
- Publisher
- npj Mental Health Research
- Date
- Jul 9, 2026
- Trust tier
- openalex
- License
- cc-by-nc-nd
SourceBenchCouncil PressDesign and Evaluation of an Interpretable Multimodal Deep Learning Framework for Early Alzheimer’s Disease Detection
- Publisher
- BenchCouncil Press
- Date
- Jun 30, 2026
- Trust tier
- crossref
- License
- https://creativecommons.org/licenses/by-nc-nd/4.0
SourceIAEME PublicationPUBLIC POLICY IN THE AGE OF AI: LINKING POLICY FORMULATION, EFFECTIVE IMPLEMENTATION, AND PEOPLE-CENTRIC POLICY BENCHMARKS IN INDIA
- Publisher
- IAEME Publication
- Date
- Jun 25, 2026
- Trust tier
- crossref
- License
- Not specified
SourceOpenAlex indexed sourceExpert Evaluation and the Limits of Human Feedback in Mental Health AI Safety Testing
- Publisher
- OpenAlex indexed source
- Date
- Jun 25, 2026
- Trust tier
- openalex
- License
- cc-by
SourceCambridge University Press (CUP)Motivation and post-design evaluations of AI usage behind AI-assisted design
- Publisher
- Cambridge University Press (CUP)
- Date
- Jul 2, 2026
- Trust tier
- crossref
- License
- https://creativecommons.org/licenses/by-nc-nd/4.0/
SourceBenchCouncil PressJAMAL: A Multidimensional Benchmark for Arabic Commonsense Reasoning Across Life-Domains and Cognitive Axes
- Publisher
- BenchCouncil Press
- Date
- Jun 30, 2026
- Trust tier
- crossref
- License
- https://creativecommons.org/licenses/by-nc-nd/4.0
SourceCornell UniversityYuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety
- Publisher
- Cornell University
- Date
- Jun 23, 2026
- Trust tier
- openalex
- License
- cc-by
SourceCarlson Research LLCExplanations of the Fermi Paradox and the Drake Equation
- Publisher
- Carlson Research LLC
- Date
- Jul 16, 2026
- Trust tier
- crossref
- License
- https://creativecommons.org/licenses/by/4.0
SourceEuropean Organization for Nuclear ResearchBU213 AI Alignment Regime Switching under B_U BU213 B_U 体系下 AI 对齐中的 Regime Switching
- Publisher
- European Organization for Nuclear Research
- Date
- Jul 1, 2026
- Trust tier
- openalex
- License
- cc-by
SourceOpenAlex indexed sourceThe 2025 AI Agent Index: Documenting Technical and Safety Features of Deployed Agentic AI Systems
- Publisher
- OpenAlex indexed source
- Date
- Jun 23, 2026
- Trust tier
- openalex
- License
- cc-by
SourceOpenAlex indexed sourceStrategic Polysemy in AI Discourse: A Philosophical Analysis of Language, Hype, and Power
- Publisher
- OpenAlex indexed source
- Date
- Jun 23, 2026
- Trust tier
- openalex
- License
- cc-by
SourceCornell UniversityPluralis v0.1: Towards a Multicultural, Multimodal, Multilingual Benchmark for AI Risk and Reliability
- Publisher
- Cornell University
- Date
- Jul 7, 2026
- Trust tier
- openalex
- License
- Not specified
SourceCARI Journals LimitedOperationalizing Federated Healthcare AI: Design Patterns, Benchmarks, and Policy
- Publisher
- CARI Journals Limited
- Date
- Jun 23, 2026
- Trust tier
- crossref
- License
- https://creativecommons.org/licenses/by/4.0
SourceThe Korea English Language Testing AssociationHow Consistent Is AI Writing Feedback? Dimension-Level Reliability Across Repeated CEFR-Based Evaluations
- Publisher
- The Korea English Language Testing Association
- Date
- Jun 30, 2026
- Trust tier
- crossref
- License
- Not specified
SourceAcademy of ManagementGenerative AI as a Tool for Content Validation: Establishing Model-Specific Interpretive Benchmarks
- Publisher
- Academy of Management
- Date
- Jul 1, 2026
- Trust tier
- crossref
- License
- Not specified
SourceBenchCouncil PressA Hybrid MCDM Framework for Assessing Financial Resilience and Trend Dynamics in Indian Commercial Banks
- Publisher
- BenchCouncil Press
- Date
- Jun 30, 2026
- Trust tier
- crossref
- License
- https://creativecommons.org/licenses/by-nc-nd/4.0
SourceCornell UniversityMolSafeEval: A Benchmark for Uncovering Safety Risks in AI-Generated Molecules
- Publisher
- Cornell University
- Date
- Jul 1, 2026
- Trust tier
- openalex
- License
- Not specified
SourceSpringer Science and Business Media LLCCompliance without coherence: fluent failure and the ethics of alignment evaluation in multi-agent language models
- Publisher
- Springer Science and Business Media LLC
- Date
- Jul 6, 2026
- Trust tier
- crossref
- License
- https://www.springernature.com/gp/researchers/text-and-data-mining
SourceUniversity of LjubljanaOff-Centre AI
- Publisher
- University of Ljubljana
- Date
- Jul 10, 2026
- Trust tier
- crossref
- License
- https://creativecommons.org/licenses/by-sa/4.0
SourceEuropean Organization for Nuclear ResearchA Prompt-Based AI Safety Evaluation of Gender Bias in Large Language Models: A Comparative Study of ChatGPT and Gemini
- Publisher
- European Organization for Nuclear Research
- Date
- Jul 6, 2026
- Trust tier
- openalex
- License
- cc-by
SourceAssociation for Computing MachineryDefining and Evaluating Physical Safety for Large Language Models
- Publisher
- Association for Computing Machinery
- Date
- Jul 13, 2026
- Trust tier
- openalex
- License
- cc-by
SourceUniversity of Maryland, College ParkPRINCIPLED FRAMEWORKS FOR AI ALIGNMENT: FROM POST-TRAINING TO INFERENCE
- Publisher
- University of Maryland, College Park
- Date
- Jul 2, 2026
- Trust tier
- openalex
- License
- Not specified