科学迷社区

像看英超和懂球帝一样看 AI 与 Biotech:最新论文、科研成果、人物、实验室、证据来源、商业化信号和关系图谱集中在一个信息流里。

已审核信号
601
证据来源
721
追踪领域
12
全球总榜人工智能安全与评测37 条匹配动态
publicationAIeditor reviewed
Rising signal74

From reactive filtering to proactive moral architecture: rethinking ethical alignment in large language models

原始来源文本: Publication signal: From reactive filtering to proactive moral architecture: rethinking ethical alignment in large language models

AISci 已验证editor reviewed / 1 证据来源
publicationAIeditor reviewed
Rising signal66

Can AI help reduce prejudice? Evaluating the effectiveness of AI-powered personalized persuasion on support for transgender rights

原始来源文本: Publication signal: Can AI help reduce prejudice? Evaluating the effectiveness of AI-powered personalized persuasion on support for transgender rights

AISci 已验证editor reviewed / 1 证据来源
publicationAIeditor reviewed
Rising signal59

AI Alignment and Safety of Large Language Models: A Survey of RLHF, Constitutional AI, Red-Teaming, and Value Learning

原始来源文本: Publication signal: AI Alignment and Safety of Large Language Models: A Survey of RLHF, Constitutional AI, Red-Teaming, and Value Learning

AISci 已验证editor reviewed / 1 证据来源
publicationAIeditor reviewed
Rising signal59

AI Alignment and Safety of Large Language Models: A Survey of RLHF, Constitutional AI, Red-Teaming, and Value Learning

原始来源文本: Publication signal: AI Alignment and Safety of Large Language Models: A Survey of RLHF, Constitutional AI, Red-Teaming, and Value Learning

AISci 已验证editor reviewed / 1 证据来源
publicationAIeditor reviewed
Tracked signal49

A Prompt-Based AI Safety Evaluation of Gender Bias in Large Language Models: A Comparative Study of ChatGPT and Gemini

原始来源文本: Publication signal: A Prompt-Based AI Safety Evaluation of Gender Bias in Large Language Models: A Comparative Study of ChatGPT and Gemini

AISci 已验证editor reviewed / 1 证据来源
publicationAIeditor reviewed
Tracked signal49

A Prompt-Based AI Safety Evaluation of Gender Bias in Large Language Models: A Comparative Study of ChatGPT and Gemini

原始来源文本: Publication signal: A Prompt-Based AI Safety Evaluation of Gender Bias in Large Language Models: A Comparative Study of ChatGPT and Gemini

AISci 已验证editor reviewed / 1 证据来源
publicationAIeditor reviewed
Tracked signal49

Compliance without coherence: fluent failure and the ethics of alignment evaluation in multi-agent language models

原始来源文本: Publication signal: Compliance without coherence: fluent failure and the ethics of alignment evaluation in multi-agent language models

AISci 已验证editor reviewed / 1 证据来源