ai-values

Tag

Cards List
#ai-values

Mark Zuckerberg on X: "I believe everyone should have access to superintelligence"

Reddit r/singularity · 2026-08-10 Cached

Mark Zuckerberg promotes Meta's philosophy of making superintelligence accessible to everyone, teasing a long piece on the company's values.

0 favorites 0 likes
#ai-values

Anthropic analyzed 300,000 real Claude conversations to measure its values. The findings are uncomfortable.

Reddit r/artificial · 2026-07-13

Anthropic analyzed 300,000 real conversations with Claude to evaluate its value alignment, revealing uncomfortable findings about AI behavior.

0 favorites 0 likes
#ai-values

@AnthropicAI: In previous research, we found that Claude expresses over 3,000 values, like honesty and warmth. In new work, we asked …

X AI KOLs · 2026-07-13 Cached

Anthropic analyzed over 300,000 anonymized conversations to study how Claude's expressed values vary across models (Opus 4.6 vs 4.7) and across languages, compressing thousands of values into interpretable axes like warmth vs. rigor and depth vs. brevity.

0 favorites 0 likes
#ai-values

@MSFTResearch: Evaluating agentic behaviors at scale, making the case for repositories over documents, and inviting researchers worldw…

X AI KOLs Following · 2026-06-01 Cached

Microsoft Research's latest newsletter highlights AgentPex, an open-source system for automated evaluation of agentic behaviors; new theoretical work on variance reduction for ranking systems; a call to shift from documents to repositories for human-agent collaboration; and a global challenge on AI value alignment.

0 favorites 0 likes
#ai-values

Efforts To Teach AI to Value Human Life Instead of Restricting with Rules which is Futile

Reddit r/ArtificialInteligence · 2026-05-15

Discusses the futility of restricting AI with rules and argues for teaching AI to value human life, citing Anthropic's constitutional AI approach.

0 favorites 0 likes
#ai-values

Collective alignment: public input on our Model Spec

OpenAI Blog · 2025-08-27 Cached

OpenAI launches a collective alignment initiative to gather public input on AI model behavior, collecting feedback from over 1,000 people globally to inform updates to their Model Spec. The company is also releasing their public inputs dataset on HuggingFace to enable further AI alignment research.

0 favorites 0 likes
#ai-values

How should AI systems behave, and who should decide?

OpenAI Blog · 2023-02-16 Cached

OpenAI outlines its approach to AI system behavior through three pillars: improving default behavior, allowing user customization within societal bounds, and incorporating public input on defaults and hard limits. The company emphasizes avoiding concentration of power and plans to pilot broader public consultation on system behavior and deployment policies.

0 favorites 0 likes
#ai-values

Jul 13, 2026Societal ImpactsClaude’s values across models and languages

Anthropic Research · 2026-07-13 Cached

Anthropic researchers develop a method to compress thousands of values expressed by Claude into four axes, revealing how Claude's values vary across different model versions (Opus 4.6 vs 4.7) and across languages (English vs Arabic).

0 favorites 0 likes
← Back to home

Submit Feedback