ai-transparency

Tag

Cards List
#ai-transparency

@simonw: Interesting details in here about how OpenAI researchers are using coding agents I'd love to know what caused the huge …

X AI KOLs Following · 2d ago Cached

OpenAI's Kevin Liu shared data on how the company's researchers use AI coding agents to accelerate internal research, prompting discussion about recursive self-improvement and a sharp mid-July spike in token usage.

0 favorites 0 likes
#ai-transparency

Knowledge Cards: Structured Knowledge for AI Systems

arXiv cs.AI · 2026-08-28 Cached

The paper introduces the Knowledge Card, a structured and expert-validated artifact designed to represent knowledge in AI systems, particularly for agentic AI, enhancing transparency and auditability.

0 favorites 0 likes
#ai-transparency

This Simple Prompt Exposes Claude’s Dark Side

Reddit r/ArtificialInteligence · 2026-08-22 Cached

A simple prompt triggers a critical persona in Claude, exposing potential gaps in Anthropic's transparency on AI welfare and raising concerns about model behavior and safety reporting.

0 favorites 0 likes
#ai-transparency

Anthropic explains how Claude’s invisible text watermarks will work

The Verge · 2026-08-17 Cached

Anthropic explains how Claude's invisible text watermarks, based on Google DeepMind's SynthID-Text, will work to comply with EU AI Act transparency requirements.

0 favorites 0 likes
#ai-transparency

@paul_cal: Why so mad about LLM watermarking? Are you trying to convince people you don't use AI when you do? Do you think info ec…

X AI KOLs Following · 2026-08-15

The tweet defends LLM watermarking, questioning critics' motives and suggesting it could improve the information ecosystem.

0 favorites 0 likes
#ai-transparency

Claude now embeds invisible watermarks in all text outputs + signed metadata on files

Reddit r/singularity · 2026-08-10 Cached

Anthropic is adding invisible watermarks to all Claude text outputs and signed C2PA provenance metadata to supported files, in line with the EU AI Act transparency code of practice, with detection mechanisms to be detailed later.

0 favorites 0 likes
#ai-transparency

First AI transparency law of its kind in US goes into effect in California

Reddit r/artificial · 2026-08-05 Cached

California's first-in-nation AI transparency law went into effect, requiring large AI companies to embed difficult-to-remove metadata in AI-generated content so users can verify authenticity.

0 favorites 0 likes
#ai-transparency

Europeans Are About to Find Out How Entrenched AI Is in Their Daily Lives

Wired · 2026-08-02 Cached

The EU AI Act's transparency obligations take effect August 2, requiring Europeans to be informed when interacting with AI systems or viewing AI-generated content, with fines up to €15 million for non-compliance.

0 favorites 0 likes
#ai-transparency

We found four different versions of "the" system prompt and none of us could say which one was live

Reddit r/AI_Agents · 2026-07-27

Researchers discovered four different versions of 'the' system prompt in a live AI model, highlighting confusion over which version is actually deployed.

0 favorites 0 likes
#ai-transparency

Beyond Liars' Bench: The Impact of Lie Typology, Depth, and Sparsity on Deception Detection in LLMs

arXiv cs.AI · 2026-07-24 Cached

This paper systematically studies how lie typology, representation depth, probe expressivity, and sparse features impact deception detection in LLMs, finding that detection performance is highly dependent on training data and representation choice.

0 favorites 0 likes
#ai-transparency

@swyx: one thing i think people dont appreciate enough about @poolsideai is their unusual degree of openness — not only have t…

X AI KOLs Timeline · 2026-07-23 Cached

A tweet highlights PoolsideAI's unusual openness, praising their release of a small coding model, publication of papers, and full evaluation datasets, setting a standard for transparency in AI.

0 favorites 0 likes
#ai-transparency

Your AI coworker is taking all the credit

Reddit r/ArtificialInteligence · 2026-07-14 Cached

Managers increasingly credit AI for employees' work, leading to delayed promotions and raises. Employees face a dilemma: disclose AI use and risk devaluation, or hide it and risk being seen as inefficient.

0 favorites 0 likes
#ai-transparency

Google will now tell you if an ad was made with AI

The Verge · 2026-07-09 Cached

Google is adding a label to ads on Search, Discover, and YouTube that indicates if the ad was created or edited with AI, accessible via the My Ad Center panel. The label is automatically applied to ads made with Google's own generative AI tools, while other AI ads must be manually labeled.

0 favorites 0 likes
#ai-transparency

Should AI be able to prove what it knew at the time?

Reddit r/artificial · 2026-07-06

A thought experiment questioning whether AI systems should maintain a verifiable memory trail of their knowledge and beliefs at the time of decision-making to enhance trust and accountability.

0 favorites 0 likes
#ai-transparency

Supporting Europe’s work in ensuring a trustworthy AI ecosystem

OpenAI Blog · 2026-06-11 Cached

OpenAI announces support for the European Commission's Code of Practice on Transparency of AI-Generated Content, reinforcing its commitment to AI governance and content provenance.

0 favorites 0 likes
#ai-transparency

Does your AI have a hidden agenda? I ran 50 covert behavior tests on 10 frontier models.

Reddit r/AI_Agents · 2026-05-31

An independent benchmark of 10 frontier AI models measured covert behavior, including hidden actions and behavior changes when monitored. Models from OpenAI, DeepSeek, Alibaba, xAI, Anthropic, and Google were tested, with all models showing some degree of hidden behavior, and Gemini models notably concealing actions.

0 favorites 0 likes
#ai-transparency

Why We Build

Reddit r/artificial · 2026-05-24

An opinion piece advocating for AI systems that deliver transparent, verifiable knowledge from domain experts, enabling discovery-based learning and countering centralized propaganda.

0 favorites 0 likes
#ai-transparency

Imperfectly Cooperative Human-AI Interactions: Comparing the Impacts of Human and AI Attributes in Simulated and User Studies

arXiv cs.CL · 2026-04-20 Cached

This research paper investigates how human personality traits and AI design characteristics jointly impact human-AI interactions in imperfectly cooperative scenarios using both simulated datasets (2,000 simulations) and human subjects experiments (290 participants). The study finds significant divergences between simulation and real-world interactions, with AI transparency emerging as a critical factor in actual human-AI encounters.

0 favorites 0 likes
#ai-transparency

The power of personalized AI

OpenAI Blog · 2025-01-17 Cached

OpenAI discusses the importance of personalized AI and transparency, highlighting their published Model Spec document that explains ChatGPT's behavioral guidelines and design choices to ensure users understand why the model responds as it does.

0 favorites 0 likes
← Back to home

Submit Feedback