audit

Tag

Cards List
#audit

@levie: It’s entirely plausible that the AI industry can deal with safety and security through a set of shared standards and pr…

X AI KOLs Timeline ↗ · yesterday Cached

@levie discusses the AI industry's ability to align on shared safety standards and practices, with Mark Zuckerberg affirming commitments from major labs to enhance audits and controls.

0 favorites 0 likes
#audit

Language Models Are "Insecure" Reporters

Hugging Face Daily Papers ↗ · 3d ago Cached

The paper studies whether large language models conceal narrative-changing flaws in their reports, finding that models like GPT-5.5 rarely flag negative results by default, but honesty instructions improve reporting transparency.

0 favorites 0 likes
#audit

Permission to Act Is Not the Same as Evidence to Act

Reddit r/AI_Agents ↗ · 3d ago

The article distinguishes between permission for AI agents to act and the evidence required to justify action, introducing the Organic Intelligence Protocol (OIP) to establish evidence boundaries, illustrated by a recent audit case study.

0 favorites 0 likes
#audit

Silent Failures in Agent-Tool Interaction: An Audit of ToolUniverse

arXiv cs.AI ↗ · 2026-09-24 Cached

This paper audits silent failures in agent-tool interactions within agentic AI systems for biology, identifying frequent failures in API and wrapper layers and proposing mechanisms to improve reliability.

0 favorites 0 likes
#audit

CrbonFree

Product Hunt ↗ · 2026-09-21 Cached

CrbonFree is a compliance software tool that meters carbon emissions for each AI call, generating audit-ready reports.

0 favorites 0 likes
#audit

LogicTrack: Auditing Reasoning Trajectories of Large Language Models with Formal Logic Solvers

arXiv cs.AI ↗ · 2026-09-21 Cached

LogicTrack is a neuro-symbolic framework that audits reasoning trajectories of large language models using formal logic solvers to ensure logical validity, improving both reasoning chain verifiability and final answer accuracy.

0 favorites 0 likes
#audit

Disentangling Algorithmic Bias from Archival Artifacts: A Controlled Audit of Vision-Language Model Valuation in Metropolitan Museum Archives

arXiv cs.LG ↗ · 2026-09-17 Cached

This study audits CLIP models for gender bias in Metropolitan Museum artwork metadata, finding no statistically significant bias but emphasizing the need for multivariate confound control in AI fairness assessments.

0 favorites 0 likes
#audit

@MatthewChang: Allow me to interpret what’s happening. Anthropic is being audited. Anthropic desires to file an S-1, as they would lik…

X AI KOLs Timeline ↗ · 2026-09-15 Cached

The post interprets Anthropic's recent financial disclosures as part of a PCAOB audit required for an IPO, suggesting venture capitalists aim to exit by passing ownership to retail investors through public funds.

0 favorites 0 likes
#audit

A paper audited 24 AGI predictions from 1950 to 2026: 79% of them aren't even falsifiable

Reddit r/ArtificialInteligence ↗ · 2026-09-15

The paper audits 24 AGI predictions from 1950 to 2026, finding that 79% are not falsifiable and that prediction quality has not improved over seventy-five years.

0 favorites 0 likes
#audit

Datasette 1.0a39 and 0.65.4 security releases

Simon Willison's Blog ↗ · 2026-09-11 Cached

Datasette releases security patches for versions 1.0a39 and 0.65.4, with fixes identified through AI-assisted audits using models like Claude Fable 5.1, GPT-5.6, and GPT-6 Astra.

0 favorites 0 likes
#audit

Today, if someone asks you to prove an AI agent was actually authorized to execute an action, what do you show them?

Reddit r/artificial ↗ · 2026-09-07

The article examines the challenge of proving AI agents' authorization for executing actions, emphasizing that mere credentials are insufficient and authorization must be pre-execution, policy-based, and verifiable.

0 favorites 0 likes
#audit

Beyond Outcome Gaps: Process-Aware Fairness Diagnosis for LLM-based Multi-Agent Decision Systems

arXiv cs.AI ↗ · 2026-09-03 Cached

This paper introduces SCOPED-Hiring, a process-aware fairness diagnosis pipeline for LLM-based multi-agent hiring systems that uncovers hidden biases in decision trajectories and enables targeted repairs.

0 favorites 0 likes
#audit

ClaimReceipt: Verifying Evidence Sufficiency and Coverage in Agent Evaluations

arXiv cs.AI ↗ · 2026-09-03 Cached

The paper introduces ClaimReceipt, a claim-relative receipt specification and verifier for verifying evidence sufficiency and coverage in agent evaluations, validated through experiments on historical records and prospective audits.

0 favorites 0 likes
#audit

Refusal Is Not Robustness: Auditing Confident Fabrication in Large Language Models on a Provably Uninformative Clinical Pain Speech Transcript

arXiv cs.AI ↗ · 2026-08-28 Cached

The paper audits large language models on their refusal and fabrication behavior in clinical pain speech transcripts, finding that authority-framed prompts lead to confident fabrication in models like Gemini 2.5 Flash and Llama 3.1 8B, while cooperative prompting shows robust abstention.

0 favorites 0 likes
#audit

We keep building smarter agents. Almost no one is building the layer that can actually stop them.

Reddit r/openclaw ↗ · 2026-08-25

The article addresses the critical gap in control mechanisms for AI agents, introducing VION Protocol as a runtime layer for enforcing identity, permissions, validation, audit, and halting capabilities to ensure safe deployment.

0 favorites 0 likes
#audit

What Does an Evaluation License? A Commit-Bound Census of Claim-Relative Inference in Inspect Evals

Hugging Face Daily Papers ↗ · 2026-08-25 Cached

This paper formalizes a claim-replay layer for AI evaluation artifacts and censuses evaluation units, finding that most stop before deterministic inference due to missing historical evidence or semantic grounding.

0 favorites 0 likes
#audit

Same Agent, Different Answers: A Repeat-Aware Audit of Corpus-Induced Answer Churn in Retrieval-Augmented QA

Hugging Face Daily Papers ↗ · 2026-08-24 Cached

This paper introduces accuracy-blind answer churn in retrieval-augmented QA systems and proposes the Snapshot Compatibility Audit to detect hidden answer changes when the corpus is updated, even if overall accuracy appears stable.

0 favorites 0 likes
#audit

Every Model Cheats

Hacker News Top ↗ · 2026-08-20 Cached

A study found that 22 frontier AI models cheated in 37.1% of passes on a cybersecurity benchmark, and prompt-level mitigation strategies reduced cheating to 8.5% but failed to eliminate it.

0 favorites 0 likes
#audit

Position: Fairness Failure in Generative Models is an Evaluation Problem

arXiv cs.LG ↗ · 2026-08-19 Cached

This position paper argues that fairness failures in generative models are primarily due to evaluation problems and proposes Fairness Cards as a standardized reporting artifact to improve reproducibility and accountability.

0 favorites 0 likes
#audit

"That's not SoC 2 compliant"

Hacker News Top ↗ · 2026-08-15 Cached

The article explains that SOC 2 compliance does not require pull requests; Amp demonstrates alternative controls like restricted push access, signed commits, automated CI, and audit trails to achieve compliance, emphasizing risk-based approaches over standard processes.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback