Tag
ASIRF is an agentic framework that retrieves domain-specific definitions for sensitive information redaction at inference time, achieving higher recall than trained classifiers across various domains with minimal expert input.
The learn-from-materials Agent Skill converts PDFs, EPUBs, Word docs, and PPTs into interactive learning websites with source tracing and knowledge base building.
ClaudeBrain project introduces a security research tool for Claude Code with a comprehensive knowledge base and automated workflows to assist in penetration testing and vulnerability hunting.
Opyt is a free, MIT-licensed tool that creates a personalized knowledge base from your followed content on X, Substack, and other sources, using AI to enhance exploration and track new developments.
MACE is a multi-agent candidate event acquisition method that uses evidence-specialized LLM agents to refine event structure, improving accuracy in event linking tasks without modifying existing models.
The author has built Global Fail Map, an open-source interactive tool documenting historical failures in science, technology, and business to serve as a research layer for AI agents and developers.
An AI agency shares their experience implementing AI in a client's support system, detailing benefits like improved self-service and agent efficiency, but noting that AI cannot fix underlying policy issues and requires ongoing maintenance.
Grok, an AI model, has created a configurable and downloadable encyclopedia for agentic engineering, including a field guide and integrated lessons, showcasing AI's ability to generate comprehensive educational resources.
A web developer shares experiences building systems around AI coding agents for reliable development and seeks advice on workflows and orchestration tools.
This AI screenwriting knowledge base, named 'screenwriting-skills', distills skills from 19 classic screenwriting books, provides 13 screenwriting skills, has 392 GitHub stars, and aims to outperform AI in scriptwriting.
The paper introduces a two-stage LLM pipeline using fine-tuned Qwen3-4B with Hyper-Parallel Decoding to extract purchase-discriminative attributes in e-commerce, achieving 85% accuracy with 92% cost reduction.
This article describes a comprehensive screenwriting knowledge base derived from 19 classic books, designed to assist in AI-assisted script writing and learning, covering various stages and cultural methods.
The paper introduces PACE, a dataset for evaluating whether AI models can identify hidden conflicts in user requests by retrieving implicit knowledge base facts, and proposes PaceMaker, a multi-agent framework to enhance conflict-aware decision-making.
This paper introduces LexIssue, a benchmark for identifying disputed legal issues in Chinese civil litigation, constructed from real cases and expert annotations, and shows that retrieval-augmented generation enhances performance.
This article provides a comprehensive guide from concept to implementation for enterprise AI knowledge bases, emphasizing the importance of data governance, access control, testing, and validation, and includes an implementation checklist.
Built an open-source fact-checker for AI agents that verifies claims by fetching real-time sources and providing truth and confidence scores, useful for both public and internal documents.
Google Research's WikiSkill preprint demonstrates that compiling execution history into skill files allows a 9B AI model to outperform a 27B model in benchmarks, highlighting the role of procedural memory and skill transfer in agent performance.
Virgil is an interactive system designed to help users discover and compare explainability tools for transformer-based language models through a curated knowledge base and unified interface.
WikiSkill introduces a framework that co-evolves agent skills with a persistent knowledge base to systematically accumulate experience and improve performance across models, demonstrating benefits over state-of-the-art skill-evolution methods.
Cabinet is a self-hosted AI workspace that allows companies to own their knowledge base and automate tasks with AI agents, using models like Claude and GPT without vendor lock-in.