Launch HN: Context.dev (YC S26) – API to get structured data from any website

Hacker News Top Products

Summary

Context.dev is a YC-backed API that allows developers and AI agents to scrape, crawl, and extract structured data from any website, with features like markdown, HTML, sitemaps, screenshots, and brand intelligence, aiming to simplify web data integration.

Hi Hacker News, I’m Yahia. I built Context.dev (<a href="https:&#x2F;&#x2F;www.context.dev&#x2F;">https:&#x2F;&#x2F;www.context.dev&#x2F;</a>) to make it really easy to integrate web data into your products and agents.<p>Here’s a demo video: <a href="https:&#x2F;&#x2F;www.tella.tv&#x2F;video&#x2F;build-faster-with-context-dev-apis-2cgl" rel="nofollow">https:&#x2F;&#x2F;www.tella.tv&#x2F;video&#x2F;build-faster-with-context-dev-api...</a><p>Since it’s an API, here are the docs: <a href="https:&#x2F;&#x2F;docs.context.dev&#x2F;quickstart">https:&#x2F;&#x2F;docs.context.dev&#x2F;quickstart</a>.<p>You can send us a URL and get back clean Markdown, rendered HTML, screenshots, extracted images, etc.. You can also send us a domain and get company or brand context: name, description, logos, colors, fonts, social links, screenshots, style information, and related metadata. For more custom use cases, you can send a URL plus a JSON Schema and ask us to extract structured data from the site into that shape. For example, you might ask for pricing plans, product categories, office locations, support links, integration partners, or anything else that is visible on the public site.<p>The goal is to give developers the output they actually want. Raw HTML is rarely the useful thing; the useful thing is usually Markdown for a model, JSON for an application, a logo for a UI, or a structured company profile for an agent.<p>Before, I worked at Amazon and Sunrun, and co-founded StockAlarm.io &amp; essense.io, both of which were acquired. Also, I built knifegeek.io, which scraped pocket knives from across the internet and listed them easily. The project is outdated now (coming back soon) but back then it hit the frontpage of hacker news and people seemed to like it: <a href="https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=34604281">https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=34604281</a>.<p>Just before Context.dev, I built Brand.dev. The idea was that your software product should automatically know about your customer if they sign up with a corporate email. The API pulled brand data such as logos, backdrops, name, description, industry, and more from the public web and surfaced it to your product to integrate as part of their onboarding experience. That’s worth doing because conversion rates on onboarding improve dramatically when you go from “enter all this info” to “confirm all this info” (and there was never any privacy concern all the information is public).<p>That was a nifty niche, but the more customers used it, it became obvious that “brand data” was only one slice of a larger need. People started asking for things like screenshots, structured extraction, and LLM ready data. So I expanded to Context.dev, and applied to YC (got rejected after an interview), then kept going and re-applied at which point I got in as a solo founder.<p>People use Context.dev in more ways than I can list, but here are some: keeping context up to date on customer websites for chatbots - building beautiful brand assets&#x2F;ads for customers - enrichment flows using agent harnesses like eve.dev - crawling customer websites into chatbot knowledge bases - turning GitHub repos into branded docs sites - academic journal and PDF crawling. There are a ton more examples at <a href="https:&#x2F;&#x2F;www.context.dev&#x2F;customers">https:&#x2F;&#x2F;www.context.dev&#x2F;customers</a>.<p>We know that many crawlers are not behaving like good citizens on the web, and the entire space has a bad reputation as a result. At the same time, customers are not usually trying to buy “scraping”. They are trying to make a support bot work, personalize onboarding, enrich CRM records, generate docs, monitor leads, or let an agent research a company. There are lots of legit use cases. We want to satisfy those while being respectful of everyone involved.<p>We maintain a caching layer and avoid hammering websites. Customers can configure the cache, but if we find we’re sending too many requests to a url in a certain amount of time, we step in and tone it down. Websites can opt out of our service, and we respect these requests and add them to our block list.<p>We focus on customers who want to build cool things for their users. Enriching onboarding is a popular use case. So is integrating context about their own websites (things like support bots), and building agents that can automatically reason about complex tasks involving the internet.<p>We only allow customers to use brand data to identify a specific customer on their software, you cannot use it in your own materials or to imply endorsement.<p>I&#x27;d love to hear your feedback about the product in the comments, thanks!
Original Article
View Cached Full Text

Cached at: 07/09/26, 04:36 PM

# Context.dev: Web Scraping & Crawl API for AI Agents Source: [https://www.context.dev/](https://www.context.dev/) ![](https://www.context.dev/_next/image?url=%2Fhero-glass-texture.webp&w=3840&q=75&dpl=dpl_9WdAgKhEKUCXoti2FuJBj7GLq46t) Poweringproductsandagentsat ## From maintaining data infrastructure toshipping product features ## Set it upyour way\. Do it yourself in the dashboard, or let your AI agent handle the whole thing\. ### Do it yourself 1. 1\. Sign up and verify your email\. 2. 2\. Copy your API key from the dashboard\. 3. 3\. Install the SDK and start calling the API\. ## The\{APIs\}powering your next feature Scrape any URL, crawl an entire website, and extract structured data in seconds: markdown, HTML, sitemaps, screenshots, and brand intelligence\. Connect Your Agent To The Web ## From web scraping to brand extraction inone API\. Scrape, crawl, extract, and enrich web data with AI, all from one provider\. Here's what your agent can do\. ## Get web data into your productin just three steps [Get API key](https://www.context.dev/signup) 01 ### Choose the data your product needs Pick from web content, brand, company, and website APIs, all under one key\. 02 ### Connect it to your workflow Integrate in minutes with developer\-friendly SDKs, clear docs, and example code\. 03 ### Ship the feature Power onboarding, enrichment, AI agents, and automation with live web data\. ## Whatdeveloperssay Feedback from the engineers shipping with us every day\. ![John Acosta](https://www.context.dev/_next/image?url=%2Ftestimonials%2Fjohnacosta.jpeg&w=128&q=75&dpl=dpl_9WdAgKhEKUCXoti2FuJBj7GLq46t) John Acosta > “As soon as you type in the email, we use Context\.dev to collect information about your brand and different logos\.**Context\.dev saves us a lot of time right now**\.” Founder @ UsePropane\.ai ![Nick Khami](https://www.context.dev/_next/image?url=%2Ftestimonials%2Fnickkhami.png&w=128&q=75&dpl=dpl_9WdAgKhEKUCXoti2FuJBj7GLq46t) Nick Khami > “Onboarding couldn't be simpler\. Self\-serve sign\-up gives you an API key immediately, the documentation is thorough, and**we were integrating within 10 minutes**\.” Engineering Manager @ Mintlify ![Luke Ramsden](https://www.context.dev/_next/image?url=%2Ftestimonials%2Flukeramsden.webp&w=128&q=75&dpl=dpl_9WdAgKhEKUCXoti2FuJBj7GLq46t) Luke Ramsden > “**Getting started is very simple**\. API docs are great and sign\-up is self serve, with an API key generated immediately\. Took 10 minutes to start integrating\.” CPTO @ Architect \(tryarchitect\.com\) ![Aaron Edwards](https://www.context.dev/_next/image?url=%2Ftestimonials%2Faaronedwards.webp&w=128&q=75&dpl=dpl_9WdAgKhEKUCXoti2FuJBj7GLq46t) Aaron Edwards > “We're seeing**much higher activation rates**for our free trials and sign\-ups because of it\. It's made it a lot simpler for customers to get started with DocsBot and to get to that**'aha' value moment**very quickly\.” Founder @ DocsBot ![Vlad Veselukha](https://www.context.dev/_next/image?url=%2Ftestimonials%2Fvlad.webp&w=128&q=75&dpl=dpl_9WdAgKhEKUCXoti2FuJBj7GLq46t) Vlad Veselukha > “Context\.dev offers**extensive documentation**making the setup super easy\. On top of this, there is a**dedicated Slack channel**in case you have follow\-up questions\. Overall the API integration process was very smooth\.” Senior Data Engineer @ Vizzy CUSTOMERS ## See how our customers do it From scrappy startups to Fortune 500s, teams ship brand\-powered features in days instead of quarters\.[See all customers](https://www.context.dev/customers) ## The Blog Engineering deep dives, product updates, and practical guides for building with brand data\. ## Frequently askedquestions Everything you need to know about integrating and scaling with context\.dev\. ## Ship an agent that actuallyknows things\. Free tier, 10\-minute integration, and the same API powering agents at Mintlify, daily\.dev, and Propane\. No credit card to start\.

Similar Articles

Context.dev

Product Hunt

Context.dev provides a single API for scraping, enriching, and extracting data from the internet.

Show HN: A working reference implementation of context engineering

Hacker News Top

A working reference implementation of context engineering — a discipline for designing, retrieving, and injecting organizational context into AI systems to produce accurate, domain-specific outputs. The repo demonstrates five components (corpus, retrieval, injection, output, enforcement) running against Amazon Bedrock with Claude.