<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:dc="http://purl.org/dc/elements/1.1/">
  <channel>
    <title>DEV Community: manish Gudimetla</title>
    <description>The latest articles on DEV Community by manish Gudimetla (@manish_gudimetla_6d37f97a).</description>
    <link>https://dev.to/manish_gudimetla_6d37f97a</link>
    <image>
      <url>https://media2.dev.to/dynamic/image/width=90,height=90,fit=cover,gravity=auto,format=auto/https:%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Fuser%2Fprofile_image%2F4162636%2Fc4d5541e-e22b-4523-ae7e-25d2b6da8f2c.jpg</url>
      <title>DEV Community: manish Gudimetla</title>
      <link>https://dev.to/manish_gudimetla_6d37f97a</link>
    </image>
    <atom:link rel="self" type="application/rss+xml" href="https://dev.to/feed/manish_gudimetla_6d37f97a"/>
    <language>en</language>
    <item>
      <title>Screenshot APIs all send your page content to their servers — built a self-hosted one, looking for feedback</title>
      <dc:creator>manish Gudimetla</dc:creator>
      <pubDate>Mon, 05 Oct 2026 01:17:04 +0000</pubDate>
      <link>https://dev.to/manish_gudimetla_6d37f97a/screenshot-apis-all-send-your-page-content-to-their-servers-built-a-self-hosted-one-looking-for-5371</link>
      <guid>https://dev.to/manish_gudimetla_6d37f97a/screenshot-apis-all-send-your-page-content-to-their-servers-built-a-self-hosted-one-looking-for-5371</guid>
      <description>&lt;p&gt;Ran into this building an agent that needed to look at web pages.&lt;/p&gt;

&lt;p&gt;Every screenshot API is SaaS, which is fine, until you notice the&lt;br&gt;
page-understanding part runs on their model. So the text of every page you&lt;br&gt;
capture gets sent to a third party and processed there. For a public marketing&lt;br&gt;
page, who cares. For anything behind a login with real data on it, that's a&lt;br&gt;
problem I couldn't get around.&lt;/p&gt;

&lt;p&gt;Looked for a self-hostable one. The MCP screenshot servers that exist (Urlbox,&lt;br&gt;
ScreenshotOne) are good but all point at the vendor's cloud. Didn't find one you&lt;br&gt;
run yourself, so I built it.&lt;/p&gt;

&lt;p&gt;Runs with docker compose. The extraction part runs on whatever model you point&lt;br&gt;
at it — including Ollama on localhost, in which case the page text never leaves&lt;br&gt;
the machine that rendered it. AGPL-3.0.&lt;/p&gt;

&lt;p&gt;Honest about where it's weak: no security audit, extraction quality is decent&lt;br&gt;
not amazing, and you can't capture your own private-IP hosts yet because the&lt;br&gt;
SSRF guard blocks RFC1918 with no opt-out (allowlist is the next thing I'm&lt;br&gt;
building). I also benchmarked it against two hosted APIs and the latency margins&lt;br&gt;
weren't reproducible across runs, so I'm not claiming it's faster — the data and&lt;br&gt;
the caveats are both in the repo.&lt;/p&gt;

&lt;p&gt;github.com/route1-ai/shotbase&lt;/p&gt;

&lt;p&gt;The thing I'm unsure about: is "run it yourself so page content stays in your&lt;br&gt;
network" actually the useful part here, or do people mostly just want&lt;br&gt;
screenshots and I've over-thought it? Genuinely asking — I built this for my own&lt;br&gt;
use case and I don't know if it generalizes.&lt;/p&gt;

</description>
      <category>docker</category>
      <category>llm</category>
      <category>opensource</category>
      <category>privacy</category>
    </item>
  </channel>
</rss>
