{
  "$schema_doc": "https://gtmstacker.com/registry/schema/entry.schema.json",
  "stability": "emerging",
  "generator": "agentic-media-registry",
  "generated_at": "2026-09-29T00:00:00Z",
  "id": "com.gtmstacker.registry/tool/textsnap",
  "type": "tool",
  "slug": "textsnap",
  "canonical_url": "https://gtmstacker.com/registry/tool/textsnap/",
  "title": "textsnap",
  "description": "textsnap converts any image, screenshot, or webpage into plaintext locally via OCR, running fully offline after the first model download. Open source: yes (MIT project; models Apache-2.0); self-hostable (pip install, ONNX runtime, no GPU, no cloud, no API keys). Free; ~186 stars; maker kouhxp.",
  "category": "ai-infrastructure",
  "tags": [
    "ai-infrastructure",
    "ocr",
    "offline",
    "local-inference",
    "self-hostable",
    "text-extraction"
  ],
  "status": "active",
  "revision": 1,
  "content_hash": "637f134b3ba787234e0fed9c660b1a179a6919e34b7ec6ad4f85670365f4d941",
  "date_published": "2026-09-29T00:00:00Z",
  "date_modified": "2026-09-29T00:00:00Z",
  "source": {
    "name": "kouhxp · GitHub",
    "url": "https://github.com/kouhxp/textsnap"
  },
  "license": "MIT",
  "one_liner": "textsnap turns any image, screenshot or webpage into plaintext on-device via ONNX OCR, running fully offline after one model download — no GPU or API keys.",
  "open_source": "yes",
  "self_hostable": "yes",
  "pricing_model": "free",
  "who_its_for": "A builder feeding text into an agent or document pipeline who needs OCR that runs entirely on-device — no cloud OCR API, no GPU, no keys — so image, screenshot, and webpage content becomes plaintext without data leaving the machine.",
  "who_its_not_for": "A team that needs cloud-scale batch OCR with layout reconstruction, table structure, or handwriting accuracy guarantees, or that prefers a managed OCR API over running a local ONNX model.",
  "aliases": [
    "textsnap",
    "kouhxp/textsnap"
  ],
  "alternatives": [
    "tenderness"
  ],
  "secondary_categories": [],
  "last_verified": "2026-09-29",
  "evidence": {
    "claim_type": "vendor-claim",
    "source_id": "https://github.com/kouhxp/textsnap",
    "note": "MIT for the project, Apache-2.0 for the models, per repo; ~186 stars. Converts any image/screenshot/webpage into plaintext via OCR. Self-hostable: `pip install`, ONNX runtime, no GPU, no cloud, no API keys, fully offline after the first model download (kouhxp/textsnap, verified 2026-09-29)."
  },
  "caveats": "Verified from the primary repo (2026-09-29); no independent accuracy testing here. OCR quality on low-resolution, skewed, handwritten, or dense-layout inputs is not benchmarked here — evaluate on your own documents. Dual license: MIT for the project, Apache-2.0 for the bundled models. It downloads the model on first run, so the very first use requires network access even though later use is offline.",
  "lead": "textsnap converts any image, screenshot, or webpage into plaintext locally via OCR, and after the first model download it runs fully offline. Open source: yes (MIT for the project; the models are Apache-2.0); self-hostable with a `pip install` on the ONNX runtime — no GPU, no cloud, and no API…",
  "chunks": [
    {
      "index": 0,
      "heading_path": [],
      "est_tokens": 90,
      "text": "textsnap converts any image, screenshot, or webpage into plaintext locally via OCR, and after the first model download it runs fully offline. Open source: yes (MIT for the project; the models are Apache-2.0); self-hostable with a `pip install` on the ONNX runtime — no GPU, no cloud, and no API keys. It is free, has ~186 stars, and is maintained by kouhxp."
    },
    {
      "index": 1,
      "heading_path": [
        null,
        "What it does"
      ],
      "est_tokens": 227,
      "text": "and after the first model download it runs fully offline. Open source: yes (MIT for the project; the models are Apache-2.0); self-hostable with a `pip install` on the ONNX runtime — no GPU, no cloud, and no API keys. It is free, has ~186 stars, and is maintained by kouhxp.\n\ntextsnap is an on-device OCR primitive aimed at builders who need to turn visual content into text without calling a cloud OCR service. You point it at an image, a screenshot, or a webpage and it returns plaintext, running the model through the ONNX runtime on the local CPU — no GPU required and no API keys to manage. After the initial model download it works entirely offline, so nothing you OCR ever leaves the machine. It installs with `pip`, which makes it easy to drop into an agent or document-ingestion pipeline as the step that feeds text downstream. Open source: yes (MIT project; Apache-2.0 models); self-hostable; free."
    },
    {
      "index": 2,
      "heading_path": [
        null,
        "Provenance"
      ],
      "est_tokens": 187,
      "text": "manage. After the initial model download it works entirely offline, so nothing you OCR ever leaves the machine. It installs with `pip`, which makes it easy to drop into an agent or document-ingestion pipeline as the step that feeds text downstream. Open source: yes (MIT project; Apache-2.0 models); self-hostable; free.\n\n- MIT for the project, Apache-2.0 for the models, per repo; ~186 stars.\n- Converts any image/screenshot/webpage into plaintext via OCR.\n- Self-hostable: `pip install`, ONNX runtime, no GPU, no cloud, no API keys; fully offline after the first model download (kouhxp/textsnap, verified 2026-09-29).\n- Surfaced via the GTM Stacker studio daily pull (2026-09-29 pass); license/facts verified against the primary repo 2026-09-29."
    },
    {
      "index": 3,
      "heading_path": [
        null,
        "Why it matters for a GTM stack"
      ],
      "est_tokens": 253,
      "text": "stars. - Converts any image/screenshot/webpage into plaintext via OCR. - Self-hostable: `pip install`, ONNX runtime, no GPU, no cloud, no API keys; fully offline after the first model download (kouhxp/textsnap, verified 2026-09-29). - Surfaced via the GTM Stacker studio daily pull (2026-09-29 pass); license/facts verified against the primary repo 2026-09-29.\n\nGTM pipelines constantly hit text that is trapped in images — a screenshot of a pricing page, a scanned contract, a chart in a deck, a competitor's site rendered as an image. textsnap turns that into plaintext an agent can read, and it does it locally, which matters when the source contains confidential prospect or deal information you would rather not send to a third-party OCR API. The honest read: OCR accuracy on messy real-world inputs — skew, low resolution, dense tables, handwriting — is not benchmarked here, so test it on your actual documents; and while it runs offline afterward, the first run needs network access to fetch the model."
    }
  ],
  "alternates": {
    "markdown": "https://gtmstacker.com/registry/tool/textsnap/index.md",
    "html": "https://gtmstacker.com/registry/tool/textsnap/",
    "json": "https://gtmstacker.com/registry/tool/textsnap/index.json",
    "server_json": "https://gtmstacker.com/registry/tool/textsnap/server.json"
  },
  "jsonld": {
    "@context": "https://schema.org",
    "@graph": [
      {
        "@type": "WebSite",
        "@id": "https://gtmstacker.com/#website",
        "url": "https://gtmstacker.com/",
        "name": "GTM Stacker Agent Registry",
        "description": "A daily-updated, agent-native registry of open-source tool discoveries, tool updates, and curated news for the go-to-market / RevOps engineering niche. Machine-readable first: agents can discover, parse, page, and delta-sync it without scraping HTML.",
        "inLanguage": "en",
        "publisher": {
          "@id": "https://gtmstacker.com/#organization"
        }
      },
      {
        "@type": "Organization",
        "@id": "https://gtmstacker.com/#organization",
        "name": "GTM Stacker",
        "url": "https://gtmstacker.com",
        "description": "The growth-systems practice of Theo Popov: AI-native enrichment, outbound, content engines and internal tooling for startups and venture programs. Its agent-native media property, the GTM Stacker Agent Registry, maintains a daily-updated catalog of open-source go-to-market and RevOps tools that both people and AI engines can discover, compare, and cite.",
        "foundingDate": "2024-08",
        "knowsAbout": [
          "go-to-market engineering",
          "RevOps",
          "sales automation",
          "marketing operations",
          "open-source software",
          "AI agents"
        ],
        "founder": {
          "@type": "Person",
          "@id": "https://gtmstacker.com/#founder",
          "name": "Theo Popov",
          "jobTitle": "Growth Operations & GTM Systems",
          "url": "https://gtmstacker.com/about/",
          "sameAs": [
            "https://www.linkedin.com/in/theo-popov",
            "https://x.com/Theo_Popov",
            "https://github.com/theopopov"
          ],
          "worksFor": {
            "@id": "https://gtmstacker.com/#organization"
          }
        },
        "sameAs": [
          "https://www.linkedin.com/company/gtmstacker",
          "https://www.youtube.com/@gtmstacker",
          "https://www.instagram.com/gtmstacker/",
          "https://www.tiktok.com/@gtmstacker"
        ],
        "mainEntityOfPage": "https://gtmstacker.com/registry/about/"
      },
      {
        "@type": "SoftwareApplication",
        "@id": "https://gtmstacker.com/registry/tool/textsnap/#software",
        "name": "textsnap",
        "identifier": "kouhxp/textsnap",
        "description": "textsnap converts any image, screenshot, or webpage into plaintext locally via OCR, running fully offline after the first model download. Open source: yes (MIT project; models Apache-2.0); self-hostable (pip install, ONNX runtime, no GPU, no cloud, no API keys). Free; ~186 stars; maker kouhxp.",
        "applicationCategory": "DeveloperApplication",
        "url": "https://gtmstacker.com/registry/tool/textsnap/",
        "datePublished": "2026-09-29T00:00:00Z",
        "dateModified": "2026-09-29T00:00:00Z",
        "isPartOf": {
          "@id": "https://gtmstacker.com/#website"
        },
        "license": "https://spdx.org/licenses/MIT.html",
        "codeRepository": "https://github.com/kouhxp/textsnap",
        "keywords": "ai-infrastructure, ocr, offline, local-inference, self-hostable, text-extraction",
        "author": {
          "@type": "Organization",
          "name": "kouhxp",
          "url": "https://github.com/kouhxp",
          "sameAs": [
            "https://github.com/kouhxp/textsnap"
          ]
        },
        "offers": {
          "@type": "Offer",
          "price": 0,
          "priceCurrency": "USD"
        },
        "isSimilarTo": [
          {
            "@type": "SoftwareApplication",
            "name": "Tenderness",
            "url": "https://gtmstacker.com/registry/tool/tenderness/",
            "applicationCategory": "DeveloperApplication",
            "offers": {
              "@type": "Offer",
              "price": 0,
              "priceCurrency": "USD"
            }
          }
        ]
      },
      {
        "@type": "BreadcrumbList",
        "@id": "https://gtmstacker.com/registry/tool/textsnap/#breadcrumb",
        "itemListElement": [
          {
            "@type": "ListItem",
            "position": 1,
            "name": "GTM Stacker Registry",
            "item": "https://gtmstacker.com/registry/"
          },
          {
            "@type": "ListItem",
            "position": 2,
            "name": "AI Infrastructure",
            "item": "https://gtmstacker.com/registry/category/ai-infrastructure/"
          },
          {
            "@type": "ListItem",
            "position": 3,
            "name": "textsnap",
            "item": "https://gtmstacker.com/registry/tool/textsnap/"
          }
        ]
      }
    ]
  },
  "tool": {
    "name": "kouhxp/textsnap",
    "repository": {
      "url": "https://github.com/kouhxp/textsnap",
      "source": "github"
    }
  }
}
