S3 can list your files but can't look inside them. Point StackGrep at a prefix: we index every text object, keep up as files change, and answer searches over all of it in milliseconds. Your data stays in your bucket; we only read it.
PUT /api/collections/logs/source
{"bucket": "acme-logs", "prefix": "app/2026/", "region": "us-east-1"}
GET /api/collections/logs/search?q=OutOfMemoryError&filter=
→ every object that mentions it, with the linesearch_collection(query: "88213")count_collection(query: "\"ssn\"")search_collection(query: "max_connections\s*=\s*([5-9][0-9]{2}|[0-9]{4,})", regex: true)Read-only access to the one prefix, granted by a bucket policy we show you, plus a verify file you write there so nobody can point us at a bucket that isn't theirs. Remove the policy and we stop.
Every text object under the prefix, its key as the document id. Gzipped files are unzipped; binary files and objects over 4 MB are skipped.
We read it to build a compact index, kept in our object storage and cached where searches run. Delete the source and the collection goes with it.
StackGrep Engine is onboarding teams from the waitlist. Bring your documents or your bucket.