different species of crabsoft-shell crab
4
15 Comments

A tool I reviewed audited my own site and found a bug I'd missed

Started Solo Stack Lab a couple weeks ago — reviews of AI/SEO/automation tools for solo people running content or affiliate sites. Built mostly with Claude Code (an AI coding agent). Rule: sign up and actually use the tool before writing about it.

Ran Frase (one of the tools I reviewed) against my own site. It flagged zero og:image tags, site-wide — every link I'd ever shared had no preview image. I'd written "social previews done" in my notes days earlier. Wrong. Fixed same day, put it in the review instead of quietly patching it.

Other stuff from the first couple weeks, no particular order: Surfer SEO's "start for free" button charges $59/mo immediately, no free tier. Make logged "success" on an automation while an email silently never sent because of an OAuth scope thing. n8n rejected my affiliate application with zero explanation, while three other programs approved same day.

Current traffic per GSC: 19 impressions, 0 clicks, average position 36, three months. Early.

Anyone done the early-content-site thing — did backlinks or more content move the needle first for you?

posted to Icon for group Building in Public
Building in Public
on August 21, 2026
  1. 1

    This captures something that silently kills shipping velocity: the gap between "I noted it as done" and "it measurably works in production."

    The og:image thing is sharp because it's not a missing feature - it's a completed task on a checklist. You marked it done, moved on. But your measurement system was the person-who-wrote-the-note, not the automated verifier.

    I notice teams that move fast maintain parallel measurement systems: human checklist (fast, forgetful) and automated audit (slower to set up, reliable, finds three more things). The audit doesn't feel like "more work" - it's the difference between "we think we solved it" and "we know we did."

    And that gap compounds across 50 features. Your checklist says 50/50 done. The tool says 22/50. Which measurement system determines what you fix next?

  2. 1

    The two findings that make this post work are both claimed-versus-actual gaps: your notes said social previews done while Frase found zero tags, and Make logged success while the email never sent. I would make that exact gap a fixed section of every review instead of a side lesson. 'What the tool says happened versus what we could verify' is harder for competitors to copy than opinion, and it gives readers a reusable question instead of another rating. On the traffic question: at position 36 with 19 impressions, I would publish more pages first, but only pages with a distinct verified finding; thin filler will not give you anything worth ranking.

  3. 1

    The 'Make logged success while the email silently never sent' line is the one that would haunt me, because a tool reporting success on a partial failure is worse than an honest error. I run automations daily and now treat every green checkmark as unverified until I check the actual downstream artifact, the sent message or the written row, not the tool's own status. On content vs backlinks: at position 36 with single-digit impressions, links mostly lift pages you already rank 5 to 15 for, so more indexed pages usually give Google more surface to find those first. What's the smallest publishing cadence you can actually sustain without burning out?

    1. 1

      Realistically about one new piece every day or two right now, but that's front-loaded since I'm testing a backlog of untested tools. Honest sustainable pace once that backlog runs out: probably one solid piece a week. Rather publish less and keep the firsthand-testing rule than pad it out with filler.

  4. 1

    This is why I paid for a Zarek audit before PH. Found 8 issues, 3 critical. The "expert eye" catches what the builder is blind to after 100 hours in the code.

    Best $0 I spent (it was a free community audit). Worth 10x more than any ads.

    1. 1

      what's zarek? not heard of that one before.

      1. 1

        Zarek is an AI launch coach — basically an automated audit that scans your landing page for conversion-killing issues before you go live. Found 8 problems on mine, 3 critical.

        I learned the hard way: builders are blind to their own product after 100 hours in the code. An outside eye (even AI) catches what you don't.

        I'm building ThumbRank (AI thumbnail scoring for YouTubers) — launching PH Aug 25. Would love to trade feedback if you're launching soon too. My design eye is fresh, your landing page eye is fresh. Win-win.

  5. 1

    The part about writing “social previews done” in your notes and then discovering they were completely broken is painfully relatable. 😅

    I think that's also why actually using a tool before reviewing it is such a good rule. A feature can look perfect on paper while something completely different breaks once it touches a real workflow.

    And honestly, 19 impressions / 0 clicks after only a couple of weeks doesn't sound particularly worrying to me. I'd probably prioritize publishing consistently first, then use GSC to see which pages/queries start getting impressions before going hard on backlinks.

    Curious to see how the traffic changes once you've got a bigger sample size.

    1. 1

      yeah that's basically the plan. more published pages before going hard on backlinks. will report back once there's an actual sample size worth looking at instead of single digits.

  6. 1

    The "og:image done" → "actually zero tags" gap is pure measurement clarity in action. Your notes were a mental model. The tool was reality.

    Most founders never run this kind of audit because it's easier to trust our own annotations. But what you're doing here is outsourcing verification to something that doesn't know how to lie - it just counts what exists.

    That 19 impressions / 0 clicks is interesting because it actually tells you something specific: you're getting search visibility but the click-through is catastrophic. Most people would blame "position 36" and build more content. But your stack audit already found a concrete friction point (og:image, and presumably those "success" logging failures).

    The measurement system determines what you can see. You built one that surfaces friction instead of just celebrating views.

    1. 1

      yeah, that's the part that surprised me too. expected the tool to catch keyword stuff, not "your own notes were wrong about what shipped."

  7. 1

    The “I already marked this as done” part is the most interesting to me. It shows how easily our own notes become assumptions instead of evidence. Automated audits are valuable not just because they find bugs, but because they challenge what we think we've already verified. Curious whether you'll start treating these audits as a recurring pre-publish check.

    1. 1

      planning to, yeah. cheap insurance compared to publishing something broken and not noticing for a week.

  8. 1

    The strongest part is the willingness to let the product audit the founder’s own assumptions. That makes the “review” model more credible because the process is producing findings, not just content about tools.

  9. 1

    This comment was deleted 4 hours ago.

  10. 1

    This comment was deleted 4 hours ago.