Discussion about this post

User's avatar
Anlam Kuyusu's avatar

How did your friend upload the entire text of the book to Claude? Can he do the same for your books?

Duane McMullen's avatar

What Ariel Krakowski did requires advanced knowledge of how to use AI. It would be very interesting were he to comment on his process. I doubt it was a one shot query with an answer. In my experience, AIs go off the rails pretty quickly if under supervised.

While I would be delighted for Ariel to show how this is wrong, for his process I would imagine:

- ask for a chapter by chapter list of verifiable claims, output in some kind of machine legible format. The completion of this step might take a while.

- with that output, ask the AI to verify each claim and classify the result as 'verified', 'factual error', 'framing problem', 'contradiction', or whatever categorization you think best; the results go in an updated output machine readable file that adds the results of the search to each verifiable claim in step one. The completion of this step will definitely take 'a while' (depending on your plan, how much you want to pay extra for speed etc etc).

- with that output, ask it to analyze the results and present them as a single-file interactive HTML dashboard displaying summary cards, category filters, expandable details, and citations.

In Claude Opus, that would cost ~$20 (a rough guess). In an open weights model, perhaps $0.20 but more babysitting might be required through the steps.

To improve process, before you executed the plan you'd first give it to the AI and ask it to critique and refine. For instance, perhaps the AI has a better idea about how to categorize the errors. Frequently, the AI will tell me something I'd not have thought of that changes the plan ('clever!), or reminds me of something stupid and obvious that I'd overlooked (slaps forehead). With that critique, you update the plan and then execute.

As another improvement, you'd take the file of claims, errors and results from step 2 and give it to another AI to ask it to verify and critique the results. You'd then decide how to change the error file before proceeding to step 3, producing the HTML dashboard.

This is not to say that Ariel should have done more than he did. What he did was exceptionally impressive for someone who has not been deeply working with AI. However, once one is at Ariel's level of competence at working with AIs, then you always know how you could spend more money/time to get even better results.

18 more comments...

No posts

Ready for more?