Prepare Your Dataset
Load your list of records (e.g., company names, user accounts) that you suspect contain duplicates. This can be a large JSON file, a CSV, or a database query result.
Efficiently clean up large databases by identifying and merging duplicate or near-duplicate records. Use a decision model to get confidence scores for potential matches, allowing for automated merging or manual review.

Step by step
Follow the sequence once, then adapt the prompts, checks, and handoffs to your own setup.
6 steps
Load your list of records (e.g., company names, user accounts) that you suspect contain duplicates. This can be a large JSON file, a CSV, or a database query result.
Feed the list of records to Jev. Configure it to compare entries and identify potential duplicates, such as 'Cedar Grove Office Products' vs 'Cedar Grove Office'.
Test on a representative sample first, then batch candidate comparisons and measure cost and accuracy before scaling.
For each potential match identified, Jev should return a confidence score indicating the likelihood that the two records are duplicates. This score is key to automating the next step.
Define a confidence score threshold to determine what to do with the matches. For example, evaluate a 99% confidence threshold against labeled examples before enabling automatic merges; keep a backup and review uncertain matches.
For pairs with a lower confidence score (e.g., between 75% and 99%), flag them for manual review by a human operator. This creates a human-in-the-loop system that balances automation with accuracy.
For the pairs flagged for review, pass them to a cheap generative model (such as one accessed through OpenRouter). Prompt the model to explain why it thinks the pair is a match, providing a qualitative reason to assist the human reviewer.
Turn an idea into a PRD, user stories, and a plan.
Keep building

Create a custom, real-time incident command dashboard using ChatGPT Sites. This tool pulls live, user-specific data from Slack, Notion, and Google Calendar to keep you informed without disrupting your engineering team.

Create a fluid, hands-free to-do list that updates instantly as you speak. Use a multi-step classification process to understand commands, match tasks, and execute functions in real-time without needing to pause.

Create an intelligent command bar (Command-K) that not only finds tools but executes actions within them from a single natural language query. Use a layered decision model to route requests to the right app and then to the right function.
Join 100,000+ product managers who use ChatPRD to write better docs, align teams faster, and build products users love.