Systems Lab

Agent skill

list-dedup

Deduplicate and clean contact/company lists before enrichment so you never pay ColdIQ credits twice for the same record.

activeSelf-containedInstructions only343 words

Filed under CRM and RevOps.

From Cold-IQ/coldiq-marketplace-skills · 17 skill entries · 1 · pushed 2026-09-22

What it does when it runs

Deduplicate and clean contact/company lists before enrichment so you never pay ColdIQ credits twice for the same record. Use when combining lists from multiple sources, removing duplicate contacts or companies, normalizing LinkedIn URLs, applying per-company contact caps, or cross-referencing against a TAM. Triggers on "dedup", "deduplicate", "remove duplicates", "clean the list", "combine sources", "merge lists", "contact cap per company", "normalize LinkedIn URLs". Do NOT use for enrichment/email-finding (see contact-enrichment), search (see coldiq-search-enrich), or scoring (see tam-scoring).

Automated analysis of the skill and the 1 file bundled beside it. A skill’s own description is written to be selected by an agent, so it describes the job and not the dependencies.

Keys and connectors you must supply
None found.
Hosts it reaches
No third-party host appears in the skill or its bundled files.
Tool permissions it declares
No allowed-tools in the frontmatter. It only issues instructions, so there is nothing to bound.
Actions present in the files
None. Instructions only.

Ask about list-dedup

Opens your assistant with this page's verified links already in the prompt.

Is this safe to install?ClaudeChatGPT
Adapt it to my stackClaudeChatGPT
What else do I need for it to workClaudeChatGPT
Rather ask a human? Talk to Cheetah
git clone --depth 1 --filter=blob:none --sparse https://github.com/Cold-IQ/coldiq-marketplace-skills.git /tmp/coldiq-marketplace-skills
git -C /tmp/coldiq-marketplace-skills sparse-checkout set "skills/list-dedup"
mkdir -p ~/.claude/skills/list-dedup
cp -R "/tmp/coldiq-marketplace-skills/skills/list-dedup/." ~/.claude/skills/list-dedup/

Picked up without a restart. A project skill of the same name is shadowed by your personal one. For one repository only, swap ~/.claude/skills for .claude/skills. Claude Code docs ↗

Or take the whole library

This repo ships a .claude-plugin manifest, so Claude Code can install all 17 skills at once. Plugin skills are invoked as /<plugin>:<skill>, so they never collide with your own.

/plugin marketplace add Cold-IQ/coldiq-marketplace-skills
/plugin

The folder is the same in every client that implements the format — 46 of them — so if yours is not above, only the destination changes.

Reproduced in full from Cold-IQ/coldiq-marketplace-skills/blob/495c1b4256db8c6683ac1ff59a093899228eca73/skills/list-dedup/SKILL.md, which is licensed MIT (repository). 343 words, 8 headings.

List Dedup

A pure data-processing utility: dedup and clean BEFORE any ColdIQ enrichment. This is the single biggest credit saver — every duplicate you remove is a paid find/enrich call you don't make. See resources/credit-optimization.md. This skill makes no API calls.

When to dedup

  • Before enrichment (never pay to enrich the same person twice)
  • After combining multiple sources (database search + enrichment + ad scrape)
  • After scoring (collapse duplicates, keep the highest-scored row)

Primary key: LinkedIn URL (normalize first)

import re
def normalize_linkedin(url):
    if not url: return ""
    return re.sub(r'\?.*$', '', url).rstrip('/').lower().replace('https://www.', 'https://')

def dedup_people(rows):
    seen, out = set(), []
    for r in rows:
        key = normalize_linkedin(r.get('linkedin_url', '')) or (r.get('email','').lower())
        if key and key not in seen:
            seen.add(key); out.append(r)
    return out

Secondary key: company name / company LinkedIn URL

Normalize company names (strip Inc/LLC/Ltd suffixes, lowercase) or, better, dedup on the company LinkedIn URL when present.

Multi-source combine pattern

  1. Normalize columns — map each source's columns to one standard schema (full_name, title, linkedin_url, email, company_name, company_domain, company_linkedin, source, score).
  2. Load all source CSVs.
  3. Dedup people by normalized LinkedIn URL (fallback email); companies by company LinkedIn / name.
  4. Sort & export — Tier 1 first, then by persona, then company name.

Per-company contact caps

Cap 5–10 contacts per company so you don't blast one account. Rank by title priority (C-level/VP/Director > Manager > IC), keep the top N per company_domain.

Cross-reference against TAM

Left-join the contact list against the scored TAM on company_domain to attach Tier, and drop contacts at DQ'd companies before enrichment.

Checklist

  • Normalize LinkedIn URLs before comparing
  • Dedup people by LinkedIn URL (fallback email)
  • Dedup companies by LinkedIn URL / normalized name
  • Map all sources to one schema before combining
  • Keep highest-scored row on duplicate
  • Apply per-company contact cap (5–10)
  • Cross-reference against TAM, drop DQ companies
  • THEN enrich (see contact-enrichment)

Files bundled with it

These load only when the skill asks for them, so they cost nothing until it runs.

Other skills for the same job

Different authors, same problem. Matched on the words in the skill name, across every library in the catalogue except this one.

Need help setting it up?

This page tells you what list-dedup does and what it needs. Cheetah builds the agent setup it runs inside: data, CRM, sequencing and the guardrails.

Book a call →

The directory stays free. There is nothing gated behind this.