Bow Tie Kreative Intel System

File 06

PR, Reputation and Narrative Tool Library

Research snapshot: 2026-09-01. “Free” describes access; “open source” describes code. Neither automatically authorizes commercial collection, retention, training, republication, or lead generation.

1. Access hierarchy

  1. official API or first-party export;
  2. publisher RSS/Atom/JSON Feed;
  3. licensed archive/dataset;
  4. publisher permission;
  5. manual citation-only review;
  6. skip the source.
Source → purpose check → access authorization → collect → rights classify
→ normalize → deduplicate → verify → cluster → score → human review
→ evidence → outreach

2. News, media and web mentions

06.2-news-media-and-web-mentions.t1
Tool/sourceStatusInterface and best useLimits/cautionsOfficial source
GDELTfree/open-access data service; not an OSI software license for corpusAPIs, bulk and BigQuery → URLs, events, GKG, timelines, tone/context; global discovery and historybroad but uneven; 15-minute updates; dedupe syndication; “tone” ≠ public opinion; publisher rights persistProject · Data
Media Cloudaccount/API; official Python client Apache-2.0REST/Python → story counts/metadata/URLs/source collections; comparative panels/SOVcurated archive, not universal/social firehose; client license ≠ content rightsAPI guide · Client
Guardian Open Platformfree developer key for noncommercial useREST → Guardian search/metadata/text where enabled1 request/sec, 500/day; commercial mining/sentiment requires arrangementAccess
NewsAPIproprietary dev planREST → headlines/snippets/URLsfree: 100/day, 24-hour delay, one-month history, localhost/development only; not production/commercialPricing
GNews APIproprietary dev/noncommercial planREST → article metadata/snippets/URLsfree: 100/day, 10 results/request, 12-hour delay, 30 days, truncated contentPricing
Google Alertsfree UI/emailexpressions → email discoveryno supported API/completeness/count; do not scrape search resultsHelp
Talkwalker Alertsfree proprietary alert serviceBoolean query → email/RSS across news/blogs/forums/web; vendor advertises Xnot paid Talkwalker coverage; no general API, opaque sampling; validate source claimsAlerts
Google Trendsfree UI/CSV/RSS; alpha API limitedsearch-interest series/trending queries → relative interestsampled/relative, not volume/SOV/sales/sentiment; alpha not generally availableOfficial API
Common Crawlfree corpus; downloader MIT/Apache-2.0WARC/WAT/WET/index → historic pages/metadata/textincomplete/not live; inclusion does not erase copyright/privacy/deletion dutiesProject
Wayback Machinefree public archiveURL/history → snapshotsincomplete/removable/throttled; source/capture time required; no bulk republicationDevelopers · Terms

3. Social and review sources

06.3-social-and-review-sources.t1
SourceStatus/interfaceAppropriate useLimits and stop rulesOfficial source
Reddit Data APIOAuth API; free access conditionalapproved-purpose subreddit/search/post/comment discoveryroughly 100 queries/min OAuth baseline but headers govern; commercial uses may need agreement; deletion synchronization; no private communities/web scraping/limit evasionData API Terms · Developer Terms
Bluesky/AT Protocol Jetstreamdual MIT/Apache-2.0 codeWebSocket/repo stream → public JSON events; strong narrative discoveryrate rules and deletion handling; public does not erase privacy/context dutiesJetstream · Repo
MastodonAGPL-3.0 server; per-instance APIspublic per-instance accounts/hashtags/statuses → JSON/streamdefault around 300 requests/5 min but instance-specific; partial fediverse visibility; exclude private/unlisted targetsAPI · Limits
YouTube Data APIproprietary API with free quotasearch/channel/video/comment metadatasearch uses high quota; daily pool and endpoint costs apply; no general transcript scrape; caption download usually owner-authorizedGetting started · Costs
X APIproprietary pay-per-useapproved search/read useno free core; policy restricts redistribution, deletion sync, benchmarking and some repurposing; never scrape x.comPricing · Policy
Meta Content Library/APIcontrolled research accessqualified academic/nonprofit public-interest researchnot a general commercial connector; CrowdTangle discontinued 2024-08-14; no general FB/IG scrapingDocs
TikTok Research APIcontrolled nonprofit research accessqualified research on public videos/users/commentsnot commercial prospect profiling; eligibility and daily/request caps; no bypass via scrapingGetting started
TikTok Commercial Content APIapproval-gated transparencyads/commercial-content metadatajurisdiction scope (Europe-focused) and not a general comment/firehose sourceDocs
Google Places APIbilling required, monthly no-charge capsplace identity/rating and up to five review samplestiny relevance-ranked sample; attribution/caching/display rules; not corpus sentimentPlace resource · Pricing
Trustpilot APIskey/OAuth and contractual termsapproved business-unit/review retrieval including repliespublic access ≠ content license; deletion/refresh/rate rules; commercial partner use may require agreement; no scrapingPortal · Limits
Yelp Placestrial/paid proprietary APIpermitted business display/limited excerptsordinary integration not commercial review analysis; caching limits; site scraping prohibited; Yelp Insights for analysisIntro · Scraping policy
BBBmanual spot-check only unless licensed/permittednarrow citation of public recordterms restrict aggregation/republication/sales use; submissions not necessarily verified; no scraping/lead databaseTerms
Glassdoormanual citation or licensed dataqualitative employment-reputation context where relevant2026 terms restrict bots/scraping/mining/competitive use; do not automate or profile authorsTerms
LinkedInmanual public business review; approved owned-asset APIscurrent executive/company context from public business materialno broad people/listening API; user agreement prohibits scraping; do not harvest profiles/emails/relationshipsUser Agreement · Developer docs

4. Narrative verification

06.4-narrative-verification.t1
ToolStatus/licenseUseLimitationOfficial source
Google Fact Check Tools APIproprietary API-key servicequery → ClaimReview records and reviewer ratingspublisher assessment, not universal truth; preserve reviewer/sourceAPI
InVID-WeVerify/vera.ai pluginMITvideo keyframes/metadata/forensic launchers → analyst evidence bundledoes not autonomously prove truth; external services have termsProject · Repo
Junkipediacontrolled public-interest platformmulti-platform narrative/actor research for eligible usersdo not assume commercial eligibility or downstream rightsSite
HoaxyGPL-3.0 frontend; archived 2023historical diffusion design patternX dependency and archive status make it unsuitable for productionRepo

Keep these separate: observable statement, claim, named fact-check, corroboration, semantic narrative cluster, coordination indicator, verified fact. Similar timing/wording/hashtags do not prove coordinated inauthentic behavior.

5. Open-source ingestion and monitoring

06.5-open-source-ingestion-and-monitoring.t1
ComponentLicenseRoleRights/operations noteOfficial source
RSSHubAGPL-3.0route adapters → RSS/Atomadapter code does not authorize target collection; allow-list reviewed routesRepo
FreshRSSAGPL-3.0feed/OPML → analyst inbox/APIpreserve publisher links/rights; avoid unauthorized full-page fetchRepo
MinifluxApache-2.0lightweight feed store/API on PostgreSQLsame publisher content/deletion constraintsRepo
changedetection.ioApache-2.0permitted URL + selectors → diffs/webhooks/RSSpolite allow-list; no CAPTCHA/stealth bypass or login/personal pagesRepo
ScrapyBSD-3-Clausepermitted URLs + extraction rules → structured recordsframework license is not scraping permissionRepo
Trafilaturacurrent Apache-2.0; older releases differallowed HTML/feeds/sitemaps → cleaned text/metadatano paywall/auth/anti-bot bypass; pin/version licenseRepo

Use Miniflux or FreshRSS—not both—as the canonical feed store. Add RSSHub only for reviewed adapters.

6. NLP, clustering and network analysis

06.6-nlp-clustering-and-network-analysis.t1
LibraryLicenseBest roleLimitationOfficial source
VADERMITtransparent English short-text sentiment baselinepoor on jargon, languages, sarcasm, quotations/mixed sentimentRepo
spaCyMIT libraryNER, PII redaction, rules/classifierspipeline/model licenses and language quality varyRepo
TransformersApache-2.0 libraryselected sentiment/stance/NER/summarization modelseach model has separate license/data/bias/compute; some noncommercialRepo
Sentence TransformersApache-2.0 libraryembeddings, dedupe, clustering and semantic comparisonmodel license/threshold/language calibrationRepo
BERTopicMITcomplaint/narrative topic discoverytopics depend on corpus/model/parameters/seedRepo
KeyBERTMITcandidate keyphrases/query expansionrelevance ≠ demand/importance; human label reviewRepo
scikit-learnBSD-3-Clausetransparent calibrated domain classifiers/evaluationneeds representative labels and drift testsRepo
NetworkXBSD-3-Clausecitation/co-mention/diffusion graphscentrality ≠ influence, intent or coordinationRepo

Maintain a model registry: package/model version, license, intended use, languages, evaluation set, precision/recall/F1 and error slices, retirement date.

7. Sources not to scrape

Do not build automated commercial collectors for LinkedIn profiles/posts, Glassdoor reviews, BBB complaints, Yelp pages/reviews, Google Search/Maps HTML, general Facebook/Instagram content, general TikTok feed/comments, x.com, Reddit web pages/private communities, login-walled/private/deleted content, or leaked credential/dark-web dumps. Use official APIs/exports, permission/licensing, manual citation-only review, or exclusion.

8. PR metrics

Comparison set = target + up to three selected comparators

Unweighted SOV = target deduplicated qualifying story clusters
               / all comparison-set qualifying story clusters

Weighted story value = source quality × prominence × capped reach proxy × relevance
Weighted SOV = target weighted story value / all comparison-set weighted story value

Net sentiment = (positive − negative) / (positive + negative + neutral + uncertain)

Issue share = negative items assigned to issue / all relevant negative items

Crisis velocity z = (current negative volume − baseline mean) / baseline standard deviation
                    only when baseline standard deviation > 0;
                    otherwise report absolute and percentage change without a z-score

Response gap = actionable negative items unresolved beyond SLA / actionable negative items

Bootstrap uncertainty over story clusters or source-days, not syndicated URLs. Report denominator, languages, source panel, duplicate rate, missing periods, classifier F1, false positives and likely blind spots.

When a metric denominator is zero, report not estimable rather than zero, infinity, or an imputed rate.

If issue coding is multi-label, call the metric Issue prevalence in collected items; its category shares may sum above 100%. Use Issue share only for a mutually exclusive coding scheme.

9. LAKA ladder

06.9-laka-ladder.t1
StateReputation intelligence system
Baselinealerts/RSS and manual weekly evidence review
Minordaily permitted APIs/feeds, canonicalization and duplicate clustering
Majormultilingual entity resolution, calibrated sentiment/stance, issue taxonomy and competitor panels
Structuralevent-sourced evidence graph, rights registry, deletion sync, method versioning and CRM evidence links
Paradigmnarrative digital twin with scenario/intervention measurement; no causal claim without experimental/causal design

Every opportunity carries four separate labels: observed fact, analytic inference, financial assumption and recommended action.


06-pr-reputation-tool-library.md · 136 lines · 15280 bytes · SHA-256 8784acae48cb8359