Urgent.News

What's breaking now, across thousands of outlets.

Tech

A 10-Point Scorecard for Vetting Telegram Channels Before Adding Them to Your Collection Set

Every Telegram OSINT collection set starts as a pile of channels someone pasted from a Twitter thread. Most of that pile is junk, and the junk is expensive: every junk channel in your collection set adds review load, dedup noise, and false-alarm alerts forever. I classify candidate channels on five axes before they enter the set. Score each 0-2; keep at 7+, review at 5-6, reject below. 1. Origin…

Creating a reliable Telegram OSINT collection set begins with carefully vetting the channels you add. To ensure the quality of your collection, I recommend using a 10-point scorecard that assesses candidate channels across five key criteria. Each criterion is scored on a scale of 0 to 2, with a minimum overall score of 7 for inclusion in the set. Channels are then reviewed again with scores between 5 and 6, while anything below that threshold is rejected.

The scorecard covers five main axes: origin, language consistency, cadence regularity, audience quality proxy, and link hygiene. To evaluate the origin of a channel, check if it creates original reporting or simply republishes content. Look for original photos, videos, and consistent watermarks, as well as posts that other channels quote and link to. Channels that only republish content without adding original reporting should score 0. Pure republishers are not useless, but they are not valuable sources either.

Language consistency is another important factor. A well-run reporting channel should consistently report in one language with a clear target audience. If a channel switches languages frequently in an attempt to attract more engagement, it is likely a content farm. You can automate this assessment using a language-detect script to flag any language inconsistencies.

Cadence regularity assesses the posting frequency of the channel. Real operational channels, such as military administration, emergency services, or shipping feeds, tend to post in bursts with structured rhythms and quiet periods. In contrast, content farms typically post at a consistent, machine-like cadence throughout the day, including odd hours like 06:00 UTC. Analyzing the posting series using autocorrelation can help separate these types of channels effectively.

The audience quality proxy evaluates the quality of the channel's audience. For channels with more than 1000 subscribers, you can analyze the view counts provided in the public preview. Compute the view-to-subscriber ratio and assess the variance of per-post views. Genuine audiences will show high variance, with some posts being more important than others. Purchased or artificially inflated audiences, on the other hand, will display suspiciously uniform view counts across posts.

Lastly, link hygiene examines the outbound links shared by the channel. Scan the last 20 posts for links to suspicious domains, such as crypto giveaway sites, URL shorteners, or gambling affiliates. These types of links should score 0. Instead, look for links to primary sources like official statements, maps, or registries and award a score of 2 for those. This criterion ensures that the channel is sharing valuable and credible information, rather than spreading misinformation or promoting dubious content.

By applying this scoring system to each channel, you can create a curated collection set that is both high-quality and reliable. The scorecard can be implemented using publicly available information, without requiring any API keys or paid data sources. The full 11-point version, including additional rules for admin provenance, cross-post graph position, content-farm fingerprints, and more, is available in the Telegram and Web OSINT Bundle for $5.

A free sample brief is also provided for reference. This approach ensures that your collection set delivers high-quality and trustworthy information, ultimately enhancing your alert quality.

Written by urgent.news from Dev.to's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.

Read the original at dev.to →

More in Tech

"Your file never leaves your device": how to test that claim (and the Chromium catch)

Lots of web tools now say "runs in your browser, nothing is uploaded". I make one of them ( Drible , 53 file tools, 48 of which run client-side), and at some point I realised the claim was just a…

  • Many web tools claim client-side operation, but claim can be tested.
  • Test involves capturing network requests for file signatures and non-GET requests.
  • Chromium fails test due to not exposing file upload bodies to Playwright.

The Evil of Optional Parameters: When Having too Many Options is Self Destructing

You are the main developer of agents.com, an app that lets an organization manage an army of AI agents. You’re designing an admin page where an admin within an org can see each agent and the skills…

  • Optional parameters create contradictory and impossible states, confusing developers.
  • Defensive code required for consumers to interpret responses, increasing complexity.

Extracting Table Data from PDFs Using Python

PDF is one of the most common document formats in daily work, but its "tables" are a different thing from tables in Excel: a PDF file itself does not store structured table data.

  • Use Spire.PDF library in Python to extract PDF tables
  • Load document with PdfDocument object, create PdfTableExtractor
  • Extract tables page by page using ExtractTable() method

More from Saturday 10 October →