← Back to Portfolio

URL Based Spam Detection System

A Java project that reads a message, finds the link inside it and runs it through seven checks (whitelist, HTTPS, risky words, URL pattern, repeated links, domain ending and domain age). Every check adds risk points; the total decides the verdict.

How this demo works

Nothing is faked in the browser. Your text is sent to a backend, which runs the real Java program and returns its answer.

  1. Browseryou type, JavaScript sends it
  2. Backend APIvalidates the input
  3. Java URL filterruns the 7 checks
  4. Resultsent back and shown here

The 7 checks

  1. −80
    Whitelist Domain is in a list of ~1,600 trusted domains.
  2. +30
    HTTP / HTTPS Plain http:// (or no scheme) is risky.
  3. +20
    Risky keywords Words like urgent, verify, password.
  4. +30
    URL pattern Very long URL, ?, % or many dashes.
  5. +20
    Multiple URLs More than one link, or the same link repeated.
  6. +30
    Domain ending Unknown TLD (not .com, .in, .org …).
  7. +40
    Domain age WHOIS lookup: unknown or under 30 days is risky; very old domains lower the score.

Score 60 or more = potential spam · 30–59 = suspicious · below 30 = safe. Domains that are not whitelisted also get +20.

JavaRegexWHOIS (Python)

Try it

Checking backend…
0 / 1000
Try:
← Back to Portfolio