Regex Tester & Debugger→Specialized Version
🔍

URL Slug Regex Tester

Test and explain the url slug tester pattern

//gm
Flags:
Examples:
how-to-write-regex png-to-webp How-To-Write trailing- double--dash
#MatchIndexGroups
1how-to-write-regex0—
2png-to-webp19—

The Pattern

``regex ^[a-z0-9]+(?:-[a-z0-9]+)*$ `

Flags: gm — multiline, so ^ and $ anchor to each line rather than the whole input.

What It Accepts and Rejects

InputResult
how-to-write-regex✅ matches
png-to-webp✅ matches
How-To-Write❌ rejected
trailing-❌ rejected
double--dash❌ rejected

One Group, Repeated

^[a-z0-9]+(?:-[a-z0-9]+)*$ reads as "a run of lowercase alphanumerics, then zero or more occurrences of a hyphen followed by another run". That structure makes the three common slug defects impossible by construction:

  • A leading or trailing hyphen cannot match, because every hyphen must be followed by a run.
  • A double hyphen cannot match, for the same reason.
  • Uppercase cannot match at all, so How-To-Write is rejected.
The alternative — ^[a-z0-9-]+$ — accepts -, -- and slug- without complaint.

The Limits of This Pattern

The pattern validates a slug but cannot create one. Generating a slug from a title needs transliteration (é to e, ü to ue or u depending on language), whitespace collapsing, and truncation at a word boundary — none of which a regex does. Validate with this; generate with a real slugify function.

Testing Before Shipping

A regex that has only been tried against inputs you expect to match is untested. Every pattern needs three kinds of case:

1. Valid inputs that should match, including the awkward-but-legal ones. 2. Invalid inputs that should not, especially near-misses that differ by one character. 3. Adversarial inputs — very long strings, unusual Unicode, and nesting that could trigger catastrophic backtracking.

Paste your own examples into the tester above and watch which lines highlight. A pattern that matches everything you throw at it is usually too permissive rather than correct.

Catastrophic Backtracking

Nested quantifiers over overlapping character classes — (a+)+, (\w+\s?)* — can take exponential time on a non-matching input. On a server that is a denial-of-service bug, not a performance issue. If a pattern is applied to user input, bound the input length first and prefer explicit alternation over nested repetition.

Anchors, Greediness and Backtracking

Three behaviours account for most regex surprises, and this pattern shows all three.

Anchors. ^ and $ pin the match to the start and end of the input. Without them, \d{3} matches the 123 inside abc123def. With m in the flags — as here — they pin to each *line* instead, which is what lets one pattern be tested against a list.

Greediness. .* takes as much as it can and gives back only when forced; .*? takes as little as possible. On , the pattern <.*> matches the whole string and <.*?> matches just .

Backtracking. When a match fails, the engine reverses and tries other splits. Nested quantifiers like (a+)+ make that exponential, and a 30-character input can hang a server — a class of denial of service known as ReDoS. Avoid nesting quantifiers, and prefer explicit character classes over . wherever you can.

Testing It Properly

`javascript const pattern = /^[a-z0-9]+(?:-[a-z0-9]+)*$/gm;

// A global regex keeps lastIndex between calls, so reusing one across // test() calls returns alternating true/false on the same input. pattern.lastIndex = 0;

// Named groups make the result readable const named = /(?\d{4})-(?\d{2})/; const { groups } = '2026-08'.match(named); `

Write the failing cases first. A pattern that accepts everything valid is easy; one that also rejects everything invalid is the hard half, and it is where the bugs are.

When Not to Use a Regex

Structured formats have parsers, and the parser is always more correct: new URL() for URLs, DOMParser for HTML, JSON.parse` for JSON, a date library for dates. Reach for a regex to *find* things in unstructured text, not to validate something a parser understands.

Frequently Asked Questions

Does this pattern handle every valid case?

The pattern validates a slug but cannot create one. Generating a slug from a title needs transliteration (é to e, ü to ue or u depending on language), whitespace collapsing, and truncation at a word boundary — none of which a regex does. Validate with this; generate with a real slugify function.

Why does the pattern use the flags it uses?

Flags `gm`. The `g` flag finds every match rather than stopping at the first, and `m` makes `^` and `$` match at each line boundary so a multi-line test string can be checked line by line. Changing them changes the results, so test with the flags you will ship.

Is validating this with a regex the right approach?

For a format check before doing real work, usually yes. For anything security-critical, a regex confirms shape and nothing else — parse the value with a real parser, or verify it against the system that owns it, before trusting it.

Related Tools

Explore other tools you might find useful:

More Regex Tester & Debugger tools

You might also need