Skip to content

Glossary URL

What Is a URL?

Definition

A URL (Uniform Resource Locator) is the address that identifies a resource on the web, such as a page, an image or a file, and says how to reach it: with which protocol, on which server and under which path.

On this page 5
  1. What URL means
  2. How it works
  3. Why it matters
  4. Best practices
  5. Common mistakes
In brief

A URL is the unique address of a resource on the web. Its parts, how Google and browsers read it, and the mistakes that create duplicate pages.

What URL means

URL stands for Uniform Resource Locator. It is the address typed into the browser bar, the address behind a link and the address Google stores in its index. Technically, a URL is a kind of URI (Uniform Resource Identifier), the format defined by the RFC 3986 standard: an identifier for a resource. A URL is the kind that also says where that resource can be found.

Three neighbouring terms often get mixed up. The domain is only one part of the URL, the name of the server (for example, zds.es). The slug is the last part of the path, the one that describes the page in words (for this entry, what-is-a-url). The URL is the whole thing: protocol, domain, path and, if present, parameters and fragment.

To a search engine, every different URL is at first a different document. Google says so explicitly in its guide to URL structure: it treats /APPLE and /apple as two separate addresses, each with its own content. Much of the trouble with duplicate content starts here, with the same page reachable under several URLs.

How it works

RFC 3986 describes the generic syntax in five components: scheme, authority, path, query and fragment. In https://zds.es/que-es-sitemap?utm_source=newsletter#faq, the scheme is https, the authority is zds.es, the path is /que-es-sitemap, the query is utm_source=newsletter and the fragment is faq. The scheme and the path are required; the rest is optional.

The parts do not all behave the same way. The scheme and the server name are not case-sensitive, although the standard recommends writing them in lower case. The path, on the other hand, is case-sensitive unless the server decides otherwise. Characters the standard does not allow as they are, such as a space or a "ü", are encoded with the percent sign: the "ü" travels as %C3%BC.

The fragment, the part after "#", never reaches the server. The browser uses it to jump to a section of the page. That is why Google says it generally does not support fragments and advises against using them to change a page's content; for that it recommends the JavaScript History API. Browsers, for their part, read URLs according to the WHATWG URL Standard, a living standard last updated on 7 October 2026.

Capital letters only count from the path onward.

Why it matters

A well-built URL helps two readers at once. A person who sees it in a search result or a link understands what the page is about before opening it. The search engine uses it as a stable identifier to crawl, index and collect signals. If the same page answers at several URLs, those signals are split, and the search engine has to choose which version is the main one.

The same holds for AI search. When an assistant cites a source, it links a specific URL. If the page exists in several versions, for instance with and without a trailing slash, or with campaign parameters, the citation can be spread across them or point to the version you did not want.

In this glossary we handle it like this, checked on 8 October 2026 with the sitemap entry: with a trailing slash, zds.es/que-es-sitemap/ redirects with a 301 to the version without one. With a campaign parameter, ?utm_source=…, the same page is served, but with a canonical tag pointing to the clean URL. And with the Spanish slug under the German folder, /de/que-es-sitemap, it redirects with a 301 to /de/was-ist-sitemap. Written in capitals, /Que-Es-Sitemap, it also redirects with a 301 to the lower-case form. Four variants, one URL that counts.

Best practices

  • Use readable words instead of long ID numbers, in the language of the people who will read the page.
  • Separate words with hyphens, not underscores, as Google recommends.
  • Keep paths in lower case throughout, and redirect upper-case variants with a 301 if the server accepts them.
  • Pick one form, with or without a trailing slash, and redirect the other.
  • Drop parameters that do not change the content, and for the ones that stay, use the usual encoding: "=" between key and value and "&" between parameters.
  • Mark the main URL with a canonical tag when a page also has to exist with parameters.

Common mistakes

  • Changing URLs without redirecting the old ones, so links and citations end up on a 404.
  • Using fragments ("#") to load different content that should be indexable.
  • Mixing upper and lower case in internal links, so the same page exists twice.
  • Creating new URLs for every filter combination in a shop or listing without controlling which ones get indexed.
  • Translating a page's text but leaving the slug in another language, or the reverse, without keeping track of which versions belong together.
Manuel Riveiro Rodriguez CEO & Digital Strategist

A technical audit covers this and everything else in one pass.

Request an audit

Frequently asked

What is the difference between a URL and a URI?

URI is the general term in RFC 3986 for any identifier of a resource. A URL is a URI that also says how to locate the resource, for example via https and a server. In everyday use, and in search engines and browsers, web addresses are called URLs, and the distinction has little practical effect.

Are URLs case-sensitive?

The path is. The scheme and the server name are not. Google treats /APPLE and /apple as two different URLs, each with its own content. If your server delivers the same page under both spellings, the right move is to pick one, usually lower case, and redirect the other with a 301.

Does the URL affect rankings?

Google asks for simple, readable URLs in the audience's language, because they help it understand how a site is structured. In practice, what matters more is that each page has exactly one stable URL. Duplicate variants split signals and force the search engine to choose which one to index.

What happens to the part of a URL after "#"?

That is the fragment. The browser uses it to jump to a section of the page, and it is not sent to the server. Google says it generally does not support fragments, so content that only appears when the fragment changes may not get indexed. To change content, the History API is the better route.

Can a URL contain accented letters?

Yes. Google accepts words in the audience's language, including non-ASCII characters, as long as they are encoded in UTF-8; in links, those characters travel percent-encoded. Removing accents from the slug is another option, and it stops the same address from appearing sometimes encoded and sometimes not.

Sources

  1. Google Search Central, "URL structure best practices for Google Search", last updated 10 December 2025. Source for the recommendations on readable words, the audience's language, hyphens, parameters and encoding, for case sensitivity (/APPLE versus /apple) and for the lack of support for fragments.
  2. IETF, RFC 3986 "Uniform Resource Identifier (URI): Generic Syntax" (STD 66). Source for the five components of the generic syntax, the lower-case recommendation for the server name and the definition of the fragment.
  3. WHATWG, "URL Standard", living standard, last updated 7 October 2026. Source for the standard browsers use to read URLs.