TL;DR

  • Converting to Lowercase -> lower()
  • Caseless String matching/comparison -> casefold()

casefold() is a text normalization function like lower() that is specifically designed to remove upper- or lower-case distinctions for the purposes of comparison. It is another form of normalizing text that may initially appear to be very similar to lower() because generally, the results are the same. As of Unicode 13.0.0, only ~300 of ~150,000 characters produced differing results when passed through lower() and casefold(). @dlukes' answer has the code to identify the characters that generate those differing results.

To answer your other two questions:

  • use lower() when you specifically want to ensure a character is lowercase, like for presenting to users or persisting data
  • use casefold() when you want to compare that result to another casefold-ed value.

Other Material

I suggest you take a closer look into what case folding actually is, so here's a good start: W3 Case Folding Wiki

Another source: Elastic.co Case Folding

Edit: I just recently found another very good related answer to a slightly different question here on SO (doing a case-insensitive string comparison)


Performance

Using this snippet, you can get a sense for the performance between the two:

import sys
from timeit import timeit

unicode_codepoints = tuple(map(chr, range(sys.maxunicode)))

def compute_lower():
    return tuple(codepoint.lower() for codepoint in unicode_codepoints)

def compute_casefold():
    return tuple(codepoint.casefold() for codepoint in unicode_codepoints)

timer_repeat = 1000

print(f"time to compute lower on unicode namespace: {timeit(compute_lower, number = timer_repeat) / timer_repeat} seconds")
print(f"time to compute casefold on unicode namespace: {timeit(compute_casefold, number = timer_repeat) / timer_repeat} seconds")

print(f"number of distinct characters from lower: {len(set(compute_lower()))}")
print(f"number of distinct characters from casefold: {len(set(compute_casefold()))}")

Running this, you'll get the results that the two are overwhelmingly the same in both performance and the number of distinct characters returned

time to compute lower on unicode namespace: 0.137255663 seconds
time to compute casefold on unicode namespace: 0.136321374 seconds
number of distinct characters from lower: 1112719
number of distinct characters from casefold: 1112694

If you run the numbers, that means it takes about 1.6e-07 seconds to run the computation on a single character for either function, so there isn't a performance benefit either way.

Answer from David Culbreth on Stack Overflow
🌐
W3Schools
w3schools.com › python › ref_string_casefold.asp
Python String casefold() Method
Remove List Duplicates Reverse ... Python Interview Q&A Python Training ... The casefold() method returns a string where all the characters are lower case....
Top answer
1 of 5
111

TL;DR

  • Converting to Lowercase -> lower()
  • Caseless String matching/comparison -> casefold()

casefold() is a text normalization function like lower() that is specifically designed to remove upper- or lower-case distinctions for the purposes of comparison. It is another form of normalizing text that may initially appear to be very similar to lower() because generally, the results are the same. As of Unicode 13.0.0, only ~300 of ~150,000 characters produced differing results when passed through lower() and casefold(). @dlukes' answer has the code to identify the characters that generate those differing results.

To answer your other two questions:

  • use lower() when you specifically want to ensure a character is lowercase, like for presenting to users or persisting data
  • use casefold() when you want to compare that result to another casefold-ed value.

Other Material

I suggest you take a closer look into what case folding actually is, so here's a good start: W3 Case Folding Wiki

Another source: Elastic.co Case Folding

Edit: I just recently found another very good related answer to a slightly different question here on SO (doing a case-insensitive string comparison)


Performance

Using this snippet, you can get a sense for the performance between the two:

import sys
from timeit import timeit

unicode_codepoints = tuple(map(chr, range(sys.maxunicode)))

def compute_lower():
    return tuple(codepoint.lower() for codepoint in unicode_codepoints)

def compute_casefold():
    return tuple(codepoint.casefold() for codepoint in unicode_codepoints)

timer_repeat = 1000

print(f"time to compute lower on unicode namespace: {timeit(compute_lower, number = timer_repeat) / timer_repeat} seconds")
print(f"time to compute casefold on unicode namespace: {timeit(compute_casefold, number = timer_repeat) / timer_repeat} seconds")

print(f"number of distinct characters from lower: {len(set(compute_lower()))}")
print(f"number of distinct characters from casefold: {len(set(compute_casefold()))}")

Running this, you'll get the results that the two are overwhelmingly the same in both performance and the number of distinct characters returned

time to compute lower on unicode namespace: 0.137255663 seconds
time to compute casefold on unicode namespace: 0.136321374 seconds
number of distinct characters from lower: 1112719
number of distinct characters from casefold: 1112694

If you run the numbers, that means it takes about 1.6e-07 seconds to run the computation on a single character for either function, so there isn't a performance benefit either way.

2 of 5
42

Both .lower() and .casefold() act on the full range of Unicode codepoints

There's some confusion in the existing answers, even the accepted one (EDIT: I was referring to this currently outdated version; the current one is fine). The distinction between .lower() and .casefold() has nothing to do with ASCII vs. Unicode, both act on the whole Unicode range of codepoints, just in slightly different ways. But both perform relatively complex mappings which they need to look up in the Unicode database, for instance:

>>> "Ť".lower()
'ť'

Both can involve single-to-multiple codepoint mappings, like we saw with "ß".casefold(). Look what happens to ß when you apply .lower()'s counterpart .upper():

>>> "ß".upper()
'SS'

And the one example I found where .lower() also does this:

>>> list("İ".lower())
['i', '̇']

So the performance claims, like "lower() will require less memory or less time because there are no lookups, and it's only dealing with 26 characters it has to transform", are simply not true.

The vast majority of the time, both operations yield the same thing, but there are a few cases (297 as of Unicode 13.0.0) where they don't. You can identify them like this:

import sys
import unicodedata as ud

print("Unicode version:", ud.unidata_version, "\n")
total = 0
for codepoint in map(chr, range(sys.maxunicode)):
    lower, casefold = codepoint.lower(), codepoint.casefold()
    if lower != casefold:
        total += 1
        for conversion, converted in zip(
            ("orig", "lower", "casefold"),
            (codepoint, lower, casefold)
        ):
            print(conversion, [ud.name(cp) for cp in converted], converted)
        print()
print("Total differences:", total)

When to use which

The Unicode standard covers lowercasing as part of Default Case Conversion in Section 3.13, and Default Case Folding is described right below that. The first paragraph says:

Case folding is related to case conversion. However, the main purpose of case folding is to contribute to caseless matching of strings, whereas the main purpose of case conversion is to put strings into a particular cased form.

My rule of thumb based on this:

  • Want to display a lowercased version of a string to users? Use .lower().
  • Want to do case-insensitive string comparison? Use .casefold().

(As a sidenote, I routinely break this rule of thumb and use .lower() across the board, just because it's shorter to type, the output is overwhelmingly the same, and what differences there are don't affect the languages I typically come across and work with. Don't be like me though ;) )

Just to hammer home that in terms of complexity, both operations are basically the same, they just use slightly different mappings -- this is Unicode's abstract definition of lowercasing:

R2 toLowercase(X): Map each character C in X to Lowercase_Mapping(C).

And this is its abstract definition of case folding:

R4 toCasefold(X): Map each character C in X to Case_Folding(C).

In Python's official documentation

The Python docs are quite clear that this is what the respective methods do, they even point the user to the aforementioned Section 3.13.

They describe .lower() as converting cased characters to lowercase, where cased characters are "those with general category property being one of “Lu” (Letter, uppercase), “Ll” (Letter, lowercase), or “Lt” (Letter, titlecase)". Same with .upper() and uppercase.

With .casefold(), the docs explicitly state that it's meant for "caseless matching", and that it's "similar to lowercasing but more aggressive because it is intended to remove all case distinctions in a string".

🌐
Programiz
programiz.com › python-programming › methods › string › casefold
Python String casefold() (with Examples)
# convert all characters to lowercase lowercased_string = text.casefold() print(lowercased_string) # Output: python
🌐
Reddit
reddit.com › r/learnpython › casefold() vs lower()
r/learnpython on Reddit: casefold() vs lower()
October 12, 2024 -

I just learnt about casefold() and decided to google it to see what it did. Here is w3schools definition:

Definition and Usage The casefold() method returns a string where all the characters are lower case.

This method is similar to the lower() method, but the casefold() method is stronger, more aggressive, meaning that it will convert more characters into lower case, and will find more matches when comparing two strings and both are converted using the casefold() method.

How does one “more aggressively” convert strings to lower case? Meaning, what more can/does it do than lower()?

🌐
GeeksforGeeks
geeksforgeeks.org › python › python-string-casefold-method
Python String casefold() Method - GeeksforGeeks
April 11, 2023 - Python String casefold() method is used to convert string to lowercase.
🌐
Codecademy
codecademy.com › docs › python › strings › .casefold()
Python | Strings | .casefold() | Codecademy
December 21, 2021 - The .casefold() method returns a copy of a string with all characters in lowercase. It is similar to .lower(), but whereas that method deals purely with ASCII text, .casefold() can also convert Unicode characters.
🌐
Tutorial Gateway
tutorialgateway.org › python-casefold
Python casefold
May 14, 2019 - The Python casefold function converts all the characters in a given string into lowercase letters. Although casefold() is the same as the lower function, it is more aggressive and stronger than the lower.
Find elsewhere
🌐
Vultr Docs
docs.vultr.com › python › standard library › str › casefold()
Python str casefold() - Case Insensitive Comparison
December 30, 2024 - The str.casefold() method in Python is essential for performing case-insensitive text comparisons or lookups. This method is particularly useful when comparing text strings where case variation is irrelevant, such as usernames or hashtags, making ...
🌐
Toppr
toppr.com › guides › python-guide › references › methods-and-functions › methods › string › casefold › python-string-casefold
Python casefold() function | Why do we use Python string casefold()? |
August 26, 2021 - Python casefold() function follows the following syntax: ... The casefold() function does not accept any parameter. It returns the case folded string, which is the string that has been changed to lower case.
🌐
Educative
educative.io › answers › what-is-casefold-in-python
What is casefold() in Python?
The casefold() method in Python converts all the uppercase letters in a string to lowercase letters.
🌐
AskPython
askpython.com › python › string › python-string-casefold
Python String casefold() - AskPython
August 6, 2022 - The Python String casefold() method returns a case folded string, and we use this to remove all case distinctions in a string.
🌐
Medium
medium.com › @yeaske › wait-there-is-casefold-in-python-9beb538c5a33
Mastering Python’s ‘casefold()’: A beginner’s guide | by Arun Suresh Kumar | Medium
May 14, 2024 - The casefold() method is a very powerful string method in Python designed to convert all characters in a string to lowercase, while also removing any case distinctions.
🌐
Tutorialspoint
tutorialspoint.com › python › casefold_method.htm
Python String casefold() Method
The Python string casefold() method converts all characters in the string to lowercase. It also performs additional transformations to handle special cases, such as certain Unicode characters that have different representations in uppercase and
🌐
Medium
vishvajitrao.medium.com › python-string-casefold-method-cbde417e8b8f
Python String casefold() Method. Python string casefold() method is a… | by Vishvajit Rao | Medium
October 26, 2020 - string casefold() method is a python string built-in method that is used to convert string to casefold string or lower case.
🌐
Python
docs.python.org › 3 › builtins › stdtypes.html
Built-in Types — Python 3.14.7 documentation
Casefolding is similar to lowercasing but more aggressive because it is intended to remove all case distinctions in a string. For example, the German lowercase letter 'ß' is equivalent to "ss".
🌐
GoLinuxCloud
golinuxcloud.com › home › programming › python › python casefold() string method
Python casefold(): String casefold() Method, Examples, and casefold vs lower (2026)
June 20, 2026 - Apply casefold() per element when normalizing a collection for lookup or sorting keys. A list comprehension keeps the intent clear; tuples and other iterables use the same pattern.