# An Informal Introduction to Python  
Topics Covered in this tutorial are:  
    * Numbers  
    * Strings  
    * Lists  

Comments in Python start with the hash character, #, and extend to the end of the physical line. A comment may appear at the start of a line or following whitespace or code, but not within a string literal. A hash character within a string literal is just a hash character. Since comments are to clarify code and are not interpreted by Python, they may be omitted when typing in examples.

In [2]:
# this is the first comment
spam = 1  # and this is the second comment
          # ... and now a third!
text = "# This is not a comment because it's inside quotes."

In [3]:
print(spam)

1


In [4]:
print(text)

# This is not a comment because it's inside quotes.


In [5]:
"""
This is the first line of multi comment,
this is second line and
this is third line.
"""

# this is the first comment
spam = 1  # and this is the second comment
          # ... and now a third!
text = "# This is not a comment because it's inside quotes."

## Using Python as Calculator

## Numbers  
The interpreter acts as a simple calculator: you can type an expression at it and it will write the value. Expression syntax is straightforward: the operators +, -, * and / work just like in most other languages (for example, Pascal or C); parentheses (()) can be used for grouping. For example:

In [6]:
2 + 2

4

In [7]:
50 - 5*6

20

In [8]:
(50 - 5*6) / 4

5.0

In [9]:
8 / 5  # division always returns a floating point number

1.6

In [10]:
type(1.6)

float

In [11]:
type(5)

int

The integer numbers (e.g. 2, 4, 20) have type int, the ones with a fractional part (e.g. 5.0, 1.6) have type float. We will see more about numeric types later in the tutorial.

Division (/) always returns a float. To do floor division and get an integer result (discarding any fractional result) you can use the // operator; to calculate the remainder you can use %:

>17 / 3  # classic division returns a float  
>17 // 3  # floor division discards the fractional part  
>17 % 3  # the % operator returns the remainder of the division  
>5 * 3 + 2  # result * divisor + remainder  

In [12]:
17/3

5.666666666666667

In [13]:
17//3

5

In [14]:
17%3

2

In [15]:
5*3+2

17

With Python, it is possible to use the ** operator to calculate powers 

In [16]:
2 ** 5 # 2 to the power of 5

32

In [17]:
5 ** 2 # 5 to the power of 2

25

The equal sign (=) is used to assign a value to a variable. Afterwards, no result is displayed before the next interactive prompt:

In [18]:
width = 20
height = 5 * 9
width * height

900

In [19]:
20 = width

SyntaxError: cannot assign to literal (<ipython-input-19-5390a46bdca6>, line 1)

If a variable is not “defined” (assigned a value), trying to use it will give you an error:

In [20]:
n

NameError: name 'n' is not defined

There is full support for floating point; operators with mixed type operands convert the integer operand to floating point:

In [21]:
4 * 3.75 - 1

14.0

In interactive mode, the last printed expression is assigned to the variable _. This means that when you are using Python as a desk calculator, it is somewhat easier to continue calculations, for example:

In [22]:
tax = 12.5 / 100
price = 100.50
price * tax

12.5625

In [23]:
price + _

113.0625

round(value, round_count)

In [24]:
round(_,3)

113.062

In [25]:
round(_,2)

113.06

This variable should be treated as read-only by the user. Don’t explicitly assign a value to it — you would create an independent local variable with the same name masking the built-in variable with its magic behavior.

## Strings

Besides numbers, Python can also manipulate strings, which can be expressed in several ways. They can be enclosed in single quotes ('...') or double quotes ("...") with the same result. \ can be used to escape quotes:

In [26]:
'spam eggs' # single quotes

'spam eggs'

In [27]:
'doesn't'

SyntaxError: invalid syntax (<ipython-input-27-f70cf6904edb>, line 1)

In [28]:
'doesn\'t' # use \' to escape the single quote...

"doesn't"

In [29]:
"doesn't" # ...or use double quotes instead

"doesn't"

In [30]:
'"Yes," they said.'

'"Yes," they said.'

In [31]:
"\"Yes,\" they said."

'"Yes," they said.'

In [32]:
'"Isn\'t," they said.'

'"Isn\'t," they said.'

In the interactive interpreter, the output string is enclosed in quotes and special characters are escaped with backslashes. While this might sometimes look different from the input (the enclosing quotes could change), the two strings are equivalent. The string is enclosed in double quotes if the string contains a single quote and no double quotes, otherwise it is enclosed in single quotes. The print() function produces a more readable output, by omitting the enclosing quotes and by printing escaped and special characters:

In [33]:
'"Isn\'t," they said.'

'"Isn\'t," they said.'

In [34]:
print('"Isn\'t," they said.')

"Isn't," they said.


In [35]:
s = 'First line.\nSecond line.'  # \n means newline
s  # without print(), \n is included in the output

'First line.\nSecond line.'

In [36]:
print(s)  # with print(), \n produces a new line

First line.
Second line.


If you don’t want characters prefaced by \ to be interpreted as special characters, you can use raw strings by adding an r before the first quote:

In [37]:
print('C:\some\name')  # here \n means newline!

C:\some
ame


In [38]:
print(r'C:\some\name')  # note the r before the quote

C:\some\name


String literals can span multiple lines. One way is using triple-quotes: """...""" or '''...'''. End of lines are automatically included in the string, but it’s possible to prevent this by adding a \ at the end of the line. The following example:

In [39]:
print("""\
Usage: thingy [OPTIONS]
     -h                        Display this usage message
     -H hostname               Hostname to connect to
""")

Usage: thingy [OPTIONS]
     -h                        Display this usage message
     -H hostname               Hostname to connect to



In [40]:
print("""
Usage: thingy [OPTIONS]
     -h                        Display this usage message
     -H hostname               Hostname to connect to
""")


Usage: thingy [OPTIONS]
     -h                        Display this usage message
     -H hostname               Hostname to connect to



Strings can be concatenated (glued together) with the + operator, and repeated with *:

In [41]:
# 3 times 'un', followed by 'ium'
3 * 'un' + 'ium'

'unununium'

Two or more string literals (i.e. the ones enclosed between quotes) next to each other are automatically concatenated.

In [42]:
'Py' 'thon'

'Python'

This feature is particularly useful when you want to break long strings:

In [43]:
text = ('Put several strings within parentheses ' 
        'to have them joined together.')

In [44]:
text

'Put several strings within parentheses to have them joined together.'

This only works with two literals though, not with variables or expressions:

In [45]:
prefix = 'Py'

In [46]:
prefix 'thon'  # can't concatenate a variable and a string literal

SyntaxError: invalid syntax (<ipython-input-46-c5901e312aa3>, line 1)

In [47]:
('un' * 3) 'ium'

SyntaxError: invalid syntax (<ipython-input-47-f4764cbe42a8>, line 1)

If you want to concatenate variables or a variable and a literal, use +:

In [48]:
prefix + 'thon'

'Python'

Strings can be indexed (subscripted), with the first character having index 0. There is no separate character type; a character is simply a string of size one:

|P|y|t|h|o|n|
|-|-|-|-|-|-|
|0|1|2|3|4|5|

In [49]:
word = 'Python'

In [50]:
word[0] # character in position 0

'P'

In [51]:
word[5] # characte in position 5

'n'

Indices may also be negative numbers, to start counting from the right:

|P|y|t|h|o|n|
|-|-|-|-|-|-|
|0|1|2|3|4|5|
|-6|-5|-4|-3|-2|-1|


In [52]:
word[-1] # last character

'n'

In [53]:
word[-2] # second-last character

'o'

In [54]:
word[-6]

'P'

Note that since -0 is the same as 0, negative indices start from -1.

In addition to indexing, slicing is also supported. While indexing is used to obtain individual characters, slicing allows you to obtain substring:

In [55]:
word[0:2]  # characters from position 0 (included) to 2 (excluded)

'Py'

In [56]:
word[2:5]  # characters from position 2 (included) to 5 (excluded)

'tho'

Note how the start is always included, and the end always excluded. This makes sure that s[:i] + s[i:] is always equal to s:

In [57]:
word[:2] + word[2:]

'Python'

In [58]:
word[:4] + word[4:]

'Python'

Slice indices have useful defaults; an omitted first index defaults to zero, an omitted second index defaults to the size of the string being sliced.

In [59]:
word[:2]   # character from the beginning to position 2 (excluded)

'Py'

In [60]:
word[4:]   # characters from position 4 (included) to the end

'on'

In [61]:
word[-2:]  # characters from the second-last (included) to the end

'on'

In [62]:
word[:-2]

'Pyth'

One way to remember how slices work is to think of the indices as pointing between characters, with the left edge of the first character numbered 0. Then the right edge of the last character of a string of n characters has index n, for example:

|P|y|t|h|o|n|
|-|-|-|-|-|-|
|0|1|2|3|4|5|
|-6|-5|-4|-3|-2|-1|

The first row of numbers gives the position of the indices 0…6 in the string; the second row gives the corresponding negative indices. The slice from i to j consists of all characters between the edges labeled i and j, respectively.

For non-negative indices, the length of a slice is the difference of the indices, if both are within bounds. For example, the length of word[1:3] is 2.

Attempting to use an index that is too large will result in an error:

In [63]:
word[42]  # the word only has 6 characters

IndexError: string index out of range

However, out of range slice indexes are handled gracefully when used for slicing:

In [64]:
word[4:42]

'on'

In [65]:
word[42:]

''

Python strings cannot be changed — they are immutable. Therefore, assigning to an indexed position in the string results in an error:

In [66]:
word[0] = 'J'

TypeError: 'str' object does not support item assignment

In [67]:
word[2:] = 'py'

TypeError: 'str' object does not support item assignment

If you need a different string, you should create a new one:

In [68]:
'J' + word[1:]

'Jython'

In [69]:
word[:2] + 'py'

'Pypy'

The built-in function len() returns the length of a string:

In [70]:
s = 'supercalifragilisticexpialidocious'

In [71]:
len(s)

34

**String Methods:**


Strings implement all of the common sequence operations, along with the additional methods described below.

In [72]:
string = "Hello, Everyöne!"

**str.capitalize()**  
Return a copy of the string with its first character capitalized and the rest lowercased.

In [73]:
str.capitalize("hello, everyone")

'Hello, everyone'

**str.casefold()**  
Return a casefolded copy of the string. Casefolded strings may be used for caseless matching.

In [74]:
str.casefold("Hello, Everyone")

'hello, everyone'

**str.center(width[, fillchar])**  
Return centered in a string of length width. Padding is done using the specified fillchar (default is an ASCII space). The original string is returned if width is less than or equal to len(s).

In [75]:
len(string)

16

In [76]:
string.center(30)

'       Hello, Everyöne!       '

In [77]:
string.center(10)

'Hello, Everyöne!'

In [78]:
string.center(30, '-')

'-------Hello, Everyöne!-------'

**str.count(sub[, start[, end]])**  
Return the number of non-overlapping occurrences of substring sub in the range [start, end]. Optional arguments start and end are interpreted as in slice notation.

|H|e|l|l|o|,| |E|v|e|r|y|o|n|e|!|
|-|-|-|-|-|-|-|-|-|-|-|-|-|-|-|-|
|0|1|2|3|4|5|6|7|8|9|10|11|12|13|14|15|

In [79]:
string

'Hello, Everyöne!'

In [80]:
string.count('e')

3

In [81]:
string.count('e',6)

2

In [82]:
string.count('e',6,10)

1

In [83]:
'banana'.count('na')

2

**str.encode(encoding="utf-8", errors="strict")**  
Return an encoded version of the string as a bytes object. Default encoding is 'utf-8'. errors may be given to set a different error handling scheme. The default for errors is 'strict', meaning that encoding errors raise a UnicodeError.

However, it takes two parameters:

* **encoding** - the encoding type a string has to be encoded to
* **errors** - response when encoding fails. There are six types of error response
    - strict - default response which raises a UnicodeDecodeError exception on failure
    - ignore - ignores the unencodable unicode from the result
    - replace - replaces the unencodable unicode to a question mark ?
    - xmlcharrefreplace - inserts XML character reference instead of unencodable unicode
    - backslashreplace - inserts a \uNNNN escape sequence instead of unencodable unicode
    - namereplace - inserts a \N{...} escape sequence instead of unencodable unicode

In [84]:
string_utf = string.encode()

In [85]:
string

'Hello, Everyöne!'

In [86]:
string_utf

b'Hello, Every\xc3\xb6ne!'

In [87]:
print('The encoded version (with ignore) is:', string.encode("ascii", "ignore"))

The encoded version (with ignore) is: b'Hello, Everyne!'


In [88]:
print('The encoded version (with replace) is:', string.encode("ascii", "replace"))

The encoded version (with replace) is: b'Hello, Every?ne!'


**str.endswith(suffix[, start[, end]])**  
Return True if the string ends with the specified suffix, otherwise return False. suffix can also be a tuple of suffixes to look for. With optional start, test beginning at that position. With optional end, stop comparing at that position.

In [89]:
string

'Hello, Everyöne!'

In [90]:
string.endswith('s')

False

In [91]:
string.endswith('!')

True

In [92]:
string.endswith('!',4,10)

False

**str.find(sub[, start[, end]])**  
Return the lowest index in the string where substring sub is found within the slice s[start:end]. Optional arguments start and end are interpreted as in slice notation. Return -1 if sub is not found.

In [93]:
string.find('llo')

2

In [94]:
string.find('e')

1

**Note:** The find() method should be used only if you need to know the position of sub. To check if sub is a substring or not, use the in operator:

In [95]:
'Py' in 'Python'

True

**str.format(*args, **kwargs)**  
Perform a string formatting operation. The string on which this method is called can contain literal text or replacement fields delimited by braces {}. Each replacement field contains either the numeric index of a positional argument, or the name of a keyword argument. Returns a copy of the string where each replacement field is replaced with the string value of the corresponding argument.

In [96]:
"The sum of 1 + 2 is {0}".format(1+2)

'The sum of 1 + 2 is 3'

In [97]:
"The sum of 1 + 2 is {0} and product of 1 * 2 is {1}".format(1+2, 1*2)

'The sum of 1 + 2 is 3 and product of 1 * 2 is 2'

**str.isalnum()**  
Return True if all characters in the string are alphanumeric and there is at least one character, False otherwise. A character c is alphanumeric if one of the following returns True: c.isalpha(), c.isdecimal(), c.isdigit(), or c.isnumeric().

In [98]:
name = "M234onica"
print(name.isalnum())

True


In [99]:
# contains whitespace
name = "M3onica Gell22er "
print(name.isalnum())

False


In [100]:
name = "Mo3nicaGell22er"
print(name.isalnum())

True


**str.isalpha()**  
Return True if all characters in the string are alphabetic and there is at least one character, False otherwise. 

In [101]:
name = "Monica"
print(name.isalpha())

True


In [102]:
# contains whitespace
name = "Monica Geller"
print(name.isalpha())

False


In [103]:
# contains number
name = "Mo3nicaGell22er"
print(name.isalpha())

False


**str.isdecimal()**  
Return True if all characters in the string are decimal characters and there is at least one character, False otherwise.

In [104]:
s = "28212"
print(s.isdecimal())

True


In [105]:
# contains alphabets
s = "32ladk3"
print(s.isdecimal())

False


In [106]:
# contains alphabets and spaces
s = "Mo3 nicaG el l22er"
print(s.isdecimal())

False


**str.isascii()**  
Return True if the string is empty or all characters in the string are ASCII, False otherwise. ASCII characters have code points in the range U+0000-U+007F.

In [107]:
s = 'ABCabc#$%.'
s.isascii()

True

In [108]:
''.isascii()

True

In [109]:
' '.isascii()

True

In [110]:
'∑'.isascii()

False

**str.isdigit()**  
Return True if all characters in the string are digits and there is at least one character, False otherwise.

In [111]:
s = '123456'
s.isdigit()

True

In [112]:
s = '123abc'
s.isdigit()

False

**str.islower()**  
Return True if all cased characters in the string are lowercase and there is at least one cased character, False otherwise.

In [113]:
s = 'asdlkjgadb'
s.islower()

True

In [114]:
s = 'asdlkJgadb'
s.islower()

False

**str.isupper()**  
Return True if all cased characters 4 in the string are uppercase and there is at least one cased character, False otherwise.

In [115]:
s = 'ABCDEFGH'
s.isupper()

True

In [116]:
s = 'ABCDEFgH'
s.isupper()

False

**str.isspace()**  
Return True if there are only whitespace characters in the string and there is at least one character, False otherwise.

In [117]:
s = 'a \n b'
s.isspace()

False

In [118]:
s = '\t\n '
s.isspace()

True

**str.isnumeric()**  
Return True if all characters in the string are numeric characters, and there is at least one character, False otherwise.


In [119]:
s = "1234"
s.isnumeric()

True

In [120]:
s = "123abc4"
s.isnumeric()

False

**str.istitle()**  
Return True if the string is a titlecased string and there is at least one character, for example uppercase characters may only follow uncased characters and lowercase characters only cased ones. Return False otherwise.

In [121]:
s = 'The Sun Also Rises'
s.istitle()

True

In [122]:
s = "Bob's Burgers!"
s.istitle()

False

**str.ljust(width[, fillchar])**  
Return the string left justified in a string of length width. Padding is done using the specified fillchar (default is an ASCII space). The original string is returned if width is less than or equal to len(s).

**str.rjust(width[, fillchar])**  
Return the string right justified in a string of length width. Padding is done using the specified fillchar (default is an ASCII space). The original string is returned if width is less than or equal to len(s).

In [123]:
s = "Python"
print(s.ljust(10,'*'))
print(s.rjust(10,'*'))

Python****
****Python


**str.lower()**  
Return a copy of the string with all the cased characters converted to lowercase.

**str.upper()**  
Return a copy of the string with all the cased characters converted to uppercase.

In [124]:
s = "PyThon"
print(s.lower())
print(s.upper())

python
PYTHON


**str.lstrip([chars])**  
Return a copy of the string with leading characters removed. The chars argument is a string specifying the set of characters to be removed. If omitted or None, the chars argument defaults to removing whitespace. The chars argument is not a prefix; rather, all combinations of its values are stripped:

**str.rstrip([chars])**  
Return a copy of the string with trailing characters removed. The chars argument is a string specifying the set of characters to be removed. If omitted or None, the chars argument defaults to removing whitespace. The chars argument is not a suffix; rather, all combinations of its values are stripped:

**str.strip([chars])**  
Return a copy of the string with the leading and trailing characters removed. The chars argument is a string specifying the set of characters to be removed. If omitted or None, the chars argument defaults to removing whitespace. The chars argument is not a prefix or suffix; rather, all combinations of its values are stripped:

In [125]:
s = '   spacious    '

In [126]:
s.lstrip()

'spacious    '

In [127]:
s.rstrip()

'   spacious'

In [128]:
s.strip()

'spacious'

**str.replace(old, new[, count])**  
Return a copy of the string with all occurrences of substring old replaced by new. If the optional argument count is given, only the first count occurrences are replaced.

In [129]:
s = "Hello, Everyone!"

In [130]:
s.replace('Hello', 'Hi')

'Hi, Everyone!'

In [131]:
s.replace('e', '--')

'H--llo, Ev--ryon--!'

**str.split(sep=None, maxsplit=-1)**  
Return a list of the words in the string, using sep as the delimiter string. If maxsplit is given, at most maxsplit splits are done (thus, the list will have at most maxsplit+1 elements). If maxsplit is not specified or -1, then there is no limit on the number of splits (all possible splits are made).

If sep is given, consecutive delimiters are not grouped together and are deemed to delimit empty strings (for example, '1,,2'.split(',') returns ['1', '', '2']). The sep argument may consist of multiple characters (for example, '1<>2<>3'.split('<>') returns ['1', '2', '3']). Splitting an empty string with a specified separator returns [''].

In [132]:
'1,2,3'.split(',')

['1', '2', '3']

In [133]:
'1,2,3'.split(',', maxsplit=1)

['1', '2,3']

In [134]:
'1,2,,3,'.split(',')

['1', '2', '', '3', '']

If sep is not specified or is None, a different splitting algorithm is applied: runs of consecutive whitespace are regarded as a single separator, and the result will contain no empty strings at the start or end if the string has leading or trailing whitespace. Consequently, splitting an empty string or a string consisting of just whitespace with a None separator returns [].

In [135]:
'1 2 3'.split()

['1', '2', '3']

In [136]:
'1 2 3'.split(maxsplit=1)

['1', '2 3']

In [137]:
'   1   2   3   '.split()

['1', '2', '3']

### Formatted string literals

A formatted string literal or f-string is a string literal that is prefixed with 'f' or 'F'. These strings may contain replacement fields, which are expressions delimited by curly braces {}. While other string literals always have a constant value, formatted strings are really expressions evaluated at run time.

Some examples of formatted string literals:

In [138]:
name = "Fred"

In [139]:
f"He said his name is {name!r}."

"He said his name is 'Fred'."

In [140]:
f"He said his name is {repr(name)}."  # repr() is equivalent to !r

"He said his name is 'Fred'."

**repr(object)**  
Return a string containing a printable representation of an object. For many types, this function makes an attempt to return a string that would yield an object with the same value when passed to eval(), otherwise the representation is a string enclosed in angle brackets that contains the name of the type of the object together with additional information often including the name and address of the object.

In [141]:
import decimal
width = 10
precision = 4
value = decimal.Decimal("12.34567")

In [142]:
f"result: {value:{width}.{precision}}"  # nested fields

'result:      12.35'

In [143]:
from datetime import datetime
today = datetime(year=2017, month=1, day=27)

In [144]:
f"{today:%B %d, %Y}"  # using date format specifier

'January 27, 2017'

In [145]:
f"{today=:%B %d, %Y}" # using date format specifier and debugging

'today=January 27, 2017'

In [146]:
number = 1024

In [147]:
foo = "bar"

In [148]:
f"{ foo = }" # preserves whitespace

" foo = 'bar'"

In [149]:
line = "The mill's closed"

In [150]:
f"{line = }"

'line = "The mill\'s closed"'

In [151]:
f"{line = :20}"

"line = The mill's closed   "

In [152]:
f"{line = !r:20}"

'line = "The mill\'s closed" '

## Lists

Python knows a number of compound data types, used to group together other values. The most versatile is the list, which can be written as a list of comma-separated values (items) between square brackets. Lists might contain items of different types, but usually the items all have the same type.

In [153]:
squares = [1, 4, 9, 16, 25]

In [154]:
squares

[1, 4, 9, 16, 25]

Like strings (and all other built-in sequence types), lists can be indexed and sliced:

In [155]:
squares[0]  # indexing returns the item

1

In [156]:
squares[-1]

25

In [157]:
squares[-3:]  # slicing returns a new list

[9, 16, 25]

All slice operations return a new list containing the requested elements. This means that the following slice returns a shallow copy of the list:

In [158]:
squares[:]

[1, 4, 9, 16, 25]

Lists also support operations like concatenation:

In [159]:
squares + [36, 49, 64, 81, 100]

[1, 4, 9, 16, 25, 36, 49, 64, 81, 100]

Unlike strings, which are immutable, lists are a mutable type, i.e. it is possible to change their content:

In [160]:
cubes = [1, 8, 27, 65, 125]  # something's wrong here

In [161]:
4 ** 3

64

In [162]:
cubes[3] = 64  # replace the wrong value

In [163]:
cubes

[1, 8, 27, 64, 125]

You can also add new items at the end of the list, by using the append() method (we will see more about methods later):

In [164]:
cubes.append(216)  # add the cube of 6

In [165]:
cubes.append(7 ** 3)  # and the cube of 7

In [166]:
cubes

[1, 8, 27, 64, 125, 216, 343]

Assignment to slices is also possible, and this can even change the size of the list or clear it entirely:

In [167]:
letters = ['a', 'b', 'c', 'd', 'e', 'f', 'g']

In [168]:
letters

['a', 'b', 'c', 'd', 'e', 'f', 'g']

In [169]:
letters[2:5] = ['C', 'D', 'E']

In [170]:
letters

['a', 'b', 'C', 'D', 'E', 'f', 'g']

In [171]:
letters[2:5] = [] # remove the letters

In [172]:
letters

['a', 'b', 'f', 'g']

In [173]:
letters[:] = [] # clear the list by replacing all the elements with an empty list

In [174]:
letters

[]

The built-in function len() also applies to lists:

In [175]:
letters = ['a', 'b', 'c', 'd']

In [176]:
len(letters)

4

It is possible to nest lists (create lists containing other lists), for example:

In [177]:
a = ['a', 'b', 'c']
n = [1, 2, 3]
x = [a, n]

In [178]:
x

[['a', 'b', 'c'], [1, 2, 3]]

In [179]:
x[0]

['a', 'b', 'c']

In [180]:
x[0][1]

'b'