## Tutorial 11: Data Structures

These notes are adapted from the Python tutorial available at: https://docs.python.org/3/tutorial/.

Additional data structures for storing collections of objects.

## Tuples and Sequences

We saw that lists and strings have many common properties, such as indexing and slicing operations. They are two examples of sequence data types (see Sequence Types — list, tuple, range). Since Python is an evolving language, other sequence data types may be added. There is also another standard sequence data type: the tuple.

A tuple consists of a number of values separated by commas, for instance:

In [None]:
t = 12345, 54321, 'hello!'
t[0]

Often they will be written with a paranthesis around the elements, 
because that is how Python prints them out:

In [None]:
t = (12345, 54321, 'hello!')
t

As you see, on output tuples are always enclosed in parentheses, so that nested tuples are interpreted correctly; they may be input with or without surrounding parentheses, although often parentheses are necessary anyway (if the tuple is part of a larger expression). It is not possible to assign to the individual items of a tuple, however it is possible to create tuples which contain mutable objects, such as lists.

Unlike lists, Tuples are immutable (the following gives a TypeError):

In [None]:
t[0] = 88888

Though tuples may seem similar to lists, they are often used in different situations and for different purposes. Tuples are immutable, and usually contain a heterogeneous sequence of elements that are accessed via unpacking or indexing. Lists are mutable, and their elements are usually homogeneous and are accessed by iterating over the list.

## Sets

Python also includes a data type for sets. A set is an unordered collection with no duplicate elements. Basic uses include membership testing and eliminating duplicate entries. Set objects also support mathematical operations like union, intersection, difference, and symmetric difference.

Curly braces or the `set()` function can be used to create sets. Note: to create an empty set you have to use `set()`, not `{}`; the latter creates an empty dictionary, a data structure that we discuss in the next section.

Here is a brief demonstration:

In [None]:
basket = {'apple', 'orange', 'apple', 'pear', 'orange', 'banana'}
print(basket)  # show that duplicates have been removed

In [None]:
'orange' in basket 

In [None]:
a = set('abracadabra')
b = set('alacazam')

In [None]:
a       # unique letters in a

In [None]:
a - b   # letters in a but not in b

In [None]:
a | b   # letters in a or b or both

In [None]:
a & b   # letters in both a and b

In [None]:
a ^ b   # letters in a or b but not both

Similarly to list comprehensions, set comprehensions are also supported:

In [None]:
a = {x for x in 'abracadabra' if x not in 'abc'}
a

## Dictionaries

Another useful data type built into Python is the dictionary. Dictionaries are sometimes found in other languages as “associative memories” or “associative arrays”. Unlike sequences, which are indexed by a range of numbers, dictionaries are indexed by keys, which can be any immutable type; strings and numbers can always be keys. Tuples can be used as keys if they contain only strings, numbers, or tuples; if a tuple contains any mutable object either directly or indirectly, it cannot be used as a key. You can’t use lists as keys, since lists can be modified in place using index assignments, slice assignments, or methods like `append()` and `extend()`.

It is best to think of a dictionary as a set of key: value pairs, with the requirement that the keys are unique (within one dictionary). A pair of braces creates an empty dictionary: `{}`. Placing a comma-separated list of key:value pairs within the braces adds initial key:value pairs to the dictionary; this is also the way dictionaries are written on output.

The main operations on a dictionary are storing a value with some key and extracting the value given the key. It is also possible to delete a key:value pair with del. If you store using a key that is already in use, the old value associated with that key is forgotten. It is an error to extract a value using a non-existent key.

Performing list(d) on a dictionary returns a list of all the keys used in the dictionary, in insertion order (if you want it sorted, just use sorted(d) instead). To check whether a single key is in the dictionary, use the in keyword.

Here is a small example using a dictionary:

In [None]:
tel = {'jack': 4098, 'sape': 4139}
tel['guido'] = 4127
tel

In [None]:
tel['jack']

In [None]:
tel['irv'] = 4127
tel

In [None]:
list(tel)

In [None]:
sorted(tel)

In [None]:
'guido' in tel

## Collections

There is a standard library Python module called **collections** that contains
many  other common data types such as `Hashable`, `Generator`, and
`OrderedDict`. One that I find particularly useful is
`Counter`, which creates an ordered dictionary from hashable
items (such as a list).

In [None]:
import collections

x = ['a', 'b', 'c', 'a', 'a', 'b', 'z']
collections.Counter(x)

## Looping techniques

When looping through dictionaries, the key and corresponding value can be retrieved at the same time using the `items()` method.

In [None]:
knights = {'gallahad': 'the pure', 'robin': 'the brave'}
for k, v in knights.items():
    print(k, v)

When looping through a sequence, the position index and corresponding value can be retrieved at the same time using the `enumerate()` function.

In [None]:
for i, v in enumerate(['tic', 'tac', 'toe']):
    print(i, v)

To loop over two or more sequences at the same time, the entries can be paired with the `zip()` function.

In [None]:
questions = ['name', 'quest', 'favorite color']
answers = ['lancelot', 'the holy grail', 'blue']
for q, a in zip(questions, answers):
    print('What is your {0}?  It is {1}.'.format(q, a))

-------

## Practice