Python Tutorial
NumPy Random Numbers
Use the Generator API for reproducible random integers, floats, and choices.
Generator
import numpy as np
rng = np.random.default_rng(seed=42)
print(rng.integers(0, 100, size=5))
print(rng.random(3))
print(rng.choice([10, 20, 30], size=4, replace=True))The same seed produces the same sequence — useful for tests and tutorials.
Distributions
print(rng.normal(loc=0, scale=1, size=5))
print(rng.uniform(1, 10, size=(2, 3)))Shuffle
arr = np.array([1, 2, 3, 4, 5])
rng.shuffle(arr) # in place
print(arr)
print(rng.permutation([1, 2, 3, 4, 5])) # new array📘 Real-World Deep Dive
Knowing <strong>NumPy Random (NumPy)</strong> well is what turns NumPy from a curiosity into a daily tool — you'll reach for it in nearly every real project.
Real-Life Scenario
An end-to-end usage of NumPy Random that you'd actually see in a data pipeline or analytics notebook.
Real-Life Example
import numpy as np
rng = np.random.default_rng(42)
print(rng.integers(1, 7, size=5)) # 5 dice rolls
print(rng.normal(loc=0, scale=1, size=4))Expected Output
(see source)Common mistakes
- NumPy uses 0-based, C-order indexing — the rightmost axis is the *fastest-varying* one. Mixing it with Fortran-order arrays is a common surprise.
np.array([[1,2],[3,4]], dtype=int)is fine, but a ragged Python list produces dtype=object and silently disables vectorisation.- In-place ops (
a *= 2) sometimes break views instead of returning a new array; usenp.multiply(a, 2, out=...)if explicitness matters. - Treating NumPy Random as a black box without reading the docs — the API has subtle defaults that bite when you scale.
🚀 Performance & Best Practices
- Vectorise: replace Python
forloops with ufuncs; you can expect 10–100× speedups. - Pre-allocate output arrays with
np.emptyinstead of growing them withnp.append. - Keep data in float32 unless you need float64 precision — half the memory, double the cache locality.
- When working with NumPy, prefer vectorised / batched operations over Python loops.
🧪 Try It Yourself
- Reproduce the snippet on a representative slice of your own data.
- Profile the snippet with
cProfileortimeitand find the single biggest improvement. - Generalise the snippet into a small, reusable function you can drop into future projects.