Ikke's blog

'cause this is what I do

Python ‘all’ odity

[update] Question solved, see bottom of post.

Since Python 2.5 the language got a new built-in method ‘all’ (and it’s nephew ‘any’). I wanted to play around with this a little, combined with generators, so I created a little testcase to test performance.

Here’s the test-case: take a list L of X random numbers in a given range [A, B], and check whether

all elements in L are >= A
all elements in L are >= (A + Z) where Z is a number in [0, (B - A)]

The first test should always result True, the second test could result to False.

Here’s the output of a test-run:

In [1]: import random, sys

In [2]: a = [random.randint(100, sys.maxint) for i in xrange(2000000)]

In [3]: len(a)
Out[3]: 2000000

In [4]: #Check whether all elements are >= 100 

In [5]: %timeit all(i >= 100 for i in a)
10 loops, best of 3: 515 ms per loop

In [6]: %timeit any(i < 100 for i in a)
10 loops, best of 3: 454 ms per loop

In [7]: def f(l):
   ...:     for i in l:
   ...:         if i < 100:
   ...:             return False
   ...:     return True
   ...: 

In [8]: %timeit f(a)
10 loops, best of 3: 292 ms per loop

In [9]: #Same thing for 100000, since now the list shouldn't be completely iterated

In [10]: %timeit all(i >= 100000 for i in a)
100 loops, best of 3: 4.73 ms per loop

In [11]: %timeit any(i < 100000 for i in a)
100 loops, best of 3: 4.29 ms per loop

In [12]: def g(l):
   ....:     for i in l:
   ....:         if i < 100000:
   ....:             return False
   ....:     return True
   ....: 

In [13]: %timeit g(a)
100 loops, best of 3: 2.82 ms per loop

In [14]: #For reference

In [15]: %timeit False in (i >= 100 for i in a)
10 loops, best of 3: 531 ms per loop

In [16]: %timeit False in (i >= 100000 for i in a)
100 loops, best of 3: 5.03 ms per loop

It’s as if ‘all’, ‘any’ or ‘in’ don’t break/return when a first occurence of False (or True, obviously) is found. Is this the desired behaviour, and if it is, why? The calculation time difference between using all/any/in or a custom-made function (which is, unlike all etc, not written in C) which breaks whenever it can, is pretty astonishing.

[update] Question solved. It’s pretty normal the function-based approach performs better, since it combines what ‘all’ and the generator provided to ‘all’ do, taking away the generator function-call overhead. Damn

Posted in Development.

Tagged with performance, python.

By Nicolas – May 1, 2008

3 Responses

Stay in touch with the conversation, subscribe to the RSS feed for comments on this post.

Manlome says

Hmm, hier snap ik niets van, maar bedankt voor je berichtje op mn blog en succes met de cello! (Staat ook nog op mijn lijstje)

May 31, 2008, 09:54 Reply
Nicke says

Ik kan goed programmeren, maar toch niet in deze taal. I like

Bedankt voor je berichtje

July 3, 2008, 14:15 Reply
jwm says

Today I was suspecting the same thing, that all() did not bail early.
Here is some code that demonstrates that it does bail on the first False

x = [] def testnum5(n): x.append(1) return n==5
print all( (testnum5(i) for i in [5,5,5,1,5,5]) ) print 'testnum5() called:',sum(x) # should be 4 if all() bails early, which it is

October 19, 2010, 22:06 Reply

« Python if/else in lambda GUADEC »

@donsbot Any reason Data.ByteString.zipWith' isn't exported/public API? 12:30:27 AM December 17, 2011 from web in reply to donsbot Reply Retweet Favorite
RT @raichoo: "#Ocaml has an OOP extension… that nobody uses […] not even its inventor" -Yaron Minsky, Janestreet. 09:28:04 PM December 16, 2011 from web Reply Retweet Favorite
Whenever you launch a tech startup, don't look for office space: head to Starbucks. Coffee, power & free wifi, all you need is a laptop! 06:30:52 PM December 16, 2011 from web Reply Retweet Favorite
@viktorklang Bit me a couple of time in the Haskell bindings as well. Simple approach to make a library "threadsafe"? 03:19:38 PM December 16, 2011 from web in reply to viktorklang Reply Retweet Favorite
W00t ^_^ RT @johtib Screenshot teaser for something I've been working on lately: http://t.co/OyGog474 01:19:31 PM December 16, 2011 from web Reply Retweet Favorite

@eikke

Proudly powered by WordPress and Carrington.

Python ‘all’ odity

3 Responses

About Ikke's blog

Meta

Friends

Me

Planets

License

Me @ Twitter

Python ‘all’ odity

3 Responses

Subscribe

About Ikke's blog

Meta

Friends

Me

Planets

License

Me @ Twitter

Tags