Efficient way to find the index of the max upper triangular entry in a numpy array?

Tags:

More specifically, I have a list of rows/columns that need to be ignored when choosing the max entry. In other words, when choosing the max upper triangular entry, certain indices need to skipped. In that case, what is the most efficient way to find the location of the max upper triangular entry?

For example:

>>> a
array([[0, 1, 1, 1],
       [1, 2, 3, 4],
       [4, 5, 6, 6],
       [4, 5, 6, 7]])
>>> indices_to_skip = [0,1,2]

I need to find the index of the min element among all elements in the upper triangle except for the entries a[0,1], a[0,2], and a[1,2].

965

asked Sep 01 '13 18:09

methane

1 Answers

You can use np.triu_indices_from:

>>> np.vstack(np.triu_indices_from(a,k=1)).T
array([[0, 1],
       [0, 2],
       [0, 3],
       [1, 2],
       [1, 3],
       [2, 3]])

>>> inds=inds[inds[:,1]>2] #Or whatever columns you want to start from.
>>> inds
array([[0, 3],
       [1, 3],
       [2, 3]])


>>> a[inds[:,0],inds[:,1]]
array([1, 4, 6])

>>> max_index = np.argmax(a[inds[:,0],inds[:,1]])
>>> inds[max_index]
array([2, 3]])

Or:

>>> inds=np.triu_indices_from(a,k=1)
>>> mask = (inds[1]>2) #Again change 2 for whatever columns you want to start at.
>>> a[inds][mask]
array([1, 4, 6])

>>> max_index = np.argmax(a[inds][mask])
>>> inds[mask][max_index]
array([2, 3]])

For the above you can use inds[0] to skip certains rows.

To skip specific rows or columns:

def ignore_upper(arr, k=0, skip_rows=None, skip_cols=None):
    rows, cols = np.triu_indices_from(arr, k=k)

    if skip_rows != None:
        row_mask = ~np.in1d(rows, skip_rows)
        rows = rows[row_mask]
        cols = cols[row_mask]

    if skip_cols != None:
        col_mask = ~np.in1d(cols, skip_cols)
        rows = rows[col_mask]
        cols = cols[col_mask]

    inds=np.ravel_multi_index((rows,cols),arr.shape)
    return np.take(arr,inds)

print ignore_upper(a, skip_rows=1, skip_cols=2) #Will also take numpy arrays for skipping.
[0 1 1 6 7]

The two can be combined and creative use of boolean indexing can help speed up specific cases.

Something interesting that I ran across, a faster way to take upper triu indices:

def fast_triu_indices(dim,k=0):

    tmp_range = np.arange(dim-k)
    rows = np.repeat(tmp_range,(tmp_range+1)[::-1])

    cols = np.ones(rows.shape[0],dtype=np.int)
    inds = np.cumsum(tmp_range[1:][::-1]+1)

    np.put(cols,inds,np.arange(dim*-1+2+k,1))
    cols[0] = k
    np.cumsum(cols,out=cols)
    return (rows,cols)

Its about ~6x faster although it does not work for k<0:

dim=5000
a=np.random.rand(dim,dim)

k=50
t=time.time()
rows,cols=np.triu_indices(dim,k=k)
print time.time()-t
0.913508892059

t=time.time()
rows2,cols2,=fast_triu_indices(dim,k=k)
print time.time()-t
0.16515994072

print np.allclose(rows,rows2)
True

print np.allclose(cols,cols2)
True

194

answered Oct 29 '22 11:10

Daniel

Related questions
                            
                                How to add a command hook into setuptools setup?
                            
                                How to load an object into memory for entire django project to see?
                            
                                Custom django tag returning a list?
                            
                                EOF when using pexpect and pxssh
                            
                                Django's AUTH_PROFILE_MODULE changing the login success url?
                            
                                Homebrew: Python version still 2.7.4 after homebrew install
                            
                                Relative Import Confusion in Python
                            
                                Windows named pipes in practice
                            
                                psycopg2.ProgrammingError: syntax error at or near "\"
                            
                                Exception signature
                            
                                Apply function to a MultiIndex dataframe with pandas/python
                            
                                Reordering columns/rows of a pivot_table?
                            
                                How to find backlinks in a website with python [closed]
                            
                                How can I specify the version of Python that Perl's Inline::Python module is using?
                            
                                Resizing a 3D image (and resampling)
                            
                                python xmpp simple client error
                            
                                How do I start with Django ORM to easily switch to SQLAlchemy?
                            
                                Determine arguments where two numpy arrays intersect in Python
                            
                                Python - Test if Raw-Input has no entry
                            
                                How to execute awk command by python code

Donate For Us

If you love us? You can donate to us via Paypal or buy me a coffee so we can maintain and grow! Thank you!

Donate Us With

Efficient way to find the index of the max upper triangular entry in a numpy array?

Tags:

python

arrays

numpy

methane

People also ask

1 Answers

Daniel

Recent Activity

Donate For Us