Graph-based community clustering approach to extract protein domains from a predicted aligned error matrix

Last update: Nov 23, 2022

Related tags

Overview

pae_to_domains

Graph-based community clustering approach to extract protein domains from a predicted aligned error matrix

Overview

Using a predicted aligned error matrix corresponding to an AlphaFold2 model (e.g. as downloaded from https://alphafold.ebi.ac.uk/), returns a series of lists of residue indices, where each list corresponds to a set of residues clustering together into a pseudo-rigid domain.

Requirements

Python >=3.7
NetworkX >= 2.6.2

Known Issues

Due to an internal implementation issue in NetworkX (Issue #4992) some combinations of PAE matrix and resolution can lead to a KeyError. Solutions to this are being explored, and it will hopefully be fixed in the next NetworkX release.

Usage

While primarily intended as a code snippet to be incorporated into larger projects, this can also be called from the command line. At its simplest:

python pae_to_domains.py pae_file.json

... will yield a .csv file with each line providing the indices for one residue cluster. Full help for the command-line version:

positional arguments:
  pae_file              Name of the PAE JSON file.

optional arguments:
  -h, --help            show this help message and exit
  --output_file OUTPUT_FILE
                        Name of output file (comma-delimited text format.
                        Default: clusters.csv
  --pae_power PAE_POWER
                        Graph edges will be weighted as 1/pae**pae_power.
                        Default: 1.0
  --pae_cutoff PAE_CUTOFF
                        Graph edges will only be created for residue pairs
                        with pae



Example
Using https://alphafold.ebi.ac.uk/entry/Q9HBA0 as an example case...
resolution=0.5: 
resolution=1.0: 
resolution=2.0:

Graph-based community clustering approach to extract protein domains from a predicted aligned error matrix

Related tags

Overview

pae_to_domains

Overview

Requirements

Known Issues

Usage

Example

Owner

Tristan Croll

BalaGAN: Image Translation Between Imbalanced Domains via Cross-Modal Transfer

Self-supervised spatio-spectro-temporal represenation learning for EEG analysis

根据midi文件演奏“风物之诗琴”的脚本 "Windsong Lyre" auto play

BEAMetrics: Benchmark to Evaluate Automatic Metrics in Natural Language Generation

Self-Adaptable Point Processes with Nonparametric Time Decays

A pytorch implementation of faster RCNN detection framework (Use detectron2, it's a masterpiece)

Towards Calibrated Model for Long-Tailed Visual Recognition from Prior Perspective

StyleGAN - Official TensorFlow Implementation

PyTorch implementation of "A Two-Stage End-to-End System for Speech-in-Noise Hearing Aid Processing"

Explanatory Learning: Beyond Empiricism in Neural Networks

Algorithm to texture 3D reconstructions from multi-view stereo images

FLAVR is a fast, flow-free frame interpolation method capable of single shot multi-frame prediction

Rethinking Semantic Segmentation from a Sequence-to-Sequence Perspective with Transformers

Code release for BlockGAN: Learning 3D Object-aware Scene Representations from Unlabelled Images

Software for Multimodalty 2D+3D Facial Expression Recognition (FER) UI

Mesh TensorFlow: Model Parallelism Made Easier

Code for the ICCV 2021 Workshop paper: A Unified Efficient Pyramid Transformer for Semantic Segmentation.

Ensemble Learning Priors Driven Deep Unfolding for Scalable Snapshot Compressive Imaging [PyTorch]

Framework web SnakeServer.

Pytorch reimplementation of the Mixer (MLP-Mixer: An all-MLP Architecture for Vision)