# Random AIR exclusions

**URL:** <https://ask.bioexcel.eu/t/random-air-exclusions/2198>\
**Category:** HADDOCK\
**Created:** [June 20, 2020, 9:15pm UTC](https://ask.bioexcel.eu/t/random-air-exclusions/2198 "2020-06-20T21:15:17Z")\
**Posts on this page:** 7\
**Page:** 1

<div class="post-metadata">

**Author:** ![gapm](https://avatars.discourse-cdn.com/v4/letter/g/f05b48/32.png) [@gapm](https://ask.bioexcel.eu/u/gapm)\
**Post date:** [June 20, 2020, 9:15pm UTC](https://ask.bioexcel.eu/t/random-air-exclusions/2198/1 "2020-06-20T21:15:17Z")

</div>

I am using Haddock to compare docking simulations of mutants of a protein with peptides.

I am using the same AIRs (NMR derived binding residues for each docking job). Although I have found run to run I get slightly variable results. In that haddock scores and cluster, sizes, etc differ.Even if I submit the same files with the same options (with random AIR exclusion turned on as a default). Just trying to understand what I am doing wrong or could change to remedy this?

I have been considering if the Random AIRs are actually hindering me in my particular use I know I can turn them off but I first wanted to understand if I am understanding correctly how they work. Each time I submit a job I specify a list of say 10 AIRs the random exclusion default means 50% are excluded so AIR 1,5,7,9,10 in 1 submission but in the next it could be 2,5,7,8, and 9. Or do the random exclusions vary within each job?

---

<div class="post-metadata">

**Author:** ![amjjbonvin](https://dub1.discourse-cdn.com/flex013/user_avatar/ask.bioexcel.eu/amjjbonvin/32/23_2.png) [@amjjbonvin](https://ask.bioexcel.eu/u/amjjbonvin)\
**Post date:** [June 20, 2020, 9:35pm UTC](https://ask.bioexcel.eu/t/random-air-exclusions/2198/2 "2020-06-20T21:35:09Z")

</div>

Are you using the web server? Jobs are distributed on a worldwide grid of computers. You will only get exactly the same results if you run on exactly the same hardware, which is not the case.

So small fluctuations are to be expected - but the overall picture should remain consistent.

---

<div class="post-metadata">

**Author:** ![gapm](https://avatars.discourse-cdn.com/v4/letter/g/f05b48/32.png) [@gapm](https://ask.bioexcel.eu/u/gapm)\
**Post date:** [June 20, 2020, 9:54pm UTC](https://ask.bioexcel.eu/t/random-air-exclusions/2198/3 "2020-06-20T21:54:48Z")

</div>

In some cases the differences are more than slight. I am comparing binding orders as in a recent covid ace2 example. I am using the web server. I am thinking of trialling with AIR random exclusion off but wanted to check I understood correctly?

Thanks for the quick response!

> ![](https://dub1.discourse-cdn.com/flex013/user_avatar/ask.bioexcel.eu/amjjbonvin/45/23_2.png "amjjbonvin") | [amjjbonvin](https://ask.bioexcel.eu/u/amjjbonvin)  
> June 20 |
> 
> - | - |

Are you using the web server? Jobs are distributed on a worldwide grid of computers. You will only get exactly the same results if you run on exactly the same hardware, which is not the case.

So small fluctuations are to be expected - but the overall picture should remain consistent.

---

<div class="post-metadata">

**Author:** ![amjjbonvin](https://dub1.discourse-cdn.com/flex013/user_avatar/ask.bioexcel.eu/amjjbonvin/32/23_2.png) [@amjjbonvin](https://ask.bioexcel.eu/u/amjjbonvin)\
**Post date:** [June 21, 2020, 6:50am UTC](https://ask.bioexcel.eu/t/random-air-exclusions/2198/4 "2020-06-21T06:50:13Z")

</div>

Some of there recent work on binding mutations fo ACE2 is based on only running the refinement… Not full docking.

Be careful in interpreting docking scores as binding affinities… Not much correlation there.

---

<div class="post-metadata">

**Author:** ![gapm](https://avatars.discourse-cdn.com/v4/letter/g/f05b48/32.png) [@gapm](https://ask.bioexcel.eu/u/gapm)\
**Post date:** [June 21, 2020, 9:58am UTC](https://ask.bioexcel.eu/t/random-air-exclusions/2198/5 "2020-06-21T09:58:11Z")

</div>

I was using the scores as an indication of the order of binding is this not viable? Can you provide a link or reference outlining how to use the refinement interface? I am needing to dock to get a structure to submit. I think I need to review my strategy…

---

<div class="post-metadata">

**Author:** ![amjjbonvin](https://dub1.discourse-cdn.com/flex013/user_avatar/ask.bioexcel.eu/amjjbonvin/32/23_2.png) [@amjjbonvin](https://ask.bioexcel.eu/u/amjjbonvin)\
**Post date:** [June 21, 2020, 10:37am UTC](https://ask.bioexcel.eu/t/random-air-exclusions/2198/6 "2020-06-21T10:37:17Z")

</div>

We have published quite some papers about binding affinity and scores… Check our publication list on [bonvinlab.org](http://bonvinlab.org)

You can only use refinement if you have a starting complex and only want to introduce some mutations.

---

<div class="post-metadata">

**Author:** ![gapm](https://avatars.discourse-cdn.com/v4/letter/g/f05b48/32.png) [@gapm](https://ask.bioexcel.eu/u/gapm)\
**Post date:** [June 21, 2020, 11:07am UTC](https://ask.bioexcel.eu/t/random-air-exclusions/2198/7 "2020-06-21T11:07:27Z")

</div>

Thanks I think I now understand. I will dock each of my peptides with my WT protein to get a WT complex. Then introduce mutations into the WT complex PDBs. Then use the refinement interface to see if the mutations have any effect on the scores for each peptide complex variant. I will turn off the random AIR for each full dock though as I want the same residues to be used.
