If I want to know whether signal of protein X at protein Y peaks is significantly/not significant higher/lower value than random regions in the genome,
I have 1000 peaks of protein Y with length 1000bp in mouse genome and calculate the average signal of protein X ChIP-seq at these peaks
As control I shuffle these peaks 1000 times randomly in the genome and calculate the average signal of protein X ChIP-seq at these shuffled peaks
Therefore I have observation table:
Peak Signal
Peak_A 100
Peak_B 150
Peak_C 50
etc
And shuffle table:
Peak Signal_1 Signal_2 Signal_3 ... Signal_1000
Peak_A 99 94 14 ... 50
Peak_B 50 90 100 ... 25
Peak_C 25 60 35 ... 40
What is the best method and how to calculate p-value? Why?
Thank you
I have 1000 peaks of protein Y with length 1000bp in mouse genome and calculate the average signal of protein X ChIP-seq at these peaks
As control I shuffle these peaks 1000 times randomly in the genome and calculate the average signal of protein X ChIP-seq at these shuffled peaks
Therefore I have observation table:
Peak Signal
Peak_A 100
Peak_B 150
Peak_C 50
etc
And shuffle table:
Peak Signal_1 Signal_2 Signal_3 ... Signal_1000
Peak_A 99 94 14 ... 50
Peak_B 50 90 100 ... 25
Peak_C 25 60 35 ... 40
What is the best method and how to calculate p-value? Why?
Thank you
Comment