Feature representations for new Proteins in DiG
Chưa có ai nhận issue này.
Đánh giá
- Độ khó
- 4/5
- Thời gian dự kiến
- 3-5 ngày
- Mức phù hợp với người mới
- 25/100
- Loại issue
- Tài liệu
- Độ rõ ràng
- Cần làm rõ
- Mức độ hoạt động
- Đình trệ
- Công nghệ
- python
- Lĩnh vực
- bioinformatics, machine-learning
Hướng nghiên cứu
Bắt đầu với Phụ lục B.1 của bài báo và vị trí model/model.py được liên kết của OpenFold, sau đó so sánh các đầu ra Evoformer được báo cáo với các tệp pickle protein được mô tả trong issue. Ghi lại chính xác các bước FASTA, MSA, đầu ra mô hình và tuần tự hóa cần thiết để tái tạo các biểu diễn single và pair.
Do mô hình lập chỉ mục viết ra từ nội dung của issue.
Mô tả
Hi,
This is regarding protein generation in DiG.
I wanted to know how you obtained the features present in the protein pickle files. As per Appendix B.1 of the paper, the single and pair representations are simply outputs of a pre-trained Evoformer model from AlphaFold given the corresponding protein's Fasta sequence and MSAs.
I set up OpenFold on our systems and saved the representations from Evoformer in a pickle file for the corresponding protein. I used the single and pair keys in the output dictionary in this link. Also, to get the MSAs for the fasta sequence I queried the ColabFold server.
Unfortunately, the representations I received from OpenFold's Evoformer and the representations in the dataset's pickle file were quite different.
Can you please let me know the exact method you used to obtain the single and pair representations for the respective protein fasta sequence?
- Ngôn ngữ chính
- Python
- Star
- 2.5k
- Fork
- 374
- Chỉ số merge pull request
- Không có pull request nào được merge trong 30 ngày
Hướng dẫn đóng góp
Chưa lập chỉ mục được hướng dẫn đóng góp cho kho mã nguồn này
Bắt đầu từ đâu
- Đọc hết issue, rồi đọc hướng dẫn đóng góp của dự án.
- Bình luận trên issue rằng bạn sẽ nhận — tránh hai người làm cùng một việc.
- Fork repository và làm thay đổi trên một nhánh.
- Mở pull request có tham chiếu số hiệu của issue.
Issue khác của microsoft/Graphormer
-
Độ khó 2/5 1-3 giờ Mức phù hợp với người mới 72/100
microsoft/Graphormer#211 ·
-
SAS updata Đang mở
Độ khó 5/5 Hơn một tuần Mức phù hợp với người mới 20/100
microsoft/Graphormer#210 ·
-
Interpolation between states Đang mở
Độ khó 5/5 Hơn một tuần Mức phù hợp với người mới 25/100
microsoft/Graphormer#208 ·
-
CUDA out of memory Đang mở
Độ khó 4/5 3-5 ngày Mức phù hợp với người mới 20/100
microsoft/Graphormer#207 · 1 bình luận ·
-
How do I prepare OC20 dataset? Đang mở
Độ khó 3/5 1-2 ngày Mức phù hợp với người mới 35/100
microsoft/Graphormer#205 ·
Tất cả issue của microsoft/Graphormer
Issue tương tự
-
Độ khó 1/5 Dưới một giờ Mức phù hợp với người mới 75/100
-
hcocena Đang mởpolicies-accepted pre-review precheck-passed
Độ khó 1/5 Dưới một giờ Mức phù hợp với người mới 88/100
Bioconductor/BiocContributions#214 · 5 bình luận ·
-
Độ khó 1/5 Dưới một giờ Mức phù hợp với người mới 92/100
TencentCloud/Octop#1169 · 1 bình luận ·
-
[开源推荐] 在老板拷问你之前,先让 AI 灵魂拷问你 Đang mở
Độ khó 2/5 1-3 giờ Mức phù hợp với người mới 70/100
521xueweihan/HelloGitHub#3778 ·
-
The version checker's trailing attribute region has no control for a less-than inside a quoted value Đang mởarea: dashboard area: tests bug perceived difficulty: 2 python
Độ khó 2/5 1-3 giờ Mức phù hợp với người mới 84/100
Nitjsefnie-Harness-Commons/daedalus#1105 · 1 bình luận ·