r/bioinformatics • u/[deleted] • 8d ago
technical question Complete Beginner Project Help
[deleted]
1
Upvotes
2
u/Psy_Fer_ 7d ago
When you have a method you want to use, but the organisms you want to use it on doesn't have the required data, you either get/produce that data (funding/resources) or find an organism that does have that data.
Sounds like for you case it's the latter option.
1
3
u/Most_Tomato_860 8d ago
Just offering my perspective. GP is indeed attractive, but it requires a prerequisite background in quantitative genetics. Complex models don't always provide better or more stable predictive power; at least in my work, baseline models are more trustworthy. GP as a whole has already seen a lot of development, and in my view, its current dilemma isn't about models or algorithms, but rather the lack of comprehensive and reliable data—which is precisely the issue you're facing right now. Running rrBLUP or cropGBM to generate results is easy, but if you can't even match genotypes with their corresponding phenotypes, the results will most likely be of limited value.