data-analysis

Affinity Groups

Logo	Name	Description	Tags	Join
	Launch	Launch is a regional computational resource that supports researchers incorporating computational and data-enabled approaches in their scientific workflows at eleven under-resourced institutions in…	ai data-analysis	Login to join

Announcements

Title	Date
Register for HTC25: Connect with the High Throughput Computing Community	03/18/25
CI Pathways: Leading the Way to Effective CI Use	02/20/25
NSF NCAR's Derecho Expands Resources For Earth Science	02/06/25

Upcoming Events & Trainings

Title	Date
ACES: Python for Data Science	4/11/25
Maximize Allocations Proposal Window Open	6/15/25
CyberInfrastruture-Enabled Machine Learning (CIML) Summer Institute 2025	6/24/25

Topics from Ask.CI

Knowledge Base Resources

Title	Category	Tags	Skill Level
AI Institutes Cyberinfrastructure Documents: SAIL Meeting	Learning	ACCESS-account ai data-analysis +1 more tags	Beginner, Intermediate, Advanced
An Introduction to the Julia Programming Language	Learning	ai data-analysis machine-learning +1 more tags	Beginner
Applications of Machine Learning in Engineering and Parameter Tuning Tutorial	Learning	data-analysis machine-learning python	Beginner, Intermediate

Engagements

Prediction of Polymerization of the Yersinia Pestis Type III Secretion System

Nova Southeastern University

Yersinia pestis, the bacterium that causes the bubonic plague, uses a type III secretion system (T3SS) to inject toxins into host cells. The structure of the Y. pestis T3SS needle has not been modeled using AI or cryo-EM. T3SS in homologous bacteria have been solved using cryo-EM. Previously, we created possible hexamers of the Y. pestis T3SS needle protein, YscF, using CollabFold and AlphaFold2 Colab on Google Colab in an effort to understand more about the needle structure and calcium regulation of secretion. Hexamers and mutated hexamers were designed using data from a wet lab experiment by Torruellas et. al (2005). T3SS structures in homologous organisms show a 22 or 23mer structure where the rings of hexamers interlocked in layers. When folding was attempted with more than six monomers, we observed larger single rings of monomers. This revealed the inaccuracies of these online systems. To create a more accurate complete needle structure, a different computer software capable of creating a helical polymerized needle is required. The number of atoms in the predicted final needle is very high and more than our computational infrastructure can handle. For that reason, we need the computational resources of a supercomputer. We have hypothesized two ways to direct the folding that have the potential to result in a more accurate needle structure. The first option involves fusing the current hexamer structure into one protein chain, so that the software recognizes the hexamer as one protein. This will make it easier to connect multiple hexamers together. Alternatively, or additionally the cryo-EM structures of the T3SS of Shigella flexneri and Salmonella enterica Typhimurium can be used as models to guide the construction of the Y. pestis T3SS needle. The full AlphaFold library or a program like RoseTTAFold could help us predict protein-protein interactions more accurately for large structures. Based on our needs we have identified the TAMU ACES, Rockfish and Stampede-2 as promising resources for this project. The generated model of the Y. pestis T3SS YscF needle will provide insight into a possible structure of the needle.

Status: In Progress