Built by CATH, TÜM and NVIDIA, ProFam-1 is our new open-source protein family language model (pfLM) designed to generate functional protein variants and predict fitness using in-context example sequences.
CATH-Gene3D
@cathgene3d.bsky.social
CATH/Gene3D at University College London Evolutionary relationships and classification of protein domains. https://cathdb.info https://ted.cathdb.info
It was lovely to speak at the CATH 30 symposium, celebrating 30 years of the @cathgene3d.bsky.social protein structure classification database. I was presenting recent work on our new generative protein-family language model: preprint coming soon.
Kickstarting our symposium “Protein Annotations in the age of AI” at UCL!
CATH turns 30 years old this year! We are organising a 1-day symposium on September 16th at UCL, highlighting recent AI-based developments to enhance protein family classifications, annotations and analyses. www.eventbrite.co.uk/e/protein-an...
Protein Annotations in the age of AI
A not-for-profit symposium hosted at UCL - more details about speakers and venue below.
eventbrite.co.uk
Another CATH outing at Greenwich Park after a lovely cruise along the Thames and a pub lunch!
Our latest preprint is out on bioRxiv! A collaboration between the groups of @martinsteinegger.bsky.social , David Jones and Christine Orengo, we clustered AlphaFold Database and ESMatlas, a whopping 821 million proteins! We reveal biome-specific groups & over 11k novel domain combinations.
Metagenomic-scale analysis of the predicted protein structure universe
Protein structure prediction breakthroughs, notably AlphaFold2 and ESMfold, have led to an unprecedented influx of computationally derived structures. The AlphaFold Protein Structure Database now prov...
biorxiv.org
TED is a collaborative project between the structural bioinformatics groups of Professor David Jones & Professor Christine Orengo @cathgene3d.bsky.social at @ucl.ac.uk. The TED integration is set to enhance the interpretability and usability of #AlphaFold predictions. Is this useful in your work?
🚀 #AlphaFold Database update AlphaFold DB now integrates The Encyclopedia of Domains (TED) – a resource designed to systematically identify & classify structural domains within AlphaFold-predicted protein structures. www.ebi.ac.uk/about/news/u... @pdbeurope.bsky.social
A new version of CATH, v4.4, is out! 🎉 Here’s a link to the manuscript in NAR.
CATH v4.4: major expansion of CATH by experimental and predicted structural data
Abstract. CATH (https://www.cathdb.info) is a structural classification database that assigns domains to the structures in the Protein Data Bank (PDB) and
academic.oup.com
Hi TED! 👋 Out now in Science, we are finally ready to make our classification of all domains in AFDB v4 available to all! A joint effort with the Jones group at UCL Computer Science, here's a link to the article www.science.org/doi/10.1126/... 1/6 🧵⤵️
Exploring structural diversity across the protein universe with The Encyclopedia of Domains
The AlphaFold Protein Structure Database (AFDB) contains more than 214 million predicted protein structures composed of domains, which are independently folding units found in multiple structural and ...
science.org
Hello 🦋! We are CATH! We classify protein domains in homologous superfamilies from the Protein Data Bank and AlphaFold Database! Follow us for our latest research!