Sign in

Antonio Camargo

@apcamargo.bsky.social
814 followers 136 following 65 posts
PostsRepliesMedia
Antonio Camargo @apcamargo.bsky.social · 15/04/2026
It’s that wonderful time of year again. A new GTDB release is out :)
0144
Reposted by Antonio Camargo
Sebastian Deorowicz @sdeorowicz.bsky.social · 14/04/2026
10 years after the first FAMSA paper, its successor is now published in Nat Biotech! We believe that FAMSA2 can enable analyses of large protein collections that were previously unattainable. Thank you, Andrzej and Cedric, for great collaboration www.nature.com/articles/s41...
nature.com
Fast and accurate multiple-protein-sequence alignment at scale with FAMSA2 - Nature Biotechnology
FAMSA2 accurately aligns millions of protein sequences at high speed.
35923
Reposted by Antonio Camargo
Simon Roux @simrouxvirus.bsky.social · 03/12/2025
🦠🧪🧬🚨 New paper and database alert: the new IMG/VR release is now MetaVR ! We have a new website - meta-virome.org - with quick search capabilities for the >24M viruses, >12M vOTUs, and >42M protein clusters (including >790k with predicted structures !). academic.oup.com/nar/advance-...
academic.oup.com
Meta-virus resource (MetaVR): expanding the frontiers of viral diversity with 24 million uncultivated virus genomes
Abstract. Viruses are ubiquitous in all environments and impact host metabolism, evolution, and ecology, although our knowledge of their biodiversity is st
16443
Reposted by Antonio Camargo
Michael Kuhn @biocs.bsky.social · 15/08/2025
We're very happy to release our new database Metalog metalog.embl.de ! It offers manually curated and harmonised contextual data for 110k metagenomics samples across the globe, incl. precomputed taxonomic profiles, for interactive browsing and for download 🧵 1/7 #microsky
metalog.embl.de
Metalog
Metalog is a repository of manually annotated metadata (or contextual data) for metagenomic sequencing data from across the globe.
37345
Reposted by Antonio Camargo
Axel Visel @axelvisel.bsky.social · 18/11/2025
Soils contain an amazing diversity of functions encoded in plasmids. The Global Soil Plasmidome Resource: 98,728 soil plasmids from 6,860 samples. Led by @mattlabguy.bsky.social and @apcamargo.bsky.social at @jgi.doe.gov @biosci.lbl.gov @berkeleylab.lbl.gov www.nature.com/articles/s41...
A stylized infographic showing the workflow for building a global soil plasmidome resource on the left and a textured world map on the right. The workflow depicts three input data streams from metagenomic datasets and isolate plasmids, which pass through steps like quality control, clustering, functional annotation, CRISPR analysis, host assignment, and detection of gene categories such as biosynthetic clusters, antimicrobial resistance, antimicrobial peptides, and CAZymes. All outputs feed into a central SQL database. The world map shows sample locations across the globe as teal circles of varying size, highlighting regions from many plasmids were recovered. Adapted from Fig 1A and 1B in Fiamenghi et al., doi:10.1038/s41467-025-65102-6
0127
Antonio Camargo @apcamargo.bsky.social · 06/11/2025
🚨New preprint out! We present a foundational genomic resource of human gut microbiome viruses. It delivers high-quality, deeply curated data spanning taxonomy, predicted hosts, structures, and functions, providing a reference for gut virome research. (1/8) www.biorxiv.org/content/10.1...
49047
Reposted by Antonio Camargo
Yunha Hwang @microyunha.bsky.social · 28/10/2025
We're thrilled to announce SeqHub, an AI-enabled platform for biological sequence analysis. SeqHub brings together sequence search, genome annotation, and data sharing in one place.
34919
Reposted by Antonio Camargo
ace-gtdb.bsky.social @ace-gtdb.bsky.social · 22/10/2025
Our @narjournal.bsky.social manuscript is out! It explores the growth of the GTDB (gtdb.ecogenomic.org) since its inception, as well as updates to the website, methodology, policies, and major taxonomic and nomenclatural changes over the past three years. academic.oup.com/nar/advance-...
academic.oup.com
GTDB release 10: a complete and systematic taxonomy for 715 230 bacterial and 17 245 archaeal genomes
Abstract. The Genome Taxonomy Database (GTDB; https://gtdb.ecogenomic.org) provides a phylogenetically consistent and rank normalized genome-based taxonomy
06946
Antonio Camargo @apcamargo.bsky.social · 06/09/2025
A BLAST update adding support for compressed files and csv output with headers is a Good Friday night surprise! blast.ncbi.nlm.nih.gov/doc/blast-ne...
blast.ncbi.nlm.nih.gov
2025 BLAST NEWS — BlastNews 0.1.1 documentation
03514
Reposted by Antonio Camargo
Rayan Chikhi @rayanchikhi.bsky.social · 03/09/2025
🌎👩‍🔬 For 15+ years biology has accumulated petabytes (million gigabytes) of🧬DNA sequencing data🧬 from the far reaches of our planet.🦠🍄🌵 Logan now democratizes efficient access to the world’s most comprehensive genetics dataset. Free and open. doi.org/10.1101/2024...
3218118
Reposted by Antonio Camargo
Ben J Woodcroft @benjwoodcroft.bsky.social · 16/07/2025
Out in @natbiotech.nature.com: Metagenome taxonomy profilers usually ignore unknown species. SingleM is an accurate profiler which doesn't, even detecting phyla with no MAGs. Profiles of 700,000 metagenomes at sandpiper.qut.edu.au. A 🧵
Logo for the Sandpiper website
713471
Reposted by Antonio Camargo
Bonsai Sequence Bioinformatics @bonsaiseqbioinfo.bsky.social · 02/07/2025
Preprint alert from the group 🚨 super fast grep-like sequence selection
065
Reposted by Antonio Camargo
Cameron Thrash @jcamthrash.bsky.social · 02/07/2025
Scaling laws of bacterial and archaeal plasmids www.nature.com/articles/s41... #jcampubs
nature.com
Scaling laws of bacterial and archaeal plasmids - Nature Communications
The capacity of a plasmid to express genes is constrained by parameters such as its length and copy number. Here, Maddamsetti et al. present a computational method that enables rapid and accurate dete...
13713
Reposted by Antonio Camargo
Robert Aboukhalil @robert.bio · 10/06/2025
Excited to announce our first interactive article on sandbox.bio, about genomic ranges: sandbox.bio/concepts/gen... Move & resize the ranges to see how that affects bedtools operations like merge and intersect in real time!
15018
Reposted by Antonio Camargo
Simon Roux @simrouxvirus.bsky.social · 13/06/2025
New pre-print out \o/ All about CRISPR, metagenomes, and what you learn when you collect (a lot of) spacers from natural communities, with @apcamargo.bsky.social @urineri.bsky.social @lhug.bsky.social but also Uri Gophna, Nikhil George (not on Bsky I think) & others at JGI doi.org/10.1101/2025...
doi.org
39245
Reposted by Antonio Camargo
Sebastian Schmidt @tsbschm.bsky.social · 03/06/2025
The Koonin Law of Computatoinal Biology: Whenever you think you have a great idea in computational or evolutionary biology, it will already have been published by Eugene Koonin in the mid 90ies.
36910
Reposted by Antonio Camargo
Yunha Hwang @microyunha.bsky.social · 02/06/2025
At Tatta Bio, we have been thinking deeply about the sequence-to-function problem. We believe that before AI can power functional prediction, we first need to rethink how we curate, manage, and share sequence data. Here, we share our initial ideas on what we are building next:
tattabio.substack.com
Today's sequence data infrastructure is set up for failure in the age of AI.
Building an open and collaborative sequence platform for both Human and AI scientists.
184
Reposted by Antonio Camargo
Jim Shaw @jimshaw.bsky.social · 28/05/2025
Announcing myloasm, a new long-read (ONT R10/PacBio) metagenome assembler that I've been working on during my postdoc in the Heng Li lab (@lh3lh3.bsky.social). myloasm-docs.github.io
myloasm-docs.github.io
myloasm - metagenomic assembly with (noisy) long reads
513378
Reposted by Antonio Camargo
Tábita Hünemeier @hunemeier.bsky.social · 15/05/2025
Our new paper is out in @science.org! By exploring the rich genetic diversity of Brazil, we show how fine-scale genomic analyses reveal that this diversity, rooted in Indigenous ancestry and centuries of complex demographic history, plays a key role in population health.
3236
Reposted by Antonio Camargo
Eduardo Amorim @cegamorim.bsky.social · 15/05/2025
Very proud of my colleagues and friends for their amazing publication out in Science today! Nunes et al. "Admixture’s impact on Brazilian population evolution and health" www.science.org/doi/10.1126/... @hunemeier.bsky.social @macscastro.bsky.social 👏
science.org
Admixture’s impact on Brazilian population evolution and health
Brazil, the largest Latin American country, is underrepresented in genomic research despite boasting the world’s largest recently admixed population. In this study, we generated 2723 high-coverage who...
02110
Reposted by Antonio Camargo
STCmicrobeblog @stcmicrobeblog.bsky.social · 28/04/2025
schaechter.asmblog.org/schaechter/2... #MicroSky #Archaea #ArchaeaSky #SymbioSky
schaechter.asmblog.org
Cultivating the Ancestors (4|4)
by Christoph — Having one or two 'Asgard' archaea under the microscope – after having cultivated them with great effort and even more patience – and looking them in the face is exciting, but a bit uns...
12011
Reposted by Antonio Camargo
Oliver Schwengers @oschwengers.bsky.social · 28/04/2025
We happily present: “Bakta Web – rapid and standardized genome annotation on scalable infrastructures” @OxUniPress NAR’s Web Server issue doi.org/10.1093/nar/... Easy to use, no registration, fast, scalable, various visualizations, in sync with Bakta CLI: bakta.computational.bio (1/5)
doi.org
Bakta Web – rapid and standardized genome annotation on scalable infrastructures
Abstract. The Bakta command line application is widely used and one of the most established tools for bacterial genome annotation. It balances comprehensiv
14228
Reposted by Antonio Camargo
Alex Crits-Christoph @acritschristoph.bsky.social · 25/04/2025
Unique investigation of some errors in long read assemblers. In particular these remarkably chimeric contigs are 😱, if rare.... Improving long read assemblers is definitely the space to be in when it comes to the future of metagenomics, as short reads won't be part of it 😉
36438
Reposted by Antonio Camargo
Ben J Woodcroft @benjwoodcroft.bsky.social · 17/04/2025
A 1.0 release for Sandpiper. 700,000 microbial community profiles (3x the last version, 4.7 Pbp metaG), searchable via the @ace-gtdb.bsky.social R226 taxonomy that just dropped. MetaGs are going exponential, but we are still nowhere near a MAG for all species. sandpiper.qut.edu.au #microsky 🧬🖥️ 1/2
Growth of public metagenomes over time
34834
Reposted by Antonio Camargo
ace-gtdb.bsky.social @ace-gtdb.bsky.social · 18/04/2025
GTDB release 10 based on RefSeq 226 (R10-RS226) is live at gtdb.ecogenomic.org. This release covers 732,475 genomes (22% increase) and has 143,6141 species clusters (37% increase). Release notes at: forum.gtdb.ecogenomic.org/t/announcing.... Release statistics at: gtdb.ecogenomic.org/stats/r226.
gtdb.ecogenomic.org
GTDB - Genome Taxonomy Database
The Genome Taxonomy Database (GTDB) is an initiative to establish a standardised microbial taxonomy based on genome phylogeny.
02717
Reposted by Antonio Camargo
Itai Yanai @itaiyanai.bsky.social · 11/04/2025
Is the genome just a bag of genes? A new paper in Science now reports that for two thirds of an organisms' genes the position along the chromosome is actually very tightly constrained! Amazing work from my favorite night scientist Martin Lercher and his team! www.science.org/doi/10.1126/...
110230
Reposted by Antonio Camargo
Cameron Thrash @jcamthrash.bsky.social · 09/04/2025
CoverM is published! CoverM: Read alignment statistics for metagenomics academic.oup.com/bioinformati... #jcampubs
academic.oup.com
CoverM: Read alignment statistics for metagenomics
AbstractSummary. Genome-centric analysis of metagenomic samples is a powerful method for understanding the function of microbial communities. Calculating r
17231
Reposted by Antonio Camargo
Kaçar Lab at UW-Madison @kacarlab.bsky.social · 24/03/2025
🚨New paper!🚨 A comprehensive take on the origins and evolution of translation factors & how these essential players evolved across the tree of life. 🌍🧬 Led by Evrim Fer @uwmadisonmdtp.bsky.social grad student! 👏 Free access: www.sciencedirect.com/science/arti... @cp-trendsgenetics.bsky.social
39446
Reposted by Antonio Camargo
RdRp Summit @rdrpsummit.bsky.social · 19/02/2025
The abstract submission & registration for #RdRpSummit2025 is NOW OPEN!!! 🎉🙌 Do you work on RNA virus discovery using RdRps? Are you joining the #ViBioM2025 in Portugal? Consider submitting your research in one of our sessions! More info: RdRp.io
01413
Antonio Camargo @apcamargo.bsky.social · 18/02/2025
Publishing a paper in Nature while based in Brazil is an incredible achievement. Huge congratulations to the authors! I'm excited to give this a proper read :)
030
Antonio Camargo @apcamargo.bsky.social · 22/01/2025
One of the great things about bioinformatics is how open it can be. Contributing to a project you admire can actually get you involved in it :)
081
Reposted by Antonio Camargo
A. Murat Eren (Meren) @merenbey.bsky.social · 17/01/2025
I'm very happy to report that Matt Schechter's PhD work is now online as a pre-print: "Ribosomal protein phylogeography offers quantitative insights into the efficacy of genome-resolved surveys of microbial communities", www.biorxiv.org/content/10.1... Here's a little 🧵 about it.
13520
Reposted by Antonio Camargo
Eduardo Rocha @epcrocha.bsky.social · 18/01/2025
This preprint makes a point that is valid for a lot of machine learning approaches. Organisms or genes are linked by evolutionary history; they are not independent. This results in correlation between learning and test sets, and often in over-optimistic evaluations of the methods' outcome.
35927
Reposted by Antonio Camargo
Joint Genome Institute @jgi.doe.gov · 17/01/2025
In Science Advances: "We took a deep dive into over 1.8 million bacterial and archaeal genomes to see how much of their diversity we’ve actually captured. Turns out that despite all the genomes we’ve sequenced, we’ve only scratched the surface." -Dongying Wu 🖥️🧬🦠 biosciences.lbl.gov/2025/01/17/t...
biosciences.lbl.gov
Taking Stock of the Known and Unknown Microbial Space - Biosciences Area
Using publicly available genome sequence data generated over the past three decades, JGI researchers assess the known fraction of microbial diversity.
24315
Reposted by Antonio Camargo
Simon Roux @simrouxvirus.bsky.social · 07/01/2025
🚨Postdoc position(s) alert🚨 If you like to develop new bioinformatic tools for *human* microbiome analysis, are interested by virus/phages, we (JGI Viral Genomics and Microbiome Data groups) may have a great position for you ! No official job posting yet, but email or DM if you are interested !
07476
Reposted by Antonio Camargo
Knut Drescher @knutdrescher.bsky.social · 06/01/2025
We found that many bacterial species use exogenous peptidoglycan fragments - released by lysis of neighboring cells - as a general danger signal, triggering a danger response that protects bacteria against many dangers: biofilm formation. Details here 👇 www.nature.com/articles/s41...
nature.com
Bacteria use exogenous peptidoglycan as a danger signal to trigger biofilm formation - Nature Microbiology
Peptidoglycan released by neighbouring kin or non-kin cell lysis induces physiological changes that protect from a range of stresses, including phage predation.
014256
Reposted by Antonio Camargo
Ryan Wick @rrwick.bsky.social · 31/12/2024
New year, new assemblies! I'm excited to announce Autocycler, my new tool for consensus assembly of long-read bacterial genomes! It's the successor to Trycycler, designed to be faster and less reliant on user intervention. Check it out: github.com/rrwick/Autoc... (1/5)
github.com
Home
A tool for generating consensus long-read assemblies for bacterial genomes - rrwick/Autocycler
215697
Reposted by Antonio Camargo
Mart Krupovic @mkrupovic.bsky.social · 30/12/2024
RNA virologists, check out "The protein structurome of Orthornavirae and its dark matter" by Pascal Mutz, Valerian Dolja, Eugene Koonin et al (including @simrouxvirus.bsky.social, @apcamargo.bsky.social, @urineri.bsky.social, @anamarijabutkovic.bsky.social) journals.asm.org/doi/10.1128/...
journals.asm.org
The protein structurome of Orthornavirae and its dark matter | mBio
Advanced methods for protein structure prediction, such as AlphaFold2, greatly expand our capability to identify protein domains and infer their likely functions and evolutionary relationships. This i...
02210
Reposted by Antonio Camargo
Gav Armstrong @nchemgav.bsky.social · 10/12/2024
Don't use red and green data lines/surfaces in the same panel please #chemsky. It can be difficult for some colorblind readers to differentiate them. I've accepted (in principle) 2 papers today, and both sets of authors were asked to remove red/green colour contrasts www.nature.com/articles/d41...
nature.com
Colour me better: fixing figures for colour blindness
Images can be made more accessible by choosing hues, shapes and textures carefully.
411347
Reposted by Antonio Camargo
Mart Krupovic @mkrupovic.bsky.social · 17/12/2024
An interesting paper for the EV (extracellular vesicle) fans by Patel et al. @ahnaskop.bsky.social lab: "Extracellular vesicles, including large translating vesicles called midbody remnants, are released during the cell cycle" www.molbiolcell.org/doi/10.1091/...
0113
Reposted by Antonio Camargo
Yunha Hwang @microyunha.bsky.social · 17/12/2024
Can LLM agents discover novel protein functions? Introducing Gaia Agent 🌎 🤖: an AI biologist capable of reasoning across genomic contexts to predict functions of proteins! Gaia Agent is now integrated with Gaia Search at gaia.tatta.bio
23813
Reposted by Antonio Camargo
Yunha Hwang @microyunha.bsky.social · 15/12/2024
If you are at #NeurIPS2024 don't miss @ancornman1.bsky.social's talk on OMG/gLM2 at 9AM! @workshopmlsb.bsky.social East meeting room 11,12
0123
Reposted by Antonio Camargo
Yunha Hwang @microyunha.bsky.social · 10/12/2024
Excited to be at #NeurIPS this week. @ancornman1.bsky.social will give a spotlight talk at the @workshopmlsb.bsky.social on gLM2/OMG! Please reach out if you want to chat about gLM2/OMG/Gaia and our latest projects😇 www.biorxiv.org/content/10.1...
biorxiv.org
The OMG dataset: An Open MetaGenomic corpus for mixed-modality genomic language modeling
Biological language model performance depends heavily on pretraining data quality, diversity, and size. While metagenomic datasets feature enormous biological diversity, their utilization as pretraini...
093
Reposted by Antonio Camargo
Rob Patro @robp.bsky.social · 04/12/2024
Very cool work from Yang Lu et al. demonstrating miscalibration of BLASTP’s E-values and generating well-calibrated values via a knockoff-based approach (cc @mikelove.bsky.social) - academic.oup.com/bioinformati...! More analyses could benefit from knockoff-based approaches.
academic.oup.com
A BLAST from the past: revisiting blastp’s E-value
AbstractMotivation. The Basic Local Alignment Search Tool, BLAST, is an indispensable tool for genomic research. BLAST established itself as the canonical
23722
Reposted by Antonio Camargo
Pierre Peterlongo @pierrepeterlongo.bsky.social · 11/11/2024
🧬🔍There are 50 petabases of freely-available DNA sequencing data. We introducing Logan Search which allows you to search for any DNA sequence in minutes, bringing Earth’s largest genomic resource to your fingertips. 🏔️ logan-search.org 🏔️ #Genomics #Bioinformatics #OpenScience
211156
Reposted by Antonio Camargo
Yunha Hwang @microyunha.bsky.social · 19/11/2024
Hello 🦋 #protein / #microbio / #BioML community! We are excited to release Gaia🌎, a context-aware protein search tool, extending protein search and discovery capabilities beyond sequence and structure, to include *genomic context*. Search your favorite protein sequences with on gaia.tatta.bio
1023875
Reposted by Antonio Camargo
Martin Steinegger 🇺🇦 @martinsteinegger.bsky.social · 15/11/2024
New GPU-based MMseqs2: 20x faster searches on a single L40S (approx. as fast as a RTX 4090) vs. a 128-core CPU. This work enables to set up a very cost-efficient ColabFold MSA GPU server. 🧵🧵 📄 www.biorxiv.org/content/10.1... 💾 mmseqs.com 🗞️ developer.nvidia.com/blog/boost-a...
217359
Reposted by Antonio Camargo
Alex Crits-Christoph @acritschristoph.bsky.social · 02/10/2024
The OMG dataset: An Open MetaGenomic corpus for mixed-modality genomic language modeling From friends at Tatta Bio GitHub: github.com/TattaBio/OMG www.biorxiv.org/content/10.1...
biorxiv.org
The OMG dataset: An Open MetaGenomic corpus for mixed-modality genomic language modeling
Biological language model performance depends heavily on pretraining data quality, diversity, and size. While metagenomic datasets feature enormous biological diversity, their utilization as pretraini...
052
Reposted by Antonio Camargo
Ákos T Kovács @evolvedbiofilm.bsky.social · 28/08/2024
Terrabacteria: redefining bacterial envelope diversity, biogenesis and evolution #NatureRevMicro from Simonetta Gribaldo www.nature.com/articles/s41...
043
Reposted by Antonio Camargo
Alex Crits-Christoph @acritschristoph.bsky.social · 22/08/2024
AntiDefenseFinder! And it is available also as an option with DefenseFinder: defensefinder.mdmlab.fr Exploring the diversity of anti-defense systems across prokaryotes, phages, and mobile genetic elements www.biorxiv.org/content/10.1...
biorxiv.org
Exploring the diversity of anti-defense systems across prokaryotes, phages, and mobile genetic elements
bioRxiv - the preprint server for biology, operated by Cold Spring Harbor Laboratory, a research and educational institution
11616