A Protein Embedding Is Not an Explanation: Protein Language Models in 2026
A practical guide to choosing, extracting, adapting and validating protein language-model representations without mistaking a system score for biological generalisation.
Pragmatic advisory and empirical analysis for machine learning in genomics, proteins, and drug discovery — what holds up on a held-out test set, what doesn't, and what it means if you build with it. The writing here is the practice thinking in the open.
Protein generators propose different biological objects. A defensible design starts with the assay, records the complete computational stack, and preserves every experimental denominator.
A practical guide to choosing, extracting, adapting and validating protein language-model representations without mistaking a system score for biological generalisation.
A task-first guide to antibody representations, structure prediction, CDR design, humanisation and developability, with the evidence and reproducibility checks that model scores leave out.
Suppose somebody hands you a FASTA file and asks for “the structure.” First decide whether the target is one chain, a protein assembly or a mixed complex containing ligands or nucleic acids.
The pitch for genomic foundation models is that one pretrained network now beats task-specific tools across the board, from regulatory annotation to clinical variant interpretation.
Carbon-3B matches Evo2-7B on sequence recovery, variant-effect prediction, and motif-perturbation discrimination while generating DNA over 150 times faster.
A 1-billion-parameter model conditioned on RNA chemical-probing reactivity folds a held-out transcript to an F1 of 0.987 against an experimentally guided reference.
Sequence models, variant-effect prediction, and the work to read non-coding DNA.
Single-sequence folding, protein language models, and structure without alignment.
Contrastive screening, protein–ligand interaction, and design that survives the wet lab.
What models encode, where they fail, and how evaluation and oversight hold up.
rewire.it works with biotech, pharma, and research teams deciding whether and how to adopt a given AI-in-bio method — feasibility reviews, evaluation design, and held-out benchmarking. Independent, candid about where machine learning helps and where it doesn't.